Distributed AI Systems: A practical guide to building scalable training, inference, and serving systems for production AI
R 2,411
Evolution of Programming Languages: Unleashing the Potential of Software (Technology 101 Book 6)
R 1,082
CUDA C++ in Practice: A Complete Developer's Guide to GPU Architecture, Parallel Algorithms, and Real-World Performance Optimization
R 1,123
Rust Atomics and Locks: Low-Level Concurrency in Practice
R 2,189
Asynchronous Programming with C++: Build blazing-fast software with multithreading and asynchronous programming for ultimate efficiency
R 1,771
Concurrent Go: A field book on goroutines, channels, sync, and context
R 1,123
GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization for High-Throughput AI Production Systems (AI Infrastructure, Hardware & Compiler Engineering Series)
R 1,042
CUDA C++ Optimization: Coding Faster GPU Kernels (Generative AI LLM Programming)
R 962
CUDA C++ Debugging: Safer GPU Kernel Programming (Generative AI LLM Programming)
R 962
Building Next-Generation GPU Software with CUDA 13.3 and C++26: A Comprehensive Guide to Parallel Algorithms, GPU Performance Engineering, Memory ... ... (MasterWorks Technology Series Book 5)
R 1,324