llm inference engineering

(38 offers*)
Filter
GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization for High-Throughput AI Production Systems (AI Infrastructure, Hardware & Compiler Engineering Series)
GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization for High-Throughput AI Production Systems (AI Infrastructure, Hardware & Compiler Engineering Series)
£7.48
Go to shop
amazon.co.uk
Delivery from £2.99
AI Performance Engineering: From GPU Kernels to LLM Inference
AI Performance Engineering: From GPU Kernels to LLM Inference
£32.00
Compare 3 prices
amazon.co.uk
Free Delivery
AI Performance Engineering: From GPU Kernels to LLM Inference
AI Performance Engineering: From GPU Kernels to LLM Inference
£27.00
Compare 2 prices
amazon.co.uk
Free Delivery
Production AI Inference Engineering with Claude: Build High-Performance AI Applications with Claude and LLM Infrastructure
Production AI Inference Engineering with Claude: Build High-Performance AI Applications with Claude and LLM Infrastructure
£20.00
Go to shop
amazon.co.uk
Free Delivery
LLM INFERENCE ENGINEERING: Optimizing Large Language Models on NVIDIA GPUs
LLM INFERENCE ENGINEERING: Optimizing Large Language Models on NVIDIA GPUs
£23.83
Go to shop
amazon.co.uk
Free Delivery
Production AI Inference Engineering with Claude: Build High-Performance AI Applications with Claude and LLM Infrastructure
Production AI Inference Engineering with Claude: Build High-Performance AI Applications with Claude and LLM Infrastructure
£40.00
Go to shop
amazon.co.uk
Free Delivery
AI, Volume Three: Advanced LLM Engineering - RLHF & DPO, Reasoning Models, Distributed Training, Quantization, Fast Inference, Agents, MCP, and Evaluating AI Apps
AI, Volume Three: Advanced LLM Engineering - RLHF & DPO, Reasoning Models, Distributed Training, Quantization, Fast Inference, Agents, MCP, and Evaluating AI Apps
£12.98
Go to shop
amazon.co.uk
Free Delivery
Master LLM Engineering with Rust: Build Scalable AI Applications, RAG Pipelines, Agents, and Efficient Inference Systems
Master LLM Engineering with Rust: Build Scalable AI Applications, RAG Pipelines, Agents, and Efficient Inference Systems
£28.09
Go to shop
amazon.co.uk
Free Delivery
LLM Engineer's Handbook : Master the art of engineering large language models from concept to production
LLM Engineer's Handbook : Master the art of engineering large language models from concept to production
£45.99
Go to shop
Whsmith.co.uk
Free Delivery
LLM Inference in C++: Building High-Throughput Engines with PagedAttention and CUDA Kernels (High-Performance C++ Engineering)
LLM Inference in C++: Building High-Throughput Engines with PagedAttention and CUDA Kernels (High-Performance C++ Engineering)
£22.61
Go to shop
amazon.co.uk
Free Delivery
LLM Inference in C++: Building High-Throughput Engines with PagedAttention and CUDA Kernels (High-Performance C++ Engineering)
LLM Inference in C++: Building High-Throughput Engines with PagedAttention and CUDA Kernels (High-Performance C++ Engineering)
£25.52
Go to shop
amazon.co.uk
Free Delivery
Building and Customizing Inference Engines for LLMs: From First Principles to Production - A Complete Guide to Designing High-Performance, Efficient, and Scalable LLM Inference Systems
Building and Customizing Inference Engines for LLMs: From First Principles to Production - A Complete Guide to Designing High-Performance, Efficient, and Scalable LLM Inference Systems
£18.72
Go to shop
amazon.co.uk
Free Delivery
Local LLM Inference Optimization: A Comprehensive Guide to Quantization, Hardware Acceleration, and Efficient Private AI Deployment
Local LLM Inference Optimization: A Comprehensive Guide to Quantization, Hardware Acceleration, and Efficient Private AI Deployment
£12.56
Go to shop
amazon.co.uk
Free Delivery
Mastering vLLM: Build, Optimize, and Scale High-Performance Large Language Model Inference for Production AI Applications
Mastering vLLM: Build, Optimize, and Scale High-Performance Large Language Model Inference for Production AI Applications
£13.81
Go to shop
amazon.co.uk
Free Delivery
Local LLM Inference Optimization: A Comprehensive Guide to Quantization, Hardware Acceleration, and Efficient Private AI Deployment
Local LLM Inference Optimization: A Comprehensive Guide to Quantization, Hardware Acceleration, and Efficient Private AI Deployment
£19.21
Go to shop
amazon.co.uk
Free Delivery
GGUF Model Packaging: Managing Quantized LLM Artifacts for Local Inference
GGUF Model Packaging: Managing Quantized LLM Artifacts for Local Inference
£30.31
Go to shop
amazon.co.uk
Free Delivery
LLM Performance Optimization for Engineers: A Practical System for Faster Inference, Lower Costs, Better Accuracy, and Production-Ready AI Workflows
LLM Performance Optimization for Engineers: A Practical System for Faster Inference, Lower Costs, Better Accuracy, and Production-Ready AI Workflows
£8.23
Go to shop
amazon.co.uk
Delivery from £2.99
LLM INFERENCE ENGINEERING: Optimizing Large Language Models on NVIDIA GPUs
LLM INFERENCE ENGINEERING: Optimizing Large Language Models on NVIDIA GPUs
£17.87
Go to shop
amazon.co.uk
Free Delivery

🤖 Ask ChatGPT

Informations about "llm inference engineering"

Having searched the market for the cheapest offers, 38 bids were found for comparison.

Furthermore, a large number of offers in 18 relevant categories with a price range from £7.42 to £87.00 were found.

About "llm inference engineering"

  • Overall, our search showed 3 different online shops for your product "llm inference engineering", including amazon.co.uk, Amazon-marketplace.co.uk and Whsmith.co.uk.
  • If you would prefer an item from a particular brands, you can find 0 online shops for this product. If you have not yet made a decision, you can also filter your favourite manufacturers and choose between 0 manufacturers.
  • The most bids (4) were found in the price range from £7.00 to £7.99.
  • Furthermore, other users were also interested in the following product: .
  • Personalise your product by choosing one of the 0 coloration.
Don't forget your voucher code: