Skip to content
AI Landscape

Infrastructure

Inference & Compute

43 tools · 5 subcategories

The infrastructure layer that runs AI: GPU cloud providers, inference optimization libraries, distributed training infrastructure, purpose-built chips, and local LLM runtimes like llama.cpp and LM Studio. Renting compute? Compare the GPU clouds. Serving models in production? Start with vLLM and the inference optimizers. Running LLMs locally or at the edge? The local inference section covers runtimes and on-device tools. If it makes models run faster or cheaper, it's here.