# Triton

> NVIDIA's open-source inference serving software

- Category: [Inference & Compute](https://ailandscape.org/category/inference-compute) › Cloud & GPU Providers
- Homepage: https://developer.nvidia.com/triton-inference-server
- Repository: https://github.com/triton-inference-server/server
- Crunchbase: https://www.crunchbase.com/organization/nvidia
- Tags: serving, nvidia, inference
- Added to the landscape: 2026-03-18

## Similar tools in Cloud & GPU Providers

- [Akash Network](https://ailandscape.org/tool/akash-network): Decentralized cloud compute marketplace for AI workloads on GPU nodes
- [AWS Bedrock](https://ailandscape.org/tool/aws-bedrock): AWS managed service for foundation models from Anthropic, Meta, and Mistral
- [AWS SageMaker](https://ailandscape.org/tool/aws-sagemaker): Amazon's fully managed ML platform
- [Azure ML](https://ailandscape.org/tool/azure-ml): Microsoft Azure's cloud ML service
- [BentoML](https://ailandscape.org/tool/bentoml): Framework for building ML-powered services
- [Cloudflare AI](https://ailandscape.org/tool/cloudflare-ai): AI inference at the edge via Cloudflare Workers AI with zero cold starts
- [CoreWeave](https://ailandscape.org/tool/coreweave): GPU-accelerated cloud for AI/ML workloads
- [Databricks](https://ailandscape.org/tool/databricks): Unified data analytics and AI platform

---

Part of [AI Landscape](https://ailandscape.org), an open map of the AI ecosystem. Web page: https://ailandscape.org/tool/triton · Index for AI assistants: https://ailandscape.org/llms.txt
