
OpenRouter
LLM Orchestration
OpenRouter is a unified API and routing layer for accessing, comparing, and managing multiple AI models through one developer-facing endpoint.
PriceUsage-based
Model APIs, inference platforms, and orchestration to take an AI feature from demo to production.

LLM Orchestration
OpenRouter is a unified API and routing layer for accessing, comparing, and managing multiple AI models through one developer-facing endpoint.
PriceUsage-based

AI Observability
Langfuse is an open-source platform for tracing, prompt management, evaluations, and observability across LLM applications and agent systems.
Price$29+ / month
Price position: below this category's displayed range.
AI Observability
LangSmith is LangChain's platform for debugging, testing, evaluating, and monitoring LLM applications and agent workflows.
Price$39+ / seat / month
Starting price: below the category median.
AI Infrastructure
AI compute company building wafer-scale chips for the fastest training and inference of large language models.
Price$50+ / month
Starting price: below the category median.
AI Infrastructure
Developer platform for running AI models and applications with simple APIs and optimized serving infrastructure.
PriceUsage-based

LLM Ops
LLM operations platform for managing prompts, models, and deployments in production with collaboration tools for AI product teams.
Price$5+ / month

AI Evaluation and Testing
End-to-end AI evaluation and monitoring platform for LLM product teams with prompt management and datasets.
Price$39+ / user / month
Starting price: below the category median.
AI Infrastructure
Fireworks AI provides inference infrastructure and model serving for teams that need high-performance AI workloads without operating the full stack themselves.
AI Agents
Open-source model for generating accurate API calls from natural language, reducing LLM hallucination on tools.
MLOps Platforms
Kubernetes-based model inference platform for standardized serverless serving of machine learning models.
AI/ML Infrastructure
Hugging Face text generation inference server optimized for serving large language models in production.
Cloud Infrastructure
Global GPU cloud and AI infrastructure for training and low-latency inference.

AI Infrastructure
Open and frontier LLMs with embedding APIs.
AI Infrastructure
Databricks' Mosaic AI suite for building, tuning, and serving LLMs with RAG, evaluation, and monitoring.
PriceContact sales

AI Infrastructure
Open-source LLM serving framework with PagedAttention for high-throughput and memory-efficient inference.

AI/ML Infrastructure
Groq provides the fastest AI inference platform using custom LPU hardware, delivering ultra-low latency responses for LLM applications at competitive per-token pricing.

AI Infrastructure
Baseten provides GPU infrastructure for deploying machine learning models as production APIs, with support for open-source models, custom models, and automatic scaling.

AI Infrastructure
Serverless inference platform for running popular open-source AI models with per-token pricing and low latency.