Braintrust
AI Evaluation and Testing
Braintrust is an AI evaluation platform for testing prompts, models, and application behavior with production-like datasets and scoring workflows.
Pricing$249 / month
Curated software listings.
AI Evaluation and Testing
Braintrust is an AI evaluation platform for testing prompts, models, and application behavior with production-like datasets and scoring workflows.
Pricing$249 / month
AI Observability
LangSmith is LangChain's platform for debugging, testing, evaluating, and monitoring LLM applications and agent workflows.
Pricing$39+ / seat / month
AI Evaluation and Testing
LLM testing and evaluation platform with continuous testing pipelines.
Pricing$299+ / month
AI Evaluation and Testing
LLM evaluation platform providing automated testing and benchmarking for AI applications with custom metrics.
PricingPricing unavailable
AI Evaluation and Testing
Open-source evaluation tool for testing and benchmarking LLM applications in CI/CD pipelines.
PricingPricing unavailable
AI Evaluation and Testing
Open-source platform for prompt engineering, evaluation, and deployment of LLM applications with team collaboration.
PricingPricing unavailable
Talent Marketplace
AI recruiting platform that uses video interviews to match candidates with tech companies.
PricingPricing unavailable
AI Observability
LLM observability and evaluation platform for AI teams with prompt testing and production monitoring.
PricingPricing unavailable
AI Evaluation and Testing
Open-source tool for testing and evaluating LLM prompts, models, and RAG pipelines with side-by-side comparisons.
PricingPricing unavailable
AI Developer Tools
Developer platform for building, evaluating, and deploying generative AI applications to production.
PricingPricing unavailable
LLM Ops
Open-source library for evaluating and tracking LLM and RAG application quality with feedback functions.
PricingPricing unavailable
AI Evaluation and Testing
Open-source framework for evaluating retrieval-augmented generation pipelines with automated quality metrics.
PricingPricing unavailable
AI Evaluation and Testing
AI application development platform for testing, evaluation, and deployment of LLM systems.
PricingPricing unavailable
AI Evaluation and Testing
AI testing platform for debugging and evaluating ML and generative AI systems.
PricingPricing unavailable
AI Observability
LLM observability and evaluation platform for monitoring, debugging, and improving AI application performance.
PricingPricing unavailable