Modal
Cloud Hosting
Modal is a serverless cloud platform for running AI inference, data processing, and compute-intensive tasks with automatic scaling, GPU access, and pay-per-second pricing.
PricingUsage-based
Products tagged Usage-Based Pricing across the Cubbie marketplace.
Cloud Hosting
Modal is a serverless cloud platform for running AI inference, data processing, and compute-intensive tasks with automatic scaling, GPU access, and pay-per-second pricing.
PricingUsage-based
AI Agents & Sidekicks
Replicate is a platform for running open-source machine learning models in the cloud with a simple API, enabling developers to deploy AI without managing infrastructure.
PricingUsage-based
Cloud Hosting
Baseten provides GPU infrastructure for deploying machine learning models as production APIs, with support for open-source models, custom models, and automatic scaling.
PricingUsage-based
AI Agents & Sidekicks
Groq provides the fastest AI inference platform using custom LPU hardware, delivering ultra-low latency responses for LLM applications at competitive per-token pricing.
PricingUsage-based
AI Agents & Sidekicks
Browserbase provides a reliable, managed browser infrastructure for AI agents, enabling web scraping, automated testing, and agent-based web interactions at scale.
PricingUsage-based
Cloud Hosting
Zeabur is a cloud application platform for deploying services and databases with a simpler operational experience than assembling every cloud primitive yourself.
PricingUsage-based
AI Infrastructure
Fireworks AI provides inference infrastructure and model serving for teams that need high-performance AI workloads without operating the full stack themselves.
PricingUsage-based
Cloud Hosting
Runpod provides GPU infrastructure and deployment workflows for AI teams training, fine-tuning, and serving models without managing raw cloud complexity.
PricingUsage-based
Network Performance Monitoring
Honeycomb is an observability platform focused on high-cardinality analysis, distributed tracing, and faster debugging for complex production systems.
PricingUsage-based
Cloud Hosting
Northflank is a cloud application platform for deploying services, jobs, databases, and internal tooling with a more productized platform experience.
PricingUsage-based
Speech-to-Text Platforms
AssemblyAI provides speech-to-text, real-time transcription, and voice intelligence models for teams shipping voice AI and audio-driven workflows.
PricingUsage-based
AI Evaluation and Testing
Braintrust is an AI evaluation platform for testing prompts, models, and application behavior with production-like datasets and scoring workflows.
PricingUsage-based
AI Agents & Sidekicks
Exa is an AI-native search engine and API for applications and agents that need fresh web results, structured extraction, and research workflows.
PricingUsage-based
ETL & Data Integration
PeerDB helps teams replicate Postgres data into warehouses and downstream systems with lower-latency sync than traditional batch-heavy pipelines.
PricingUsage-based
RAG Platforms
Ragie is a managed retrieval platform for building AI products that need ingestion, indexing, syncing, and production-ready retrieval workflows.
PricingUsage-based
Speech-to-Text Platforms
Deepgram provides speech-to-text and voice understanding APIs for teams building transcription, voice workflows, and AI-powered audio products.
PricingUsage-based
LLM Orchestration
OpenRouter is a unified API and routing layer for accessing, comparing, and managing multiple AI models through one developer-facing endpoint.
PricingUsage-based
Vector Databases
Turbopuffer is a retrieval and vector database platform designed for high-scale search, embeddings, and AI application data access patterns.
PricingUsage-based