Cohere
75Enterprise LLM family (Command, Embed, Rerank) sold API-first with VPC, on-prem and dedicated hosting options
49 tools in AI Infrastructure & Dev Platforms charge this way, each one checked against its own pricing page.
Freemium tools are free to keep using, with a paid tier that lifts limits or unlocks features. The free tier is the product, not a countdown.
Paid plans start between $0.99 and $250 a month.
47 of them expose an API.
All AI Infrastructure & Dev Platforms tools, any price →
Enterprise LLM family (Command, Embed, Rerank) sold API-first with VPC, on-prem and dedicated hosting options
Connects AI agents to over 1,000 apps and tens of thousands of actions through one credential layer.
OpenAI-compatible gateway to 400+ models from one key, with routing and provider fallback.
Open-source backend platform (auth, database, storage, functions) increasingly pitched as the data layer under AI agents and RAG apps.
Flat monthly plans for LLM API access with no per-token metering, built on open models.
Turns documents, images and forms into structured JSON, with roughly 2,800 pretrained document types included
Serverless GPU cloud with per-second billing for inference endpoints, training jobs and sandboxes
Checkpoint and LoRA hub with an on-site generator paid in Buzz
A routing and verification layer that sits between your apps and language models to cut cost and catch ungrounded answers.
Turns documents and past sessions into a graph-plus-vector memory that AI agents can query across runs.
Gives AI agents authenticated, rate-limited access to over 1,000 external tools and apps through one API.
Open-source control plane that provisions and schedules GPU jobs across clouds, Kubernetes and on-prem.
EU-hosted LLM gateway with an OpenAI-compatible API and routing by price, quality or latency.
Policy gate that checks every AI agent tool call before it runs and writes a signed audit trail
Enterprise control plane for AI: guardrails, observability and governance for LLM agents and ML models
Runs ComfyUI workflows in a browser on rented GPUs, billing separately for edit time and generation time.
Testing and monitoring platform for AI agents, catching hallucinations before and after they ship.
Open-source machine learning with an enterprise cloud
Hosting hub for models, datasets and demo apps
Simulates user scenarios against LLM agents, scores the answers and traces every step in production
A serverless platform for building, deploying, and monitoring AI agents across 600-plus LLMs.
Spreadsheet-style workbench for writing, testing and shipping prompts, with a firewall on the outputs
Browser Studios with on-demand GPUs (T4 to H200) for training, tuning and serving PyTorch models
Tracing, analytics and prompt versioning for LLM apps, with a self-hostable community edition.
Email inboxes and delivery infrastructure for autonomous agents
Open-source TypeScript framework for building, running, and observing AI agents in production.
Open-source search engine with vector storage and hybrid keyword-plus-semantic retrieval, self-hosted or run as a cloud service
A versioned filesystem built for AI coding agents, adding Git-style branching, checkpoints, and rollback to agent file storage.
French model lab offering open and commercial LLMs, the Vibe assistant, and an EU-hosted API and cloud
API of specialized models for coding agents—fast code-apply, code search and context compression—priced per million tokens.
Runs open-weight language models locally from one command, with a hosted tier for larger models
ML platform native to Snowflake that trains, monitors, and retrains models continually instead of from scratch each cycle.
Managed vector database for similarity search over billions of embeddings, with serverless and BYOC options
Simulation-driven testing, evaluation, and guardrails platform for AI agents before and after production.
Postgres backend bundled with a RAG pipeline, agent runtime and a visual workflow builder
Stores and organizes AI prompts across ChatGPT, Claude, Gemini and other models with tags, versioning and sharing.
Prompt versioning, evals and tracing for LLM engineering teams
Privacy-first LLM gateway that routes chat and API requests through trusted hardware.
Open-source workspace where domain experts and engineers build test sets for AI agents together.
Dataset labeling and deployment for computer vision
Tencent's model family and studio covering text, image, video and 3D, with many weights open-sourced
Registry and governance layer for the skills your coding agents run - versioned, security-scanned and evaluated
One OpenAI-compatible API that routes requests across GPT, Claude, Gemini, and other models.
Experiment tracking and evaluation for ML teams
Social data API and hosted MCP server for X, Instagram, TikTok and Reddit without platform API keys
Pay-per-call API gateway giving one endpoint to 79+ image, video, music, and text generation models.
Cloud computers for running Claude Code, Codex, and other coding agents, reachable by browser, SSH, or iMessage.
Gives each AI coding agent its own persistent cloud computer — disk, shell, and network — reachable via MCP from Claude Code or Cursor.
One REST endpoint for image, video, audio, chat and text models, billed from a prepaid credit balance