Cohere
75Enterprise LLM family (Command, Embed, Rerank) sold API-first with VPC, on-prem and dedicated hosting options
97 verified tools · 14 subcategories
The layer for people building AI: models, inference, vector stores, observability, and GPUs.
Showing 1–24 of 97
Enterprise LLM family (Command, Embed, Rerank) sold API-first with VPC, on-prem and dedicated hosting options
Pay-as-you-go single API for text, image, and video models from multiple AI providers.
One API key and credit balance across Veo, Sora, Kling, Seedance, Wan, Suno and image models.
One OpenAI-compatible endpoint fronting a thousand hosted models
Pay-as-you-go gateway that resells GPT, Claude, and DeepSeek access through one OpenAI-compatible API.
Single API endpoint that routes requests across 100+ AI models from OpenAI, Anthropic, Gemini, and others with automatic failover.
Prepaid-balance workspace and API for generating images, video, audio, and text from many AI models.
Connects AI agents to over 1,000 apps and tens of thousands of actions through one credential layer.
OpenAI-compatible gateway to 400+ models from one key, with routing and provider fallback.
One API key routes requests to more than 100 models from Claude, GPT, Gemini, and others.
Open-source backend platform (auth, database, storage, functions) increasingly pitched as the data layer under AI agents and RAG apps.
Flat monthly plans for LLM API access with no per-token metering, built on open models.
Turns documents, images and forms into structured JSON, with roughly 2,800 pretrained document types included
Serverless GPU cloud with per-second billing for inference endpoints, training jobs and sandboxes
Enterprise gateway serving 300+ models over one encrypted, zero-retention endpoint with per-token billing
Checkpoint and LoRA hub with an on-site generator paid in Buzz
A routing and verification layer that sits between your apps and language models to cut cost and catch ungrounded answers.
Turns documents and past sessions into a graph-plus-vector memory that AI agents can query across runs.
Gives AI agents authenticated, rate-limited access to over 1,000 external tools and apps through one API.
GPU cloud provider renting Kubernetes-native compute by the hour for AI training and inference.
Automated machine learning with model governance
Compiles a codebase into architecture docs, symbol tables, and dependency maps that AI coding agents query over MCP or an API.
Open-source control plane that provisions and schedules GPU jobs across clouds, Kubernetes and on-prem.
EU-hosted LLM gateway with an OpenAI-compatible API and routing by price, quality or latency.