Awan LLM
71Flat monthly plans for LLM API access with no per-token metering, built on open models.
15 verified tools
Model hosting and fast inference.
Showing 1–15 of 15
Flat monthly plans for LLM API access with no per-token metering, built on open models.
Enterprise gateway serving 300+ models over one encrypted, zero-retention endpoint with per-token billing
Trains and serves open and custom language models with per-token API pricing.
Links the machines you already own into an encrypted mesh that runs, trains and versions models locally
Ahead-of-time GPU compiler and serverless inference for ML models
On-device AI inference stack for Apple Silicon, aimed at fast, private, local model execution.
One API over 500+ image, video and audio models, billed per generation with no charge for failed jobs
Pay-as-you-go APIs for 200-plus open models, rented GPUs, and a sandbox for running agents.
Runs open-weight language models locally from one command, with a hosted tier for larger models
Runs LLMs and vision models on encrypted GPUs so enterprise data never leaves the customer's environment.
Pay-per-job inference API serving 50+ image and video models with sub-200 ms image latency
Free platform for optimizing, quantizing and deploying PyTorch or ONNX models to run on Qualcomm phones, PCs, cars and IoT chips.
Serverless API for running 100+ open image, video and audio models, billed per second
Pay-per-call hosting for 1,000+ image, video, audio and language models behind one API
Developer platform serving 1,000+ generative image, video and audio models through one metered API