EnglishEN
Try OPTES
Model catalog

Providers

The labs, clouds and routers whose models are in the catalog. The prices shown are OPTES prices, not theirs.

OPTES

Model labs

Chinese labs, international platforms

Cloud platforms

Inference providers

AkashMLOpen models on the decentralised Akash GPU network.7 modelsBasetenAmong the fastest providers for gpt-oss-120b and Kimi K3, with near-lowest prices.8 modelsCerebrasThe fastest published speeds in the research: about 3,000 tokens a second on gpt-oss-120b.2 modelsChutesDecentralised inference where every model runs inside a trusted execution environment (TEE).9 modelsCloudflare Workers AIModels served on Cloudflare's network, priced per 1M tokens from its Neuron units.13 modelsCoreWeaveGPU cloud with serverless inference (formerly W&B Inference) and a very recent catalog.14 modelsCrusoeManaged inference with low cached-input prices: DeepSeek, Kimi, GLM, gpt-oss and Nemotron.9 modelsDeepInfraAmong the lowest prices for gpt-oss, GLM-5.3, DeepSeek V4 Flash and MiniMax M3.15 modelsfalA marketplace of image, video and audio generation models.3 modelsFeatherlessMore than 4,000 open models and fine-tunes from Hugging Face behind one API.12 modelsFireworks AIFast serverless inference with Standard and Fast tiers for the big open models.7 modelsFriendliAIFast inference engine; a focused serverless list around GLM-5.x, DeepSeek and MiniMax.8 modelsGMI CloudGPU cloud with an inference engine for DeepSeek, Kimi, GLM, Qwen and MiniMax.11 modelsGroqVery fast inference on its own LPU chips; a small public catalog.6 modelsinference.netInference platform with open models and its own Schematron extraction models.10 modelsio.netDecentralised GPU network with IO Intelligence, an API for open models.12 modelsModalServerless GPU platform with a shared per-token API for Kimi K3, GLM-5.3 and more.4 modelsNebius Token FactoryEuropean inference cloud that publishes a measured speed for every model.14 modelsNovita AIA broad catalog of 120 open models for text, images, video and audio.17 modelsNscaleBritish AI cloud with serverless endpoints for gpt-oss, Llama and Qwen.7 modelsOVHcloud AI EndpointsEuropean serverless endpoints for gpt-oss, Llama 3.3 and Qwen.6 modelsParasailDay-zero support for big open models and some of the lowest DeepSeek and gpt-oss prices.17 modelsReplicateA large catalog of image, video and audio models; a few language models.2 modelsRunpodGPU cloud with public endpoints for Kimi models, including Kimi K3 with 1M context.4 modelsSambaNovaInference on SambaNova's own chips: DeepSeek, Llama, gpt-oss, MiniMax and Gemma.7 modelsSiliconFlowInternational platform with some of the lowest prices for DeepSeek V4.1 Flash and gpt-oss.14 modelsTogether AIOne of the widest catalogs of open-weights models: DeepSeek, Kimi, GLM, Qwen, Llama and more.12 models

Routers and gateways

The OPTES API, sign-up and keys open at launch on Friday, 2 October 2026. Third-party models are coming soon after that.