DeutschDE
Modellkatalog

Anbieter

Die Labore, Clouds und Router, deren Modelle im Katalog stehen. Die gezeigten Preise sind OPTES-Preise, nicht ihre.

OPTES

Modelllabore

Chinesische Labore, internationale Plattformen

Cloud-Plattformen

Inferenzanbieter

AkashMLOpen models on the decentralised Akash GPU network.7 ModelleBasetenAmong the fastest providers for gpt-oss-120b and Kimi K3, with near-lowest prices.8 ModelleCerebrasThe fastest published speeds in the research: about 3,000 tokens a second on gpt-oss-120b.2 ModelleChutesDecentralised inference where every model runs inside a trusted execution environment (TEE).9 ModelleCloudflare Workers AIModels served on Cloudflare's network, priced per 1M tokens from its Neuron units.13 ModelleCoreWeaveGPU cloud with serverless inference (formerly W&B Inference) and a very recent catalog.14 ModelleCrusoeManaged inference with low cached-input prices: DeepSeek, Kimi, GLM, gpt-oss and Nemotron.9 ModelleDeepInfraAmong the lowest prices for gpt-oss, GLM-5.3, DeepSeek V4 Flash and MiniMax M3.15 ModellefalA marketplace of image, video and audio generation models.3 ModelleFeatherlessMore than 4,000 open models and fine-tunes from Hugging Face behind one API.12 ModelleFireworks AIFast serverless inference with Standard and Fast tiers for the big open models.7 ModelleFriendliAIFast inference engine; a focused serverless list around GLM-5.x, DeepSeek and MiniMax.8 ModelleGMI CloudGPU cloud with an inference engine for DeepSeek, Kimi, GLM, Qwen and MiniMax.11 ModelleGroqVery fast inference on its own LPU chips; a small public catalog.6 Modelleinference.netInference platform with open models and its own Schematron extraction models.10 Modelleio.netDecentralised GPU network with IO Intelligence, an API for open models.12 ModelleModalServerless GPU platform with a shared per-token API for Kimi K3, GLM-5.3 and more.4 ModelleNebius Token FactoryEuropean inference cloud that publishes a measured speed for every model.14 ModelleNovita AIA broad catalog of 120 open models for text, images, video and audio.17 ModelleNscaleBritish AI cloud with serverless endpoints for gpt-oss, Llama and Qwen.7 ModelleOVHcloud AI EndpointsEuropean serverless endpoints for gpt-oss, Llama 3.3 and Qwen.6 ModelleParasailDay-zero support for big open models and some of the lowest DeepSeek and gpt-oss prices.17 ModelleReplicateA large catalog of image, video and audio models; a few language models.2 ModelleRunpodGPU cloud with public endpoints for Kimi models, including Kimi K3 with 1M context.4 ModelleSambaNovaInference on SambaNova's own chips: DeepSeek, Llama, gpt-oss, MiniMax and Gemma.7 ModelleSiliconFlowInternational platform with some of the lowest prices for DeepSeek V4.1 Flash and gpt-oss.14 ModelleTogether AIOne of the widest catalogs of open-weights models: DeepSeek, Kimi, GLM, Qwen, Llama and more.12 Modelle

Router und Gateways

Die OPTES-API, die Registrierung und die Schlüssel öffnen zum Start am Montag, 5. Oktober 2026. Modelle anderer Anbieter folgen kurz danach.