PortuguêsPT
Catálogo de modelos

Provedores

Os laboratórios, nuvens e roteadores cujos modelos estão no catálogo. Os preços mostrados são da OPTES, não deles.

OPTES

Laboratórios de modelos

Laboratórios chineses, plataformas internacionais

Plataformas de nuvem

Provedores de inferência

AkashMLOpen models on the decentralised Akash GPU network.7 modelosBasetenAmong the fastest providers for gpt-oss-120b and Kimi K3, with near-lowest prices.8 modelosCerebrasThe fastest published speeds in the research: about 3,000 tokens a second on gpt-oss-120b.2 modelosChutesDecentralised inference where every model runs inside a trusted execution environment (TEE).9 modelosCloudflare Workers AIModels served on Cloudflare's network, priced per 1M tokens from its Neuron units.13 modelosCoreWeaveGPU cloud with serverless inference (formerly W&B Inference) and a very recent catalog.14 modelosCrusoeManaged inference with low cached-input prices: DeepSeek, Kimi, GLM, gpt-oss and Nemotron.9 modelosDeepInfraAmong the lowest prices for gpt-oss, GLM-5.3, DeepSeek V4 Flash and MiniMax M3.15 modelosfalA marketplace of image, video and audio generation models.3 modelosFeatherlessMore than 4,000 open models and fine-tunes from Hugging Face behind one API.12 modelosFireworks AIFast serverless inference with Standard and Fast tiers for the big open models.7 modelosFriendliAIFast inference engine; a focused serverless list around GLM-5.x, DeepSeek and MiniMax.8 modelosGMI CloudGPU cloud with an inference engine for DeepSeek, Kimi, GLM, Qwen and MiniMax.11 modelosGroqVery fast inference on its own LPU chips; a small public catalog.6 modelosinference.netInference platform with open models and its own Schematron extraction models.10 modelosio.netDecentralised GPU network with IO Intelligence, an API for open models.12 modelosModalServerless GPU platform with a shared per-token API for Kimi K3, GLM-5.3 and more.4 modelosNebius Token FactoryEuropean inference cloud that publishes a measured speed for every model.14 modelosNovita AIA broad catalog of 120 open models for text, images, video and audio.17 modelosNscaleBritish AI cloud with serverless endpoints for gpt-oss, Llama and Qwen.7 modelosOVHcloud AI EndpointsEuropean serverless endpoints for gpt-oss, Llama 3.3 and Qwen.6 modelosParasailDay-zero support for big open models and some of the lowest DeepSeek and gpt-oss prices.17 modelosReplicateA large catalog of image, video and audio models; a few language models.2 modelosRunpodGPU cloud with public endpoints for Kimi models, including Kimi K3 with 1M context.4 modelosSambaNovaInference on SambaNova's own chips: DeepSeek, Llama, gpt-oss, MiniMax and Gemma.7 modelosSiliconFlowInternational platform with some of the lowest prices for DeepSeek V4.1 Flash and gpt-oss.14 modelosTogether AIOne of the widest catalogs of open-weights models: DeepSeek, Kimi, GLM, Qwen, Llama and more.12 modelos

Roteadores e gateways

A API da OPTES, o cadastro e as chaves abrem no lançamento, segunda-feira, 5 de outubro de 2026. Os modelos de terceiros chegam logo depois.