Novidades da OpenRouter nos últimos 30 dias: lançamentos, mudanças de preço e remoções.
Atualizado em 23/09/2026 às 06:00383 modelos monitorados47 novos14 mudanças de preço0 removidos
Novos nas últimas semanas
Cohere: Command A+
Cohere · cohere/command-a-plus
Contexto
192K tokens
Entrada
US$ 0,30 /1M
Saída
US$ 1,50 /1M
Detectado
23/09/2026
Command A+ is Cohere's flagship model for enterprise agentic workflows. It accepts text and image inputs with a 192K context window, supports native tool calling with strict tool schemas, structured...
OpenAI: GPT-6 Luna Pro
OpenAI · openai/gpt-6-luna-pro
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,50 /1M
Detectado
22/09/2026
GPT-6 Luna Pro is the same underlying model as [GPT-6 Luna](https://openrouter.ai/openai/gpt-6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks.
Learn more in OpenAI's docs:…
OpenAI: GPT-6 Luna
OpenAI · openai/gpt-6-luna
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,50 /1M
Detectado
22/09/2026
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
OpenAI: GPT-6 Sol Pro
OpenAI · openai/gpt-6-sol-pro
Contexto
1.05M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 10,00 /1M
Detectado
22/09/2026
GPT-6 Sol Pro is the same underlying model as [GPT-6 Sol](https://openrouter.ai/openai/gpt-6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks.
Learn more in OpenAI's docs: ht…
OpenAI: GPT-6 Sol
OpenAI · openai/gpt-6-sol
Contexto
1.05M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 10,00 /1M
Detectado
22/09/2026
GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...
Anthropic: Claude Opus 5.5
Anthropic · anthropic/claude-opus-5.5
Contexto
1M tokens
Entrada
US$ 4,00 /1M
Saída
US$ 20,00 /1M
Detectado
22/09/2026
Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...
Xiaomi: MiMo-V2.6-Pro-UltraSpeed
Xiaomi · xiaomi/mimo-v2.6-pro-ultraspeed
Contexto
1.05M tokens
Entrada
US$ 4,35 /1M
Saída
US$ 8,70 /1M
Detectado
21/09/2026
MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro. Built from the same 1T MiMo-V2.6-Pro checkpoint, it matches the original model in quality while delivering roughly …
Xiaomi: MiMo-V2.6-Flash
Xiaomi · xiaomi/mimo-v2.6-flash
Contexto
1.05M tokens
Entrada
US$ 0,14 /1M
Saída
US$ 0,28 /1M
Detectado
21/09/2026
MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...
Xiaomi: MiMo-V2.6-Pro
Xiaomi · xiaomi/mimo-v2.6-pro
Contexto
1.05M tokens
Entrada
US$ 0,43 /1M
Saída
US$ 0,87 /1M
Detectado
21/09/2026
MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding...
SpaceXAI: Grok 4.7
xAI · x-ai/grok-4.7
Contexto
500K tokens
Entrada
US$ 1,60 /1M
Saída
US$ 4,80 /1M
Detectado
21/09/2026
Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...
Qwen: Qwen3.8 Omni Flash
Qwen (Alibaba) · qwen/qwen3.8-omni-flash
Contexto
1M tokens
Entrada
US$ 0,15 /1M
Saída
US$ 0,47 /1M
Detectado
21/09/2026
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...
PrismML: Ternary Bonsai 2 27B
PrismML · prism-ml/ternary-bonsai-2-27b
Contexto
262K tokens
Entrada
US$ 0,07 /1M
Saída
US$ 0,50 /1M
Detectado
18/09/2026
Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B. It supports coding, mathematics, tool calling, and image understanding with a 262K-token context window. Ternary compression shrinks.…
Z.ai: GLM 5.3 FlashX
Z.ai (Zhipu) · z-ai/glm-5.3-flashx
Contexto
1.05M tokens
Entrada
US$ 0,37 /1M
Saída
US$ 1,25 /1M
Detectado
18/09/2026
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
Pareto
unbiased · unbiased/pareto
Contexto
262K tokens
Entrada
US$ 2,50 /1M
Saída
US$ 7,50 /1M
Detectado
17/09/2026
Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad range of general-purpose tasks.
DeepSeek: DeepSeek Pro Latest
DeepSeek · ~deepseek/deepseek-pro-latest
Contexto
1.05M tokens
Entrada
US$ 0,40 /1M
Saída
US$ 4,30 /1M
Detectado
14/09/2026
This model always redirects to the latest model in the DeepSeek Pro family.
DeepSeek: DeepSeek Flash Latest
DeepSeek · ~deepseek/deepseek-flash-latest
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,50 /1M
Detectado
14/09/2026
This model always redirects to the latest model in the DeepSeek Flash family.
Inference.net: Schematron V2 Turbo
Inference.net · inference-net/schematron-v2-turbo
Contexto
128K tokens
Entrada
US$ 0,03 /1M
Saída
US$ 0,15 /1M
Detectado
11/09/2026
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in re…
Inference.net: Schematron V2 Small
Inference.net · inference-net/schematron-v2-small
Contexto
128K tokens
Entrada
US$ 0,05 /1M
Saída
US$ 0,23 /1M
Detectado
11/09/2026
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema…
OpenAI: GPT Astra Latest
OpenAI · ~openai/gpt-astra-latest
Contexto
1.05M tokens
Entrada
US$ 10,00 /1M
Saída
US$ 50,00 /1M
Detectado
11/09/2026
This model always redirects to the latest model in the GPT Astra family.
OpenAI: GPT Sol Latest
OpenAI · ~openai/gpt-sol-latest
Contexto
1.05M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 10,00 /1M
Detectado
11/09/2026
This model always redirects to the latest model in the GPT Sol family.
OpenAI: GPT Terra Latest
OpenAI · ~openai/gpt-terra-latest
Contexto
1.05M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 12,00 /1M
Detectado
11/09/2026
This model always redirects to the latest model in the GPT Terra family.
OpenAI: GPT Luna Latest
OpenAI · ~openai/gpt-luna-latest
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,50 /1M
Detectado
11/09/2026
This model always redirects to the latest model in the GPT Luna family.
Sakana: Fugu Ultra v2
Sakana · sakana/fugu-ultra-v2
Contexto
1M tokens
Entrada
US$ 5,00 /1M
Saída
US$ 30,00 /1M
Detectado
11/09/2026
Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...
Sakana: Fugu Max
Sakana · sakana/fugu-max
Contexto
1M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 6,00 /1M
Detectado
11/09/2026
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
inclusionAI: Ling 3.0 Flash VL
inclusionAI · inclusionai/ling-3.0-flash-vl
Contexto
131K tokens
Entrada
US$ 0,06 /1M
Saída
US$ 0,18 /1M
Detectado
10/09/2026
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
inclusionAI: Ling 3.0 Flash VL (free)GRÁTIS
inclusionAI · inclusionai/ling-3.0-flash-vl:free
Contexto
262K tokens
Entrada
US$ 0 /1M
Saída
US$ 0 /1M
Detectado
10/09/2026
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
DeepSeek: DeepSeek V4.1 Flash
DeepSeek · deepseek/deepseek-v4.1-flash
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,50 /1M
Detectado
10/09/2026
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
Inception: Mercury 2.5
Inception · inception/mercury-2.5
Contexto
260K tokens
Entrada
US$ 0,04 /1M
Saída
US$ 0,15 /1M
Detectado
08/09/2026
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Nex AGI: Nex-N2.5-Mini
Nex AGI · nex-agi/nex-n2.5-mini
Contexto
262K tokens
Entrada
US$ 0,03 /1M
Saída
US$ 0,10 /1M
Detectado
08/09/2026
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
Nex AGI: Nex-N2.5-Mini (free)GRÁTIS
Nex AGI · nex-agi/nex-n2.5-mini:free
Contexto
262K tokens
Entrada
US$ 0 /1M
Saída
US$ 0 /1M
Detectado
08/09/2026
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
Nex AGI: Nex-N2.5-Pro
Nex AGI · nex-agi/nex-n2.5-pro
Contexto
262K tokens
Entrada
US$ 0,07 /1M
Saída
US$ 0,25 /1M
Detectado
08/09/2026
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
Nex AGI: Nex-N2.5-Pro (free)GRÁTIS
Nex AGI · nex-agi/nex-n2.5-pro:free
Contexto
262K tokens
Entrada
US$ 0 /1M
Saída
US$ 0 /1M
Detectado
08/09/2026
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
OpenAI: GPT-6 Astra
OpenAI · openai/gpt-6-astra
Contexto
1.05M tokens
Entrada
US$ 10,00 /1M
Saída
US$ 50,00 /1M
Detectado
04/09/2026
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-hor…
OpenAI: GPT-6 Astra Pro
OpenAI · openai/gpt-6-astra-pro
Contexto
1.05M tokens
Entrada
US$ 10,00 /1M
Saída
US$ 50,00 /1M
Detectado
04/09/2026
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks.
Learn more in OpenAI's do…
Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...
Qwen: Qwen3.8 Max (0902)
Qwen (Alibaba) · qwen/qwen3.8-max-0902
Contexto
1M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 6,00 /1M
Detectado
03/09/2026
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
Meta: Muse Spark 1.3 Contributor
Meta · meta/muse-spark-1.3-contributor
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,20 /1M
Detectado
02/09/2026
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track in…
Meta: Muse Spark 1.3
Meta · meta/muse-spark-1.3
Contexto
1.05M tokens
Entrada
US$ 1,25 /1M
Saída
US$ 4,25 /1M
Detectado
02/09/2026
Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of information across extended tasks, work through...
Google: Gemini 3.8 Flash
Google · google/gemini-3.8-flash
Contexto
1.05M tokens
Entrada
US$ 0,75 /1M
Saída
US$ 3,75 /1M
Detectado
02/09/2026
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Anthropic: Claude Fable 5.1
Anthropic · anthropic/claude-fable-5.1
Contexto
1M tokens
Entrada
US$ 10,00 /1M
Saída
US$ 50,00 /1M
Detectado
01/09/2026
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
IBM: Granite 4.2 8B
IBM · ibm-granite/granite-4.2-8b
Contexto
131K tokens
Entrada
US$ 0,06 /1M
Saída
US$ 0,25 /1M
Detectado
31/08/2026
Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...
Tencent: Hy4 preview
Tencent · tencent/hy4-preview
Contexto
1.05M tokens
Entrada
US$ 0,83 /1M
Saída
US$ 2,50 /1M
Detectado
28/08/2026
Tencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total. It is designed for coding agents, complex tool-use workflows, and productivity tasks that...
inclusionAI: Ling 3.0 Flash Fin
inclusionAI · inclusionai/ling-3.0-flash-fin
Contexto
262K tokens
Entrada
US$ 0,06 /1M
Saída
US$ 0,18 /1M
Detectado
27/08/2026
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
inclusionAI: Ling 3.0 Flash Fin (free)GRÁTIS
inclusionAI · inclusionai/ling-3.0-flash-fin:free
Contexto
262K tokens
Entrada
US$ 0 /1M
Saída
US$ 0 /1M
Detectado
27/08/2026
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
Z.ai: GLM Flash Latest
Z.ai (Zhipu) · ~z-ai/glm-flash-latest
Contexto
1.31M tokens
Entrada
US$ 0,07 /1M
Saída
US$ 0,25 /1M
Detectado
27/08/2026
This model always redirects to the latest model in the GLM Flash family.
Qwen: Qwen3.8 Flash
Qwen (Alibaba) · qwen/qwen3.8-flash
Contexto
1M tokens
Entrada
US$ 0,15 /1M
Saída
US$ 0,47 /1M
Detectado
26/08/2026
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video…
Z.ai: GLM 5.3 Flash
Z.ai (Zhipu) · z-ai/glm-5.3-flash
Contexto
1.31M tokens
Entrada
US$ 0,15 /1M
Saída
US$ 0,50 /1M
Detectado
26/08/2026
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Nenhum modelo novo nos últimos 30 dias.
Mudanças de preço
Data
Modelo
Mudança (por 1M tokens)
23/09/2026
DeepSeek: DeepSeek Pro Latest DeepSeek · ~deepseek/deepseek-pro-latest