Radar de Modelos — José IA

Novidades da OpenRouter nos últimos 30 dias: lançamentos, mudanças de preço e remoções.

Atualizado em 23/09/2026 às 06:00 383 modelos monitorados 47 novos 14 mudanças de preço 0 removidos

Novos nas últimas semanas

Cohere: Command A+

Cohere · cohere/command-a-plus
Contexto
192K tokens
Entrada
US$ 0,30 /1M
Saída
US$ 1,50 /1M
Detectado
23/09/2026

Command A+ is Cohere's flagship model for enterprise agentic workflows. It accepts text and image inputs with a 192K context window, supports native tool calling with strict tool schemas, structured...

OpenAI: GPT-6 Luna Pro

OpenAI · openai/gpt-6-luna-pro
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,50 /1M
Detectado
22/09/2026

GPT-6 Luna Pro is the same underlying model as [GPT-6 Luna](https://openrouter.ai/openai/gpt-6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs:…

OpenAI: GPT-6 Luna

OpenAI · openai/gpt-6-luna
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,50 /1M
Detectado
22/09/2026

GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...

OpenAI: GPT-6 Sol Pro

OpenAI · openai/gpt-6-sol-pro
Contexto
1.05M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 10,00 /1M
Detectado
22/09/2026

GPT-6 Sol Pro is the same underlying model as [GPT-6 Sol](https://openrouter.ai/openai/gpt-6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: ht…

OpenAI: GPT-6 Sol

OpenAI · openai/gpt-6-sol
Contexto
1.05M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 10,00 /1M
Detectado
22/09/2026

GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...

Anthropic: Claude Opus 5.5

Anthropic · anthropic/claude-opus-5.5
Contexto
1M tokens
Entrada
US$ 4,00 /1M
Saída
US$ 20,00 /1M
Detectado
22/09/2026

Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...

Xiaomi: MiMo-V2.6-Pro-UltraSpeed

Xiaomi · xiaomi/mimo-v2.6-pro-ultraspeed
Contexto
1.05M tokens
Entrada
US$ 4,35 /1M
Saída
US$ 8,70 /1M
Detectado
21/09/2026

MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro. Built from the same 1T MiMo-V2.6-Pro checkpoint, it matches the original model in quality while delivering roughly …

Xiaomi: MiMo-V2.6-Flash

Xiaomi · xiaomi/mimo-v2.6-flash
Contexto
1.05M tokens
Entrada
US$ 0,14 /1M
Saída
US$ 0,28 /1M
Detectado
21/09/2026

MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...

Xiaomi: MiMo-V2.6-Pro

Xiaomi · xiaomi/mimo-v2.6-pro
Contexto
1.05M tokens
Entrada
US$ 0,43 /1M
Saída
US$ 0,87 /1M
Detectado
21/09/2026

MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding...

SpaceXAI: Grok 4.7

xAI · x-ai/grok-4.7
Contexto
500K tokens
Entrada
US$ 1,60 /1M
Saída
US$ 4,80 /1M
Detectado
21/09/2026

Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...

Qwen: Qwen3.8 Omni Flash

Qwen (Alibaba) · qwen/qwen3.8-omni-flash
Contexto
1M tokens
Entrada
US$ 0,15 /1M
Saída
US$ 0,47 /1M
Detectado
21/09/2026

Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...

PrismML: Ternary Bonsai 2 27B

PrismML · prism-ml/ternary-bonsai-2-27b
Contexto
262K tokens
Entrada
US$ 0,07 /1M
Saída
US$ 0,50 /1M
Detectado
18/09/2026

Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B. It supports coding, mathematics, tool calling, and image understanding with a 262K-token context window. Ternary compression shrinks.…

Z.ai: GLM 5.3 FlashX

Z.ai (Zhipu) · z-ai/glm-5.3-flashx
Contexto
1.05M tokens
Entrada
US$ 0,37 /1M
Saída
US$ 1,25 /1M
Detectado
18/09/2026

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

Pareto

unbiased · unbiased/pareto
Contexto
262K tokens
Entrada
US$ 2,50 /1M
Saída
US$ 7,50 /1M
Detectado
17/09/2026

Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad range of general-purpose tasks.

DeepSeek: DeepSeek Pro Latest

DeepSeek · ~deepseek/deepseek-pro-latest
Contexto
1.05M tokens
Entrada
US$ 0,40 /1M
Saída
US$ 4,30 /1M
Detectado
14/09/2026

This model always redirects to the latest model in the DeepSeek Pro family.

DeepSeek: DeepSeek Flash Latest

DeepSeek · ~deepseek/deepseek-flash-latest
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,50 /1M
Detectado
14/09/2026

This model always redirects to the latest model in the DeepSeek Flash family.

Inference.net: Schematron V2 Turbo

Inference.net · inference-net/schematron-v2-turbo
Contexto
128K tokens
Entrada
US$ 0,03 /1M
Saída
US$ 0,15 /1M
Detectado
11/09/2026

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in re…

Inference.net: Schematron V2 Small

Inference.net · inference-net/schematron-v2-small
Contexto
128K tokens
Entrada
US$ 0,05 /1M
Saída
US$ 0,23 /1M
Detectado
11/09/2026

Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema…

OpenAI: GPT Astra Latest

OpenAI · ~openai/gpt-astra-latest
Contexto
1.05M tokens
Entrada
US$ 10,00 /1M
Saída
US$ 50,00 /1M
Detectado
11/09/2026

This model always redirects to the latest model in the GPT Astra family.

OpenAI: GPT Sol Latest

OpenAI · ~openai/gpt-sol-latest
Contexto
1.05M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 10,00 /1M
Detectado
11/09/2026

This model always redirects to the latest model in the GPT Sol family.

OpenAI: GPT Terra Latest

OpenAI · ~openai/gpt-terra-latest
Contexto
1.05M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 12,00 /1M
Detectado
11/09/2026

This model always redirects to the latest model in the GPT Terra family.

OpenAI: GPT Luna Latest

OpenAI · ~openai/gpt-luna-latest
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,50 /1M
Detectado
11/09/2026

This model always redirects to the latest model in the GPT Luna family.

Sakana: Fugu Ultra v2

Sakana · sakana/fugu-ultra-v2
Contexto
1M tokens
Entrada
US$ 5,00 /1M
Saída
US$ 30,00 /1M
Detectado
11/09/2026

Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...

Sakana: Fugu Max

Sakana · sakana/fugu-max
Contexto
1M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 6,00 /1M
Detectado
11/09/2026

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

inclusionAI: Ling 3.0 Flash VL

inclusionAI · inclusionai/ling-3.0-flash-vl
Contexto
131K tokens
Entrada
US$ 0,06 /1M
Saída
US$ 0,18 /1M
Detectado
10/09/2026

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

inclusionAI: Ling 3.0 Flash VL (free)GRÁTIS

inclusionAI · inclusionai/ling-3.0-flash-vl:free
Contexto
262K tokens
Entrada
US$ 0 /1M
Saída
US$ 0 /1M
Detectado
10/09/2026

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

DeepSeek: DeepSeek V4.1 Flash

DeepSeek · deepseek/deepseek-v4.1-flash
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,50 /1M
Detectado
10/09/2026

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

Inception: Mercury 2.5

Inception · inception/mercury-2.5
Contexto
260K tokens
Entrada
US$ 0,04 /1M
Saída
US$ 0,15 /1M
Detectado
08/09/2026

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

Nex AGI: Nex-N2.5-Mini

Nex AGI · nex-agi/nex-n2.5-mini
Contexto
262K tokens
Entrada
US$ 0,03 /1M
Saída
US$ 0,10 /1M
Detectado
08/09/2026

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...

Nex AGI: Nex-N2.5-Mini (free)GRÁTIS

Nex AGI · nex-agi/nex-n2.5-mini:free
Contexto
262K tokens
Entrada
US$ 0 /1M
Saída
US$ 0 /1M
Detectado
08/09/2026

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...

Nex AGI: Nex-N2.5-Pro

Nex AGI · nex-agi/nex-n2.5-pro
Contexto
262K tokens
Entrada
US$ 0,07 /1M
Saída
US$ 0,25 /1M
Detectado
08/09/2026

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...

Nex AGI: Nex-N2.5-Pro (free)GRÁTIS

Nex AGI · nex-agi/nex-n2.5-pro:free
Contexto
262K tokens
Entrada
US$ 0 /1M
Saída
US$ 0 /1M
Detectado
08/09/2026

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...

OpenAI: GPT-6 Astra

OpenAI · openai/gpt-6-astra
Contexto
1.05M tokens
Entrada
US$ 10,00 /1M
Saída
US$ 50,00 /1M
Detectado
04/09/2026

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-hor…

OpenAI: GPT-6 Astra Pro

OpenAI · openai/gpt-6-astra-pro
Contexto
1.05M tokens
Entrada
US$ 10,00 /1M
Saída
US$ 50,00 /1M
Detectado
04/09/2026

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's do…

inclusionAI: Ling 3.0 Flash Sante (free)GRÁTIS

inclusionAI · inclusionai/ling-3.0-flash-sante:free
Contexto
262K tokens
Entrada
US$ 0 /1M
Saída
US$ 0 /1M
Detectado
04/09/2026

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...

Qwen: Qwen3.8 Max (0902)

Qwen (Alibaba) · qwen/qwen3.8-max-0902
Contexto
1M tokens
Entrada
US$ 2,00 /1M
Saída
US$ 6,00 /1M
Detectado
03/09/2026

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

Meta: Muse Spark 1.3 Contributor

Meta · meta/muse-spark-1.3-contributor
Contexto
1.05M tokens
Entrada
US$ 0,10 /1M
Saída
US$ 0,20 /1M
Detectado
02/09/2026

Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track in…

Meta: Muse Spark 1.3

Meta · meta/muse-spark-1.3
Contexto
1.05M tokens
Entrada
US$ 1,25 /1M
Saída
US$ 4,25 /1M
Detectado
02/09/2026

Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of information across extended tasks, work through...

Google: Gemini 3.8 Flash

Google · google/gemini-3.8-flash
Contexto
1.05M tokens
Entrada
US$ 0,75 /1M
Saída
US$ 3,75 /1M
Detectado
02/09/2026

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Anthropic: Claude Fable 5.1

Anthropic · anthropic/claude-fable-5.1
Contexto
1M tokens
Entrada
US$ 10,00 /1M
Saída
US$ 50,00 /1M
Detectado
01/09/2026

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

IBM: Granite 4.2 8B

IBM · ibm-granite/granite-4.2-8b
Contexto
131K tokens
Entrada
US$ 0,06 /1M
Saída
US$ 0,25 /1M
Detectado
31/08/2026

Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...

Tencent: Hy4 preview

Tencent · tencent/hy4-preview
Contexto
1.05M tokens
Entrada
US$ 0,83 /1M
Saída
US$ 2,50 /1M
Detectado
28/08/2026

Tencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total. It is designed for coding agents, complex tool-use workflows, and productivity tasks that...

inclusionAI: Ling 3.0 Flash Fin

inclusionAI · inclusionai/ling-3.0-flash-fin
Contexto
262K tokens
Entrada
US$ 0,06 /1M
Saída
US$ 0,18 /1M
Detectado
27/08/2026

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

inclusionAI: Ling 3.0 Flash Fin (free)GRÁTIS

inclusionAI · inclusionai/ling-3.0-flash-fin:free
Contexto
262K tokens
Entrada
US$ 0 /1M
Saída
US$ 0 /1M
Detectado
27/08/2026

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

Z.ai: GLM Flash Latest

Z.ai (Zhipu) · ~z-ai/glm-flash-latest
Contexto
1.31M tokens
Entrada
US$ 0,07 /1M
Saída
US$ 0,25 /1M
Detectado
27/08/2026

This model always redirects to the latest model in the GLM Flash family.

Qwen: Qwen3.8 Flash

Qwen (Alibaba) · qwen/qwen3.8-flash
Contexto
1M tokens
Entrada
US$ 0,15 /1M
Saída
US$ 0,47 /1M
Detectado
26/08/2026

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video…

Z.ai: GLM 5.3 Flash

Z.ai (Zhipu) · z-ai/glm-5.3-flash
Contexto
1.31M tokens
Entrada
US$ 0,15 /1M
Saída
US$ 0,50 /1M
Detectado
26/08/2026

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Mudanças de preço

DataModeloMudança (por 1M tokens)
23/09/2026DeepSeek: DeepSeek Pro Latest
DeepSeek · ~deepseek/deepseek-pro-latest
saída: US$ 1,20 → US$ 4,30 ▲
23/09/2026Z.ai: GLM Latest
Z.ai (Zhipu) · ~z-ai/glm-latest
saída: US$ 1,76 → US$ 2,50 ▲
23/09/2026Z.ai: GLM 5.3
Z.ai (Zhipu) · z-ai/glm-5.3
entrada: US$ 0,56 → US$ 0,84 ▲ · saída: US$ 1,76 → US$ 2,64 ▲
23/09/2026DeepSeek: DeepSeek V4 Pro 0813
DeepSeek · deepseek/deepseek-v4-pro-0813
entrada: US$ 0,66 → US$ 1,32 ▲ · saída: US$ 1,98 → US$ 3,96 ▲
23/09/2026DeepSeek: DeepSeek V4 Flash Latest
DeepSeek · ~deepseek/deepseek-v4-flash-latest
entrada: US$ 0,03 → US$ 0,04 ▲ · saída: US$ 0,80 → US$ 0,55 ▼
23/09/2026Tencent: Hy3
Tencent · tencent/hy3
entrada: US$ 0,08 → US$ 0,13 ▲ · saída: US$ 0,33 → US$ 0,53 ▲
23/09/2026MoonshotAI: Kimi Latest
Moonshot · ~moonshotai/kimi-latest
saída: US$ 7,50 → US$ 10,76 ▲
23/09/2026DeepSeek: DeepSeek V4 Pro 0423
DeepSeek · deepseek/deepseek-v4-pro
entrada: US$ 0,86 → US$ 0,96 ▲ · saída: US$ 1,73 → US$ 1,91 ▲
23/09/2026DeepSeek: DeepSeek V4 Flash 0423
DeepSeek · deepseek/deepseek-v4-flash
entrada: US$ 0,05 → US$ 0,09 ▲ · saída: US$ 0,10 → US$ 0,18 ▲
22/09/2026DeepSeek: DeepSeek Flash Latest
DeepSeek · ~deepseek/deepseek-flash-latest
entrada: US$ 0,12 → US$ 0,10 ▼
22/09/2026DeepSeek: DeepSeek V4.1 Flash
DeepSeek · deepseek/deepseek-v4.1-flash
entrada: US$ 0,15 → US$ 0,10 ▼ · saída: US$ 0,60 → US$ 0,50 ▼
22/09/2026Z.ai: GLM Latest
Z.ai (Zhipu) · ~z-ai/glm-latest
entrada: US$ 0,65 → US$ 0,56 ▼ · saída: US$ 2,05 → US$ 1,76 ▼
22/09/2026Z.ai: GLM 5.3
Z.ai (Zhipu) · z-ai/glm-5.3
entrada: US$ 0,65 → US$ 0,56 ▼ · saída: US$ 2,05 → US$ 1,76 ▼
22/09/2026MoonshotAI: Kimi Latest
Moonshot · ~moonshotai/kimi-latest
saída: US$ 13,00 → US$ 7,50 ▼

Removidos

Nenhum modelo removido nos últimos 30 dias.