Model Information: Explore the capabilities, specifications, and pricing of all available models in botscanner.

Alibaba

Qwen3.6 35B A3B
Thinking
Tier 1
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
Parameters Unknown
Context 262,144 tokens
Released 2026-04-27
API Type openrouter
This model is available for answering and ranking.
qwen alibaba
Learn more – Model Terms of Service
Qwen3.7 Flash
Thinking
Tier Free
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...
Parameters Unknown
Context 1,000,000 tokens
Released 2026-07-27
API Type openrouter
This model is available for answering and ranking.
qwen alibaba
Learn more – Model Terms of Service
Qwen3.7 Plus
Thinking
Tier 1
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
Parameters Unknown
Context 1,000,000 tokens
Released 2026-06-03
API Type openrouter
This model is available for answering and ranking.
qwen alibaba
Learn more – Model Terms of Service
Qwen3.8 2.4T A95B
Thinking
Tier 1
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
Parameters Unknown
Context 1,000,000 tokens
Released 2026-08-12
API Type openrouter
This model is available for answering and ranking.
qwen alibaba
Learn more – Model Terms of Service
Qwen3.8 27B
Thinking
Tier 1
Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...
Parameters Unknown
Context 1,000,000 tokens
Released 2026-08-14
API Type openrouter
This model is available for answering and ranking.
qwen alibaba
Learn more – Model Terms of Service
Qwen3.8 Flash
Thinking
Tier Free
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
Parameters Unknown
Context 1,000,000 tokens
Released 2026-08-26
API Type openrouter
This model is available for answering and ranking.
qwen alibaba
Learn more – Model Terms of Service
Qwen3.8 Max (0902)
Thinking
Tier 1
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
Parameters Unknown
Context 1,000,000 tokens
Released 2026-09-03
API Type openrouter
This model is available for answering and ranking.
qwen alibaba
Learn more – Model Terms of Service
Qwen3.8 Max Prime
Thinking
Tier 1
Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...
Parameters Unknown
Context 1,000,000 tokens
Released 2026-09-23
API Type openrouter
This model is available for answering.
qwen alibaba
Learn more – Model Terms of Service

Amazon

Nova 2 Lite v1.0
Thinking
Tier 1
Powerful and cost-effective LLM for lightweight tasks. Nova 2 Lite is an advanced multimodal reasoning model that intelligently balances performance and efficiency by dynamically adjusting reasoning depth based on task complexity.
Parameters Unknown
Context 1,000,000 tokens
Released 2025-12-02
API Type openrouter
This model is available for answering and ranking.
nova amazon cost-effective thinking
Learn more – Model Terms of Service
Nova Premier 1.0
Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.
Parameters Unknown
Context 1,000,000 tokens
Released 2025-10-31
API Type openrouter
This model is available for answering.
amazon nova
Learn more – Model Terms of Service

Anthropic

Claude Fable 5.1
Thinking
Tier 1
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Parameters Unknown
Context 1,000,000 tokens
Released 2026-09-01
API Type openrouter
This model is available for answering.
anthropic claude
Learn more – Model Terms of Service
Claude Haiku 4.5
Thinking
Tier 1
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Parameters Unknown
Context 200,000 tokens
Released 2025-10-15
API Type openrouter
This model is available for answering and ranking.
anthropic claude
Learn more – Model Terms of Service
Claude Opus 5.5
Thinking
Tier 1
Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...
Parameters Unknown
Context 1,000,000 tokens
Released 2026-09-22
API Type openrouter
This model is available for answering.
anthropic claude
Learn more – Model Terms of Service
Claude Sonnet 5.5
Thinking
Tier 1
Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a direct upgrade. It is especially strong at building features, fixing bugs, and producing...
Parameters Unknown
Context 1,000,000 tokens
Released 2026-09-28
API Type openrouter
This model is available for answering and ranking.
anthropic claude
Learn more – Model Terms of Service

DeepSeek

R1 0528
Thinking
Tier 1
May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...
Parameters Unknown
Context 163,840 tokens
Released 2025-05-28
API Type openrouter
This model is available for answering and ranking.
deepseek
Learn more – Model Terms of Service
DeepSeek V3.1 Terminus
Thinking
Tier 1
DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...
Parameters Unknown
Context 131,072 tokens
Released 2025-09-22
API Type openrouter
This model is available for answering and ranking.
deepseek
Learn more – Model Terms of Service
DeepSeek V3.2
Thinking
Tier Free
DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
Parameters Unknown
Context 131,072 tokens
Released 2025-12-01
API Type openrouter
This model is available for answering and ranking.
deepseek
Learn more – Model Terms of Service
DeepSeek V4 Pro 0813
Thinking
Tier 1
DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.
Parameters Unknown
Context 1,048,576 tokens
Released 2026-08-12
API Type openrouter
This model is available for answering and ranking.
deepseek
Learn more – Model Terms of Service
DeepSeek V4.1 Flash
Thinking
Tier 1
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
Parameters Unknown
Context 1,048,576 tokens
Released 2026-09-10
API Type openrouter
This model is available for answering and ranking.
deepseek
Learn more – Model Terms of Service

Google

Gemini 3.1 Pro Preview
Thinking
Tier 1
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
Parameters Unknown
Context 1,048,576 tokens
Released 2026-02-19
API Type openrouter
This model is available for answering.
google gemini
Learn more – Model Terms of Service
Gemini 3.5 Flash Lite
Thinking
Tier 1
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
Parameters Unknown
Context 1,048,576 tokens
Released 2026-07-21
API Type openrouter
This model is available for answering and ranking.
google gemini
Learn more – Model Terms of Service
Gemini 3.8 Flash
Thinking
Tier 1
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Parameters Unknown
Context 1,048,576 tokens
Released 2026-09-02
API Type openrouter
This model is available for answering and ranking.
google gemini
Learn more – Model Terms of Service
Gemma 4 26B A4B
Thinking
Tier Free
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
Parameters Unknown
Context 262,144 tokens
Released 2026-04-03
API Type openrouter
This model is available for answering and ranking.
google gemini
Learn more – Model Terms of Service
Gemma 4 31B
Thinking
Tier Free
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Parameters Unknown
Context 262,144 tokens
Released 2026-04-02
API Type openrouter
This model is available for answering and ranking.
google gemini
Learn more – Model Terms of Service

Meta

Llama 4 Maverick
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Parameters Unknown
Context 128,000 tokens
Released 2025-04-05
API Type openrouter
This model is available for answering and ranking.
meta llama
Learn more – Model Terms of Service
Llama 4 Scout
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
Parameters Unknown
Context 327,680 tokens
Released 2025-04-05
API Type openrouter
This model is available for answering and ranking.
meta llama
Learn more – Model Terms of Service
Muse Glimmer 30B
Thinking
Tier 1
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
Parameters Unknown
Context 131,072 tokens
Released 2026-08-09
API Type openrouter
This model is available for answering and ranking.
meta llama
Learn more – Model Terms of Service
Muse Spark 1.3
Thinking
Tier 1
Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of information across extended tasks, work through...
Parameters Unknown
Context 1,048,576 tokens
Released 2026-09-02
API Type openrouter
This model is available for answering and ranking.
meta llama
Learn more – Model Terms of Service

MiniMax

MiniMax M1
Thinking
Tier 1
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...
Parameters Unknown
Context 1,000,000 tokens
Released 2025-06-17
API Type openrouter
This model is available for answering and ranking.
minimax
Learn more – Model Terms of Service
MiniMax M3
Thinking
Tier 1
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Parameters Unknown
Context 524,288 tokens
Released 2026-05-31
API Type openrouter
This model is available for answering and ranking.
minimax
Learn more – Model Terms of Service

Mistral AI

Ministral 3 14B 2512
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...
Parameters Unknown
Context 262,144 tokens
Released 2025-12-02
API Type openrouter
This model is available for answering and ranking.
mistral
Learn more – Model Terms of Service
Ministral 3 3B 2512
The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.
Parameters Unknown
Context 131,072 tokens
Released 2025-12-02
API Type openrouter
This model is available for answering and ranking.
mistral
Learn more – Model Terms of Service
Ministral 3 8B 2512
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
Parameters Unknown
Context 262,144 tokens
Released 2025-12-02
API Type openrouter
This model is available for answering and ranking.
mistral
Learn more – Model Terms of Service
Mistral Large 3 2512
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Parameters Unknown
Context 262,144 tokens
Released 2025-12-01
API Type openrouter
This model is available for answering and ranking.
mistral
Learn more – Model Terms of Service
Mistral Medium 3.5
Thinking
Tier 1
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
Parameters Unknown
Context 262,144 tokens
Released 2026-04-30
API Type openrouter
This model is available for answering and ranking.
mistral
Learn more – Model Terms of Service
Mistral Small 4
Thinking
Tier Free
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
Parameters Unknown
Context 262,144 tokens
Released 2026-03-16
API Type openrouter
This model is available for answering and ranking.
mistral
Learn more – Model Terms of Service
Mistral Small 3.2 24B
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...
Parameters Unknown
Context 256,000 tokens
Released 2025-06-20
API Type openrouter
This model is available for answering and ranking.
mistral
Learn more – Model Terms of Service

MoonShot AI

Kimi K2 Thinking
Thinking
Tier 1
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...
Parameters Unknown
Context 262,144 tokens
Released 2025-11-06
API Type openrouter
This model is available for answering and ranking.
moonshot kimi
Learn more – Model Terms of Service
Kimi K3
Thinking
Tier 1
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Parameters Unknown
Context 1,048,576 tokens
Released 2026-07-16
API Type openrouter
This model is available for answering.
moonshot kimi
Learn more – Model Terms of Service

NVIDIA

Nemotron 3 Nano 30B A3B
Thinking
Tier Free
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
Parameters Unknown
Context 262,144 tokens
Released 2025-12-14
API Type openrouter
This model is available for answering and ranking.
nvidia nemotron
Learn more – Model Terms of Service
Nemotron 3 Super
Thinking
Tier Free
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
Parameters Unknown
Context 262,144 tokens
Released 2026-03-11
API Type openrouter
This model is available for answering and ranking.
nvidia nemotron
Learn more – Model Terms of Service
Nemotron 3 Ultra
Thinking
Tier 1
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Parameters Unknown
Context 262,144 tokens
Released 2026-06-04
API Type openrouter
This model is available for answering and ranking.
nvidia nemotron
Learn more – Model Terms of Service
Nemotron 3.5 Lightning
Thinking
Tier Free
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
Parameters Unknown
Context 262,144 tokens
Released 2026-08-11
API Type openrouter
This model is available for answering and ranking.
nvidia nemotron
Learn more – Model Terms of Service

OpenAI

GPT-5.4 Nano
Thinking
Tier 1
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
Parameters Unknown
Context 400,000 tokens
Released 2026-03-17
API Type openrouter
This model is available for answering and ranking.
openai gpt
Learn more – Model Terms of Service
GPT-5.6 Terra
Thinking
Tier 1
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
Parameters Unknown
Context 1,050,000 tokens
Released 2026-07-09
API Type openrouter
This model is available for answering.
openai gpt
Learn more – Model Terms of Service
GPT-6 Astra
Thinking
Tier 1
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
Parameters Unknown
Context 1,050,000 tokens
Released 2026-09-04
API Type openrouter
This model is available for answering.
openai gpt
Learn more – Model Terms of Service
GPT-6 Luna
Thinking
Tier Free
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
Parameters Unknown
Context 1,050,000 tokens
Released 2026-09-22
API Type openrouter
This model is available for answering and ranking.
openai gpt
Learn more – Model Terms of Service
GPT-6.1 Sol
Thinking
Tier 1
GPT-6.1 Sol is an upgrade to GPT-6 Sol from OpenAI, positioned below the flagship GPT-6 Astra in the GPT-6 series. It is suited for agentic coding, computer use, document-heavy professional...
Parameters Unknown
Context 1,050,000 tokens
Released 2026-09-29
API Type openrouter
This model is available for answering and ranking.
openai gpt
Learn more – Model Terms of Service
gpt-oss-120b
Thinking
Tier Free
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
Parameters Unknown
Context 131,072 tokens
Released 2025-08-05
API Type openrouter
This model is available for answering and ranking.
openai gpt
Learn more – Model Terms of Service
gpt-oss-20b
Thinking
Tier Free
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
Parameters Unknown
Context 131,072 tokens
Released 2025-08-05
API Type openrouter
This model is available for answering and ranking.
openai gpt
Learn more – Model Terms of Service

Zai Org

GLM 4.5 Air
Thinking
Tier Free
GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...
Parameters Unknown
Context 131,072 tokens
Released 2025-07-25
API Type openrouter
This model is available for answering and ranking.
zai glm
Learn more – Model Terms of Service
GLM 5 Turbo
Thinking
Tier 1
GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...
Parameters Unknown
Context 202,752 tokens
Released 2026-03-15
API Type openrouter
This model is available for answering and ranking.
zai glm
Learn more – Model Terms of Service
GLM 5.3
Thinking
Tier 1
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
Parameters Unknown
Context 1,048,576 tokens
Released 2026-08-18
API Type openrouter
This model is available for answering and ranking.
zai glm
Learn more – Model Terms of Service
GLM 5.3 Flash
Thinking
Tier Free
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Parameters Unknown
Context 1,048,575 tokens
Released 2026-08-26
API Type openrouter
This model is available for answering and ranking.
zai glm
Learn more – Model Terms of Service
GLM 5.3 FlashX
Thinking
Tier 1
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
Parameters Unknown
Context 1,048,576 tokens
Released 2026-09-18
API Type openrouter
This model is available for answering and ranking.
zai glm
Learn more – Model Terms of Service
GLM 5.3 Prime
Thinking
Tier 1
GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...
Parameters Unknown
Context 1,000,000 tokens
Released 2026-09-23
API Type openrouter
This model is available for answering and ranking.
zai glm
Learn more – Model Terms of Service

xAI

Grok 4.7
Thinking
Tier 1
Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...
Parameters Unknown
Context 500,000 tokens
Released 2026-09-21
API Type openrouter
This model is available for answering and ranking.
xai grok
Learn more – Model Terms of Service