Advanced search

Type: Model

18 316 entries

Pareto Code Router

openrouter.ai

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) coding percentiles. Set min_coding_score between 0 and 1 on the [pareto-router plugin](https://openrouter.ai/docs/guides/routing/routers/pareto-router#the-min_coding_score-parameter) to control how...

Model Models & platforms Self-hosted cloud paid

This model always redirects to the latest model in the Claude Opus family.

Model Models & platforms Self-hosted cloud paid

OpenAI: GPT-5.4 Image 2

openrouter.ai

[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image generation capabilities from GPT Image 2. It enables rich multimodal workflows, allowing users to seamlessly move between reasoning, coding, and...

Model Models & platforms Self-hosted cloud paid

Xiaomi: MiMo-V2.5

openrouter.ai

MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

Model Models & platforms Self-hosted cloud paid

Xiaomi: MiMo-V2.5-Pro

openrouter.ai

MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....

Model Models & platforms Self-hosted cloud paid

Tencent: Hy3 preview

openrouter.ai

Hy3 preview is a high-efficiency Mixture-of-Experts model from Tencent designed for agentic workflows and production use. It supports configurable reasoning levels across disabled, low, and high modes, allowing it to...

Model Models & platforms Self-hosted cloud paid

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

Model Models & platforms Self-hosted cloud paid

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

Model Models & platforms Self-hosted cloud paid

OpenAI: GPT-5.5

openrouter.ai

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

Model Models & platforms Self-hosted cloud paid

OpenAI: GPT-5.5 Pro

openrouter.ai

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

Model Models & platforms Self-hosted cloud paid

Qwen: Qwen3.6 27B

openrouter.ai

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

Model Models & platforms Self-hosted cloud paid

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...

Model Models & platforms Self-hosted cloud paid

Qwen: Qwen3.6 35B A3B

openrouter.ai

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

Model Models & platforms Self-hosted cloud paid

Qwen: Qwen3.6 Flash

openrouter.ai

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

Model Models & platforms Self-hosted cloud paid

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

Model Models & platforms Self-hosted cloud paid

OpenAI GPT Latest

openrouter.ai

This model always redirects to the latest model in the OpenAI GPT family.

Model Models & platforms Self-hosted cloud paid

This model always redirects to the latest model in the Anthropic Claude Sonnet family.

Model Models & platforms Self-hosted cloud paid

This model always redirects to the latest model in the Google Gemini Flash family.

Model Models & platforms Self-hosted cloud paid

MoonshotAI Kimi Latest

openrouter.ai

This model always redirects to the latest model in the MoonshotAI Kimi family.

Model Models & platforms Self-hosted cloud paid

Google Gemini Pro Latest

openrouter.ai

This model always redirects to the latest model in the Google Gemini Pro family.

Model Models & platforms Self-hosted cloud paid

OpenAI GPT Mini Latest

openrouter.ai

This model always redirects to the latest model in the OpenAI GPT Mini family.

Model Models & platforms Self-hosted cloud paid

This model always redirects to the latest model in the Anthropic Claude Haiku family.

Model Models & platforms Self-hosted cloud paid

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

Model Models & platforms Self-hosted cloud free

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

Model Models & platforms Self-hosted cloud paid

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

Model Models & platforms Self-hosted cloud paid

IBM: Granite 4.1 8B

openrouter.ai

Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...

Model Models & platforms Self-hosted cloud paid

SpaceXAI: Grok 4.3

openrouter.ai

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

Model Models & platforms Self-hosted cloud paid

OpenAI: GPT Chat Latest

openrouter.ai

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

Model Models & platforms Self-hosted cloud paid

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

Model Models & platforms Self-hosted cloud paid

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

Model Models & platforms Self-hosted cloud paid

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...

Model Models & platforms Self-hosted cloud paid

Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

Model Models & platforms Self-hosted cloud paid

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

Model Models & platforms Self-hosted cloud paid

Google: Gemini 3.5 Flash

openrouter.ai

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

Model Models & platforms Self-hosted cloud paid

SpaceXAI: Grok Build 0.1

openrouter.ai

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

Model Models & platforms Self-hosted cloud paid

Qwen: Qwen3.7 Max

openrouter.ai

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

Model Models & platforms Self-hosted cloud paid

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

Model Models & platforms Self-hosted cloud paid

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

Model Models & platforms Self-hosted cloud paid

Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

Model Models & platforms Self-hosted cloud paid

StepFun: Step 3.7 Flash

openrouter.ai

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

Model Models & platforms Self-hosted cloud paid

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

Model Models & platforms Self-hosted cloud free

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

Model Models & platforms Self-hosted cloud paid

MiniMax: MiniMax M3

openrouter.ai

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

Model Models & platforms Self-hosted cloud paid

Qwen: Qwen3.7 Plus

openrouter.ai

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

Model Models & platforms Self-hosted cloud paid

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

Model Models & platforms Self-hosted cloud free

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

Model Models & platforms Self-hosted cloud paid

NVIDIA: Nemotron 3 Ultra

openrouter.ai

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

Model Models & platforms Self-hosted cloud paid

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

Model Models & platforms Self-hosted cloud free

Submit a site to the catalog

Just send the link — we will work out the rest.

We will review what you send and add it to the catalog if it fits.

Not sure how to implement it? We can help

Tell us about your task — we will pick the tools and suggest where to start.

0 / 5000
Verification code

Fields marked with an asterisk are required. Your data is used only to reply.