101 production models from 14 publishers behind a single OpenAI-compatible endpoint — api.tunerouter.com. Same key, same SDK, every modality.
DeepSeek-V4-Flash-Vision-Exp is an experimental multimodal model for visual understanding. Its text-only capabilities, including agent tasks, reasonin…
DeepSeek-V4.1-Flash is the smallest member of DeepSeek’s new model architecture family and natively supports multimodal visual understanding. The arch…

MiniMax H3 Max Turbo supports text-to-video and image-to-video at 480P/768P, with native audio.

OpenAI GPT-6 Astra for complex reasoning, coding, computer use, research, and document creation.

MiniMax H3 Fast supports text-to-video, image-to-video, and reference-to-video at 480P, with native audio.

Google Gemini 3.8 Flash, configured for reasoning, multimodal, tool, structured-output, and cache workloads.
Anthropic Claude Fable 5.1 for demanding reasoning, long-horizon agentic coding, multistep research, and complex knowledge work.

MiniMax H3 Max supports text-to-video, image-to-video, and reference-to-video at 480P/768P, with native audio.

Faster Berry 1.0 with super-resolution output at 1080p, 2K, and 4K, based on Wan 3.0 Prime.

Faster CarrotHub Berry 1.0 video generation, based on Wan 3.0 Prime.

Berry 1.0 with super-resolution output at 1080p, 2K, and 4K, based on Wan 3.0 Standard.

CarrotHub Berry 1.0 all-in-one video generation, based on Wan 3.0 Standard.

Alibaba Wan 3.0 Video Prime provides the same unified text, frame, and multimodal reference workflows as Wan 3.0 with significantly faster end-to-end …

Tencent Hunyuan Hy4 Preview is an open-source hybrid-reasoning MoE model for coding, agentic tasks, tool use, and long-context workloads.

Qwen3.8 Flash is Qwen's cost-efficient multimodal reasoning model with a 1M-token context window, 131K maximum output, tool calling, structured output…

Z.AI GLM-5.3-Flash is a cost-efficient native multimodal model for coding, agents, and professional workflows, with a 1M-token context window.

Alibaba Wan 3.0 unified video generation model for text, first/last-frame, and reference image, video, and audio workflows, with generated audio.

Z.AI GLM-5.3 flagship reasoning model for complex software engineering and long-horizon agent tasks, with a 1M-token context window.
DeepSeek V4 Pro official GA release (0813). Supports thinking mode; up to 1M context tokens.

Google Veo 3.1 Lite cost-efficient video and audio generation on Agent Platform.

Google Veo 3.1 Fast video generation on Agent Platform.

Google Veo 3.1 high-fidelity video generation on Agent Platform.

Google Gemini 3.7 Flash, verified on Google Agent Platform and configured for agentic, reasoning, multimodal, tool, structured-output, and cache workl…

xAI Grok 4.6 is a frontier multimodal reasoning model focused on long-running agents, coding, knowledge work, and interactive visual work.

Google Gemini 2.5 Pro model

Google's GA, speed-optimized frontier model for agentic coding, multimodal reasoning, tool use, and structured output.

xAI Grok 4.5 is a frontier multimodal reasoning model for coding, agentic tasks, and knowledge work, served through AtlasCloud with a 500K context win…
ByteDance's next-generation multimodal video generation, editing, and extension model with long-form, multi-asset, and multilingual support.

MiniMax M3

MiniMax reasoning model with a 200K context window and tool use.

MiniMax reasoning model with a 200K context window and tool use.

MiniMax H3 supports text-to-video, image-to-video, and reference-to-video at 480P/768P/2K, with native audio.

Kimi K2.7 Code long-context coding and agentic reasoning model.

Kimi K3 is an open-weight, native multimodal model developed by Moonshot AI. It has 2.8 trillion total parameters and is built on Kimi Delta Attention…

Tencent Hy3
DeepSeek V4 Flash official release (0731 build).

Qwen3.8 Max is Qwen's flagship multimodal reasoning model with a 1M-token context window, 131K maximum output, tool calling, structured output, and op…
Claude Fable 5
Claude Opus 5
Seedance 2.0 supports multimodal input (images, videos, audios, texts) with capabilities including video generation, video editing, and video extensio…

Flagship model for complex professional work. GPT-5.6 Sol is the top tier of the GPT-5.6 family—strongest at complex reasoning, coding, and agentic wo…
Seedream 5.0 Pro image generation and editing through BytePlus Ark.
Dreamina Seedance 2.0 Mini cost-effective multimodal video generation model
Seedance 2.0 supports multimodal input (images, videos, audios, texts) with capabilities including video generation, video editing, and video extensio…
Claude Sonnet 5

GPT-5.6 model that balances intelligence and cost. Terra sits between flagship Sol and cost-efficient Luna—suited for everyday coding, reasoning, and …

Wan 2.7 image generation and editing pro model
ByteDance image generation model with built-in reasoning, web search, and multi-image editing up to 3K resolution

GPT-5.6 model optimized for cost-sensitive workloads. Luna is the fastest, most affordable GPT-5.6 tier—built for high-volume, latency-sensitive chat,…

Moonshot AI general-purpose multimodal model with a 262K context window, long-horizon coding, self-correction, reasoning, and tool use.

A new class of intelligence for coding and professional work. GPT-5.5 is OpenAI's flagship model for the most complex professional tasks, with stronge…

GLM-5.2 is Zhipu AI's (brand Z.AI) coding-focused flagship, launched June 13, 2026, a 744B-parameter mixture-of-experts model with about 40B active pa…
Gemini 3.5 Flash is Google DeepMind's mid-2026 Flash-tier reasoning model, shipped May 19, 2026 at Google I/O on the Gemini 3 Flash foundation. It is …
Gemini 3.1 Pro Preview is the preview release of Google DeepMind's flagship Gemini 3.1 Pro, launched February 19, 2026 as a Transformer-based Mixture-…
Anthropic's flagship model for long-horizon agentic coding.

Qwen3.7-Max is Alibaba's proprietary flagship large language model, launched May 2026 and built for long-horizon agentic work, coding and complex reas…

DeepSeek V4 Pro is DeepSeek's flagship Mixture-of-Experts model, released open-source under the MIT license on April 24, 2026, with 1.6 trillion total…
ByteDance Seedream 4.5 image generation model with text-to-image, image-to-image, and sequential image generation support
ByteDance Seedream 4.0 image generation model with text-to-image, image-to-image, and sequential image generation support
ByteDance Dola Seed 2.1 Pro multimodal reasoning model for complex agentic, coding, and visual understanding tasks

xAI Grok 4.3, OpenAI-compatible chat model with text and image input.
Claude Opus 4.7 is an Anthropic frontier model released April 16, 2026, positioned as a notable step up from Opus 4.6 in advanced software engineering…
Claude Opus 4.6 is an Anthropic flagship model released February 5, 2026, building on Opus 4.5 with higher reliability and precision for coding, agent…

Z.AI CogVideoX-3 video generation model
Claude Opus 4.5 is an Anthropic frontier model launched November 24, 2025, built for high-intelligence coding, agents and complex reasoning. It suppor…
Claude Sonnet 4.6 is Anthropic's most capable Sonnet model, released February 17, 2026, with upgrades across coding, computer use, long-context reason…

State-of-the-art image generation model. GPT Image 2 is OpenAI's current image model for fast, high-quality generation and editing, with flexible size…
Claude Sonnet 4.5 is an Anthropic model released September 2025, positioned at launch as among the best in the world for real-world agents, coding and…

DeepSeek V4 Flash is the efficiency-optimized member of DeepSeek's V4 family, released open-source under the MIT license on April 24, 2026. Per its mo…

Shengshu Technology text-to-video model with cinematic quality and 1080p rendering
Claude Haiku 4.5 is Anthropic's small, fast model released in 2025, delivering strong coding, tool use and reasoning at one-third the cost and more th…

Happy Horse 1.1

Our previous image generation model. GPT Image 1.5 offers strong instruction following and prompt adherence for image generation and editing.

GLM-5.1 is Z.AI's flagship reasoning and agentic-engineering model released April 7, 2026, a 754B-parameter hybrid MoE (GlmMoeDSA) combining linear an…

GLM-5 is Zhipu AI's (Z.AI) flagship foundation model launched February 11, 2026, a 744B-parameter mixture-of-experts model with about 40B active param…

Zhipu AI high-quality image generation model

GLM-ASR-2512 is Zhipu AI's (Z.AI) next-generation cloud speech recognition model for real-time conversion of speech into high-quality text. It reports…

GLM-5V-Turbo is Z.AI's first native multimodal agent foundation model, released April 1, 2026, built on the GLM-5 base and handling image, video and t…

Moonshot AI multimodal reasoning model with a 262K context window, tool use, structured output, and text, image, and video input.
BytePlus Seed Audio 1.0 non-streaming audio generation model.
Doubao-Seed-2.0 Pro is the top-tier agentic reasoning model in ByteDance's Doubao/Seed 2.0 LLM series, launched February 14, 2026 (the Dola prefix is …
Gemini 3.1 Flash-Lite is Google DeepMind's lowest-cost, lowest-latency model in the Gemini 3.1 series, generally available in 2026 and priced around $…

Gemini 3 Flash is Google DeepMind's Gemini 3 series Flash-tier model, rolled out around mid-December 2025, with the preview being the early developer-…

Gemini 2.5 Flash-Lite is Google DeepMind's fastest and lowest-cost model in the stable Gemini 2.5 family, priced at $0.10/1M input and $0.40/1M output…

Gemini 2.5 Flash is Google DeepMind's price-performance workhorse model, built for speed and low cost while handling text, audio, image and video inpu…

Nano Banana Pro

Nano Banana 2
ByteDance video model with native audio generation including dialogue, sound effects, and multi-language support

Qwen3.6-Flash is the speed-optimized, cost-efficient tier of Alibaba's Qwen3.6 model family, reported as released in late April 2026 for high-throughp…

Shengshu Technology image-to-video model that animates a single input image

A more affordable model for coding and professional work. GPT-5.4 is a previous flagship for complex professional tasks, with a 1.05M context window, …

Strong intelligence for cost-sensitive, low-latency, high-volume workloads. GPT-5 Mini is a faster, more cost-efficient version of GPT-5, best for wel…

Nano Banana image generation model

Qwen3.5-Flash is the cost-optimized, lower-latency tier of Alibaba's Qwen3.5 series, reported as built on a 35B-A3B architecture for efficient inferen…

Qwen3.5-Plus is the higher-capability tier in Alibaba's Qwen3.5 series, a natively multimodal LLM line built around an efficient hybrid architecture f…

Qwen3.6-Plus is the balanced mid-tier model in Alibaba's Qwen3.6 family, released around April 2026 and positioned for real-world enterprise agent wor…

Wan 2.6 image generation and editing model

Wan 2.7 image generation and editing model
Anthropic Claude Opus 4.1 snapshot released on 2025-08-05, for advanced reasoning, coding, and complex agentic work.
Anthropic Claude Sonnet 4 snapshot released on 2025-05-14, for capable and efficient reasoning, coding, and agentic work.
Anthropic Claude Opus 4 snapshot released on 2025-05-14, for advanced reasoning, coding, and complex agentic work.
Pricing in USD per 1M tokens unless noted · catalog mirrored from the live API · JSON list