VastLLM Model Catalog

101 production models from 14 publishers behind a single OpenAI-compatible endpoint — api.tunerouter.com. Same key, same SDK, every modality.

DeepSeek V4 Flash Vision ExpDeepSeek

DeepSeek-V4-Flash-Vision-Exp is an experimental multimodal model for visual understanding. Its text-only capabilities, including agent tasks, reasonin…

$0.19 in · $0.78 out / 1Mtext, image → text
DeepSeek V4.1 FlashDeepSeek

DeepSeek-V4.1-Flash is the smallest member of DeepSeek’s new model architecture family and natively supports multimodal visual understanding. The arch…

$0.19 in · $0.78 out / 1Mtext, image → text
MiniMax H3 Max TurboMiniMax

MiniMax H3 Max Turbo supports text-to-video and image-to-video at 480P/768P, with native audio.

$0.0518 / calltext, image → video
GPT-6 AstraOpenAI

OpenAI GPT-6 Astra for complex reasoning, coding, computer use, research, and document creation.

$12.95 in · $64.75 out / 1Mtext, images → text
MiniMax H3 FastMiniMax

MiniMax H3 Fast supports text-to-video, image-to-video, and reference-to-video at 480P, with native audio.

$0.0648 / calltext, image, video, audio → video
Gemini 3.8 FlashGoogle DeepMind

Google Gemini 3.8 Flash, configured for reasoning, multimodal, tool, structured-output, and cache workloads.

$0.97 in · $4.86 out / 1Mtext, image, video, audio → text
Claude Fable 5.1Anthropic

Anthropic Claude Fable 5.1 for demanding reasoning, long-horizon agentic coding, multistep research, and complex knowledge work.

$12.95 in · $64.75 out / 1Mtext, images → text
MiniMax H3 MaxMiniMax

MiniMax H3 Max supports text-to-video, image-to-video, and reference-to-video at 480P/768P, with native audio.

$0.1036 / calltext, image, video, audio → video
Berry 1.0 Pro TurboCarrotHub

Faster Berry 1.0 with super-resolution output at 1080p, 2K, and 4K, based on Wan 3.0 Prime.

$0.3367 / calltext, image, video, audio → video, audio
Berry 1.0 TurboCarrotHub

Faster CarrotHub Berry 1.0 video generation, based on Wan 3.0 Prime.

$0.3626 / calltext, image, video, audio → video, audio
Berry 1.0 ProCarrotHub

Berry 1.0 with super-resolution output at 1080p, 2K, and 4K, based on Wan 3.0 Standard.

$0.2331 / calltext, image, video, audio → video, audio
Berry 1.0CarrotHub

CarrotHub Berry 1.0 all-in-one video generation, based on Wan 3.0 Standard.

$0.259 / calltext, image, video, audio → video, audio
Wan 3.0 Video PrimeAlibaba Cloud

Alibaba Wan 3.0 Video Prime provides the same unified text, frame, and multimodal reference workflows as Wan 3.0 with significantly faster end-to-end …

$0.3626 / calltext, image, video, audio → video
Hy4 PreviewTencent Hunyuan

Tencent Hunyuan Hy4 Preview is an open-source hybrid-reasoning MoE model for coding, agentic tasks, tool use, and long-context workloads.

$1.08 in · $3.24 out / 1Mtext → text
Qwen3.8 FlashAlibaba(Qwen)

Qwen3.8 Flash is Qwen's cost-efficient multimodal reasoning model with a 1M-token context window, 131K maximum output, tool calling, structured output…

$0.19 in · $0.61 out / 1Mtext, image, video → text
GLM-5.3-FlashZ.AI

Z.AI GLM-5.3-Flash is a cost-efficient native multimodal model for coding, agents, and professional workflows, with a 1M-token context window.

$0.19 in · $0.65 out / 1Mtext, image → text
Wan 3.0 VideoAlibaba Cloud

Alibaba Wan 3.0 unified video generation model for text, first/last-frame, and reference image, video, and audio workflows, with generated audio.

$0.14 / calltext, image, video, audio → video, audio
GLM-5.3Z.AI

Z.AI GLM-5.3 flagship reasoning model for complex software engineering and long-horizon agent tasks, with a 1M-token context window.

$1.81 in · $5.7 out / 1Mtext → text
DeepSeek V4 Pro 0813DeepSeek

DeepSeek V4 Pro official GA release (0813). Supports thinking mode; up to 1M context tokens.

$0.85 in · $2.56 out / 1Mtext → text
Veo 3.1 LiteGoogle DeepMind

Google Veo 3.1 Lite cost-efficient video and audio generation on Agent Platform.

$0.0389 / calltext, image → video
Veo 3.1 FastGoogle DeepMind

Google Veo 3.1 Fast video generation on Agent Platform.

$0.1036 / calltext, image → video
Veo 3.1Google DeepMind

Google Veo 3.1 high-fidelity video generation on Agent Platform.

$0.259 / calltext, image → video
Gemini 3.7 FlashGoogle DeepMind

Google Gemini 3.7 Flash, verified on Google Agent Platform and configured for agentic, reasoning, multimodal, tool, structured-output, and cache workl…

$0.97 in · $4.86 out / 1Mtext, image, video, audio → text
Grok 4.6xAI

xAI Grok 4.6 is a frontier multimodal reasoning model focused on long-running agents, coding, knowledge work, and interactive visual work.

$5.18 in · $15.54 out / 1Mtext, image → text
Gemini 2.5 ProGoogle DeepMind

Google Gemini 2.5 Pro model

$1.62 in · $12.95 out / 1Mtext → text
Gemini 3.6 FlashGoogle DeepMind

Google's GA, speed-optimized frontier model for agentic coding, multimodal reasoning, tool use, and structured output.

$1.94 in · $9.71 out / 1Mtext, image, video, audio → text
Grok 4.5xAI

xAI Grok 4.5 is a frontier multimodal reasoning model for coding, agentic tasks, and knowledge work, served through AtlasCloud with a 500K context win…

$5.18 in · $15.54 out / 1Mtext, image → text
Seedance 2.5ByteDance

ByteDance's next-generation multimodal video generation, editing, and extension model with long-form, multi-asset, and multilingual support.

$0 in · $13.86 out / 1Mtext, image, video, audio → video
MiniMax M3MiniMax

MiniMax M3

$0.39 in · $1.55 out / 1Mtext → text
MiniMax M2.7MiniMax

MiniMax reasoning model with a 200K context window and tool use.

$0.39 in · $1.55 out / 1Mtext → text
MiniMax M2.5MiniMax

MiniMax reasoning model with a 200K context window and tool use.

$0.39 in · $1.55 out / 1Mtext → text
MiniMax H3MiniMax

MiniMax H3 supports text-to-video, image-to-video, and reference-to-video at 480P/768P/2K, with native audio.

$0.1683 / calltext, image, video, audio → video
Kimi K2.7 CodeMoonshot AI

Kimi K2.7 Code long-context coding and agentic reasoning model.

$1.23 in · $5.18 out / 1Mtext → text
Kimi K3Moonshot AI

Kimi K3 is an open-weight, native multimodal model developed by Moonshot AI. It has 2.8 trillion total parameters and is built on Kimi Delta Attention…

$3.88 in · $19.43 out / 1Mtext, image, video → text
Tencent Hy3Tencent Hunyuan

Tencent Hy3

$0.17 in · $0.68 out / 1Mtext → text
DeepSeek V4 Flash 0731DeepSeek

DeepSeek V4 Flash official release (0731 build).

$0.28 in · $0.85 out / 1Mtext → text
Qwen3.8 MaxAlibaba(Qwen)

Qwen3.8 Max is Qwen's flagship multimodal reasoning model with a 1M-token context window, 131K maximum output, tool calling, structured output, and op…

$2.59 in · $7.77 out / 1Mtext, image, video → text
Claude Fable 5Anthropic

Claude Fable 5

$12.95 in · $64.75 out / 1Mtext, image → text
Claude Opus 5Anthropic

Claude Opus 5

$6.47 in · $32.38 out / 1Mtext, image → text
Seedance 2.0ByteDance

Seedance 2.0 supports multimodal input (images, videos, audios, texts) with capabilities including video generation, video editing, and video extensio…

$0 in · $9.97 out / 1Mtext, image, video, audio → video
GPT-5.6 SolOpenAI

Flagship model for complex professional work. GPT-5.6 Sol is the top tier of the GPT-5.6 family—strongest at complex reasoning, coding, and agentic wo…

$5.18 in · $25.9 out / 1Mtext, images → text
Seedream 5.0 ProByteDance

Seedream 5.0 Pro image generation and editing through BytePlus Ark.

$0.0583 / calltext, image → image
Seedance 2.0 MiniByteDance

Dreamina Seedance 2.0 Mini cost-effective multimodal video generation model

$0 in · $4.53 out / 1Mtext, image, video, audio → video
Seedance 2.0 FastByteDance

Seedance 2.0 supports multimodal input (images, videos, audios, texts) with capabilities including video generation, video editing, and video extensio…

$0 in · $7.25 out / 1Mtext, image, video, audio → video
Claude Sonnet 5Anthropic

Claude Sonnet 5

$2.59 in · $12.95 out / 1Mtext, image → text
GPT-5.6 TerraOpenAI

GPT-5.6 model that balances intelligence and cost. Terra sits between flagship Sol and cost-efficient Luna—suited for everyday coding, reasoning, and …

$2.59 in · $15.54 out / 1Mtext, images → text
Wan 2.7 Image ProAlibaba Cloud

Wan 2.7 image generation and editing pro model

$0.0971 / calltext, image → image
Seedream 5.0 LiteByteDance

ByteDance image generation model with built-in reasoning, web search, and multi-image editing up to 3K resolution

$0.0453 / calltext, image → image
GPT-5.6 LunaOpenAI

GPT-5.6 model optimized for cost-sensitive workloads. Luna is the fastest, most affordable GPT-5.6 tier—built for high-volume, latency-sensitive chat,…

$0.26 in · $1.55 out / 1Mtext, images → text
Kimi K2.6Moonshot AI

Moonshot AI general-purpose multimodal model with a 262K context window, long-horizon coding, self-correction, reasoning, and tool use.

$1.23 in · $5.18 out / 1Mtext, image, video → text
GPT 5.5OpenAI

A new class of intelligence for coding and professional work. GPT-5.5 is OpenAI's flagship model for the most complex professional tasks, with stronge…

$6.47 in · $38.85 out / 1Mtext → text
GLM 5.2Z.AI

GLM-5.2 is Zhipu AI's (brand Z.AI) coding-focused flagship, launched June 13, 2026, a 744B-parameter mixture-of-experts model with about 40B active pa…

$1.81 in · $5.7 out / 1Mtext → text
Gemini 3.5 FlashGoogle DeepMind

Gemini 3.5 Flash is Google DeepMind's mid-2026 Flash-tier reasoning model, shipped May 19, 2026 at Google I/O on the Gemini 3 Flash foundation. It is …

$1.94 in · $11.65 out / 1Mtext, image → text
Gemini 3.1 Pro PreviewGoogle DeepMind

Gemini 3.1 Pro Preview is the preview release of Google DeepMind's flagship Gemini 3.1 Pro, launched February 19, 2026 as a Transformer-based Mixture-…

$2.59 in · $15.54 out / 1Mtext, image → text
Claude Opus 4.8Anthropic

Anthropic's flagship model for long-horizon agentic coding.

$6.47 in · $32.38 out / 1Mtext, image → text
Qwen3.7 MaxAlibaba(Qwen)

Qwen3.7-Max is Alibaba's proprietary flagship large language model, launched May 2026 and built for long-horizon agentic work, coding and complex reas…

$2.14 in · $6.41 out / 1Mimage, text → text
DeepSeek V4 ProDeepSeek

DeepSeek V4 Pro is DeepSeek's flagship Mixture-of-Experts model, released open-source under the MIT license on April 24, 2026, with 1.6 trillion total…

$0.85 in · $2.56 out / 1Mtext → text
Seedream 4.5ByteDance

ByteDance Seedream 4.5 image generation model with text-to-image, image-to-image, and sequential image generation support

$0.0518 / calltext, image → image
Seedream 4.0ByteDance

ByteDance Seedream 4.0 image generation model with text-to-image, image-to-image, and sequential image generation support

$0.0389 / calltext, image → image
Seed 2.1 ProByteDance

ByteDance Dola Seed 2.1 Pro multimodal reasoning model for complex agentic, coding, and visual understanding tasks

$0.65 in · $3.24 out / 1Mtext, image, video → text
Grok 4.3xAI

xAI Grok 4.3, OpenAI-compatible chat model with text and image input.

$3.24 in · $6.47 out / 1Mtext, image → text
Claude Opus 4.7Anthropic

Claude Opus 4.7 is an Anthropic frontier model released April 16, 2026, positioned as a notable step up from Opus 4.6 in advanced software engineering…

$6.47 in · $32.38 out / 1Mtext, image → text
Claude Opus 4.6Anthropic

Claude Opus 4.6 is an Anthropic flagship model released February 5, 2026, building on Opus 4.5 with higher reliability and precision for coding, agent…

$6.47 in · $32.38 out / 1MTEXT, IMAGE → TEXT
CogVideoX-3Z.AI

Z.AI CogVideoX-3 video generation model

$0 in · $0 out / 1Mtext, image → video
Claude Opus 4.5Anthropic

Claude Opus 4.5 is an Anthropic frontier model launched November 24, 2025, built for high-intelligence coding, agents and complex reasoning. It suppor…

$6.47 in · $32.38 out / 1Mtext, image → text
Claude Sonnet 4.6Anthropic

Claude Sonnet 4.6 is Anthropic's most capable Sonnet model, released February 17, 2026, with upgrades across coding, computer use, long-context reason…

$3.88 in · $19.43 out / 1Mtext, image → text
GPT Image 2OpenAI

State-of-the-art image generation model. GPT Image 2 is OpenAI's current image model for fast, high-quality generation and editing, with flexible size…

$0 in · $0 out / 1Mtext, image → image
Claude Sonnet 4.5Anthropic

Claude Sonnet 4.5 is an Anthropic model released September 2025, positioned at launch as among the best in the world for real-world agents, coding and…

$3.88 in · $19.43 out / 1Mtext, image → text
DeepSeek V4 FlashDeepSeek

DeepSeek V4 Flash is the efficiency-optimized member of DeepSeek's V4 family, released open-source under the MIT license on April 24, 2026. Per its mo…

$0.28 in · $0.85 out / 1Mtext → text
Vidu Q1 TextShengshu Technology

Shengshu Technology text-to-video model with cinematic quality and 1080p rendering

$0 in · $0 out / 1Mtext → video
Claude Haiku 4.5Anthropic

Claude Haiku 4.5 is Anthropic's small, fast model released in 2025, delivering strong coding, tool use and reasoning at one-third the cost and more th…

$1.29 in · $6.47 out / 1MTEXT, IMAGE → TEXT
Happy Horse 1.1Alibaba Cloud

Happy Horse 1.1

$0.2331 / calltext, image → video
GPT Image 1.5OpenAI

Our previous image generation model. GPT Image 1.5 offers strong instruction following and prompt adherence for image generation and editing.

$0 in · $0 out / 1Mtext, image → image
GLM 5.1Z.AI

GLM-5.1 is Z.AI's flagship reasoning and agentic-engineering model released April 7, 2026, a 754B-parameter hybrid MoE (GlmMoeDSA) combining linear an…

$1.81 in · $5.7 out / 1Mtext → text
GLM 5Z.AI

GLM-5 is Zhipu AI's (Z.AI) flagship foundation model launched February 11, 2026, a 744B-parameter mixture-of-experts model with about 40B active param…

$1.29 in · $4.14 out / 1Mtext → text
GLM-ImageZ.AI

Zhipu AI high-quality image generation model

$0.0194 / calltext → image
GLM ASRZ.AI

GLM-ASR-2512 is Zhipu AI's (Z.AI) next-generation cloud speech recognition model for real-time conversion of speech into high-quality text. It reports…

$0 in · $0.04 out / 1Maudio → text
GLM 5V TurboZ.AI

GLM-5V-Turbo is Z.AI's first native multimodal agent foundation model, released April 1, 2026, built on the GLM-5 base and handling image, video and t…

$1.55 in · $5.18 out / 1Mtext, image, video, file → text
Kimi K2.5Moonshot AI

Moonshot AI multimodal reasoning model with a 262K context window, tool use, structured output, and text, image, and video input.

$0.78 in · $3.88 out / 1Mtext, image, video → text
Seed Audio 1.0ByteDance

BytePlus Seed Audio 1.0 non-streaming audio generation model.

$0.0032 / calltext → audio
Seed 2.0 ProByteDance

Doubao-Seed-2.0 Pro is the top-tier agentic reasoning model in ByteDance's Doubao/Seed 2.0 LLM series, launched February 14, 2026 (the Dola prefix is …

$1.29 in · $7.77 out / 1Mtext → text
Gemini 3.1 Flash LiteGoogle DeepMind

Gemini 3.1 Flash-Lite is Google DeepMind's lowest-cost, lowest-latency model in the Gemini 3.1 series, generally available in 2026 and priced around $…

$0.32 in · $1.94 out / 1Mtext → text
Gemini 3 Flash PreviewGoogle DeepMind

Gemini 3 Flash is Google DeepMind's Gemini 3 series Flash-tier model, rolled out around mid-December 2025, with the preview being the early developer-…

$0.65 in · $3.88 out / 1Mtext → text
Gemini 2.5 Flash LiteGoogle DeepMind

Gemini 2.5 Flash-Lite is Google DeepMind's fastest and lowest-cost model in the stable Gemini 2.5 family, priced at $0.10/1M input and $0.40/1M output…

$0.13 in · $0.52 out / 1Mtext → text
Gemini 2.5 FlashGoogle DeepMind

Gemini 2.5 Flash is Google DeepMind's price-performance workhorse model, built for speed and low cost while handling text, audio, image and video inpu…

$0.39 in · $3.24 out / 1Mtext → text
Nano Banana ProGoogle DeepMind

Nano Banana Pro

$2.59 in · $0 out / 1Mtext, image → image
Nano Banana 2Google DeepMind

Nano Banana 2

$0.65 in · $0 out / 1Mtext, image → image
Seedance 1.5 ProByteDance

ByteDance video model with native audio generation including dialogue, sound effects, and multi-language support

$0 in · $3.11 out / 1Mimage, text → video
Qwen3.6 FlashAlibaba(Qwen)

Qwen3.6-Flash is the speed-optimized, cost-efficient tier of Alibaba's Qwen3.6 model family, reported as released in late April 2026 for high-throughp…

$0.85 in · $5.13 out / 1Mtext, image, audio → text
Vidu Q1 ImageShengshu Technology

Shengshu Technology image-to-video model that animates a single input image

$0 in · $0 out / 1Mimage, text → video
GPT 5.4OpenAI

A more affordable model for coding and professional work. GPT-5.4 is a previous flagship for complex professional tasks, with a 1.05M context window, …

$3.24 in · $19.43 out / 1Mtext, image → text
GPT 5 MiniOpenAI

Strong intelligence for cost-sensitive, low-latency, high-volume workloads. GPT-5 Mini is a faster, more cost-efficient version of GPT-5, best for wel…

$0.32 in · $2.59 out / 1Mtext, image → text
Nano BananaGoogle DeepMind

Nano Banana image generation model

$0.39 in · $0 out / 1Mtext, image → image
Qwen3.5 FlashAlibaba(Qwen)

Qwen3.5-Flash is the cost-optimized, lower-latency tier of Alibaba's Qwen3.5 series, reported as built on a 35B-A3B architecture for efficient inferen…

$0.22 in · $2.23 out / 1Mtext, image → text
Qwen3.5 PlusAlibaba(Qwen)

Qwen3.5-Plus is the higher-capability tier in Alibaba's Qwen3.5 series, a natively multimodal LLM line built around an efficient hybrid architecture f…

$0.15 in · $0.89 out / 1Mtext, image, video → text
Qwen3.6 PlusAlibaba(Qwen)

Qwen3.6-Plus is the balanced mid-tier model in Alibaba's Qwen3.6 family, released around April 2026 and positioned for real-world enterprise agent wor…

$0.36 in · $2.14 out / 1Mtext, image, video → text
Wan 2.6 ImageAlibaba Cloud

Wan 2.6 image generation and editing model

$0.0389 / calltext, image → image
Wan 2.7 ImageAlibaba Cloud

Wan 2.7 image generation and editing model

$0.0389 / calltext, image → image
Claude Opus 4.1Anthropic

Anthropic Claude Opus 4.1 snapshot released on 2025-08-05, for advanced reasoning, coding, and complex agentic work.

$19.43 in · $97.12 out / 1Mtext → text
Claude Sonnet 4Anthropic

Anthropic Claude Sonnet 4 snapshot released on 2025-05-14, for capable and efficient reasoning, coding, and agentic work.

$3.88 in · $19.43 out / 1Mtext → text
Claude Opus 4Anthropic

Anthropic Claude Opus 4 snapshot released on 2025-05-14, for advanced reasoning, coding, and complex agentic work.

$19.43 in · $97.12 out / 1Mtext → text

Pricing in USD per 1M tokens unless noted · catalog mirrored from the live API · JSON list