Skip to content
Not available in this workspace
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube

Models

Models

CompareDiscover Models
Favicon for anthropic
Favicon for openai
CompareDiscover Models
Favicon for anthropic
Favicon for openai
  • Favicon for heygen
    HeyGen: Avatar IVAvatar IV
    <1 hour

    HeyGen: Avatar IV is an image-to-video model that animates a single photo into an expressive, lip-synced talking-head video. Rather than only matching mouth shapes to words, it interprets the vocal tone, rhythm, and emotion of the audio to drive head motion, facial expression, and gestures, producing output at up to 1080p. The spoken audio comes from one of two inputs: a text script, which the model voices with HeyGen text-to-speech, or a supplied audio track, which the image is lip-synced to directly. Passthrough parameters let you choose a voice, tune voice settings, set expressiveness, prompt specific motion, replace or remove the background, add captions, and title the video.

    by heygenAug 24, 2026$0.05/second
  • Favicon for meta
    Meta: Muse Spark 1.2 ContributorMuse Spark 1.2 Contributor
    18+
    84.9B tokens

    Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark 1.2. Your prompts and outputs may be used to improve Meta’s products, making it ideal for experimentation, learning, and early-stage projects without worry about spend. It is a reasoning model from Meta, tailored for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context window. The model is built to support multi-agent workflows, whether as either a main agent that plans and delegates or as a subagent executing in parallel. It works across multiple coding harnesses and supports structured output, parallel function calling, and configurable reasoning effort. In Meta’s testing, it performs well on multi-file refactors, extended debugging sessions, whole-repository generation, and tasks that stretch well past a single prompt.

    by metaAug 21, 20261.05M context$0.10/M input tokens$0.20/M output tokens
  • Favicon for deepseek
    DeepSeek: DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp
    76.2B tokens

    DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of DeepSeek V4 Flash 0731 from DeepSeek, adding image understanding while matching the base model on text capabilities including agents, reasoning, and world knowledge. It is a sparse mixture-of-experts model with 13B active parameters out of 284B total. It is suited for document and chart understanding, visual question answering, and multimodal agent workflows that interleave text and images.

    by deepseekAug 21, 20261.05M context$0.22/M input tokens$0.66/M output tokens
  • Favicon for stealth
    Ox AlphaOx Alpha
    16.5T tokens

    Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with visual context. Ox Alpha is a stealth model. It is developed and operated by a third-party provider who has chosen to remain anonymous during this preview. OpenRouter routes requests to it and is not its developer, owner, or provider. Prompts and completions are retained by the provider and are not used for training; all other use is governed by the Stealth Model Terms.

    by stealthAug 20, 20261.05M context$0/M input tokens$0/M output tokens
  • Favicon for tencent
    Tencent: Hy-MT2-1.8BHy-MT2-1.8B
    31.1M tokens

    Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.

    by tencentAug 20, 20268K context$0.044/M input tokens$0.177/M output tokens
  • Favicon for tencent
    Tencent: Hy-MT2-30B-A3BHy-MT2-30B-A3B
    45.5M tokens

    Hy-MT2-30B-A3B is Tencent's flagship translation model in the Hy-MT2 family. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation. It uses 3B active parameters out of 30B total.

    by tencentAug 20, 20268K context$0.074/M input tokens$0.295/M output tokens
  • Favicon for black-forest-labs
    Black Forest Labs: FLUX Video UpscaleFLUX Video Upscale

    FLUX Video Upscale is a video upscaling model from Black Forest Labs. It enlarges a single source video by 1.5× to 3× while preserving its duration, with an optional prompt and precise or creative processing modes.

    by black-forest-labsAug 19, 2026from $0.075/megapixel-second
  • Favicon for z-ai
    Z.ai: GLM LatestGLM Latest

    This model always redirects to the latest GLM model from Z.ai.

    by z-aiAug 19, 20261.05M context$1.40/M input tokens$4.40/M output tokens
  • Favicon for tencent
    Tencent: Hy-MT2-7BHy-MT2-7B
    5.6M tokens

    Hy-MT2-7B is a 7B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.

    by tencentAug 19, 20268K context$0.074/M input tokens$0.295/M output tokens
  • Favicon for z-ai
    Z.ai: GLM 5.3GLM 5.3
    597B tokens

    GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves on GLM-5.2 in coding and in the balance between performance and token efficiency. Reasoning is always on and cannot be disabled. Reasoning efforts low, high, and max are supported; max is the default.

    by z-aiAug 18, 20261.05M context$1.40/M input tokens$4.40/M output tokens
  • Favicon for liquid
    LiquidAI: LFM2.5-Embedding-350M (free)LFM2.5-Embedding-350M (free)Free variant
    350M tokens

    LFM2.5-Embedding-350M is a text embedding model from Liquid AI. It produces 1,024-dimensional embeddings for retrieval and semantic search. Successful OpenRouter requests and embeddings may be retained and used to train Liquid models.

    by liquidAug 18, 2026512 context$0/M input tokens$0/M output tokens
  • Favicon for qwen
    Qwen: Qwen3.8 27BQwen3.8 27B
    152B tokens

    Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be enabled or disabled.

    by qwenAug 14, 20261M context$0.35/M input tokens$2.75/M output tokens
  • Favicon for dots-studio
    Dots Studio: Dots3-Note Preview (free)Dots3-Note Preview (free)Free variant
    Going away September 30, 2026
    123B tokens

    Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is suited for reasoning, coding, multimodal understanding, long-context processing, and multi-step agent workflows.

    by dots-studioAug 14, 2026512K context$0/M input tokens$0/M output tokens
  • Favicon for nvidia
    NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6BNemotron 3.5 ASR Streaming Multilingual 0.6B
    6.86M characters

    Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice agents, and multilingual transcription pipelines.

    by nvidiaAug 13, 2026$0.000003/second
  • Favicon for mistralai
    Mistral: Voxtral Small 24B 2507 STTVoxtral Small 24B 2507 STT
    2.08M characters

    Voxtral Small 24B 2507 STT is a speech transcription model from Mistral AI. It is suited for transcription, translation, and audio understanding workloads that benefit from its larger model capacity.

    by mistralaiAug 13, 2026$0.00005/second
  • Favicon for mistralai
    Mistral: Voxtral Mini 3B 2507Voxtral Mini 3B 2507
    3.54M characters

    Voxtral Mini 3B 2507 is a speech and audio understanding model from Mistral AI. It is suited for transcription, translation, and compact audio processing workloads.

    by mistralaiAug 13, 2026$0.000017/second
  • Favicon for bytedance-seed
    ByteDance Seed: Seedream 5.0 LiteSeedream 5.0 Lite
    335M tokens

    Seedream 5.0 Lite is an image generation model from ByteDance Seed. It is suited for professional visual creation that benefits from web-connected retrieval, complex-prompt comprehension, visual references, and broad knowledge coverage.

    by bytedance-seedAug 13, 2026$0.035/image
  • Favicon for google
    Google: Gemini 3.7 Flash (batch)Gemini 3.7 Flash (batch)
    75% off
    Batch variant
    8.6B tokens

    Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step problem solving.

    by googleAug 13, 20261.05M context$0.1875/M input tokens$0.9375/M output tokens
  • Favicon for google
    Google: Gemini 3.7 FlashGemini 3.7 Flash
    75% off
    2.09T tokens

    Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step problem solving.

    by googleAug 13, 20261.05M context$0.375/M input tokens$1.875/M output tokens
  • Favicon for voyageai
    VoyageAI by MongoDB: voyage-code-4voyage-code-4
    160M tokens

    voyage-code-4 is a code embedding model from Voyage AI, a MongoDB company. It is designed for coding agents and code retrieval, with Matryoshka embeddings at 2048, 1024, 512, and 256 dimensions and multiple quantization options. Learn more about voyage-code-4 here: blog.voyageai.com/2026/08/13/voyage-code-4

    by voyageaiAug 13, 202632K context$0.12/M tokens