Skip to content
provod.ai
RU

/v1/models

Model catalog with the details needed to choose

Each row identifies a public model and its documented request capabilities.

72 models

  • Z.ai: GLM 4.5

    glm-4.5

    GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...

    Provider
    Z.ai
    Context
    131.07K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 50.73
    Output RUB / 1M
    Output RUB / 1M: 186
  • Z.ai: GLM 4.5 Air

    glm-4.5-air

    GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

    Provider
    Z.ai
    Context
    131.07K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 16.91
    Output RUB / 1M
    Output RUB / 1M: 93
  • Z.ai: GLM 4.5V

    glm-4.5v

    GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

    Provider
    Z.ai
    Context
    65.54K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 50.73
    Output RUB / 1M
    Output RUB / 1M: 152.18
  • Z.ai: GLM 4.6

    glm-4.6

    Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 50.73
    Output RUB / 1M
    Output RUB / 1M: 186
  • Z.ai: GLM 4.6V

    glm-4.6v

    GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...

    Provider
    Z.ai
    Context
    131.07K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 25.36
    Output RUB / 1M
    Output RUB / 1M: 76.09
  • Z.ai: GLM 4.7

    glm-4.7

    GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 50.73
    Output RUB / 1M
    Output RUB / 1M: 186
  • Z.ai: GLM 5

    glm-5

    GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 84.54
    Output RUB / 1M
    Output RUB / 1M: 270.54
  • Z.ai: GLM 5 Turbo

    glm-5-turbo

    GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

    Provider
    Z.ai
    Context
    202.75K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 101.45
    Output RUB / 1M
    Output RUB / 1M: 338.18
  • Z.ai: GLM 5V Turbo

    glm-5v-turbo

    GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

    Provider
    Z.ai
    Context
    202.75K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 101.45
    Output RUB / 1M
    Output RUB / 1M: 338.18
  • ByteDance: Seedance 2.5

    seedance-2.5

    Seedance 2.5 is a video generation model from ByteDance. It is suited for long-form storytelling, multimodal reference-based generation, video editing, and video extension. It supports first-frame and first-and-last-frame control, up...

    Provider
    ByteDance
    Capabilities
    Video
    Catalog updated
  • Black Forest Labs: FLUX.3 Video

    flux-3-video

    FLUX.3 Video is a video generation model from Black Forest Labs. It supports text-to-video, image-guided generation with opening and closing keyframes, and video continuation workflows, making it suited for controlled...

    Provider
    Black Forest Labs
    Capabilities
    Video
    Catalog updated
  • Claude Opus 5

    claude-opus-5

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,113.62
  • DeepSeek: DeepSeek V4 Flash 0731

    deepseek-v4-flash-0731

    DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.

    Provider
    DeepSeek
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 11.84
    Output RUB / 1M
    Output RUB / 1M: 23.67
  • MiniMax: H3

    hailuo-3

    MiniMax H3 is a lightweight, open-weights video generation model from MiniMax. It is designed for precise multimodal editing and controlled content generation, including instruction-guided edits, text and brand rendering, and...

    Provider
    MiniMax
    Capabilities
    Video
    Catalog updated
  • Runway: Aleph 2.0

    aleph-2

    Runway Aleph 2.0 is an in-context video editing model from Runway. It applies text instructions and keyframe-guided edits across existing footage while preserving details that are not meant to change....

    Provider
    Runway
    Capabilities
    Video
    Catalog updated
  • Runway: Gen-4.5

    gen-4.5

    Runway Gen-4.5 is a video generation model from Runway for text-to-video and image-to-video workflows. It is designed for cinematic scene creation with strong motion quality, visual fidelity, and prompt adherence....

    Provider
    Runway
    Capabilities
    Video
    Catalog updated
  • SpaceXAI: Grok Imagine Video 1.5

    grok-imagine-video-1.5

    Grok Imagine Video 1.5 is a video generation model from SpaceXAI. It creates videos from text prompts, with an optional starting image to guide the scene. It can direct subject...

    Provider
    xAI
    Capabilities
    Video
    Catalog updated
  • Anthropic: Claude Sonnet 5

    claude-sonnet-5

    Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 169.09
    Output RUB / 1M
    Output RUB / 1M: 845.45
  • MoonshotAI: Kimi K3

    kimi-k3

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

    Provider
    Moonshot AI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 253.63
    Output RUB / 1M
    Output RUB / 1M: 1,268.17
  • Google: Gemini 3 Pro Image

    gemini-3-pro-image

    Advanced Google image model for detailed generation and editing at resolutions up to 4K.

    Provider
    Google
    Capabilities
    Image
    Catalog updated
  • Google: Gemini 3.1 Flash Image

    gemini-3.1-flash-image

    Google image model for generation and editing with multiple output resolutions and reference images.

    Provider
    Google
    Capabilities
    Image
    Catalog updated
  • Google: Gemini 3.1 Flash Lite Image

    gemini-3.1-flash-lite-image

    Efficient Google image model for fast generation and editing across common aspect ratios.

    Provider
    Google
    Capabilities
    Image
    Catalog updated
  • xAI: Grok 4.5

    grok-4.5

    Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

    Provider
    xAI
    Context
    500K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 169.09
    Output RUB / 1M
    Output RUB / 1M: 507.27
  • OpenAI: GPT-5.6 Luna

    gpt-5.6-luna

    GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 16.91
    Output RUB / 1M
    Output RUB / 1M: 101.45
  • OpenAI: GPT-5.6 Sol

    gpt-5.6-sol

    GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,536.35
  • OpenAI: GPT-5.6 Terra

    gpt-5.6-terra

    GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 211.36
    Output RUB / 1M
    Output RUB / 1M: 1,268.17
  • Google: Gemini 2.5 Flash

    gemini-2.5-flash

    Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 25.36
    Output RUB / 1M
    Output RUB / 1M: 211.36
  • Google: Gemini 2.5 Flash Lite

    gemini-2.5-flash-lite

    Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 8.45
    Output RUB / 1M
    Output RUB / 1M: 33.82
  • MoonshotAI: Kimi K2.7 Code

    kimi-k2.7-code

    MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

    Provider
    Moonshot AI
    Context
    262.14K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 80.32
    Output RUB / 1M
    Output RUB / 1M: 338.18
  • Z.ai: GLM 5.2

    glm-5.2

    GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

    Provider
    Z.ai
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 118.36
    Output RUB / 1M
    Output RUB / 1M: 372
  • Alibaba: HappyHorse 1.1

    happyhorse-1.1

    HappyHorse 1.1 is a video generation model from Alibaba. It generates short videos from a text prompt, a single starting image, or a set of reference images, with output up...

    Provider
    Alibaba
    Capabilities
    Video
    Catalog updated
  • Alibaba: HappyHorse 1.0

    happyhorse-1.0

    HappyHorse 1.0 is a video generation model from Alibaba. It generates short videos from a text prompt, a single starting image, or a set of reference images, with output up...

    Provider
    Alibaba
    Capabilities
    Video
    Catalog updated
  • Anthropic: Claude Fable 5

    claude-fable-5

    Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 845.45
    Output RUB / 1M
    Output RUB / 1M: 4,227.25
  • Anthropic: Claude Opus 4.8

    claude-opus-4.8

    Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,113.62
  • Anthropic: Claude Opus 4.6

    claude-opus-4.6

    Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,113.62
  • Anthropic: Claude Opus 4.7

    claude-opus-4.7

    Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,113.62
  • Anthropic: Claude Sonnet 4.6

    claude-sonnet-4.6

    Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 253.63
    Output RUB / 1M
    Output RUB / 1M: 1,268.17
  • DeepSeek: DeepSeek V4 Flash 0423

    deepseek-v4-flash

    DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

    Provider
    DeepSeek
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 11.84
    Output RUB / 1M
    Output RUB / 1M: 23.67
  • DeepSeek: DeepSeek V4 Pro

    deepseek-v4-pro

    DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

    Provider
    DeepSeek
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 36.78
    Output RUB / 1M
    Output RUB / 1M: 73.55
  • Google: Gemini 3 Flash Preview

    gemini-3-flash-preview

    Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 42.27
    Output RUB / 1M
    Output RUB / 1M: 253.63
  • Google: Gemini 3.1 Flash Lite

    gemini-3.1-flash-lite

    Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 21.14
    Output RUB / 1M
    Output RUB / 1M: 126.82
  • Google: Gemini 3.1 Pro Preview

    gemini-3.1-pro-preview

    Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 169.09
    Output RUB / 1M
    Output RUB / 1M: 1,014.54
  • Google: Gemini 3.5 Flash

    gemini-3.5-flash

    Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 126.82
    Output RUB / 1M
    Output RUB / 1M: 760.9
  • MiniMax: MiniMax M2.7

    minimax-m2.7

    MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

    Provider
    MiniMax
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 25.36
    Output RUB / 1M
    Output RUB / 1M: 101.45
  • MiniMax: MiniMax M3

    minimax-m3

    MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

    Provider
    MiniMax
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 25.36
    Output RUB / 1M
    Output RUB / 1M: 101.45
  • MoonshotAI: Kimi K2.6

    kimi-k2.6

    Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

    Provider
    Moonshot AI
    Context
    262.14K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 80.32
    Output RUB / 1M
    Output RUB / 1M: 338.18
  • OpenAI: GPT-5.4

    gpt-5.4

    GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 211.36
    Output RUB / 1M
    Output RUB / 1M: 1,268.17
  • OpenAI: GPT-5.4 Mini

    gpt-5.4-mini

    GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

    Provider
    OpenAI
    Context
    400K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 63.41
    Output RUB / 1M
    Output RUB / 1M: 380.45
  • OpenAI: GPT-5.4 Nano

    gpt-5.4-nano

    GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

    Provider
    OpenAI
    Context
    400K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 16.91
    Output RUB / 1M
    Output RUB / 1M: 105.68
  • OpenAI: GPT-5.5

    gpt-5.5

    GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,536.35
  • Qwen: Qwen3.7 Max

    qwen3.7-max

    Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

    Provider
    Qwen
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 101.45
    Output RUB / 1M
    Output RUB / 1M: 507.27
  • xAI: Grok 4.1 Fast

    grok-4.1-fast

    Grok 4.1 Fast is xAI's best agentic tool calling model that shines in real-world use cases like customer support and deep research. 2M context window. Reasoning can be enabled/disabled using...

    Provider
    xAI
    Context
    128K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 105.68
    Output RUB / 1M
    Output RUB / 1M: 211.36
  • xAI: Grok 4.3

    grok-4.3

    Grok 4.3 is a reasoning model from xAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

    Provider
    xAI
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 105.68
    Output RUB / 1M
    Output RUB / 1M: 211.36
  • xAI: Grok Build 0.1

    grok-build-0.1

    Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

    Provider
    xAI
    Context
    256K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 105.68
    Output RUB / 1M
    Output RUB / 1M: 211.36
  • Xiaomi: MiMo-V2.5

    mimo-v2.5

    MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

    Provider
    Xiaomi
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 11.84
    Output RUB / 1M
    Output RUB / 1M: 23.67
  • Xiaomi: MiMo-V2.5-Pro

    mimo-v2.5-pro

    MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....

    Provider
    Xiaomi
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 36.78
    Output RUB / 1M
    Output RUB / 1M: 73.55
  • Z.ai: GLM 5.1

    glm-5.1

    GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 118.36
    Output RUB / 1M
    Output RUB / 1M: 372
  • OpenAI: GPT Image 2

    gpt-image-2

    OpenAI image model for image generation and editing with reference images and masks.

    Provider
    OpenAI
    Capabilities
    Image
    Catalog updated
  • SpaceXAI: Grok Imagine Video

    grok-imagine-video

    Grok Imagine Video is SpaceXAI's fast, text-, image-, and reference-conditioned video generation model. It produces short videos (1–15 seconds, 24 fps) at 480p or 720p across seven aspect ratios -...

    Provider
    xAI
    Capabilities
    Video
    Catalog updated
  • Kling: Video v3.0 Pro

    kling-v3.0-pro

    Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...

    Provider
    Kuaishou
    Capabilities
    Video
    Catalog updated
  • Kling: Video v3.0 Standard

    kling-v3.0-std

    Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...

    Provider
    Kuaishou
    Capabilities
    Video
    Catalog updated
  • Google: Veo 3.1 Fast

    veo-3.1-fast

    Google's mid-tier video generation model balancing speed and quality. Veo 3.1 Fast generates high-quality video from text or image prompts with native synchronized audio, offering faster turnaround than Veo 3.1...

    Provider
    Google
    Capabilities
    Video
    Catalog updated
  • Google: Veo 3.1 Lite

    veo-3.1-lite

    Google's most cost-effective video generation model, designed for high-volume applications and rapid iteration. Veo 3.1 Lite generates 720p and 1080p video from text or image prompts with native synchronized audio...

    Provider
    Google
    Capabilities
    Video
    Catalog updated
  • Kling: Video O1

    kling-video-o1

    Kling Video O1 is a video generation model from Kuaishou. It supports text and image inputs with video output, enabling text-to-video and image-to-video workflows. It is suited for cinematic content...

    Provider
    Kuaishou
    Capabilities
    Video
    Catalog updated
  • MiniMax: Hailuo 2.3

    hailuo-2.3

    Hailuo 2.3 is a video generation model from MiniMax. It accepts text prompts and reference images as input and generates video output, supporting both text-to-video and image-to-video workflows. It is...

    Provider
    MiniMax
    Capabilities
    Video
    Catalog updated
  • Alibaba: Wan 2.7

    wan-2.7

    Wan 2.7 is a video generation model from Alibaba. It supports text-to-video, image-to-video with first and last frame control, and reference-to-video, where multiple reference images guide the style and content...

    Provider
    Alibaba
    Capabilities
    Video
    Catalog updated
  • ByteDance: Seedance 2.0

    seedance-2.0

    Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...

    Provider
    ByteDance
    Capabilities
    Video
    Catalog updated
  • ByteDance: Seedance 2.0 Fast

    seedance-2.0-fast

    Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...

    Provider
    ByteDance
    Capabilities
    Video
    Catalog updated
  • Alibaba: Wan 2.6

    wan-2.6

    Alibaba's most advanced video generation model, supporting over 10 visual creation capabilities in a unified system. Wan 2.6 generates 1080p video at 24fps from text, images, reference videos, or audio,...

    Provider
    Alibaba
    Capabilities
    Video
    Catalog updated
  • ByteDance: Seedance 1.5 Pro

    seedance-1-5-pro

    ByteDance's next-generation audio-visual generation model with a 4.5B parameter Dual-Branch Diffusion Transformer architecture. Seedance 1.5 Pro generates video and audio simultaneously in a single unified pass — eliminating the timing...

    Provider
    ByteDance
    Capabilities
    Video
    Catalog updated
  • OpenAI: Sora 2 Pro

    sora-2-pro

    OpenAI's flagship video generation model, delivering production-quality video with physics-accurate motion, synchronized audio, and world-state persistence across shots. Sora 2 Pro follows intricate multi-shot instructions while maintaining consistent spatial relationships...

    Provider
    OpenAI
    Capabilities
    Video
    Catalog updated
  • Google: Veo 3.1

    veo-3.1

    Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —...

    Provider
    Google
    Capabilities
    Video
    Catalog updated

From catalog to request

Move from a public model ID to a visible request path

Create a key in the cabinet, point a compatible client to the endpoint, then choose a public model from this catalog. Usage and published pricing stay visible alongside the work.

  1. 01Create a platform API key.
  2. 02Set the compatible endpoint.
  3. 03Choose a public model ID.
  4. 04Review usage and pricing.
Data source
Live public catalog
Catalog updated

Choosing and using models

Use the live catalog for availability and request capabilities.

Check its unavailable reason in the current catalog. Personal and shared workspaces can follow different access rules.

Use the public catalog or model discovery. Model IDs and availability can change, so avoid relying on an old list.

Review the selected model’s published modalities, limits, and supported request parameters before sending a request.

Start with the required modality and capability, then compare current context, limits, availability, and price.

Read the full FAQ