Skip to content
provod.ai

/v1/models

Model catalog with the details needed to choose

Each row identifies a public model and its documented request capabilities.

74 models

  • Gemini 2.5 Pro

    gemini-2.5-pro

    Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    104.45
    Output RUB / 1M
    835.59
  • Nemotron 3 Ultra

    nemotron-3-ultra-550b-a55b

    NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

    Provider
    NVIDIA
    Context
    202.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    50.14
    Output RUB / 1M
    200.54
  • Hy4 preview

    hy4-preview

    Tencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total. It is designed for coding agents, complex tool-use workflows, and productivity tasks that...

    Provider
    tencent
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    139.38
    Output RUB / 1M
    417.96
  • GPT-4o-mini

    gpt-4o-mini

    GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

    Provider
    OpenAI
    Context
    128K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    25.07
    Output RUB / 1M
    100.27
  • GPT-4.1 Nano

    gpt-4.1-nano

    For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    16.71
    Output RUB / 1M
    66.85
  • GPT-4.1 Mini

    gpt-4.1-mini

    GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    66.85
    Output RUB / 1M
    267.39
  • GPT-4.1

    gpt-4.1

    GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    334.24
    Output RUB / 1M
    1,336.94
  • Qwen3.8 Flash

    qwen3.8-flash

    Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

    Provider
    Qwen
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    25.07
    Output RUB / 1M
    78.55
  • Grok 4.6

    grok-4.6

    Grok 4.6 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM. It is succeeded by [Grok 4.7](/x-ai/grok-4.7).

    Provider
    xAI
    Context
    500K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    167.12
    Output RUB / 1M
    501.35
  • Gemini 3.7 Flash

    gemini-3.7-flash

    Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    62.67
    Output RUB / 1M
    313.35
  • Gemini 3.6 Flash

    gemini-3.6-flash

    Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    62.67
    Output RUB / 1M
    313.35
  • Gemini 3.5 Flash Lite

    gemini-3.5-flash-lite

    Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    25.07
    Output RUB / 1M
    208.90
  • DeepSeek V4 Pro 0813

    deepseek-v4-pro-0813

    DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

    Provider
    DeepSeek
    Context
    1.02M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    77.21
    Output RUB / 1M
    231.62
  • Claude Haiku 4.5

    claude-haiku-4.5

    Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

    Provider
    Anthropic
    Context
    200K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    83.56
    Output RUB / 1M
    417.79
  • GLM 5.3 FlashX

    glm-5.3-flashx

    GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

    Provider
    Z.ai
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    30.92
    Output RUB / 1M
    104.45
  • Qwen3.8 Max (0902)

    qwen3.8-max-0902

    Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

    Provider
    Qwen
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    167.12
    Output RUB / 1M
    501.35
  • Grok 4.7

    grok-4.7

    Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...

    Provider
    xAI
    Context
    500K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    167.12
    Output RUB / 1M
    501.35
  • MiMo-V2.6-Flash

    mimo-v2.6-flash

    MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...

    Provider
    Xiaomi
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    11.70
    Output RUB / 1M
    23.40
  • MiMo-V2.6-Pro

    mimo-v2.6-pro

    MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding...

    Provider
    Xiaomi
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    36.35
    Output RUB / 1M
    72.70
  • DeepSeek V4.1 Flash

    deepseek-v4.1-flash

    DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

    Provider
    DeepSeek
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    25.07
    Output RUB / 1M
    100.27
  • Claude Fable 5.1

    claude-fable-5.1

    Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    835.59
    Output RUB / 1M
    4,177.94
  • Gemini 3.8 Flash

    gemini-3.8-flash

    Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    62.67
    Output RUB / 1M
    313.35
  • GPT-6 Astra

    gpt-6-astra

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    835.59
    Output RUB / 1M
    4,177.94
  • GPT-6 Luna

    gpt-6-luna

    GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    8.36
    Output RUB / 1M
    41.78
  • GPT-6 Sol

    gpt-6-sol

    GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    167.12
    Output RUB / 1M
    835.59
  • Claude Opus 5.5

    claude-opus-5.5

    Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    334.24
    Output RUB / 1M
    1,671.18
  • Z.ai: GLM 5.3 Flash

    glm-5.3-flash

    GLM 5.3 Flash is Z.ai’s efficient multimodal reasoning model for long-context and agent workflows.

    Provider
    Z.ai
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    12.53
    Output RUB / 1M
    41.78
  • Z.ai: GLM 5.3

    glm-5.3

    GLM 5.3 is Z.ai’s reasoning model for long-context text and agent workflows.

    Provider
    Z.ai
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    116.98
    Output RUB / 1M
    367.66
  • Qwen: Qwen3 Max Preview

    qwen3-max-preview

    Qwen3-Max-Preview is the flagship model of the Qwen3 generation, built for complex agentic, coding, reasoning, multilingual, retrieval, and tool-use workloads. This route provides text input and output, function calling, structured outputs, streaming, and automatic prefix caching.

    Provider
    Qwen
    Context
    262.14K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    100.27
    Output RUB / 1M
    501.35
  • Z.ai: GLM 4.5

    glm-4.5

    GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...

    Provider
    Z.ai
    Context
    131.07K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    50.14
    Output RUB / 1M
    183.83
  • Z.ai: GLM 4.5 Air

    glm-4.5-air

    GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

    Provider
    Z.ai
    Context
    131.07K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    16.71
    Output RUB / 1M
    91.91
  • Z.ai: GLM 4.5V

    glm-4.5v

    GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

    Provider
    Z.ai
    Context
    65.54K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    50.14
    Output RUB / 1M
    150.41
  • Z.ai: GLM 4.6

    glm-4.6

    Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    50.14
    Output RUB / 1M
    183.83
  • Z.ai: GLM 4.6V

    glm-4.6v

    GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...

    Provider
    Z.ai
    Context
    131.07K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    25.07
    Output RUB / 1M
    75.20
  • Z.ai: GLM 4.7

    glm-4.7

    GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    50.14
    Output RUB / 1M
    183.83
  • Z.ai: GLM 5

    glm-5

    GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    83.56
    Output RUB / 1M
    267.39
  • Z.ai: GLM 5 Turbo

    glm-5-turbo

    GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

    Provider
    Z.ai
    Context
    202.75K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    100.27
    Output RUB / 1M
    334.24
  • Z.ai: GLM 5V Turbo

    glm-5v-turbo

    GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

    Provider
    Z.ai
    Context
    202.75K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    100.27
    Output RUB / 1M
    334.24
  • Claude Opus 5

    claude-opus-5

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    417.79
    Output RUB / 1M
    2,088.97
  • DeepSeek: DeepSeek V4 Flash 0731

    deepseek-v4-flash-0731

    DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.

    Provider
    DeepSeek
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    55.15
    Output RUB / 1M
    165.45
  • Anthropic: Claude Sonnet 5

    claude-sonnet-5

    Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    167.12
    Output RUB / 1M
    835.59
  • MoonshotAI: Kimi K3

    kimi-k3

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

    Provider
    Moonshot AI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    250.68
    Output RUB / 1M
    1,253.38
  • xAI: Grok 4.5

    grok-4.5

    Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

    Provider
    xAI
    Context
    500K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    167.12
    Output RUB / 1M
    501.35
  • OpenAI: GPT-5.6 Luna

    gpt-5.6-luna

    GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    16.71
    Output RUB / 1M
    100.27
  • OpenAI: GPT-5.6 Sol

    gpt-5.6-sol

    GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    417.79
    Output RUB / 1M
    2,506.76
  • OpenAI: GPT-5.6 Terra

    gpt-5.6-terra

    GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    208.90
    Output RUB / 1M
    1,253.38
  • Google: Gemini 2.5 Flash

    gemini-2.5-flash

    Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    25.07
    Output RUB / 1M
    208.90
  • Google: Gemini 2.5 Flash Lite

    gemini-2.5-flash-lite

    Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    8.36
    Output RUB / 1M
    33.42
  • MoonshotAI: Kimi K2.7 Code

    kimi-k2.7-code

    MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

    Provider
    Moonshot AI
    Context
    262.14K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    79.38
    Output RUB / 1M
    334.24
  • Z.ai: GLM 5.2

    glm-5.2

    GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

    Provider
    Z.ai
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    80.72
    Output RUB / 1M
    253.68
  • Anthropic: Claude Fable 5

    claude-fable-5

    Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    835.59
    Output RUB / 1M
    4,177.94
  • Anthropic: Claude Opus 4.8

    claude-opus-4.8

    Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    417.79
    Output RUB / 1M
    2,088.97
  • Anthropic: Claude Opus 4.6

    claude-opus-4.6

    Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    417.79
    Output RUB / 1M
    2,088.97
  • Anthropic: Claude Opus 4.7

    claude-opus-4.7

    Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    417.79
    Output RUB / 1M
    2,088.97
  • Anthropic: Claude Sonnet 4.6

    claude-sonnet-4.6

    Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    250.68
    Output RUB / 1M
    1,253.38
  • DeepSeek: DeepSeek V4 Flash 0423

    deepseek-v4-flash

    DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

    Provider
    DeepSeek
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    55.15
    Output RUB / 1M
    165.45
  • DeepSeek: DeepSeek V4 Pro

    deepseek-v4-pro

    DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

    Provider
    DeepSeek
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    165.45
    Output RUB / 1M
    496.34
  • Google: Gemini 3 Flash Preview

    gemini-3-flash-preview

    Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    41.78
    Output RUB / 1M
    250.68
  • Google: Gemini 3.1 Flash Lite

    gemini-3.1-flash-lite

    Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    20.89
    Output RUB / 1M
    125.34
  • Google: Gemini 3.1 Pro Preview

    gemini-3.1-pro-preview

    Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    167.12
    Output RUB / 1M
    1,002.71
  • Google: Gemini 3.5 Flash

    gemini-3.5-flash

    Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    125.34
    Output RUB / 1M
    752.03
  • MiniMax: MiniMax M2.7

    minimax-m2.7

    MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

    Provider
    MiniMax
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    25.07
    Output RUB / 1M
    100.27
  • MiniMax: MiniMax M3

    minimax-m3

    MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

    Provider
    MiniMax
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    25.07
    Output RUB / 1M
    100.27
  • MoonshotAI: Kimi K2.6

    kimi-k2.6

    Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

    Provider
    Moonshot AI
    Context
    262.14K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    79.38
    Output RUB / 1M
    334.24
  • OpenAI: GPT-5.4

    gpt-5.4

    GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    208.90
    Output RUB / 1M
    1,253.38
  • OpenAI: GPT-5.4 Mini

    gpt-5.4-mini

    GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

    Provider
    OpenAI
    Context
    400K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    62.67
    Output RUB / 1M
    376.01
  • OpenAI: GPT-5.4 Nano

    gpt-5.4-nano

    GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

    Provider
    OpenAI
    Context
    400K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    16.71
    Output RUB / 1M
    104.45
  • OpenAI: GPT-5.5

    gpt-5.5

    GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    417.79
    Output RUB / 1M
    2,506.76
  • Qwen: Qwen3.7 Max

    qwen3.7-max

    Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

    Provider
    Qwen
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    166.11
    Output RUB / 1M
    506.78
  • xAI: Grok 4.3

    grok-4.3

    Grok 4.3 is a reasoning model from xAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

    Provider
    xAI
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    104.45
    Output RUB / 1M
    208.90
  • xAI: Grok Build 0.1

    grok-build-0.1

    Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

    Provider
    xAI
    Context
    256K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    104.45
    Output RUB / 1M
    208.90
  • Xiaomi: MiMo-V2.5

    mimo-v2.5

    MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

    Provider
    Xiaomi
    Context
    262.14K
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    11.70
    Output RUB / 1M
    23.40
  • Xiaomi: MiMo-V2.5-Pro

    mimo-v2.5-pro

    MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....

    Provider
    Xiaomi
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    36.35
    Output RUB / 1M
    72.70
  • Z.ai: GLM 5.1

    glm-5.1

    GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    116.98
    Output RUB / 1M
    367.66

From catalog to request

Move from a public model ID to a visible request path

Create a key in the cabinet, point a compatible client to the endpoint, then choose a public model from this catalog. Usage and published pricing stay visible alongside the work.

  1. 01Create a platform API key.
  2. 02Set the compatible endpoint.
  3. 03Choose a public model ID.
  4. 04Review usage and pricing.
Data source
Live public catalog
Catalog updated

Choosing and using models

Use the live catalog for availability and request capabilities.

Check its unavailable reason in the current catalog. Personal and shared workspaces can follow different access rules.

Use the public catalog or model discovery. Model IDs and availability can change, so avoid relying on an old list.

Review the selected model’s published modalities, limits, and supported request parameters before sending a request.

Start with the required modality and capability, then compare current context, limits, availability, and price.

Read the full FAQ