Skip to content
provod.ai
RU

/v1/models

Model catalog with the details needed to choose

Each row identifies a public model and its documented request capabilities.

46 models

  • Z.ai: GLM 4.5

    glm-4.5

    GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...

    Provider
    Z.ai
    Context
    131.07K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 50.73
    Output RUB / 1M
    Output RUB / 1M: 186
  • Z.ai: GLM 4.5 Air

    glm-4.5-air

    GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

    Provider
    Z.ai
    Context
    131.07K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 16.91
    Output RUB / 1M
    Output RUB / 1M: 93
  • Z.ai: GLM 4.5V

    glm-4.5v

    GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

    Provider
    Z.ai
    Context
    65.54K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 50.73
    Output RUB / 1M
    Output RUB / 1M: 152.18
  • Z.ai: GLM 4.6

    glm-4.6

    Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 50.73
    Output RUB / 1M
    Output RUB / 1M: 186
  • Z.ai: GLM 4.6V

    glm-4.6v

    GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...

    Provider
    Z.ai
    Context
    131.07K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 25.36
    Output RUB / 1M
    Output RUB / 1M: 76.09
  • Z.ai: GLM 4.7

    glm-4.7

    GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 50.73
    Output RUB / 1M
    Output RUB / 1M: 186
  • Z.ai: GLM 5

    glm-5

    GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 84.54
    Output RUB / 1M
    Output RUB / 1M: 270.54
  • Z.ai: GLM 5 Turbo

    glm-5-turbo

    GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

    Provider
    Z.ai
    Context
    202.75K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 101.45
    Output RUB / 1M
    Output RUB / 1M: 338.18
  • Z.ai: GLM 5V Turbo

    glm-5v-turbo

    GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

    Provider
    Z.ai
    Context
    202.75K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 101.45
    Output RUB / 1M
    Output RUB / 1M: 338.18
  • Claude Opus 5

    claude-opus-5

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,113.62
  • DeepSeek: DeepSeek V4 Flash 0731

    deepseek-v4-flash-0731

    DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.

    Provider
    DeepSeek
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 11.84
    Output RUB / 1M
    Output RUB / 1M: 23.67
  • Anthropic: Claude Sonnet 5

    claude-sonnet-5

    Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 169.09
    Output RUB / 1M
    Output RUB / 1M: 845.45
  • MoonshotAI: Kimi K3

    kimi-k3

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

    Provider
    Moonshot AI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 253.63
    Output RUB / 1M
    Output RUB / 1M: 1,268.17
  • xAI: Grok 4.5

    grok-4.5

    Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

    Provider
    xAI
    Context
    500K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 169.09
    Output RUB / 1M
    Output RUB / 1M: 507.27
  • OpenAI: GPT-5.6 Luna

    gpt-5.6-luna

    GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 16.91
    Output RUB / 1M
    Output RUB / 1M: 101.45
  • OpenAI: GPT-5.6 Sol

    gpt-5.6-sol

    GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,536.35
  • OpenAI: GPT-5.6 Terra

    gpt-5.6-terra

    GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 211.36
    Output RUB / 1M
    Output RUB / 1M: 1,268.17
  • Google: Gemini 2.5 Flash

    gemini-2.5-flash

    Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 25.36
    Output RUB / 1M
    Output RUB / 1M: 211.36
  • Google: Gemini 2.5 Flash Lite

    gemini-2.5-flash-lite

    Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 8.45
    Output RUB / 1M
    Output RUB / 1M: 33.82
  • MoonshotAI: Kimi K2.7 Code

    kimi-k2.7-code

    MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

    Provider
    Moonshot AI
    Context
    262.14K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 80.32
    Output RUB / 1M
    Output RUB / 1M: 338.18
  • Z.ai: GLM 5.2

    glm-5.2

    GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

    Provider
    Z.ai
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 118.36
    Output RUB / 1M
    Output RUB / 1M: 372
  • Anthropic: Claude Fable 5

    claude-fable-5

    Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 845.45
    Output RUB / 1M
    Output RUB / 1M: 4,227.25
  • Anthropic: Claude Opus 4.8

    claude-opus-4.8

    Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,113.62
  • Anthropic: Claude Opus 4.6

    claude-opus-4.6

    Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,113.62
  • Anthropic: Claude Opus 4.7

    claude-opus-4.7

    Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,113.62
  • Anthropic: Claude Sonnet 4.6

    claude-sonnet-4.6

    Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

    Provider
    Anthropic
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 253.63
    Output RUB / 1M
    Output RUB / 1M: 1,268.17
  • DeepSeek: DeepSeek V4 Flash 0423

    deepseek-v4-flash

    DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

    Provider
    DeepSeek
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 11.84
    Output RUB / 1M
    Output RUB / 1M: 23.67
  • DeepSeek: DeepSeek V4 Pro

    deepseek-v4-pro

    DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

    Provider
    DeepSeek
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 36.78
    Output RUB / 1M
    Output RUB / 1M: 73.55
  • Google: Gemini 3 Flash Preview

    gemini-3-flash-preview

    Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 42.27
    Output RUB / 1M
    Output RUB / 1M: 253.63
  • Google: Gemini 3.1 Flash Lite

    gemini-3.1-flash-lite

    Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 21.14
    Output RUB / 1M
    Output RUB / 1M: 126.82
  • Google: Gemini 3.1 Pro Preview

    gemini-3.1-pro-preview

    Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 169.09
    Output RUB / 1M
    Output RUB / 1M: 1,014.54
  • Google: Gemini 3.5 Flash

    gemini-3.5-flash

    Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

    Provider
    Google
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 126.82
    Output RUB / 1M
    Output RUB / 1M: 760.9
  • MiniMax: MiniMax M2.7

    minimax-m2.7

    MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

    Provider
    MiniMax
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 25.36
    Output RUB / 1M
    Output RUB / 1M: 101.45
  • MiniMax: MiniMax M3

    minimax-m3

    MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

    Provider
    MiniMax
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 25.36
    Output RUB / 1M
    Output RUB / 1M: 101.45
  • MoonshotAI: Kimi K2.6

    kimi-k2.6

    Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

    Provider
    Moonshot AI
    Context
    262.14K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 80.32
    Output RUB / 1M
    Output RUB / 1M: 338.18
  • OpenAI: GPT-5.4

    gpt-5.4

    GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 211.36
    Output RUB / 1M
    Output RUB / 1M: 1,268.17
  • OpenAI: GPT-5.4 Mini

    gpt-5.4-mini

    GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

    Provider
    OpenAI
    Context
    400K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 63.41
    Output RUB / 1M
    Output RUB / 1M: 380.45
  • OpenAI: GPT-5.4 Nano

    gpt-5.4-nano

    GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

    Provider
    OpenAI
    Context
    400K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 16.91
    Output RUB / 1M
    Output RUB / 1M: 105.68
  • OpenAI: GPT-5.5

    gpt-5.5

    GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

    Provider
    OpenAI
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 422.72
    Output RUB / 1M
    Output RUB / 1M: 2,536.35
  • Qwen: Qwen3.7 Max

    qwen3.7-max

    Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

    Provider
    Qwen
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 101.45
    Output RUB / 1M
    Output RUB / 1M: 507.27
  • xAI: Grok 4.1 Fast

    grok-4.1-fast

    Grok 4.1 Fast is xAI's best agentic tool calling model that shines in real-world use cases like customer support and deep research. 2M context window. Reasoning can be enabled/disabled using...

    Provider
    xAI
    Context
    128K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 105.68
    Output RUB / 1M
    Output RUB / 1M: 211.36
  • xAI: Grok 4.3

    grok-4.3

    Grok 4.3 is a reasoning model from xAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

    Provider
    xAI
    Context
    1M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 105.68
    Output RUB / 1M
    Output RUB / 1M: 211.36
  • xAI: Grok Build 0.1

    grok-build-0.1

    Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

    Provider
    xAI
    Context
    256K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 105.68
    Output RUB / 1M
    Output RUB / 1M: 211.36
  • Xiaomi: MiMo-V2.5

    mimo-v2.5

    MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

    Provider
    Xiaomi
    Context
    1.05M
    Capabilities
    Text · Transcription
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 11.84
    Output RUB / 1M
    Output RUB / 1M: 23.67
  • Xiaomi: MiMo-V2.5-Pro

    mimo-v2.5-pro

    MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....

    Provider
    Xiaomi
    Context
    1.05M
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 36.78
    Output RUB / 1M
    Output RUB / 1M: 73.55
  • Z.ai: GLM 5.1

    glm-5.1

    GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

    Provider
    Z.ai
    Context
    204.8K
    Capabilities
    Text
    Catalog updated
    Input RUB / 1M
    Input RUB / 1M: 118.36
    Output RUB / 1M
    Output RUB / 1M: 372

From catalog to request

Move from a public model ID to a visible request path

Create a key in the cabinet, point a compatible client to the endpoint, then choose a public model from this catalog. Usage and published pricing stay visible alongside the work.

  1. 01Create a platform API key.
  2. 02Set the compatible endpoint.
  3. 03Choose a public model ID.
  4. 04Review usage and pricing.
Data source
Live public catalog
Catalog updated

Choosing and using models

Use the live catalog for availability and request capabilities.

Check its unavailable reason in the current catalog. Personal and shared workspaces can follow different access rules.

Use the public catalog or model discovery. Model IDs and availability can change, so avoid relying on an old list.

Review the selected model’s published modalities, limits, and supported request parameters before sending a request.

Start with the required modality and capability, then compare current context, limits, availability, and price.

Read the full FAQ