/v1/models
Model catalog with the details needed to choose
Each row identifies a public model and its documented request capabilities.
102 models
GPT Image 2.5 Flare
gpt-image-2.5-flareOpenAI image generation and editing. Speed-oriented tier.
- Provider
- OpenAI
- Capabilities
- Image
- Catalog updated
GPT Image 2.5 Sunburst
gpt-image-2.5-sunburstOpenAI image generation and editing. Precision-oriented tier.
- Provider
- OpenAI
- Capabilities
- Image
- Catalog updated
Gemini 2.5 Pro
gemini-2.5-proGemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
- Provider
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 104.45
- Output RUB / 1M
- 835.59
Nemotron 3 Ultra
nemotron-3-ultra-550b-a55bNVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
- Provider
- NVIDIA
- Context
- 202.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 50.14
- Output RUB / 1M
- 200.54
Hy4 preview
hy4-previewTencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total. It is designed for coding agents, complex tool-use workflows, and productivity tasks that...
- Provider
- tencent
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 139.38
- Output RUB / 1M
- 417.96
GPT-4o-mini
gpt-4o-miniGPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
- Provider
- OpenAI
- Context
- 128K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 25.07
- Output RUB / 1M
- 100.27
GPT-4.1 Nano
gpt-4.1-nanoFor tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 16.71
- Output RUB / 1M
- 66.85
GPT-4.1 Mini
gpt-4.1-miniGPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 66.85
- Output RUB / 1M
- 267.39
GPT-4.1
gpt-4.1GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 334.24
- Output RUB / 1M
- 1,336.94
Qwen3.8 Flash
qwen3.8-flashQwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
- Provider
- Qwen
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 25.07
- Output RUB / 1M
- 78.55
Grok 4.6
grok-4.6Grok 4.6 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM. It is succeeded by [Grok 4.7](/x-ai/grok-4.7).
- Provider
- xAI
- Context
- 500K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 167.12
- Output RUB / 1M
- 501.35
Gemini 3.7 Flash
gemini-3.7-flashGemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
- Provider
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 62.67
- Output RUB / 1M
- 313.35
Gemini 3.6 Flash
gemini-3.6-flashGemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
- Provider
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 62.67
- Output RUB / 1M
- 313.35
Gemini 3.5 Flash Lite
gemini-3.5-flash-liteGemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
- Provider
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 25.07
- Output RUB / 1M
- 208.90
DeepSeek V4 Pro 0813
deepseek-v4-pro-0813DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.
- Provider
- DeepSeek
- Context
- 1.02M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 77.21
- Output RUB / 1M
- 231.62
Claude Haiku 4.5
claude-haiku-4.5Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
- Provider
- Anthropic
- Context
- 200K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 83.56
- Output RUB / 1M
- 417.79
GLM 5.3 FlashX
glm-5.3-flashxGLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
- Provider
- Z.ai
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 30.92
- Output RUB / 1M
- 104.45
Qwen3.8 Max (0902)
qwen3.8-max-0902Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
- Provider
- Qwen
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 167.12
- Output RUB / 1M
- 501.35
Grok 4.7
grok-4.7Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...
- Provider
- xAI
- Context
- 500K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 167.12
- Output RUB / 1M
- 501.35
MiMo-V2.6-Flash
mimo-v2.6-flashMiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...
- Provider
- Xiaomi
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- 11.70
- Output RUB / 1M
- 23.40
MiMo-V2.6-Pro
mimo-v2.6-proMiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding...
- Provider
- Xiaomi
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- 36.35
- Output RUB / 1M
- 72.70
DeepSeek V4.1 Flash
deepseek-v4.1-flashDeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
- Provider
- DeepSeek
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 25.07
- Output RUB / 1M
- 100.27
Claude Fable 5.1
claude-fable-5.1Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 835.59
- Output RUB / 1M
- 4,177.94
Gemini 3.8 Flash
gemini-3.8-flashGemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- 62.67
- Output RUB / 1M
- 313.35
GPT-6 Astra
gpt-6-astraGPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 835.59
- Output RUB / 1M
- 4,177.94
GPT-6 Luna
gpt-6-lunaGPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 8.36
- Output RUB / 1M
- 41.78
GPT-6 Sol
gpt-6-solGPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 167.12
- Output RUB / 1M
- 835.59
Claude Opus 5.5
claude-opus-5.5Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 334.24
- Output RUB / 1M
- 1,671.18
Z.ai: GLM 5.3 Flash
glm-5.3-flashGLM 5.3 Flash is Z.ai’s efficient multimodal reasoning model for long-context and agent workflows.
- Provider
- Z.ai
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 12.53
- Output RUB / 1M
- 41.78
Z.ai: GLM 5.3
glm-5.3GLM 5.3 is Z.ai’s reasoning model for long-context text and agent workflows.
- Provider
- Z.ai
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 116.98
- Output RUB / 1M
- 367.66
Qwen: Qwen3 Max Preview
qwen3-max-previewQwen3-Max-Preview is the flagship model of the Qwen3 generation, built for complex agentic, coding, reasoning, multilingual, retrieval, and tool-use workloads. This route provides text input and output, function calling, structured outputs, streaming, and automatic prefix caching.
- Provider
- Qwen
- Context
- 262.14K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 100.27
- Output RUB / 1M
- 501.35
Z.ai: GLM 4.5
glm-4.5GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
- Provider
- Z.ai
- Context
- 131.07K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 50.14
- Output RUB / 1M
- 183.83
Z.ai: GLM 4.5 Air
glm-4.5-airGLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...
- Provider
- Z.ai
- Context
- 131.07K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 16.71
- Output RUB / 1M
- 91.91
Z.ai: GLM 4.5V
glm-4.5vGLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
- Provider
- Z.ai
- Context
- 65.54K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 50.14
- Output RUB / 1M
- 150.41
Z.ai: GLM 4.6
glm-4.6Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
- Provider
- Z.ai
- Context
- 204.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 50.14
- Output RUB / 1M
- 183.83
Z.ai: GLM 4.6V
glm-4.6vGLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...
- Provider
- Z.ai
- Context
- 131.07K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 25.07
- Output RUB / 1M
- 75.20
Z.ai: GLM 4.7
glm-4.7GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...
- Provider
- Z.ai
- Context
- 204.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 50.14
- Output RUB / 1M
- 183.83
Z.ai: GLM 5
glm-5GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
- Provider
- Z.ai
- Context
- 204.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 83.56
- Output RUB / 1M
- 267.39
Z.ai: GLM 5 Turbo
glm-5-turboGLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...
- Provider
- Z.ai
- Context
- 202.75K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 100.27
- Output RUB / 1M
- 334.24
Z.ai: GLM 5V Turbo
glm-5v-turboGLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
- Provider
- Z.ai
- Context
- 202.75K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 100.27
- Output RUB / 1M
- 334.24
Seedance 2.5 is a video generation model from ByteDance. It is suited for long-form storytelling, multimodal reference-based generation, video editing, and video extension. It supports first-frame and first-and-last-frame control, up...
- Provider
- ByteDance
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 7.70 ₽/s · Minimum duration: 4 s
FLUX.3 Video is a video generation model from Black Forest Labs. It supports text-to-video, image-guided generation with opening and closing keyframes, and video continuation workflows, making it suited for controlled...
- Provider
- Black Forest Labs
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 21.31 ₽/s · Minimum duration: 5 s
Claude Opus 5
claude-opus-5Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 417.79
- Output RUB / 1M
- 2,088.97
DeepSeek: DeepSeek V4 Flash 0731
deepseek-v4-flash-0731DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.
- Provider
- DeepSeek
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 55.15
- Output RUB / 1M
- 165.45
MiniMax H3 is a lightweight, open-weights video generation model from MiniMax. It is designed for precise multimodal editing and controlled content generation, including instruction-guided edits, text and brand rendering, and...
- Provider
- MiniMax
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 16.29 ₽/s · Minimum duration: 5 s
Runway: Aleph 2.0
aleph-2Runway Aleph 2.0 is an in-context video editing model from Runway. It applies text instructions and keyframe-guided edits across existing footage while preserving details that are not meant to change....
- Provider
- Runway
- Capabilities
- Video
- Catalog updated
Runway Gen-4.5 is a video generation model from Runway for text-to-video and image-to-video workflows. It is designed for cinematic scene creation with strong motion quality, visual fidelity, and prompt adherence....
- Provider
- Runway
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 15.04 ₽/s · Minimum duration: 2 s
Grok Imagine Video 1.5 is a video generation model from SpaceXAI. It creates videos from text prompts, with an optional starting image to guide the scene. It can direct subject...
- Provider
- xAI
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 10.03 ₽/s · Minimum duration: 1 s
Anthropic: Claude Sonnet 5
claude-sonnet-5Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 167.12
- Output RUB / 1M
- 835.59
MoonshotAI: Kimi K3
kimi-k3Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
- Provider
- Moonshot AI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 250.68
- Output RUB / 1M
- 1,253.38
Google: Gemini 3 Pro Image
gemini-3-pro-imageAdvanced Google image model for detailed generation and editing at resolutions up to 4K.
- Provider
- Capabilities
- Image
- Catalog updated
Google: Gemini 3.1 Flash Image
gemini-3.1-flash-imageGoogle image model for generation and editing with multiple output resolutions and reference images.
- Provider
- Capabilities
- Image
- Catalog updated
Google: Gemini 3.1 Flash Lite Image
gemini-3.1-flash-lite-imageEfficient Google image model for fast generation and editing across common aspect ratios.
- Provider
- Capabilities
- Image
- Catalog updated
xAI: Grok 4.5
grok-4.5Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
- Provider
- xAI
- Context
- 500K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 167.12
- Output RUB / 1M
- 501.35
OpenAI: GPT-5.6 Luna
gpt-5.6-lunaGPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 16.71
- Output RUB / 1M
- 100.27
OpenAI: GPT-5.6 Sol
gpt-5.6-solGPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 417.79
- Output RUB / 1M
- 2,506.76
OpenAI: GPT-5.6 Terra
gpt-5.6-terraGPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 208.90
- Output RUB / 1M
- 1,253.38
Google: Gemini 2.5 Flash
gemini-2.5-flashGemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- 25.07
- Output RUB / 1M
- 208.90
Google: Gemini 2.5 Flash Lite
gemini-2.5-flash-liteGemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- 8.36
- Output RUB / 1M
- 33.42
MoonshotAI: Kimi K2.7 Code
kimi-k2.7-codeMoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
- Provider
- Moonshot AI
- Context
- 262.14K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 79.38
- Output RUB / 1M
- 334.24
Z.ai: GLM 5.2
glm-5.2GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
- Provider
- Z.ai
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 80.72
- Output RUB / 1M
- 253.68
HappyHorse 1.1 is a video generation model from Alibaba. It generates short videos from a text prompt, a single starting image, or a set of reference images, with output up...
- Provider
- Alibaba
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 12.38 ₽/s · Minimum duration: 3 s
HappyHorse 1.0 is a video generation model from Alibaba. It generates short videos from a text prompt, a single starting image, or a set of reference images, with output up...
- Provider
- Alibaba
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 12.38 ₽/s · Minimum duration: 3 s
Anthropic: Claude Fable 5
claude-fable-5Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 835.59
- Output RUB / 1M
- 4,177.94
Anthropic: Claude Opus 4.8
claude-opus-4.8Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 417.79
- Output RUB / 1M
- 2,088.97
Anthropic: Claude Opus 4.6
claude-opus-4.6Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 417.79
- Output RUB / 1M
- 2,088.97
Anthropic: Claude Opus 4.7
claude-opus-4.7Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 417.79
- Output RUB / 1M
- 2,088.97
Anthropic: Claude Sonnet 4.6
claude-sonnet-4.6Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 250.68
- Output RUB / 1M
- 1,253.38
DeepSeek: DeepSeek V4 Flash 0423
deepseek-v4-flashDeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
- Provider
- DeepSeek
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 55.15
- Output RUB / 1M
- 165.45
DeepSeek: DeepSeek V4 Pro
deepseek-v4-proDeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
- Provider
- DeepSeek
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 165.45
- Output RUB / 1M
- 496.34
Google: Gemini 3 Flash Preview
gemini-3-flash-previewGemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- 41.78
- Output RUB / 1M
- 250.68
Google: Gemini 3.1 Flash Lite
gemini-3.1-flash-liteGemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- 20.89
- Output RUB / 1M
- 125.34
Google: Gemini 3.1 Pro Preview
gemini-3.1-pro-previewGemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- 167.12
- Output RUB / 1M
- 1,002.71
Google: Gemini 3.5 Flash
gemini-3.5-flashGemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- 125.34
- Output RUB / 1M
- 752.03
MiniMax: MiniMax M2.7
minimax-m2.7MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
- Provider
- MiniMax
- Context
- 204.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 25.07
- Output RUB / 1M
- 100.27
MiniMax: MiniMax M3
minimax-m3MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
- Provider
- MiniMax
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 25.07
- Output RUB / 1M
- 100.27
MoonshotAI: Kimi K2.6
kimi-k2.6Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
- Provider
- Moonshot AI
- Context
- 262.14K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 79.38
- Output RUB / 1M
- 334.24
OpenAI: GPT-5.4
gpt-5.4GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 208.90
- Output RUB / 1M
- 1,253.38
OpenAI: GPT-5.4 Mini
gpt-5.4-miniGPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
- Provider
- OpenAI
- Context
- 400K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 62.67
- Output RUB / 1M
- 376.01
OpenAI: GPT-5.4 Nano
gpt-5.4-nanoGPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
- Provider
- OpenAI
- Context
- 400K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 16.71
- Output RUB / 1M
- 104.45
OpenAI: GPT-5.5
gpt-5.5GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 417.79
- Output RUB / 1M
- 2,506.76
Qwen: Qwen3.7 Max
qwen3.7-maxQwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...
- Provider
- Qwen
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 166.11
- Output RUB / 1M
- 506.78
xAI: Grok 4.3
grok-4.3Grok 4.3 is a reasoning model from xAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
- Provider
- xAI
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 104.45
- Output RUB / 1M
- 208.90
xAI: Grok Build 0.1
grok-build-0.1Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...
- Provider
- xAI
- Context
- 256K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 104.45
- Output RUB / 1M
- 208.90
Xiaomi: MiMo-V2.5
mimo-v2.5MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...
- Provider
- Xiaomi
- Context
- 262.14K
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- 11.70
- Output RUB / 1M
- 23.40
Xiaomi: MiMo-V2.5-Pro
mimo-v2.5-proMiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
- Provider
- Xiaomi
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 36.35
- Output RUB / 1M
- 72.70
Z.ai: GLM 5.1
glm-5.1GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
- Provider
- Z.ai
- Context
- 204.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- 116.98
- Output RUB / 1M
- 367.66
OpenAI: GPT Image 2
gpt-image-2OpenAI image model for image generation and editing with reference images and masks.
- Provider
- OpenAI
- Capabilities
- Image
- Catalog updated
Grok Imagine Video is SpaceXAI's fast, text-, image-, and reference-conditioned video generation model. It produces short videos (1–15 seconds, 24 fps) at 480p or 720p across seven aspect ratios -...
- Provider
- xAI
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 6.27 ₽/s · Minimum duration: 1 s
Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...
- Provider
- Kuaishou
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 14.04 ₽/s · Minimum duration: 3 s
Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...
- Provider
- Kuaishou
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 10.53 ₽/s · Minimum duration: 3 s
Google's mid-tier video generation model balancing speed and quality. Veo 3.1 Fast generates high-quality video from text or image prompts with native synchronized audio, offering faster turnaround than Veo 3.1...
- Provider
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 10.03 ₽/s · Minimum duration: 4 s
Google's most cost-effective video generation model, designed for high-volume applications and rapid iteration. Veo 3.1 Lite generates 720p and 1080p video from text or image prompts with native synchronized audio...
- Provider
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 3.76 ₽/s · Minimum duration: 4 s
Kling Video O1 is a video generation model from Kuaishou. It supports text and image inputs with video output, enabling text-to-video and image-to-video workflows. It is suited for cinematic content...
- Provider
- Kuaishou
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 14.04 ₽/s · Minimum duration: 5 s
Hailuo 2.3 is a video generation model from MiniMax. It accepts text prompts and reference images as input and generates video output, supporting both text-to-video and image-to-video workflows. It is...
- Provider
- MiniMax
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 10.24 ₽/s · Minimum duration: 6 s
Wan 2.7 is a video generation model from Alibaba. It supports text-to-video, image-to-video with first and last frame control, and reference-to-video, where multiple reference images guide the style and content...
- Provider
- Alibaba
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 12.53 ₽/s · Minimum duration: 2 s
Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...
- Provider
- ByteDance
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 2.91 ₽/s · Minimum duration: 4 s
Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...
- Provider
- ByteDance
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 1.68 ₽/s · Minimum duration: 4 s
Alibaba's most advanced video generation model, supporting over 10 visual creation capabilities in a unified system. Wan 2.6 generates 1080p video at 24fps from text, images, reference videos, or audio,...
- Provider
- Alibaba
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 10.03 ₽/s · Minimum duration: 5 s
ByteDance's next-generation audio-visual generation model with a 4.5B parameter Dual-Branch Diffusion Transformer architecture. Seedance 1.5 Pro generates video and audio simultaneously in a single unified pass — eliminating the timing...
- Provider
- ByteDance
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 0.81 ₽/s · Minimum duration: 4 s
OpenAI's flagship video generation model, delivering production-quality video with physics-accurate motion, synchronized audio, and world-state persistence across shots. Sora 2 Pro follows intricate multi-shot instructions while maintaining consistent spatial relationships...
- Provider
- OpenAI
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 37.60 ₽/s · Minimum duration: 4 s
Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —...
- Provider
- Capabilities
- Video
- Catalog updated
- Starting video price
- Starting video price: 25.07 ₽/s · Minimum duration: 4 s
From catalog to request
Move from a public model ID to a visible request path
Create a key in the cabinet, point a compatible client to the endpoint, then choose a public model from this catalog. Usage and published pricing stay visible alongside the work.
- 01Create a platform API key.
- 02Set the compatible endpoint.
- 03Choose a public model ID.
- 04Review usage and pricing.
- Data source
- Live public catalog
- Catalog updated
Choosing and using models
Use the live catalog for availability and request capabilities.
Check its unavailable reason in the current catalog. Personal and shared workspaces can follow different access rules.
Use the public catalog or model discovery. Model IDs and availability can change, so avoid relying on an old list.
Review the selected model’s published modalities, limits, and supported request parameters before sending a request.
Start with the required modality and capability, then compare current context, limits, availability, and price.