/v1/models
Model catalog with the details needed to choose
Each row identifies a public model and its documented request capabilities.
46 models
Z.ai: GLM 4.5
glm-4.5GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
- Provider
- Z.ai
- Context
- 131.07K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 50.73
- Output RUB / 1M
- Output RUB / 1M: 186
Z.ai: GLM 4.5 Air
glm-4.5-airGLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...
- Provider
- Z.ai
- Context
- 131.07K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 16.91
- Output RUB / 1M
- Output RUB / 1M: 93
Z.ai: GLM 4.5V
glm-4.5vGLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
- Provider
- Z.ai
- Context
- 65.54K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 50.73
- Output RUB / 1M
- Output RUB / 1M: 152.18
Z.ai: GLM 4.6
glm-4.6Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
- Provider
- Z.ai
- Context
- 204.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 50.73
- Output RUB / 1M
- Output RUB / 1M: 186
Z.ai: GLM 4.6V
glm-4.6vGLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...
- Provider
- Z.ai
- Context
- 131.07K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 25.36
- Output RUB / 1M
- Output RUB / 1M: 76.09
Z.ai: GLM 4.7
glm-4.7GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...
- Provider
- Z.ai
- Context
- 204.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 50.73
- Output RUB / 1M
- Output RUB / 1M: 186
Z.ai: GLM 5
glm-5GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
- Provider
- Z.ai
- Context
- 204.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 84.54
- Output RUB / 1M
- Output RUB / 1M: 270.54
Z.ai: GLM 5 Turbo
glm-5-turboGLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...
- Provider
- Z.ai
- Context
- 202.75K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 101.45
- Output RUB / 1M
- Output RUB / 1M: 338.18
Z.ai: GLM 5V Turbo
glm-5v-turboGLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
- Provider
- Z.ai
- Context
- 202.75K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 101.45
- Output RUB / 1M
- Output RUB / 1M: 338.18
Claude Opus 5
claude-opus-5Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 422.72
- Output RUB / 1M
- Output RUB / 1M: 2,113.62
DeepSeek: DeepSeek V4 Flash 0731
deepseek-v4-flash-0731DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.
- Provider
- DeepSeek
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 11.84
- Output RUB / 1M
- Output RUB / 1M: 23.67
Anthropic: Claude Sonnet 5
claude-sonnet-5Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 169.09
- Output RUB / 1M
- Output RUB / 1M: 845.45
MoonshotAI: Kimi K3
kimi-k3Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
- Provider
- Moonshot AI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 253.63
- Output RUB / 1M
- Output RUB / 1M: 1,268.17
xAI: Grok 4.5
grok-4.5Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
- Provider
- xAI
- Context
- 500K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 169.09
- Output RUB / 1M
- Output RUB / 1M: 507.27
OpenAI: GPT-5.6 Luna
gpt-5.6-lunaGPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 16.91
- Output RUB / 1M
- Output RUB / 1M: 101.45
OpenAI: GPT-5.6 Sol
gpt-5.6-solGPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 422.72
- Output RUB / 1M
- Output RUB / 1M: 2,536.35
OpenAI: GPT-5.6 Terra
gpt-5.6-terraGPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 211.36
- Output RUB / 1M
- Output RUB / 1M: 1,268.17
Google: Gemini 2.5 Flash
gemini-2.5-flashGemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 25.36
- Output RUB / 1M
- Output RUB / 1M: 211.36
Google: Gemini 2.5 Flash Lite
gemini-2.5-flash-liteGemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 8.45
- Output RUB / 1M
- Output RUB / 1M: 33.82
MoonshotAI: Kimi K2.7 Code
kimi-k2.7-codeMoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
- Provider
- Moonshot AI
- Context
- 262.14K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 80.32
- Output RUB / 1M
- Output RUB / 1M: 338.18
Z.ai: GLM 5.2
glm-5.2GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
- Provider
- Z.ai
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 118.36
- Output RUB / 1M
- Output RUB / 1M: 372
Anthropic: Claude Fable 5
claude-fable-5Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 845.45
- Output RUB / 1M
- Output RUB / 1M: 4,227.25
Anthropic: Claude Opus 4.8
claude-opus-4.8Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 422.72
- Output RUB / 1M
- Output RUB / 1M: 2,113.62
Anthropic: Claude Opus 4.6
claude-opus-4.6Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 422.72
- Output RUB / 1M
- Output RUB / 1M: 2,113.62
Anthropic: Claude Opus 4.7
claude-opus-4.7Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 422.72
- Output RUB / 1M
- Output RUB / 1M: 2,113.62
Anthropic: Claude Sonnet 4.6
claude-sonnet-4.6Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
- Provider
- Anthropic
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 253.63
- Output RUB / 1M
- Output RUB / 1M: 1,268.17
DeepSeek: DeepSeek V4 Flash 0423
deepseek-v4-flashDeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
- Provider
- DeepSeek
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 11.84
- Output RUB / 1M
- Output RUB / 1M: 23.67
DeepSeek: DeepSeek V4 Pro
deepseek-v4-proDeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
- Provider
- DeepSeek
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 36.78
- Output RUB / 1M
- Output RUB / 1M: 73.55
Google: Gemini 3 Flash Preview
gemini-3-flash-previewGemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 42.27
- Output RUB / 1M
- Output RUB / 1M: 253.63
Google: Gemini 3.1 Flash Lite
gemini-3.1-flash-liteGemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 21.14
- Output RUB / 1M
- Output RUB / 1M: 126.82
Google: Gemini 3.1 Pro Preview
gemini-3.1-pro-previewGemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 169.09
- Output RUB / 1M
- Output RUB / 1M: 1,014.54
Google: Gemini 3.5 Flash
gemini-3.5-flashGemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 126.82
- Output RUB / 1M
- Output RUB / 1M: 760.9
MiniMax: MiniMax M2.7
minimax-m2.7MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
- Provider
- MiniMax
- Context
- 204.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 25.36
- Output RUB / 1M
- Output RUB / 1M: 101.45
MiniMax: MiniMax M3
minimax-m3MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
- Provider
- MiniMax
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 25.36
- Output RUB / 1M
- Output RUB / 1M: 101.45
MoonshotAI: Kimi K2.6
kimi-k2.6Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
- Provider
- Moonshot AI
- Context
- 262.14K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 80.32
- Output RUB / 1M
- Output RUB / 1M: 338.18
OpenAI: GPT-5.4
gpt-5.4GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 211.36
- Output RUB / 1M
- Output RUB / 1M: 1,268.17
OpenAI: GPT-5.4 Mini
gpt-5.4-miniGPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
- Provider
- OpenAI
- Context
- 400K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 63.41
- Output RUB / 1M
- Output RUB / 1M: 380.45
OpenAI: GPT-5.4 Nano
gpt-5.4-nanoGPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
- Provider
- OpenAI
- Context
- 400K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 16.91
- Output RUB / 1M
- Output RUB / 1M: 105.68
OpenAI: GPT-5.5
gpt-5.5GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
- Provider
- OpenAI
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 422.72
- Output RUB / 1M
- Output RUB / 1M: 2,536.35
Qwen: Qwen3.7 Max
qwen3.7-maxQwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...
- Provider
- Qwen
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 101.45
- Output RUB / 1M
- Output RUB / 1M: 507.27
xAI: Grok 4.1 Fast
grok-4.1-fastGrok 4.1 Fast is xAI's best agentic tool calling model that shines in real-world use cases like customer support and deep research. 2M context window. Reasoning can be enabled/disabled using...
- Provider
- xAI
- Context
- 128K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 105.68
- Output RUB / 1M
- Output RUB / 1M: 211.36
xAI: Grok 4.3
grok-4.3Grok 4.3 is a reasoning model from xAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
- Provider
- xAI
- Context
- 1M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 105.68
- Output RUB / 1M
- Output RUB / 1M: 211.36
xAI: Grok Build 0.1
grok-build-0.1Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...
- Provider
- xAI
- Context
- 256K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 105.68
- Output RUB / 1M
- Output RUB / 1M: 211.36
Xiaomi: MiMo-V2.5
mimo-v2.5MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...
- Provider
- Xiaomi
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 11.84
- Output RUB / 1M
- Output RUB / 1M: 23.67
Xiaomi: MiMo-V2.5-Pro
mimo-v2.5-proMiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
- Provider
- Xiaomi
- Context
- 1.05M
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 36.78
- Output RUB / 1M
- Output RUB / 1M: 73.55
Z.ai: GLM 5.1
glm-5.1GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
- Provider
- Z.ai
- Context
- 204.8K
- Capabilities
- Text
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 118.36
- Output RUB / 1M
- Output RUB / 1M: 372
From catalog to request
Move from a public model ID to a visible request path
Create a key in the cabinet, point a compatible client to the endpoint, then choose a public model from this catalog. Usage and published pricing stay visible alongside the work.
- 01Create a platform API key.
- 02Set the compatible endpoint.
- 03Choose a public model ID.
- 04Review usage and pricing.
- Data source
- Live public catalog
- Catalog updated
Choosing and using models
Use the live catalog for availability and request capabilities.
Check its unavailable reason in the current catalog. Personal and shared workspaces can follow different access rules.
Use the public catalog or model discovery. Model IDs and availability can change, so avoid relying on an old list.
Review the selected model’s published modalities, limits, and supported request parameters before sending a request.
Start with the required modality and capability, then compare current context, limits, availability, and price.