/v1/models
Model catalog with the details needed to choose
Each row identifies a public model and its documented request capabilities.
7 models
Google: Gemini 2.5 Flash
gemini-2.5-flashGemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 25.36
- Output RUB / 1M
- Output RUB / 1M: 211.36
Google: Gemini 2.5 Flash Lite
gemini-2.5-flash-liteGemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 8.45
- Output RUB / 1M
- Output RUB / 1M: 33.82
Google: Gemini 3 Flash Preview
gemini-3-flash-previewGemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 42.27
- Output RUB / 1M
- Output RUB / 1M: 253.63
Google: Gemini 3.1 Flash Lite
gemini-3.1-flash-liteGemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 21.14
- Output RUB / 1M
- Output RUB / 1M: 126.82
Google: Gemini 3.1 Pro Preview
gemini-3.1-pro-previewGemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 169.09
- Output RUB / 1M
- Output RUB / 1M: 1,014.54
Google: Gemini 3.5 Flash
gemini-3.5-flashGemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
- Provider
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 126.82
- Output RUB / 1M
- Output RUB / 1M: 760.9
Xiaomi: MiMo-V2.5
mimo-v2.5MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...
- Provider
- Xiaomi
- Context
- 1.05M
- Capabilities
- Text · Transcription
- Catalog updated
- Input RUB / 1M
- Input RUB / 1M: 11.84
- Output RUB / 1M
- Output RUB / 1M: 23.67
From catalog to request
Move from a public model ID to a visible request path
Create a key in the cabinet, point a compatible client to the endpoint, then choose a public model from this catalog. Usage and published pricing stay visible alongside the work.
- 01Create a platform API key.
- 02Set the compatible endpoint.
- 03Choose a public model ID.
- 04Review usage and pricing.
- Data source
- Live public catalog
- Catalog updated
Choosing and using models
Use the live catalog for availability and request capabilities.
Check its unavailable reason in the current catalog. Personal and shared workspaces can follow different access rules.
Use the public catalog or model discovery. Model IDs and availability can change, so avoid relying on an old list.
Review the selected model’s published modalities, limits, and supported request parameters before sending a request.
Start with the required modality and capability, then compare current context, limits, availability, and price.