/v1/models
Model catalog with the details needed to choose
Each row identifies a public model and its documented request capabilities.
22 models
ByteDance: Seedance 2.5
seedance-2.5Seedance 2.5 is a video generation model from ByteDance. It is suited for long-form storytelling, multimodal reference-based generation, video editing, and video extension. It supports first-frame and first-and-last-frame control, up...
- Provider
- ByteDance
- Capabilities
- Video
- Catalog updated
Black Forest Labs: FLUX.3 Video
flux-3-videoFLUX.3 Video is a video generation model from Black Forest Labs. It supports text-to-video, image-guided generation with opening and closing keyframes, and video continuation workflows, making it suited for controlled...
- Provider
- Black Forest Labs
- Capabilities
- Video
- Catalog updated
MiniMax: H3
hailuo-3MiniMax H3 is a lightweight, open-weights video generation model from MiniMax. It is designed for precise multimodal editing and controlled content generation, including instruction-guided edits, text and brand rendering, and...
- Provider
- MiniMax
- Capabilities
- Video
- Catalog updated
Runway: Aleph 2.0
aleph-2Runway Aleph 2.0 is an in-context video editing model from Runway. It applies text instructions and keyframe-guided edits across existing footage while preserving details that are not meant to change....
- Provider
- Runway
- Capabilities
- Video
- Catalog updated
Runway: Gen-4.5
gen-4.5Runway Gen-4.5 is a video generation model from Runway for text-to-video and image-to-video workflows. It is designed for cinematic scene creation with strong motion quality, visual fidelity, and prompt adherence....
- Provider
- Runway
- Capabilities
- Video
- Catalog updated
SpaceXAI: Grok Imagine Video 1.5
grok-imagine-video-1.5Grok Imagine Video 1.5 is a video generation model from SpaceXAI. It creates videos from text prompts, with an optional starting image to guide the scene. It can direct subject...
- Provider
- xAI
- Capabilities
- Video
- Catalog updated
Alibaba: HappyHorse 1.1
happyhorse-1.1HappyHorse 1.1 is a video generation model from Alibaba. It generates short videos from a text prompt, a single starting image, or a set of reference images, with output up...
- Provider
- Alibaba
- Capabilities
- Video
- Catalog updated
Alibaba: HappyHorse 1.0
happyhorse-1.0HappyHorse 1.0 is a video generation model from Alibaba. It generates short videos from a text prompt, a single starting image, or a set of reference images, with output up...
- Provider
- Alibaba
- Capabilities
- Video
- Catalog updated
SpaceXAI: Grok Imagine Video
grok-imagine-videoGrok Imagine Video is SpaceXAI's fast, text-, image-, and reference-conditioned video generation model. It produces short videos (1–15 seconds, 24 fps) at 480p or 720p across seven aspect ratios -...
- Provider
- xAI
- Capabilities
- Video
- Catalog updated
Kling: Video v3.0 Pro
kling-v3.0-proKling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...
- Provider
- Kuaishou
- Capabilities
- Video
- Catalog updated
Kling: Video v3.0 Standard
kling-v3.0-stdKling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...
- Provider
- Kuaishou
- Capabilities
- Video
- Catalog updated
Google: Veo 3.1 Fast
veo-3.1-fastGoogle's mid-tier video generation model balancing speed and quality. Veo 3.1 Fast generates high-quality video from text or image prompts with native synchronized audio, offering faster turnaround than Veo 3.1...
- Provider
- Capabilities
- Video
- Catalog updated
Google: Veo 3.1 Lite
veo-3.1-liteGoogle's most cost-effective video generation model, designed for high-volume applications and rapid iteration. Veo 3.1 Lite generates 720p and 1080p video from text or image prompts with native synchronized audio...
- Provider
- Capabilities
- Video
- Catalog updated
Kling: Video O1
kling-video-o1Kling Video O1 is a video generation model from Kuaishou. It supports text and image inputs with video output, enabling text-to-video and image-to-video workflows. It is suited for cinematic content...
- Provider
- Kuaishou
- Capabilities
- Video
- Catalog updated
MiniMax: Hailuo 2.3
hailuo-2.3Hailuo 2.3 is a video generation model from MiniMax. It accepts text prompts and reference images as input and generates video output, supporting both text-to-video and image-to-video workflows. It is...
- Provider
- MiniMax
- Capabilities
- Video
- Catalog updated
Alibaba: Wan 2.7
wan-2.7Wan 2.7 is a video generation model from Alibaba. It supports text-to-video, image-to-video with first and last frame control, and reference-to-video, where multiple reference images guide the style and content...
- Provider
- Alibaba
- Capabilities
- Video
- Catalog updated
ByteDance: Seedance 2.0
seedance-2.0Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...
- Provider
- ByteDance
- Capabilities
- Video
- Catalog updated
ByteDance: Seedance 2.0 Fast
seedance-2.0-fastSeedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...
- Provider
- ByteDance
- Capabilities
- Video
- Catalog updated
Alibaba: Wan 2.6
wan-2.6Alibaba's most advanced video generation model, supporting over 10 visual creation capabilities in a unified system. Wan 2.6 generates 1080p video at 24fps from text, images, reference videos, or audio,...
- Provider
- Alibaba
- Capabilities
- Video
- Catalog updated
ByteDance: Seedance 1.5 Pro
seedance-1-5-proByteDance's next-generation audio-visual generation model with a 4.5B parameter Dual-Branch Diffusion Transformer architecture. Seedance 1.5 Pro generates video and audio simultaneously in a single unified pass — eliminating the timing...
- Provider
- ByteDance
- Capabilities
- Video
- Catalog updated
OpenAI: Sora 2 Pro
sora-2-proOpenAI's flagship video generation model, delivering production-quality video with physics-accurate motion, synchronized audio, and world-state persistence across shots. Sora 2 Pro follows intricate multi-shot instructions while maintaining consistent spatial relationships...
- Provider
- OpenAI
- Capabilities
- Video
- Catalog updated
Google: Veo 3.1
veo-3.1Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —...
- Provider
- Capabilities
- Video
- Catalog updated
From catalog to request
Move from a public model ID to a visible request path
Create a key in the cabinet, point a compatible client to the endpoint, then choose a public model from this catalog. Usage and published pricing stay visible alongside the work.
- 01Create a platform API key.
- 02Set the compatible endpoint.
- 03Choose a public model ID.
- 04Review usage and pricing.
- Data source
- Live public catalog
- Catalog updated
Choosing and using models
Use the live catalog for availability and request capabilities.
Check its unavailable reason in the current catalog. Personal and shared workspaces can follow different access rules.
Use the public catalog or model discovery. Model IDs and availability can change, so avoid relying on an old list.
Review the selected model’s published modalities, limits, and supported request parameters before sending a request.
Start with the required modality and capability, then compare current context, limits, availability, and price.