Model catalog
One endpoint, every model. Compare context windows, capabilities and real per-token pricing side by side, then swap the model string in your request — nothing else changes.
- 351
- Models
- 35
- Providers
- 11
- Free to try
- $0.0001
- Cheapest input / 1M
8 of 351 models
- Video
Kling v3.0 Motion Control
klingai/kling-v3.0-motion-control
Kling 3.0 delivers a major leap in character fidelity for motion-driven generation, with stable facial features across multi-angle and long-duration motion, accurate complex emotions from multi-image face references, identity preservation through partial occlusions (hats, hands, fans), and steady clarity as the camera zooms, pans, or tracks.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Video generation
- Video
Kling v3.0 Image-to-Video
klingai/kling-v3.0-i2v
Build upon an All-in-One product framework, the Kling 3.0 model series supports full multimodal input and output spanning text, images, audio, and video, bringing the understanding, generation, and editing of video together in one streamlined AI workflow. The models integrate multiple tasks, including text-to-video, image-to-video, reference-to-video, and in-video editing, into a single, native multimodal architecture, enabling the models to follow complex narrative logic, deliver precise shot control, and maintain strong prompt adherence.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Video
Kling v3.0 Text-to-Video
klingai/kling-v3.0-t2v
Build upon an All-in-One product framework, the Kling 3.0 model series supports full multimodal input and output spanning text, images, audio, and video, bringing the understanding, generation, and editing of video together in one streamlined AI workflow. The models integrate multiple tasks, including text-to-video, image-to-video, reference-to-video, and in-video editing, into a single, native multimodal architecture, enabling the models to follow complex narrative logic, deliver precise shot control, and maintain strong prompt adherence.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Video
Kling v2.6 Motion Control
klingai/kling-v2.6-motion-control
Kling 2.6 introduces a groundbreaking "Native Audio" capability, enabling the generation of complete videos in a single go, including natural voice, action sound effects, and environmental ambient sounds, providing an immersive "what you see if what you hear" experience.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Video
Kling v2.6 Image-to-Video
klingai/kling-v2.6-i2v
Kling 2.6 introduces a groundbreaking "Native Audio" capability, enabling the generation of complete videos in a single go, including natural voice, action sound effects, and environmental ambient sounds, providing an immersive "what you see if what you hear" experience.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Video
Kling v2.6 Text-to-Video
klingai/kling-v2.6-t2v
Kling 2.6 introduces a groundbreaking "Native Audio" capability, enabling the generation of complete videos in a single go, including natural voice, action sound effects, and environmental ambient sounds, providing an immersive "what you see if what you hear" experience.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Video
Kling v2.5 Turbo Image-to-Video
klingai/kling-v2.5-turbo-i2v
Kling 2.5 Turbo is a major update to the AI video generation model focused on significantly improving speed, video quality, temporal stability, and creative control for creators, making professional-grade AI-generated video faster, more coherent, and easier to direct from text prompts.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Video
Kling v2.5 Turbo Text-to-Video
klingai/kling-v2.5-turbo-t2v
Kling 2.5 Turbo is a major update to the AI video generation model focused on significantly improving speed, video quality, temporal stability, and creative control for creators, making professional-grade AI-generated video faster, more coherent, and easier to direct from text prompts.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
