Model Catalog

Explore 240 AI Models

Browse, compare, and integrate the best AI models for video, image, music, audio, and text generation — all through one unified API.

All models available · Real-time pricing

REST API
# Available Endpoints

POST   /v1/chat/completions         # LLM
POST   /v1/images/generations       # Image
POST   /api/v1/kling/text_to_video  # Video
POST   /v1/audio/speech             # TTS
POST   /v1/audio/transcriptions     # STT
POST   /v1/suno/text_to_music       # Music

Leading AI models, one API

ElevenLabsKlingGPT Image 2Veo 3.1HailuoLumaPixVerseRecraftFish Audio
14 models
Provider
Video

Kling

Kuaishou

Kling video generation for cinematic text-to-video, image-to-video, motion control, and avatar clips.

from $0.070 / secondView
Video

Seedance

Bytedance

Seedance 2.5 delivers precise choreography and camera control for text- and image-driven video.

from $0.090 / secondView
Video

Veo 3.1

Google

Google Veo 3.1 for native audio-video generation, extension, and upscaling with strong prompt adherence.

from $0.150 / secondView
Video

Gemini Omni

Google

Gemini Omni API access for voice, character, and multimodal video resources in agent media workflows.

from $0.0000 / callView
Video

Hailuo

MiniMax

MiniMax Hailuo video models for expressive character motion and physics-aware rendering.

from $0.060 / secondView
Video

HappyHorse

Alibaba

HappyHorse video models optimised for fast draft iterations and edit-in-place workflows.

from $0.030 / secondView
Video

InfiniteTalk

MeiGen-AI

InfiniteTalk turns a single portrait plus audio into a talking-head video with lip sync.

from $0.180 / callView
Video

Lip Sync

Bytedance

Volcengine lip sync re-dubs an existing video to new audio while preserving the original performance.

from $0.120 / callView
Video

Luma

Luma

Luma Ray models for smooth motion synthesis, keyframe control, and video modification.

from $0.055 / secondView
Video

MiniMax H3

MiniMax

MiniMax H3 for high-fidelity long-form generation with multi-shot consistency.

from $0.080 / secondView
Video

OmniHuman

Bytedance

OmniHuman audio-to-video avatars with subject detection and human identification helpers.

from $0.200 / callView
Video

PixVerse

PixVerse

PixVerse for stylised video generation, transitions, extension, and clip editing.

from $0.045 / secondView
Video

Runway

Runway

Runway Gen-4 and Aleph for text-to-video generation and instruction-based video editing.

from $0.120 / secondView
Video

Wan Video

Alibaba

Alibaba Wan video models for text-to-video, image-to-video, speech-driven avatars, and editing.

from $0.035 / secondView

Not sure which model to pick?

Every model page lists the exact model IDs, per-unit pricing, and a runnable request you can copy.

POST/v1/chat/completions
POST/v1/images/generations
POST/v1/kling/text_to_video