Explore 240 AI Models
Browse, compare, and integrate the best AI models for video, image, music, audio, and text generation — all through one unified API.
All models available · Real-time pricing
# Available Endpoints
POST /v1/chat/completions # LLM
POST /v1/images/generations # Image
POST /api/v1/kling/text_to_video # Video
POST /v1/audio/speech # TTS
POST /v1/audio/transcriptions # STT
POST /v1/suno/text_to_music # MusicLeading AI models, one API
ElevenLabs
ElevenLabs
ElevenLabs API access for voice synthesis, text-to-speech, sound effects, speech-to-text, and audio isolation.
Suno
Suno
Suno v5.5 generates full songs with vocals, lyrics, and stems — no official API available elsewhere.
Fish Audio
Fish Audio
Fish Audio API access for expressive multilingual and production-grade text-to-speech with managed MP3 or WAV output.
Gemini TTS
Gemini TTS API access for multi-speaker dialogue with configurable voices, accents, delivery styles, and pacing.
OpenAI TTS
OpenAI
OpenAI text-to-speech with low-latency streaming voices for real-time assistants.
Producer
Producer
Producer creates loopable beds, stems, and adaptive music cues for games and video.
Transcription
OpenAI
Whisper-based speech-to-text with timestamps, diarisation, and translation for long recordings.
Not sure which model to pick?
Every model page lists the exact model IDs, per-unit pricing, and a runnable request you can copy.