API Tools
9 endpoints across 2 categories
AI Audio
— 6 tools/v1/audio/classify
Detect audio events, language, and acoustic properties using GPU-accelerated classification.
View details → 1 credit/v1/audio/convert
GPU-accelerated audio format conversion between WAV, MP3, FLAC, OGG, and AAC.
View details → 1 credit/v1/audio/enhance
Reduce background noise and normalize audio levels using GPU acceleration.
View details → 1 credit/v1/audio/separate
Separate audio into individual stems (vocals, drums, bass, other) using Demucs on GPU.
View details → 1 credit/v1/audio/synthesize
Convert text to natural-sounding speech using Kokoro TTS on GPU — multiple voices, adjustable speed.
View details → 1 credit/v1/audio/transcribe
Transcribe speech to text using Whisper on GPU — 99+ languages, word timestamps, subtitle output.
View details → 1 creditEditing
— 3 tools/v1/audio/merge
Concatenate multiple audio files into one continuous track using ffmpeg.
View details → 1 credit/v1/audio/speed
Change audio playback speed with optional pitch correction using ffmpeg.
View details → 1 credit/v1/audio/trim
Cut an audio segment by start and end time using ffmpeg.
View details → 1 credit