Skip to main content
npm version Node 16+

Installation

Quick Start

Authentication


Stem Separation

Extract individual audio components from mixed recordings.

Available Modes

Examples


Transcription

Convert audio to text with speaker diarization.

Voice Cloning & TTS

Create custom voices and generate speech.

Music Generation (AudioMusic V2)

Generate AI music with AudioMusic V2 — text-to-music, covers, style transfer, audio analysis, and more.

Audiobook

Turn a manuscript into a finished, ACX-compliant audiobook. Use produce() for the whole pipeline in one call, or drive each step. See the Audiobook API reference for the full surface.
Generate an original book from a prompt instead of uploading:

Noise Reduction

Clean up audio by removing background noise.

Speaker Separation

Identify and separate multiple speakers.

API Wallet

Check balance and manage billing.

Error Handling


TypeScript Support

Full TypeScript support with exported types:

Environment Variables


Resources

npm Package

View on npm

GitHub

Source code

Get API Key

Generate your API key

API Reference

Raw API documentation