Back to models

MiniMax Music 3

AudioText to audio
minimax/music-3View on fal.ai

Parameters

Dictation ready.

Music description: style, mood, vocals, instrumentation and arrangement. For precise control use a Structured Caption with global metadata (genre, BPM, key, emotional progression), vocal details, and a section-by-section arrangement.

Upper bound on the generated audio length in seconds. The model may stop earlier; the actual duration is returned in the output.

Dictation ready.

The lyrics to sing. Structure tags such as [intro], [verse], [pre-chorus], [chorus], [post-chorus], [bridge], [instrumental], [solo] and [outro] must each be on their own line; text on the same line as a leading tag is dropped by the model's input contract.

Result

Run the model to see the result here.