MiniMax Music 2.5+: Unlock Your Exclusive "Castle in the Sky"
Published · Mar 5 · Thu Source · MiniMax稀宇科技 (CN)

MiniMax Music 2.5+: Unlock Your Exclusive "Castle in the Sky"

MiniMax releases the Music 2.5+ version, adding pure music creation capabilities. The model supports diverse styles such as classical orchestral, minimalism, modern electronic, and ambient sounds. It can generate complete works ranging from zero-instrument natural sounds to multi-track instrumental arrangements, suitable for meditation, sleep aid, advertising, game scoring, and film scoring scenarios.

KeywordsMiniMaxMusicUnlockYourExclusiveCastleSkyThe

Today, we officially release the latest generation music model — MiniMax Music 2.0. In this version, the model's understanding and expression of music have achieved a true leap: whether it is the delicate emotions of vocals or the dynamic tension of instruments, they can all be accurately captured and restored.

It understands rhythm, and it understands emotion. In the interweaving of vocals and instruments, it becomes that "singing producer".

From now on, expressing yourself through music is no longer the privilege of a few, but a joy that everyone can obtain.

Let inspiration become flowing musical phrases, Feel the rhythm, let the music belong to you.

1. Vocals are agile, mastering different singing styles

You don't need to accept vocal training, you can also use your favorite voice, apply techniques and styles, and sing the melody in your heart.

In terms of vocal texture, Music 2.0 timbre is infinitely close to real human voices. In addition, the model is like a seasoned "singer", capable of mastering various singing styles and emotional styles; the appropriate handling of musical phrases, rhythm, and breathing clearly possesses a "singing quotient" comparable to real humans.

Urban: Cooler, more Chill

The model supports precise control of vocal timbre, and through Prompts, while maintaining core timbre consistency, it allows the same voice to switch between different singing styles, achieving one voice with thousand changes, AI can also transform into a "versatile singer".

One Person Multiple Singing: Jump Blues-Rock-Electronic

(The same female voice can freely switch between different styles such as Jump Blues, Rock, Electronic)

In addition to easily mastering common singing styles such as pop, jazz, Blues, rock, folk, etc., the model also supports styles such as male-female duets, a cappella, etc.

Jazz Duet: Different frequencies but tacit understanding

The connection effect of different male and female lead singers can achieve a sense of dialogue, dynamic duet with changes in strength and weakness

A Cappella: Worried? Only vocals left is fresher

Without accompaniment, it can also present rich melodies

2. Catchy melodies, precise instrument control

You don't need to be an arranger, you can also build a complete movement belonging to yourself like an arranger.

Music 2.0 inherits the advantage of the previous generation model's complete structure, able to generate songs with clear logic and complete structure including verses, choruses, bridges, etc. Single song duration can reach 5 min. In addition, the melodies generated by the new model are easier to remember, able to quickly catch the ear.

Pop: Just want this one that can be hummed at first listen

Hook part melodies are easy to remember, more possessing the melodic habits of real human creation

In different style expressions, the model can follow precise instruction control, independently control and adjust various instruments in the accompaniment, achieving rich layers and natural rhythm arrangement.

Jazz: As if coming to Bluenote

Saxophone, trombone, trumpet, jazz drum kit, piano appear in order, as if being at a master jazz live scene

3. Professional-grade sound quality experience

The new model also brings a comprehensive sound quality upgrade, whether it is vocal track texture or instrument spatial sense are more enhanced, bringing you an immersive auditory experience.

Disco: Power on, restart the golden age

Retro disco dance floor, energetic vocal performance and 80s classic instrument performance, take you back to that dancing golden age.

One More Thing

When we were testing Music 2.0, we surprisingly found that we can also accurately describe vocal emotions, voice scenes, and other factors through Prompts, to generate film-level scoring monologues. Layer-by-layer progressive emotions and music layering, making people feel as if they "heard" the color of the picture.

Surprised, we realized that this benefits from the model's accurate understanding of semantics, and precise control of vocal expressiveness — this is exactly the perfect combination of the model's semantic understanding and acoustic expressiveness, letting sound possess variable emotional contours.

Film Scoring Monologue: When the Tide is Calm

Scoring Poetry Recitation: Only the Curtain Opens.

This page provides an editorial summary based on publicly available information. It is not a republished article. Use the source link below for the original report.