Grom Zvuk

The first music model in the Grom lineup.

Launch offer: 25% off every track

Listen

How Grom Zvuk sounds

Twelve tracks the model wrote end to end: from jazz and pop to nu metal, punk, reggae and hardcore. Hit a record — it plays right here.

Technology

How a track is made

First the model floods the whole length of the future song with noise and marks out where the verses go, where the chorus lands and where the bridge sits. Then, over several passes, it turns that noise into sound worth listening to.

IntroVerseChorusVerseChorusBridgeOutro

There is no partial result: the track arrives whole, once it is finished.

Architecture

How the pipeline works

Two experts under shared attention: the first writes music as symbols, the second unfolds those symbols into sound.

CREATEstyle and lyricsCOVERmelody lifted from the sourceEDITedits to a finished score
  1. yPromptstyle · lyrics
  2. sScoreABC notation
  3. cSemantics24 Hz
  4. zAcoustic latents24 Hz × 64
  5. 48 kHzFinished recordVAE decoder · stereo
AR · CAUSALPlannercausal token prediction · y → s → c

Walks through the music step by step: decides where the melody leads and where the harmony turns.

NAR · BIDIRECTIONALSynthesiserflow-velocity prediction · z_t → z

Works across the full length at once: refines the sound as a whole, not piece by piece.

HYBRID SELF-ATTENTIONshared memory of the track

The diagram is simplified: it explains the principle rather than repeating the model's internals.

Comparison

Where Grom Zvuk sits among music models

Two axes at once: how the track sounds and how closely the model sings the given lyrics. On lyric accuracy Grom Zvuk is in the top three; on sound it holds its place among the closed services.

02550751000255075100Sound quality →Lyric accuracy →Mureka 9Suno v5Suno v4.5Grom ZvukSuno v5.5Suno v6MiniMax Music 3ACE-Step 1.5LeVo 2MuseDiffRhythm 2SongBloom

sound / lyrics

  • Suno v588 / 90
  • Mureka 990 / 78
  • Grom Zvuk83 / 85
  • Suno v680 / 85
  • Suno v5.582 / 82
  • Suno v4.584 / 78
  • ACE-Step 1.567 / 74
  • MiniMax Music 370 / 64
  • Muse59 / 62
  • LeVo 265 / 42
  • DiffRhythm 237 / 53
  • SongBloom9 / 9

Blind comparison across 150 prompts: identical briefs, finished tracks rated on sound and on how well they follow the given lyrics, with model names hidden from the raters. Values are normalised to a 0–100 scale. Hover a dot to see its name.

Under the hood

What sits behind a generation

48 kHz, stereo

Not a compressed draft but a studio sample rate and a real stereo image.

Tracks up to five minutes

A full song with verses, choruses and an ending, not a fifteen-second loop.

A single pass

Every part of the song is computed together rather than in turn — hence the unified sound.

Our own hardware

The model runs on our own fleet of accelerators — we scale capacity to the client's volumes.

Languages

Sings in any language

From English to Chinese and Japanese. The model reads the lyrics, the stress and the phrasing — the vocal never turns into accented mush.

English中文日本語EspañolDeutschFrançais한국어ItalianoPortuguêsTürkçeҚазақшаالعربيةहिन्दीРусский

The list is open-ended: these are not all the languages the model sings in.

What it does

Six ways to get a track

The idea of a track lives in the score, so a song can be edited in parts: swap the style, an instrument or the voice without rewriting it.

From scratch

Describe the style and the mood, hand over your lines — the model writes the melody, builds the arrangement and sings it.

Cover

Upload a song and name a new sound: the melody stays recognisable, the arrangement changes completely.

Your own lyrics

Upload a finished song and drop in your own lines: the music stays recognisable, the words are yours.

Style swap

Same track, same lyrics, different genre: a ballad turns into punk, hip-hop into jazz.

Vocals off

Ask for a cover with no lyrics and leave the style empty — the same arrangement remains, just without the voice.

Instrumental

Ambient, lo-fi, a theme for a video — music without words from scratch, from a single description.

Need your own stream of music?

Get in touch — we'll go over volumes, limits and pricing ahead of launch. The model runs on our own hardware, so load and timelines are fixed by contract and capacity scales to your volumes.

Questions about Grom Zvuk

A finished track: vocals, melody, arrangement and mix. No editing pass is required — the file can go straight into a video or be published as is.

Try Grom Zvuk

One model for songs, covers and instrumentals. The balance is shared with every other GPTunneL model, and every track is 25% off at launch.