Neural Analog Restoration Models

See how each audio restoration model in Neural Analog works: its architecture, source paper and code, and the settings you can adjust.

MP3 Music Restoration (Apollo)

music

This model was trained to regenerate high-quality music from its low-quality MP3 version. It fills gaps in the spectrogram and restores missing frequencies above 16 kHz. Perfect for your old 128 kbps MP3s downloaded from the internet and lost media.

AI Cover - Stable Audio 3 (Stable Audio 3 Medium)

ai_remix

This instrumental music model generates an AI cover using the input audio as a reference. The settings let you make the result more or less similar to the original. Does not work with singing! Remix your instrumentals, turn noise into musical sounds, and repair low frequencies.

AI Cover - ACEStep 1.5 XL (ACE-Step 1.5 XL)

ai_remixUp to 2x

This instrumental music model generates an AI cover using the input audio as a reference. Generations often sound very close to the reference unless you greatly reduce the original's influence. Recommended for improving bass tracks, reconstructing missing notes, and correcting specific artifacts.

High-Frequency Restoration - AudioSR (AudioSR)

music2x

This model regenerates high frequencies above the selected cutoff and keeps the original audio below it. The lower the cutoff, the more of the sound it reconstructs. Great for historical recordings, poor-quality recordings, or recordings with unwanted high-frequency noise. It does not change the bass (use an AI cover model instead).

High-Frequency Restoration - UniverSR (UniverSR)

musicUp to 9x

An upgraded version of AudioSR for music, voice, and sound effects. UniverSR reconstructs missing or poor-sounding high frequencies to add brightness. Try it on dull voices, old recordings (such as cassettes), or AI music from SUNO. It does not change the low frequencies (use an AI cover model instead).

Speech Enhancement - NovaSR Upscaler (NovaSR)

voice

This upscaling model was trained on English speech. Use it to remove the walkie-talkie sound from your podcasts, voice-overs, or isolated vocal tracks.

High-Frequency Restoration - FlashSR (FlashSR)

music2x

Similar to AudioSR, this model reconstructs high frequencies more quickly. Handy for quickly upscaling audio from 44.1 kHz to 48 kHz. It does not change the low frequencies (use an AI cover model instead).

Correct Tempo Variations (Constant BPM) (Beat This!)

music

SUNO tracks often have an unstable tempo, subtly speeding up or slowing down throughout the music. This model detects the tempo and corrects these variations. Ideal before importing a track into a DAW or playing it in a DJ set. Also useful for live recordings. Works best when there are drums: EDM, rock, pop, techno, house, etc.

Neural Reconstruction (DACVAE)

This model passes the sound through a high-quality encoder-decoder model. The effect is very subtle: it can help correct certain artifacts (especially hi-hats in your drum'n'bass tracks).

Correct Clipping (Declip) (Declipping restoration)

Heavily compressed or limited audio sometimes has a waveform that is flat in places (clipping). This digital artifact, caused during recording or mixing, creates unpleasant noise. Use this model to correct it.

Noise Reduction (Denoising separator)

Up to 6x

Remove hiss, hum, background noise, rumble, or animal and crowd sounds. Choose the variant based on the noise present and the sound you want to keep.

Remove Reverb (Dereverb separator)

Get a dry sound from a recording drenched in reverb or echo, such as one recorded in a cellar or a church. Also useful for removing the slight echo of a poorly soundproofed room.

Crowd Noise Reduction (Crowd-noise separator)

Up to 6x

Improve your live recordings! Keep the performance and remove audience noise: coughing, whistling, shouting, applause, etc. Use the Before/After slider to control the ambience level.

Keep Only Center Mono (MDX23C CenterWide)

This model keeps the mono center of a track and removes the content spread across the sides. Handy for isolating a bass, kick drum, or podcast voice. You can also use it to remove phaser, chorus, or flanger from an instrument.

Speech Enhancement - RE-USE Upscaler (RE-USE)

voice6x

This model dramatically improves speech quality in multiple languages. It recovers high-frequency detail and reduces noise, reverb, and audio defects. For the most difficult restorations of podcasts, voice messages, or singing. Hallucinations are common: use UniverSR or blend with the original audio to improve the result.

MP3 Singing Restoration (Apollo Voice)

voice

An MP3 restoration model specialized in singing. Handy for improving a cappella vocals downloaded as MP3s from the internet or correcting certain compression artifacts.