CoolFace
Modelpublic

DisturbingTheField/ACE-Step-v1.5-raspy-vocal-and-instrumental-5-LoRAs

sourceHugging Facemitupdated 7mo agoView on Hugging Face
3likes70downloads
Model Card

Date: 2/20/2026

Rename these LoRA files to only adaptermodel.safetensors when using. The same adapterconfig.json can be used for all LoRA files.<br>

LoRA file: malevocalsadapter_model.safetensors

Sounds: Me singing when I had a cold and a very hoarse voice<br> Training: 7 self recorded wav files, 58 MB, 1200 epochs<br>

LoRA file: instrumentaladaptermodel.safetensors

Sounds: Instrumental songs made by myself. Include electric guitar, distorted guitar, bass, drums, piano, synth etc.<br> Training: 21 files, 766 MB, 800 epochs<br>

Both trained with acestep-v15-base and acestep-5Hz-lm-4B.<br> Dataset, preprocessed tensors with both, and training with only acestep-v15-base.<br> Both trained on a laptop, Nvidia RTX 3080 Ti with 16 GB VRAM.<br> When making songs I believe I used acestep-v15-turbo with acestep-5Hz-lm-1.7B.<br>

Merged LoRA files

Use the Python script MERGE-LORA.py.txt, read some more info in the script. I made these 3:<br> Strength: voc 0.8 inst 0.8 (adaptermodel.safetensors)<br> Strength: voc 0.6 inst 1.4 (voc06inst14__adaptermodel.safetensors)<br> Strength: voc 1.4 inst 0.6 (voc14inst06_adaptermodel.safetensors)<br> You can use the same adapter_config.json with all LoRAs.<br>

Description

There are 2 LoRA adapters here. The first is pure vocal, trained on my own voice when I had a cold and a very hoarse voice<br> with uncontrollable pitch. The second is pure instrumental, trained on 21 instrumental tracks that I made myself.<br> The music styles vary: rock, acoustic guitar, distorted guitar, ambient, etc.<br> Instruments include clean guitar, distorted guitar, bass, drums, piano, synth, and more.<br>

Use these two as a kind of filter — they work best with low LoRA Scale values between 0.2 and 0.7.<br> Also test them with “Think” both on and off, and try varying the LM Temperature and LM CFG Scale values.<br>

There is also a Python script included called MERGE-LORA.py.txt, which can be used to combine two LoRA adapters.<br> You set the strength for each of them. The two LoRA adapters you combine must contain the same layers — typically<br> meaning they were created from the same base model. The script checks this before generating the merged version.<br> See the script for more information if you want to use it.<br>

Additionally, there are 3 more LoRA adapters included, which are simply three different combinations of the two main LoRA adapters.<br> These provide both vocal and instrumental effects.<br>

I’ve included some demo MP3 files as well. These are not necessarily polished songs, but rather examples so you can<br> hear the kinds of sounds/effects you can achieve. There are a lot of possibilities here, so I can’t<br> test everything — you’ll just have to experiment yourself.<br>

Again, think of these two main LoRA adapters more as filters that adjust aspects of the vocals and instruments.<br> The vocal LoRA adapter will affect both the vocals, the instruments, and the overall song.<br> Songs become calmer, almost sadder, when used with stronger values.<br>

You can store everything in one folder, but make sure you have a JSON file named “adapterconfig.json” and a<br> LoRA file named “adaptermodel.safetensors”. They must have these names in order to be loaded properly.<br> So rename the LoRA files according to which one you want to load.<br>

3 captions for testing:

  • —90s dance feel-good vibe:<br> Upbeat, feel-good dance music inspired by the smooth European club sound of the late 1990s. Male singer. The groove is driven by steady four-on-the-floor<br> drums and a warm, rounded bassline that locks tightly with the rhythm. Clean electric guitar adds light, funky chord stabs and rhythmic accents,<br> giving the track a fresh, organic touch against a backdrop of shimmering synth pads, bright keyboard hooks, and subtle electronic textures.<br> The production is polished and uplifting, blending disco-influenced grooves with pop sensibility and dancefloor energy.<br> The overall vibe is sunny, nostalgic, and effortlessly catchy—music designed to feel carefree, stylish, and movement-driven.<br> Male singer. Funky clean guitar chords. Synth pads soft in the background.<br>
  • —Rock:<br> Melodic British-style rock with a polished yet organic sound, driven by clean, articulate lead guitar lines and a steady, radio-friendly mid-tempo groove.<br> The arrangement features tight rhythm guitar, supportive bass lines, crisp drums, and subtle dynamic builds.<br> The vocal delivery comes from a male singer with a dark, slightly raspy voice, combining understated intensity with a conversational, storytelling approach.<br> He has a deep, dark baritone with a gravelly, rough-edged texture. There’s a raw, raspy quality to his voice.<br> The overall feel is catchy and accessible, blending classic rock sensibility with pop structure and memorable hooks. High-fidelity, studio-polished.<br>
  • —Electronica, ambient, dance etc:<br> Music combines dark, moody atmospheres with melodic electronic pop. Male voice. He has a deep, dark baritone with a gravelly, rough-edged texture.<br> There’s a raw, raspy quality to his voice. Synthesizers dominate the sound, layering rich textures, pulsating basslines, and hypnotic arpeggios.<br> The vocals are expressive, sometimes melancholic or brooding, and often carry a sense of intimacy or vulnerability. Guitar parts occasionally cut through,<br> adding grit or accentuating climactic moments. The song explore themes of love, desire, pain, and introspection, often tinged with darkness.<br> Effects like reverb, delay, and subtle distortion enhance the emotional atmosphere, giving the music a cinematic quality. Syth adds melody hooks.<br> The overall sound is both danceable and haunting, blending electronic sophistication with raw human emotion. The lyrics are almost spoken sometimes,<br> with deep dark voice.<br>

Lyrics for testing:

[Intro]<br>

[Verse 1]<br> Sun is rising slowly, light upon the floor<br> Coffee’s on the table, no alarms, no chores<br> Kids are laughing softly, in the morning glow<br> Nothing on the schedule, nowhere we need to go<br>

[Chorus]<br> Oh, Sundays feel like heaven, hearts are running free<br> Time to love, time to linger, just my family and me<br> The world can wait a little, we’ve got our own parade<br> Oh, Sundays feel like magic, every moment we have made<br>

[Verse 2]<br> Stories in the kitchen, songs drift through the air<br> Moments like these linger, precious and rare<br> The clock is just a number, the hours drift away<br> Wrapped up in each other, there’s nothing left to say<br>

[Chorus]<br> Oh, Sundays feel like heaven, hearts are running free<br> Time to love, time to linger, just my family and me<br> The world can wait a little, we’ve got our own parade<br> Oh, Sundays feel like magic, every moment we have made<br>

[Outro]<br>

<br>

Model Card for Model ID:

5 LoRA adapters with raspy male vocals, instrumental and mix of both for ACE-Step/Ace-Step1.5

Model Description:<br> <br> If yu use "startgradioui.bat" then edit the file:<br><br> set INITSERVICE=--initservice false<br> So you get access to the LoRA loading part in the Web UI.<br>

Model Sources:<br> https://huggingface.co/ACE-Step/Ace-Step1.5<br>

Uses:<br> 5 LoRA adapters with raspy male vocals, instrumental and mix of both for ACE-Step/Ace-Step1.5<br>

Training Details:

LoRA file: malevocalsadapter_model.safetensors<br> Sounds: My singing when I had a cold and a very hoarse voice<br> Training: 7 self recorded wav files, 58 MB, 1200 epochs<br>

LoRA file: instrumentaladaptermodel.safetensors<br> Sounds: Instrumental songs made by myself. Include electric guitar, distorted guitar, bass, drums, piano, synth etc.<br> Training: 21 files, 766 MB, 800 epochs<br>

Both trained with acestep-v15-base and acestep-5Hz-lm-4B.<br> Dataset, preprocessed tensors with both, and training with only acestep-v15-base.<br> Both trained on a laptop, Nvidia RTX 3080 Ti with 16 GB VRAM.<br>