AceStep 1.5 LoRA Trained on Modern Talking Albums
Train custom LoRA on Ace-step 1.5 with 1000 epochs, dynamic LR peaking at 3e-4, targeting loss ~0.084.
Use Ace-step 1.5 as base, train LoRA with 1000 epochs, dynamic LR, monitor loss, aim for ~0.084.
Summary
A user trained a custom LoRA on the Ace-step 1.5 base model using a dataset composed solely of the first two Modern Talking albums. The training ran for 1000 epochs, totaling 4000 steps, with a dynamic learning rate that peaked at 3e-4 and tapered to approximately 3e-6. The best loss achieved was 0.107 at epoch 879, with a final loss of 0.084. The resulting LoRA was used to generate a track titled "Midnight Phantom" at epoch 800, capturing an 80s synth‑pop vibe. The training logs also indicate a strict 10‑syllable iambic structure for lyric generation. The experiment demonstrates the feasibility of fine‑tuning music style with a relatively small dataset and moderate compute. The user shared the LoRA weights and training logs for community review.
This approach can be replicated for other musical styles by curating a focused dataset and adjusting the dynamic learning rate schedule. The final loss metrics provide a benchmark for evaluating similar LoRA projects.
Key changes
- Base model: Ace-step 1.5
- Max epochs: 1000 (4000 steps)
- Dynamic LR peak: 3e-4, drop to ~3e-6
- Best loss: 0.107 at epoch 879
- Final loss: 0.084
- Dataset: first two Modern Talking albums
- Generated track "Midnight Phantom" at epoch 800