Keel is an original studio in this repo. It is not Suno, does not call Suno’s API, and does not ship unlicensed celebrity voices. Duration is a hard budget. Famous names are refused.
Same six notes, 6.5 beats. The music does not move. Your words have to.
Syllable count roughly matches the melody. This is the only case where “perfect timing” and “new lyrics” can both be true without chewing the vowels.
Voice conversion keeps the words. Singing synthesis needs a piano roll. Lyric editing must not move the mix. Glue them and each one corrupts the others: convert after an edit and the new consonants smear; synthesize a cover and the drums shift; inpaint a line and the room tone dies.
Speech TTS can stretch a sentence. Singing cannot. Vowels are the notes. Consonants are the onsets. “Stay” and “put your name in the chorus tonight” are not interchangeable on the same MIDI. The 2026 papers even generate evaluation lyrics with an LLM so the new line is duration-feasible. Arbitrary fan lyrics fail that test.
“Perfect timing” usually means sample-lock to the original performance, which already leans, scoops, and breathes off the click. Forced aligners trained on speech miss long vowels. Score-based singers lock to MIDI and then sound like MIDI. You pick one clock. You do not get both for free.
Pick a song in the library, then press Play.