← All projects

ragafied

A singing-practice app I first built in 2012-13, before browsers could do the audio work it needed, and rebuilt in 2026 on AI models, offline and real-time signal processing, and algorithms of my own: upload a recording and it becomes a track you can sing, with the vocal removed, key and tempo moving independently, and phrase-by-phrase call-and-response practice.

I first built Ragafied in 2012-13, around an idea that has not changed since: Hindustani vocal music is learned by imitation, so a recording of a great singer should become something you can sing back to, one phrase at a time. The browser of the time could not carry that idea. Its audio support was young and uneven from one browser to the next, custom sound processing could not run off the page's main thread, there was no way to run near-native code or a neural model inside a web page, and microphone access was patchy. What I could build then was the scaffold around the idea - accounts, ragas, gurus, lessons - and not the instrument at its centre.

The 2026 version is that instrument, and it runs at ragafied.com, where singers keep their tracks in playlists that nest like folders, share them with other singers, and practise with the lyrics open beside the player. A demo player on the home page lets a first-time visitor try a track straight away.

What it is built on

  • AI models. Neural source separation lifts the lead vocal away from the accompaniment. A neural pitch detector reads the teacher's melody ahead of time and the learner's voice live in the browser. A neural beat tracker finds the pulse across the whole recording. The heavy models run on cloud graphics hardware that scales to zero when nothing is waiting.
  • Offline signal processing. Every upload goes through a pipeline on the server - separation, rhythm, pitch, phrases, playback preparation - with each result remembered by the audio's own fingerprint, so no work is ever done twice.
  • Online signal processing. In the browser, in real time: key and tempo move independently of each other, changes cross over mid-phrase without a gap, the room is measured and the accompaniment cancelled out of the microphone, and a chosen tempo and key are rendered ahead and kept on the device.
  • Algorithms written for this music. A pitch-and-tempo processor that rebuilds a slowed voice from the singer's own pitch periods, so a phrase at a fifth of its speed still sounds like a person singing. Phrase boundaries placed where the singer actually breathed. Ornaments - meend, gamak, khatka, andolan - read from the motion between notes, and a slow-down that stretches those ornaments far more than the held notes around them.