About Dhwani

Every word, right on time.

Dhwani (ध्वनि, Sanskrit for “sound” or “resonance”) turns raw audio into karaoke-style, word-synced lyrics. Upload a recording in Kannada, Hindi, or another Indian language, and Dhwani transcribes it, times every word, and lets you read along as it plays — just like karaoke, but for spoken audio.

What you get

01
Automatic transcription
Speech-to-text models (Sarvam AI, OpenAI Whisper) turn your audio into a phrase-level transcript with timestamps.
02
Karaoke-style playback
Word timings are interpolated from phrase timestamps so the transcript highlights in sync with the audio as it plays.
03
Titles, summaries & tags
Every clip can be auto-titled, summarized, and organized with tags so a growing library stays easy to browse and search.
04
Shareable clips
Each clip gets its own page with a deep-linkable timestamp and an auto-generated preview card for sharing.

How it works

1
Upload audio
Drop in a recording and pick a speech-to-text model.
2
Transcribe & time
Dhwani transcribes the audio and times every word.
3
Review & refine
Edit the title, summary, and tags before publishing.
4
Play & share
Read along in sync, then share the clip's own page.