Lesson introductions
Framing passages before a lesson, the use our note names. One sentence per line, one caption per sentence.

Our note gives Ananya lesson introductions, and the published list gives her the standard Hindi record: a C grade, B target quality, 10 to 100 minutes of training audio, from a Hindi set that totals one to ten hours across four voices. Unlike the Mandarin and Japanese voices, she returns word-level timing, so captions can follow each word.
Hear it
On the record
All four Hindi voices carry the same line on the list: C, target B, tens of minutes. The grade is about data, not delivery. What sets the Hindi set apart from the other non-English voices is timing: the engine reports where each word starts, so subtitles can highlight word by word.
Lesson introductions are short framing passages before the content, a use where an even opening and a clean caption matter more than range. Keep each sentence on its own line and each becomes one cue.
Length the engine likes. The published guidance puts the sweet spot at roughly 100 to 200 tokens, a few sentences; very short lines are the weak spot and very long ones can rush.
Where we would use it
Framing passages before a lesson, the use our note names. One sentence per line, one caption per sentence.
Word-level SRT lets a learner follow each word as it is read; slow the speed to half for shadowing.
Any English voice can sit behind her; the pairings below are where we would start.
Blends
No recipe we kept includes Ananya. Each card is a pairing to try; she keeps the Hindi pronunciation and the second voice lends its timbre.
A pairing to try. Nora is our English lesson-introduction voice; listen for a teaching timbre carried into Hindi.
50% Nora
A pairing to try. Our grade A voice at 60 percent; listen for whether the clarity carries.
60% Maya
A pairing to try, two Hindi voices with teaching notes: Meera reads instructional scripts by our note.
50% Meera
A pairing to try. Delaney's throaty edge worked behind Mandarin, Japanese, and French primaries in kept recipes; listen for it under Hindi.
50% Delaney
A fixed sample at this mix, rendered once and cached. No credits used.
A cross-language blend: Ananya keeps the pronunciation, the second voice only lends its timbre.
Every blend works the same way: the first voice sets the pronunciation and pacing, the second voice lends only its timbre, and you choose the ratio between 10 and 90 percent. See the three steps on the voice library page
At a glance
| Language | Hindi |
|---|---|
| Voice | Female |
| Overall grade | C |
| Target quality | B |
| Training audio | 10 to 100 minutes |
| Captions | Word-level captions |
| Cost | 1 credit per job, room for about 3,300 characters |
| Formats | MP3 free, WAV on paid plans |
| Commercial use | Included on every plan |
| Preview | Free, no sign-up |
Grade, target quality, and training audio come from the engine's published voice list and estimate training data, not how a voice sounds.
FAQ
Yes. Hindi voices return word-level timing, so the SRT and WebVTT can highlight word by word, unlike the sentence-level cues for Mandarin and Japanese.
Billing counts UTF-8 bytes and Devanagari characters are three bytes each, so one credit carries about 3,300 characters. The cost is shown before you generate.
Ten to a hundred minutes of training audio at B target quality. It is the same line for all four Hindi voices and measures data, not sound.
Yes, the speed slider goes down to half speed with the pitch unchanged, and the captions stay aligned.
Similar voices
Paste a script, keep Ananya or blend a second voice, and download the MP3 with subtitles.