Navana.ai Releases Indian-Language Voice Model at ₹12 per 10,000 Characters

Bodhi TTS gives businesses tools to correct names and numbers without retraining. Its speed, deployment and rival-price comparisons remain company claims.

By 3 min read
Navana.ai Releases Indian-Language Voice Model at ₹12 per 10,000 Characters
Navana.ai Releases Indian-Language Voice Model at ₹12 per 10,000 Characters

Listen to this story

The audio brief

About 1:26
0:001:26
Read transcript
Navana.ai has launched Bodhi TTS, a speech model for businesses that need automated voices to get Indian names and financial details right. It costs twelve rupees per ten thousand characters, and includes more than fifty voices across ten languages, including Hindi, Tamil, Telugu and English. The practical hook is control. Businesses can upload dictionaries for customer, product or branch names, then correct how the model says a word without retraining it. The company also says a few seconds of reference audio can create a custom voice. That could help teams tailor calls without building a new voice model for every use. Navana.ai says audio starts in under one hundred milliseconds. That is the time to first audio, not the time to finish a spoken response. It also offers deployment in a company’s own data center, which it says can keep audio, transcripts and personal data within the enterprise environment; the actual protections depend on how a system is set up. The company compares its rate with thirty rupees for Sarvam AI’s Bulbul v3 and about ninety-five rupees for ElevenLabs. Those are Navana.ai’s figures, not independently normalized comparisons, so they don’t establish equivalent quality or total cost. The key test is still in real calls: whether Bodhi TTS handles unfamiliar names and money details clearly, and whether quick first audio makes the exchange feel responsive.

Story brief

3 key points

Navana.ai’s Bodhi TTS is now available for businesses seeking Indian-language speech generation, with a list price of ₹12 per 10,000 characters and more than 50 voices. Its commercial pitch centers on operational control: teams can correct pronunciations through dictionaries and optionally deploy the model in their own data center. Navana.ai claims audio begins in under 100 milliseconds, but that is first-audio...

  1. 01

    Coverage spans Hindi, Telugu, Tamil, Marathi, Bengali, Odia, Malayalam, Kannada, English and Gujarati.

  2. 02

    Navana.ai says teams can upload name dictionaries and correct pronunciations without retraining; voice cloning uses a few seconds of reference audio.

  3. 03

    The company compares ₹12 per 10,000 characters with ₹30 for Sarvam AI’s Bulbul v3 and about ₹95 for ElevenLabs; these are vendor-supplied, not independently normalized estimates.

A bank’s automated caller can sound fluent and still get a customer’s name or loan details wrong. Navana.ai has launched Bodhi TTS, a model that turns text into speech across 10 languages, with controls for how particular words are pronounced. The company lists it at ₹12 per 10,000 characters, pitching a way for businesses to make large volumes of customer calls without paying separately for custom voices.

The release covers Hindi, Telugu, Tamil, Marathi, Bengali, Odia, Malayalam, Kannada, English and Gujarati, with more than 50 voices. Navana.ai is aiming the model at real-time applications: it says speech starts in under 100 milliseconds, while the rest of the sentence is still being generated. That measures the start of audio, not how long a complete call response takes.

Built around the words a bank actually says

Navana.ai says Bodhi TTS is designed to speak currency amounts, phone numbers, PAN numbers, dates, loan installments, scheme names and account numbers in forms familiar to Indian businesses. A company can upload a dictionary of product, branch and customer names to steer pronunciation. It says a correction does not require retraining the model, which would let a team address a name that comes out wrong without rebuilding its speech system.

The other customization tool is voice cloning. Navana.ai says a developer can submit a few seconds of reference audio through its API and generate a custom voice without additional training. It says the reference is cached so that using the voice adds no delay during a call. Whether that holds across customer deployments is a different question from whether the feature is available.

The company’s price comparison
₹12Bodhi TTS

Navana.ai lists Bodhi TTS at ₹12 per 10,000 characters, with all voices and features included.

₹30Sarvam AI Bulbul v3

Navana.ai cites ₹30 per 10,000 characters for Sarvam AI’s Bulbul v3. This rival-price figure comes from Navana.ai’s comparison.

About ₹95ElevenLabs

Navana.ai puts comparable ElevenLabs pricing at about ₹95 per 10,000 characters. This is its own comparison, not a verified like-for-like service-cost estimate.

The price is only part of the pitch

Navana.ai says every Bodhi TTS voice and feature is covered by that list price, rather than a premium tier. It estimates the charge at about ₹0.72 per minute of generated speech. The rival figures above are also supplied by Navana.ai; a buyer comparing services would still need to weigh voice quality and the amount of speech its own calls generate. The stated character rate is a clearer starting point than a projected per-call bill.

For a bank deciding where customer information goes, Navana.ai makes a separate deployment claim. It says Bodhi TTS can run inside an enterprise’s own data center, keeping audio, transcripts and personal data within that organization’s security environment. The company says the model is roughly five times smaller than the nearest comparable architecture, but does not identify that architecture in its comparison. An in-house installation is an option it describes, not a guarantee about every customer setup.

Access is open; the call test comes next

Bodhi TTS is available through Navana.ai’s platform, with ₹1,000 in signup credits for new users. The start date is less tidy than the access status: Moneycontrol says it was available from September 23, while a company-supplied release published by ANI says public access opened September 24. Neither account changes the central development: customers can now try the model rather than wait for a planned release.

The next useful evidence will come from the calls Bodhi TTS is meant to serve: whether its voices handle unfamiliar names and financial details clearly, and whether fast first audio leads to a responsive exchange. The launch supplies features, a price and Navana.ai’s performance claims, but not results from customers testing those claims across their own calls. Voice cloning raises another practical choice: when, if ever, should a business use a copied voice to speak with a customer?

Sources

  1. moneycontrol.comNavana.ai launches Bodhi TTS voice AI model in 10 Indian languages- Moneycontrol.com
  2. aninews.inNavana.ai launches ‘Bodhi TTS’, Indian Text-to-Speech Voice AI Model in 10 Indian Languages; Priced at one-eighth of ElevenLabs, 60% below Sarvam

Loading discussion...

YOUR READING SPACE

Notifications