Premium

‘Low latency critical for enterprise-grade voice AI assistants’: Gnani.ai CEO Ganesh Gopalan

Amid AI Impact Summit 2026, Ganesh Gopalan, CEO, Gnani.ai, discusses building AI voice-generation models to scale and the challenges of sourcing voice training data.

Gnani AIGnani.ai booth at India-AI Impact Summit 2026. (Express Image)
Written by: Karan Mahadik
9 min readNew DelhiFeb 19, 2026 09:06 AM IST First published on: Feb 18, 2026 at 11:59 AM IST

With AI-powered voice assistants set to reshape customer support interactions, Indian AI startups are racing to build foundational AI voice-generation models that go beyond simple text-to-speech or speech-to-speech translation. By focusing on phonetics, prosody, semantics, and intent, these homegrown conversational AI tools aim to generate speech that preserves tone, emotion, pacing, and pauses in order to make these interactions more natural and human-like.

One such speech-to-speech AI model was released under research preview amid the ongoing AI Impact Summit 2026 hosted by India at Bharat Mandapam in New Delhi. The 5 billion-parameter closed model is part of a sovereign AI stack called Inya VoiceOS developed by Gnani.ai, which is among the cohort of Indian AI startups selected under the Centre’s Rs 10,372-crore IndiaAI Mission to build sovereign AI models and strengthen India’s position in the global AI race. Gnani.ai on Wednesday, February 18, launched another model called Vachana TTS (text-to-speech) designed for voice cloning in 12 Indian languages, including Hindi, Bengali, Tamil, Telugu, Kannada, Malayalam, Gujarati, Marathi, Punjabi, Odia, Assamese, and Indian English.

Karan Mahadik is a Tech Correspondent for The Indian Express based in Delhi... Read More

Latest Comment
Post Comment
Read Comments