The Editorial Staff at AIChief is a team of professional content writers with extensive experience in AI and marketing. Founded in 2025, AIChief has quickly grown into the largest free AI resource hub in the industry.
Sep 13
OpenAI Unveils Enhanced AI Models for Transcription and Voice
OpenAI introduces new transcription and voice AI models, enhancing user interaction and accuracy for developers and users alike.

Originally reported bytechcrunch
OpenAI has released upgraded transcription and voice-generating AI models, boasting significant improvements over previous versions. Aligning with its vision of more autonomous systems, OpenAI aims to create models that can independently perform tasks for users. Olivier Godement, OpenAI's Head of Product, describes these models as a step towards building chatbots capable of engaging in meaningful conversations with customers.
The new text-to-speech model, called “gpt-4o-mini-tts,” is designed to produce more realistic and nuanced speech while allowing developers to customize tones and styles. Users can guide the model on how to deliver lines, whether adopting a quirky style reminiscent of a “mad scientist” or a calm demeanor like a mindfulness teacher. Jeff Harris, a member of OpenAI’s product team, emphasizes the goal of enabling developers to craft both voice experience and emotional context for better user interactions.
In addition to the text-to-speech upgrades, OpenAI has introduced “gpt-4o-transcribe” and “gpt-4o-mini-transcribe” as replacements for the outdated Whisper transcription model. These new models were developed using diverse and high-quality audio datasets to improve the accuracy of speech recognition, especially in challenging environments. Harris notes a reduction in errors, claiming these models will not fabricate information as Whisper often did.
Despite these advancements, challenges remain, particularly for Indic and Dravidian languages, where OpenAI reports a word error rate nearing 30%. Interestingly, the new transcription models will not be available for open-source usage, diverging from OpenAI's tradition of releasing models under open licenses. Harris explains that these larger models are not suitable for local machine operation, prompting a more considered approach to their release.
#news
ES
Editorial StaffEditor
View all posts
Filter:
No comments yet. Be the first to comment!
Continue reading
View all newsRelated stories
Photon held a funeral for mobile apps. Now it has $4.5M to help replace them with agents
#ainews
The AI startupPhotonis so sure that agents will eventually come to replace mobile apps that it held a funeral for the latter — yes, a real funeral in a church, with speeches and everything. Photon, wh...
5 min read
2h ago
Legato Unveils AI-Powered Smart Glasses for Better Hearing
#ainews#hearingtechnology#hearingloss#smartglasses#open-eardesign
A hearing technology startup named Legato revealed on Thursday that its flagship AI-based hearing glasses are now on the market. The new eyewear, known as Legato Frames, incorporates the company’s pat...
2 min read
3h ago
Satlyt, founded by a former Google and SpaceX product manager, raises $8M to run AI on satellites
#ainews
The way Rama Afullo tells it, he saw orbital data centers coming. The Kenyan American engineer previously worked in Google’s cloud computing business before a brief stint at SpaceX’s Starlink communic...
4 min read
4h ago