Google DeepMind has introduced Gemini 3.8 Live and Extended Thinking, two new audio models that developers can access via the Gemini API. The Gemini 3.8 Live model is revolutionary for voice agents due to its ability to support API calls in the background, handle visual inputs, and maintain conversation simultaneously. It supports over 97 languages, offering developers a competitive tool for creating voice-driven applications. While it has achieved impressive benchmarks, particularly on the Speech-to-Speech Leaderboard, Google has opted for a lower pricing strategy compared to OpenAI. This move positions Gemini as an affordable alternative, costing developers approximately $1.38 per hour of voice conversation, significantly less than the $3.00 minimum for OpenAI's equivalent services. However, the trade-off appears to be quality, as OpenAI's models are noted for more natural-sounding conversations and superior performance, especially in full duplex interactions.
NewsBite reading:Google unveils Gemini 3.8 Live to rival OpenAI's GPT-Live-1 at lower costs
Google DeepMind has launched new audio models that provide powerful features at a lower cost compared to competitors.
Unchanged: OpenAI continues to offer high-quality models, albeit at a higher price point.
The announcement carries a cautious optimism as it introduces innovative features and competitive pricing, but there are underlying concerns about the quality implications.
The release enhances the AI landscape by increasing competition in audio model offerings.
Lower costs for cloud audio services provide more options for developers in building applications.
New APIs extend opportunities for developers to innovate in voice technology.
Launches competitive products aimed at increasing market share in AI.
Faces intensified competition in audio models and possible market share loss.
The launch signifies a competitive push in the AI audio model space, with Google targeting affordability while positioning itself to challenge the status quo set by OpenAI. This move could shift developer preferences towards Google's offerings, particularly in cost-sensitive applications.
They gain access to a competitive audio API at a lower price, allowing for cost-effective development of voice applications.
The introduction of affordable AI solutions has widespread implications for developers worldwide.
May need to adjust pricing or enhance model features to retain competitive edge.
Potential vulnerabilities in new software releases need scrutiny.
Must ensure user data protection while deploying models.
Risks associated with performance quality may affect brand perception.
Risks tied to successful adoption and integration of new models.
Reliance on cloud technology requires robust support systems.
No significant geopolitical implications identified.
Current models comply with existing regulations.
Dependence on AI training data and models can face challenges.
Shift towards affordable solutions could impact jobs in higher-end model development.
Low likelihood of liability in current model usage.