Google launches Gemini 3.8 Live and 3.8 Live Extended Thinking, its voice-first dialogue models for real-time reasoning
- Google launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking on Sep 15, 2026, available today through the Gemini API, Google Workspace, the Gemini app, and Search.
- 3.8 Live is built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding; 3.8 Live Extended Thinking targets high-complexity tasks with multi-step reasoning.
- Both models handle complex reasoning, real-time visual context, and background task execution without interrupting the conversation, and the article frames them as the building blocks for production voice agents.
- The post is credited to Tom Ouyang (Principal Engineer) and Malini Jaganathan (Member of Technical Staff) on behalf of the Gemini Audio Team, and the page's own summaries are generated by Google AI.
Hacker News opinions
Their flagship demo has the model walking straight into the most common checkmate pattern in chess. Rough look for something they're calling their most advanced live model.
It's still fine for a live model. You can just hand it engine analysis and it plays at whatever strength you want, no correlation between its understanding of the position and the move it picks.
For an LLM, finishing a full game without illegal moves or inventing pieces that aren't on the board is already an achievement, and even more so for a live one.
Gemini is underrated for prose. It's the only model whose output I can actually stand reading.
I use Astra for heavy work, but Gemini is way more fun for rabbit holes and brainstorming. Worried they'll lose that natural tone chasing SOTA.
It's also the only model that gets translation and localization right. Coding is subpar, but the natural language side is top tier.
We run our analysis on DeepSeek v4.1 Flash and pass the output through Gemini 3.8 Flash just for readability. Works.
I find Gemini's style the most sycophantic and annoying, personally.
Try Gemini Live in a multilingual room. It picks out speakers and live translates to you. Genuinely underrated.
Terrible for code, amazing for prose. Documentation sub agent runs on Gemini, code agent on Luna.
I work at Google and we choose between Gemini and Opus. Opus is slightly better than Gemini Flash but the style is unbearable, like a pedantic grad student.
Our Workspace Business seat still only shows 3.6 Flash and 3.6 Thinking in the Gemini app. Anyone actually seeing 3.7 or 3.8 roll out?
I have both. Benchmarks on my task are actually better on 3.7. Still an improvement over the rest.
Basic Workspace only includes 3.6 Flash and 3.1 Pro. Your admin has to upgrade the seat, it's 24/mo starting Jan 2027.
Live Mode already beats GPT Voice for me even when it was dumber. Feels like talking to a person, ChatGPT just hums along with weird voices. Annoyed they didn't ship 3.8 to AI Plus yet.
When is this coming to Vertex?
Vertex is dead, and for good reason.
Everyday user here: 3.8 Flash is already good enough for most people, only Astra Max edges it on vision. Fast and compute light.
Google has no incentive to fight in coding models. They make money routing everyone through search, ads and YouTube. Everything else is a sideshow.
Coding agents are a race to the bottom anyway. Zero switching cost, hard to differentiate. Focusing on the Gemini app and search integration is the shrewder move.