FM News
Founder Mode reads
Google’s new models demand a choice between speed and depth
◆ 95RelevanceOn a story from Google DeepMind2h ago
The split between 'Live' and 'Extended Thinking' signals that a single model no longer fits all your product's needs. You must now decide which features require instant latency and which benefit from slower, high-reasoning compute.
Takeaways
- Use 'Extended Thinking' for complex logic tasks that can tolerate higher latency.
- 'Live' is optimized for real-time interaction and low-latency user experiences.
- Re-evaluate your model routing to match task complexity with the right variant.
Read the original at deepmind.google
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
fmode.me/n/introducing-gemini-38-live-and-38-live-extended-thinking
Written by Founder Mode using gemini-3-flash-preview, from the publisher's own summary. We link the original rather than reproduce it — the reporting belongs to Google DeepMind.
More from FM News
Your Agent Aced the Task. Will It Do It Again?1 pts · Hugging FaceAEO startup Profound hits unicorn valuation, raises $180M Series D 7 months after last round1 pts · TechCrunch AIHow to attach an owner to every cloud resource you find1 pts · The New Stack[AINews] AEF-1 standard emerges for Third Party Evaluators, as Xai, OpenAI, and Anthropic all cosign1 pts · Latent Space