01
Gemini 3.8 Live talks and thinks in the background
DeepMind ships Gemini 3.8 Live and Live Extended Thinking: conversational models that talk, reason, and keep tasks running without breaking your flow.
What they showed / shipped
- Google DeepMind introduces Gemini 3.8 Live and 3.8 Live Extended Thinking as its best conversational AI: talk, think, and handle background tasks without breaking flow (@GoogleDeepMind).
- Official Google blog + HN front page on the same drop (blog.google; HN ♥292).
- 📌 On the radar: Sam Altman teases "big 🚢 this week and then for DevDay 🚢🚢🚢…" (@sama).
Why it matters
- Background tasking + extended thinking is the agent primitive people actually feel in a chat product, not another bench number.
Sources
02
Odyssey-3 world model drives robots and cars
Odyssey unveils Odyssey-3, a foundation world model that can control robots, humanoids, cars on Indian roads, drones, training loops, and even games.
What they showed / shipped
- Odyssey unveils Odyssey-3 as a big step for foundation world models: control robots, power humanoids, drive cars on Indian roads, train AIs, pilot drones, and play video games (@odysseyml).
- 📌 On the radar: Agility's new humanoid will stop and squat to avoid harming human coworkers (Ars); Boston Dynamics Spot + Orbit get agentic factory workflows (@BostonDynamics); Reddit flags OM-1 as a fast embodied AI stack (r/singularity).
Why it matters
- World models that act in the physical world are the door after chat agents.
Sources
03
Perplexity agents build CobbleDB in two months
Two Perplexity engineers plus hundreds of Computer agents shipped CobbleDB, an in-house DynamoDB-class store, aiming to save up to $100M a year, while Computer lands preinstalled on HP ZBook.
What they showed / shipped
- Arav: two engineers + hundreds of persistent Computer agents built a DynamoDB replacement for fast web content fetches in two months; migrating in-house could save up to $100M yearly (@AravSrinivas).
- Perplexity publishes the CobbleDB research writeup: proactive always-on agents built core search infra (@perplexity_ai).
- Perplexity Computer now comes preinstalled on HP ZBook Ultra G3a, with Autodesk among hundreds of tools (@perplexity_ai).
Why it matters
- This is the receipt for agent-built production infra, not a toy demo — two humans, agent swarm, real cost target.
Sources
04
Claude for Financial Advisors ships real connectors
Anthropic's Claude for Financial Advisors is a Cowork plugin with live wealth-stack connectors, not a new model and not unsupervised advice.
What they showed / shipped
- Min Choi: Claude for Financial Advisors = Cowork plugin with connectors to Charles Schwab, BlackRock Advisor Center, Addepar, Envestnet/Tamarac, Orion, Wealthbox, iCapital, and more, plus skills for meeting prep, notes, rebalance review, and compliance docs (@minchoi).
- Framing: not a new model, not ChatGPT's financial product, not unsupervised advice — human still approves the regulated stuff. Available today on the enterprise/RIA path.
Why it matters
- Agents + workflow connectors beat "another chat with a finance prompt."
Sources
05
OpenArt Arena ranks creative models by real jobs
OpenArt launches Arena, a blind creative leaderboard judged by pros across Film, Ads, Animation, Motion Design, and more, so "best model" finally means best for a job.
What they showed / shipped
- OpenArt Arena: creative pros judge image and video outputs blind — no logos, no model names — with separate boards for Film, Ads, Animation, Motion Design, Video Editing, Lip Sync, and more (@minchoi).
- 📌 On the radar: Cartesian is AI 3D modeling for design (Formas; HN ♥90); Runway shows a Solaris "playable internet" concept (@runwayml).
Why it matters
- Blind pro judging is a better eval shape than vibe screenshots.
Sources
06
Paying for frontier buys 4 months at 5x cost
Ars reports open Chinese models have closed so much of the gap that paying for frontier mostly buys a ~4-month head start at about 5x the cost.
What they showed / shipped
- Ars Technica exclusive: paying for frontier AI models buys roughly a 4-month head start at about 5x the cost as open Chinese models close the gap (Ars).
- 📌 On the radar: a 44M-parameter quantized LLM trained from scratch on 45B tokens ships in 19.8 MB and runs ~1,900 tok/s on CPU (r/MachineLearning); Reddit says US gov RAG is using a Qwen embedding model (r/singularity).
Why it matters
- Hard tradeoff you can decide with — latency to frontier vs cost of good-enough open weights.
Sources
07
Local LLM payback math and an M4 Linux GPU driver
Sunk Cost estimates when a local LLM rig pays for itself, while a builder ships a Linux GPU driver for the M4 Mac Mini in one month.
What they showed / shipped
- Show HN: Sunk Cost — calculator for how long until a local LLM rig pays for itself (sunkcost.ai; HN ♥46).
- Cody Ho: built a Linux GPU driver for the M4 Mac Mini in one month (codyho.dev; HN ♥152).
- 📌 On the radar: IEEE Spectrum on the 2026 inference hardware revolution (Spectrum).
Why it matters
- Local AI stops being vibes when you can price the payback and run Apple Silicon under Linux.
Sources
08
Data centers are unpopular and gas-hungry
Polls show AI data centers losing with the public, while projections say US data centers could burn more natural gas than Germany and Japan combined by 2035.
What they showed / shipped
- The Verge / NYT midterm poll framing: AI and data centers are incredibly unpopular across polls (Verge).
- TechCrunch: US data centers could consume more natural gas than Germany and Japan combined by 2035 (TC).
- TechCrunch: the AI data center boom is colliding with cities scarred by big industry (TC).
- 📌 On the radar: NVIDIA DSX AI Factory Platform pitches megawatt-aware AI factories (@nvidia); Jensen says leave safety to industry, not regulation (TC).
Why it matters
- Power + politics are becoming the real capacity constraint, not just GPUs.
Sources