GPT-6 rollout with Intelligent UI & Anthropic Claude Haiku 5.5 launch - AI News (Oct 9, 2026)
GPT-6 reaches every ChatGPT user, Claude Haiku 5.5 launches, and Ethereum researchers warn AI could crack crypto wallets before quantum computers do.
Our Sponsors
Today's AI News Topics
-
GPT-6 rollout with Intelligent UI
— OpenAI rolls out GPT-6 to all ChatGPT users with Intelligent UI, adding interactive charts, forms and on-the-fly tools to answers. It also starts responding before it finishes reasoning and brings stronger jailbreak resistance. -
Anthropic Claude Haiku 5.5 launch
— Anthropic released Claude Haiku 5.5, its cheapest and fastest small model, aimed at high-volume agentic tasks. The release also halves cache-read pricing for Sonnet 5.5. -
Grok routes tasks to rivals
— Elon Musk says Grok Bot will route tasks to outside models like Claude Opus 5.5, Midjourney and Suno, choosing whichever model is likely to give the best result for each request. -
StepFun Step 5 open weights
— StepFun's Step 5 Preview, a 600-billion-parameter mixture-of-experts model built for agentic coding and finance work, will have its open weights released on October 15. -
Microsoft and Liquid AI model updates
— Microsoft's MAI-Code-1.1-Flash brings cheaper, more efficient coding to GitHub Copilot. Liquid AI added two open-weight d1 decision models, including a multimodal version for edge devices. -
AI visual intuition benchmark gap
— Scale Labs' Humanity's Sixth Sense benchmark shows humans scoring 93% on intuitive visual reasoning, while the best multimodal model, GPT-6-astra, reaches only about 54%. -
Pac-Man benchmark and Perplexity retrieval
— Opper's JEVMAN benchmark tests AI decision-making through Pac-Man games. Perplexity launched late-interaction embedding models for multimodal and agentic search. -
AI threat to crypto wallets
— Ethereum's Vitalik Buterin and Justin Drake warn that AI-driven math breakthroughs could weaken elliptic curve cryptography before quantum computers do, and suggest a cautious move to fresh wallet addresses. -
Former OpenAI researchers safety letter
— Recently fired OpenAI researchers urged the company's board to preserve chain-of-thought monitoring and to work with outside AI safety auditors. -
Anthropic governance transparency questions
— An analysis of Anthropic's pre-IPO structure calls for publishing the Long Term Benefit Trust agreement, which shapes control of its board. -
Google SynthID Detector goes global
— Google expanded its SynthID Detector to everyone worldwide, helping users spot AI-generated images, video and audio through watermark detection. -
NVIDIA and Microsoft on-device AI
— NVIDIA and Microsoft unveiled RTX Spark, Windows Execution Containers and DGX Station for Windows to run AI agents locally on PCs. -
Debt financing fuels AI buildout
— Oracle, Broadcom and SpaceX are pursuing debt deals worth tens of billions of dollars to fund AI hardware, including Broadcom's custom chip with OpenAI. -
Texas grid slows data centers
— Texas and ERCOT are slowing approvals for large data centers after interconnection requests jumped to 474 GW, prompting tighter rules and more on-site power generation. -
Tony Fadell on AI gadgets
— Nest founder Tony Fadell says the Humane Ai Pin and Rabbit R1 failed by putting technology ahead of real needs, and predicts trusted, on-device assistants will come out ahead. -
Biohub $1.8B virtual biology push
— Biohub's Virtual Biology Initiative grows to a $1.8 billion commitment with the Department of Energy, NIH, DeepMind and Meta to build AI-ready biological data. -
Google Playground and open-source tools
— Google Playground lets users create games without code. New open-source projects include an on-screen arrow tool that lets AI agents point users to the right spot, and Rembrandt, a private Lightroom alternative.
Sources & AI News References
- → Tony Fadell Says First-Wave AI Gadgets Failed Over Weak Use Cases
- → Biohub Expands Virtual Biology Initiative With $1.8 Billion Commitment
- → OpenAI Rolls Out GPT-6 With Intelligent UI
- → Opper Launches JEVMAN, a Pac-Man Benchmark for AI Decision Models
- → StepFun Launches Step-Audio 3 Voice AI Suite
- → Google Launches Playground, a No-Code Game Creation Platform
- → LlamaIndex Launches OpenDocRouter for Unified Document Parsing
- → Anthropic Launches Claude Haiku 5.5 as Its Fastest, Cheapest Small Model
- → Microsoft launches MAI-Code-1.1-Flash for faster, cheaper coding
- → Grok Bot to Use Claude, Midjourney, and Suno Models
- → Ethereum Researchers Warn AI Could Threaten Crypto Wallet Security
- → Oracle, Broadcom and SpaceX Pursue Massive AI-Financing Deals
- → StepFun Launches Step 5 Preview for Agentic Work
- → Scale Labs Benchmark Exposes Gap in Human-Like Visual Reasoning ([labs.scale.com](https://labs.scale.com/papers/humanitys-sixth-sense))
- → NVIDIA and Microsoft Unveil RTX Spark and Windows AI Agent Push
- → Why Texas Is Slowing Data Center Power Connections
- → Perplexity Introduces Late-Interaction Multimodal Embeddings
- → Liquid AI releases open d1 edge decision models
- → Anthropic’s Hidden Governance Structure
- → GitHub Project Brings On-Screen Arrows to Guide Mac AI Agents
- → Rembrandt: Open-Source, On-Device Photo Editor
- → Former OpenAI Researchers Urge Transparency Into AI Reasoning
- → Google expands global access to SynthID Detector
Full Episode Transcript: GPT-6 rollout with Intelligent UI & Anthropic Claude Haiku 5.5 launch
What if the real threat to your crypto wallet isn't a quantum computer, but AI doing math? Some of Ethereum's top minds now think that could arrive first. Welcome to The Automated Daily, AI News edition. The podcast created by generative AI. I'm TrendTeller, and today is October 9th, 2026. It's a busy day, with new models, sobering benchmarks, power grid strain and a few useful tools, so let's get started.
GPT-6 rollout with Intelligent UI
First, an update on a story we've been following. OpenAI has started rolling out GPT-6 in ChatGPT, beginning with paid tiers and reaching Free and Go users a day later. The main new feature is called Intelligent UI. Instead of plain text, answers can now include charts, buttons, forms and small custom tools built on the fly. GPT-6 can also start replying before it has finished reasoning, so waits should be shorter. OpenAI says safety training is tougher against jailbreaks, while it also tries to cut down on pointless refusals.
Anthropic Claude Haiku 5.5 launch
Anthropic answered with Claude Haiku 5.5, its cheapest, fastest and most capable small model so far. It targets high-volume work like summaries, classification and running sub-agents inside coding workflows. Anthropic also halved cache-read pricing for Sonnet 5.5. The aim is clear: make running agents at scale cheaper.
Grok routes tasks to rivals
Elon Musk says Grok Bot will begin sending some tasks to rival models, including Anthropic's Claude Opus 5.5, Midjourney and Suno, picking whichever model is likely to do best. When even xAI leans on competitors' models, it suggests the future may belong to model routers rather than single models.
StepFun Step 5 open weights
In China, StepFun previewed Step 5, a large mixture-of-experts model built for long, multi-step agentic work in coding and finance. It is already available through StepFun's API, and the company says open weights will follow on October 15.
Microsoft and Liquid AI model updates
Two quick model updates. Microsoft's new MAI-Code-1.1-Flash is now live in GitHub Copilot, and Microsoft says it writes better code at about a quarter of the cost of its June model. Liquid AI also released two small open-weight decision models. They make choices in a single pass instead of generating text, which makes them a good fit for robotics and edge devices.
AI visual intuition benchmark gap
Now for a humbling result. Scale Labs released a benchmark called Humanity's Sixth Sense, which tests the quick, intuitive reading of images and video that people do without thinking. Humans scored about 93 percent. The best model, GPT-6-astra, managed roughly 54 percent, even at maximum reasoning effort. Bigger models clearly haven't solved intuition yet.
Pac-Man benchmark and Perplexity retrieval
On a lighter note, Opper released JEVMAN, an open-source benchmark where AI models play Pac-Man against the classic ghosts, with only two seconds to choose a direction at each turn. It's a playful way to test real decision-making instead of language skills. Meanwhile, Perplexity launched new embedding models that compare queries and documents in finer detail, and it reports stronger results for image, long-document and agentic search.
AI threat to crypto wallets
Now to that crypto warning. Ethereum researcher Justin Drake argues that AI-driven breakthroughs in mathematics could weaken the cryptography that protects wallets, possibly before quantum computers become a real threat. Vitalik Buterin agrees. He suggests that people who can do so safely should move funds to fresh addresses that have never signed a transaction, but he warns that rushed migrations can cause losses of their own. He also questions whether some quantum-resistant schemes are as safe as people assume.
Former OpenAI researchers safety letter
On AI safety, recently fired OpenAI researchers have written to the company's board urging it to keep the ability to monitor how models reason, and to bring in outside auditors. Separately, an analysis of Anthropic's pre-IPO structure points out that the agreements behind its Long Term Benefit Trust, which appoints four of seven board seats, are still not public. The author argues they should be released before the IPO.
Anthropic governance transparency questions
On transparency, Google is opening its SynthID Detector to everyone worldwide, in English. The tool checks images, video and audio for invisible watermarks that mark AI-generated content. Google says these watermarks are already in more than 180 billion images and videos.
Google SynthID Detector goes global
On hardware, NVIDIA and Microsoft are teaming up to bring AI agents to Windows. Microsoft introduced Execution Containers, which let agents run safely under the operating system's control. NVIDIA showed RTX Spark, a platform for running large models locally on laptops and small desktops, which ships October 16. The shared message is that more AI work should run on your own machine instead of the cloud.
NVIDIA and Microsoft on-device AI
Paying for all of this is getting expensive. The Wall Street Journal reports that Oracle, Broadcom and SpaceX are seeking debt deals worth tens of billions of dollars each, including more than 50 billion dollars for Broadcom's custom chip with OpenAI. More and more of the AI boom is being funded with borrowed money.
Debt financing fuels AI buildout
Power is the other limit. In Texas, requests to connect to the ERCOT grid jumped from 63 gigawatts to 474 gigawatts in about eighteen months, mostly from data centers. Regulators are now slowing approvals, adding fees and screening out speculative projects. Developers are turning to on-site gas plants to get electricity faster.
Texas grid slows data centers
Nest and iPod veteran Tony Fadell has a view on AI devices too. He says gadgets like the Humane Ai Pin and Rabbit R1 failed because they put flashy technology ahead of real needs and never earned people's trust. He expects successful assistants to run mostly on the device, and he thinks Apple is well placed even without a top in-house model.
Tony Fadell on AI gadgets
In science, Biohub's Virtual Biology Initiative has grown into a 1.8 billion dollar effort. The Department of Energy is putting in over 500 million dollars, and DeepMind, Isomorphic Labs and Meta are adding 300 million. The goal is to build open, AI-ready biological data for virtual models of cells and disease, which could speed up medical discovery.
Biohub $1.8B virtual biology push
Finally, a few tools. Google launched Playground, an experiment that lets US adults create and share games by describing them in plain language. On GitHub, a clever macOS utility lets AI agents draw a big arrow on your screen when they need you to approve something. And Rembrandt is a free, open-source Lightroom alternative with on-device AI editing and no account or tracking.
That's it for today's AI News edition. One thread ran through many of these stories: running AI locally, on hardware you control, came up again and again. Links to all the stories can be found in the episode notes. I'm TrendTeller. Thanks for listening, and see you tomorrow.
More from AI News
- October 7, 2026 AI Self-Models and Emergent Misalignment & OpenAI's $30 Billion Funding Talks
- October 6, 2026 Claude flags threats, Florida arrest & Sam Altman accepts AI harms
- October 5, 2026 Google pauses open-source bug bounty & Claude misused in African influence ops
- October 4, 2026 OpenAI safety culture criticized & LeCun rejects AI doom
- October 3, 2026 The Pause Becomes Real & the Bill Gets a Prospectus