Context Windows Got Huge and Nobody Knows What to Do With Them
Million-token context windows were supposed to change everything. In practice, most AI products ignore them entirely.
Read more →Your daily snap of tech news
36 posts
Million-token context windows were supposed to change everything. In practice, most AI products ignore them entirely.
Read more →
The numbers AI companies publish to prove their models are smarter keep going up. The actual experience of using those models tells a different story.
Read more →
Autonomous AI agents promised to collapse your tool stack. Instead, they've quietly become the most expensive middleware you didn't ask for.
Read more →
Agents can book flights and write code, but ask one to handle a multi-step refund and watch it unravel completely.
Read more →
Open-source models have nearly caught up to frontier AI - but the companies celebrating that fact are the ones who should be worried.
Read more →
Every major AI lab is racing to release autonomous agents, but the definition of 'agent' is still being argued in real time.
Read more →
Persistent memory in AI assistants sounds like a feature. In practice, it mostly learns your bad habits.
Read more →
Persistent memory in AI assistants sounds like a win. In practice, it surfaces a problem nobody designed for.
Read more →
Every time OpenAI cuts API prices, the narrative is democratization. The actual story is about locking in developers before the margin improves.
Read more →
The numbers labs publish to sell their models tell you almost nothing about how those models perform on your actual work.
Read more →
Million-token context windows were supposed to change everything. In practice, most people are still pasting in three paragraphs.
Read more →
Reasoning models cost significantly more to run than standard models - and most products built on them haven't figured out who absorbs that.
Read more →
The labs are optimizing their models to ace evals instead of solving real problems, and the gap between scores and usefulness keeps widening.
Read more →
Claude's maker has built its brand on AI safety. That's admirable - until you notice how conveniently it maps onto competitive positioning.
Read more →
Expanding context windows hasn't made AI assistants more capable of knowing you. It's made it easier to confuse size with understanding.
Read more →
OpenAI keeps building tools that make more sense as infrastructure than as consumer software. That's not accidental.
Read more →
OpenAI wants AI agents to act on your behalf. The problem is that trust, not technology, is the actual barrier.
Read more →
AI agents that book restaurants and file forms sound useful - until you realise the friction they remove was never the real obstacle.
Read more →
Giving AI agents access to your browser and accounts sounds powerful. The real question is who benefits when something goes wrong.
Read more →
OpenAI wants AI agents doing your tasks. That's less about helping you and more about making you impossible to leave.
Read more →
OpenAI wants GPT-4o to book your flights and manage your inbox. The gap between that pitch and reality is still enormous.
Read more →
AI agents that browse, book, and buy on your behalf sound useful. But the trust infrastructure to make that safe barely exists.
Read more →
Dreams of Violets, a $2,000 AI-generated film dramatizing Iran's crackdown on protesters, will premiere at the Tribeca Festival next month.
Read more →
Anthropic has closed a $65B Series H round at a $965B valuation, potentially its last private raise before an IPO, with backing from Sequoia, Amazon, and others.
Read more →
Anthropic says its Mythos-class Claude model could reach general users within weeks, as the company works to finalize cybersecurity safeguards ahead of a public release.
Read more →
Anthropic's Claude Opus 4.8 is built to better acknowledge when it's uncertain, rather than presenting unsupported conclusions with false confidence.
Read more →
Anthropic released Opus 4.8 with a new Dynamic Workflows feature for coordinating large numbers of parallel subagents, plus improved handling of uncertain data.
Read more →
Bloomberg's Mark Gurman has shared an early look at Apple's Siri overhaul, expected at WWDC 2026, including a Dynamic Island home and a new chatbot-style interface.
Read more →
Asana acquired no-code agent builder Stack AI for $75 million, adding its founders and technology to Asana's growing suite of AI workflow tools.
Read more →
OpenAI's new Trusted Contact feature lets ChatGPT alert a nominated friend if a user shows signs of being at serious risk of self-harm, with human review required before any notification is sent.
Read more →
Major financial exchanges, including Shanghai's futures market and CME Group, are developing derivative products tied to AI compute tokens and GPU rental costs.
Read more →
Paris is emerging as a major hub for AI research, startups, and policy - challenging Silicon Valley's long-held dominance in the global tech conversation.
Read more →
Running an open-weight AI chatbot locally on an iPhone can cost as little as a one-time $5 purchase, with added benefits of offline use and stronger privacy.
Read more →
Microsoft is rolling out a redesigned Microsoft 365 Copilot with a faster interface, structured responses, and a context-sensitive UI approach called progressive disclosure.
Read more →
Microsoft has redesigned Copilot for Microsoft 365 with a minimal, largely black-and-white interface aimed at consistency across Word, Excel, and PowerPoint.
Read more →
YouTube is rolling out a prompt-based AI feature that lets US users generate a personalized Home feed, with open questions about what it means for creators.
Read more →