Your daily AI news digest
Anthropic launched Claude Frontier Academy, a first-of-its-kind training program from an AI company, aimed at what it calls the hardest talent to find in enterprise AI: engineers who can take Claude from an idea to a system running in production. The Academy's first program, the Frontier Deployed Engineer Residency, follows what Anthropic calls the medical model. Engineers learn from practitioners, practice on realistic cases, and are assessed before they work alone. Organizations nominate their strongest engineers, each arriving with a named Claude project to lead when they return, and Accenture, Bain, Capgemini, Commonwealth Bank of Australia, Deloitte, McKinsey, Morgan Stanley, and Novo Nordisk are in the first cohorts.
The path starts with a multi-day in-person program that includes a simulated enterprise deployment, from choosing the use case through security review to handover, and ends with a graded practical. Those who pass earn a Claude Resident Engineer badge and move into a 12-week residency leading a real Claude use case at their own organization. A second assessment at the end of the 12 weeks earns the Frontier Deployed Engineer badge, with the first expected in early 2027. Cohorts are running now, and participation is by nomination through Anthropic account teams.
An interactive thought experiment, explicitly not a measurement of any current model. At an imagined million tokens a second, 1,000 story endings take about one second to write, but testing, evidence, and judgment keep their own clocks, so the scarce skill becomes choosing among the options.
Apple announced new controls around Full Disk Access, a setting built for backups that gives an app reach into files, mail, messages, and browsing history. It says capable AI agents make that access riskier and wants users to grant it only through "very explicit user action." The move follows a claim, which Meta disputed, that Muse read a columnist's private messages.
OpenAI says it parted ways with three safety researchers after an investigation found they mishandled sensitive company information, which the WSJ says was shared with a third-party AI safety organization. The report did not name them. It landed two days after the New York Times reported that executives brushed aside employees' safety warnings.
Argon is rolling out first to a select group of cyber partners through Google's Fairwind Program. Google says it was trained for defensive security and can "autonomously find, validate, and patch critical software vulnerabilities," and cites the Vals index to claim it leads GPT-6 Astra and Anthropic's Fable and Opus models.
A team from Carnegie Mellon, MIT, NYU, and Stanford built Ataraxos, which beat Pim Niemeijer 15 games to one with four draws. It trained on 163 million self-play games using 16 GPUs and a few thousand dollars, and a second network that guesses hidden pieces lets it plan ahead.
Barclays expects Claude Code adoption to reach 50% of its developers by the end of 2026 and a majority of software engineers in 2027. Its Claude-powered Colleague Knowledge Assistant, live since 2025, has been adopted by more than 16,000 colleagues and has handled over one million searches.
Merchants chat with Shopify's Sidekick agent while the store's real code renders live, so interactivity and animation can be tested as changes happen. Third-party themes, app blocks, markets, and translations are not in the first release, and Canvas is desktop-only.
OpenAI launched a Try On button in ChatGPT shopping results, plus a Favorites library, both built on the new ChatGPT Images 2.5 model. Users upload a selfie or full-body photo. Google launched its own virtual try-on last year, and OpenAI earlier pivoted away from an instant checkout feature that underperformed.
Photon helps developers build agents that work over iMessage, WhatsApp, SMS, and email. It says it has signed up more than 40,000 developers and grown revenue 10x in four months, though the open source version still accounts for 98% of use.
Legato Frames start at $999 and are built for adults with up to moderate hearing loss. An on-device AI separates voices from background noise, and a dual-speaker design cuts leaked sound by 99% a few inches from the ear. The company says there are no cameras and no recording.
Founder Rama Afullo calls Satlyt the "Android" to SpaceX's iPhone-style orbital data centers: open software that works across many companies' satellites. Its software launches this week on a SpaceX rocket alongside the first prototype of Google's Project Suncatcher.
Flow raised a $50 million Series B for AI agents that align CAD drawings with product requirements and test results. Customers include Anduril, Rivian, Joby Aviation, and Stoke Space, and Roelof Botha joined the board.
Labelbox, known for training environments for frontier labs, now sells Recursion, agents that start on schedules, alerts, or events and work across business apps. Its demo shows seven agents on an acquisition diligence job, with a grader agent scoring the outcome.
Uppu argues general intelligence will become a commodity while "personal intelligence," a person's intuition, judgment, and taste, becomes scarce and valuable. He says people will own and train their own, and that AI companies become platforms where those personal models are chosen the way models are chosen today.
Anthropic's new academy for 10,000 engineers is built on a simple admission: the scarce resource in enterprise AI is no longer the model, it is the person who can make it work inside a real business. Capgemini put it bluntly in its launch quote, saying clients care about "knowing firsthand how to solve real-world problems," not "certification counts or training volume." Today's stories keep returning to that gap between what AI can produce and who is qualified to direct, check, and trust it.
Put the Academy next to the OpenAI dismissals and a pattern shows up: the industry is now competing on trusted people, not just trusted models. Anthropic is spending $100 million to certify humans who can deploy Claude, and the Accenture quote says why, describing "judgment tested against a simulated enterprise deployment, not technical skills in isolation." OpenAI, meanwhile, fired three safety researchers for moving sensitive information outside approved channels, days after a Times report that executives waved off safety warnings. Both stories are about who gets to be the human in the loop, and what happens to the loop when that person is a liability or a bottleneck.
The consumer side has the same shape. Apple's Full Disk Access change exists because an agent that can read your mail is only as safe as the permission screen that preceded it, and it follows the Muse and ChatGPT Mac app reports. Photon's pitch, agents inside your messages, and Shopify's Canvas, an agent editing live store code, widen what a single "yes" can authorize. Labelbox's answer is telling: put a grader agent on every run and let the next run read the notes. That is supervision by software, which is useful, but notice that the 0.93 score in its demo still ends in a memo "for the deal lead's review." The human is still the last gate.
The echohive essay supplies the underlying economics. If writing costs nearly nothing, then checking is where the money goes. Barclays' numbers hint at the same thing: adoption by half its developers is not the headline, the 16,000 colleagues using a knowledge assistant that has handled a million searches is, because retrieval-and-review work is where trust gets built. Ataraxos adds a wry footnote. Farina says machines find it easy to act without being swayed by a secret they hold, which is exactly the discipline we ask of humans and often fail at.
Zoom out and the AI market is sorting itself into layers, each with its own scarce input. Models are converging, with Google claiming the top of the Vals index the same week Anthropic and OpenAI trade releases, and both Gemini and ChatGPT sitting near a billion monthly users. When capability converges, the differentiators move up and down the stack: down to chips and orbit, with Satlyt and Google's Suncatcher betting on computing in space, and up to people, with the Academy and Uppu's "personal intelligence" thesis that taste and judgment are what remain valuable.
The tension to watch is that autonomy is being sold faster than accountability is being built. Apple, a hardware company, is now writing agent policy for developers. A leading lab is dismissing safety staff for information handling while it ships its own agents. And the enterprise answer, per Anthropic, is a credentialed human with a 12-week residency. That is a reasonable first draft of governance, and a slow one next to a world where an agent can write a million tokens a second. The next two years will test whether training people, grading agents, and clicking "allow" can scale as fast as the thing they are meant to supervise.