|
|
|
The Anthropic Institute put its internal numbers on the record. The figure that should change your plans is not the 80%.
|
|
As of May 2026, Claude authored over 80 percent of the lines merged into Anthropic's production codebase, per the Anthropic Institute, up from low single digits before Claude Code shipped in February 2025. The bottleneck that replaced typing is human code review.
Why review is now the bottleneck →
|
|
|
|
|
Anthropic is running a wet biology lab and opening the strong models to vetted researchers
Eric Kauderer-Abrams, Anthropic's head of life sciences, confirmed to TechCrunch that the company runs a Bay Area lab where Claude and robotic equipment execute physical experiments, aimed at fundamental biology and rare diseases rather than drug discovery. The same week it launched the Life Sciences Verification Program, giving vetted researchers tiered access to capabilities the public models refuse. Why it matters. If your research keeps hitting refusals on protein or pathogen work, the fix stopped being a better prompt and became an institutional application. What the verification tiers unlock →
|
ChatGPT is inside Word now, on every plan including Free
The add-in went live September 17, installs once from Microsoft Marketplace, and covers Word, Excel, and PowerPoint together. Enterprise and Edu accounts get GPT-5.6 Sol inside Word at no extra cost through September 30, excluded from normal limits. Why it matters. The Sol window closes September 30 (your no-cost evaluation period); on October 1 the integration flips on by default for Enterprise and Edu, leaving admins eleven days to set policy before the deadline. Read the admin checklist before Oct 1 →
|
Gemini breached three real companies in May. Google confirmed it in September, after a reporter asked.
A configuration error during an authorized evaluation by Irregular exposed Gemini to the public internet instead of an isolated sandbox. It guessed passwords into one company and found credentials in public repositories for two more, then halted itself once it detected real targets. Irregular told Google in late July; confirmation came only after the Wall Street Journal asked for comment. Why it matters. Google is the fourth lab after OpenAI, Anthropic, and Meta to disclose an Irregular-linked escape inside six weeks, and every one traces to sandbox configuration, not model behavior. Test your egress boundary before you point an autonomous agent at anything holding credentials. How all four labs got here →
|
Every major lab cites Vals in its model cards. Vals just raised $40M to keep its tests secret.
Vals AI closed a $40 million Series A led by Andreessen Horowitz and grew revenue eightfold relative to all of 2025, per TechCrunch. Its test sets for law, finance, coding, cybersecurity, and biosecurity never go public, which is why OpenAI, Anthropic, Google, Meta, and xAI all cite Vals in their own model cards. Why it matters. Public leaderboards are training data now. A benchmark that stops separating models is rarely a story about capability, and the price of a private score is that you cannot audit the methodology yourself. Why buyers are paying for private evals →
|
|
•
|
Codex CLI 0.155.0 added experimental voice conversations. Live transcripts and microphone controls sit behind the /experimental flag; patch 0.155.1 disabled reasoning summaries by default after providers rejected requests. Release notes → For production voice in your own apps: ElevenLabs →
|
|
•
|
GitHub Copilot retires six models on October 19. GPT-5.5, GPT-5.4, GPT-5.4 mini, GPT-5 mini, Grok 4.5, and Gemini 3.7 Flash go dark across chat, inline edits, agent mode, and completions; orgs that disabled global defaults must manually enable replacements before the deadline. Deprecation table →
|
|
•
|
Claude Code 2.1.277 reads AGENTS.md, but only when no CLAUDE.md is present. Patch 2.1.278 moved auto mode to the server-side classifier for API and Enterprise users, removing the classifier overhead charge. Changelog → Running agents on a dedicated box: Cloudways →
|
|
•
|
BragJack: a single browser extension can hijack any AI browser agent. Demonstrated against Gemini Live (Chrome), Edge, Opera Neon, Perplexity Comet, and Claude in Chrome; Chrome (CVE-2026-0628) and Edge (CVE-2026-55945) are patched with no in-the-wild exploitation confirmed. Write-up →
|
|
•
|
Trump says he is forming an "AI Force." No budget, staffing, or home department attached yet; an AI czar to be named later. Nothing to plan against yet. Coverage →
|
|
|
|
|
Jonathan Hildebrandt
Co-founder and primary operator of Pondero. Writes the Pondero Brief.
|
|
|
Affiliate disclosure
·
Unsubscribe
·
Manage preferences
Pondero earns commissions on some links. This does not affect our editorial picks.
|
|