|
|
|
Copilot added two premium GA models that draw down your credit pool - here is how to route usage by plan before October 2 deprecations land.
|
|
Your seat caps headcount, not model spend - GPT-6 Astra and Claude Fable 5.1 went GA in Copilot this week, both billing at provider list pricing under usage-based billing rather than riding free on the seat. Route deliberately or watch the credit pool drain; four models retire October 2 and replacements need a manual admin enable. See the billing math by plan →
|
|
|
|
|
Meta undercut the frontier on input price
Muse Spark 1.3 landed September 2 in Muse Code and the Meta Model API, per Meta AI Research, at $1.25 per million input tokens and $4.25 per million output, with a 1M context window and an 88 percent prompt-cache discount, per Artificial Analysis, which scores it 48 on its Intelligence Index (13th of 202 models). Meta engineers report roughly 20 percent fewer tool calls and 25 percent fewer tokens than 1.2 on the same coding work. Worth a routing test if you run high-volume repeated-context jobs, because the cache discount is where the money actually shows up.
|
Anthropic will let regulated teams keep their own logs
Enterprise Frontier Safeguards, announced September 1, keeps monitoring data in your own Amazon S3, Azure Blob, or Google Cloud Storage account under your encryption keys and audit logging, and routes detection signals to your reviewers rather than Anthropic's. Anthropic says more than 100 customers shaped it, including the CISO group behind the largest US banks plus Comcast, KPMG, Mastercard, Salesforce, and Visa. That answers the 30-day retention policy that has kept legal, healthcare, and finance teams off Fable-class models since Fable 5. Rollout is phased and starts later this fall, so today this is a procurement conversation, not a toggle you flip. Anthropic's writeup.
|
Claude Code 2.1.260 makes prompt-cache misses visible
Three changes shipped September 3, per the changelog: a new /diff command opens a side-by-side panel while Claude is still mid-edit, /cost now names the probable cause of a cache miss (changed tool definitions, edited system prompt, idle past TTL), and org policy state surfaces in /status and claude doctor. If your agent-loop spend swings between otherwise identical runs, the cache-miss line is the first place to look. The policy line kills a whole category of "why is this blocked" tickets.
|
The DOJ told a federal court that AI training on copyrighted text is fair use
Justice filed a statement of interest on September 1 in the Southern District of New York multidistrict litigation against OpenAI, arguing that a contrary ruling would distort copyright law and weaken US competitiveness and national security, as reported by IPWatchdog on September 3. First time the federal government has taken a side in training-data litigation. It binds no judge, but if you are building on a model trained on scraped text, the downside scenario just got less likely and your indemnity clause just got less load-bearing.
|
|
•
|
Four Copilot models retire October 2. Gemini 3.5 Flash, Gemini 3.6 Flash, Kimi K2.7 Code, and Claude Opus 4.7 stop serving every Copilot surface on October 2 - that is 24 days out. Replacements are named but not automatic: a Business or Enterprise admin has to enable each one in model policy before the cutoff or pinned workflows go dark. Replacement table →
|
|
•
|
Firecrawl. One API call turns a site into clean markdown a model can read - the unglamorous half of every RAG index and research agent you will build this quarter. Try it →
|
|
•
|
Make. Visual automation for wiring model calls into the systems that already run your business, without standing up and babysitting a service to do it. Try it →
|
|
|
|
|
Jonathan Hildebrandt
Co-founder and primary operator of Pondero. Writes the Pondero Brief.
|
|
|
Affiliate disclosure
·
Unsubscribe
·
Manage preferences
Pondero earns commissions on some links. This does not affect our editorial picks.
|
|