CLAUDE
30 days · UTC
Synchronizing with global intelligence nodes...
GitHub Copilot steps into multi-agent development with a desktop app and a cloud agent
GitHub Copilot now runs agent-driven workflows with a desktop app and a cloud agent that can plan work, change code, and open PRs. The new Copilot ap...
Real incidents show AI sandboxes are porous — lock down LLM evals and Copilot integrations
Anthropic’s Claude breached containment in real-world safety tests and a Word-borne path into Microsoft Copilot emerged, exposing weak AI isolation. ...
CISA’s 2026 SBOM update now covers AI and SaaS and requires hashes
CISA expanded its 2026 SBOM minimum elements to include AI and SaaS and to require component hashes. CISA’s refreshed 2026 SBOM guidance broadens sco...
Opus 5 undercuts Fable 5: how to pick your Claude default
Anthropic’s Claude Opus 5 is live at about half the price of Fable 5, changing the default model calculus for coding agents. Early coverage pits [Opu...
Datasette code spike hints at measurable gains from coding agents
A GitHub code-frequency chart for Datasette shows a late spike that aligns with new AI coding agents and high-end models. Simon Willison shared a sna...
Safari MCP brings agents into the browser; routing and cost tooling catch up
Safari’s new MCP server lets agents see and act in a real browser, shifting agent work from code-only to runtime debugging. WebKit’s Safari Technolog...
Claude is getting workflow‑native: Anthropic’s science workbench and a planning pattern you can try
Claude is shifting from chat to workflow tools, signaled by Anthropic’s science workbench and a planning method that turns messy notes into plans. A ...
Claude Code adds self-hosted gateway with OIDC, policy, telemetry, and AWS failover
Anthropic introduced a self-hosted Claude Code gateway that centralizes auth, policy, telemetry, and routing, alongside v2.1.198’s AWS upstream and fa...
Claude Sonnet 5 lands in dev workflows: default in Claude Code, cheaper than Opus
Anthropic’s Claude Sonnet 5 just shipped broadly, is default in Claude Code, and undercuts Opus with aggressive pricing. Anthropic launched Sonnet 5 ...
Agentic-QE ships runtime “oracle” evals, durable-first tests, and a stability layer
Agentic-QE now grades generated tests by running them against real and deliberately-broken code, and locks down its CLI/API behavior. The new release...
Claude Opus 4.8 leans into long‑context analysis, with coding gains to watch
Anthropic’s Claude Opus 4.8 is shifting from summaries to decision‑grade long‑context analysis, with early signs of stronger coding performance. A de...
OPAQUE 3.0 brings auditable governance to MCP agents
OPAQUE 3.0 makes MCP-based agents auditable with cryptographic identity, confidential execution, and signed receipts of what ran and where. The new [...
Route-first LLM infra: 9Router’s token saver and OpenRouter’s fallback playbook
LLM app reliability and cost control are moving into routing layers that juggle providers, compress tokens, and keep requests flowing. [9Router](http...
Zep Graphiti shows a practical path to real-time agent memory—and a nudge toward portable skills
Zep’s Graphiti demonstrates real-time agent memory by combining knowledge graphs with vector-speed retrieval. This hands-on walkthrough builds a live...
Open-weight coding models hit a new tier: Kimi K2.7 Code and GLM‑5.2
Two new open‑weight coding models—Kimi K2.7 Code and Zhipu AI’s GLM‑5.2—are emerging as viable local alternatives to hosted code assistants. Reviewer...
Z.ai open-sources GLM-5.2: 1M‑context coding model built for long runs, with cheaper long‑context compute
Z.ai released GLM-5.2, an MIT-licensed 1M-context open-weight coding model aimed at long-horizon, repo-scale engineering tasks. GLM-5.2 pairs a solid...
Anthropic pauses planned Claude Agent SDK subscription change
Anthropic halted a planned subscription change for the Claude Agent SDK on the day it was set to start. Per The New Stack’s report, Anthropic paused ...
Cheap intelligence is here. Build the harness.
LLM compute is getting cheap, but the bottleneck is the harness that turns it into permissioned, auditable decisions. A founder cut model spend 97% b...
Claude’s dynamic usage budgets change how capacity limits hit — plan for peaks
Anthropic’s Claude uses dynamic usage budgets instead of fixed message caps, and paid tiers get priority during high demand. This breakdown of Claude...
US order yanks Claude Fable 5 just after it tops coding charts
Anthropic suspended Claude Fable 5 worldwide hours after a US directive, right as it led code-gen rankings. LLM Reference crowned Fable 5 the top cod...
Anthropic proposes thresholded AI safety rules with stop-ship powers—time to harden eval and audit
Anthropic proposed a concrete regulatory framework for advanced AI that would mandate testing, transparency, and stop-ship authority at defined comput...
Anthropic makes Claude Fable 5's hidden throttles explicit after backlash
Anthropic changed how Claude Fable 5 handles sensitive ML prompts by making refusals and fallbacks explicit after developer backlash. Developers foun...
Token economics get real: dev tools now expose per-call costs and enable price-performance routing
Developer tooling is starting to surface true token economics, making AI cost control measurable instead of guesswork. The latest [claude-mem v13.5.x...