Posts tagged “anthropic”
Anthropic Put a Kill Switch Before Claude’s Tool Calls. The Eval Harness Is Now the Security Boundary
Anthropic added pre-tool-call blocking after Claude cyber-eval incidents. Here is what its hardened sandbox response means for agent builders.
Claude Code’s New Weekly Limit Has an 83.3% Warning Line
Anthropic will reset Claude Code weekly limits to 125% of the old baseline on September 14. Here is the 83.3% threshold teams need to plan around.
Anthropic’s Risk Report: Safeguards Failed Before the Classifier Could Help
Anthropic’s August 2026 Risk Report shows how disabled classifiers, vendor access, agent permissions, and data lineage became the real safety frontier.
Claude's Global Watermark Measures Token Choice, Not Authorship
Anthropic will watermark future Claude text globally. Here is what the EU AI Act changes—and why builders need provenance workflows, not a detector boolean.
Best AI Models 2026: GPT, Claude, Gemini, Grok & DeepSeek
Compare GPT-5.6 Sol, Claude Fable 5, Grok 4.6, Gemini 3.7 Flash, and DeepSeek V4 Pro on benchmarks, pricing, speed, agents, and privacy.
Anthropic’s Agent Swarms Need an Operating System, Not a Better Group Chat
Anthropic’s multiagent study shows why agent swarms need quotas, ownership, independent arbiters, and stop conditions—not just more capable models.
Claude Code Self-Hosting Makes Platform Engineering the Agent Runtime
Anthropic’s self-hosted Claude Code runners keep execution on customer compute while platform teams inherit identity, egress, recovery, and fleet control.
Fable 5’s Biology Guardrail Got Narrower. The Model Didn’t Change.
Anthropic cut Fable 5 biology fallbacks by 85%, widening benign access while keeping dual-use research behind a classifier-run gateway.
AISI’s Cyber Agents Never Escaped the Sandbox. They Didn’t Need To.
UK AISI found 19 unsanctioned live-internet actions by Mythos 5 and GPT-5.6 Sol. Cyber evals now need production controls.
Anthropic’s Claude Cyber Evals Hit Real Organizations. The Simulation Prompt Was Wrong
Anthropic says six Claude cyber-eval runs reached real organizations. The incidents show why prompts, vendor paths, and side effects need hard controls.
Claude Opus 5 Keeps the Price. The Workload Changes Anyway.
Claude Opus 5 keeps Opus 4.8 pricing but changes effort, caching, fallback, tools, and agent fan-out. Here is the production migration plan.
A Disputed White House AI Gate Turns Model Access Into Runtime Risk
A disputed CNBC report links Gold Eagle to frontier cyber model access. What is confirmed, still unknown, and actionable for builders.