Posts tagged “agents”
Tencent Hy4 Preview: Why a 770B Open-Weight Agent Model Is Easier to Rent Than Run
Tencent’s Hy4 Preview offers 1M context, Apache-2.0 weights, and cheap agent APIs. Its real test is retrieval, caching, and route economics.
OpenAI's Hugging Face Incident Turned 1,200 Sandboxes Into One System
OpenAI's Hugging Face report shows how 1,200 nominally isolated agents turned shared caches, credentials, and retries into one attack system.
Claude Code’s New Weekly Limit Has an 83.3% Warning Line
Anthropic will reset Claude Code weekly limits to 125% of the old baseline on September 14. Here is the 83.3% threshold teams need to plan around.
Anthropic’s Risk Report: Safeguards Failed Before the Classifier Could Help
Anthropic’s August 2026 Risk Report shows how disabled classifiers, vendor access, agent permissions, and data lineage became the real safety frontier.
Gemini 3.7 Flash: Benchmarks, Pricing, and API Guide
Gemini 3.7 Flash is live. See benchmarks, API pricing, the Gemini 3.6 comparison, thinking levels, multimodal limits, and a migration guide.
Anthropic’s Agent Swarms Need an Operating System, Not a Better Group Chat
Anthropic’s multiagent study shows why agent swarms need quotas, ownership, independent arbiters, and stop conditions—not just more capable models.
LFM2.5-VL-3B Gives the Edge Eyes. Keep Its Hands Tied.
Liquid AI’s LFM2.5-VL-3B brings screen grounding, OCR, and tool calling to local devices—but builders should separate sight from action.
xAI's Grok Bot Gives Your AI Team One Computer—and One Blast Radius
xAI’s Grok Bot gives multiple AI teammates one persistent cloud computer. Shared state makes handoffs easy—and creates one shared blast radius.
Nemotron 3.5 Lightning Turns Agent Speed Into a Routing Problem
NVIDIA Nemotron 3.5 Lightning is a fast open agent worker. Its real value lies in routing, validation, and deployment economics.
OpenAI’s Agents SDK Turns a Package Upgrade Into a Runtime Migration
OpenAI’s Agents SDK defaults to GPT-5.6 Luna, negotiates MCP v2, and hardens durable runs. Here’s the production migration builders need.
GPT-5.6-Cyber Moves the Refusal Boundary Into the Security Stack
OpenAI's Daybreak Red gives trusted defenders GPT-5.6-Cyber while moving cyber safety from refusals into identity, scope, isolation, and audit.
Meta Muse Glimmer Fits a 30B Agent Stack on 24 GB—With Fine Print
Meta’s open-weight Muse Glimmer 30B targets 24 GB hardware with vision, long context, tools, and DFlash. Here is what builders should test first.