Posts tagged “coding”
Grok 4.7 Is Public. Test the State Contract, Not the 500K Window
Grok 4.7 keeps 4.6’s 500K context and pricing. Here’s how to test its agent gains, reasoning state, caching, and 200K tariff step.
Gemini 3.7 Flash: Benchmarks, Pricing, and API Guide
Gemini 3.7 Flash is live. See benchmarks, API pricing, the Gemini 3.6 comparison, thinking levels, multimodal limits, and a migration guide.
DeepSeek V4-Flash-0731 Makes the Harness Part of the Model
DeepSeek V4-Flash-0731 adds major agent gains, Codex-ready Responses API support, and open weights. Here is how builders should evaluate it.
Inkling-Small Puts Open Weights Back in the Agent Router
Thinking Machines’ Inkling-Small pairs 12B-active MoE compute with open weights. Here is where it fits—and where it should escalate.
Claude Opus 5 Keeps the Price. The Workload Changes Anyway.
Claude Opus 5 keeps Opus 4.8 pricing but changes effort, caching, fallback, tools, and agent fan-out. Here is the production migration plan.
Meta Muse Spark 1.1: The Developer API Is the Real Launch
Meta's Muse Spark 1.1 opens a developer API for coding, agents, and multimodal apps. Here is why the launch matters for builders.
Tencent Hy3 Is a Product-First AI Model, Not Just Another Open-Weight Release
Tencent Hy3 turns Hunyuan into a production AI model for agents, coding, office work, and WeChat-scale deployment. Here is why it matters.