Devlery - AI news for builders
Devlery blog
AI news for builders.
99.82% Cache Hits, the New Variable in Coding Agent Costs
The Reasonix debate shows that coding agent costs depend not only on model pricing, but on harness design that keeps prefix cache intact.
20 missed attacks, SLEIGHT-Bench warns agent security teams
SLEIGHT-Bench uses 40 synthetic attacks to show how easily LLM monitors can miss risky behavior by coding agents.
The coding agent market gets a new enterprise scorecard
Gartner and OpenAI show how AI coding agents are moving from model benchmarks toward governance, sandboxing, auditability, and cost control.
Copilot Remote Control Sets New Rules for Agent Costs
GitHub Copilot CLI remote control and automatic model routing turn coding agents from IDE helpers into long-running sessions that teams must operate.
x402 hits $24.24M, and agent wallets now need accountability
Circle Agent Stack opens a path for AI agents to pay for APIs and services with USDC, while exposing sharper responsibility and security boundaries.
15B tokens per minute, the real cost of OpenAI’s superapp strategy
OpenAI’s $122B raise is not just financing. It ties ChatGPT, Codex, API usage, enterprise adoption, and compute into one agent-first flywheel.
Two H100s Are Enough, Command A+ Targets Private Agents
Cohere Command A+ lowers the bar for private AI agents with Apache 2.0 open weights and a two-H100 deployment target.
Android Skills teach agents the rules of app development
Android Studio I/O Edition brings Agent Skills and Android CLI into the mobile workflow, turning Android development into an agent-readable knowledge layer.
Gemini CLI June 18 cutoff and the price of open source agents
Google is moving individual Gemini CLI users to Antigravity CLI. A 100K-star open source agent is being absorbed into a broader platform.
Anthropic Acquires Stainless and Takes Hold of Agent Plumbing
Anthropic’s Stainless acquisition shows why SDKs, CLIs, and MCP server generation have become strategic plumbing for AI agent platforms.
4 million Codex users, and the on-prem condition for coding agents
OpenAI and Dell’s Codex partnership shows enterprise coding-agent competition moving from model quality to internal context, governance, and deployment boundaries.
Fifty Researchers Tested Co-Scientist, and Hypothesis Ranking Changed
Google Co-Scientist and Gemini for Science shift AI research tools from answer generation toward hypothesis loops that humans can test.