Blog
Notes and analysis on AI development.
Claude subscriptions lose their free automation subsidy
Anthropic is moving Claude Agent SDK and claude -p usage into separate monthly credits on June 15, changing the economics of AI coding automation.
Codex Windows sandbox sets the baseline for local agent security
OpenAI’s Codex Windows sandbox design shows that local coding agent security is now an OS boundary problem, not only a model safety problem.
General Compute targets the GPU tax on agent inference
General Compute is making its ASIC-first inference cloud generally available, challenging GPU-centric serving for agent workloads.
OpenAI and Anthropic move the model war into deployment
OpenAI and Anthropic are building enterprise AI deployment firms. The bottleneck is moving from APIs to FDEs, integration, and governance.
Notion Workers turn the workspace into an agent runtime
Notion introduced a Developer Platform with Workers, External Agents API, Agent SDK, and CLI, moving its workspace toward AI agent runtime infrastructure.
Thinking Machines Makes AI Collaboration Real Time
Thinking Machines Interaction Models proposes full-duplex collaboration where AI can listen, see, speak, and use tools at the same time.
Claude Agent SDK credits move agent automation into metered budgets
Anthropic is separating Claude Agent SDK usage into monthly credits, making coding-agent automation a budgeted workflow rather than plain subscription usage.
Claude moves into the small business back office
Claude for Small Business packages QuickBooks, PayPal, HubSpot, Canva, and other tools into agentic workflows for SMB operations.
Baidu DAA puts a new metric on the agent era
Baidu proposed Daily Active Agents at Create 2026. The platform race is moving from token consumption toward agents that actually complete work.
Needle brings tool calling down to a 26M on-device model
Cactus Compute Needle is a 26M-parameter local model for tool calling, a small experiment that changes how agent latency, cost, and privacy should be designed.
Honeycomb turns AI agent black boxes into timelines
Honeycomb Agent Observability shows that the production bottleneck for AI agents is moving from model-call logs to handoffs, tool calls, costs, and failure reconstruction.
NVIDIA Is Targeting the RL Training Loop
NVIDIA and Ineffable Intelligence are pointing the model race toward RL infrastructure for agents that learn from experience.