Blog
Notes and analysis on AI development.
OpenAI and Anthropic move the model war into deployment
OpenAI and Anthropic are building enterprise AI deployment firms. The bottleneck is moving from APIs to FDEs, integration, and governance.
Notion Workers turn the workspace into an agent runtime
Notion introduced a Developer Platform with Workers, External Agents API, Agent SDK, and CLI, moving its workspace toward AI agent runtime infrastructure.
Thinking Machines Makes AI Collaboration Real Time
Thinking Machines Interaction Models proposes full-duplex collaboration where AI can listen, see, speak, and use tools at the same time.
Claude Agent SDK credits move agent automation into metered budgets
Anthropic is separating Claude Agent SDK usage into monthly credits, making coding-agent automation a budgeted workflow rather than plain subscription usage.
Claude moves into the small business back office
Claude for Small Business packages QuickBooks, PayPal, HubSpot, Canva, and other tools into agentic workflows for SMB operations.
Baidu DAA puts a new metric on the agent era
Baidu proposed Daily Active Agents at Create 2026. The platform race is moving from token consumption toward agents that actually complete work.
Needle brings tool calling down to a 26M on-device model
Cactus Compute Needle is a 26M-parameter local model for tool calling, a small experiment that changes how agent latency, cost, and privacy should be designed.
Honeycomb turns AI agent black boxes into timelines
Honeycomb Agent Observability shows that the production bottleneck for AI agents is moving from model-call logs to handoffs, tool calls, costs, and failure reconstruction.
NVIDIA Is Targeting the RL Training Loop
NVIDIA and Ineffable Intelligence are pointing the model race toward RL infrastructure for agents that learn from experience.
Red Hat is turning Ansible into the agent execution layer
Red Hat Summit 2026 shows how enterprise AI agents may need execution, observability, sandboxing, and governance before they can touch infrastructure.
Frontier AI predeployment review is becoming the new launch gate
CAISI is expanding predeployment evaluation work with Google DeepMind, Microsoft, and xAI, moving frontier AI launches beyond public benchmarks.
Gemini API Webhooks turns AI work into an agent runtime
Google Gemini API Webhooks moves long-running AI jobs from polling loops into event-driven backend operations.