Lyft AI Assist cuts support-agent development from six months to two weeks
Lyft showed how LangGraph and LangSmith turned customer-support agents into a self-serve platform with routing, state, evals, and prompt CI.
Lyft showed how LangGraph and LangSmith turned customer-support agents into a self-serve platform with routing, state, evals, and prompt CI.
Hugging Face explains how retokenizing tool-using agent rollouts can break gradients, and proposes TITO as a safer training-loop rule.
Anthropic added self-hosted sandboxes and MCP tunnels to Claude Managed Agents, shifting tool execution and private tool access into enterprise-controlled boundaries.
Workday and Google Cloud connected Sana to Gemini Enterprise. For HR and finance agents, approval chains, permissions, and data boundaries matter more than the model.
Robinhood opened Trading MCP and Banking MCP for AI agents. The real developer story is the permission, approval, and liability model around financial tool calls.
Ollama now supports OpenJarvis v1.0. The release shows how local personal AI changes cost, latency, and data boundaries.
jqwik 1.10.0 adds an AI-agent-facing test log message, raising new questions about stdout, prompt injection, and coding agent trust boundaries.
Fujitsu self-evolving multi-AI agents show how enterprise LLMs may keep improving through verified feedback, design-search loops, and operating controls.
React Doctor adds a post-generation audit loop for React code written by coding agents, scanning state, effects, performance, security, and accessibility.
OpenAI and Warp show that the coding-agent race is shifting from code generation to open-source verification, observability, and agent orchestration.
Google Managed Agents extends the Gemini API from model calls into sandboxed execution, shifting where agent infrastructure begins.
OpenAI Agents SDK memory and the AMP v0.1 draft turn long-term agent memory into files, Git history, MCP resources, and auditable state.