AI Infrastructure
LiteRT-LM WebGPU brings local agents into the browser
Google LiteRT-LM expands Gemma 4 local inference across Android, iOS, WebGPU, and CLI, changing where AI apps can run.
Google LiteRT-LM expands Gemma 4 local inference across Android, iOS, WebGPU, and CLI, changing where AI apps can run.
Google and Arm show how on-device generative AI is moving from model releases into CPU runtimes, quantization, memory limits, and silicon features.
General Compute is making its ASIC-first inference cloud generally available, challenging GPU-centric serving for agent workloads.
Anthropic and AWS made Claude Platform on AWS generally available, giving AWS customers a native Claude agent stack alongside Bedrock.