Tag: #on-device-ai
Writing on GPUs, LLMs, MLOps, Kubernetes — and mindset · 3 posts
llama.cpp Comes to the Browser — The Ceiling a WebGPU Backend Measured Across 16 Devices
LlamaWeb, the WebGPU backend for llama.cpp released by a UC Santa Cruz team in May 2026, left behind the widest dataset to date measuring browser LLM inference — 16 devices across 8 vendors. The results cut both ways. It
2026-07-16 · 16 min read #webgpu#llm-inference#llama-cpp#on-device-ai#browserWhy the Mac mini Became an On-Device AI Machine — What Apple's Silicon Exec Said, and What He Left Out
Apple Silicon senior product manager Doug Brooks talked to The Deep View about demand for the Mac mini and Mac Studio and where on-device AI is heading. Why developers and small teams reach for this little desktop as a l
2026-07-11 · 6 min read #apple-silicon#on-device-ai#local-llm#mac-mini#inferenceOn-Device AI 2026: Your Smartphone Becomes a Personal AI Server
On-device AI is the centerpiece of mobile innovation in 2026. With Apple Neural Engines, Snapdragon AI, and Google Tensor optimizations, smartphones now achieve both privacy protection and ultra-low latency while dramati
2026-03-16 · 9 min read #on-device-ai#edge-inference#apple-intelligence#privacy#mobile-ai