Tag: #browser-agents
Writing on GPUs, LLMs, MLOps, Kubernetes — and mindset · 2 posts
Browser and Computer-Use Agents: Where They Actually Stand, and What the Benchmarks Really Measure
"A computer-use agent hit 83.5% on OSWorld" and "even the strongest agent finishes only 20.6%" are both facts published in 2026, and both are true. The first is OSWorld 1.0; the second is OSWorld 2.0 from the same team.
2026-07-17 · 24 min read #ai#computer-use#browser-agents#benchmark#prompt-injectionWeb Agents Read Page Text as Commands — Cross-Site Prompt Injection and Prismata's Confinement
Prismata, posted to arXiv on July 9, 2026 by Villa, Ozdarendeli, Tan, and Popa, tackles the oldest weakness of autonomous web agents. Because agents interpret natural language as instructions, third-party and user-genera
2026-07-11 · 5 min read #prompt-injection#web-agents#ai-security#least-privilege#browser-agents