2026-06-15

AI Agents

Today's work zeroes in on how agents fail. We're seeing a wave of new benchmarks designed to find subtle but critical vulnerabilities: decomposition attacks, social engineering in PRs, latent planning errors, and even denial-of-service against the guardrails themselves. This shift from capability to reliability is driving more robust architectural patterns, from privacy-preserving UI brokers to specialized sub-agents for efficiency.

15 papers 2 news 0 blogs 8 appendix 489 considered

Papers

15

HarnessX: A Composable, Adaptive, and Evolvable Agent Harness Foundry

arxiv:arxiv-agents — Tingyang Chen, Shuo Lu, Kang Zhao, Weicheng Meng, Hanlin Teng, Tianhao Li, Chao Li, Xule Liu, Jian Liang, Zhizhong Zhang, Yuan Xie, Heng Qu, Kun Shao, Jian Luan 10/10 frameworksevalsresearch
HarnessX: A Composable, Adaptive, and Evolvable Agent Harness Foundry

News

2

Why Is Claude Turning into an a**Hole?

hn:hn-claude — drob518 9/10 safetyobservability
A practitioner's first-hand account of Claude models becoming overly-censorious and unhelpful, even for benign prompts. This isn't just a complaint; it's a documented case of model degradation or a safety-alignment change with real-world consequences for developers relying on model stability. The post catalogs specific examples of previously working prompts now failing, highlighting the operational risk of undocumented behavioral drift in foundation models.
8 more items the ranker flagged but didn't feature