Emergence World: Adversarial Stress-Testing of Long-Horizon Multi-Agent Systems Paper • 2609.17320 • Published 5 days ago • 2 • 3
No Free Checker: A Survey of Verifiers for Robot Policies Paper • 2609.09250 • Published 12 days ago • 1 • 2
MANIGUARD: A Benchmark and Data Suite for Specification-Grounded Safety Evaluation and Improvement of Robotic Manipulation Paper • 2608.17386 • Published Aug 18 • 2 • 1
Vision-Language-Action Safety: Threats, Challenges, Evaluations, and Mechanisms Paper • 2604.23775 • Published Apr 26 • 45 • 3
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 264 • 6
Silent Failures in Physical AI: A Literature Review of Runtime Action Authorization for Autonomous Systems Paper • 2606.00090 • Published May 23 • 3 • 4
DRIFT: Derailing Denoising Trajectories of Flow-Matching VLAs with Adversarial Patch Attack Paper • 2608.03207 • Published Aug 4 • 4 • 3