Papers

Filtered to multi-agent systems · clear filter

Browse by term

continual learning 33reinforcement learning 21large language models 9vision-language models 8language models 6benchmarking 5autoregressive models 4diffusion transformers 4generative models 4multimodal learning 4robotics 4video generation 4benchmarks 3computer vision 3policy optimization 3self-distillation 3vision-language-action models 3world modeling 3agent-based systems 2autonomous agents 2calibration 2coding agents 2diffusion models 2foundation models 2image synthesis 2in-context learning 2knowledge graphs 2LLMs 2multimodal evaluation 2multimodal large language models 2

Matching papers

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation

2 upvotes · 20 JUL 2026 · Jasmine Brazilek, Maheep Chaudhary, Zoe Lu et al.

This paper creates a benchmark to test how AI systems respond when an authority figure tries to get an unwilling subordinate to complete a task, and it explores how different levels of authority affect the outcome. Practitioners might care because it helps them understand and manage the complex interactions between AI systems.