Attributing Emergence in Million-Agent Systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Tang, Ling, Mei, Jilin, Chen, Qian, Ren, Qihan, Zhang, Linfeng, Zhang, Quanshi, Shao, Jing, Hu, Xia, Liu, Dongrui |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Interpreting Emergent Extreme Events in Multi-Agent Systems
di: Tang, Ling, et al.
Pubblicazione: (2026)
di: Tang, Ling, et al.
Pubblicazione: (2026)
What Do EEG Foundation Models Capture from Human Brain Signals?
di: Tang, Ling, et al.
Pubblicazione: (2026)
di: Tang, Ling, et al.
Pubblicazione: (2026)
The Why Behind the Action: Unveiling Internal Drivers via Agentic Attribution
di: Qian, Chen, et al.
Pubblicazione: (2026)
di: Qian, Chen, et al.
Pubblicazione: (2026)
Rethinking Generalization in Reasoning SFT: A Conditional Analysis on Optimization, Data, and Model Capability
di: Ren, Qihan, et al.
Pubblicazione: (2026)
di: Ren, Qihan, et al.
Pubblicazione: (2026)
Where We Have Arrived in Proving the Emergence of Sparse Symbolic Concepts in AI Models
di: Ren, Qihan, et al.
Pubblicazione: (2023)
di: Ren, Qihan, et al.
Pubblicazione: (2023)
Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents
di: Shao, Shuai, et al.
Pubblicazione: (2025)
di: Shao, Shuai, et al.
Pubblicazione: (2025)
Towards the Dynamics of a DNN Learning Symbolic Interactions
di: Ren, Qihan, et al.
Pubblicazione: (2024)
di: Ren, Qihan, et al.
Pubblicazione: (2024)
Towards Self-Evolving Benchmarks: Synthesizing Agent Trajectories via Test-Time Exploration under Validate-by-Reproduce Paradigm
di: Guo, Dadi, et al.
Pubblicazione: (2025)
di: Guo, Dadi, et al.
Pubblicazione: (2025)
TradeTrap: Are LLM-based Trading Agents Truly Reliable and Faithful?
di: Yan, Lewen, et al.
Pubblicazione: (2025)
di: Yan, Lewen, et al.
Pubblicazione: (2025)
Entropy-Gradient Inversion: Moving Toward Internal Mechanism of Large Reasoning Models
di: Yang, Junyao, et al.
Pubblicazione: (2026)
di: Yang, Junyao, et al.
Pubblicazione: (2026)
LED-Merging: Mitigating Safety-Utility Conflicts in Model Merging with Location-Election-Disjoint
di: Ma, Qianli, et al.
Pubblicazione: (2025)
di: Ma, Qianli, et al.
Pubblicazione: (2025)
Revisiting Generalization Power of a DNN in Terms of Symbolic Interactions
di: Cheng, Lei, et al.
Pubblicazione: (2025)
di: Cheng, Lei, et al.
Pubblicazione: (2025)
ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis
di: Li, Yu, et al.
Pubblicazione: (2026)
di: Li, Yu, et al.
Pubblicazione: (2026)
REEF: Representation Encoding Fingerprints for Large Language Models
di: Zhang, Jie, et al.
Pubblicazione: (2024)
di: Zhang, Jie, et al.
Pubblicazione: (2024)
The Tug of War Within: Mitigating the Fairness-Privacy Conflicts in Large Language Models
di: Qian, Chen, et al.
Pubblicazione: (2024)
di: Qian, Chen, et al.
Pubblicazione: (2024)
RiOSWorld: Benchmarking the Risk of Multimodal Computer-Use Agents
di: Yang, Jingyi, et al.
Pubblicazione: (2025)
di: Yang, Jingyi, et al.
Pubblicazione: (2025)
Evaluating the Correctness of Inference Patterns Used by LLMs for Judgment
di: Chen, Lu, et al.
Pubblicazione: (2024)
di: Chen, Lu, et al.
Pubblicazione: (2024)
Conditional Advantage Estimation for Reinforcement Learning in Large Reasoning Models
di: Chen, Guanxu, et al.
Pubblicazione: (2025)
di: Chen, Guanxu, et al.
Pubblicazione: (2025)
Are Your Agents Upward Deceivers?
di: Guo, Dadi, et al.
Pubblicazione: (2025)
di: Guo, Dadi, et al.
Pubblicazione: (2025)
Identifying Semantic Induction Heads to Understand In-Context Learning
di: Ren, Jie, et al.
Pubblicazione: (2024)
di: Ren, Jie, et al.
Pubblicazione: (2024)
Towards Attributions of Input Variables in a Coalition
di: Zheng, Xinhao, et al.
Pubblicazione: (2023)
di: Zheng, Xinhao, et al.
Pubblicazione: (2023)
The Interaction Bottleneck of Deep Neural Networks: Discovery, Proof, and Modulation
di: Deng, Huiqi, et al.
Pubblicazione: (2025)
di: Deng, Huiqi, et al.
Pubblicazione: (2025)
Attribution Explanations for Deep Neural Networks: A Theoretical Perspective
di: Deng, Huiqi, et al.
Pubblicazione: (2025)
di: Deng, Huiqi, et al.
Pubblicazione: (2025)
Loop as a Bridge: Can Looped Transformers Truly Link Representation Space and Natural Language Outputs?
di: Chen, Guanxu, et al.
Pubblicazione: (2026)
di: Chen, Guanxu, et al.
Pubblicazione: (2026)
AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security
di: Liu, Dongrui, et al.
Pubblicazione: (2026)
di: Liu, Dongrui, et al.
Pubblicazione: (2026)
INFA-Guard: Mitigating Malicious Propagation via Infection-Aware Safeguarding in LLM-Based Multi-Agent Systems
di: Zhou, Yijin, et al.
Pubblicazione: (2026)
di: Zhou, Yijin, et al.
Pubblicazione: (2026)
Towards the Resistance of Neural Network Watermarking to Fine-tuning
di: Tang, Ling, et al.
Pubblicazione: (2025)
di: Tang, Ling, et al.
Pubblicazione: (2025)
COLLEAGUE.SKILL: Automated AI Skill Generation via Expert Knowledge Distillation
di: Zhou, Tianyi, et al.
Pubblicazione: (2026)
di: Zhou, Tianyi, et al.
Pubblicazione: (2026)
LPS-Bench: Benchmarking Safety Awareness of Computer-Use Agents in Long-Horizon Planning under Benign and Adversarial Scenarios
di: Chen, Tianyu, et al.
Pubblicazione: (2026)
di: Chen, Tianyu, et al.
Pubblicazione: (2026)
ReasonAny: Incorporating Reasoning Capability to Any Model via Simple and Effective Model Merging
di: Yang, Junyao, et al.
Pubblicazione: (2026)
di: Yang, Junyao, et al.
Pubblicazione: (2026)
Demystifying Reasoning Dynamics with Mutual Information: Thinking Tokens are Information Peaks in LLM Reasoning
di: Qian, Chen, et al.
Pubblicazione: (2025)
di: Qian, Chen, et al.
Pubblicazione: (2025)
RvB: Automating AI System Hardening via Iterative Red-Blue Games
di: Huang, Lige, et al.
Pubblicazione: (2026)
di: Huang, Lige, et al.
Pubblicazione: (2026)
Explaining Generalization Power of a DNN Using Interactive Concepts
di: Zhou, Huilin, et al.
Pubblicazione: (2023)
di: Zhou, Huilin, et al.
Pubblicazione: (2023)
RetroAgent: From Solving to Evolving via Retrospective Dual Intrinsic Feedback
di: Zhang, Xiaoying, et al.
Pubblicazione: (2026)
di: Zhang, Xiaoying, et al.
Pubblicazione: (2026)
Towards Tracing Trustworthiness Dynamics: Revisiting Pre-training Period of Large Language Models
di: Qian, Chen, et al.
Pubblicazione: (2024)
di: Qian, Chen, et al.
Pubblicazione: (2024)
Does a Neural Network Really Encode Symbolic Concepts?
di: Li, Mingjie, et al.
Pubblicazione: (2023)
di: Li, Mingjie, et al.
Pubblicazione: (2023)
Technical Note: Defining and Quantifying AND-OR Interactions for Faithful and Concise Explanation of DNNs
di: Li, Mingjie, et al.
Pubblicazione: (2023)
di: Li, Mingjie, et al.
Pubblicazione: (2023)
Decouple-Then-Merge: Finetune Diffusion Models as Multi-Task Learning
di: Ma, Qianli, et al.
Pubblicazione: (2024)
di: Ma, Qianli, et al.
Pubblicazione: (2024)
PACEbench: A Framework for Evaluating Practical AI Cyber-Exploitation Capabilities
di: Liu, Zicheng, et al.
Pubblicazione: (2025)
di: Liu, Zicheng, et al.
Pubblicazione: (2025)
Alita: Generalist Agent Enabling Scalable Agentic Reasoning with Minimal Predefinition and Maximal Self-Evolution
di: Qiu, Jiahao, et al.
Pubblicazione: (2025)
di: Qiu, Jiahao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Interpreting Emergent Extreme Events in Multi-Agent Systems
di: Tang, Ling, et al.
Pubblicazione: (2026) -
What Do EEG Foundation Models Capture from Human Brain Signals?
di: Tang, Ling, et al.
Pubblicazione: (2026) -
The Why Behind the Action: Unveiling Internal Drivers via Agentic Attribution
di: Qian, Chen, et al.
Pubblicazione: (2026) -
Rethinking Generalization in Reasoning SFT: A Conditional Analysis on Optimization, Data, and Model Capability
di: Ren, Qihan, et al.
Pubblicazione: (2026) -
Where We Have Arrived in Proving the Emergence of Sparse Symbolic Concepts in AI Models
di: Ren, Qihan, et al.
Pubblicazione: (2023)