LangMARL: Natural Language Multi-Agent Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Yao, Huaiyuan, Da, Longchao, Liu, Xiaoou, Fleming, Charles, Chen, Tianlong, Wei, Hua |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
por: Chen, Tiejin, et al.
Publicado: (2026)
por: Chen, Tiejin, et al.
Publicado: (2026)
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology
por: Da, Longchao, et al.
Publicado: (2025)
por: Da, Longchao, et al.
Publicado: (2025)
CoMAL: Collaborative Multi-Agent Large Language Models for Mixed-Autonomy Traffic
por: Yao, Huaiyuan, et al.
Publicado: (2024)
por: Yao, Huaiyuan, et al.
Publicado: (2024)
Do LLMs have a Gender (Entropy) Bias?
por: Prabhune, Sonal, et al.
Publicado: (2025)
por: Prabhune, Sonal, et al.
Publicado: (2025)
Black-box Context-free Grammar Inference for Readable & Natural Grammars
por: Arefin, Mohammad Rifat, et al.
Publicado: (2025)
por: Arefin, Mohammad Rifat, et al.
Publicado: (2025)
Multi-Agent-as-Judge: Aligning LLM-Agent-Based Automated Evaluation with Multi-Dimensional Human Evaluation
por: Chen, Jiaju, et al.
Publicado: (2025)
por: Chen, Jiaju, et al.
Publicado: (2025)
Drama Engine: A Framework for Narrative Agents
por: Pichlmair, Martin, et al.
Publicado: (2024)
por: Pichlmair, Martin, et al.
Publicado: (2024)
A Multi-Agent Framework for Medical AI: Leveraging Fine-Tuned GPT, LLaMA, and DeepSeek R1 for Evidence-Based and Bias-Aware Clinical Query Processing
por: Nourmohammadi, Naeimeh, et al.
Publicado: (2026)
por: Nourmohammadi, Naeimeh, et al.
Publicado: (2026)
Communicative Agents for Slideshow Storytelling Video Generation based on LLMs
por: Fan, Jingxing, et al.
Publicado: (2025)
por: Fan, Jingxing, et al.
Publicado: (2025)
GE-Chat: A Graph Enhanced RAG Framework for Evidential Response Generation of LLMs
por: Da, Longchao, et al.
Publicado: (2025)
por: Da, Longchao, et al.
Publicado: (2025)
SAGED: A Holistic Bias-Benchmarking Pipeline for Language Models with Customisable Fairness Calibration
por: Guan, Xin, et al.
Publicado: (2024)
por: Guan, Xin, et al.
Publicado: (2024)
SmartWalkCoach: An AI Companion for End-to-End Walking Guidance, Motivation, and Reflection
por: Zhang, Xianzhe, et al.
Publicado: (2026)
por: Zhang, Xianzhe, et al.
Publicado: (2026)
FUAS-Agents: Autonomous Multi-Modal LLM Agents for Treatment Planning in Focused Ultrasound Ablation Surgery
por: Zhao, Lina, et al.
Publicado: (2025)
por: Zhao, Lina, et al.
Publicado: (2025)
AgentBreeder: Mitigating the AI Safety Risks of Multi-Agent Scaffolds via Self-Improvement
por: Rosser, J, et al.
Publicado: (2025)
por: Rosser, J, et al.
Publicado: (2025)
Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution
por: Liu, Xiaoou, et al.
Publicado: (2026)
por: Liu, Xiaoou, et al.
Publicado: (2026)
MATRIX: Multi-Agent simulaTion fRamework for safe Interactions and conteXtual clinical conversational evaluation
por: Lim, Ernest, et al.
Publicado: (2025)
por: Lim, Ernest, et al.
Publicado: (2025)
Universal Conditional Logic: A Formal Language for Prompt Engineering
por: Mikinka, Anthony
Publicado: (2025)
por: Mikinka, Anthony
Publicado: (2025)
A Framework for Scalable Heterogeneous Multi-Agent Adversarial Reinforcement Learning in IsaacLab
por: Peterson, Isaac, et al.
Publicado: (2025)
por: Peterson, Isaac, et al.
Publicado: (2025)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
por: Jia, Xiao
Publicado: (2026)
por: Jia, Xiao
Publicado: (2026)
Tool Receipts, Not Zero-Knowledge Proofs: Practical Hallucination Detection for AI Agents
por: Basu, Abhinaba
Publicado: (2026)
por: Basu, Abhinaba
Publicado: (2026)
Tatarstan Toponyms: A Bilingual Dataset and Hybrid RAG System for Geospatial Question Answering
por: Arabov, Mullosharaf K.
Publicado: (2026)
por: Arabov, Mullosharaf K.
Publicado: (2026)
Sparse Autoencoders Can Capture Language-Specific Concepts Across Diverse Languages
por: Andrylie, Lyzander Marciano, et al.
Publicado: (2025)
por: Andrylie, Lyzander Marciano, et al.
Publicado: (2025)
LLMs as Deceptive Agents: How Role-Based Prompting Induces Semantic Ambiguity in Puzzle Tasks
por: Yoo, Seunghyun
Publicado: (2025)
por: Yoo, Seunghyun
Publicado: (2025)
Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach
por: Liu, Linyu, et al.
Publicado: (2024)
por: Liu, Linyu, et al.
Publicado: (2024)
Who's Asking? Investigating Bias Through the Lens of Disability Framed Queries in LLMs
por: Hari, Vishnu, et al.
Publicado: (2025)
por: Hari, Vishnu, et al.
Publicado: (2025)
HInter: Exposing Hidden Intersectional Bias in Large Language Models
por: Souani, Badr, et al.
Publicado: (2025)
por: Souani, Badr, et al.
Publicado: (2025)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
por: Huo, Dongjie, et al.
Publicado: (2026)
por: Huo, Dongjie, et al.
Publicado: (2026)
Rainbow Delay Compensation: A Multi-Agent Reinforcement Learning Framework for Mitigating Delayed Observation
por: Fu, Songchen, et al.
Publicado: (2025)
por: Fu, Songchen, et al.
Publicado: (2025)
COPAL-ID: Indonesian Language Reasoning with Local Culture and Nuances
por: Wibowo, Haryo Akbarianto, et al.
Publicado: (2023)
por: Wibowo, Haryo Akbarianto, et al.
Publicado: (2023)
Knowledge Editing for Large Language Model with Knowledge Neuronal Ensemble
por: Li, Yongchang, et al.
Publicado: (2024)
por: Li, Yongchang, et al.
Publicado: (2024)
Knesset-DictaBERT: A Hebrew Language Model for Parliamentary Proceedings
por: Goldin, Gili, et al.
Publicado: (2024)
por: Goldin, Gili, et al.
Publicado: (2024)
Fishing for Magikarp: Automatically Detecting Under-trained Tokens in Large Language Models
por: Land, Sander, et al.
Publicado: (2024)
por: Land, Sander, et al.
Publicado: (2024)
A Survey on Hypothesis Generation for Scientific Discovery in the Era of Large Language Models
por: Alkan, Atilla Kaan, et al.
Publicado: (2025)
por: Alkan, Atilla Kaan, et al.
Publicado: (2025)
Math Natural Language Inference: this should be easy!
por: de Paiva, Valeria, et al.
Publicado: (2025)
por: de Paiva, Valeria, et al.
Publicado: (2025)
Bridging the Language Gap: Enhancing Multilingual Prompt-Based Code Generation in LLMs via Zero-Shot Cross-Lingual Transfer
por: Li, Mingda, et al.
Publicado: (2024)
por: Li, Mingda, et al.
Publicado: (2024)
Surfing the modeling of PoS taggers in low-resource scenarios
por: Ferro, Manuel Vilares, et al.
Publicado: (2024)
por: Ferro, Manuel Vilares, et al.
Publicado: (2024)
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
por: Wang, Youkang, et al.
Publicado: (2025)
por: Wang, Youkang, et al.
Publicado: (2025)
A Legal Framework for Natural Language Processing Model Training in Portugal
por: Almeida, Rúben, et al.
Publicado: (2024)
por: Almeida, Rúben, et al.
Publicado: (2024)
A Survey on Natural Language Counterfactual Generation
por: Wang, Yongjie, et al.
Publicado: (2024)
por: Wang, Yongjie, et al.
Publicado: (2024)
Attention is also needed for form design
por: Sankar, B., et al.
Publicado: (2025)
por: Sankar, B., et al.
Publicado: (2025)
Ejemplares similares
-
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
por: Chen, Tiejin, et al.
Publicado: (2026) -
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology
por: Da, Longchao, et al.
Publicado: (2025) -
CoMAL: Collaborative Multi-Agent Large Language Models for Mixed-Autonomy Traffic
por: Yao, Huaiyuan, et al.
Publicado: (2024) -
Do LLMs have a Gender (Entropy) Bias?
por: Prabhune, Sonal, et al.
Publicado: (2025) -
Black-box Context-free Grammar Inference for Readable & Natural Grammars
por: Arefin, Mohammad Rifat, et al.
Publicado: (2025)