MALT: Improving Reasoning with Multi-Agent LLM Training
Fuente:
arXiv
Salvato in:
| Autori principali: | Motwani, Sumeet Ramesh, Smith, Chandler, Das, Rocktim Jyoti, Rafailov, Rafael, Laptev, Ivan, Torr, Philip H. S., Pizzati, Fabio, Clark, Ronald, de Witt, Christian Schroeder |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
BLAZER: Bootstrapping LLM-based Manipulation Agents with Zero-Shot Data Generation
di: Das, Rocktim Jyoti, et al.
Pubblicazione: (2025)
di: Das, Rocktim Jyoti, et al.
Pubblicazione: (2025)
h1: Bootstrapping LLMs to Reason over Longer Horizons via Reinforcement Learning
di: Motwani, Sumeet Ramesh, et al.
Pubblicazione: (2025)
di: Motwani, Sumeet Ramesh, et al.
Pubblicazione: (2025)
Secret Collusion among AI Agents: Multi-Agent Deception via Steganography
di: Motwani, Sumeet Ramesh, et al.
Pubblicazione: (2024)
di: Motwani, Sumeet Ramesh, et al.
Pubblicazione: (2024)
MALMM: Multi-Agent Large Language Models for Zero-Shot Robotics Manipulation
di: Singh, Harsh, et al.
Pubblicazione: (2024)
di: Singh, Harsh, et al.
Pubblicazione: (2024)
PSyDUCK: Training-Free Steganography for Latent Diffusion
di: Mahfuz, Aqib, et al.
Pubblicazione: (2025)
di: Mahfuz, Aqib, et al.
Pubblicazione: (2025)
LongCoT: Benchmarking Long-Horizon Chain-of-Thought Reasoning
di: Motwani, Sumeet Ramesh, et al.
Pubblicazione: (2026)
di: Motwani, Sumeet Ramesh, et al.
Pubblicazione: (2026)
Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
di: Putta, Pranav, et al.
Pubblicazione: (2024)
di: Putta, Pranav, et al.
Pubblicazione: (2024)
Unelicitable Backdoors in Language Models via Cryptographic Transformer Circuits
di: Draguns, Andis, et al.
Pubblicazione: (2024)
di: Draguns, Andis, et al.
Pubblicazione: (2024)
SLAT-Phys: Fast Material Property Field Prediction from Structured 3D Latents
di: Das, Rocktim Jyoti, et al.
Pubblicazione: (2026)
di: Das, Rocktim Jyoti, et al.
Pubblicazione: (2026)
Toward Robust Real-World Audio Deepfake Detection: Closing the Explainability Gap
di: Channing, Georgia, et al.
Pubblicazione: (2024)
di: Channing, Georgia, et al.
Pubblicazione: (2024)
MessyKitchens: Contact-rich object-level 3D scene reconstruction
di: Ansari, Junaid Ahmed, et al.
Pubblicazione: (2026)
di: Ansari, Junaid Ahmed, et al.
Pubblicazione: (2026)
LikePhys: Evaluating Intuitive Physics Understanding in Video Diffusion Models via Likelihood Preference
di: Yuan, Jianhao, et al.
Pubblicazione: (2025)
di: Yuan, Jianhao, et al.
Pubblicazione: (2025)
Towards Reliable Identification of Diffusion-based Image Manipulations
di: Costanzino, Alex, et al.
Pubblicazione: (2025)
di: Costanzino, Alex, et al.
Pubblicazione: (2025)
Specify and Edit: Overcoming Ambiguity in Text-Based Image Editing
di: Iakovleva, Ekaterina, et al.
Pubblicazione: (2024)
di: Iakovleva, Ekaterina, et al.
Pubblicazione: (2024)
AutoOR: Scalably Post-training LLMs to Autoformalize Operations Research Problems
di: Motwani, Sumeet Ramesh, et al.
Pubblicazione: (2026)
di: Motwani, Sumeet Ramesh, et al.
Pubblicazione: (2026)
Towards Understanding Multimodal Fine-Tuning: Spatial Features
di: Naghashyar, Lachin, et al.
Pubblicazione: (2026)
di: Naghashyar, Lachin, et al.
Pubblicazione: (2026)
ActCam: Zero-Shot Joint Camera and 3D Motion Control for Video Generation
di: Khalifi, Omar El, et al.
Pubblicazione: (2026)
di: Khalifi, Omar El, et al.
Pubblicazione: (2026)
Video Motion Transfer with Diffusion Transformers
di: Pondaven, Alexander, et al.
Pubblicazione: (2024)
di: Pondaven, Alexander, et al.
Pubblicazione: (2024)
Learning to Generate Rigid Body Interactions with Video Diffusion Models
di: Romero, David, et al.
Pubblicazione: (2025)
di: Romero, David, et al.
Pubblicazione: (2025)
Synergizing In-context Learning with Hints for End-to-end Task-oriented Dialog Systems
di: Saley, Vishal Vivek, et al.
Pubblicazione: (2024)
di: Saley, Vishal Vivek, et al.
Pubblicazione: (2024)
Beyond Size: How Gradients Shape Pruning Decisions in Large Language Models
di: Das, Rocktim Jyoti, et al.
Pubblicazione: (2023)
di: Das, Rocktim Jyoti, et al.
Pubblicazione: (2023)
MatchDiffusion: Training-free Generation of Match-cuts
di: Pardo, Alejandro, et al.
Pubblicazione: (2024)
di: Pardo, Alejandro, et al.
Pubblicazione: (2024)
DreamPolisher: Towards High-Quality Text-to-3D Generation via Geometric Diffusion
di: Lin, Yuanze, et al.
Pubblicazione: (2024)
di: Lin, Yuanze, et al.
Pubblicazione: (2024)
SAGE: Scalable Ground Truth Evaluations for Large Sparse Autoencoders
di: Venhoff, Constantin, et al.
Pubblicazione: (2024)
di: Venhoff, Constantin, et al.
Pubblicazione: (2024)
DEEDEE: Fast and Scalable Out-of-Distribution Dynamics Detection
di: Aljaafari, Tala, et al.
Pubblicazione: (2025)
di: Aljaafari, Tala, et al.
Pubblicazione: (2025)
SynthCLIP: Are We Ready for a Fully Synthetic CLIP Training?
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024)
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024)
OpenSanctions Pairs: Large-Scale Entity Matching with LLMs
di: Smith, Chandler, et al.
Pubblicazione: (2026)
di: Smith, Chandler, et al.
Pubblicazione: (2026)
On Pretraining Data Diversity for Self-Supervised Learning
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024)
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024)
Latent Guard: a Safety Framework for Text-to-image Generation
di: Liu, Runtao, et al.
Pubblicazione: (2024)
di: Liu, Runtao, et al.
Pubblicazione: (2024)
MediTOD: An English Dialogue Dataset for Medical History Taking with Comprehensive Annotations
di: Saley, Vishal Vivek, et al.
Pubblicazione: (2024)
di: Saley, Vishal Vivek, et al.
Pubblicazione: (2024)
EXAMS-V: A Multi-Discipline Multilingual Multimodal Exam Benchmark for Evaluating Vision Language Models
di: Das, Rocktim Jyoti, et al.
Pubblicazione: (2024)
di: Das, Rocktim Jyoti, et al.
Pubblicazione: (2024)
ActionParty: Multi-Subject Action Binding in Generative Video Games
di: Pondaven, Alexander, et al.
Pubblicazione: (2026)
di: Pondaven, Alexander, et al.
Pubblicazione: (2026)
STARC: A General Framework For Quantifying Differences Between Reward Functions
di: Skalse, Joar, et al.
Pubblicazione: (2023)
di: Skalse, Joar, et al.
Pubblicazione: (2023)
PhysMoDPO: Physically-Plausible Humanoid Motion with Preference Optimization
di: Zhang, Yangsong, et al.
Pubblicazione: (2026)
di: Zhang, Yangsong, et al.
Pubblicazione: (2026)
Detecting Multi-Agent Collusion Through Multi-Agent Interpretability
di: Rose, Aaron, et al.
Pubblicazione: (2026)
di: Rose, Aaron, et al.
Pubblicazione: (2026)
LLM Agents Are the Antidote to Walled Gardens
di: Marro, Samuele, et al.
Pubblicazione: (2025)
di: Marro, Samuele, et al.
Pubblicazione: (2025)
Model Merging and Safety Alignment: One Bad Model Spoils the Bunch
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024)
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024)
MAD-Sherlock: Multi-Agent Debate for Visual Misinformation Detection
di: Lakara, Kumud, et al.
Pubblicazione: (2024)
di: Lakara, Kumud, et al.
Pubblicazione: (2024)
HorizonMath: Measuring AI Progress Toward Mathematical Discovery with Automatic Verification
di: Wang, Erik Y., et al.
Pubblicazione: (2026)
di: Wang, Erik Y., et al.
Pubblicazione: (2026)
Single LLM Debate, MoLaCE: Mixture of Latent Concept Experts Against Confirmation Bias
di: Kim, Hazel, et al.
Pubblicazione: (2025)
di: Kim, Hazel, et al.
Pubblicazione: (2025)
Documenti analoghi
-
BLAZER: Bootstrapping LLM-based Manipulation Agents with Zero-Shot Data Generation
di: Das, Rocktim Jyoti, et al.
Pubblicazione: (2025) -
h1: Bootstrapping LLMs to Reason over Longer Horizons via Reinforcement Learning
di: Motwani, Sumeet Ramesh, et al.
Pubblicazione: (2025) -
Secret Collusion among AI Agents: Multi-Agent Deception via Steganography
di: Motwani, Sumeet Ramesh, et al.
Pubblicazione: (2024) -
MALMM: Multi-Agent Large Language Models for Zero-Shot Robotics Manipulation
di: Singh, Harsh, et al.
Pubblicazione: (2024) -
PSyDUCK: Training-Free Steganography for Latent Diffusion
di: Mahfuz, Aqib, et al.
Pubblicazione: (2025)