TRAM: Test-Time Risk Adaptation with Mixture of Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chehade, Mohamad Fares El Hajj, Bedi, Amrit Singh, Zhang, Amy, Zhu, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LEVIS: Large Exact Verifiable Input Spaces for Neural Networks
von: Chehade, Mohamad Fares El Hajj, et al.
Veröffentlicht: (2024)
von: Chehade, Mohamad Fares El Hajj, et al.
Veröffentlicht: (2024)
SINBAD: Saliency-informed detection of breakage caused by ad blocking
von: Chehade, Saiid El Hajj, et al.
Veröffentlicht: (2024)
von: Chehade, Saiid El Hajj, et al.
Veröffentlicht: (2024)
RL with Learnable Textual Feedback: A Bilevel Approach
von: Singh, Utsav, et al.
Veröffentlicht: (2026)
von: Singh, Utsav, et al.
Veröffentlicht: (2026)
Should we use model-free or model-based control? A case study of battery management systems
von: Chehade, Mohamad Fares El Hajj, et al.
Veröffentlicht: (2024)
von: Chehade, Mohamad Fares El Hajj, et al.
Veröffentlicht: (2024)
Improved Sample Complexity For Diffusion Model Training Without Empirical Risk Minimizer Access
von: Gaur, Mudit, et al.
Veröffentlicht: (2025)
von: Gaur, Mudit, et al.
Veröffentlicht: (2025)
Achieving Zero Constraint Violation for Constrained Reinforcement Learning via Conservative Natural Policy Gradient Primal-Dual Algorithm
von: Bai, Qinbo, et al.
Veröffentlicht: (2022)
von: Bai, Qinbo, et al.
Veröffentlicht: (2022)
On The Global Convergence Of Online RLHF With Neural Parametrization
von: Gaur, Mudit, et al.
Veröffentlicht: (2024)
von: Gaur, Mudit, et al.
Veröffentlicht: (2024)
Analysis of a degenerate parabolic system for cell dynamics in intestinal crypts
von: Hajj, Ahmad El, et al.
Veröffentlicht: (2026)
von: Hajj, Ahmad El, et al.
Veröffentlicht: (2026)
Closing the Gap: Achieving Global Convergence (Last Iterate) of Actor-Critic under Markovian Sampling with Neural Network Parametrization
von: Gaur, Mudit, et al.
Veröffentlicht: (2024)
von: Gaur, Mudit, et al.
Veröffentlicht: (2024)
Why Pass@k Optimization Can Degrade Pass@1: Prompt Interference in LLM Post-training
von: Barakat, Anas, et al.
Veröffentlicht: (2026)
von: Barakat, Anas, et al.
Veröffentlicht: (2026)
On The Sample Complexity Bounds In Bilevel Reinforcement Learning
von: Gaur, Mudit, et al.
Veröffentlicht: (2025)
von: Gaur, Mudit, et al.
Veröffentlicht: (2025)
Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time
von: Chehade, Mohamad, et al.
Veröffentlicht: (2025)
von: Chehade, Mohamad, et al.
Veröffentlicht: (2025)
Draft-Conditioned Constrained Decoding for Structured Generation in LLMs
von: Reddy, Avinash, et al.
Veröffentlicht: (2026)
von: Reddy, Avinash, et al.
Veröffentlicht: (2026)
Test-Time Scaling in Diffusion LLMs via Hidden Semi-Autoregressive Experts
von: Lee, Jihoon, et al.
Veröffentlicht: (2025)
von: Lee, Jihoon, et al.
Veröffentlicht: (2025)
MIRA: Towards Mitigating Reward Hacking in Inference-Time Alignment of T2I Diffusion Models
von: Zhai, Kevin, et al.
Veröffentlicht: (2025)
von: Zhai, Kevin, et al.
Veröffentlicht: (2025)
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2023)
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2023)
PIPER: Primitive-Informed Preference-based Hierarchical Reinforcement Learning via Hindsight Relabeling
von: Singh, Utsav, et al.
Veröffentlicht: (2024)
von: Singh, Utsav, et al.
Veröffentlicht: (2024)
On the Global Optimality of Policy Gradient Methods in General Utility Reinforcement Learning
von: Barakat, Anas, et al.
Veröffentlicht: (2024)
von: Barakat, Anas, et al.
Veröffentlicht: (2024)
Multi-LLM QA with Embodied Exploration
von: Patel, Bhrij, et al.
Veröffentlicht: (2024)
von: Patel, Bhrij, et al.
Veröffentlicht: (2024)
Towards Global Optimality for Practical Average Reward Reinforcement Learning without Mixing Time Oracles
von: Patel, Bhrij, et al.
Veröffentlicht: (2024)
von: Patel, Bhrij, et al.
Veröffentlicht: (2024)
TRAM: Bridging Trust Regions and Sharpness Aware Minimization
von: Sherborne, Tom, et al.
Veröffentlicht: (2023)
von: Sherborne, Tom, et al.
Veröffentlicht: (2023)
Monitoring Risks in Test-Time Adaptation
von: Schirmer, Mona, et al.
Veröffentlicht: (2025)
von: Schirmer, Mona, et al.
Veröffentlicht: (2025)
VARP: Reinforcement Learning from Vision-Language Model Feedback with Agent Regularized Preferences
von: Singh, Anukriti, et al.
Veröffentlicht: (2025)
von: Singh, Anukriti, et al.
Veröffentlicht: (2025)
Code Comprehension then Auditing for Unsupervised LLM Evaluation
von: Patel, Bhrij, et al.
Veröffentlicht: (2024)
von: Patel, Bhrij, et al.
Veröffentlicht: (2024)
Interpretable Neural Causal Models with TRAM-DAGs
von: Sick, Beate, et al.
Veröffentlicht: (2025)
von: Sick, Beate, et al.
Veröffentlicht: (2025)
DIPPER: Direct Preference Optimization to Accelerate Primitive-Enabled Hierarchical Reinforcement Learning
von: Singh, Utsav, et al.
Veröffentlicht: (2024)
von: Singh, Utsav, et al.
Veröffentlicht: (2024)
FACT or Fiction: Can Truthful Mechanisms Eliminate Federated Free Riding?
von: Bornstein, Marco, et al.
Veröffentlicht: (2024)
von: Bornstein, Marco, et al.
Veröffentlicht: (2024)
Test-Time Adaptation for LLM Agents via Environment Interaction
von: Chen, Arthur, et al.
Veröffentlicht: (2025)
von: Chen, Arthur, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Answer Selection for Improved Reasoning in Multi-LLM Systems
von: Agrawal, Aakriti, et al.
Veröffentlicht: (2025)
von: Agrawal, Aakriti, et al.
Veröffentlicht: (2025)
PARL: A Unified Framework for Policy Alignment in Reinforcement Learning from Human Feedback
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2023)
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2023)
SafeR-CLIP: Mitigating NSFW Content in Vision-Language Models While Preserving Pre-Trained Knowledge
von: Yousaf, Adeel, et al.
Veröffentlicht: (2025)
von: Yousaf, Adeel, et al.
Veröffentlicht: (2025)
Align-Pro: A Principled Approach to Prompt Optimization for LLM Alignment
von: Trivedi, Prashant, et al.
Veröffentlicht: (2025)
von: Trivedi, Prashant, et al.
Veröffentlicht: (2025)
Generative Modeling with Continuous Flows: Sample Complexity of Flow Matching
von: Gaur, Mudit, et al.
Veröffentlicht: (2025)
von: Gaur, Mudit, et al.
Veröffentlicht: (2025)
Variational Continual Test-Time Adaptation
von: Lyu, Fan, et al.
Veröffentlicht: (2024)
von: Lyu, Fan, et al.
Veröffentlicht: (2024)
Controllable Continual Test-Time Adaptation
von: Shi, Ziqi, et al.
Veröffentlicht: (2024)
von: Shi, Ziqi, et al.
Veröffentlicht: (2024)
Structural Alignment Improves Graph Test-Time Adaptation
von: Hsu, Hans Hao-Hsun, et al.
Veröffentlicht: (2025)
von: Hsu, Hans Hao-Hsun, et al.
Veröffentlicht: (2025)
Reliability-Gated Source Anchoring for Continual Test-Time Adaptation
von: Singh, Vikash, et al.
Veröffentlicht: (2026)
von: Singh, Vikash, et al.
Veröffentlicht: (2026)
On the Adversarial Risk of Test Time Adaptation: An Investigation into Realistic Test-Time Data Poisoning
von: Su, Yongyi, et al.
Veröffentlicht: (2024)
von: Su, Yongyi, et al.
Veröffentlicht: (2024)
PROPS: Progressively Private Self-alignment of Large Language Models
von: Teku, Noel, et al.
Veröffentlicht: (2025)
von: Teku, Noel, et al.
Veröffentlicht: (2025)
Transfer Q Star: Principled Decoding for LLM Alignment
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2024)
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
LEVIS: Large Exact Verifiable Input Spaces for Neural Networks
von: Chehade, Mohamad Fares El Hajj, et al.
Veröffentlicht: (2024) -
SINBAD: Saliency-informed detection of breakage caused by ad blocking
von: Chehade, Saiid El Hajj, et al.
Veröffentlicht: (2024) -
RL with Learnable Textual Feedback: A Bilevel Approach
von: Singh, Utsav, et al.
Veröffentlicht: (2026) -
Should we use model-free or model-based control? A case study of battery management systems
von: Chehade, Mohamad Fares El Hajj, et al.
Veröffentlicht: (2024) -
Improved Sample Complexity For Diffusion Model Training Without Empirical Risk Minimizer Access
von: Gaur, Mudit, et al.
Veröffentlicht: (2025)