Amulet: ReAlignment During Test Time for Personalized Preference Adaptation of LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Zhaowei, Bai, Fengshuo, Chen, Qizhi, Ma, Chengdong, Wang, Mingzhi, Sun, Haoran, Zheng, Zilong, Yang, Yaodong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PoliCon: Evaluating LLMs on Achieving Diverse Political Consensus Objectives
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2025)
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2025)
Guided Speculative Inference for Efficient Test-Time Alignment of LLMs
von: Geuter, Jonathan, et al.
Veröffentlicht: (2025)
von: Geuter, Jonathan, et al.
Veröffentlicht: (2025)
Roadmap on Incentive Compatibility for AI Alignment and Governance in Sociotechnical Systems
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2024)
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2024)
PersonalLLM: Tailoring LLMs to Individual Preferences
von: Zollo, Thomas P., et al.
Veröffentlicht: (2024)
von: Zollo, Thomas P., et al.
Veröffentlicht: (2024)
ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment
von: Bian, Zhipeng, et al.
Veröffentlicht: (2026)
von: Bian, Zhipeng, et al.
Veröffentlicht: (2026)
Latent Personality Alignment: Improving Harmlessness Without Mentioning Harms
von: Le, Linh, et al.
Veröffentlicht: (2026)
von: Le, Linh, et al.
Veröffentlicht: (2026)
Cross-lingual Human-Preference Alignment for Neural Machine Translation with Direct Quality Optimization
von: Uhlig, Kaden, et al.
Veröffentlicht: (2024)
von: Uhlig, Kaden, et al.
Veröffentlicht: (2024)
Does LLM Alignment Really Need Diversity? An Empirical Study of Adapting RLVR Methods for Moral Reasoning
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2026)
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2026)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
von: Saji, Alan, et al.
Veröffentlicht: (2025)
von: Saji, Alan, et al.
Veröffentlicht: (2025)
RAG-Optimized Tibetan Tourism LLMs: Enhancing Accuracy and Personalization
von: Qi, Jinhu, et al.
Veröffentlicht: (2024)
von: Qi, Jinhu, et al.
Veröffentlicht: (2024)
Automatic Task Detection and Heterogeneous LLM Speculative Decoding
von: Ge, Danying, et al.
Veröffentlicht: (2025)
von: Ge, Danying, et al.
Veröffentlicht: (2025)
Steer-MoE: Efficient Audio-Language Alignment with a Mixture-of-Experts Steering Module
von: Feng, Ruitao, et al.
Veröffentlicht: (2025)
von: Feng, Ruitao, et al.
Veröffentlicht: (2025)
Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation
von: Resck, Lucas, et al.
Veröffentlicht: (2026)
von: Resck, Lucas, et al.
Veröffentlicht: (2026)
Test-Time Scaling of Reasoning Models for Machine Translation
von: Li, Zihao, et al.
Veröffentlicht: (2025)
von: Li, Zihao, et al.
Veröffentlicht: (2025)
Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs
von: Easley, Eric, et al.
Veröffentlicht: (2026)
von: Easley, Eric, et al.
Veröffentlicht: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
von: Sorstkins, Andrejs
Veröffentlicht: (2025)
von: Sorstkins, Andrejs
Veröffentlicht: (2025)
Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2025)
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2025)
Unleashing LLMs in Bayesian Optimization: Preference-Guided Framework for Scientific Discovery
von: Yuan, Xinzhe, et al.
Veröffentlicht: (2026)
von: Yuan, Xinzhe, et al.
Veröffentlicht: (2026)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
MIRIAD: Augmenting LLMs with millions of medical query-response pairs
von: Zheng, Qinyue, et al.
Veröffentlicht: (2025)
von: Zheng, Qinyue, et al.
Veröffentlicht: (2025)
BitCal-TTS: Bit-Calibrated Test-Time Scaling for Quantized Reasoning Models
von: Patarlapalli, Sai Babu, et al.
Veröffentlicht: (2026)
von: Patarlapalli, Sai Babu, et al.
Veröffentlicht: (2026)
Whose Facts Win? LLM Source Preferences under Knowledge Conflicts
von: Schuster, Jakob, et al.
Veröffentlicht: (2026)
von: Schuster, Jakob, et al.
Veröffentlicht: (2026)
Rubrik's Cube: Testing a New Rubric for Evaluating Explanations on the CUBE dataset
von: Galvan-Sosa, Diana, et al.
Veröffentlicht: (2025)
von: Galvan-Sosa, Diana, et al.
Veröffentlicht: (2025)
SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2026)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2026)
Strategy Adaptation in Large Language Model Werewolf Agents
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025)
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025)
Generative Active Testing: Efficient LLM Evaluation via Proxy Task Adaptation
von: Ramakrishnan, Aashish Anantha, et al.
Veröffentlicht: (2026)
von: Ramakrishnan, Aashish Anantha, et al.
Veröffentlicht: (2026)
Aligning LLMs on a Budget: Inference-Time Alignment with Heuristic Reward Models
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
von: Sun, Yuhui, et al.
Veröffentlicht: (2025)
von: Sun, Yuhui, et al.
Veröffentlicht: (2025)
Adaptive Interviewing for Persona Simulation in LLMs: Evidence-Grounded Reasoning Improves Decision Alignment
von: Su, Ruoxi, et al.
Veröffentlicht: (2026)
von: Su, Ruoxi, et al.
Veröffentlicht: (2026)
Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels
von: Rath, Plawan Kumar, et al.
Veröffentlicht: (2026)
von: Rath, Plawan Kumar, et al.
Veröffentlicht: (2026)
How Do Large Language Models Acquire Factual Knowledge During Pretraining?
von: Chang, Hoyeon, et al.
Veröffentlicht: (2024)
von: Chang, Hoyeon, et al.
Veröffentlicht: (2024)
Evaluating Relational Reasoning in LLMs with REL
von: Fesser, Lukas, et al.
Veröffentlicht: (2026)
von: Fesser, Lukas, et al.
Veröffentlicht: (2026)
LoRS: Efficient Low-Rank Adaptation for Sparse Large Language Model
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
Large Language Models as Oracles for Ontology Alignment
von: Lushnei, Sviatoslav, et al.
Veröffentlicht: (2025)
von: Lushnei, Sviatoslav, et al.
Veröffentlicht: (2025)
Personality, Role, and Expressive Style in Large Language Models: An Interactionist Analysis
von: Nagao, Moe, et al.
Veröffentlicht: (2026)
von: Nagao, Moe, et al.
Veröffentlicht: (2026)
Multilingual LLMs Inherently Reward In-Language Time-Sensitive Semantic Alignment for Low-Resource Languages
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2024)
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2024)
Constructing Benchmarks and Interventions for Combating Hallucinations in LLMs
von: Simhi, Adi, et al.
Veröffentlicht: (2024)
von: Simhi, Adi, et al.
Veröffentlicht: (2024)
LLMs and the Human Condition
von: Wallis, Peter
Veröffentlicht: (2024)
von: Wallis, Peter
Veröffentlicht: (2024)
Ähnliche Einträge
-
PoliCon: Evaluating LLMs on Achieving Diverse Political Consensus Objectives
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2025) -
Guided Speculative Inference for Efficient Test-Time Alignment of LLMs
von: Geuter, Jonathan, et al.
Veröffentlicht: (2025) -
Roadmap on Incentive Compatibility for AI Alignment and Governance in Sociotechnical Systems
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2024) -
PersonalLLM: Tailoring LLMs to Individual Preferences
von: Zollo, Thomas P., et al.
Veröffentlicht: (2024) -
ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment
von: Bian, Zhipeng, et al.
Veröffentlicht: (2026)