Amulet: ReAlignment During Test Time for Personalized Preference Adaptation of LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Zhaowei, Bai, Fengshuo, Chen, Qizhi, Ma, Chengdong, Wang, Mingzhi, Sun, Haoran, Zheng, Zilong, Yang, Yaodong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PoliCon: Evaluating LLMs on Achieving Diverse Political Consensus Objectives
por: Zhang, Zhaowei, et al.
Publicado: (2025)
por: Zhang, Zhaowei, et al.
Publicado: (2025)
Guided Speculative Inference for Efficient Test-Time Alignment of LLMs
por: Geuter, Jonathan, et al.
Publicado: (2025)
por: Geuter, Jonathan, et al.
Publicado: (2025)
Roadmap on Incentive Compatibility for AI Alignment and Governance in Sociotechnical Systems
por: Zhang, Zhaowei, et al.
Publicado: (2024)
por: Zhang, Zhaowei, et al.
Publicado: (2024)
PersonalLLM: Tailoring LLMs to Individual Preferences
por: Zollo, Thomas P., et al.
Publicado: (2024)
por: Zollo, Thomas P., et al.
Publicado: (2024)
ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment
por: Bian, Zhipeng, et al.
Publicado: (2026)
por: Bian, Zhipeng, et al.
Publicado: (2026)
Latent Personality Alignment: Improving Harmlessness Without Mentioning Harms
por: Le, Linh, et al.
Publicado: (2026)
por: Le, Linh, et al.
Publicado: (2026)
Cross-lingual Human-Preference Alignment for Neural Machine Translation with Direct Quality Optimization
por: Uhlig, Kaden, et al.
Publicado: (2024)
por: Uhlig, Kaden, et al.
Publicado: (2024)
Does LLM Alignment Really Need Diversity? An Empirical Study of Adapting RLVR Methods for Moral Reasoning
por: Zhang, Zhaowei, et al.
Publicado: (2026)
por: Zhang, Zhaowei, et al.
Publicado: (2026)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
por: Saji, Alan, et al.
Publicado: (2025)
por: Saji, Alan, et al.
Publicado: (2025)
RAG-Optimized Tibetan Tourism LLMs: Enhancing Accuracy and Personalization
por: Qi, Jinhu, et al.
Publicado: (2024)
por: Qi, Jinhu, et al.
Publicado: (2024)
Automatic Task Detection and Heterogeneous LLM Speculative Decoding
por: Ge, Danying, et al.
Publicado: (2025)
por: Ge, Danying, et al.
Publicado: (2025)
Steer-MoE: Efficient Audio-Language Alignment with a Mixture-of-Experts Steering Module
por: Feng, Ruitao, et al.
Publicado: (2025)
por: Feng, Ruitao, et al.
Publicado: (2025)
Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation
por: Resck, Lucas, et al.
Publicado: (2026)
por: Resck, Lucas, et al.
Publicado: (2026)
Test-Time Scaling of Reasoning Models for Machine Translation
por: Li, Zihao, et al.
Publicado: (2025)
por: Li, Zihao, et al.
Publicado: (2025)
Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs
por: Easley, Eric, et al.
Publicado: (2026)
por: Easley, Eric, et al.
Publicado: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
por: Ashuach, Tomer, et al.
Publicado: (2025)
por: Ashuach, Tomer, et al.
Publicado: (2025)
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
por: Sorstkins, Andrejs
Publicado: (2025)
por: Sorstkins, Andrejs
Publicado: (2025)
Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
por: Bajpai, Ashutosh, et al.
Publicado: (2025)
por: Bajpai, Ashutosh, et al.
Publicado: (2025)
Unleashing LLMs in Bayesian Optimization: Preference-Guided Framework for Scientific Discovery
por: Yuan, Xinzhe, et al.
Publicado: (2026)
por: Yuan, Xinzhe, et al.
Publicado: (2026)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
por: Fadli, Samih
Publicado: (2025)
por: Fadli, Samih
Publicado: (2025)
MIRIAD: Augmenting LLMs with millions of medical query-response pairs
por: Zheng, Qinyue, et al.
Publicado: (2025)
por: Zheng, Qinyue, et al.
Publicado: (2025)
BitCal-TTS: Bit-Calibrated Test-Time Scaling for Quantized Reasoning Models
por: Patarlapalli, Sai Babu, et al.
Publicado: (2026)
por: Patarlapalli, Sai Babu, et al.
Publicado: (2026)
Whose Facts Win? LLM Source Preferences under Knowledge Conflicts
por: Schuster, Jakob, et al.
Publicado: (2026)
por: Schuster, Jakob, et al.
Publicado: (2026)
Rubrik's Cube: Testing a New Rubric for Evaluating Explanations on the CUBE dataset
por: Galvan-Sosa, Diana, et al.
Publicado: (2025)
por: Galvan-Sosa, Diana, et al.
Publicado: (2025)
SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation
por: Kim, Seoyeon, et al.
Publicado: (2026)
por: Kim, Seoyeon, et al.
Publicado: (2026)
Strategy Adaptation in Large Language Model Werewolf Agents
por: Nakamori, Fuya, et al.
Publicado: (2025)
por: Nakamori, Fuya, et al.
Publicado: (2025)
Generative Active Testing: Efficient LLM Evaluation via Proxy Task Adaptation
por: Ramakrishnan, Aashish Anantha, et al.
Publicado: (2026)
por: Ramakrishnan, Aashish Anantha, et al.
Publicado: (2026)
Aligning LLMs on a Budget: Inference-Time Alignment with Heuristic Reward Models
por: Nakamura, Mason, et al.
Publicado: (2025)
por: Nakamura, Mason, et al.
Publicado: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
por: Peters, Sydney, et al.
Publicado: (2025)
por: Peters, Sydney, et al.
Publicado: (2025)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
por: Sun, Yuhui, et al.
Publicado: (2025)
por: Sun, Yuhui, et al.
Publicado: (2025)
Adaptive Interviewing for Persona Simulation in LLMs: Evidence-Grounded Reasoning Improves Decision Alignment
por: Su, Ruoxi, et al.
Publicado: (2026)
por: Su, Ruoxi, et al.
Publicado: (2026)
Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels
por: Rath, Plawan Kumar, et al.
Publicado: (2026)
por: Rath, Plawan Kumar, et al.
Publicado: (2026)
How Do Large Language Models Acquire Factual Knowledge During Pretraining?
por: Chang, Hoyeon, et al.
Publicado: (2024)
por: Chang, Hoyeon, et al.
Publicado: (2024)
Evaluating Relational Reasoning in LLMs with REL
por: Fesser, Lukas, et al.
Publicado: (2026)
por: Fesser, Lukas, et al.
Publicado: (2026)
LoRS: Efficient Low-Rank Adaptation for Sparse Large Language Model
por: Hu, Yuxuan, et al.
Publicado: (2025)
por: Hu, Yuxuan, et al.
Publicado: (2025)
Large Language Models as Oracles for Ontology Alignment
por: Lushnei, Sviatoslav, et al.
Publicado: (2025)
por: Lushnei, Sviatoslav, et al.
Publicado: (2025)
Personality, Role, and Expressive Style in Large Language Models: An Interactionist Analysis
por: Nagao, Moe, et al.
Publicado: (2026)
por: Nagao, Moe, et al.
Publicado: (2026)
Multilingual LLMs Inherently Reward In-Language Time-Sensitive Semantic Alignment for Low-Resource Languages
por: Bajpai, Ashutosh, et al.
Publicado: (2024)
por: Bajpai, Ashutosh, et al.
Publicado: (2024)
Constructing Benchmarks and Interventions for Combating Hallucinations in LLMs
por: Simhi, Adi, et al.
Publicado: (2024)
por: Simhi, Adi, et al.
Publicado: (2024)
LLMs and the Human Condition
por: Wallis, Peter
Publicado: (2024)
por: Wallis, Peter
Publicado: (2024)
Ejemplares similares
-
PoliCon: Evaluating LLMs on Achieving Diverse Political Consensus Objectives
por: Zhang, Zhaowei, et al.
Publicado: (2025) -
Guided Speculative Inference for Efficient Test-Time Alignment of LLMs
por: Geuter, Jonathan, et al.
Publicado: (2025) -
Roadmap on Incentive Compatibility for AI Alignment and Governance in Sociotechnical Systems
por: Zhang, Zhaowei, et al.
Publicado: (2024) -
PersonalLLM: Tailoring LLMs to Individual Preferences
por: Zollo, Thomas P., et al.
Publicado: (2024) -
ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment
por: Bian, Zhipeng, et al.
Publicado: (2026)