Aligning LLM Agents by Learning Latent Preference from User Edits
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Ge, Taymanov, Alexey, Salinas, Eduardo, Mineiro, Paul, Misra, Dipendra |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Principled Fine-tuning of LLMs from User-Edits: A Medley of Preference, Supervision, and Reward
by: Misra, Dipendra, et al.
Published: (2026)
by: Misra, Dipendra, et al.
Published: (2026)
From Volume to Value: Preference-Aligned Memory Construction for On-Device RAG
by: Lee, Changmin, et al.
Published: (2026)
by: Lee, Changmin, et al.
Published: (2026)
Separating and Learning Latent Confounders to Enhancing User Preferences Modeling
by: Xu, Hangtong, et al.
Published: (2023)
by: Xu, Hangtong, et al.
Published: (2023)
Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory Framework
by: Zhang, Zeyu, et al.
Published: (2025)
by: Zhang, Zeyu, et al.
Published: (2025)
IDGenRec: LLM-RecSys Alignment with Textual ID Learning
by: Tan, Juntao, et al.
Published: (2024)
by: Tan, Juntao, et al.
Published: (2024)
Context Retrieval via Normalized Contextual Latent Interaction for Conversational Agent
by: Liu, Junfeng, et al.
Published: (2023)
by: Liu, Junfeng, et al.
Published: (2023)
Holistic Utility Preference Learning for Listwise Alignment
by: Zhou, Jiacong, et al.
Published: (2024)
by: Zhou, Jiacong, et al.
Published: (2024)
Human-Inspired Memory Architecture for LLM Agents
by: Kerestecioglu, Doga, et al.
Published: (2026)
by: Kerestecioglu, Doga, et al.
Published: (2026)
AriadneMem: Threading the Maze of Lifelong Memory for LLM Agents
by: Zhu, Wenhui, et al.
Published: (2026)
by: Zhu, Wenhui, et al.
Published: (2026)
NextMem: Towards Latent Factual Memory for LLM-based Agents
by: Zhang, Zeyu, et al.
Published: (2026)
by: Zhang, Zeyu, et al.
Published: (2026)
CKnowEdit: A New Chinese Knowledge Editing Dataset for Linguistics, Facts, and Logic Error Correction in LLMs
by: Fang, Jizhan, et al.
Published: (2024)
by: Fang, Jizhan, et al.
Published: (2024)
Adapting Job Recommendations to User Preference Drift with Behavioral-Semantic Fusion Learning
by: Han, Xiao, et al.
Published: (2024)
by: Han, Xiao, et al.
Published: (2024)
SkillGraph: Graph Foundation Priors for LLM Agent Tool Sequence Recommendation
by: Liu, Hao, et al.
Published: (2026)
by: Liu, Hao, et al.
Published: (2026)
Agent Learning via Early Experience
by: Zhang, Kai, et al.
Published: (2025)
by: Zhang, Kai, et al.
Published: (2025)
Preference Discerning with LLM-Enhanced Generative Retrieval
by: Paischer, Fabian, et al.
Published: (2024)
by: Paischer, Fabian, et al.
Published: (2024)
High Fidelity Textual User Representation over Heterogeneous Sources via Reinforcement Learning
by: Arora, Rajat, et al.
Published: (2026)
by: Arora, Rajat, et al.
Published: (2026)
OneEdit: A Neural-Symbolic Collaboratively Knowledge Editing System
by: Zhang, Ningyu, et al.
Published: (2024)
by: Zhang, Ningyu, et al.
Published: (2024)
AD-Bench: A Real-World, Trajectory-Aware Advertising Analytics Benchmark for LLM Agents
by: Hu, Lingxiang, et al.
Published: (2026)
by: Hu, Lingxiang, et al.
Published: (2026)
RankPO: Preference Optimization for Job-Talent Matching
by: Zhang, Yafei, et al.
Published: (2025)
by: Zhang, Yafei, et al.
Published: (2025)
User Preference Modeling for Conversational LLM Agents: Weak Rewards from Retrieval-Augmented Interaction
by: Hao, Yuren, et al.
Published: (2026)
by: Hao, Yuren, et al.
Published: (2026)
Policy-Gradient Training of Language Models for Ranking
by: Gao, Ge, et al.
Published: (2023)
by: Gao, Ge, et al.
Published: (2023)
User Embedding Model for Personalized Language Prompting
by: Doddapaneni, Sumanth, et al.
Published: (2024)
by: Doddapaneni, Sumanth, et al.
Published: (2024)
An Interpretable Alternative to Neural Representation Learning for Rating Prediction -- Transparent Latent Class Modeling of User Reviews
by: Serra, Giuseppe, et al.
Published: (2024)
by: Serra, Giuseppe, et al.
Published: (2024)
Aligning Dense Retrievers with LLM Utility via Distillation
by: Sandhu, Rajinder, et al.
Published: (2026)
by: Sandhu, Rajinder, et al.
Published: (2026)
ComMer: a Framework for Compressing and Merging User Data for Personalization
by: Zeldes, Yoel, et al.
Published: (2025)
by: Zeldes, Yoel, et al.
Published: (2025)
Solving the Content Gap in Roblox Game Recommendations: LLM-Based Profile Generation and Reranking
by: Wang, Chen, et al.
Published: (2025)
by: Wang, Chen, et al.
Published: (2025)
ARL2: Aligning Retrievers for Black-box Large Language Models via Self-guided Adaptive Relevance Labeling
by: Zhang, Lingxi, et al.
Published: (2024)
by: Zhang, Lingxi, et al.
Published: (2024)
DeepAgent: A General Reasoning Agent with Scalable Toolsets
by: Li, Xiaoxi, et al.
Published: (2025)
by: Li, Xiaoxi, et al.
Published: (2025)
Science Consultant Agent
by: K, Karthikeyan, et al.
Published: (2025)
by: K, Karthikeyan, et al.
Published: (2025)
Solving the Cold Start Problem on One's Own as an End User via Preference Transfer
by: Sato, Ryoma
Published: (2025)
by: Sato, Ryoma
Published: (2025)
CuriousLLM: Elevating Multi-Document Question Answering with LLM-Enhanced Knowledge Graph Reasoning
by: Yang, Zukang, et al.
Published: (2024)
by: Yang, Zukang, et al.
Published: (2024)
OneKE: A Dockerized Schema-Guided LLM Agent-based Knowledge Extraction System
by: Luo, Yujie, et al.
Published: (2024)
by: Luo, Yujie, et al.
Published: (2024)
MTRec: Learning to Align with User Preferences via Mental Reward Models
by: Zhao, Mengchen, et al.
Published: (2025)
by: Zhao, Mengchen, et al.
Published: (2025)
Session-based Recommender Systems: User Interest as a Stochastic Process in the Latent Space
by: Balcer, Klaudia, et al.
Published: (2025)
by: Balcer, Klaudia, et al.
Published: (2025)
Sequential LLM Framework for Fashion Recommendation
by: Liu, Han, et al.
Published: (2024)
by: Liu, Han, et al.
Published: (2024)
LLM-Enhanced Linear Autoencoders for Recommendation
by: Moon, Jaewan, et al.
Published: (2025)
by: Moon, Jaewan, et al.
Published: (2025)
Scaling Generalist Data-Analytic Agents
by: Qiao, Shuofei, et al.
Published: (2025)
by: Qiao, Shuofei, et al.
Published: (2025)
Personas within Parameters: Fine-Tuning Small Language Models with Low-Rank Adapters to Mimic User Behaviors
by: Thakur, Himanshu, et al.
Published: (2025)
by: Thakur, Himanshu, et al.
Published: (2025)
STRUM-LLM: Attributed and Structured Contrastive Summarization
by: Gunel, Beliz, et al.
Published: (2024)
by: Gunel, Beliz, et al.
Published: (2024)
Learning Unified User Quantized Tokenizers for User Representation
by: He, Chuan, et al.
Published: (2025)
by: He, Chuan, et al.
Published: (2025)
Similar Items
-
Principled Fine-tuning of LLMs from User-Edits: A Medley of Preference, Supervision, and Reward
by: Misra, Dipendra, et al.
Published: (2026) -
From Volume to Value: Preference-Aligned Memory Construction for On-Device RAG
by: Lee, Changmin, et al.
Published: (2026) -
Separating and Learning Latent Confounders to Enhancing User Preferences Modeling
by: Xu, Hangtong, et al.
Published: (2023) -
Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory Framework
by: Zhang, Zeyu, et al.
Published: (2025) -
IDGenRec: LLM-RecSys Alignment with Textual ID Learning
by: Tan, Juntao, et al.
Published: (2024)