Efficient and Accurate Memorable Conversation Model using DPO based on sLLM
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Seo, Youngkyung, Heo, Yoonseok, Koh, Jun-Seok, Chang, Du-Seong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Key-Element-Informed sLLM Tuning for Document Summarization
von: Ryu, Sangwon, et al.
Veröffentlicht: (2024)
von: Ryu, Sangwon, et al.
Veröffentlicht: (2024)
Not All Personas Are Worth It: Culture-Reflective Persona Data Augmentation
von: Han, Ji-Eun, et al.
Veröffentlicht: (2025)
von: Han, Ji-Eun, et al.
Veröffentlicht: (2025)
daDPO: Distribution-Aware DPO for Distilling Conversational Abilities
von: Zhang, Zhengze, et al.
Veröffentlicht: (2025)
von: Zhang, Zhengze, et al.
Veröffentlicht: (2025)
Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis
von: Djiré, Albérick Euraste, et al.
Veröffentlicht: (2025)
von: Djiré, Albérick Euraste, et al.
Veröffentlicht: (2025)
Enhancing Psychotherapy Counseling: A Data Augmentation Pipeline Leveraging Large Language Models for Counseling Conversations
von: Kim, Jun-Woo, et al.
Veröffentlicht: (2024)
von: Kim, Jun-Woo, et al.
Veröffentlicht: (2024)
SP^2DPO: An LLM-assisted Semantic Per-Pair DPO Generalization
von: He, Chaoyue, et al.
Veröffentlicht: (2026)
von: He, Chaoyue, et al.
Veröffentlicht: (2026)
Data Compressibility Quantifies LLM Memorization
von: Huang, Yizhan, et al.
Veröffentlicht: (2025)
von: Huang, Yizhan, et al.
Veröffentlicht: (2025)
PSYDIAL: Personality-based Synthetic Dialogue Generation using Large Language Models
von: Han, Ji-Eun, et al.
Veröffentlicht: (2024)
von: Han, Ji-Eun, et al.
Veröffentlicht: (2024)
Rethinking DPO: The Role of Rejected Responses in Preference Misalignment
von: Cho, Jay Hyeon, et al.
Veröffentlicht: (2025)
von: Cho, Jay Hyeon, et al.
Veröffentlicht: (2025)
Aligning Large Language Models with Counterfactual DPO
von: Butcher, Bradley
Veröffentlicht: (2024)
von: Butcher, Bradley
Veröffentlicht: (2024)
Context Memorization for Efficient Long Context Generation
von: Okoshi, Yasuyuki, et al.
Veröffentlicht: (2026)
von: Okoshi, Yasuyuki, et al.
Veröffentlicht: (2026)
Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M
von: Pant, Piyush
Veröffentlicht: (2025)
von: Pant, Piyush
Veröffentlicht: (2025)
Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning
von: Chen, Xuhang, et al.
Veröffentlicht: (2025)
von: Chen, Xuhang, et al.
Veröffentlicht: (2025)
An Empirical Study of SFT-DPO Interaction and Parameterization in Small Language Models
von: Feng, Yuming, et al.
Veröffentlicht: (2026)
von: Feng, Yuming, et al.
Veröffentlicht: (2026)
Memorization in Large Language Models in Medicine: Prevalence, Characteristics, and Implications
von: Li, Anran, et al.
Veröffentlicht: (2025)
von: Li, Anran, et al.
Veröffentlicht: (2025)
Memorizing Documents with Guidance in Large Language Models
von: Park, Bumjin, et al.
Veröffentlicht: (2024)
von: Park, Bumjin, et al.
Veröffentlicht: (2024)
Arithmetic with Language Models: from Memorization to Computation
von: Maltoni, Davide, et al.
Veröffentlicht: (2023)
von: Maltoni, Davide, et al.
Veröffentlicht: (2023)
Memorization in Fine-Tuned Large Language Models
von: Savine, Danil
Veröffentlicht: (2025)
von: Savine, Danil
Veröffentlicht: (2025)
MedMT-Bench: Can LLMs Memorize and Understand Long Multi-Turn Conversations in Medical Scenarios?
von: Yang, Lin, et al.
Veröffentlicht: (2026)
von: Yang, Lin, et al.
Veröffentlicht: (2026)
AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment
von: Xiao, Jianfei, et al.
Veröffentlicht: (2026)
von: Xiao, Jianfei, et al.
Veröffentlicht: (2026)
Cat-DPO: Category-Adaptive Safety Alignment
von: Yang, Tiankai, et al.
Veröffentlicht: (2026)
von: Yang, Tiankai, et al.
Veröffentlicht: (2026)
How Do Large Language Models Acquire Factual Knowledge During Pretraining?
von: Chang, Hoyeon, et al.
Veröffentlicht: (2024)
von: Chang, Hoyeon, et al.
Veröffentlicht: (2024)
Automated Survey Collection with LLM-based Conversational Agents
von: Kaiyrbekov, Kurmanbek, et al.
Veröffentlicht: (2025)
von: Kaiyrbekov, Kurmanbek, et al.
Veröffentlicht: (2025)
Undesirable Memorization in Large Language Models: A Survey
von: Satvaty, Ali, et al.
Veröffentlicht: (2024)
von: Satvaty, Ali, et al.
Veröffentlicht: (2024)
KoACD: The First Korean Adolescent Dataset for Cognitive Distortion Analysis via Role-Switching Multi-LLM Negotiation
von: Kim, JunSeo, et al.
Veröffentlicht: (2025)
von: Kim, JunSeo, et al.
Veröffentlicht: (2025)
DB-LLM: Accurate Dual-Binarization for Efficient LLMs
von: Chen, Hong, et al.
Veröffentlicht: (2024)
von: Chen, Hong, et al.
Veröffentlicht: (2024)
Affective-NLI: Towards Accurate and Interpretable Personality Recognition in Conversation
von: Wen, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Wen, Zhiyuan, et al.
Veröffentlicht: (2024)
Context-DPO: Aligning Language Models for Context-Faithfulness
von: Bi, Baolong, et al.
Veröffentlicht: (2024)
von: Bi, Baolong, et al.
Veröffentlicht: (2024)
Memorization in Attention-only Transformers
von: Dana, Léo, et al.
Veröffentlicht: (2024)
von: Dana, Léo, et al.
Veröffentlicht: (2024)
MixDPO: Modeling Preference Strength for Pluralistic Alignment
von: Imai, Saki, et al.
Veröffentlicht: (2026)
von: Imai, Saki, et al.
Veröffentlicht: (2026)
Mitigating Memorization In Language Models
von: Sakarvadia, Mansi, et al.
Veröffentlicht: (2024)
von: Sakarvadia, Mansi, et al.
Veröffentlicht: (2024)
A Multi-Perspective Analysis of Memorization in Large Language Models
von: Chen, Bowen, et al.
Veröffentlicht: (2024)
von: Chen, Bowen, et al.
Veröffentlicht: (2024)
Beyond Memorization: The Challenge of Random Memory Access in Language Models
von: Zhu, Tongyao, et al.
Veröffentlicht: (2024)
von: Zhu, Tongyao, et al.
Veröffentlicht: (2024)
Generalization or Memorization? Brittleness Testing for Chess-Trained Language Models
von: Tang, Ethan
Veröffentlicht: (2026)
von: Tang, Ethan
Veröffentlicht: (2026)
Retentive or Forgetful? Diving into the Knowledge Memorizing Mechanism of Language Models
von: Cao, Boxi, et al.
Veröffentlicht: (2023)
von: Cao, Boxi, et al.
Veröffentlicht: (2023)
Impact of Fine-Tuning Methods on Memorization in Large Language Models
von: Hou, Jie, et al.
Veröffentlicht: (2025)
von: Hou, Jie, et al.
Veröffentlicht: (2025)
Integrated Framework for LLM Evaluation with Answer Generation
von: Lee, Sujeong, et al.
Veröffentlicht: (2025)
von: Lee, Sujeong, et al.
Veröffentlicht: (2025)
Towards Analyzing and Understanding the Limitations of DPO: A Theoretical Perspective
von: Feng, Duanyu, et al.
Veröffentlicht: (2024)
von: Feng, Duanyu, et al.
Veröffentlicht: (2024)
sDPO: Don't Use Your Data All at Once
von: Kim, Dahyun, et al.
Veröffentlicht: (2024)
von: Kim, Dahyun, et al.
Veröffentlicht: (2024)
Teaching with Lies: Curriculum DPO on Synthetic Negatives for Hallucination Detection
von: Pandit, Shrey, et al.
Veröffentlicht: (2025)
von: Pandit, Shrey, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Key-Element-Informed sLLM Tuning for Document Summarization
von: Ryu, Sangwon, et al.
Veröffentlicht: (2024) -
Not All Personas Are Worth It: Culture-Reflective Persona Data Augmentation
von: Han, Ji-Eun, et al.
Veröffentlicht: (2025) -
daDPO: Distribution-Aware DPO for Distilling Conversational Abilities
von: Zhang, Zhengze, et al.
Veröffentlicht: (2025) -
Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis
von: Djiré, Albérick Euraste, et al.
Veröffentlicht: (2025) -
Enhancing Psychotherapy Counseling: A Data Augmentation Pipeline Leveraging Large Language Models for Counseling Conversations
von: Kim, Jun-Woo, et al.
Veröffentlicht: (2024)