Efficient and Accurate Memorable Conversation Model using DPO based on sLLM
Fuente:
arXiv
Salvato in:
| Autori principali: | Seo, Youngkyung, Heo, Yoonseok, Koh, Jun-Seok, Chang, Du-Seong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Key-Element-Informed sLLM Tuning for Document Summarization
di: Ryu, Sangwon, et al.
Pubblicazione: (2024)
di: Ryu, Sangwon, et al.
Pubblicazione: (2024)
Not All Personas Are Worth It: Culture-Reflective Persona Data Augmentation
di: Han, Ji-Eun, et al.
Pubblicazione: (2025)
di: Han, Ji-Eun, et al.
Pubblicazione: (2025)
daDPO: Distribution-Aware DPO for Distilling Conversational Abilities
di: Zhang, Zhengze, et al.
Pubblicazione: (2025)
di: Zhang, Zhengze, et al.
Pubblicazione: (2025)
Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis
di: Djiré, Albérick Euraste, et al.
Pubblicazione: (2025)
di: Djiré, Albérick Euraste, et al.
Pubblicazione: (2025)
Enhancing Psychotherapy Counseling: A Data Augmentation Pipeline Leveraging Large Language Models for Counseling Conversations
di: Kim, Jun-Woo, et al.
Pubblicazione: (2024)
di: Kim, Jun-Woo, et al.
Pubblicazione: (2024)
SP^2DPO: An LLM-assisted Semantic Per-Pair DPO Generalization
di: He, Chaoyue, et al.
Pubblicazione: (2026)
di: He, Chaoyue, et al.
Pubblicazione: (2026)
Data Compressibility Quantifies LLM Memorization
di: Huang, Yizhan, et al.
Pubblicazione: (2025)
di: Huang, Yizhan, et al.
Pubblicazione: (2025)
PSYDIAL: Personality-based Synthetic Dialogue Generation using Large Language Models
di: Han, Ji-Eun, et al.
Pubblicazione: (2024)
di: Han, Ji-Eun, et al.
Pubblicazione: (2024)
Rethinking DPO: The Role of Rejected Responses in Preference Misalignment
di: Cho, Jay Hyeon, et al.
Pubblicazione: (2025)
di: Cho, Jay Hyeon, et al.
Pubblicazione: (2025)
Aligning Large Language Models with Counterfactual DPO
di: Butcher, Bradley
Pubblicazione: (2024)
di: Butcher, Bradley
Pubblicazione: (2024)
Context Memorization for Efficient Long Context Generation
di: Okoshi, Yasuyuki, et al.
Pubblicazione: (2026)
di: Okoshi, Yasuyuki, et al.
Pubblicazione: (2026)
Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M
di: Pant, Piyush
Pubblicazione: (2025)
di: Pant, Piyush
Pubblicazione: (2025)
Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning
di: Chen, Xuhang, et al.
Pubblicazione: (2025)
di: Chen, Xuhang, et al.
Pubblicazione: (2025)
An Empirical Study of SFT-DPO Interaction and Parameterization in Small Language Models
di: Feng, Yuming, et al.
Pubblicazione: (2026)
di: Feng, Yuming, et al.
Pubblicazione: (2026)
Memorization in Large Language Models in Medicine: Prevalence, Characteristics, and Implications
di: Li, Anran, et al.
Pubblicazione: (2025)
di: Li, Anran, et al.
Pubblicazione: (2025)
Memorizing Documents with Guidance in Large Language Models
di: Park, Bumjin, et al.
Pubblicazione: (2024)
di: Park, Bumjin, et al.
Pubblicazione: (2024)
Arithmetic with Language Models: from Memorization to Computation
di: Maltoni, Davide, et al.
Pubblicazione: (2023)
di: Maltoni, Davide, et al.
Pubblicazione: (2023)
Memorization in Fine-Tuned Large Language Models
di: Savine, Danil
Pubblicazione: (2025)
di: Savine, Danil
Pubblicazione: (2025)
MedMT-Bench: Can LLMs Memorize and Understand Long Multi-Turn Conversations in Medical Scenarios?
di: Yang, Lin, et al.
Pubblicazione: (2026)
di: Yang, Lin, et al.
Pubblicazione: (2026)
AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment
di: Xiao, Jianfei, et al.
Pubblicazione: (2026)
di: Xiao, Jianfei, et al.
Pubblicazione: (2026)
Cat-DPO: Category-Adaptive Safety Alignment
di: Yang, Tiankai, et al.
Pubblicazione: (2026)
di: Yang, Tiankai, et al.
Pubblicazione: (2026)
How Do Large Language Models Acquire Factual Knowledge During Pretraining?
di: Chang, Hoyeon, et al.
Pubblicazione: (2024)
di: Chang, Hoyeon, et al.
Pubblicazione: (2024)
Automated Survey Collection with LLM-based Conversational Agents
di: Kaiyrbekov, Kurmanbek, et al.
Pubblicazione: (2025)
di: Kaiyrbekov, Kurmanbek, et al.
Pubblicazione: (2025)
Undesirable Memorization in Large Language Models: A Survey
di: Satvaty, Ali, et al.
Pubblicazione: (2024)
di: Satvaty, Ali, et al.
Pubblicazione: (2024)
KoACD: The First Korean Adolescent Dataset for Cognitive Distortion Analysis via Role-Switching Multi-LLM Negotiation
di: Kim, JunSeo, et al.
Pubblicazione: (2025)
di: Kim, JunSeo, et al.
Pubblicazione: (2025)
DB-LLM: Accurate Dual-Binarization for Efficient LLMs
di: Chen, Hong, et al.
Pubblicazione: (2024)
di: Chen, Hong, et al.
Pubblicazione: (2024)
Affective-NLI: Towards Accurate and Interpretable Personality Recognition in Conversation
di: Wen, Zhiyuan, et al.
Pubblicazione: (2024)
di: Wen, Zhiyuan, et al.
Pubblicazione: (2024)
Context-DPO: Aligning Language Models for Context-Faithfulness
di: Bi, Baolong, et al.
Pubblicazione: (2024)
di: Bi, Baolong, et al.
Pubblicazione: (2024)
Memorization in Attention-only Transformers
di: Dana, Léo, et al.
Pubblicazione: (2024)
di: Dana, Léo, et al.
Pubblicazione: (2024)
MixDPO: Modeling Preference Strength for Pluralistic Alignment
di: Imai, Saki, et al.
Pubblicazione: (2026)
di: Imai, Saki, et al.
Pubblicazione: (2026)
Mitigating Memorization In Language Models
di: Sakarvadia, Mansi, et al.
Pubblicazione: (2024)
di: Sakarvadia, Mansi, et al.
Pubblicazione: (2024)
A Multi-Perspective Analysis of Memorization in Large Language Models
di: Chen, Bowen, et al.
Pubblicazione: (2024)
di: Chen, Bowen, et al.
Pubblicazione: (2024)
Beyond Memorization: The Challenge of Random Memory Access in Language Models
di: Zhu, Tongyao, et al.
Pubblicazione: (2024)
di: Zhu, Tongyao, et al.
Pubblicazione: (2024)
Generalization or Memorization? Brittleness Testing for Chess-Trained Language Models
di: Tang, Ethan
Pubblicazione: (2026)
di: Tang, Ethan
Pubblicazione: (2026)
Retentive or Forgetful? Diving into the Knowledge Memorizing Mechanism of Language Models
di: Cao, Boxi, et al.
Pubblicazione: (2023)
di: Cao, Boxi, et al.
Pubblicazione: (2023)
Impact of Fine-Tuning Methods on Memorization in Large Language Models
di: Hou, Jie, et al.
Pubblicazione: (2025)
di: Hou, Jie, et al.
Pubblicazione: (2025)
Integrated Framework for LLM Evaluation with Answer Generation
di: Lee, Sujeong, et al.
Pubblicazione: (2025)
di: Lee, Sujeong, et al.
Pubblicazione: (2025)
Towards Analyzing and Understanding the Limitations of DPO: A Theoretical Perspective
di: Feng, Duanyu, et al.
Pubblicazione: (2024)
di: Feng, Duanyu, et al.
Pubblicazione: (2024)
sDPO: Don't Use Your Data All at Once
di: Kim, Dahyun, et al.
Pubblicazione: (2024)
di: Kim, Dahyun, et al.
Pubblicazione: (2024)
Teaching with Lies: Curriculum DPO on Synthetic Negatives for Hallucination Detection
di: Pandit, Shrey, et al.
Pubblicazione: (2025)
di: Pandit, Shrey, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Key-Element-Informed sLLM Tuning for Document Summarization
di: Ryu, Sangwon, et al.
Pubblicazione: (2024) -
Not All Personas Are Worth It: Culture-Reflective Persona Data Augmentation
di: Han, Ji-Eun, et al.
Pubblicazione: (2025) -
daDPO: Distribution-Aware DPO for Distilling Conversational Abilities
di: Zhang, Zhengze, et al.
Pubblicazione: (2025) -
Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis
di: Djiré, Albérick Euraste, et al.
Pubblicazione: (2025) -
Enhancing Psychotherapy Counseling: A Data Augmentation Pipeline Leveraging Large Language Models for Counseling Conversations
di: Kim, Jun-Woo, et al.
Pubblicazione: (2024)