DialBGM: A Benchmark for Background Music Recommendation from Everyday Multi-Turn Dialogues
Fuente:
arXiv
Saved in:
| Main Authors: | Shin, Joonhyeok, Kang, Jaehoon, Lee, Yujun, Lee, Hannah, Lee, Yejin, Park, Yoonji, Shim, Kyuhong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unlocking Fine-Grained and Within-Utterance Speaking Style Control in Prompt-Based Text-to-Speech Models
by: Kang, Jaehoon, et al.
Published: (2026)
by: Kang, Jaehoon, et al.
Published: (2026)
Whisper-CD: Accurate Long-Form Speech Recognition using Multi-Negative Contrastive Decoding
by: Ahn, Hoseong, et al.
Published: (2026)
by: Ahn, Hoseong, et al.
Published: (2026)
WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models
by: Lee, Hanna, et al.
Published: (2026)
by: Lee, Hanna, et al.
Published: (2026)
P2VA: Converting Persona Descriptions into Voice Attributes for Fair and Controllable Text-to-Speech
by: Lee, Yejin, et al.
Published: (2025)
by: Lee, Yejin, et al.
Published: (2025)
Evaluating Hallucinations in Audio-Visual Multimodal LLMs with Spoken Queries under Diverse Acoustic Conditions
by: Park, Hansol, et al.
Published: (2025)
by: Park, Hansol, et al.
Published: (2025)
Preserving Pre-trained Representation Space: On Effectiveness of Prefix-tuning for Large Multi-modal Models
by: Kim, Donghoon, et al.
Published: (2024)
by: Kim, Donghoon, et al.
Published: (2024)
Learning Primitive Relations for Compositional Zero-Shot Learning
by: Lee, Insu, et al.
Published: (2025)
by: Lee, Insu, et al.
Published: (2025)
Revealing Multi-View Hallucination in Large Vision-Language Models
by: Park, Wooje, et al.
Published: (2026)
by: Park, Wooje, et al.
Published: (2026)
ProKG-Dial: Progressive Multi-Turn Dialogue Construction with Domain Knowledge Graphs
by: Liang, Yuanyuan, et al.
Published: (2025)
by: Liang, Yuanyuan, et al.
Published: (2025)
Towards Comprehensive Scene Understanding: Integrating First and Third-Person Views for LVLMs
by: Lee, Insu, et al.
Published: (2025)
by: Lee, Insu, et al.
Published: (2025)
Unlocking Transfer Learning for Open-World Few-Shot Recognition
by: Kim, Byeonggeun, et al.
Published: (2024)
by: Kim, Byeonggeun, et al.
Published: (2024)
Chain-of-Rank: Enhancing Large Language Models for Domain-Specific RAG in Edge Device
by: Lee, Juntae, et al.
Published: (2025)
by: Lee, Juntae, et al.
Published: (2025)
PSY-STEP: Structuring Therapeutic Targets and Action Sequences for Proactive Counseling Dialogue Systems
by: Lee, Jihyun, et al.
Published: (2026)
by: Lee, Jihyun, et al.
Published: (2026)
A Temporal Graph Network Framework for Dynamic Recommendation
by: Kim, Yejin, et al.
Published: (2024)
by: Kim, Yejin, et al.
Published: (2024)
Adaptive Capacity Allocation for Vision Language Action Fine-tuning
by: Kim, Donghoon, et al.
Published: (2026)
by: Kim, Donghoon, et al.
Published: (2026)
SafeDialBench: A Fine-Grained Safety Evaluation Benchmark for Large Language Models in Multi-Turn Dialogues with Diverse Jailbreak Attacks
by: Cao, Hongye, et al.
Published: (2025)
by: Cao, Hongye, et al.
Published: (2025)
Mask2Flow-TSE: Two-Stage Target Speaker Extraction with Masking and Flow Matching
by: Moon, Junwon, et al.
Published: (2026)
by: Moon, Junwon, et al.
Published: (2026)
PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization
by: Kim, Wonjoong, et al.
Published: (2026)
by: Kim, Wonjoong, et al.
Published: (2026)
Moltbook Moderation: Uncovering Hidden Intent Through Multi-Turn Dialogue
by: Al-Lawati, Ali, et al.
Published: (2026)
by: Al-Lawati, Ali, et al.
Published: (2026)
Visually Guided Decoding: Gradient-Free Hard Prompt Inversion with Language Models
by: Kim, Donghoon, et al.
Published: (2025)
by: Kim, Donghoon, et al.
Published: (2025)
UMAIR-FPS: User-aware Multi-modal Animation Illustration Recommendation Fusion with Painting Style
by: Kang, Yan, et al.
Published: (2024)
by: Kang, Yan, et al.
Published: (2024)
A Recommender System for NFT Collectibles with Item Feature
by: Choi, Minjoo, et al.
Published: (2024)
by: Choi, Minjoo, et al.
Published: (2024)
BGM-HAN: A Hierarchical Attention Network for Accurate and Fair Decision Assessment on Semi-Structured Profiles
by: Liu, Junhua, et al.
Published: (2025)
by: Liu, Junhua, et al.
Published: (2025)
CLIP-KOA: Enhancing Knee Osteoarthritis Diagnosis with Multi-Modal Learning and Symmetry-Aware Loss Functions
by: Jeong, Yejin, et al.
Published: (2025)
by: Jeong, Yejin, et al.
Published: (2025)
OPSD Compresses What RLVR Teaches: A Post-RL Compaction Stage for Reasoning Models
by: Kim, Jaehoon, et al.
Published: (2026)
by: Kim, Jaehoon, et al.
Published: (2026)
NOVI : Chatbot System for University Novice with BERT and LLMs
by: Nam, Yoonji, et al.
Published: (2024)
by: Nam, Yoonji, et al.
Published: (2024)
ArtSplat: Feed-Forward Articulated 3D Gaussian Splatting from Sparse Multi-State Uncalibrated Views
by: Lee, Inseo, et al.
Published: (2026)
by: Lee, Inseo, et al.
Published: (2026)
DialSim: A Dialogue Simulator for Evaluating Long-Term Multi-Party Dialogue Understanding of Conversational Agents
by: Kim, Jiho, et al.
Published: (2024)
by: Kim, Jiho, et al.
Published: (2024)
Partial-Multivariate Model for Forecasting
by: Lee, Jaehoon, et al.
Published: (2024)
by: Lee, Jaehoon, et al.
Published: (2024)
No Thing, Nothing: Highlighting Safety-Critical Classes for Robust LiDAR Semantic Segmentation in Adverse Weather
by: Park, Junsung, et al.
Published: (2025)
by: Park, Junsung, et al.
Published: (2025)
FinTexTS: Financial Text-Paired Time-Series Dataset via Semantic-Based and Multi-Level Pairing
by: Lee, Jaehoon, et al.
Published: (2026)
by: Lee, Jaehoon, et al.
Published: (2026)
Beyond Task-Oriented and Chitchat Dialogues: Proactive and Transition-Aware Conversational Agents
by: Yoon, Yejin, et al.
Published: (2025)
by: Yoon, Yejin, et al.
Published: (2025)
GuruAgents: Emulating Wise Investors with Prompt-Guided LLM Agents
by: Kim, Yejin, et al.
Published: (2025)
by: Kim, Yejin, et al.
Published: (2025)
A Self-Supervised Mixture-of-Experts Framework for Multi-behavior Recommendation
by: Kim, Kyungho, et al.
Published: (2025)
by: Kim, Kyungho, et al.
Published: (2025)
Multi-Behavior Recommender Systems: A Survey
by: Kim, Kyungho, et al.
Published: (2025)
by: Kim, Kyungho, et al.
Published: (2025)
Federated Recommender System with Data Valuation for E-commerce Platform
by: Park, Jongwon, et al.
Published: (2025)
by: Park, Jongwon, et al.
Published: (2025)
Prediction of Highway Traffic Flow Based on Artificial Intelligence Algorithms Using California Traffic Data
by: Lee, Junseong, et al.
Published: (2025)
by: Lee, Junseong, et al.
Published: (2025)
Discourse Diversity in Multi-Turn Empathic Dialogue
by: Zhan, Hongli, et al.
Published: (2026)
by: Zhan, Hongli, et al.
Published: (2026)
Stock Recommendations for Individual Investors: A Temporal Graph Network Approach with Mean-Variance Efficient Sampling
by: Lee, Youngbin, et al.
Published: (2024)
by: Lee, Youngbin, et al.
Published: (2024)
The Slow Drift of Support: Boundary Failures in Multi-Turn Mental Health LLM Dialogues
by: Cheng, Youyou, et al.
Published: (2026)
by: Cheng, Youyou, et al.
Published: (2026)
Similar Items
-
Unlocking Fine-Grained and Within-Utterance Speaking Style Control in Prompt-Based Text-to-Speech Models
by: Kang, Jaehoon, et al.
Published: (2026) -
Whisper-CD: Accurate Long-Form Speech Recognition using Multi-Negative Contrastive Decoding
by: Ahn, Hoseong, et al.
Published: (2026) -
WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models
by: Lee, Hanna, et al.
Published: (2026) -
P2VA: Converting Persona Descriptions into Voice Attributes for Fair and Controllable Text-to-Speech
by: Lee, Yejin, et al.
Published: (2025) -
Evaluating Hallucinations in Audio-Visual Multimodal LLMs with Spoken Queries under Diverse Acoustic Conditions
by: Park, Hansol, et al.
Published: (2025)