Demystifying deep search: a holistic evaluation with hint-free multi-hop questions and factorised metrics
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Song, Maojia, Liu, Renhang, Wang, Xinyu, Jiang, Yong, Xie, Pengjun, Huang, Fei, Zhou, Jingren, Herremans, Dorien, Poria, Soujanya |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Video2Music: Suitable Music Generation from Videos using an Affective Multimodal Transformer model
par: Kang, Jaeyong, et autres
Publié: (2023)
par: Kang, Jaeyong, et autres
Publié: (2023)
Leveraging LLM Embeddings for Cross Dataset Label Alignment and Zero Shot Music Emotion Prediction
par: Liu, Renhang, et autres
Publié: (2024)
par: Liu, Renhang, et autres
Publié: (2024)
LLMs Can't Handle Peer Pressure: Crumbling under Multi-Agent Social Interactions
par: Song, Maojia, et autres
Publié: (2025)
par: Song, Maojia, et autres
Publié: (2025)
JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment
par: Liu, Renhang, et autres
Publié: (2025)
par: Liu, Renhang, et autres
Publié: (2025)
JamendoMaxCaps: A Large Scale Music-caption Dataset with Imputed Metadata
par: Roy, Abhinaba, et autres
Publié: (2025)
par: Roy, Abhinaba, et autres
Publié: (2025)
Mustango: Toward Controllable Text-to-Music Generation
par: Melechovsky, Jan, et autres
Publié: (2023)
par: Melechovsky, Jan, et autres
Publié: (2023)
NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards
par: Hung, Chia-Yu, et autres
Publié: (2025)
par: Hung, Chia-Yu, et autres
Publié: (2025)
BandCondiNet: Parallel Transformers-based Conditional Popular Music Generation with Multi-View Features
par: Luo, Jing, et autres
Publié: (2024)
par: Luo, Jing, et autres
Publié: (2024)
Are We There Yet? A Brief Survey of Music Emotion Prediction Datasets, Models and Outstanding Challenges
par: Kang, Jaeyong, et autres
Publié: (2024)
par: Kang, Jaeyong, et autres
Publié: (2024)
DisfluencySpeech -- Single-Speaker Conversational Speech Dataset with Paralanguage
par: Wang, Kyra, et autres
Publié: (2024)
par: Wang, Kyra, et autres
Publié: (2024)
PreBit -- A multimodal model with Twitter FinBERT embeddings for extreme price movement prediction of Bitcoin
par: Zou, Yanzhao, et autres
Publié: (2022)
par: Zou, Yanzhao, et autres
Publié: (2022)
Towards Unified Music Emotion Recognition across Dimensional and Categorical Models
par: Kang, Jaeyong, et autres
Publié: (2025)
par: Kang, Jaeyong, et autres
Publié: (2025)
Aligning Generative Music AI with Human Preferences: Methods and Challenges
par: Herremans, Dorien, et autres
Publié: (2025)
par: Herremans, Dorien, et autres
Publié: (2025)
DeepUnifiedMom: Unified Time-series Momentum Portfolio Construction via Multi-Task Learning with Multi-Gate Mixture of Experts
par: Ong, Joel, et autres
Publié: (2024)
par: Ong, Joel, et autres
Publié: (2024)
Measuring and Enhancing Trustworthiness of LLMs in RAG through Grounded Attributions and Learning to Refuse
par: Song, Maojia, et autres
Publié: (2024)
par: Song, Maojia, et autres
Publié: (2024)
APEX: Large-scale Multi-task Aesthetic-Informed Popularity Prediction for AI-Generated Music
par: Husain, Jaavid Aktar, et autres
Publié: (2026)
par: Husain, Jaavid Aktar, et autres
Publié: (2026)
Forecasting Bitcoin volatility spikes from whale transactions and CryptoQuant data using Synthesizer Transformer models
par: Herremans, Dorien, et autres
Publié: (2022)
par: Herremans, Dorien, et autres
Publié: (2022)
PromptDistill: Query-based Selective Token Retention in Intermediate Layers for Efficient Large Language Model Inference
par: Jin, Weisheng, et autres
Publié: (2025)
par: Jin, Weisheng, et autres
Publié: (2025)
Two are better than one: Context window extension with multi-grained self-injection
par: Han, Wei, et autres
Publié: (2024)
par: Han, Wei, et autres
Publié: (2024)
KARMA-MV: A Benchmark for Causal Question Answering on Music Videos
par: Ghosh, Archishman, et autres
Publié: (2026)
par: Ghosh, Archishman, et autres
Publié: (2026)
Digital Lifelong Learning in the Age of AI: Trends and Insights
par: Puri, Geeta, et autres
Publié: (2026)
par: Puri, Geeta, et autres
Publié: (2026)
MidiCaps: A large-scale MIDI dataset with text captions
par: Melechovsky, Jan, et autres
Publié: (2024)
par: Melechovsky, Jan, et autres
Publié: (2024)
MIRFLEX: Music Information Retrieval Feature Library for Extraction
par: Chopra, Anuradha, et autres
Publié: (2024)
par: Chopra, Anuradha, et autres
Publié: (2024)
Smart Timing for Mining: A Deep Learning Framework for Bitcoin Hardware ROI Prediction
par: Wickramasinghe, Sithumi, et autres
Publié: (2025)
par: Wickramasinghe, Sithumi, et autres
Publié: (2025)
MERIT: Learning Disentangled Music Representations for Audio Similarity
par: Roy, Abhinaba, et autres
Publié: (2026)
par: Roy, Abhinaba, et autres
Publié: (2026)
SonicVerse: Multi-Task Learning for Music Feature-Informed Captioning
par: Chopra, Anuradha, et autres
Publié: (2025)
par: Chopra, Anuradha, et autres
Publié: (2025)
Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment
par: Roy, Abhinaba, et autres
Publié: (2025)
par: Roy, Abhinaba, et autres
Publié: (2025)
Exact Flow Linear Attention: Exact Solution from Continuous-Time Dynamics
par: Lei, Jingdi, et autres
Publié: (2025)
par: Lei, Jingdi, et autres
Publié: (2025)
Towards Robust Instruction Tuning on Multimodal Large Language Models
par: Han, Wei, et autres
Publié: (2024)
par: Han, Wei, et autres
Publié: (2024)
PREMISE: Matching-based Prediction for Accurate Review Recommendation
par: Han, Wei, et autres
Publié: (2025)
par: Han, Wei, et autres
Publié: (2025)
M-Longdoc: A Benchmark For Multimodal Super-Long Document Understanding And A Retrieval-Aware Tuning Framework
par: Chia, Yew Ken, et autres
Publié: (2024)
par: Chia, Yew Ken, et autres
Publié: (2024)
Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder
par: Melechovsky, Jan, et autres
Publié: (2022)
par: Melechovsky, Jan, et autres
Publié: (2022)
When Drawing Is Not Enough: Exploring Spontaneous Speech with Sketch for Intent Alignment in Multimodal LLMs
par: Shi, Weiyan, et autres
Publié: (2026)
par: Shi, Weiyan, et autres
Publié: (2026)
Accent Conversion in Text-To-Speech Using Multi-Level VAE and Adversarial Training
par: Melechovsky, Jan, et autres
Publié: (2024)
par: Melechovsky, Jan, et autres
Publié: (2024)
Prevailing Research Areas for Music AI in the Era of Foundation Models
par: Wei, Megan, et autres
Publié: (2024)
par: Wei, Megan, et autres
Publié: (2024)
DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech
par: Melechovsky, Jan, et autres
Publié: (2024)
par: Melechovsky, Jan, et autres
Publié: (2024)
SonicMaster: Towards Controllable All-in-One Music Restoration and Mastering
par: Melechovsky, Jan, et autres
Publié: (2025)
par: Melechovsky, Jan, et autres
Publié: (2025)
Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
par: Bhardwaj, Rishabh, et autres
Publié: (2024)
par: Bhardwaj, Rishabh, et autres
Publié: (2024)
DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling
par: Deep, Pala Tej, et autres
Publié: (2024)
par: Deep, Pala Tej, et autres
Publié: (2024)
Epistemic Context Learning: Building Trust the Right Way in LLM-Based Multi-Agent Systems
par: Zhou, Ruiwen, et autres
Publié: (2026)
par: Zhou, Ruiwen, et autres
Publié: (2026)
Documents similaires
-
Video2Music: Suitable Music Generation from Videos using an Affective Multimodal Transformer model
par: Kang, Jaeyong, et autres
Publié: (2023) -
Leveraging LLM Embeddings for Cross Dataset Label Alignment and Zero Shot Music Emotion Prediction
par: Liu, Renhang, et autres
Publié: (2024) -
LLMs Can't Handle Peer Pressure: Crumbling under Multi-Agent Social Interactions
par: Song, Maojia, et autres
Publié: (2025) -
JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment
par: Liu, Renhang, et autres
Publié: (2025) -
JamendoMaxCaps: A Large Scale Music-caption Dataset with Imputed Metadata
par: Roy, Abhinaba, et autres
Publié: (2025)