SDIF-DA: A Shallow-to-Deep Interaction Framework with Data Augmentation for Multi-modal Intent Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Shijue, Qin, Libo, Wang, Bingbing, Tu, Geng, Xu, Ruifeng |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CroPrompt: Cross-task Interactive Prompting for Zero-shot Spoken Language Understanding
by: Qin, Libo, et al.
Published: (2024)
by: Qin, Libo, et al.
Published: (2024)
MemGuide: Intent-Driven Memory Selection for Goal-Oriented Multi-Session LLM Agents
by: Du, Yiming, et al.
Published: (2025)
by: Du, Yiming, et al.
Published: (2025)
Multi-modal Stance Detection: New Datasets and Model
by: Liang, Bin, et al.
Published: (2024)
by: Liang, Bin, et al.
Published: (2024)
CoreEval: Automatically Building Contamination-Resilient Datasets with Real-World Knowledge toward Reliable LLM Evaluation
by: Zhao, Jingqian, et al.
Published: (2025)
by: Zhao, Jingqian, et al.
Published: (2025)
Multi-modal Retrieval Augmented Multi-modal Generation: Datasets, Evaluation Metrics and Strong Baselines
by: Ma, Zi-Ao, et al.
Published: (2024)
by: Ma, Zi-Ao, et al.
Published: (2024)
Cross-modality Data Augmentation for End-to-End Sign Language Translation
by: Ye, Jinhui, et al.
Published: (2023)
by: Ye, Jinhui, et al.
Published: (2023)
Generate then Refine: Data Augmentation for Zero-shot Intent Detection
by: Lin, I-Fan, et al.
Published: (2024)
by: Lin, I-Fan, et al.
Published: (2024)
Towards On-Policy Data Evolution for Visual-Native Multimodal Deep Search Agents
by: Huang, Shijue, et al.
Published: (2026)
by: Huang, Shijue, et al.
Published: (2026)
PRISM of Opinions: A Persona-Reasoned Multimodal Framework for User-centric Conversational Stance Detection
by: Wang, Bingbing, et al.
Published: (2025)
by: Wang, Bingbing, et al.
Published: (2025)
BAIT: Boundary-Guided Disclosure Escalation via Self-Conditioned Reasoning
by: Luo, Xuan, et al.
Published: (2026)
by: Luo, Xuan, et al.
Published: (2026)
M$^3$CoT: A Novel Benchmark for Multi-Domain Multi-step Multi-modal Chain-of-Thought
by: Chen, Qiguang, et al.
Published: (2024)
by: Chen, Qiguang, et al.
Published: (2024)
MAC-SLU: Multi-Intent Automotive Cabin Spoken Language Understanding Benchmark
by: Peng, Yuezhang, et al.
Published: (2025)
by: Peng, Yuezhang, et al.
Published: (2025)
Mitigating Provenance-Role Collapse in Long-Term Agents via Typed Memory Representation
by: Jin, Zhengda, et al.
Published: (2026)
by: Jin, Zhengda, et al.
Published: (2026)
EDDA: A Encoder-Decoder Data Augmentation Framework for Zero-Shot Stance Detection
by: Ding, Daijun, et al.
Published: (2024)
by: Ding, Daijun, et al.
Published: (2024)
Learning to Describe for Predicting Zero-shot Drug-Drug Interactions
by: Zhu, Fangqi, et al.
Published: (2024)
by: Zhu, Fangqi, et al.
Published: (2024)
MIND Your Reasoning: A Meta-Cognitive Intuitive-Reflective Network for Dual-Reasoning in Multimodal Stance Detection
by: Wang, Bingbing, et al.
Published: (2025)
by: Wang, Bingbing, et al.
Published: (2025)
A Simple and Efficient Jailbreak Method Exploiting LLMs' Helpfulness
by: Luo, Xuan, et al.
Published: (2025)
by: Luo, Xuan, et al.
Published: (2025)
AMuSeD: An Attentive Deep Neural Network for Multimodal Sarcasm Detection Incorporating Bi-modal Data Augmentation
by: Gao, Xiyuan, et al.
Published: (2024)
by: Gao, Xiyuan, et al.
Published: (2024)
RoboSVG: A Unified Framework for Interactive SVG Generation with Multi-modal Guidance
by: Wang, Jiuniu, et al.
Published: (2025)
by: Wang, Jiuniu, et al.
Published: (2025)
mRAG: Elucidating the Design Space of Multi-modal Retrieval-Augmented Generation
by: Hu, Chan-Wei, et al.
Published: (2025)
by: Hu, Chan-Wei, et al.
Published: (2025)
Known Intents, New Combinations: Clause-Factorized Decoding for Compositional Multi-Intent Detection
by: Nandy, Abhilash
Published: (2026)
by: Nandy, Abhilash
Published: (2026)
Cross-modality Information Check for Detecting Jailbreaking in Multimodal Large Language Models
by: Xu, Yue, et al.
Published: (2024)
by: Xu, Yue, et al.
Published: (2024)
SARC: Sentiment-Augmented Deep Role Clustering for Fake News Detection
by: Wang, Jingqing, et al.
Published: (2025)
by: Wang, Jingqing, et al.
Published: (2025)
Evaluating Proactive Risk Awareness of Large Language Models
by: Luo, Xuan, et al.
Published: (2026)
by: Luo, Xuan, et al.
Published: (2026)
MM-StanceDet: Retrieval-Augmented Multi-modal Multi-agent Stance Detection
by: Lu, Weihai, et al.
Published: (2026)
by: Lu, Weihai, et al.
Published: (2026)
Comprehensive and Efficient Distillation for Lightweight Sentiment Analysis Models
by: Xie, Guangyu, et al.
Published: (2025)
by: Xie, Guangyu, et al.
Published: (2025)
How DDAIR you? Disambiguated Data Augmentation for Intent Recognition
by: Castillo-López, Galo, et al.
Published: (2026)
by: Castillo-López, Galo, et al.
Published: (2026)
Wrong-of-Thought: An Integrated Reasoning Framework with Multi-Perspective Verification and Wrong Information
by: Zhang, Yongheng, et al.
Published: (2024)
by: Zhang, Yongheng, et al.
Published: (2024)
X-WebAgentBench: A Multilingual Interactive Web Benchmark for Evaluating Global Agentic System
by: Wang, Peng, et al.
Published: (2025)
by: Wang, Peng, et al.
Published: (2025)
Rethinking Patient Education as Multi-turn Multi-modal Interaction
by: Yao, Zonghai, et al.
Published: (2026)
by: Yao, Zonghai, et al.
Published: (2026)
Mitigating Biases of Large Language Models in Stance Detection with Counterfactual Augmented Calibration
by: Li, Ang, et al.
Published: (2024)
by: Li, Ang, et al.
Published: (2024)
Intent Mismatch Causes LLMs to Get Lost in Multi-Turn Conversation
by: Liu, Geng, et al.
Published: (2026)
by: Liu, Geng, et al.
Published: (2026)
DS$^2$-ABSA: Dual-Stream Data Synthesis with Label Refinement for Few-Shot Aspect-Based Sentiment Analysis
by: Xu, Hongling, et al.
Published: (2024)
by: Xu, Hongling, et al.
Published: (2024)
GOSU: Retrieval-Augmented Generation with Global-Level Optimized Semantic Unit-Centric Framework
by: Zou, Xuecheng, et al.
Published: (2025)
by: Zou, Xuecheng, et al.
Published: (2025)
Exploring Description-Augmented Dataless Intent Classification
by: Hu, Ruoyu, et al.
Published: (2024)
by: Hu, Ruoyu, et al.
Published: (2024)
BlendX: Complex Multi-Intent Detection with Blended Patterns
by: Yoon, Yejin, et al.
Published: (2024)
by: Yoon, Yejin, et al.
Published: (2024)
RAMA: Retrieval-Augmented Multi-Agent Framework for Misinformation Detection in Multimodal Fact-Checking
by: Yang, Shuo, et al.
Published: (2025)
by: Yang, Shuo, et al.
Published: (2025)
Retrieve-Refine-Calibrate: A Framework for Complex Claim Fact-Checking
by: Sun, Mingwei, et al.
Published: (2026)
by: Sun, Mingwei, et al.
Published: (2026)
Freeze Deep, Train Shallow: Interpretable Layer Allocation for Continued Pre-Training
by: Wu, Yu-Hang, et al.
Published: (2026)
by: Wu, Yu-Hang, et al.
Published: (2026)
LARA: Linguistic-Adaptive Retrieval-Augmentation for Multi-Turn Intent Classification
by: Liu, Junhua, et al.
Published: (2024)
by: Liu, Junhua, et al.
Published: (2024)
Similar Items
-
CroPrompt: Cross-task Interactive Prompting for Zero-shot Spoken Language Understanding
by: Qin, Libo, et al.
Published: (2024) -
MemGuide: Intent-Driven Memory Selection for Goal-Oriented Multi-Session LLM Agents
by: Du, Yiming, et al.
Published: (2025) -
Multi-modal Stance Detection: New Datasets and Model
by: Liang, Bin, et al.
Published: (2024) -
CoreEval: Automatically Building Contamination-Resilient Datasets with Real-World Knowledge toward Reliable LLM Evaluation
by: Zhao, Jingqian, et al.
Published: (2025) -
Multi-modal Retrieval Augmented Multi-modal Generation: Datasets, Evaluation Metrics and Strong Baselines
by: Ma, Zi-Ao, et al.
Published: (2024)