Text Takes Over: A Study of Modality Bias in Multimodal Intent Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Mullick, Ankan, Sharma, Saransh, Jana, Abhik, Goyal, Pawan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
IDALC: A Semi-Supervised Framework for Intent Detection and Active Learning based Correction
by: Mullick, Ankan, et al.
Published: (2025)
by: Mullick, Ankan, et al.
Published: (2025)
Intent Detection and Entity Extraction from BioMedical Literature
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
A Pointer Network-based Approach for Joint Extraction and Detection of Multi-Label Multi-Class Intents
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
Long Dialog Summarization: An Analysis
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
Introducing Spotlight: A Novel Approach for Generating Captivating Key Information from Documents
by: Mullick, Ankan, et al.
Published: (2025)
by: Mullick, Ankan, et al.
Published: (2025)
On The Persona-based Summarization of Domain-Specific Documents
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
MatSciRE: Leveraging Pointer Networks to Automate Entity and Relation Extraction for Material Science Knowledge-base Construction
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
Leveraging the Power of LLMs: A Fine-Tuning Approach for High-Quality Aspect-Based Summarization
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
Order-Based Pre-training Strategies for Procedural Text Understanding
by: Nandy, Abhilash, et al.
Published: (2024)
by: Nandy, Abhilash, et al.
Published: (2024)
EduVidQA: Generating and Evaluating Long-form Answers to Student Questions based on Lecture Videos
by: Ray, Sourjyadip, et al.
Published: (2025)
by: Ray, Sourjyadip, et al.
Published: (2025)
A fully automated and scalable Parallel Data Augmentation for Low Resource Languages using Image and Text Analytics
by: Sharma, Prawaal, et al.
Published: (2025)
by: Sharma, Prawaal, et al.
Published: (2025)
An Empirical Study on the Characteristics of Bias upon Context Length Variation for Bangla
by: Sadhu, Jayanta, et al.
Published: (2024)
by: Sadhu, Jayanta, et al.
Published: (2024)
Breaking Token Into Concepts: Exploring Extreme Compression in Token Representation Via Compositional Shared Semantics
by: R V, Kavin, et al.
Published: (2025)
by: R V, Kavin, et al.
Published: (2025)
IL-TUR: Benchmark for Indian Legal Text Understanding and Reasoning
by: Joshi, Abhinav, et al.
Published: (2024)
by: Joshi, Abhinav, et al.
Published: (2024)
Dual Modality-Aware Gated Prompt Tuning for Few-Shot Multimodal Sarcasm Detection
by: Jana, Soumyadeep, et al.
Published: (2025)
by: Jana, Soumyadeep, et al.
Published: (2025)
An Answer is just the Start: Related Insight Generation for Open-Ended Document-Grounded QA
by: Sharma, Saransh, et al.
Published: (2026)
by: Sharma, Saransh, et al.
Published: (2026)
FiRST: Finetuning Router-Selective Transformers for Input-Adaptive Latency Reduction
by: Jain, Akriti, et al.
Published: (2024)
by: Jain, Akriti, et al.
Published: (2024)
Error-Aware Curriculum Learning for Biomedical Relation Classification
by: Chakraborty, Sinchani, et al.
Published: (2025)
by: Chakraborty, Sinchani, et al.
Published: (2025)
CrowdCounter: A benchmark type-specific multi-target counterspeech dataset
by: Saha, Punyajoy, et al.
Published: (2024)
by: Saha, Punyajoy, et al.
Published: (2024)
Enhancing Low-Resource NMT with a Multilingual Encoder and Knowledge Distillation: A Case Study
by: Roy, Aniruddha, et al.
Published: (2024)
by: Roy, Aniruddha, et al.
Published: (2024)
A Usage-centric Take on Intent Understanding in E-Commerce
by: Zhou, Wendi, et al.
Published: (2024)
by: Zhou, Wendi, et al.
Published: (2024)
VOLTAGE: A Versatile Contrastive Learning based OCR Methodology for ultra low-resource scripts through Auto Glyph Feature Extraction
by: Sharma, Prawaal, et al.
Published: (2025)
by: Sharma, Prawaal, et al.
Published: (2025)
How Robust are the Tabular QA Models for Scientific Tables? A Study using Customized Dataset
by: Ghosh, Akash, et al.
Published: (2024)
by: Ghosh, Akash, et al.
Published: (2024)
TECO: Improving Multimodal Intent Recognition with Text Enhancement through Commonsense Knowledge Extraction
by: Nguyen, Quynh-Mai Thi, et al.
Published: (2024)
by: Nguyen, Quynh-Mai Thi, et al.
Published: (2024)
Chat-Ghosting: A Comparative Study of Methods for Auto-Completion in Dialog Systems
by: Mishra, Sandeep, et al.
Published: (2025)
by: Mishra, Sandeep, et al.
Published: (2025)
Text or Pixels? It Takes Half: On the Token Efficiency of Visual Text Inputs in Multimodal LLMs
by: Li, Yanhong, et al.
Published: (2025)
by: Li, Yanhong, et al.
Published: (2025)
Lightweight Domain Adaptation of a Large Language Model for Legal Assistance in the Indian Context
by: Gupta, Jatin, et al.
Published: (2025)
by: Gupta, Jatin, et al.
Published: (2025)
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents
by: Jagadeeshan, Manoj Balaji, et al.
Published: (2025)
by: Jagadeeshan, Manoj Balaji, et al.
Published: (2025)
A Thousand Words or An Image: Studying the Influence of Persona Modality in Multimodal LLMs
by: Broomfield, Julius, et al.
Published: (2025)
by: Broomfield, Julius, et al.
Published: (2025)
CSSL: Contrastive Self-Supervised Learning for Dependency Parsing on Relatively Free Word Ordered and Morphologically Rich Low Resource Languages
by: Ray, Pretam, et al.
Published: (2024)
by: Ray, Pretam, et al.
Published: (2024)
ILSIC: Corpora for Identifying Indian Legal Statutes from Queries by Laypeople
by: Paul, Shounak, et al.
Published: (2026)
by: Paul, Shounak, et al.
Published: (2026)
SHA256 at SemEval-2025 Task 4: Selective Amnesia -- Constrained Unlearning for Large Language Models via Knowledge Isolation
by: Agrawal, Saransh, et al.
Published: (2025)
by: Agrawal, Saransh, et al.
Published: (2025)
Router-Suggest: Dynamic Routing for Multimodal Auto-Completion in Visually-Grounded Dialogs
by: Mishra, Sandeep, et al.
Published: (2026)
by: Mishra, Sandeep, et al.
Published: (2026)
On Zero-Shot Counterspeech Generation by LLMs
by: Saha, Punyajoy, et al.
Published: (2024)
by: Saha, Punyajoy, et al.
Published: (2024)
AutoIntent: AutoML for Text Classification
by: Alekseev, Ilya, et al.
Published: (2025)
by: Alekseev, Ilya, et al.
Published: (2025)
Retrieval-Augmented Reasoning for Chartered Accountancy
by: Gupta, Jatin, et al.
Published: (2026)
by: Gupta, Jatin, et al.
Published: (2026)
REFINE-AF: A Task-Agnostic Framework to Align Language Models via Self-Generated Instructions using Reinforcement Learning from Automated Feedback
by: Roy, Aniruddha, et al.
Published: (2025)
by: Roy, Aniruddha, et al.
Published: (2025)
NIM: Neuro-symbolic Ideographic Metalanguage for Inclusive Communication
by: Sharma, Prawaal, et al.
Published: (2025)
by: Sharma, Prawaal, et al.
Published: (2025)
Intent Detection in the Age of LLMs
by: Arora, Gaurav, et al.
Published: (2024)
by: Arora, Gaurav, et al.
Published: (2024)
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding
by: Ye, Junyi, et al.
Published: (2024)
by: Ye, Junyi, et al.
Published: (2024)
Similar Items
-
IDALC: A Semi-Supervised Framework for Intent Detection and Active Learning based Correction
by: Mullick, Ankan, et al.
Published: (2025) -
Intent Detection and Entity Extraction from BioMedical Literature
by: Mullick, Ankan, et al.
Published: (2024) -
A Pointer Network-based Approach for Joint Extraction and Detection of Multi-Label Multi-Class Intents
by: Mullick, Ankan, et al.
Published: (2024) -
Long Dialog Summarization: An Analysis
by: Mullick, Ankan, et al.
Published: (2024) -
Introducing Spotlight: A Novel Approach for Generating Captivating Key Information from Documents
by: Mullick, Ankan, et al.
Published: (2025)