Revealing the impact of synthetic native samples and multi-tasking strategies in Hindi-English code-mixed humour and sarcasm detection
Fuente:
arXiv
Saved in:
| Main Authors: | Mazumder, Debajyoti, Kumar, Aakash, Patro, Jasabanta |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving code-mixed hate detection by native sample mixing: A case study for Hindi-English code-mixed scenario
by: Mazumder, Debajyoti, et al.
Published: (2024)
by: Mazumder, Debajyoti, et al.
Published: (2024)
Neither Here Nor There: Cross-Lingual Representation Dynamics of Code-Mixed Text in Multilingual Encoders
by: Mazumder, Debajyoti, et al.
Published: (2026)
by: Mazumder, Debajyoti, et al.
Published: (2026)
On VLMs for Diverse Tasks in Multimodal Meme Classification
by: Gavit, Deepesh, et al.
Published: (2025)
by: Gavit, Deepesh, et al.
Published: (2025)
An Under-Explored Application for Explainable Multimodal Misogyny Detection in code-mixed Hindi-English
by: Yadav, Sargam, et al.
Published: (2026)
by: Yadav, Sargam, et al.
Published: (2026)
Entailed Opinion Matters: Improving the Fact-Checking Performance of Language Models by Relying on their Entailment Ability
by: Kumar, Gaurav, et al.
Published: (2025)
by: Kumar, Gaurav, et al.
Published: (2025)
Evaluating Cross-lingual Knowledge Consistency in Code-Mixed vis-a-vis Indian Languages using IndicKLAR
by: Mazumder, Debajyoti, et al.
Published: (2026)
by: Mazumder, Debajyoti, et al.
Published: (2026)
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
All that is English may be Hindi: Enhancing language identification through automatic ranking of likeliness of word borrowing in social media
by: Patro, Jasabanta, et al.
Published: (2017)
by: Patro, Jasabanta, et al.
Published: (2017)
HindiLLM: Large Language Model for Hindi
by: Chouhan, Sanjay, et al.
Published: (2024)
by: Chouhan, Sanjay, et al.
Published: (2024)
MultiCheck: Strengthening Web Trust with Unified Multimodal Fact Verification
by: Kishore, Aditya, et al.
Published: (2025)
by: Kishore, Aditya, et al.
Published: (2025)
Bridging the Data Gap: Creating a Hindi Text Summarization Dataset from the English XSUM
by: Katwe, Praveenkumar, et al.
Published: (2026)
by: Katwe, Praveenkumar, et al.
Published: (2026)
MedSumm: A Multimodal Approach to Summarizing Code-Mixed Hindi-English Clinical Queries
by: Ghosh, Akash, et al.
Published: (2024)
by: Ghosh, Akash, et al.
Published: (2024)
COMI-LINGUA: Expert Annotated Large-Scale Dataset for Multitask NLP in Hindi-English Code-Mixing
by: Sheth, Rajvee, et al.
Published: (2025)
by: Sheth, Rajvee, et al.
Published: (2025)
Airavata: Introducing Hindi Instruction-tuned LLM
by: Gala, Jay, et al.
Published: (2024)
by: Gala, Jay, et al.
Published: (2024)
Suvach -- Generated Hindi QA benchmark
by: Narayanan, Vaishak, et al.
Published: (2024)
by: Narayanan, Vaishak, et al.
Published: (2024)
Advantages of Domain Knowledge Injection for Legal Document Summarization: A Case Study on Summarizing Indian Court Judgments in English and Hindi
by: Datta, Debtanu, et al.
Published: (2026)
by: Datta, Debtanu, et al.
Published: (2026)
Multi-class Regret Detection in Hindi Devanagari Script
by: Sharma, Renuka, et al.
Published: (2024)
by: Sharma, Renuka, et al.
Published: (2024)
CoT-Self-Instruct: Building high-quality synthetic prompts for reasoning and non-reasoning tasks
by: Yu, Ping, et al.
Published: (2025)
by: Yu, Ping, et al.
Published: (2025)
Survey of Pseudonymization, Abstractive Summarization & Spell Checker for Hindi and Marathi
by: Ransing, Rasika, et al.
Published: (2024)
by: Ransing, Rasika, et al.
Published: (2024)
SEEK: Semantic Evidence Extraction via Adaptive ChunKing for Multilingual Fact-Checking
by: Kumar, Babu, et al.
Published: (2026)
by: Kumar, Babu, et al.
Published: (2026)
HLDC: Hindi Legal Documents Corpus
by: Kapoor, Arnav, et al.
Published: (2022)
by: Kapoor, Arnav, et al.
Published: (2022)
Multilingual LLMs Are Not Multilingual Thinkers: Evidence from Hindi Analogy Evaluation
by: Gupta, Ashray, et al.
Published: (2025)
by: Gupta, Ashray, et al.
Published: (2025)
Raising Bars, Not Parameters: LilMoo Compact Language Model for Hindi
by: Fatimah, Shiza, et al.
Published: (2026)
by: Fatimah, Shiza, et al.
Published: (2026)
AIMA at SemEval-2024 Task 10: History-Based Emotion Recognition in Hindi-English Code-Mixed Conversations
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2025)
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2025)
Assessing how hyperparameters impact Large Language Models' sarcasm detection performance
by: Gole, Montgomery, et al.
Published: (2025)
by: Gole, Montgomery, et al.
Published: (2025)
A Hybrid Supervised-LLM Pipeline for Actionable Suggestion Mining in Unstructured Customer Reviews
by: Trivedi, Aakash, et al.
Published: (2026)
by: Trivedi, Aakash, et al.
Published: (2026)
Building pre-train LLM Dataset for the INDIC Languages: a case study on Hindi
by: Parida, Shantipriya, et al.
Published: (2024)
by: Parida, Shantipriya, et al.
Published: (2024)
LLMCARE: early detection of cognitive impairment via transformer models enhanced by LLM-generated synthetic data
by: Zolnour, Ali, et al.
Published: (2025)
by: Zolnour, Ali, et al.
Published: (2025)
Causal Interventions Reveal Shared Structure Across English Filler-Gap Constructions
by: Boguraev, Sasha, et al.
Published: (2025)
by: Boguraev, Sasha, et al.
Published: (2025)
DeepRAG: Building a Custom Hindi Embedding Model for Retrieval Augmented Generation from Scratch
by: M, Nandakishor
Published: (2025)
by: M, Nandakishor
Published: (2025)
$\texttt{YC-Bench}$: Benchmarking AI Agents for Long-Term Planning and Consistent Execution
by: He, Muyu, et al.
Published: (2026)
by: He, Muyu, et al.
Published: (2026)
A new approach for fine-tuning sentence transformers for intent classification and out-of-scope detection tasks
by: Zhang, Tianyi, et al.
Published: (2024)
by: Zhang, Tianyi, et al.
Published: (2024)
Counting Without Numbers and Finding Without Words
by: Patro, Badri Narayana
Published: (2026)
by: Patro, Badri Narayana
Published: (2026)
Enhancing Hindi NER in Low Context: A Comparative study of Transformer-based models with vs. without Retrieval Augmentation
by: Singh, Sumit, et al.
Published: (2025)
by: Singh, Sumit, et al.
Published: (2025)
Human-1 by Josh Talks: A Full-Duplex Conversational Modeling Framework in Hindi using Real-World Conversations
by: Singh, Bhaskar, et al.
Published: (2026)
by: Singh, Bhaskar, et al.
Published: (2026)
WordDecipher: Enhancing Digital Workspace Communication with Explainable AI for Non-native English Speakers
by: Chen, Yuexi, et al.
Published: (2024)
by: Chen, Yuexi, et al.
Published: (2024)
'Since Lawyers are Males..': Examining Implicit Gender Bias in Hindi Language Generation by LLMs
by: Joshi, Ishika, et al.
Published: (2024)
by: Joshi, Ishika, et al.
Published: (2024)
Beyond Keywords: A Context-based Hybrid Approach to Mining Ethical Concern-related App Reviews
by: Sorathiya, Aakash, et al.
Published: (2024)
by: Sorathiya, Aakash, et al.
Published: (2024)
ProdRev: A DNN framework for empowering customers using generative pre-trained transformers
by: Gupta, Aakash, et al.
Published: (2025)
by: Gupta, Aakash, et al.
Published: (2025)
TriNER: A Series of Named Entity Recognition Models For Hindi, Bengali & Marathi
by: Dhamaskar, Mohammed Amaan, et al.
Published: (2025)
by: Dhamaskar, Mohammed Amaan, et al.
Published: (2025)
Similar Items
-
Improving code-mixed hate detection by native sample mixing: A case study for Hindi-English code-mixed scenario
by: Mazumder, Debajyoti, et al.
Published: (2024) -
Neither Here Nor There: Cross-Lingual Representation Dynamics of Code-Mixed Text in Multilingual Encoders
by: Mazumder, Debajyoti, et al.
Published: (2026) -
On VLMs for Diverse Tasks in Multimodal Meme Classification
by: Gavit, Deepesh, et al.
Published: (2025) -
An Under-Explored Application for Explainable Multimodal Misogyny Detection in code-mixed Hindi-English
by: Yadav, Sargam, et al.
Published: (2026) -
Entailed Opinion Matters: Improving the Fact-Checking Performance of Language Models by Relying on their Entailment Ability
by: Kumar, Gaurav, et al.
Published: (2025)