Automatic Normalization of Word Variations in Code-Mixed Social Media Text
Fuente:
arXiv
Saved in:
| Main Authors: | Singh, Rajat, Choudhary, Nurendra, Shrivastava, Manish |
|---|---|
| Format: | Preprint |
| Published: |
2018
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sentiment Analysis of Code-Mixed Languages leveraging Resource Rich Languages
by: Choudhary, Nurendra, et al.
Published: (2018)
by: Choudhary, Nurendra, et al.
Published: (2018)
Emotions are Universal: Learning Sentiment Based Representations of Resource-Poor Languages using Siamese Networks
by: Choudhary, Nurendra, et al.
Published: (2018)
by: Choudhary, Nurendra, et al.
Published: (2018)
Neural Network Architecture for Credibility Assessment of Textual Claims
by: Choudhary, Nurendra, et al.
Published: (2018)
by: Choudhary, Nurendra, et al.
Published: (2018)
Contrastive Learning of Emoji-based Representations for Resource-Poor Languages
by: Choudhary, Nurendra, et al.
Published: (2018)
by: Choudhary, Nurendra, et al.
Published: (2018)
Adapting Multilingual Models to Code-Mixed Tasks via Model Merging
by: Kodali, Prashant, et al.
Published: (2025)
by: Kodali, Prashant, et al.
Published: (2025)
Reasoning-Guided Claim Normalization for Noisy Multilingual Social Media Posts
by: Sharma, Manan, et al.
Published: (2025)
by: Sharma, Manan, et al.
Published: (2025)
Segmentation Beyond Defaults: Asymmetrical Byte Pair Encoding for Optimal Machine Translation Performance
by: Yadav, Saumitra, et al.
Published: (2025)
by: Yadav, Saumitra, et al.
Published: (2025)
Get away with less: Need of source side data curation to build parallel corpus for low resource Machine Translation
by: Yadav, Saumitra, et al.
Published: (2026)
by: Yadav, Saumitra, et al.
Published: (2026)
Can Constructions "SCAN" Compositionality ?
by: Katrapati, Ganesh, et al.
Published: (2025)
by: Katrapati, Ganesh, et al.
Published: (2025)
From Human Judgements to Predictive Models: Unravelling Acceptability in Code-Mixed Sentences
by: Kodali, Prashant, et al.
Published: (2024)
by: Kodali, Prashant, et al.
Published: (2024)
A Survey of using Large Language Models for Generating Infrastructure as Code
by: Srivatsa, Kalahasti Ganesh, et al.
Published: (2024)
by: Srivatsa, Kalahasti Ganesh, et al.
Published: (2024)
Exploiting Dialect Identification in Automatic Dialectal Text Normalization
by: Alhafni, Bashar, et al.
Published: (2024)
by: Alhafni, Bashar, et al.
Published: (2024)
ViLexNorm: A Lexical Normalization Corpus for Vietnamese Social Media Text
by: Nguyen, Thanh-Nhi, et al.
Published: (2024)
by: Nguyen, Thanh-Nhi, et al.
Published: (2024)
MARCUS: An Event-Centric NLP Pipeline that generates Character Arcs from Narratives
by: Bhyravajjula, Sriharsh, et al.
Published: (2025)
by: Bhyravajjula, Sriharsh, et al.
Published: (2025)
Detecting Propaganda Techniques in Code-Switched Social Media Text
by: Salman, Muhammad Umar, et al.
Published: (2023)
by: Salman, Muhammad Umar, et al.
Published: (2023)
Word Chain Generators for Prefix Normal Words
by: Adamson, Duncan, et al.
Published: (2025)
by: Adamson, Duncan, et al.
Published: (2025)
Crowd Intelligence for Early Misinformation Prediction on Social Media
by: Sundriyal, Megha, et al.
Published: (2024)
by: Sundriyal, Megha, et al.
Published: (2024)
GT2Vec: Large Language Models as Multi-Modal Encoders for Text and Graph-Structured Data
by: Lin, Jiacheng, et al.
Published: (2024)
by: Lin, Jiacheng, et al.
Published: (2024)
An Interpretable Ensemble of Graph and Language Models for Improving Search Relevance in E-Commerce
by: Choudhary, Nurendra, et al.
Published: (2024)
by: Choudhary, Nurendra, et al.
Published: (2024)
Neural Text Normalization for Luxembourgish using Real-Life Variation Data
by: Lutgen, Anne-Marie, et al.
Published: (2024)
by: Lutgen, Anne-Marie, et al.
Published: (2024)
Zero-Shot Multi-task Hallucination Detection
by: Bhamidipati, Patanjali, et al.
Published: (2024)
by: Bhamidipati, Patanjali, et al.
Published: (2024)
TeClass: A Human-Annotated Relevance-based Headline Classification and Generation Dataset for Telugu
by: Kanumolu, Gopichand, et al.
Published: (2024)
by: Kanumolu, Gopichand, et al.
Published: (2024)
Mast Kalandar at SemEval-2024 Task 8: On the Trail of Textual Origins: RoBERTa-BiLSTM Approach to Detect AI-Generated Text
by: Bafna, Jainit Sushil, et al.
Published: (2024)
by: Bafna, Jainit Sushil, et al.
Published: (2024)
HarmPot: An Annotation Framework for Evaluating Offline Harm Potential of Social Media Text
by: Kumar, Ritesh, et al.
Published: (2024)
by: Kumar, Ritesh, et al.
Published: (2024)
Lost in the Mix: Evaluating LLM Understanding of Code-Switched Text
by: Mohamed, Amr, et al.
Published: (2025)
by: Mohamed, Amr, et al.
Published: (2025)
DaVinci at SemEval-2024 Task 9: Few-shot prompting GPT-3.5 for Unconventional Reasoning
by: Mathur, Suyash Vardhan, et al.
Published: (2024)
by: Mathur, Suyash Vardhan, et al.
Published: (2024)
A Primer on Word Embeddings: AI Techniques for Text Analysis in Social Work
by: Perron, Brian E., et al.
Published: (2024)
by: Perron, Brian E., et al.
Published: (2024)
TabXEval: Why this is a Bad Table? An eXhaustive Rubric for Table Evaluation
by: Pancholi, Vihang, et al.
Published: (2025)
by: Pancholi, Vihang, et al.
Published: (2025)
Code-Mixed Probes Show How Pre-Trained Models Generalise On Code-Switched Text
by: De Leon, Frances A. Laureano, et al.
Published: (2024)
by: De Leon, Frances A. Laureano, et al.
Published: (2024)
Benchmarking Machine Translation on Chinese Social Media Texts
by: Zhao, Kaiyan, et al.
Published: (2026)
by: Zhao, Kaiyan, et al.
Published: (2026)
What talking you?: Translating Code-Mixed Messaging Texts to English
by: Ng, Lynnette Hui Xian, et al.
Published: (2024)
by: Ng, Lynnette Hui Xian, et al.
Published: (2024)
Multi-task Prompt Words Learning for Social Media Content Generation
by: Xue, Haochen, et al.
Published: (2024)
by: Xue, Haochen, et al.
Published: (2024)
Automatic Coding at Scale: Design and Deployment of a Nationwide System for Normalizing Referrals in the Chilean Public Healthcare System
by: Villena, Fabián, et al.
Published: (2023)
by: Villena, Fabián, et al.
Published: (2023)
Complex Logical Reasoning over Knowledge Graphs using Large Language Models
by: Choudhary, Nurendra, et al.
Published: (2023)
by: Choudhary, Nurendra, et al.
Published: (2023)
Prompt Engineering Using GPT for Word-Level Code-Mixed Language Identification in Low-Resource Dravidian Languages
by: Deroy, Aniket, et al.
Published: (2024)
by: Deroy, Aniket, et al.
Published: (2024)
Automatic Textual Normalization for Hate Speech Detection
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2023)
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2023)
NarrationDep: Narratives on Social Media For Automatic Depression Detection
by: Zogan, Hamad, et al.
Published: (2024)
by: Zogan, Hamad, et al.
Published: (2024)
MultiSocial: Multilingual Benchmark of Machine-Generated Text Detection of Social-Media Texts
by: Macko, Dominik, et al.
Published: (2024)
by: Macko, Dominik, et al.
Published: (2024)
Online Social Support Detection in Spanish Social Media Texts
by: Tash, Moein Shahiki, et al.
Published: (2025)
by: Tash, Moein Shahiki, et al.
Published: (2025)
IndRegBias: A Dataset for Studying Indian Regional Biases in English and Code-Mixed Social Media Comments
by: Panda, Debasmita, et al.
Published: (2026)
by: Panda, Debasmita, et al.
Published: (2026)
Similar Items
-
Sentiment Analysis of Code-Mixed Languages leveraging Resource Rich Languages
by: Choudhary, Nurendra, et al.
Published: (2018) -
Emotions are Universal: Learning Sentiment Based Representations of Resource-Poor Languages using Siamese Networks
by: Choudhary, Nurendra, et al.
Published: (2018) -
Neural Network Architecture for Credibility Assessment of Textual Claims
by: Choudhary, Nurendra, et al.
Published: (2018) -
Contrastive Learning of Emoji-based Representations for Resource-Poor Languages
by: Choudhary, Nurendra, et al.
Published: (2018) -
Adapting Multilingual Models to Code-Mixed Tasks via Model Merging
by: Kodali, Prashant, et al.
Published: (2025)