Assessing and Improving Punctuation Robustness in English-Marathi Machine Translation
Fuente:
arXiv
Guardado en:
| Autores principales: | Shejole, Kaustubh Shivshankar, Deoghare, Sourabh, Bhattacharyya, Pushpak |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
StereoDetect: Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings
por: Shejole, Kaustubh Shivshankar, et al.
Publicado: (2025)
por: Shejole, Kaustubh Shivshankar, et al.
Publicado: (2025)
Giving the Old a Fresh Spin: Quality Estimation-Assisted Constrained Decoding for Automatic Post-Editing
por: Deoghare, Sourabh, et al.
Publicado: (2025)
por: Deoghare, Sourabh, et al.
Publicado: (2025)
Together We Can: Multilingual Automatic Post-Editing for Low-Resource Languages
por: Deoghare, Sourabh, et al.
Publicado: (2024)
por: Deoghare, Sourabh, et al.
Publicado: (2024)
PSSI-MaxST: An Efficient Pixel-Segment Similarity Index Using Intensity and Smoothness Features for Maximum Spanning Tree Based Segmentation
por: Shejole, Kaustubh Shivshankar, et al.
Publicado: (2026)
por: Shejole, Kaustubh Shivshankar, et al.
Publicado: (2026)
Evaluating Extremely Low-Resource Machine Translation: A Comparative Study of ChrF++ and BLEU Metrics
por: Kumar, Sanjeev, et al.
Publicado: (2026)
por: Kumar, Sanjeev, et al.
Publicado: (2026)
A Case Study on Context-Aware Neural Machine Translation with Multi-Task Learning
por: Appicharla, Ramakrishna, et al.
Publicado: (2024)
por: Appicharla, Ramakrishna, et al.
Publicado: (2024)
How effective is Multi-source pivoting for Translation of Low Resource Indian Languages?
por: Gaikwad, Pranav, et al.
Publicado: (2024)
por: Gaikwad, Pranav, et al.
Publicado: (2024)
P-ReMIS: Pragmatic Reasoning in Mental Health and a Social Implication
por: Oram, Sneha, et al.
Publicado: (2025)
por: Oram, Sneha, et al.
Publicado: (2025)
Main Predicate and Their Arguments as Explanation Signals For Intent Classification
por: Pimparkhede, Sameer, et al.
Publicado: (2025)
por: Pimparkhede, Sameer, et al.
Publicado: (2025)
Reconsidering SMT Over NMT for Closely Related Languages: A Case Study of Persian-Hindi Pair
por: Yousofi, Waisullah, et al.
Publicado: (2024)
por: Yousofi, Waisullah, et al.
Publicado: (2024)
Facts-and-Feelings: Capturing both Objectivity and Subjectivity in Table-to-Text Generation
por: Dey, Tathagata, et al.
Publicado: (2024)
por: Dey, Tathagata, et al.
Publicado: (2024)
We Care: Multimodal Depression Detection and Knowledge Infused Mental Health Therapeutic Response Generation
por: Moon, Palash, et al.
Publicado: (2024)
por: Moon, Palash, et al.
Publicado: (2024)
GRAFT: A Graph-based Flow-aware Agentic Framework for Document-level Machine Translation
por: Dutta, Himanshu, et al.
Publicado: (2025)
por: Dutta, Himanshu, et al.
Publicado: (2025)
Punctuation Restoration for Singaporean Spoken Languages: English, Malay, and Mandarin
por: Rao, Abhinav, et al.
Publicado: (2022)
por: Rao, Abhinav, et al.
Publicado: (2022)
Improving Vietnamese-English Medical Machine Translation
por: Vo, Nhu, et al.
Publicado: (2024)
por: Vo, Nhu, et al.
Publicado: (2024)
CoSTA: Code-Switched Speech Translation using Aligned Speech-Text Interleaving
por: Shankar, Bhavani, et al.
Publicado: (2024)
por: Shankar, Bhavani, et al.
Publicado: (2024)
RoundTripOCR: A Data Generation Technique for Enhancing Post-OCR Error Correction in Low-Resource Devanagari Languages
por: Kashid, Harshvivek, et al.
Publicado: (2024)
por: Kashid, Harshvivek, et al.
Publicado: (2024)
Enhancing Food-Domain Question Answering with a Multimodal Knowledge Graph: Hybrid QA Generation and Diversity Analysis
por: B, Srihari K, et al.
Publicado: (2025)
por: B, Srihari K, et al.
Publicado: (2025)
Punctuation Restoration Improves Structure Understanding Without Supervision
por: Min, Junghyun, et al.
Publicado: (2024)
por: Min, Junghyun, et al.
Publicado: (2024)
Topic Modeling in Marathi
por: Shinde, Sanket, et al.
Publicado: (2025)
por: Shinde, Sanket, et al.
Publicado: (2025)
Improving Text Style Transfer using Masked Diffusion Language Models with Inference-time Scaling
por: Padole, Tejomay Kishor, et al.
Publicado: (2025)
por: Padole, Tejomay Kishor, et al.
Publicado: (2025)
Stereotype Detection as a Catalyst for Enhanced Bias Detection: A Multi-Task Learning Approach
por: Tomar, Aditya, et al.
Publicado: (2025)
por: Tomar, Aditya, et al.
Publicado: (2025)
Ta-G-T: Subjectivity Capture in Table to Text Generation via RDF Graphs
por: Upasham, Ronak, et al.
Publicado: (2025)
por: Upasham, Ronak, et al.
Publicado: (2025)
Pretraining Language Models Using Translationese
por: Doshi, Meet, et al.
Publicado: (2024)
por: Doshi, Meet, et al.
Publicado: (2024)
Are Language Models Agnostic to Linguistically Grounded Perturbations? A Case Study of Indic Languages
por: Ghosh, Poulami, et al.
Publicado: (2024)
por: Ghosh, Poulami, et al.
Publicado: (2024)
Towards Emotion Consistency Analysis of Large Language Models in Emotional Conversational Contexts
por: Oram, Sneha, et al.
Publicado: (2026)
por: Oram, Sneha, et al.
Publicado: (2026)
Setting up the Data Printer with Improved English to Ukrainian Machine Translation
por: Paniv, Yurii, et al.
Publicado: (2024)
por: Paniv, Yurii, et al.
Publicado: (2024)
PUB: A Pragmatics Understanding Benchmark for Assessing LLMs' Pragmatics Capabilities
por: Sravanthi, Settaluri Lakshmi, et al.
Publicado: (2024)
por: Sravanthi, Settaluri Lakshmi, et al.
Publicado: (2024)
BharatBBQ: A Multilingual Bias Benchmark for Question Answering in the Indian Context
por: Tomar, Aditya, et al.
Publicado: (2025)
por: Tomar, Aditya, et al.
Publicado: (2025)
Precision Empowers, Excess Distracts: Visual Question Answering With Dynamically Infused Knowledge In Language Models
por: Jhalani, Manas, et al.
Publicado: (2024)
por: Jhalani, Manas, et al.
Publicado: (2024)
Improving Neural Machine Translation Through Code‐Mixed Data Augmentation
por: Ramakrishna Appicharla, et al.
Publicado: (2025)
por: Ramakrishna Appicharla, et al.
Publicado: (2025)
Recon, Answer, Verify: Agents in Search of Truth
por: Shukla, Satyam, et al.
Publicado: (2025)
por: Shukla, Satyam, et al.
Publicado: (2025)
Punctuation-aware treebank tree binarization
por: Klinger, Eitan, et al.
Publicado: (2025)
por: Klinger, Eitan, et al.
Publicado: (2025)
The Role of Handling Attributive Nouns in Improving Chinese-To-English Machine Translation
por: Wang, Lisa, et al.
Publicado: (2024)
por: Wang, Lisa, et al.
Publicado: (2024)
A Morphology-Based Investigation of Positional Encodings
por: Ghosh, Poulami, et al.
Publicado: (2024)
por: Ghosh, Poulami, et al.
Publicado: (2024)
"My life is miserable, have to sign 500 autographs everyday": Exposing Humblebragging, the Brags in Disguise
por: Naganna, Sharath, et al.
Publicado: (2024)
por: Naganna, Sharath, et al.
Publicado: (2024)
Expect the unexpected: Harnessing Sentence Completion for Sarcasm Detection
por: Joshi, Aditya, et al.
Publicado: (2017)
por: Joshi, Aditya, et al.
Publicado: (2017)
Mark My Words: A Robust Multilingual Model for Punctuation in Text and Speech Transcripts
por: Pulipaka, Sidharth, et al.
Publicado: (2025)
por: Pulipaka, Sidharth, et al.
Publicado: (2025)
Punctuation and Predicates in Language Models
por: Chauhan, Sonakshi, et al.
Publicado: (2025)
por: Chauhan, Sonakshi, et al.
Publicado: (2025)
Looks can be Deceptive: Distinguishing Repetition Disfluency from Reduplication
por: Ahmad, Arif, et al.
Publicado: (2024)
por: Ahmad, Arif, et al.
Publicado: (2024)
Ejemplares similares
-
StereoDetect: Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings
por: Shejole, Kaustubh Shivshankar, et al.
Publicado: (2025) -
Giving the Old a Fresh Spin: Quality Estimation-Assisted Constrained Decoding for Automatic Post-Editing
por: Deoghare, Sourabh, et al.
Publicado: (2025) -
Together We Can: Multilingual Automatic Post-Editing for Low-Resource Languages
por: Deoghare, Sourabh, et al.
Publicado: (2024) -
PSSI-MaxST: An Efficient Pixel-Segment Similarity Index Using Intensity and Smoothness Features for Maximum Spanning Tree Based Segmentation
por: Shejole, Kaustubh Shivshankar, et al.
Publicado: (2026) -
Evaluating Extremely Low-Resource Machine Translation: A Comparative Study of ChrF++ and BLEU Metrics
por: Kumar, Sanjeev, et al.
Publicado: (2026)