Does Dependency Locality Predict Non-canonical Word Order in Hindi?
Fuente:
arXiv
Guardado en:
| Autores principales: | Ranjan, Sidharth, van Schijndel, Marten |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Code-switching in text and speech challenges information-theoretic speaker design
por: Bhattacharya, Debasmita, et al.
Publicado: (2024)
por: Bhattacharya, Debasmita, et al.
Publicado: (2024)
Semantics or spelling? Probing contextual word embeddings with orthographic noise
por: Matthews, Jacob A., et al.
Publicado: (2024)
por: Matthews, Jacob A., et al.
Publicado: (2024)
Work Smarter...Not Harder: Efficient Minimization of Dependency Length in SOV Languages
por: Ranjan, Sidharth, et al.
Publicado: (2024)
por: Ranjan, Sidharth, et al.
Publicado: (2024)
One Word Is Not Enough: Simple Prompts Improve Word Embeddings
por: Ranjan, Rajeev
Publicado: (2025)
por: Ranjan, Rajeev
Publicado: (2025)
Multilingual Gradient Word-Order Typology from Universal Dependencies
por: Baylor, Emi, et al.
Publicado: (2024)
por: Baylor, Emi, et al.
Publicado: (2024)
HindiLLM: Large Language Model for Hindi
por: Chouhan, Sanjay, et al.
Publicado: (2024)
por: Chouhan, Sanjay, et al.
Publicado: (2024)
Mark My Words: A Robust Multilingual Model for Punctuation in Text and Speech Transcripts
por: Pulipaka, Sidharth, et al.
Publicado: (2025)
por: Pulipaka, Sidharth, et al.
Publicado: (2025)
Hindi-BEIR : A Large Scale Retrieval Benchmark in Hindi
por: Acharya, Arkadeep, et al.
Publicado: (2024)
por: Acharya, Arkadeep, et al.
Publicado: (2024)
Does Incomplete Syntax Influence Korean Language Model? Focusing on Word Order and Case Markers
por: Kim, Jong Myoung, et al.
Publicado: (2024)
por: Kim, Jong Myoung, et al.
Publicado: (2024)
Benchmarking and Building Zero-Shot Hindi Retrieval Model with Hindi-BEIR and NLLB-E5
por: Acharya, Arkadeep, et al.
Publicado: (2024)
por: Acharya, Arkadeep, et al.
Publicado: (2024)
Text Detoxification as Style Transfer in English and Hindi
por: Mukherjee, Sourabrata, et al.
Publicado: (2024)
por: Mukherjee, Sourabrata, et al.
Publicado: (2024)
Word Order and World Knowledge
por: Zhao, Qinghua, et al.
Publicado: (2024)
por: Zhao, Qinghua, et al.
Publicado: (2024)
Suvach -- Generated Hindi QA benchmark
por: Narayanan, Vaishak, et al.
Publicado: (2024)
por: Narayanan, Vaishak, et al.
Publicado: (2024)
HiMed: Incentivizing Hindi Reasoning in Medical LLMs
por: Jiang, Dingfeng, et al.
Publicado: (2026)
por: Jiang, Dingfeng, et al.
Publicado: (2026)
Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers
por: Vijayakumar, Soniya, et al.
Publicado: (2024)
por: Vijayakumar, Soniya, et al.
Publicado: (2024)
Automatic Speech Recognition for Hindi
por: Saha, Anish, et al.
Publicado: (2024)
por: Saha, Anish, et al.
Publicado: (2024)
Samasāmayik: A Parallel Dataset for Hindi-Sanskrit Machine Translation
por: Karthika, N J, et al.
Publicado: (2026)
por: Karthika, N J, et al.
Publicado: (2026)
Counter Turing Test ($CT^2$): Investigating AI-Generated Text Detection for Hindi -- Ranking LLMs based on Hindi AI Detectability Index ($ADI_{hi}$)
por: Kavathekar, Ishan, et al.
Publicado: (2024)
por: Kavathekar, Ishan, et al.
Publicado: (2024)
Losing Phonotactic Distinctions in Context
por: John R. Starr, et al.
Publicado: (2025)
por: John R. Starr, et al.
Publicado: (2025)
Airavata: Introducing Hindi Instruction-tuned LLM
por: Gala, Jay, et al.
Publicado: (2024)
por: Gala, Jay, et al.
Publicado: (2024)
Breaking Language Barriers: A Question Answering Dataset for Hindi and Marathi
por: Sabane, Maithili, et al.
Publicado: (2023)
por: Sabane, Maithili, et al.
Publicado: (2023)
Modeling Romanized Hindi and Bengali: Dataset Creation and Multilingual LLM Integration
por: Gharami, Kanchon, et al.
Publicado: (2025)
por: Gharami, Kanchon, et al.
Publicado: (2025)
HLDC: Hindi Legal Documents Corpus
por: Kapoor, Arnav, et al.
Publicado: (2022)
por: Kapoor, Arnav, et al.
Publicado: (2022)
Multi-class Regret Detection in Hindi Devanagari Script
por: Sharma, Renuka, et al.
Publicado: (2024)
por: Sharma, Renuka, et al.
Publicado: (2024)
On Support Samples of Next Word Prediction
por: Li, Yuqian, et al.
Publicado: (2025)
por: Li, Yuqian, et al.
Publicado: (2025)
LAHAJA: A Robust Multi-accent Benchmark for Evaluating Hindi ASR Systems
por: Javed, Tahir, et al.
Publicado: (2024)
por: Javed, Tahir, et al.
Publicado: (2024)
Mixed-Distil-BERT: Code-mixed Language Modeling for Bangla, English, and Hindi
por: Raihan, Md Nishat, et al.
Publicado: (2023)
por: Raihan, Md Nishat, et al.
Publicado: (2023)
Emergent Word Order Universals from Cognitively-Motivated Language Models
por: Kuribayashi, Tatsuki, et al.
Publicado: (2024)
por: Kuribayashi, Tatsuki, et al.
Publicado: (2024)
On the Interplay between Positional Encodings, Morphological Complexity, and Word Order Flexibility
por: Tatariya, Kushal, et al.
Publicado: (2025)
por: Tatariya, Kushal, et al.
Publicado: (2025)
Low-Resource Counterspeech Generation for Indic Languages: The Case of Bengali and Hindi
por: Das, Mithun, et al.
Publicado: (2024)
por: Das, Mithun, et al.
Publicado: (2024)
Survey of Pseudonymization, Abstractive Summarization & Spell Checker for Hindi and Marathi
por: Ransing, Rasika, et al.
Publicado: (2024)
por: Ransing, Rasika, et al.
Publicado: (2024)
Label Words as Local Task Vectors in In-Context Learning
por: Zheng, Bowen, et al.
Publicado: (2024)
por: Zheng, Bowen, et al.
Publicado: (2024)
Toward a Better Localization of Princeton WordNet
por: Freihat, Abed Alhakim
Publicado: (2025)
por: Freihat, Abed Alhakim
Publicado: (2025)
Akal Badi ya Bias: An Exploratory Study of Gender Bias in Hindi Language Technology
por: Hada, Rishav, et al.
Publicado: (2024)
por: Hada, Rishav, et al.
Publicado: (2024)
Axis Tour: Word Tour Determines the Order of Axes in ICA-transformed Embeddings
por: Yamagiwa, Hiroaki, et al.
Publicado: (2024)
por: Yamagiwa, Hiroaki, et al.
Publicado: (2024)
Detection of Non-recorded Word Senses in English and Swedish
por: Lautenschlager, Jonathan, et al.
Publicado: (2024)
por: Lautenschlager, Jonathan, et al.
Publicado: (2024)
Non-literal Understanding of Number Words by Language Models
por: Tsvilodub, Polina, et al.
Publicado: (2025)
por: Tsvilodub, Polina, et al.
Publicado: (2025)
How Well Does First-Token Entropy Approximate Word Entropy as a Psycholinguistic Predictor?
por: Clark, Christian, et al.
Publicado: (2025)
por: Clark, Christian, et al.
Publicado: (2025)
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English
por: Anand, Avinash, et al.
Publicado: (2024)
por: Anand, Avinash, et al.
Publicado: (2024)
Multilingual LLMs Are Not Multilingual Thinkers: Evidence from Hindi Analogy Evaluation
por: Gupta, Ashray, et al.
Publicado: (2025)
por: Gupta, Ashray, et al.
Publicado: (2025)
Ejemplares similares
-
Code-switching in text and speech challenges information-theoretic speaker design
por: Bhattacharya, Debasmita, et al.
Publicado: (2024) -
Semantics or spelling? Probing contextual word embeddings with orthographic noise
por: Matthews, Jacob A., et al.
Publicado: (2024) -
Work Smarter...Not Harder: Efficient Minimization of Dependency Length in SOV Languages
por: Ranjan, Sidharth, et al.
Publicado: (2024) -
One Word Is Not Enough: Simple Prompts Improve Word Embeddings
por: Ranjan, Rajeev
Publicado: (2025) -
Multilingual Gradient Word-Order Typology from Universal Dependencies
por: Baylor, Emi, et al.
Publicado: (2024)