Multilingual Controlled Generation And Gold-Standard-Agnostic Evaluation of Code-Mixed Sentences
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gupta, Ayushman, Bhogal, Akhil, Ghosh, Kripabandhu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Code-Mixer Ya Nahi: Novel Approaches to Measuring Multilingual LLMs' Code-Mixing Capabilities
von: Gupta, Ayushman, et al.
Veröffentlicht: (2024)
von: Gupta, Ayushman, et al.
Veröffentlicht: (2024)
Applicability of Large Language Models and Generative Models for Legal Case Judgement Summarization
von: Deroy, Aniket, et al.
Veröffentlicht: (2024)
von: Deroy, Aniket, et al.
Veröffentlicht: (2024)
Structured Legal Document Generation in India: A Model-Agnostic Wrapper Approach with VidhikDastaavej
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2025)
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2025)
Mapping Clinical Doubt: Locating Linguistic Uncertainty in LLMs
von: Sridhar, Srivarshinee, et al.
Veröffentlicht: (2025)
von: Sridhar, Srivarshinee, et al.
Veröffentlicht: (2025)
From Human Judgements to Predictive Models: Unravelling Acceptability in Code-Mixed Sentences
von: Kodali, Prashant, et al.
Veröffentlicht: (2024)
von: Kodali, Prashant, et al.
Veröffentlicht: (2024)
BiST: A Gold Standard Bangla-English Bilingual Corpus for Sentence Structure and Tense Classification with Inter-Annotator Agreement
von: Shafi, Abdullah Al, et al.
Veröffentlicht: (2026)
von: Shafi, Abdullah Al, et al.
Veröffentlicht: (2026)
MARRO: Multi-headed Attention for Rhetorical Role Labeling in Legal Documents
von: Bambroo, Purbid, et al.
Veröffentlicht: (2025)
von: Bambroo, Purbid, et al.
Veröffentlicht: (2025)
Multilingual LLMs Are Not Multilingual Thinkers: Evidence from Hindi Analogy Evaluation
von: Gupta, Ashray, et al.
Veröffentlicht: (2025)
von: Gupta, Ashray, et al.
Veröffentlicht: (2025)
MCQA-Eval: Efficient Confidence Evaluation in NLG with Gold-Standard Correctness Labels
von: Liu, Xiaoou, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoou, et al.
Veröffentlicht: (2025)
ObfusQAte: A Proposed Framework to Evaluate LLM Robustness on Obfuscated Factual Question Answering
von: Ghosh, Shubhra, et al.
Veröffentlicht: (2025)
von: Ghosh, Shubhra, et al.
Veröffentlicht: (2025)
Multilingual Idioms in Sentences and Conversations Across High-, Medium-, and Low-Resource Languages
von: Almheiri, Saeed, et al.
Veröffentlicht: (2026)
von: Almheiri, Saeed, et al.
Veröffentlicht: (2026)
Code-Switching Red-Teaming: LLM Evaluation for Safety and Multilingual Understanding
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
ChiEngMixBench: Evaluating Large Language Models on Spontaneous and Natural Chinese-English Code-Mixed Generation
von: Yang, Qingyan, et al.
Veröffentlicht: (2026)
von: Yang, Qingyan, et al.
Veröffentlicht: (2026)
MORPHOGEN: A Multilingual Benchmark for Evaluating Gender-Aware Morphological Generation
von: Agarwal, Mehul, et al.
Veröffentlicht: (2026)
von: Agarwal, Mehul, et al.
Veröffentlicht: (2026)
Low-Cost Generation and Evaluation of Dictionary Example Sentences
von: Cai, Bill, et al.
Veröffentlicht: (2024)
von: Cai, Bill, et al.
Veröffentlicht: (2024)
CodeScope: An Execution-based Multilingual Multitask Multidimensional Benchmark for Evaluating LLMs on Code Understanding and Generation
von: Yan, Weixiang, et al.
Veröffentlicht: (2023)
von: Yan, Weixiang, et al.
Veröffentlicht: (2023)
MOSAIC: A Multilingual, Taxonomy-Agnostic, and Computationally Efficient Approach for Radiological Report Classification
von: Schiavone, Alice, et al.
Veröffentlicht: (2025)
von: Schiavone, Alice, et al.
Veröffentlicht: (2025)
Enhancing Multilingual Sentiment Analysis with Explainability for Sinhala, English, and Code-Mixed Content
von: Rizvi, Azmarah, et al.
Veröffentlicht: (2025)
von: Rizvi, Azmarah, et al.
Veröffentlicht: (2025)
Generating Diverse Negations from Affirmative Sentences
von: Vasquez, Darian Rodriguez, et al.
Veröffentlicht: (2024)
von: Vasquez, Darian Rodriguez, et al.
Veröffentlicht: (2024)
The Heap: A Contamination-Free Multilingual Code Dataset for Evaluating Large Language Models
von: Katzy, Jonathan, et al.
Veröffentlicht: (2025)
von: Katzy, Jonathan, et al.
Veröffentlicht: (2025)
Evaluating Multilingual and Code-Switched Alignment in LLMs via Synthetic Natural Language Inference
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
Multilingual Multimodal Software Developer for Code Generation
von: Chai, Linzheng, et al.
Veröffentlicht: (2025)
von: Chai, Linzheng, et al.
Veröffentlicht: (2025)
STRUCTSENSE: A Task-Agnostic Agentic Framework for Structured Information Extraction with Human-In-The-Loop Evaluation and Benchmarking
von: Chhetri, Tek Raj, et al.
Veröffentlicht: (2025)
von: Chhetri, Tek Raj, et al.
Veröffentlicht: (2025)
Intertwining CP and NLP: The Generation of Unreasonably Constrained Sentences
von: Bonlarron, Alexandre, et al.
Veröffentlicht: (2024)
von: Bonlarron, Alexandre, et al.
Veröffentlicht: (2024)
Legal Judgment Reimagined: PredEx and the Rise of Intelligent AI Interpretation in Indian Courts
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2024)
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2024)
Label-semantics Aware Generative Approach for Domain-Agnostic Multilabel Classification
von: Khatuya, Subhendu, et al.
Veröffentlicht: (2025)
von: Khatuya, Subhendu, et al.
Veröffentlicht: (2025)
mR3: Multilingual Rubric-Agnostic Reward Reasoning Models
von: Anugraha, David, et al.
Veröffentlicht: (2025)
von: Anugraha, David, et al.
Veröffentlicht: (2025)
WikiNER-fr-gold: A Gold-Standard NER Corpus
von: Cao, Danrun, et al.
Veröffentlicht: (2024)
von: Cao, Danrun, et al.
Veröffentlicht: (2024)
Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities
von: Liu, Zhichen, et al.
Veröffentlicht: (2026)
von: Liu, Zhichen, et al.
Veröffentlicht: (2026)
A Commonsense-Infused Language-Agnostic Learning Framework for Enhancing Prediction of Political Polarity in Multilingual News Headlines
von: Swati, Swati, et al.
Veröffentlicht: (2022)
von: Swati, Swati, et al.
Veröffentlicht: (2022)
Generating Hierarchical JSON Representations of Scientific Sentences Using LLMs
von: Nimmagadda, Satya Sri Rajiteswari, et al.
Veröffentlicht: (2026)
von: Nimmagadda, Satya Sri Rajiteswari, et al.
Veröffentlicht: (2026)
Naamah: A Large Scale Synthetic Sanskrit NER Corpus via DBpedia Seeding and LLM Generation
von: P, Akhil Rajeev, et al.
Veröffentlicht: (2026)
von: P, Akhil Rajeev, et al.
Veröffentlicht: (2026)
Lightweight Domain Adaptation of a Large Language Model for Legal Assistance in the Indian Context
von: Gupta, Jatin, et al.
Veröffentlicht: (2025)
von: Gupta, Jatin, et al.
Veröffentlicht: (2025)
E-ARMOR: Edge case Assessment and Review of Multilingual Optical Character Recognition
von: Gupta, Aryan, et al.
Veröffentlicht: (2025)
von: Gupta, Aryan, et al.
Veröffentlicht: (2025)
Measuring and Reducing LLM Hallucination without Gold-Standard Answers
von: Wei, Jiaheng, et al.
Veröffentlicht: (2024)
von: Wei, Jiaheng, et al.
Veröffentlicht: (2024)
HalluMix: A Task-Agnostic, Multi-Domain Benchmark for Real-World Hallucination Detection
von: Emery, Deanna, et al.
Veröffentlicht: (2025)
von: Emery, Deanna, et al.
Veröffentlicht: (2025)
On the Relationship between Sentence Analogy Identification and Sentence Structure Encoding in Large Language Models
von: Wijesiriwardene, Thilini, et al.
Veröffentlicht: (2023)
von: Wijesiriwardene, Thilini, et al.
Veröffentlicht: (2023)
Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs
von: Li, Bo, et al.
Veröffentlicht: (2026)
von: Li, Bo, et al.
Veröffentlicht: (2026)
MELA: Multilingual Evaluation of Linguistic Acceptability
von: Zhang, Ziyin, et al.
Veröffentlicht: (2023)
von: Zhang, Ziyin, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Code-Mixer Ya Nahi: Novel Approaches to Measuring Multilingual LLMs' Code-Mixing Capabilities
von: Gupta, Ayushman, et al.
Veröffentlicht: (2024) -
Applicability of Large Language Models and Generative Models for Legal Case Judgement Summarization
von: Deroy, Aniket, et al.
Veröffentlicht: (2024) -
Structured Legal Document Generation in India: A Model-Agnostic Wrapper Approach with VidhikDastaavej
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2025) -
Mapping Clinical Doubt: Locating Linguistic Uncertainty in LLMs
von: Sridhar, Srivarshinee, et al.
Veröffentlicht: (2025) -
From Human Judgements to Predictive Models: Unravelling Acceptability in Code-Mixed Sentences
von: Kodali, Prashant, et al.
Veröffentlicht: (2024)