FUSE : A Ridge and Random Forest-Based Metric for Evaluating MT in Indigenous Languages
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Raja, Rahul, Vats, Arpita |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Parallel Corpora for Machine Translation in Low-resource Indic Languages: A Comprehensive Review
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
Multimedia-Aware Question Answering: A Review of Retrieval and Cross-Modal Reasoning Architectures
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
Multilingual State Space Models for Structured Question Answering in Indic Languages
von: Vats, Arpita, et al.
Veröffentlicht: (2025)
von: Vats, Arpita, et al.
Veröffentlicht: (2025)
Counterfactual Risk Minimization with IPS-Weighted BPR and Self-Normalized Evaluation in Recommender Systems
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
Evaluating Generalization and Representation Stability in Small LMs via Prompting, Fine-Tuning and Out-of-Distribution Prompts
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
Bridging the Semantic Gap: Contrastive Rewards for Multilingual Text-to-SQL with GRPO
von: Kattamuri, Ashish, et al.
Veröffentlicht: (2025)
von: Kattamuri, Ashish, et al.
Veröffentlicht: (2025)
Equilibrium Dynamics and Mitigation of Gender Bias in Synthetically Generated Data
von: Kattamuri, Ashish, et al.
Veröffentlicht: (2025)
von: Kattamuri, Ashish, et al.
Veröffentlicht: (2025)
Beyond Nearest Neighbors: Semantic Compression and Graph-Augmented Retrieval for Enhanced Vector Search
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
A Comprehensive Review on Harnessing Large Language Models to Overcome Recommender System Challenges
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
Faithful Model Evaluation for Model-Based Metrics
von: Goyal, Palash, et al.
Veröffentlicht: (2023)
von: Goyal, Palash, et al.
Veröffentlicht: (2023)
Dynamic Meta-Metrics: Source-Sentence Conditioned Weighting for MT Evaluation
von: Zhang, Luke, et al.
Veröffentlicht: (2026)
von: Zhang, Luke, et al.
Veröffentlicht: (2026)
FUSE: Ensembling Verifiers with Zero Labeled Data
von: Lee, Joonhyuk, et al.
Veröffentlicht: (2026)
von: Lee, Joonhyuk, et al.
Veröffentlicht: (2026)
Exploring the Impact of Large Language Models on Recommender Systems: An Extensive Review
von: Vats, Arpita, et al.
Veröffentlicht: (2024)
von: Vats, Arpita, et al.
Veröffentlicht: (2024)
How to Evaluate Speech Translation with Source-Aware Neural MT Metrics
von: Cettolo, Mauro, et al.
Veröffentlicht: (2025)
von: Cettolo, Mauro, et al.
Veröffentlicht: (2025)
SSA-COMET: Do LLMs Outperform Learned Metrics in Evaluating MT for Under-Resourced African Languages?
von: Li, Senyu, et al.
Veröffentlicht: (2025)
von: Li, Senyu, et al.
Veröffentlicht: (2025)
PromptOptMe: Error-Aware Prompt Compression for LLM-based MT Evaluation Metrics
von: Larionov, Daniil, et al.
Veröffentlicht: (2024)
von: Larionov, Daniil, et al.
Veröffentlicht: (2024)
xCOMET-lite: Bridging the Gap Between Efficiency and Quality in Learned MT Evaluation Metrics
von: Larionov, Daniil, et al.
Veröffentlicht: (2024)
von: Larionov, Daniil, et al.
Veröffentlicht: (2024)
FUSE-ing Language Models: Zero-Shot Adapter Discovery for Prompt Optimization Across Tokenizers
von: Williams, Joshua Nathaniel, et al.
Veröffentlicht: (2024)
von: Williams, Joshua Nathaniel, et al.
Veröffentlicht: (2024)
AFRIDOC-MT: Document-level MT Corpus for African Languages
von: Alabi, Jesujoba O., et al.
Veröffentlicht: (2025)
von: Alabi, Jesujoba O., et al.
Veröffentlicht: (2025)
MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models
von: Kwan, Wai-Chung, et al.
Veröffentlicht: (2024)
von: Kwan, Wai-Chung, et al.
Veröffentlicht: (2024)
Unmask It! AI-Generated Product Review Detection in Dravidian Languages
von: De, Somsubhra, et al.
Veröffentlicht: (2025)
von: De, Somsubhra, et al.
Veröffentlicht: (2025)
How Good is Zero-Shot MT Evaluation for Low Resource Indian Languages?
von: Singh, Anushka, et al.
Veröffentlicht: (2024)
von: Singh, Anushka, et al.
Veröffentlicht: (2024)
RADAR: Mechanistic Pathways for Detecting Data Contamination in LLM Evaluation
von: Kattamuri, Ashish, et al.
Veröffentlicht: (2025)
von: Kattamuri, Ashish, et al.
Veröffentlicht: (2025)
Kreyòl-MT: Building MT for Latin American, Caribbean and Colonial African Creole Languages
von: Robinson, Nathaniel R., et al.
Veröffentlicht: (2024)
von: Robinson, Nathaniel R., et al.
Veröffentlicht: (2024)
APPLS: Evaluating Evaluation Metrics for Plain Language Summarization
von: Guo, Yue, et al.
Veröffentlicht: (2023)
von: Guo, Yue, et al.
Veröffentlicht: (2023)
LLMs Are Not Scorers: Rethinking MT Evaluation with Generation-Based Methods
von: Cui, Hyang
Veröffentlicht: (2025)
von: Cui, Hyang
Veröffentlicht: (2025)
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
von: Anugraha, David, et al.
Veröffentlicht: (2024)
von: Anugraha, David, et al.
Veröffentlicht: (2024)
Calibrating Model-Based Evaluation Metrics for Summarization
von: Liu, Hongye, et al.
Veröffentlicht: (2026)
von: Liu, Hongye, et al.
Veröffentlicht: (2026)
VerChol -- Grammar-First Tokenization for Agglutinative Languages
von: Raja, Prabhu
Veröffentlicht: (2026)
von: Raja, Prabhu
Veröffentlicht: (2026)
Alignment Quality Index (AQI) : Beyond Refusals: AQI as an Intrinsic Alignment Diagnostic via Latent Geometry, Cluster Divergence, and Layer wise Pooled Representations
von: Borah, Abhilekh, et al.
Veröffentlicht: (2025)
von: Borah, Abhilekh, et al.
Veröffentlicht: (2025)
The power of Prompts: Evaluating and Mitigating Gender Bias in MT with LLMs
von: Sant, Aleix, et al.
Veröffentlicht: (2024)
von: Sant, Aleix, et al.
Veröffentlicht: (2024)
Evaluating the Role of Large Language Models in Legal Practice in India
von: Hemrajani, Rahul
Veröffentlicht: (2025)
von: Hemrajani, Rahul
Veröffentlicht: (2025)
FUSE : Failure-aware Usage of Subagent Evidence for MultiModal Search and Recommendation
von: Vatsa, Tushar, et al.
Veröffentlicht: (2025)
von: Vatsa, Tushar, et al.
Veröffentlicht: (2025)
NLP Progress in Indigenous Latin American Languages
von: Tonja, Atnafu Lambebo, et al.
Veröffentlicht: (2024)
von: Tonja, Atnafu Lambebo, et al.
Veröffentlicht: (2024)
EthioMT: Parallel Corpus for Low-resource Ethiopian Languages
von: Tonja, Atnafu Lambebo, et al.
Veröffentlicht: (2024)
von: Tonja, Atnafu Lambebo, et al.
Veröffentlicht: (2024)
Automatic Metrics in Natural Language Generation: A Survey of Current Evaluation Practices
von: Schmidtová, Patrícia, et al.
Veröffentlicht: (2024)
von: Schmidtová, Patrícia, et al.
Veröffentlicht: (2024)
MT-LENS: An all-in-one Toolkit for Better Machine Translation Evaluation
von: Gilabert, Javier García, et al.
Veröffentlicht: (2024)
von: Gilabert, Javier García, et al.
Veröffentlicht: (2024)
MMTE: Corpus and Metrics for Evaluating Machine Translation Quality of Metaphorical Language
von: Wang, Shun, et al.
Veröffentlicht: (2024)
von: Wang, Shun, et al.
Veröffentlicht: (2024)
MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn Dialogues
von: Bai, Ge, et al.
Veröffentlicht: (2024)
von: Bai, Ge, et al.
Veröffentlicht: (2024)
Gender Bias in MT for a Genderless Language: New Benchmarks for Basque
von: Murillo, Amaia, et al.
Veröffentlicht: (2026)
von: Murillo, Amaia, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Parallel Corpora for Machine Translation in Low-resource Indic Languages: A Comprehensive Review
von: Raja, Rahul, et al.
Veröffentlicht: (2025) -
Multimedia-Aware Question Answering: A Review of Retrieval and Cross-Modal Reasoning Architectures
von: Raja, Rahul, et al.
Veröffentlicht: (2025) -
Multilingual State Space Models for Structured Question Answering in Indic Languages
von: Vats, Arpita, et al.
Veröffentlicht: (2025) -
Counterfactual Risk Minimization with IPS-Weighted BPR and Self-Normalized Evaluation in Recommender Systems
von: Raja, Rahul, et al.
Veröffentlicht: (2025) -
Evaluating Generalization and Representation Stability in Small LMs via Prompting, Fine-Tuning and Out-of-Distribution Prompts
von: Raja, Rahul, et al.
Veröffentlicht: (2025)