TASER: Translation Assessment via Systematic Evaluation and Reasoning
Fuente:
arXiv
Salvato in:
| Autori principali: | Maheswaran, Monishwaran, Carini, Marco, Federmann, Christian, Diaz, Tony |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Arbitrage: Efficient Reasoning via Advantage-Aware Speculation
di: Maheswaran, Monishwaran, et al.
Pubblicazione: (2025)
di: Maheswaran, Monishwaran, et al.
Pubblicazione: (2025)
TASER: Table Agents for Schema-guided Extraction and Recommendation
di: Cho, Nicole, et al.
Pubblicazione: (2025)
di: Cho, Nicole, et al.
Pubblicazione: (2025)
Residual Context Diffusion Language Models
di: Hu, Yuezhou, et al.
Pubblicazione: (2026)
di: Hu, Yuezhou, et al.
Pubblicazione: (2026)
Are Large Reasoning Models Good Translation Evaluators? Analysis and Performance Boost
di: Zhan, Runzhe, et al.
Pubblicazione: (2025)
di: Zhan, Runzhe, et al.
Pubblicazione: (2025)
DeepTrans: Deep Reasoning Translation via Reinforcement Learning
di: Wang, Jiaan, et al.
Pubblicazione: (2025)
di: Wang, Jiaan, et al.
Pubblicazione: (2025)
DRT: Deep Reasoning Translation via Long Chain-of-Thought
di: Wang, Jiaan, et al.
Pubblicazione: (2024)
di: Wang, Jiaan, et al.
Pubblicazione: (2024)
Reasoners or Translators? Contamination-aware Evaluation and Neuro-Symbolic Robustness in Tax Law
di: Kordjamshidi, Parisa, et al.
Pubblicazione: (2026)
di: Kordjamshidi, Parisa, et al.
Pubblicazione: (2026)
Convergences and Divergences between Automatic Assessment and Human Evaluation: Insights from Comparing ChatGPT-Generated Translation and Neural Machine Translation
di: Jiang, Zhaokun, et al.
Pubblicazione: (2024)
di: Jiang, Zhaokun, et al.
Pubblicazione: (2024)
How to Evaluate Speech Translation with Source-Aware Neural MT Metrics
di: Cettolo, Mauro, et al.
Pubblicazione: (2025)
di: Cettolo, Mauro, et al.
Pubblicazione: (2025)
Learning When to Translate for Multilingual Reasoning
di: Kang, Deokhyung, et al.
Pubblicazione: (2026)
di: Kang, Deokhyung, et al.
Pubblicazione: (2026)
Squeeze Evolve: Unified Multi-Model Orchestration for Verifier-Free Evolution
di: Maheswaran, Monishwaran, et al.
Pubblicazione: (2026)
di: Maheswaran, Monishwaran, et al.
Pubblicazione: (2026)
ExTrans: Multilingual Deep Reasoning Translation via Exemplar-Enhanced Reinforcement Learning
di: Wang, Jiaan, et al.
Pubblicazione: (2025)
di: Wang, Jiaan, et al.
Pubblicazione: (2025)
LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models
di: Parmar, Mihir, et al.
Pubblicazione: (2024)
di: Parmar, Mihir, et al.
Pubblicazione: (2024)
Improving Symbolic Translation of Language Models for Logical Reasoning
di: Thatikonda, Ramya Keerthy, et al.
Pubblicazione: (2026)
di: Thatikonda, Ramya Keerthy, et al.
Pubblicazione: (2026)
End-to-End Evaluation for Low-Latency Simultaneous Speech Translation
di: Huber, Christian, et al.
Pubblicazione: (2023)
di: Huber, Christian, et al.
Pubblicazione: (2023)
Should We be Pedantic About Reasoning Errors in Machine Translation?
di: Bao, Calvin, et al.
Pubblicazione: (2026)
di: Bao, Calvin, et al.
Pubblicazione: (2026)
Unlocking Reasoning Capability on Machine Translation in Large Language Models
di: Rajaee, Sara, et al.
Pubblicazione: (2026)
di: Rajaee, Sara, et al.
Pubblicazione: (2026)
Should I Share this Translation? Evaluating Quality Feedback for User Reliance on Machine Translation
di: Ki, Dayeon, et al.
Pubblicazione: (2025)
di: Ki, Dayeon, et al.
Pubblicazione: (2025)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
di: Choenni, Rochelle, et al.
Pubblicazione: (2024)
di: Choenni, Rochelle, et al.
Pubblicazione: (2024)
How Good Are LLMs for Literary Translation, Really? Literary Translation Evaluation with Humans and LLMs
di: Zhang, Ran, et al.
Pubblicazione: (2024)
di: Zhang, Ran, et al.
Pubblicazione: (2024)
Eliciting Critical Reasoning in Retrieval-Augmented Language Models via Contrastive Explanations
di: Ranaldi, Leonardo, et al.
Pubblicazione: (2024)
di: Ranaldi, Leonardo, et al.
Pubblicazione: (2024)
Span-Level Machine Translation Meta-Evaluation
di: Perrella, Stefano, et al.
Pubblicazione: (2026)
di: Perrella, Stefano, et al.
Pubblicazione: (2026)
Translation as a Scalable Proxy for Multilingual Evaluation
di: Issaka, Sheriff, et al.
Pubblicazione: (2026)
di: Issaka, Sheriff, et al.
Pubblicazione: (2026)
TEaR: Improving LLM-based Machine Translation with Systematic Self-Refinement
di: Feng, Zhaopeng, et al.
Pubblicazione: (2024)
di: Feng, Zhaopeng, et al.
Pubblicazione: (2024)
UTMath: Math Evaluation with Unit Test via Reasoning-to-Coding Thoughts
di: Yang, Bo, et al.
Pubblicazione: (2024)
di: Yang, Bo, et al.
Pubblicazione: (2024)
CARE-RAG - Clinical Assessment and Reasoning in RAG
di: Potluri, Deepthi, et al.
Pubblicazione: (2025)
di: Potluri, Deepthi, et al.
Pubblicazione: (2025)
Déjà Vu: Multilingual LLM Evaluation through the Lens of Machine Translation Evaluation
di: Kreutzer, Julia, et al.
Pubblicazione: (2025)
di: Kreutzer, Julia, et al.
Pubblicazione: (2025)
Beyond Correlation: Interpretable Evaluation of Machine Translation Metrics
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
Evaluation of Machine Translation Based on Semantic Dependencies and Keywords
di: Yuan, Kewei, et al.
Pubblicazione: (2024)
di: Yuan, Kewei, et al.
Pubblicazione: (2024)
FairTranslate: An English-French Dataset for Gender Bias Evaluation in Machine Translation by Overcoming Gender Binarity
di: Jourdan, Fanny, et al.
Pubblicazione: (2025)
di: Jourdan, Fanny, et al.
Pubblicazione: (2025)
PDDL-Mind: Large Language Models are Capable on Belief Reasoning with Reliable State Tracking
di: Zhu, Wang Bill, et al.
Pubblicazione: (2026)
di: Zhu, Wang Bill, et al.
Pubblicazione: (2026)
Reasoning with Natural Language Explanations
di: Valentino, Marco, et al.
Pubblicazione: (2024)
di: Valentino, Marco, et al.
Pubblicazione: (2024)
Evaluating In-Context Translation with Synchronous Context-Free Grammar Transduction
di: Petty, Jackson, et al.
Pubblicazione: (2026)
di: Petty, Jackson, et al.
Pubblicazione: (2026)
An Analysis on Automated Metrics for Evaluating Japanese-English Chat Translation
di: Rusli, Andre, et al.
Pubblicazione: (2024)
di: Rusli, Andre, et al.
Pubblicazione: (2024)
Guardians of the Machine Translation Meta-Evaluation: Sentinel Metrics Fall In!
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
di: Perrella, Stefano, et al.
Pubblicazione: (2024)
Reasoning or Not? A Comprehensive Evaluation of Reasoning LLMs for Dialogue Summarization
di: Jin, Keyan, et al.
Pubblicazione: (2025)
di: Jin, Keyan, et al.
Pubblicazione: (2025)
Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
di: Srivastava, Saurabh, et al.
Pubblicazione: (2024)
di: Srivastava, Saurabh, et al.
Pubblicazione: (2024)
PEIRCE: Unifying Material and Formal Reasoning via LLM-Driven Neuro-Symbolic Refinement
di: Quan, Xin, et al.
Pubblicazione: (2025)
di: Quan, Xin, et al.
Pubblicazione: (2025)
Socratic-PRMBench: Benchmarking Process Reward Models with Systematic Reasoning Patterns
di: Li, Xiang, et al.
Pubblicazione: (2025)
di: Li, Xiang, et al.
Pubblicazione: (2025)
Automating SPARQL Query Translations between DBpedia and Wikidata
di: Bartels, Malte Christian, et al.
Pubblicazione: (2025)
di: Bartels, Malte Christian, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Arbitrage: Efficient Reasoning via Advantage-Aware Speculation
di: Maheswaran, Monishwaran, et al.
Pubblicazione: (2025) -
TASER: Table Agents for Schema-guided Extraction and Recommendation
di: Cho, Nicole, et al.
Pubblicazione: (2025) -
Residual Context Diffusion Language Models
di: Hu, Yuezhou, et al.
Pubblicazione: (2026) -
Are Large Reasoning Models Good Translation Evaluators? Analysis and Performance Boost
di: Zhan, Runzhe, et al.
Pubblicazione: (2025) -
DeepTrans: Deep Reasoning Translation via Reinforcement Learning
di: Wang, Jiaan, et al.
Pubblicazione: (2025)