IsoChronoMeter: A simple and effective isochronic translation evaluation metric
Fuente:
arXiv
Saved in:
| Main Authors: | Rozanov, Nikolai, Pankov, Vikentiy, Mukhutdinov, Dmitrii, Vypirailenko, Dima |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
StateAct: Enhancing LLM Base Agents via Self-prompting and State-tracking
by: Rozanov, Nikolai, et al.
Published: (2024)
by: Rozanov, Nikolai, et al.
Published: (2024)
Fine-tuning with RAG for Improving LLM Learning of New Skills
by: Ibrahim, Humaid, et al.
Published: (2025)
by: Ibrahim, Humaid, et al.
Published: (2025)
SUS backprop: linear backpropagation algorithm for long inputs in transformers
by: Pankov, Sergey, et al.
Published: (2025)
by: Pankov, Sergey, et al.
Published: (2025)
Same evaluation, more tokens: On the effect of input length for machine translation evaluation using Large Language Models
by: Domhan, Tobias, et al.
Published: (2025)
by: Domhan, Tobias, et al.
Published: (2025)
Automated evaluation of LLMs for effective machine translation of Mandarin Chinese to English
by: Zhang, Yue, et al.
Published: (2026)
by: Zhang, Yue, et al.
Published: (2026)
Fineweb-Edu-Ar: Machine-translated Corpus to Support Arabic Small Language Models
by: Alrashed, Sultan, et al.
Published: (2024)
by: Alrashed, Sultan, et al.
Published: (2024)
ChronoFact: Timeline-based Temporal Fact Verification
by: Barik, Anab Maulana, et al.
Published: (2024)
by: Barik, Anab Maulana, et al.
Published: (2024)
PFluxTTS: Hybrid Flow-Matching TTS with Robust Cross-Lingual Voice Cloning and Inference-Time Model Fusion
by: Pankov, Vikentii, et al.
Published: (2026)
by: Pankov, Vikentii, et al.
Published: (2026)
Advancing LLM detection in the ALTA 2024 Shared Task: Techniques and Analysis
by: Galat, Dima
Published: (2024)
by: Galat, Dima
Published: (2024)
IsoBench: Benchmarking Multimodal Foundation Models on Isomorphic Representations
by: Fu, Deqing, et al.
Published: (2024)
by: Fu, Deqing, et al.
Published: (2024)
Chronos: Learning Temporal Dynamics of Reasoning Chains for Test-Time Scaling
by: Zhang, Kai, et al.
Published: (2026)
by: Zhang, Kai, et al.
Published: (2026)
ChronoPlay: A Framework for Modeling Dual Dynamics and Authenticity in Game RAG Benchmarks
by: He, Liyang, et al.
Published: (2025)
by: He, Liyang, et al.
Published: (2025)
Problems with Chinchilla Approach 2: Systematic Biases in IsoFLOP Parabola Fits
by: Czech, Eric, et al.
Published: (2026)
by: Czech, Eric, et al.
Published: (2026)
ChronosLex: Time-aware Incremental Training for Temporal Generalization of Legal Classification Tasks
by: Santosh, T. Y. S. S, et al.
Published: (2024)
by: Santosh, T. Y. S. S, et al.
Published: (2024)
Chronos: Temporal-Aware Conversational Agents with Structured Event Retrieval for Long-Term Memory
by: Sen, Sahil, et al.
Published: (2026)
by: Sen, Sahil, et al.
Published: (2026)
Contextual effects of sentiment deployment in human and machine translation
by: Comstock, Lindy, et al.
Published: (2025)
by: Comstock, Lindy, et al.
Published: (2025)
Sparse-IFT: Sparse Iso-FLOP Transformations for Maximizing Training Efficiency
by: Thangarasa, Vithursan, et al.
Published: (2023)
by: Thangarasa, Vithursan, et al.
Published: (2023)
Towards Generating Automatic Anaphora Annotations
by: Taji, Dima, et al.
Published: (2025)
by: Taji, Dima, et al.
Published: (2025)
LaajMeter: A Framework for LaaJ Evaluation
by: Ackerman, Samuel, et al.
Published: (2025)
by: Ackerman, Samuel, et al.
Published: (2025)
ScienceMeter: Tracking Scientific Knowledge Updates in Language Models
by: Wang, Yike, et al.
Published: (2025)
by: Wang, Yike, et al.
Published: (2025)
ChronoMagic-Bench: A Benchmark for Metamorphic Evaluation of Text-to-Time-lapse Video Generation
by: Yuan, Shenghai, et al.
Published: (2024)
by: Yuan, Shenghai, et al.
Published: (2024)
LexChronos: An Agentic Framework for Structured Event Timeline Extraction in Indian Jurisprudence
by: Tummepalli, Anka Chandrahas, et al.
Published: (2026)
by: Tummepalli, Anka Chandrahas, et al.
Published: (2026)
Automated Evaluation of Meter and Rhyme in Russian Generative and Human-Authored Poetry
by: Koziev, Ilya
Published: (2025)
by: Koziev, Ilya
Published: (2025)
How Much Is One Recurrence Worth? Iso-Depth Scaling Laws for Looped Language Models
by: Schwethelm, Kristian, et al.
Published: (2026)
by: Schwethelm, Kristian, et al.
Published: (2026)
IsoQuant: Hardware-Aligned SO(4) Isoclinic Rotations for LLM KV Cache Compression
by: Ji, Zhongping
Published: (2026)
by: Ji, Zhongping
Published: (2026)
LLM Ensemble for RAG: Role of Context Length in Zero-Shot Question Answering for BioASQ Challenge
by: Galat, Dima, et al.
Published: (2025)
by: Galat, Dima, et al.
Published: (2025)
Full Iso-recursive Types
by: Zhou, Litao, et al.
Published: (2024)
by: Zhou, Litao, et al.
Published: (2024)
ChronoSense: Exploring Temporal Understanding in Large Language Models with Time Intervals of Events
by: Islakoglu, Duygu Sezen, et al.
Published: (2025)
by: Islakoglu, Duygu Sezen, et al.
Published: (2025)
Improving the TENOR of Labeling: Re-evaluating Topic Models for Content Analysis
by: Li, Zongxia, et al.
Published: (2024)
by: Li, Zongxia, et al.
Published: (2024)
MT-Ranker: Reference-free machine translation evaluation by inter-system ranking
by: Moosa, Ibraheem Muhammad, et al.
Published: (2024)
by: Moosa, Ibraheem Muhammad, et al.
Published: (2024)
Faithfulness metric fusion: Improving the evaluation of LLM trustworthiness across domains
by: Malin, Ben, et al.
Published: (2025)
by: Malin, Ben, et al.
Published: (2025)
Understanding the effects of word-level linguistic annotations in under-resourced neural machine translation
by: Sánchez-Cartagena, Víctor M., et al.
Published: (2024)
by: Sánchez-Cartagena, Víctor M., et al.
Published: (2024)
Optimizing example selection for retrieval-augmented machine translation with translation memories
by: Bouthors, Maxime, et al.
Published: (2024)
by: Bouthors, Maxime, et al.
Published: (2024)
An evaluation of LLMs and Google Translate for translation of selected Indian languages via sentiment and semantic analyses
by: Chandra, Rohitash, et al.
Published: (2025)
by: Chandra, Rohitash, et al.
Published: (2025)
Demystifying deep search: a holistic evaluation with hint-free multi-hop questions and factorised metrics
by: Song, Maojia, et al.
Published: (2025)
by: Song, Maojia, et al.
Published: (2025)
Fabricator or dynamic translator?
by: Vasileva, Lisa, et al.
Published: (2026)
by: Vasileva, Lisa, et al.
Published: (2026)
Statistical multi-metric evaluation and visualization of LLM system predictive performance
by: Ackerman, Samuel, et al.
Published: (2025)
by: Ackerman, Samuel, et al.
Published: (2025)
Performance of diverse evaluation metrics in NLP-based assessment and text generation of consumer complaints
by: Gao, Peiheng, et al.
Published: (2025)
by: Gao, Peiheng, et al.
Published: (2025)
The illusion of a perfect metric: Why evaluating AI's words is harder than it looks
by: Oliva, Maria Paz, et al.
Published: (2025)
by: Oliva, Maria Paz, et al.
Published: (2025)
From prompting to evidence-based translation: A RAG+prompt system for Japanese-Chinese translation and its pedagogical potential
by: Gu, Wenshi
Published: (2026)
by: Gu, Wenshi
Published: (2026)
Similar Items
-
StateAct: Enhancing LLM Base Agents via Self-prompting and State-tracking
by: Rozanov, Nikolai, et al.
Published: (2024) -
Fine-tuning with RAG for Improving LLM Learning of New Skills
by: Ibrahim, Humaid, et al.
Published: (2025) -
SUS backprop: linear backpropagation algorithm for long inputs in transformers
by: Pankov, Sergey, et al.
Published: (2025) -
Same evaluation, more tokens: On the effect of input length for machine translation evaluation using Large Language Models
by: Domhan, Tobias, et al.
Published: (2025) -
Automated evaluation of LLMs for effective machine translation of Mandarin Chinese to English
by: Zhang, Yue, et al.
Published: (2026)