CATER: Leveraging LLM to Pioneer a Multidimensional, Reference-Independent Paradigm in Translation Quality Evaluation
Fuente:
arXiv
Salvato in:
| Autori principali: | IIDA, Kurando, MIMURA, Kenjiro |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Understanding the Effects of RLHF on the Quality and Detectability of LLM-Generated Texts
di: Xu, Beining, et al.
Pubblicazione: (2025)
di: Xu, Beining, et al.
Pubblicazione: (2025)
Triad: A Framework Leveraging a Multi-Role LLM-based Agent to Solve Knowledge Base Question Answering
di: Zong, Chang, et al.
Pubblicazione: (2024)
di: Zong, Chang, et al.
Pubblicazione: (2024)
The Paradox of Poetic Intent in Back-Translation: Evaluating the Quality of Large Language Models in Chinese Translation
di: Weigang, Li, et al.
Pubblicazione: (2025)
di: Weigang, Li, et al.
Pubblicazione: (2025)
Align-to-Distill: Trainable Attention Alignment for Knowledge Distillation in Neural Machine Translation
di: Jin, Heegon, et al.
Pubblicazione: (2024)
di: Jin, Heegon, et al.
Pubblicazione: (2024)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
di: Orgad, Hadas, et al.
Pubblicazione: (2024)
di: Orgad, Hadas, et al.
Pubblicazione: (2024)
Isolating LLM Lexical Bias: A Curation-Free Triangulated Metric for Preference-Stage Learning
di: Ming, Xiaoyang, et al.
Pubblicazione: (2026)
di: Ming, Xiaoyang, et al.
Pubblicazione: (2026)
HR-MultiWOZ: A Task Oriented Dialogue (TOD) Dataset for HR LLM Agent
di: Xu, Weijie, et al.
Pubblicazione: (2024)
di: Xu, Weijie, et al.
Pubblicazione: (2024)
Textual Data Bias Detection and Mitigation -- An Extensible Pipeline with Experimental Evaluation
di: Görge, Rebekka, et al.
Pubblicazione: (2025)
di: Görge, Rebekka, et al.
Pubblicazione: (2025)
WaterSeeker: Pioneering Efficient Detection of Watermarked Segments in Large Documents
di: Pan, Leyi, et al.
Pubblicazione: (2024)
di: Pan, Leyi, et al.
Pubblicazione: (2024)
Predictive Simultaneous Interpretation: Harnessing Large Language Models for Democratizing Real-Time Multilingual Communication
di: Iida, Kurando, et al.
Pubblicazione: (2024)
di: Iida, Kurando, et al.
Pubblicazione: (2024)
Advancing Explainability in Neural Machine Translation: Analytical Metrics for Attention and Alignment Consistency
di: Mishra, Anurag
Pubblicazione: (2024)
di: Mishra, Anurag
Pubblicazione: (2024)
MetaCheckGPT -- A Multi-task Hallucination Detector Using LLM Uncertainty and Meta-models
di: Mehta, Rahul, et al.
Pubblicazione: (2024)
di: Mehta, Rahul, et al.
Pubblicazione: (2024)
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
di: Miliani, Martina, et al.
Pubblicazione: (2025)
di: Miliani, Martina, et al.
Pubblicazione: (2025)
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
di: Haque, Md. Asraful, et al.
Pubblicazione: (2026)
di: Haque, Md. Asraful, et al.
Pubblicazione: (2026)
DYNAMICQA: Tracing Internal Knowledge Conflicts in Language Models
di: Marjanović, Sara Vera, et al.
Pubblicazione: (2024)
di: Marjanović, Sara Vera, et al.
Pubblicazione: (2024)
Cognitive Workspace: Active Memory Management for LLMs -- An Empirical Study of Functional Infinite Context
di: An, Tao
Pubblicazione: (2025)
di: An, Tao
Pubblicazione: (2025)
One Agent to Serve All: a Lite-Adaptive Stylized AI Assistant for Millions of Multi-Style Official Accounts
di: Fan, Xingyu, et al.
Pubblicazione: (2025)
di: Fan, Xingyu, et al.
Pubblicazione: (2025)
The Personalization Trap: How User Memory Alters Emotional Reasoning in LLMs
di: Fang, Xi, et al.
Pubblicazione: (2025)
di: Fang, Xi, et al.
Pubblicazione: (2025)
TREX: Tokenizer Regression for Optimal Data Mixture
di: Won, Inho, et al.
Pubblicazione: (2026)
di: Won, Inho, et al.
Pubblicazione: (2026)
Language corpora for the Dutch medical domain
di: van Es, B.
Pubblicazione: (2026)
di: van Es, B.
Pubblicazione: (2026)
ADE: Adaptive Dictionary Embeddings -- Scaling Multi-Anchor Representations to Large Language Models
di: Demirci, Orhan, et al.
Pubblicazione: (2026)
di: Demirci, Orhan, et al.
Pubblicazione: (2026)
MemeLens: Multilingual Multitask VLMs for Memes
di: Shahroor, Ali Ezzat, et al.
Pubblicazione: (2026)
di: Shahroor, Ali Ezzat, et al.
Pubblicazione: (2026)
Rapid Biomedical Research Classification: The Pandemic PACT Advanced Categorisation Engine
di: Rohanian, Omid, et al.
Pubblicazione: (2024)
di: Rohanian, Omid, et al.
Pubblicazione: (2024)
Give it Space! Explicit Disentangling of Positional and Semantic Representations in Encoders
di: Lequeu, Pierre-Antoine, et al.
Pubblicazione: (2026)
di: Lequeu, Pierre-Antoine, et al.
Pubblicazione: (2026)
Dynamic Demonstration Retrieval and Cognitive Understanding for Emotional Support Conversation
di: Xu, Zhe, et al.
Pubblicazione: (2024)
di: Xu, Zhe, et al.
Pubblicazione: (2024)
Tokenization Is More Than Compression
di: Schmidt, Craig W., et al.
Pubblicazione: (2024)
di: Schmidt, Craig W., et al.
Pubblicazione: (2024)
A Graph-based Approach for Multi-Modal Question Answering from Flowcharts in Telecom Documents
di: Soman, Sumit, et al.
Pubblicazione: (2025)
di: Soman, Sumit, et al.
Pubblicazione: (2025)
Revealing the Parametric Knowledge of Language Models: A Unified Framework for Attribution Methods
di: Yu, Haeun, et al.
Pubblicazione: (2024)
di: Yu, Haeun, et al.
Pubblicazione: (2024)
SpokenNativQA: Multilingual Everyday Spoken Queries for LLMs
di: Alam, Firoj, et al.
Pubblicazione: (2025)
di: Alam, Firoj, et al.
Pubblicazione: (2025)
Boundless Byte Pair Encoding: Breaking the Pre-tokenization Barrier
di: Schmidt, Craig W., et al.
Pubblicazione: (2025)
di: Schmidt, Craig W., et al.
Pubblicazione: (2025)
Obfuscation Rules for Detecting and Detoxifying Korean Toxicity
di: Lee, Yejin, et al.
Pubblicazione: (2025)
di: Lee, Yejin, et al.
Pubblicazione: (2025)
LayerTracer: A Joint Task-Particle and Vulnerable-Layer Analysis framework for Arbitrary Large Language Model Architectures
di: Wu, Yuhang, et al.
Pubblicazione: (2026)
di: Wu, Yuhang, et al.
Pubblicazione: (2026)
A Comprehensive Survey of Compression Algorithms for Language Models
di: Park, Seungcheol, et al.
Pubblicazione: (2024)
di: Park, Seungcheol, et al.
Pubblicazione: (2024)
Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts
di: Feucht, Sheridan, et al.
Pubblicazione: (2026)
di: Feucht, Sheridan, et al.
Pubblicazione: (2026)
ProSwitch: Knowledge-Guided Instruction Tuning to Switch Between Professional and Non-Professional Responses
di: Zong, Chang, et al.
Pubblicazione: (2024)
di: Zong, Chang, et al.
Pubblicazione: (2024)
RV-HATE: Reinforced Multi-Module Voting for Implicit Hate Speech Detection
di: Lee, Yejin, et al.
Pubblicazione: (2025)
di: Lee, Yejin, et al.
Pubblicazione: (2025)
Learning When to Think: Shaping Adaptive Reasoning in R1-Style Models via Multi-Stage RL
di: Tu, Songjun, et al.
Pubblicazione: (2025)
di: Tu, Songjun, et al.
Pubblicazione: (2025)
Task Complexity Matters: An Empirical Study of Reasoning in LLMs for Sentiment Analysis
di: Huang, Donghao, et al.
Pubblicazione: (2026)
di: Huang, Donghao, et al.
Pubblicazione: (2026)
Efficient Adaptive Rejection Sampling for Accelerating Speculative Decoding in Large Language Models
di: Sun, Chendong, et al.
Pubblicazione: (2025)
di: Sun, Chendong, et al.
Pubblicazione: (2025)
Topeax -- An Improved Clustering Topic Model with Density Peak Detection and Lexical-Semantic Term Importance
di: Kardos, Márton
Pubblicazione: (2026)
di: Kardos, Márton
Pubblicazione: (2026)
Documenti analoghi
-
Understanding the Effects of RLHF on the Quality and Detectability of LLM-Generated Texts
di: Xu, Beining, et al.
Pubblicazione: (2025) -
Triad: A Framework Leveraging a Multi-Role LLM-based Agent to Solve Knowledge Base Question Answering
di: Zong, Chang, et al.
Pubblicazione: (2024) -
The Paradox of Poetic Intent in Back-Translation: Evaluating the Quality of Large Language Models in Chinese Translation
di: Weigang, Li, et al.
Pubblicazione: (2025) -
Align-to-Distill: Trainable Attention Alignment for Knowledge Distillation in Neural Machine Translation
di: Jin, Heegon, et al.
Pubblicazione: (2024) -
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
di: Orgad, Hadas, et al.
Pubblicazione: (2024)