From Model-centered to Human-Centered: Revision Distance as a Metric for Text Evaluation in LLMs-based Applications
Fuente:
arXiv
Guardado en:
| Autores principales: | Ma, Yongqiang, Qing, Lizhi, Liu, Jiawei, Kang, Yangyang, Zhang, Yue, Lu, Wei, Liu, Xiaozhong, Cheng, Qikai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FlippedRAG: Black-Box Opinion Manipulation Adversarial Attacks to Retrieval-Augmented Generation Models
por: Chen, Zhuo, et al.
Publicado: (2025)
por: Chen, Zhuo, et al.
Publicado: (2025)
ListConRanker: A Contrastive Text Reranker with Listwise Encoding
por: Liu, Junlong, et al.
Publicado: (2025)
por: Liu, Junlong, et al.
Publicado: (2025)
On the Evaluation Metric for Hashing
por: Jiang, Qing-Yuan, et al.
Publicado: (2019)
por: Jiang, Qing-Yuan, et al.
Publicado: (2019)
MA-DPR: Manifold-aware Distance Metrics for Dense Passage Retrieval
por: Liu, Yifan, et al.
Publicado: (2025)
por: Liu, Yifan, et al.
Publicado: (2025)
The 2nd Workshop on Human-Centered Recommender Systems
por: Zhang, Kaike, et al.
Publicado: (2025)
por: Zhang, Kaike, et al.
Publicado: (2025)
Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models
por: Gong, Yuyang, et al.
Publicado: (2025)
por: Gong, Yuyang, et al.
Publicado: (2025)
Exploring the Potential of LLMs for Serendipity Evaluation in Recommender Systems
por: Kang, Li, et al.
Publicado: (2025)
por: Kang, Li, et al.
Publicado: (2025)
A Comparative Analysis of Faithfulness Metrics and Humans in Citation Evaluation
por: Zhang, Weijia, et al.
Publicado: (2024)
por: Zhang, Weijia, et al.
Publicado: (2024)
The 1st Workshop on Human-Centered Recommender Systems
por: Zhang, Kaike, et al.
Publicado: (2024)
por: Zhang, Kaike, et al.
Publicado: (2024)
T2S-Metrics: Unified Library for Evaluating SPARQL Queries Generated From Natural Language
por: Taghzouti, Yousouf, et al.
Publicado: (2026)
por: Taghzouti, Yousouf, et al.
Publicado: (2026)
DiscourseFlip: An Oblique Discourse-Level Opinion Manipulation Attack against Black-box Retrieval-Augmented Generation
por: Gong, Yuyang, et al.
Publicado: (2026)
por: Gong, Yuyang, et al.
Publicado: (2026)
Pointwise Metrics for Clustering Evaluation
por: van Staden, Stephan
Publicado: (2024)
por: van Staden, Stephan
Publicado: (2024)
Can LLMs Outshine Conventional Recommenders? A Comparative Evaluation
por: Liu, Qijiong, et al.
Publicado: (2025)
por: Liu, Qijiong, et al.
Publicado: (2025)
From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs
por: Wu, Yaxiong, et al.
Publicado: (2025)
por: Wu, Yaxiong, et al.
Publicado: (2025)
LLM-Driven Data Generation and a Novel Soft Metric for Evaluating Text-to-SQL in Aviation MRO
por: Sutanto, Patrick, et al.
Publicado: (2025)
por: Sutanto, Patrick, et al.
Publicado: (2025)
From Isolated Scoring to Collaborative Ranking: A Comparison-Native Framework for LLM-Based Paper Evaluation
por: Zheng, Pujun, et al.
Publicado: (2026)
por: Zheng, Pujun, et al.
Publicado: (2026)
A Sketch+Text Composed Image Retrieval Dataset for Thangka
por: Xu, Jinyu, et al.
Publicado: (2026)
por: Xu, Jinyu, et al.
Publicado: (2026)
Text2Cypher Across Languages: Evaluating and Finetuning LLMs
por: Ozsoy, Makbule Gulcin, et al.
Publicado: (2025)
por: Ozsoy, Makbule Gulcin, et al.
Publicado: (2025)
Sustainability Evaluation Metrics for Recommender Systems
por: Felfernig, Alexander, et al.
Publicado: (2025)
por: Felfernig, Alexander, et al.
Publicado: (2025)
UQABench: Evaluating User Embedding for Prompting LLMs in Personalized Question Answering
por: Liu, Langming, et al.
Publicado: (2025)
por: Liu, Langming, et al.
Publicado: (2025)
Make Any Collection Navigable: Methods for Constructing and Evaluating Hypergraph of Text
por: Alvarez, Dean E., et al.
Publicado: (2026)
por: Alvarez, Dean E., et al.
Publicado: (2026)
Enhance Robustness of Language Models Against Variation Attack through Graph Integration
por: Xiong, Zi, et al.
Publicado: (2024)
por: Xiong, Zi, et al.
Publicado: (2024)
Towards Fine-Grained Citation Evaluation in Generated Text: A Comparative Analysis of Faithfulness Metrics
por: Zhang, Weijia, et al.
Publicado: (2024)
por: Zhang, Weijia, et al.
Publicado: (2024)
Leveraging LLMs for Influence Path Planning in Proactive Recommendation
por: Wang, Mingze, et al.
Publicado: (2024)
por: Wang, Mingze, et al.
Publicado: (2024)
Survey of Query-based Text Summarization
por: Yu, Hang, et al.
Publicado: (2022)
por: Yu, Hang, et al.
Publicado: (2022)
Facet-Aware Multi-Head Mixture-of-Experts Model with Text-Enhanced Pre-training for Sequential Recommendation
por: Liu, Mingrui, et al.
Publicado: (2026)
por: Liu, Mingrui, et al.
Publicado: (2026)
You Only Evaluate Once: A Tree-based Rerank Method at Meituan
por: Wang, Shuli, et al.
Publicado: (2025)
por: Wang, Shuli, et al.
Publicado: (2025)
Intermediate Distillation: Data-Efficient Distillation from Black-Box LLMs for Information Retrieval
por: Li, Zizhong, et al.
Publicado: (2024)
por: Li, Zizhong, et al.
Publicado: (2024)
Efficient Search in Graph Edit Distance: Metric Search Trees vs. Brute Force Verification
por: Guo, Wenqi Marshall, et al.
Publicado: (2024)
por: Guo, Wenqi Marshall, et al.
Publicado: (2024)
HELM: A Human-Centered Evaluation Framework for LLM-Powered Recommender Systems
por: Mehta, Sushant
Publicado: (2026)
por: Mehta, Sushant
Publicado: (2026)
SSE: A Metric for Evaluating Search System Explainability
por: Chen, Catherine, et al.
Publicado: (2023)
por: Chen, Catherine, et al.
Publicado: (2023)
RLTP: Reinforcement Learning to Pace for Delayed Impression Modeling in Preloaded Ads
por: Wei, Penghui, et al.
Publicado: (2023)
por: Wei, Penghui, et al.
Publicado: (2023)
Training LLMs to be Better Text Embedders through Bidirectional Reconstruction
por: Su, Chang, et al.
Publicado: (2025)
por: Su, Chang, et al.
Publicado: (2025)
Text Clustering as Classification with LLMs
por: Huang, Chen, et al.
Publicado: (2024)
por: Huang, Chen, et al.
Publicado: (2024)
TF-DCon: Leveraging Large Language Models (LLMs) to Empower Training-Free Dataset Condensation for Content-Based Recommendation
por: Wu, Jiahao, et al.
Publicado: (2023)
por: Wu, Jiahao, et al.
Publicado: (2023)
Towards Context-Robust LLMs: A Gated Representation Fine-tuning Approach
por: Zeng, Shenglai, et al.
Publicado: (2025)
por: Zeng, Shenglai, et al.
Publicado: (2025)
A Survey on Deep Text Hashing: Efficient Semantic Text Retrieval with Binary Representation
por: He, Liyang, et al.
Publicado: (2025)
por: He, Liyang, et al.
Publicado: (2025)
DistRAG: Towards Distance-Based Spatial Reasoning in LLMs
por: Schneider, Nicole R, et al.
Publicado: (2025)
por: Schneider, Nicole R, et al.
Publicado: (2025)
Learning Filter-Aware Distance Metrics for Nearest Neighbor Search with Multiple Filters
por: Sutradhar, Ananya, et al.
Publicado: (2025)
por: Sutradhar, Ananya, et al.
Publicado: (2025)
Fréchet Distance for Offline Evaluation of Information Retrieval Systems with Sparse Labels
por: Arabzadeh, Negar, et al.
Publicado: (2024)
por: Arabzadeh, Negar, et al.
Publicado: (2024)
Ejemplares similares
-
FlippedRAG: Black-Box Opinion Manipulation Adversarial Attacks to Retrieval-Augmented Generation Models
por: Chen, Zhuo, et al.
Publicado: (2025) -
ListConRanker: A Contrastive Text Reranker with Listwise Encoding
por: Liu, Junlong, et al.
Publicado: (2025) -
On the Evaluation Metric for Hashing
por: Jiang, Qing-Yuan, et al.
Publicado: (2019) -
MA-DPR: Manifold-aware Distance Metrics for Dense Passage Retrieval
por: Liu, Yifan, et al.
Publicado: (2025) -
The 2nd Workshop on Human-Centered Recommender Systems
por: Zhang, Kaike, et al.
Publicado: (2025)