LINKAGE: Listwise Ranking among Varied-Quality References for Non-Factoid QA Evaluation via LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Sihui, Bi, Keping, Cui, Wanqing, Guo, Jiafeng, Cheng, Xueqi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MORE: Multi-mOdal REtrieval Augmented Generative Commonsense Reasoning
por: Cui, Wanqing, et al.
Publicado: (2024)
por: Cui, Wanqing, et al.
Publicado: (2024)
Estimating Commonsense Plausibility through Semantic Shifts
por: Cui, Wanqing, et al.
Publicado: (2025)
por: Cui, Wanqing, et al.
Publicado: (2025)
When Do LLMs Need Retrieval Augmentation? Mitigating LLMs' Overconfidence Helps Retrieval Augmentation
por: Ni, Shiyu, et al.
Publicado: (2024)
por: Ni, Shiyu, et al.
Publicado: (2024)
An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs
por: Zhang, Hengran, et al.
Publicado: (2024)
por: Zhang, Hengran, et al.
Publicado: (2024)
MinosEval: Distinguishing Factoid and Non-Factoid for Tailored Open-Ended QA Evaluation with LLMs
por: Fan, Yongqi, et al.
Publicado: (2025)
por: Fan, Yongqi, et al.
Publicado: (2025)
How Knowledge Popularity Influences and Enhances LLM Knowledge Boundary Perception
por: Ni, Shiyu, et al.
Publicado: (2025)
por: Ni, Shiyu, et al.
Publicado: (2025)
MVAM: Multi-View Attention Method for Fine-grained Image-Text Matching
por: Cui, Wanqing, et al.
Publicado: (2024)
por: Cui, Wanqing, et al.
Publicado: (2024)
Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
por: Wen, Yuchen, et al.
Publicado: (2024)
por: Wen, Yuchen, et al.
Publicado: (2024)
Towards Fully Exploiting LLM Internal States to Enhance Knowledge Boundary Perception
por: Ni, Shiyu, et al.
Publicado: (2025)
por: Ni, Shiyu, et al.
Publicado: (2025)
A Comparative Study of Specialized LLMs as Dense Retrievers
por: Zhang, Hengran, et al.
Publicado: (2025)
por: Zhang, Hengran, et al.
Publicado: (2025)
How Do LLM-Generated Texts Impact Term-Based Retrieval Models?
por: Huang, Wei, et al.
Publicado: (2025)
por: Huang, Wei, et al.
Publicado: (2025)
Annotation-Efficient Universal Honesty Alignment
por: Ni, Shiyu, et al.
Publicado: (2025)
por: Ni, Shiyu, et al.
Publicado: (2025)
Contextual Dual Learning Algorithm with Listwise Distillation for Unbiased Learning to Rank
por: Yu, Lulu, et al.
Publicado: (2024)
por: Yu, Lulu, et al.
Publicado: (2024)
Unleashing the Power of LLMs in Dense Retrieval with Query Likelihood Modeling
por: Zhang, Hengran, et al.
Publicado: (2025)
por: Zhang, Hengran, et al.
Publicado: (2025)
Bagging-Based Model Merging for Robust General Text Embeddings
por: Zhang, Hengran, et al.
Publicado: (2026)
por: Zhang, Hengran, et al.
Publicado: (2026)
Are Large Language Models More Honest in Their Probabilistic or Verbalized Confidence?
por: Ni, Shiyu, et al.
Publicado: (2024)
por: Ni, Shiyu, et al.
Publicado: (2024)
LLM-Specific Utility: A New Perspective for Retrieval-Augmented Generation
por: Zhang, Hengran, et al.
Publicado: (2025)
por: Zhang, Hengran, et al.
Publicado: (2025)
Injecting External Knowledge into the Reasoning Process Enhances Retrieval-Augmented Generation
por: Tang, Minghao, et al.
Publicado: (2025)
por: Tang, Minghao, et al.
Publicado: (2025)
Distilling a Small Utility-Based Passage Selector to Enhance Retrieval-Augmented Generation
por: Zhang, Hengran, et al.
Publicado: (2025)
por: Zhang, Hengran, et al.
Publicado: (2025)
CLIPure: Purification in Latent Space via CLIP for Adversarially Robust Zero-Shot Classification
por: Zhang, Mingkun, et al.
Publicado: (2025)
por: Zhang, Mingkun, et al.
Publicado: (2025)
Utility-Focused LLM Annotation for Retrieval and Retrieval-Augmented Generation
por: Zhang, Hengran, et al.
Publicado: (2025)
por: Zhang, Hengran, et al.
Publicado: (2025)
Beyond Relevance: Utility-Centric Retrieval in the LLM Era
por: Zhang, Hengran, et al.
Publicado: (2026)
por: Zhang, Hengran, et al.
Publicado: (2026)
Attention Grounded Enhancement for Visual Document Retrieval
por: Cui, Wanqing, et al.
Publicado: (2025)
por: Cui, Wanqing, et al.
Publicado: (2025)
Permutative Preference Alignment from Listwise Ranking of Human Judgments
por: Zhao, Yang, et al.
Publicado: (2024)
por: Zhao, Yang, et al.
Publicado: (2024)
CausalDiff: Causality-Inspired Disentanglement via Diffusion Model for Adversarial Defense
por: Zhang, Mingkun, et al.
Publicado: (2024)
por: Zhang, Mingkun, et al.
Publicado: (2024)
How Long Reasoning Chains Influence LLMs' Judgment of Answer Factuality
por: Tu, Minzhu, et al.
Publicado: (2026)
por: Tu, Minzhu, et al.
Publicado: (2026)
Continual Memorization of Factoids in Language Models
por: Chen, Howard, et al.
Publicado: (2024)
por: Chen, Howard, et al.
Publicado: (2024)
Rank-K: Test-Time Reasoning for Listwise Reranking
por: Yang, Eugene, et al.
Publicado: (2025)
por: Yang, Eugene, et al.
Publicado: (2025)
MI-PRUN: Optimize Large Language Model Pruning via Mutual Information
por: Zhang, Hao, et al.
Publicado: (2026)
por: Zhang, Hao, et al.
Publicado: (2026)
Bridging Queries and Tables through Entities in Table Retrieval
por: Li, Da, et al.
Publicado: (2025)
por: Li, Da, et al.
Publicado: (2025)
Reproducibility Analysis and Enhancements for Multi-Aspect Dense Retriever with Aspect Learning
por: Bi, Keping, et al.
Publicado: (2024)
por: Bi, Keping, et al.
Publicado: (2024)
Tailoring Table Retrieval from a Field-aware Hybrid Matching Perspective
por: Li, Da, et al.
Publicado: (2025)
por: Li, Da, et al.
Publicado: (2025)
CIR at the NTCIR-17 ULTRE-2 Task
por: Yu, Lulu, et al.
Publicado: (2023)
por: Yu, Lulu, et al.
Publicado: (2023)
Iterative Structured Pruning for Large Language Models with Multi-Domain Calibration
por: Wu, Guangxin, et al.
Publicado: (2026)
por: Wu, Guangxin, et al.
Publicado: (2026)
Controlling Risk of Retrieval-augmented Generation: A Counterfactual Prompting Framework
por: Chen, Lu, et al.
Publicado: (2024)
por: Chen, Lu, et al.
Publicado: (2024)
Class-Incremental Few-Shot Event Detection
por: Zhao, Kailin, et al.
Publicado: (2024)
por: Zhao, Kailin, et al.
Publicado: (2024)
Multilingual Non-Factoid Question Answering with Answer Paragraph Selection
por: Mishra, Ritwik, et al.
Publicado: (2024)
por: Mishra, Ritwik, et al.
Publicado: (2024)
Unbiased Learning to Rank with Query-Level Click Propensity Estimation: Beyond Pointwise Observation and Relevance
por: Yu, Lulu, et al.
Publicado: (2025)
por: Yu, Lulu, et al.
Publicado: (2025)
Towards Robust Universal Information Extraction: Benchmark, Evaluation, and Solution
por: Zhu, Jizhao, et al.
Publicado: (2025)
por: Zhu, Jizhao, et al.
Publicado: (2025)
Typed-RAG: Type-Aware Decomposition of Non-Factoid Questions for Retrieval-Augmented Generation
por: Lee, DongGeon, et al.
Publicado: (2025)
por: Lee, DongGeon, et al.
Publicado: (2025)
Ejemplares similares
-
MORE: Multi-mOdal REtrieval Augmented Generative Commonsense Reasoning
por: Cui, Wanqing, et al.
Publicado: (2024) -
Estimating Commonsense Plausibility through Semantic Shifts
por: Cui, Wanqing, et al.
Publicado: (2025) -
When Do LLMs Need Retrieval Augmentation? Mitigating LLMs' Overconfidence Helps Retrieval Augmentation
por: Ni, Shiyu, et al.
Publicado: (2024) -
An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs
por: Zhang, Hengran, et al.
Publicado: (2024) -
MinosEval: Distinguishing Factoid and Non-Factoid for Tailored Open-Ended QA Evaluation with LLMs
por: Fan, Yongqi, et al.
Publicado: (2025)