LLM Optimization Unlocks Real-Time Pairwise Reranking
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Jingyu, Shrivastava, Aditya, Zhu, Jing, Samuel, Alfy, Kumar, Anoop, Liu, Daben |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FB-RAG: Improving RAG with Forward and Backward Lookup
von: Chawla, Kushal, et al.
Veröffentlicht: (2025)
von: Chawla, Kushal, et al.
Veröffentlicht: (2025)
A Comparison of Independent and Joint Fine-tuning Strategies for Retrieval-Augmented Generation
von: Lawton, Neal Gregory, et al.
Veröffentlicht: (2025)
von: Lawton, Neal Gregory, et al.
Veröffentlicht: (2025)
Play by the Type Rules: Inferring Constraints for LLM Functions in Declarative Programs
von: Glenn, Parker, et al.
Veröffentlicht: (2025)
von: Glenn, Parker, et al.
Veröffentlicht: (2025)
Confidence-Based Response Abstinence: Improving LLM Trustworthiness via Activation-Based Uncertainty Estimation
von: Huang, Zhiqi, et al.
Veröffentlicht: (2025)
von: Huang, Zhiqi, et al.
Veröffentlicht: (2025)
Refusal Tokens: A Simple Way to Calibrate Refusals in Large Language Models
von: Jain, Neel, et al.
Veröffentlicht: (2024)
von: Jain, Neel, et al.
Veröffentlicht: (2024)
Readability Reconsidered: A Cross-Dataset Analysis of Reference-Free Metrics
von: Belem, Catarina G, et al.
Veröffentlicht: (2025)
von: Belem, Catarina G, et al.
Veröffentlicht: (2025)
Improving Consistency in Retrieval-Augmented Systems with Group Similarity Rewards
von: Hamman, Faisal, et al.
Veröffentlicht: (2025)
von: Hamman, Faisal, et al.
Veröffentlicht: (2025)
Harmonizing Diverse Models: A Layer-wise Merging Strategy for Consistent Generation
von: Peng, Xujun, et al.
Veröffentlicht: (2025)
von: Peng, Xujun, et al.
Veröffentlicht: (2025)
Alignment-Weighted DPO: A principled reasoning approach to improve safety alignment
von: Hu, Mengxuan, et al.
Veröffentlicht: (2026)
von: Hu, Mengxuan, et al.
Veröffentlicht: (2026)
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2025)
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2025)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025)
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025)
RAFFLES: Reasoning-based Attribution of Faults for LLM Systems
von: Zhu, Chenyang, et al.
Veröffentlicht: (2025)
von: Zhu, Chenyang, et al.
Veröffentlicht: (2025)
DF-RAG: Query-Aware Diversity for Retrieval-Augmented Generation
von: Khan, Saadat Hasan, et al.
Veröffentlicht: (2026)
von: Khan, Saadat Hasan, et al.
Veröffentlicht: (2026)
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning
von: Wu, Yuhang, et al.
Veröffentlicht: (2026)
von: Wu, Yuhang, et al.
Veröffentlicht: (2026)
Gumbel Reranking: Differentiable End-to-End Reranker Optimization
von: Huang, Siyuan, et al.
Veröffentlicht: (2025)
von: Huang, Siyuan, et al.
Veröffentlicht: (2025)
Efficiency-Effectiveness Reranking FLOPs for LLM-based Rerankers
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2025)
How Good are LLM-based Rerankers? An Empirical Analysis of State-of-the-Art Reranking Models
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
DGPO: Beyond Pairwise Preferences with Directional Consistent Groupwise Optimization
von: Deng, Mengyi, et al.
Veröffentlicht: (2026)
von: Deng, Mengyi, et al.
Veröffentlicht: (2026)
Knowledge Restoration-driven Prompt Optimization: Unlocking LLM Potential for Open-Domain Relational Triplet Extraction
von: Jing, Xiaonan, et al.
Veröffentlicht: (2026)
von: Jing, Xiaonan, et al.
Veröffentlicht: (2026)
A Bayesian Optimization Approach to Machine Translation Reranking
von: Cheng, Julius, et al.
Veröffentlicht: (2024)
von: Cheng, Julius, et al.
Veröffentlicht: (2024)
Investigating Test-Time Scaling with Reranking for Machine Translation
von: Tan, Shaomu, et al.
Veröffentlicht: (2025)
von: Tan, Shaomu, et al.
Veröffentlicht: (2025)
EPSVec: Efficient and Private Synthetic Data Generation via Dataset Vectors
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
ParoQuant: Pairwise Rotation Quantization for Efficient Reasoning LLM Inference
von: Liang, Yesheng, et al.
Veröffentlicht: (2025)
von: Liang, Yesheng, et al.
Veröffentlicht: (2025)
PrefPO: Pairwise Preference Prompt Optimization
von: Singhal, Rahul, et al.
Veröffentlicht: (2026)
von: Singhal, Rahul, et al.
Veröffentlicht: (2026)
SOUL: Unlocking the Power of Second-Order Optimization for LLM Unlearning
von: Jia, Jinghan, et al.
Veröffentlicht: (2024)
von: Jia, Jinghan, et al.
Veröffentlicht: (2024)
MemReranker: Reasoning-Aware Reranking for Agent Memory Retrieval
von: Li, Chunyu, et al.
Veröffentlicht: (2026)
von: Li, Chunyu, et al.
Veröffentlicht: (2026)
HAMburger: Accelerating LLM Inference via Token Smashing
von: Liu, Jingyu, et al.
Veröffentlicht: (2025)
von: Liu, Jingyu, et al.
Veröffentlicht: (2025)
MemRerank: Preference Memory for Personalized Product Reranking
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2026)
Reranker Optimization via Geodesic Distances on k-NN Manifolds
von: Gong, Wen G.
Veröffentlicht: (2026)
von: Gong, Wen G.
Veröffentlicht: (2026)
Cross-Genre Authorship Attribution via LLM-Based Retrieve-and-Rerank
von: Agarwal, Shantanu, et al.
Veröffentlicht: (2025)
von: Agarwal, Shantanu, et al.
Veröffentlicht: (2025)
Toolken+: Improving LLM Tool Usage with Reranking and a Reject Option
von: Yakovlev, Konstantin, et al.
Veröffentlicht: (2024)
von: Yakovlev, Konstantin, et al.
Veröffentlicht: (2024)
Embedding-Based Context-Aware Reranker
von: Yuan, Ye, et al.
Veröffentlicht: (2025)
von: Yuan, Ye, et al.
Veröffentlicht: (2025)
Typologically-Informed Candidate Reranking for LLM-based Translation into Low-Resource Languages
von: Abeykoon, Nipuna, et al.
Veröffentlicht: (2026)
von: Abeykoon, Nipuna, et al.
Veröffentlicht: (2026)
SciRerankBench: Benchmarking Rerankers Towards Scientific Retrieval-Augmented Generated LLMs
von: Chen, Haotian, et al.
Veröffentlicht: (2025)
von: Chen, Haotian, et al.
Veröffentlicht: (2025)
Semantic Reranking at Inference Time for Hard Examples in Rhetorical Role Labeling
von: Belfathi, Anas, et al.
Veröffentlicht: (2026)
von: Belfathi, Anas, et al.
Veröffentlicht: (2026)
Reranking Passages with Coarse-to-Fine Neural Retriever Enhanced by List-Context Information
von: Zhu, Hongyin
Veröffentlicht: (2023)
von: Zhu, Hongyin
Veröffentlicht: (2023)
Rank-K: Test-Time Reasoning for Listwise Reranking
von: Yang, Eugene, et al.
Veröffentlicht: (2025)
von: Yang, Eugene, et al.
Veröffentlicht: (2025)
Unlocking Efficient Long-to-Short LLM Reasoning with Model Merging
von: Wu, Han, et al.
Veröffentlicht: (2025)
von: Wu, Han, et al.
Veröffentlicht: (2025)
Bridging and Modeling Correlations in Pairwise Data for Direct Preference Optimization
von: Jiang, Yuxin, et al.
Veröffentlicht: (2024)
von: Jiang, Yuxin, et al.
Veröffentlicht: (2024)
The Comparative Trap: Pairwise Comparisons Amplifies Biased Preferences of LLM Evaluators
von: Jeong, Hawon, et al.
Veröffentlicht: (2024)
von: Jeong, Hawon, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FB-RAG: Improving RAG with Forward and Backward Lookup
von: Chawla, Kushal, et al.
Veröffentlicht: (2025) -
A Comparison of Independent and Joint Fine-tuning Strategies for Retrieval-Augmented Generation
von: Lawton, Neal Gregory, et al.
Veröffentlicht: (2025) -
Play by the Type Rules: Inferring Constraints for LLM Functions in Declarative Programs
von: Glenn, Parker, et al.
Veröffentlicht: (2025) -
Confidence-Based Response Abstinence: Improving LLM Trustworthiness via Activation-Based Uncertainty Estimation
von: Huang, Zhiqi, et al.
Veröffentlicht: (2025) -
Refusal Tokens: A Simple Way to Calibrate Refusals in Large Language Models
von: Jain, Neel, et al.
Veröffentlicht: (2024)