LLM Braces: Straightening Out LLM Predictions with Relevant Sub-Updates
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Ying, Huang, Lifu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reinforcement Learning-based Knowledge Distillation with LLM-as-a-Judge
by: Shen, Yiyang, et al.
Published: (2026)
by: Shen, Yiyang, et al.
Published: (2026)
Debate as Optimization: Adaptive Conformal Prediction and Diverse Retrieval for Event Extraction
by: Wang, Sijia, et al.
Published: (2024)
by: Wang, Sijia, et al.
Published: (2024)
ThinkBench: Dynamic Out-of-Distribution Evaluation for Robust LLM Reasoning
by: Huang, Shulin, et al.
Published: (2025)
by: Huang, Shulin, et al.
Published: (2025)
Benchmarking LLM-based Relevance Judgment Methods
by: Arabzadeh, Negar, et al.
Published: (2025)
by: Arabzadeh, Negar, et al.
Published: (2025)
WINELL: Wikipedia Never-Ending Updating with LLM Agents
by: Reddy, Revanth Gangi, et al.
Published: (2025)
by: Reddy, Revanth Gangi, et al.
Published: (2025)
Are LLM Belief Updates Consistent with Bayes' Theorem?
by: Imran, Sohaib, et al.
Published: (2025)
by: Imran, Sohaib, et al.
Published: (2025)
MULTISCRIPT: Multimodal Script Learning for Supporting Open Domain Everyday Tasks
by: Qi, Jingyuan, et al.
Published: (2023)
by: Qi, Jingyuan, et al.
Published: (2023)
Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations
by: Acquaye, Christabel, et al.
Published: (2026)
by: Acquaye, Christabel, et al.
Published: (2026)
LLM NL2SQL Robustness: Surface Noise vs. Linguistic Variation in Traditional and Agentic Settings
by: Tu, Lifu, et al.
Published: (2026)
by: Tu, Lifu, et al.
Published: (2026)
Evaluating the Relevance of Uncertainty Estimators for LLM Hallucination
by: Agnimo, Yedidia, et al.
Published: (2026)
by: Agnimo, Yedidia, et al.
Published: (2026)
LLM-as-RNN: A Recurrent Language Model for Memory Updates and Sequence Prediction
by: Lu, Yuxing, et al.
Published: (2026)
by: Lu, Yuxing, et al.
Published: (2026)
Targeted Augmentation for Low-Resource Event Extraction
by: Wang, Sijia, et al.
Published: (2024)
by: Wang, Sijia, et al.
Published: (2024)
SCORE: Specificity, Context Utilization, Robustness, and Relevance for Reference-Free LLM Evaluation
by: Shomee, Homaira Huda, et al.
Published: (2026)
by: Shomee, Homaira Huda, et al.
Published: (2026)
ScreenLLM: Stateful Screen Schema for Efficient Action Understanding and Prediction
by: Jin, Yiqiao, et al.
Published: (2025)
by: Jin, Yiqiao, et al.
Published: (2025)
The Effect of Document Summarization on LLM-Based Relevance Judgments
by: Mohtadi, Samaneh, et al.
Published: (2025)
by: Mohtadi, Samaneh, et al.
Published: (2025)
Human Texts Are Outliers: Detecting LLM-generated Texts via Out-of-distribution Detection
by: Zeng, Cong, et al.
Published: (2025)
by: Zeng, Cong, et al.
Published: (2025)
From Words to Worth: Newborn Article Impact Prediction with LLM
by: Zhao, Penghai, et al.
Published: (2024)
by: Zhao, Penghai, et al.
Published: (2024)
Preference-Aware Memory Update for Long-Term LLM Agents
by: Sun, Haoran, et al.
Published: (2025)
by: Sun, Haoran, et al.
Published: (2025)
PABU: Progress-Aware Belief Update for Efficient LLM Agents
by: Jiang, Haitao, et al.
Published: (2026)
by: Jiang, Haitao, et al.
Published: (2026)
Can LLM Generate Culturally Relevant Commonsense QA Data? Case Study in Indonesian and Sundanese
by: Putri, Rifki Afina, et al.
Published: (2024)
by: Putri, Rifki Afina, et al.
Published: (2024)
Foundation Models for Autonomous Robots in Unstructured Environments
by: Naderi, Hossein, et al.
Published: (2024)
by: Naderi, Hossein, et al.
Published: (2024)
Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction
by: Li, Mingchen, et al.
Published: (2026)
by: Li, Mingchen, et al.
Published: (2026)
Judging with Personality and Confidence: A Study on Personality-Conditioned LLM Relevance Assessment
by: Chen, Nuo, et al.
Published: (2026)
by: Chen, Nuo, et al.
Published: (2026)
Efficient Out-of-Scope Detection in Dialogue Systems via Uncertainty-Driven LLM Routing
by: Zaera, Álvaro, et al.
Published: (2025)
by: Zaera, Álvaro, et al.
Published: (2025)
Merging Beyond: Streaming LLM Updates via Activation-Guided Rotations
by: Yao, Yuxuan, et al.
Published: (2026)
by: Yao, Yuxuan, et al.
Published: (2026)
Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks
by: Cao, Yixin, et al.
Published: (2025)
by: Cao, Yixin, et al.
Published: (2025)
Beyond Relevance: Utility-Centric Retrieval in the LLM Era
by: Zhang, Hengran, et al.
Published: (2026)
by: Zhang, Hengran, et al.
Published: (2026)
Smoothing Out Hallucinations: Mitigating LLM Hallucination with Smoothed Knowledge Distillation
by: Nguyen, Hieu, et al.
Published: (2025)
by: Nguyen, Hieu, et al.
Published: (2025)
HyperJoin: LLM-augmented Hypergraph Link Prediction for Joinable Table Discovery
by: Liu, Shiyuan, et al.
Published: (2026)
by: Liu, Shiyuan, et al.
Published: (2026)
From Feedback Loops to Policy Updates: Reinforcement Fine-Tuning for LLM-Based Alpha Factor Discovery
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
A Tutorial on LLM Reasoning: Relevant Methods behind ChatGPT o1
by: Wang, Jun
Published: (2025)
by: Wang, Jun
Published: (2025)
A Human-AI Comparative Analysis of Prompt Sensitivity in LLM-Based Relevance Judgment
by: Arabzadeh, Negar, et al.
Published: (2025)
by: Arabzadeh, Negar, et al.
Published: (2025)
The LLM Language Network: A Neuroscientific Approach for Identifying Causally Task-Relevant Units
by: AlKhamissi, Badr, et al.
Published: (2024)
by: AlKhamissi, Badr, et al.
Published: (2024)
Improving Topic Relevance Model by Mix-structured Summarization and LLM-based Data Augmentation
by: Liu, Yizhu, et al.
Published: (2024)
by: Liu, Yizhu, et al.
Published: (2024)
The Hidden Puppet Master: Predicting Human Belief Change in Manipulative LLM Dialogues
by: Shen, Jocelyn, et al.
Published: (2026)
by: Shen, Jocelyn, et al.
Published: (2026)
Query-Document Dense Vectors for LLM Relevance Judgment Bias Analysis
by: Mohtadi, Samaneh, et al.
Published: (2026)
by: Mohtadi, Samaneh, et al.
Published: (2026)
DIRAS: Efficient LLM Annotation of Document Relevance in Retrieval Augmented Generation
by: Ni, Jingwei, et al.
Published: (2024)
by: Ni, Jingwei, et al.
Published: (2024)
LLM-as-an-Annotator: Training Lightweight Models with LLM-Annotated Examples for Aspect Sentiment Tuple Prediction
by: Hellwig, Nils Constantin, et al.
Published: (2026)
by: Hellwig, Nils Constantin, et al.
Published: (2026)
ELOQ: Resources for Enhancing LLM Detection of Out-of-Scope Questions
by: Peng, Zhiyuan, et al.
Published: (2024)
by: Peng, Zhiyuan, et al.
Published: (2024)
LLM2LLM: Boosting LLMs with Novel Iterative Data Enhancement
by: Lee, Nicholas, et al.
Published: (2024)
by: Lee, Nicholas, et al.
Published: (2024)
Similar Items
-
Reinforcement Learning-based Knowledge Distillation with LLM-as-a-Judge
by: Shen, Yiyang, et al.
Published: (2026) -
Debate as Optimization: Adaptive Conformal Prediction and Diverse Retrieval for Event Extraction
by: Wang, Sijia, et al.
Published: (2024) -
ThinkBench: Dynamic Out-of-Distribution Evaluation for Robust LLM Reasoning
by: Huang, Shulin, et al.
Published: (2025) -
Benchmarking LLM-based Relevance Judgment Methods
by: Arabzadeh, Negar, et al.
Published: (2025) -
WINELL: Wikipedia Never-Ending Updating with LLM Agents
by: Reddy, Revanth Gangi, et al.
Published: (2025)