Improving reasoning at inference time via uncertainty minimisation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Legrand, Nicolas, Enevoldsen, Kenneth, Kardos, Márton, Nielbo, Kristoffer |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Scandinavian Embedding Benchmarks: Comprehensive Assessment of Multilingual and Monolingual Text Embedding
von: Enevoldsen, Kenneth, et al.
Veröffentlicht: (2024)
von: Enevoldsen, Kenneth, et al.
Veröffentlicht: (2024)
topicwizard -- a Modern, Model-agnostic Framework for Topic Model Visualization and Interpretation
von: Kardos, Márton, et al.
Veröffentlicht: (2025)
von: Kardos, Márton, et al.
Veröffentlicht: (2025)
Naturalistic measure of social norms alignment
von: Kostiuk, Yevhen, et al.
Veröffentlicht: (2026)
von: Kostiuk, Yevhen, et al.
Veröffentlicht: (2026)
Maintaining MTEB: Towards Long Term Usability and Reproducibility of Embedding Benchmarks
von: Chung, Isaac, et al.
Veröffentlicht: (2025)
von: Chung, Isaac, et al.
Veröffentlicht: (2025)
$S^3$ -- Semantic Signal Separation
von: Kardos, Márton, et al.
Veröffentlicht: (2024)
von: Kardos, Márton, et al.
Veröffentlicht: (2024)
Topeax -- An Improved Clustering Topic Model with Density Peak Detection and Lexical-Semantic Term Importance
von: Kardos, Márton
Veröffentlicht: (2026)
von: Kardos, Márton
Veröffentlicht: (2026)
Dynaword: From One-shot to Continuously Developed Datasets
von: Enevoldsen, Kenneth, et al.
Veröffentlicht: (2025)
von: Enevoldsen, Kenneth, et al.
Veröffentlicht: (2025)
Continuous sentiment scores for literary and multilingual contexts
von: Lyngbaek, Laurits, et al.
Veröffentlicht: (2025)
von: Lyngbaek, Laurits, et al.
Veröffentlicht: (2025)
Is Sentiment Banana-Shaped? Exploring the Geometry and Portability of Sentiment Concept Vectors
von: Lyngbaek, Laurits, et al.
Veröffentlicht: (2026)
von: Lyngbaek, Laurits, et al.
Veröffentlicht: (2026)
Exposing Assumptions in AI Benchmarks through Cognitive Modelling
von: Rystrøm, Jonathan H., et al.
Veröffentlicht: (2024)
von: Rystrøm, Jonathan H., et al.
Veröffentlicht: (2024)
Enhancing reasoning accuracy in large language models during inference time
von: Sharma, Vinay, et al.
Veröffentlicht: (2026)
von: Sharma, Vinay, et al.
Veröffentlicht: (2026)
Are Chatbots Reliable Text Annotators? Sometimes
von: Kristensen-McLachlan, Ross Deans, et al.
Veröffentlicht: (2023)
von: Kristensen-McLachlan, Ross Deans, et al.
Veröffentlicht: (2023)
Improving moment tensor solutions under Earth structure uncertainty with simulation-based inference
von: Saoulis, A. A., et al.
Veröffentlicht: (2026)
von: Saoulis, A. A., et al.
Veröffentlicht: (2026)
The Coverage Illusion: From Pre-retrieval Routing Failure to Post-retrieval Cascades in a Production RAG System
von: Hussain, Zafar, et al.
Veröffentlicht: (2026)
von: Hussain, Zafar, et al.
Veröffentlicht: (2026)
Encoder vs Decoder: Comparative Analysis of Encoder and Decoder Language Models on Multilingual NLU Tasks
von: Nielsen, Dan Saattrup, et al.
Veröffentlicht: (2024)
von: Nielsen, Dan Saattrup, et al.
Veröffentlicht: (2024)
AIOptimizer - Software performance optimisation prototype for cost minimisation
von: Zambare, Noopur
Veröffentlicht: (2023)
von: Zambare, Noopur
Veröffentlicht: (2023)
MathDivide: Improved mathematical reasoning by large language models
von: Srivastava, Saksham Sahai, et al.
Veröffentlicht: (2024)
von: Srivastava, Saksham Sahai, et al.
Veröffentlicht: (2024)
Remedying uncertainty representations in visual inference through Explaining-Away Variational Autoencoders
von: Catoni, Josefina, et al.
Veröffentlicht: (2024)
von: Catoni, Josefina, et al.
Veröffentlicht: (2024)
REPEAT: Improving Uncertainty Estimation in Representation Learning Explainability
von: Wickstrøm, Kristoffer K., et al.
Veröffentlicht: (2024)
von: Wickstrøm, Kristoffer K., et al.
Veröffentlicht: (2024)
Resolving space-sharing conflicts in road user interactions through uncertainty reduction: An active inference-based computational model
von: Schumann, Julian F., et al.
Veröffentlicht: (2026)
von: Schumann, Julian F., et al.
Veröffentlicht: (2026)
Grounding Text Embeddings in Stakeholder Associations
von: Rystrøm, Jonathan, et al.
Veröffentlicht: (2026)
von: Rystrøm, Jonathan, et al.
Veröffentlicht: (2026)
AI reasoning effort predicts human decision time in content moderation
von: Davidson, Thomas
Veröffentlicht: (2025)
von: Davidson, Thomas
Veröffentlicht: (2025)
Demonstrating specification gaming in reasoning models
von: Bondarenko, Alexander, et al.
Veröffentlicht: (2025)
von: Bondarenko, Alexander, et al.
Veröffentlicht: (2025)
Causal reasoning in difference graphs
von: Assaad, Charles K.
Veröffentlicht: (2024)
von: Assaad, Charles K.
Veröffentlicht: (2024)
Mathematical reasoning and the computer
von: Buzzard, Kevin
Veröffentlicht: (2025)
von: Buzzard, Kevin
Veröffentlicht: (2025)
plingo: A system for probabilistic reasoning in clingo based on lpmln
von: Hahn, Susana, et al.
Veröffentlicht: (2022)
von: Hahn, Susana, et al.
Veröffentlicht: (2022)
Harnessing the power of LLMs for normative reasoning in MASs
von: Savarimuthu, Bastin Tony Roy, et al.
Veröffentlicht: (2024)
von: Savarimuthu, Bastin Tony Roy, et al.
Veröffentlicht: (2024)
Improving Zero-Shot Offline RL via Behavioral Task Sampling
von: Bendib, Nazim, et al.
Veröffentlicht: (2026)
von: Bendib, Nazim, et al.
Veröffentlicht: (2026)
Reinforcing privacy reasoning in LLMs via normative simulacra from fiction
von: Franchi, Matt, et al.
Veröffentlicht: (2026)
von: Franchi, Matt, et al.
Veröffentlicht: (2026)
Deliberative Alignment is Deep, but Uncertainty Remains: Inference time safety improvement in reasoning via attribution of unsafe behavior to base model
von: Pathmanathan, Pankayaraj, et al.
Veröffentlicht: (2026)
von: Pathmanathan, Pankayaraj, et al.
Veröffentlicht: (2026)
CP or DP? Why Not Both: A Case Study in the Partial Shop Scheduling Problem
von: Legrand, Emma, et al.
Veröffentlicht: (2026)
von: Legrand, Emma, et al.
Veröffentlicht: (2026)
CoT-Self-Instruct: Building high-quality synthetic prompts for reasoning and non-reasoning tasks
von: Yu, Ping, et al.
Veröffentlicht: (2025)
von: Yu, Ping, et al.
Veröffentlicht: (2025)
A transformer architecture alteration to incentivise externalised reasoning
von: Pavlova, Elizabeth, et al.
Veröffentlicht: (2026)
von: Pavlova, Elizabeth, et al.
Veröffentlicht: (2026)
Logos: An evolvable reasoning engine for rational molecular design
von: Wen, Haibin, et al.
Veröffentlicht: (2026)
von: Wen, Haibin, et al.
Veröffentlicht: (2026)
Effects of structure on reasoning in instance-level Self-Discover
von: Gunasekara, Sachith, et al.
Veröffentlicht: (2025)
von: Gunasekara, Sachith, et al.
Veröffentlicht: (2025)
From Flexibility to Manipulation: The Slippery Slope of XAI Evaluation
von: Wickstrøm, Kristoffer, et al.
Veröffentlicht: (2024)
von: Wickstrøm, Kristoffer, et al.
Veröffentlicht: (2024)
Limited Reasoning Space: The cage of long-horizon reasoning in LLMs
von: Li, Zhenyu, et al.
Veröffentlicht: (2026)
von: Li, Zhenyu, et al.
Veröffentlicht: (2026)
TIM-PRM: Verifying multimodal reasoning with Tool-Integrated PRM
von: Kuang, Peng, et al.
Veröffentlicht: (2025)
von: Kuang, Peng, et al.
Veröffentlicht: (2025)
Unsupervised decoding of encoded reasoning using language model interpretability
von: Fang, Ching, et al.
Veröffentlicht: (2025)
von: Fang, Ching, et al.
Veröffentlicht: (2025)
Controllable and explainable personality sliders for LLMs at inference time
von: Hoppe, Florian, et al.
Veröffentlicht: (2026)
von: Hoppe, Florian, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
The Scandinavian Embedding Benchmarks: Comprehensive Assessment of Multilingual and Monolingual Text Embedding
von: Enevoldsen, Kenneth, et al.
Veröffentlicht: (2024) -
topicwizard -- a Modern, Model-agnostic Framework for Topic Model Visualization and Interpretation
von: Kardos, Márton, et al.
Veröffentlicht: (2025) -
Naturalistic measure of social norms alignment
von: Kostiuk, Yevhen, et al.
Veröffentlicht: (2026) -
Maintaining MTEB: Towards Long Term Usability and Reproducibility of Embedding Benchmarks
von: Chung, Isaac, et al.
Veröffentlicht: (2025) -
$S^3$ -- Semantic Signal Separation
von: Kardos, Márton, et al.
Veröffentlicht: (2024)