What properties of reasoning supervision are associated with improved downstream model quality?
Fuente:
arXiv
Saved in:
| Main Authors: | Langner, Mikołaj, Pihulski, Dzmitry, Eliasz, Jan, Rajkowski, Michał, Kazienko, Przemysław, Piasecki, Maciej, Kocoń, Jan, Ferdinan, Teddy |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AggTruth: Contextual Hallucination Detection using Aggregated Attention Scores in LLMs
by: Matys, Piotr, et al.
Published: (2025)
by: Matys, Piotr, et al.
Published: (2025)
Into the Unknown: Self-Learning Large Language Models
by: Ferdinan, Teddy, et al.
Published: (2024)
by: Ferdinan, Teddy, et al.
Published: (2024)
Language, Culture, and Ideology: Personalizing Offensiveness Detection in Political Tweets with Reasoning LLMs
by: Pihulski, Dzmitry, et al.
Published: (2025)
by: Pihulski, Dzmitry, et al.
Published: (2025)
Fortifying NLP models - dataset + code
by: Ferdinan, Teddy, et al.
Published: (2025)
by: Ferdinan, Teddy, et al.
Published: (2025)
LLMSQL: Upgrading WikiSQL for the LLM Era of Text-to-SQL
by: Pihulski, Dzmitry, et al.
Published: (2025)
by: Pihulski, Dzmitry, et al.
Published: (2025)
Divide, Cache, Conquer: Dichotomic Prompting for Efficient Multi-Label LLM-Based Classification
by: Langner, Mikołaj, et al.
Published: (2025)
by: Langner, Mikołaj, et al.
Published: (2025)
Personalized Large Language Models
by: Woźniak, Stanisław, et al.
Published: (2024)
by: Woźniak, Stanisław, et al.
Published: (2024)
Self-training Large Language Models through Knowledge Detection
by: Yeo, Wei Jie, et al.
Published: (2024)
by: Yeo, Wei Jie, et al.
Published: (2024)
Backtranslation and paraphrasing in the LLM era? Comparing data augmentation methods for emotion classification
by: Radliński, Łukasz, et al.
Published: (2025)
by: Radliński, Łukasz, et al.
Published: (2025)
Unraveling SITT: Social Influence Technique Taxonomy and Detection with LLMs
by: Mieleszczenko-Kowszewicz, Wiktoria, et al.
Published: (2025)
by: Mieleszczenko-Kowszewicz, Wiktoria, et al.
Published: (2025)
Chunking Methods on Retrieval-Augmented Generation - Effectiveness Evaluation Against Computational Cost and Limitations
by: Śmigielski, Mateusz, et al.
Published: (2026)
by: Śmigielski, Mateusz, et al.
Published: (2026)
SSN and the Hodge Conjecture – Project 1: Millennium Problems
by: Nowak, Eliasz Przemysław
Published: (2025)
by: Nowak, Eliasz Przemysław
Published: (2025)
Typology of Image Crises Using Large Language Models: A Novel Approach to Crisis Classification
by: Grzegorz Chodak, et al.
Published: (2025)
by: Grzegorz Chodak, et al.
Published: (2025)
Sociodemographic Biases in Educational Counselling by Large Language Models
by: Adamczyk, Tomasz, et al.
Published: (2026)
by: Adamczyk, Tomasz, et al.
Published: (2026)
The PLLuM Instruction Corpus
by: Pęzik, Piotr, et al.
Published: (2025)
by: Pęzik, Piotr, et al.
Published: (2025)
Predicting stock prices with ChatGPT-annotated Reddit sentiment
by: Kmak, Mateusz, et al.
Published: (2025)
by: Kmak, Mateusz, et al.
Published: (2025)
STEP-Parts: Geometric Partitioning of Boundary Representations for Large-Scale CAD Processing
by: Fan, Shen, et al.
Published: (2026)
by: Fan, Shen, et al.
Published: (2026)
Eagle and Finch: RWKV with Matrix-Valued States and Dynamic Recurrence
by: Peng, Bo, et al.
Published: (2024)
by: Peng, Bo, et al.
Published: (2024)
BEIR-PL: Zero Shot Information Retrieval Benchmark for the Polish Language
by: Wojtasik, Konrad, et al.
Published: (2023)
by: Wojtasik, Konrad, et al.
Published: (2023)
How Annotation Trains Annotators: Competence Development in Social Influence Recognition
by: Markiewicz, Maciej, et al.
Published: (2026)
by: Markiewicz, Maciej, et al.
Published: (2026)
Provable unlearning in topic modeling and downstream tasks
by: Wei, Stanley, et al.
Published: (2024)
by: Wei, Stanley, et al.
Published: (2024)
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain
by: Kosowski, Adrian, et al.
Published: (2025)
by: Kosowski, Adrian, et al.
Published: (2025)
SupResDiffGAN a new approach for the Super-Resolution task
by: Kopeć, Dawid, et al.
Published: (2025)
by: Kopeć, Dawid, et al.
Published: (2025)
FlatCAD: Fast Curvature Regularization of Neural SDFs for CAD Models
by: Yin, Haotian, et al.
Published: (2025)
by: Yin, Haotian, et al.
Published: (2025)
survex: an R package for explaining machine learning survival models
by: Spytek, Mikołaj, et al.
Published: (2023)
by: Spytek, Mikołaj, et al.
Published: (2023)
Projected Compression: Trainable Projection for Efficient Transformer Compression
by: Stefaniak, Maciej, et al.
Published: (2025)
by: Stefaniak, Maciej, et al.
Published: (2025)
Multi-step retrieval and reasoning improves radiology question answering with large language models
by: Wind, Sebastian, et al.
Published: (2025)
by: Wind, Sebastian, et al.
Published: (2025)
Enhancing AI Face Realism: Cost-Efficient Quality Improvement in Distilled Diffusion Models with a Fully Synthetic Dataset
by: Wasala, Jakub, et al.
Published: (2025)
by: Wasala, Jakub, et al.
Published: (2025)
Generative Diffusion Models for Fast Simulations of Particle Collisions at CERN
by: Kita, Mikołaj, et al.
Published: (2024)
by: Kita, Mikołaj, et al.
Published: (2024)
Demonstrating specification gaming in reasoning models
by: Bondarenko, Alexander, et al.
Published: (2025)
by: Bondarenko, Alexander, et al.
Published: (2025)
Developing PUGG for Polish: A Modern Approach to KBQA, MRC, and IR Dataset Construction
by: Sawczyn, Albert, et al.
Published: (2024)
by: Sawczyn, Albert, et al.
Published: (2024)
CoT-Self-Instruct: Building high-quality synthetic prompts for reasoning and non-reasoning tasks
by: Yu, Ping, et al.
Published: (2025)
by: Yu, Ping, et al.
Published: (2025)
What's in a Lie? How Researchers Judge the Justifiability of Deception
by: Kamiel Verbeke, et al.
Published: (2025)
by: Kamiel Verbeke, et al.
Published: (2025)
When Does Non-Uniform Replay Matter in Reinforcement Learning?
by: Korniak, Michal, et al.
Published: (2026)
by: Korniak, Michal, et al.
Published: (2026)
Joint MoE Scaling Laws: Mixture of Experts Can Be Memory Efficient
by: Ludziejewski, Jan, et al.
Published: (2025)
by: Ludziejewski, Jan, et al.
Published: (2025)
Alignment-Weighted DPO: A principled reasoning approach to improve safety alignment
by: Hu, Mengxuan, et al.
Published: (2026)
by: Hu, Mengxuan, et al.
Published: (2026)
Disentangling perception and reasoning for improving data efficiency in learning cloth manipulation without demonstrations
by: Delehelle, Donatien, et al.
Published: (2026)
by: Delehelle, Donatien, et al.
Published: (2026)
Local Intrinsic Dimension Unveils Hallucinations in Diffusion Models
by: Sobieski, Bartlomiej, et al.
Published: (2026)
by: Sobieski, Bartlomiej, et al.
Published: (2026)
Fair Indivisible Payoffs through Shapley Value
by: Czarnecki, Mikołaj, et al.
Published: (2025)
by: Czarnecki, Mikołaj, et al.
Published: (2025)
Deep Generative Models for Proton Zero Degree Calorimeter Simulations in ALICE, CERN
by: Będkowski, Patryk, et al.
Published: (2024)
by: Będkowski, Patryk, et al.
Published: (2024)
Similar Items
-
AggTruth: Contextual Hallucination Detection using Aggregated Attention Scores in LLMs
by: Matys, Piotr, et al.
Published: (2025) -
Into the Unknown: Self-Learning Large Language Models
by: Ferdinan, Teddy, et al.
Published: (2024) -
Language, Culture, and Ideology: Personalizing Offensiveness Detection in Political Tweets with Reasoning LLMs
by: Pihulski, Dzmitry, et al.
Published: (2025) -
Fortifying NLP models - dataset + code
by: Ferdinan, Teddy, et al.
Published: (2025) -
LLMSQL: Upgrading WikiSQL for the LLM Era of Text-to-SQL
by: Pihulski, Dzmitry, et al.
Published: (2025)