Your Model Diversity, Not Method, Determines Reasoning Strategy
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Choraria, Moulik, Gerogiannis, Argyrios, Das, Anirban, Chakraborty, Supriyo, Kapusuzoglu, Berkcan, Lee, Chia-Hsuan, Balasubramaniam, Kartik, Zhang, Shi-Xiong, Sahu, Sambit |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Critique-Guided Distillation for Robust Reasoning via Refinement
von: Kapusuzoglu, Berkcan, et al.
Veröffentlicht: (2025)
von: Kapusuzoglu, Berkcan, et al.
Veröffentlicht: (2025)
Optimizing Reasoning Efficiency through Prompt Difficulty Prediction
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
SPEAR-MM: Selective Parameter Evaluation and Restoration via Model Merging for Efficient Financial LLM Adaptation
von: Kapusuzoglu, Berkcan, et al.
Veröffentlicht: (2025)
von: Kapusuzoglu, Berkcan, et al.
Veröffentlicht: (2025)
Context-Gated Associative Retrieval: From Theory to Transformers
von: Choraria, Moulik, et al.
Veröffentlicht: (2026)
von: Choraria, Moulik, et al.
Veröffentlicht: (2026)
CoT-Guard: Small Models for Strong Monitoring
von: Diwan, Nirav, et al.
Veröffentlicht: (2026)
von: Diwan, Nirav, et al.
Veröffentlicht: (2026)
LLM Surgery: Efficient Knowledge Unlearning and Editing in Large Language Models
von: Veldanda, Akshaj Kumar, et al.
Veröffentlicht: (2024)
von: Veldanda, Akshaj Kumar, et al.
Veröffentlicht: (2024)
Leveraging Parameter Space Symmetries for Reasoning Skill Transfer in LLMs
von: Horoi, Stefan, et al.
Veröffentlicht: (2025)
von: Horoi, Stefan, et al.
Veröffentlicht: (2025)
Hiding in Plain Sight: Detectability-Aware Antidistillation of Reasoning Models
von: Hartman, Max, et al.
Veröffentlicht: (2026)
von: Hartman, Max, et al.
Veröffentlicht: (2026)
Decomposing the Delta: What Do Models Actually Learn from Preference Pairs?
von: Lee, Chia-Hsuan, et al.
Veröffentlicht: (2026)
von: Lee, Chia-Hsuan, et al.
Veröffentlicht: (2026)
Watermarking Discrete Diffusion Language Models
von: Bagchi, Avi, et al.
Veröffentlicht: (2025)
von: Bagchi, Avi, et al.
Veröffentlicht: (2025)
Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models
von: Hartman, Max, et al.
Veröffentlicht: (2025)
von: Hartman, Max, et al.
Veröffentlicht: (2025)
DARLING: Detection Augmented Reinforcement Learning with Non-Stationary Guarantees
von: Gerogiannis, Argyrios, et al.
Veröffentlicht: (2026)
von: Gerogiannis, Argyrios, et al.
Veröffentlicht: (2026)
Is Prior-Free Black-Box Non-Stationary Reinforcement Learning Feasible?
von: Gerogiannis, Argyrios, et al.
Veröffentlicht: (2024)
von: Gerogiannis, Argyrios, et al.
Veröffentlicht: (2024)
Training Dynamics Underlying Language Model Scaling Laws: Loss Deceleration and Zero-Sum Learning
von: Mircea, Andrei, et al.
Veröffentlicht: (2025)
von: Mircea, Andrei, et al.
Veröffentlicht: (2025)
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
von: Panda, Ashwinee, et al.
Veröffentlicht: (2025)
von: Panda, Ashwinee, et al.
Veröffentlicht: (2025)
T1: A Tool-Oriented Conversational Dataset for Multi-Turn Agentic Planning
von: Chakraborty, Amartya, et al.
Veröffentlicht: (2025)
von: Chakraborty, Amartya, et al.
Veröffentlicht: (2025)
RainbowPO: A Unified Framework for Combining Improvements in Preference Optimization
von: Zhao, Hanyang, et al.
Veröffentlicht: (2024)
von: Zhao, Hanyang, et al.
Veröffentlicht: (2024)
Preference Tuning with Human Feedback on Language, Speech, and Vision Tasks: A Survey
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
Interview of the famous female impersonator of Indian Theatre: Chapal Bhaduri
von: Supriyo Chakraborty
Veröffentlicht: (2019)
von: Supriyo Chakraborty
Veröffentlicht: (2019)
Detection Augmented Bandit Procedures for Piecewise Stationary MABs: A Modular Approach
von: Huang, Yu-Han, et al.
Veröffentlicht: (2025)
von: Huang, Yu-Han, et al.
Veröffentlicht: (2025)
DAL: A Practical Prior-Free Black-Box Framework for Non-Stationary Bandits
von: Gerogiannis, Argyrios, et al.
Veröffentlicht: (2025)
von: Gerogiannis, Argyrios, et al.
Veröffentlicht: (2025)
Learning Where to Look: UCB-Driven Controlled Sensing for Quickest Change Detection
von: Huang, Yu-Han, et al.
Veröffentlicht: (2026)
von: Huang, Yu-Han, et al.
Veröffentlicht: (2026)
Editorial Commentary: Periacetabular Osteotomy for Hip Dysplasia May Result in Paralabral Cyst Resolution Despite Persistent Labral Tearing
von: Berkcan Akpinar
Veröffentlicht: (2024)
von: Berkcan Akpinar
Veröffentlicht: (2024)
Editorial Commentary : Time to Recovery Varies After Hip Gluteus Repair: Patient Expectations Should Be Set Accordingly
von: Berkcan Akpinar
Veröffentlicht: (2024)
von: Berkcan Akpinar
Veröffentlicht: (2024)
Routing with Generated Data: Annotation-Free LLM Skill Estimation and Expert Selection
von: Niu, Tianyi, et al.
Veröffentlicht: (2026)
von: Niu, Tianyi, et al.
Veröffentlicht: (2026)
Track-MDP: Reinforcement Learning for Target Tracking with Controlled Sensing
von: Subramaniam, Adarsh M., et al.
Veröffentlicht: (2024)
von: Subramaniam, Adarsh M., et al.
Veröffentlicht: (2024)
One-Dimensional Inverse Scattering Problem in Acoustics
von: Demetrios Gerogiannis
Veröffentlicht: (2011)
von: Demetrios Gerogiannis
Veröffentlicht: (2011)
DeepInsert: Early Layer Bypass for Efficient and Performant Multimodal Understanding
von: Choraria, Moulik, et al.
Veröffentlicht: (2025)
von: Choraria, Moulik, et al.
Veröffentlicht: (2025)
Semantically Grounded QFormer for Efficient Vision Language Understanding
von: Choraria, Moulik, et al.
Veröffentlicht: (2023)
von: Choraria, Moulik, et al.
Veröffentlicht: (2023)
Continual Pre-training of MoEs: How robust is your router?
von: Thérien, Benjamin, et al.
Veröffentlicht: (2025)
von: Thérien, Benjamin, et al.
Veröffentlicht: (2025)
On Understanding of the Dynamics of Model Capacity in Continual Learning
von: Chakraborty, Supriyo, et al.
Veröffentlicht: (2025)
von: Chakraborty, Supriyo, et al.
Veröffentlicht: (2025)
DIAL-SUMMER: A Structured Evaluation Framework of Hierarchical Errors in Dialogue Summaries
von: Ramnath, Sahana, et al.
Veröffentlicht: (2026)
von: Ramnath, Sahana, et al.
Veröffentlicht: (2026)
Probabilities and Potentialities of SHG-led Women Entrepreneurship in the Fish Processing and Value Addition Sector
von: Dey, Suman, et al.
Veröffentlicht: (2025)
von: Dey, Suman, et al.
Veröffentlicht: (2025)
Intrinsic ductility enhancement in Mg alloys elucidated via large-scale ab-initio calculations
von: Das, Sambit, et al.
Veröffentlicht: (2026)
von: Das, Sambit, et al.
Veröffentlicht: (2026)
SALSA: Speech Aware LLM Adaptation via Learned Steering Activation Vectors
von: Yegorova, Yekaterina, et al.
Veröffentlicht: (2026)
von: Yegorova, Yekaterina, et al.
Veröffentlicht: (2026)
AVSD: Adaptive-View Self-Distillation by Balancing Consensus and Teacher-Specific Privileged Signals
von: Nguyen, Duy, et al.
Veröffentlicht: (2026)
von: Nguyen, Duy, et al.
Veröffentlicht: (2026)
Facebook and Consumer Research: A Review, AI‐Driven Thematic Visualisation, and Research Agenda
von: Moulik Zaveri, et al.
Veröffentlicht: (2024)
von: Moulik Zaveri, et al.
Veröffentlicht: (2024)
Rebels in Search of Champions: Envisioning the Library of the Future.
von: Moulik, Amal, et al.
Veröffentlicht: (1992)
von: Moulik, Amal, et al.
Veröffentlicht: (1992)
Cosmological constraints on mass-varying dark matter
von: Chakraborty, Amlan, et al.
Veröffentlicht: (2024)
von: Chakraborty, Amlan, et al.
Veröffentlicht: (2024)
A Dataset for Multi-intensity Continuous Human Activity Recognition through Passive Sensing
von: Sen, Argha, et al.
Veröffentlicht: (2024)
von: Sen, Argha, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Critique-Guided Distillation for Robust Reasoning via Refinement
von: Kapusuzoglu, Berkcan, et al.
Veröffentlicht: (2025) -
Optimizing Reasoning Efficiency through Prompt Difficulty Prediction
von: Zhao, Bo, et al.
Veröffentlicht: (2025) -
SPEAR-MM: Selective Parameter Evaluation and Restoration via Model Merging for Efficient Financial LLM Adaptation
von: Kapusuzoglu, Berkcan, et al.
Veröffentlicht: (2025) -
Context-Gated Associative Retrieval: From Theory to Transformers
von: Choraria, Moulik, et al.
Veröffentlicht: (2026) -
CoT-Guard: Small Models for Strong Monitoring
von: Diwan, Nirav, et al.
Veröffentlicht: (2026)