Your Model Diversity, Not Method, Determines Reasoning Strategy
Fuente:
arXiv
Guardado en:
| Autores principales: | Choraria, Moulik, Gerogiannis, Argyrios, Das, Anirban, Chakraborty, Supriyo, Kapusuzoglu, Berkcan, Lee, Chia-Hsuan, Balasubramaniam, Kartik, Zhang, Shi-Xiong, Sahu, Sambit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Critique-Guided Distillation for Robust Reasoning via Refinement
por: Kapusuzoglu, Berkcan, et al.
Publicado: (2025)
por: Kapusuzoglu, Berkcan, et al.
Publicado: (2025)
Optimizing Reasoning Efficiency through Prompt Difficulty Prediction
por: Zhao, Bo, et al.
Publicado: (2025)
por: Zhao, Bo, et al.
Publicado: (2025)
SPEAR-MM: Selective Parameter Evaluation and Restoration via Model Merging for Efficient Financial LLM Adaptation
por: Kapusuzoglu, Berkcan, et al.
Publicado: (2025)
por: Kapusuzoglu, Berkcan, et al.
Publicado: (2025)
Context-Gated Associative Retrieval: From Theory to Transformers
por: Choraria, Moulik, et al.
Publicado: (2026)
por: Choraria, Moulik, et al.
Publicado: (2026)
CoT-Guard: Small Models for Strong Monitoring
por: Diwan, Nirav, et al.
Publicado: (2026)
por: Diwan, Nirav, et al.
Publicado: (2026)
LLM Surgery: Efficient Knowledge Unlearning and Editing in Large Language Models
por: Veldanda, Akshaj Kumar, et al.
Publicado: (2024)
por: Veldanda, Akshaj Kumar, et al.
Publicado: (2024)
Leveraging Parameter Space Symmetries for Reasoning Skill Transfer in LLMs
por: Horoi, Stefan, et al.
Publicado: (2025)
por: Horoi, Stefan, et al.
Publicado: (2025)
Hiding in Plain Sight: Detectability-Aware Antidistillation of Reasoning Models
por: Hartman, Max, et al.
Publicado: (2026)
por: Hartman, Max, et al.
Publicado: (2026)
Decomposing the Delta: What Do Models Actually Learn from Preference Pairs?
por: Lee, Chia-Hsuan, et al.
Publicado: (2026)
por: Lee, Chia-Hsuan, et al.
Publicado: (2026)
Watermarking Discrete Diffusion Language Models
por: Bagchi, Avi, et al.
Publicado: (2025)
por: Bagchi, Avi, et al.
Publicado: (2025)
Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models
por: Hartman, Max, et al.
Publicado: (2025)
por: Hartman, Max, et al.
Publicado: (2025)
DARLING: Detection Augmented Reinforcement Learning with Non-Stationary Guarantees
por: Gerogiannis, Argyrios, et al.
Publicado: (2026)
por: Gerogiannis, Argyrios, et al.
Publicado: (2026)
Is Prior-Free Black-Box Non-Stationary Reinforcement Learning Feasible?
por: Gerogiannis, Argyrios, et al.
Publicado: (2024)
por: Gerogiannis, Argyrios, et al.
Publicado: (2024)
Training Dynamics Underlying Language Model Scaling Laws: Loss Deceleration and Zero-Sum Learning
por: Mircea, Andrei, et al.
Publicado: (2025)
por: Mircea, Andrei, et al.
Publicado: (2025)
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
por: Panda, Ashwinee, et al.
Publicado: (2025)
por: Panda, Ashwinee, et al.
Publicado: (2025)
T1: A Tool-Oriented Conversational Dataset for Multi-Turn Agentic Planning
por: Chakraborty, Amartya, et al.
Publicado: (2025)
por: Chakraborty, Amartya, et al.
Publicado: (2025)
RainbowPO: A Unified Framework for Combining Improvements in Preference Optimization
por: Zhao, Hanyang, et al.
Publicado: (2024)
por: Zhao, Hanyang, et al.
Publicado: (2024)
Preference Tuning with Human Feedback on Language, Speech, and Vision Tasks: A Survey
por: Winata, Genta Indra, et al.
Publicado: (2024)
por: Winata, Genta Indra, et al.
Publicado: (2024)
Interview of the famous female impersonator of Indian Theatre: Chapal Bhaduri
por: Supriyo Chakraborty
Publicado: (2019)
por: Supriyo Chakraborty
Publicado: (2019)
Detection Augmented Bandit Procedures for Piecewise Stationary MABs: A Modular Approach
por: Huang, Yu-Han, et al.
Publicado: (2025)
por: Huang, Yu-Han, et al.
Publicado: (2025)
DAL: A Practical Prior-Free Black-Box Framework for Non-Stationary Bandits
por: Gerogiannis, Argyrios, et al.
Publicado: (2025)
por: Gerogiannis, Argyrios, et al.
Publicado: (2025)
Learning Where to Look: UCB-Driven Controlled Sensing for Quickest Change Detection
por: Huang, Yu-Han, et al.
Publicado: (2026)
por: Huang, Yu-Han, et al.
Publicado: (2026)
Editorial Commentary: Periacetabular Osteotomy for Hip Dysplasia May Result in Paralabral Cyst Resolution Despite Persistent Labral Tearing
por: Berkcan Akpinar
Publicado: (2024)
por: Berkcan Akpinar
Publicado: (2024)
Editorial Commentary : Time to Recovery Varies After Hip Gluteus Repair: Patient Expectations Should Be Set Accordingly
por: Berkcan Akpinar
Publicado: (2024)
por: Berkcan Akpinar
Publicado: (2024)
Routing with Generated Data: Annotation-Free LLM Skill Estimation and Expert Selection
por: Niu, Tianyi, et al.
Publicado: (2026)
por: Niu, Tianyi, et al.
Publicado: (2026)
Track-MDP: Reinforcement Learning for Target Tracking with Controlled Sensing
por: Subramaniam, Adarsh M., et al.
Publicado: (2024)
por: Subramaniam, Adarsh M., et al.
Publicado: (2024)
One-Dimensional Inverse Scattering Problem in Acoustics
por: Demetrios Gerogiannis
Publicado: (2011)
por: Demetrios Gerogiannis
Publicado: (2011)
DeepInsert: Early Layer Bypass for Efficient and Performant Multimodal Understanding
por: Choraria, Moulik, et al.
Publicado: (2025)
por: Choraria, Moulik, et al.
Publicado: (2025)
Semantically Grounded QFormer for Efficient Vision Language Understanding
por: Choraria, Moulik, et al.
Publicado: (2023)
por: Choraria, Moulik, et al.
Publicado: (2023)
Continual Pre-training of MoEs: How robust is your router?
por: Thérien, Benjamin, et al.
Publicado: (2025)
por: Thérien, Benjamin, et al.
Publicado: (2025)
On Understanding of the Dynamics of Model Capacity in Continual Learning
por: Chakraborty, Supriyo, et al.
Publicado: (2025)
por: Chakraborty, Supriyo, et al.
Publicado: (2025)
DIAL-SUMMER: A Structured Evaluation Framework of Hierarchical Errors in Dialogue Summaries
por: Ramnath, Sahana, et al.
Publicado: (2026)
por: Ramnath, Sahana, et al.
Publicado: (2026)
Probabilities and Potentialities of SHG-led Women Entrepreneurship in the Fish Processing and Value Addition Sector
por: Dey, Suman, et al.
Publicado: (2025)
por: Dey, Suman, et al.
Publicado: (2025)
Intrinsic ductility enhancement in Mg alloys elucidated via large-scale ab-initio calculations
por: Das, Sambit, et al.
Publicado: (2026)
por: Das, Sambit, et al.
Publicado: (2026)
SALSA: Speech Aware LLM Adaptation via Learned Steering Activation Vectors
por: Yegorova, Yekaterina, et al.
Publicado: (2026)
por: Yegorova, Yekaterina, et al.
Publicado: (2026)
AVSD: Adaptive-View Self-Distillation by Balancing Consensus and Teacher-Specific Privileged Signals
por: Nguyen, Duy, et al.
Publicado: (2026)
por: Nguyen, Duy, et al.
Publicado: (2026)
Facebook and Consumer Research: A Review, AI‐Driven Thematic Visualisation, and Research Agenda
por: Moulik Zaveri, et al.
Publicado: (2024)
por: Moulik Zaveri, et al.
Publicado: (2024)
Rebels in Search of Champions: Envisioning the Library of the Future.
por: Moulik, Amal, et al.
Publicado: (1992)
por: Moulik, Amal, et al.
Publicado: (1992)
Cosmological constraints on mass-varying dark matter
por: Chakraborty, Amlan, et al.
Publicado: (2024)
por: Chakraborty, Amlan, et al.
Publicado: (2024)
A Dataset for Multi-intensity Continuous Human Activity Recognition through Passive Sensing
por: Sen, Argha, et al.
Publicado: (2024)
por: Sen, Argha, et al.
Publicado: (2024)
Ejemplares similares
-
Critique-Guided Distillation for Robust Reasoning via Refinement
por: Kapusuzoglu, Berkcan, et al.
Publicado: (2025) -
Optimizing Reasoning Efficiency through Prompt Difficulty Prediction
por: Zhao, Bo, et al.
Publicado: (2025) -
SPEAR-MM: Selective Parameter Evaluation and Restoration via Model Merging for Efficient Financial LLM Adaptation
por: Kapusuzoglu, Berkcan, et al.
Publicado: (2025) -
Context-Gated Associative Retrieval: From Theory to Transformers
por: Choraria, Moulik, et al.
Publicado: (2026) -
CoT-Guard: Small Models for Strong Monitoring
por: Diwan, Nirav, et al.
Publicado: (2026)