Salvato in:
| Autori principali: | Nayak, Nihal V., Rodriguez-Diaz, Paula, Hulkund, Neha, Beery, Sara, Alvarez-Melis, David |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.14696 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Reinforcement Learning for Reasoning in Small LLMs: What Works and What Doesn't
di: Dang, Quy-Anh, et al.
Pubblicazione: (2025)
di: Dang, Quy-Anh, et al.
Pubblicazione: (2025)
What is the Right Notion of Distance between Predict-then-Optimize Tasks?
di: Rodriguez-Diaz, Paula, et al.
Pubblicazione: (2024)
di: Rodriguez-Diaz, Paula, et al.
Pubblicazione: (2024)
Boomerang Distillation Enables Zero-Shot Model Size Interpolation
di: Kangaslahti, Sara, et al.
Pubblicazione: (2025)
di: Kangaslahti, Sara, et al.
Pubblicazione: (2025)
Machine Unlearning Doesn't Do What You Think: Lessons for Generative AI Policy and Research
di: Cooper, A. Feder, et al.
Pubblicazione: (2024)
di: Cooper, A. Feder, et al.
Pubblicazione: (2024)
MedCalc-Bench Doesn't Measure What You Think: A Benchmark Audit and the Case for Open-Book Evaluation
di: Krohn-Grimberghe, Artus
Pubblicazione: (2026)
di: Krohn-Grimberghe, Artus
Pubblicazione: (2026)
Testing Autonomous Driving Systems -- What Really Matters and What Doesn't
di: Li, Changwen, et al.
Pubblicazione: (2025)
di: Li, Changwen, et al.
Pubblicazione: (2025)
Revisiting Intermediate-Layer Matching in Knowledge Distillation: Layer-Selection Strategy Doesn't Matter (Much)
di: Yu, Zony, et al.
Pubblicazione: (2025)
di: Yu, Zony, et al.
Pubblicazione: (2025)
Semantics at an Angle: When Cosine Similarity Works Until It Doesn't
di: You, Kisung
Pubblicazione: (2025)
di: You, Kisung
Pubblicazione: (2025)
Privacy-preserving data release leveraging optimal transport and particle gradient descent
di: Donhauser, Konstantin, et al.
Pubblicazione: (2024)
di: Donhauser, Konstantin, et al.
Pubblicazione: (2024)
Infinite Width Models That Work: Why Feature Learning Doesn't Matter as Much as You Think
di: Sernau, Luke
Pubblicazione: (2024)
di: Sernau, Luke
Pubblicazione: (2024)
Recurrent Off-Policy Deep Reinforcement Learning Doesn't Have to be Slow
di: Clark, Tyler, et al.
Pubblicazione: (2025)
di: Clark, Tyler, et al.
Pubblicazione: (2025)
When More Data Doesn't Help: Limits of Adaptation in Multitask Learning
di: Hanneke, Steve, et al.
Pubblicazione: (2026)
di: Hanneke, Steve, et al.
Pubblicazione: (2026)
Teach AI What It Doesn't Know
di: Sean Du
Pubblicazione: (2026)
di: Sean Du
Pubblicazione: (2026)
Library Designs Revisited: What Works--What Doesn't.
di: Metz, T. John, et al.
Pubblicazione: (1987)
di: Metz, T. John, et al.
Pubblicazione: (1987)
Library Learning Doesn't: The Curious Case of the Single-Use "Library"
di: Berlot-Attwell, Ian, et al.
Pubblicazione: (2024)
di: Berlot-Attwell, Ian, et al.
Pubblicazione: (2024)
Explorations of the Softmax Space: Knowing When the Neural Network Doesn't Know
di: Sikar, Daniel, et al.
Pubblicazione: (2025)
di: Sikar, Daniel, et al.
Pubblicazione: (2025)
Continuous Language Model Interpolation for Dynamic and Controllable Text Generation
di: Kangaslahti, Sara, et al.
Pubblicazione: (2024)
di: Kangaslahti, Sara, et al.
Pubblicazione: (2024)
Understanding the Role of Functional Diversity in Weight-Ensembling with Ingredient Selection and Multidimensional Scaling
di: Rojas, Alex, et al.
Pubblicazione: (2024)
di: Rojas, Alex, et al.
Pubblicazione: (2024)
In-Service and the School Library Media Specialist: What Works and What Doesn't.
di: Turner, Philip M.
Pubblicazione: (1988)
di: Turner, Philip M.
Pubblicazione: (1988)
When Structure Doesn't Help: LLMs Do Not Read Text-Attributed Graphs as Effectively as We Expected
di: Xu, Haotian, et al.
Pubblicazione: (2025)
di: Xu, Haotian, et al.
Pubblicazione: (2025)
Learning to Generate Instruction Tuning Datasets for Zero-Shot Task Adaptation
di: Nayak, Nihal V., et al.
Pubblicazione: (2024)
di: Nayak, Nihal V., et al.
Pubblicazione: (2024)
One LR Doesn't Fit All: Heavy-Tail Guided Layerwise Learning Rates for LLMs
di: He, Di, et al.
Pubblicazione: (2026)
di: He, Di, et al.
Pubblicazione: (2026)
DataS^3: Dataset Subset Selection for Specialization
di: Hulkund, Neha, et al.
Pubblicazione: (2025)
di: Hulkund, Neha, et al.
Pubblicazione: (2025)
Curated Synthetic Data Doesn't Have to Collapse: A Theoretical Study of Generative Retraining with Pluralistic Preferences
di: Falahati, Ali, et al.
Pubblicazione: (2026)
di: Falahati, Ali, et al.
Pubblicazione: (2026)
Why Prompt Optimization Works, and Why It Sometimes Doesn't: A Causal-Inspired Edit-Level Analysis
di: Gong, Shuzhi, et al.
Pubblicazione: (2026)
di: Gong, Shuzhi, et al.
Pubblicazione: (2026)
Strongly Isomorphic Neural Optimal Transport Across Incomparable Spaces
di: Sotiropoulou, Athina, et al.
Pubblicazione: (2024)
di: Sotiropoulou, Athina, et al.
Pubblicazione: (2024)
Self-Training Doesn't Flatten Language -- It Restructures It: Surface Markers Amplify While Deep Syntax Dies
di: Liu, Ming
Pubblicazione: (2026)
di: Liu, Ming
Pubblicazione: (2026)
Decomposing Elements of Problem Solving: What "Math" Does RL Teach?
di: Qin, Tian, et al.
Pubblicazione: (2025)
di: Qin, Tian, et al.
Pubblicazione: (2025)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
di: Chanda, Prateek, et al.
Pubblicazione: (2025)
di: Chanda, Prateek, et al.
Pubblicazione: (2025)
Consensus-Driven Active Model Selection
di: Kay, Justin, et al.
Pubblicazione: (2025)
di: Kay, Justin, et al.
Pubblicazione: (2025)
Consciousness Doesn't Do That
di: Matthias Michel
Pubblicazione: (2026)
di: Matthias Michel
Pubblicazione: (2026)
When Online Instruction Doesn't Measure Up: How Can You Tell, and What Should You Do?
di: Rapchak, Marcia
Pubblicazione: (2019)
di: Rapchak, Marcia
Pubblicazione: (2019)
"Something Comes through or It Doesn't": Intensive Reading in Post-Qualitative Inquiry
di: Maggie MacLure
Pubblicazione: (2024)
di: Maggie MacLure
Pubblicazione: (2024)
Do Large Language Model Benchmarks Test Reliability?
di: Vendrow, Joshua, et al.
Pubblicazione: (2025)
di: Vendrow, Joshua, et al.
Pubblicazione: (2025)
What Matters in Data for DPO?
di: Pan, Yu, et al.
Pubblicazione: (2025)
di: Pan, Yu, et al.
Pubblicazione: (2025)
AVEX: What Matters for Animal Vocalization Encoding
di: Miron, Marius, et al.
Pubblicazione: (2025)
di: Miron, Marius, et al.
Pubblicazione: (2025)
How to Connect Speech Foundation Models and Large Language Models? What Matters and What Does Not
di: Verdini, Francesco, et al.
Pubblicazione: (2024)
di: Verdini, Francesco, et al.
Pubblicazione: (2024)
Distributional Dataset Distillation with Subtask Decomposition
di: Qin, Tian, et al.
Pubblicazione: (2024)
di: Qin, Tian, et al.
Pubblicazione: (2024)
Forget What Matters, Keep the Rest: Selective Unlearning of Informative Tokens
di: Koh, Seunghee, et al.
Pubblicazione: (2026)
di: Koh, Seunghee, et al.
Pubblicazione: (2026)
Aggregation Hides Out-of-Distribution Generalization Failures from Spurious Correlations
di: Salaudeen, Olawale, et al.
Pubblicazione: (2025)
di: Salaudeen, Olawale, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Reinforcement Learning for Reasoning in Small LLMs: What Works and What Doesn't
di: Dang, Quy-Anh, et al.
Pubblicazione: (2025) -
What is the Right Notion of Distance between Predict-then-Optimize Tasks?
di: Rodriguez-Diaz, Paula, et al.
Pubblicazione: (2024) -
Boomerang Distillation Enables Zero-Shot Model Size Interpolation
di: Kangaslahti, Sara, et al.
Pubblicazione: (2025) -
Machine Unlearning Doesn't Do What You Think: Lessons for Generative AI Policy and Research
di: Cooper, A. Feder, et al.
Pubblicazione: (2024) -
MedCalc-Bench Doesn't Measure What You Think: A Benchmark Audit and the Case for Open-Book Evaluation
di: Krohn-Grimberghe, Artus
Pubblicazione: (2026)