Subset Selection for Fine-Tuning: A Utility-Diversity Balanced Approach for Mathematical Domain Adaptation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kotecha, Madhav, Vaishya, Vijendra Kumar, Gautam, Smita, Racha, Suraj |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A LoRA-Based Approach to Fine-Tuning LLMs for Educational Guidance in Resource-Constrained Settings
von: Hosen, Md Millat
Veröffentlicht: (2025)
von: Hosen, Md Millat
Veröffentlicht: (2025)
Deep Reinforcement Learning Xiangqi Player with Monte Carlo Tree Search
von: Yilmaz, Berk, et al.
Veröffentlicht: (2025)
von: Yilmaz, Berk, et al.
Veröffentlicht: (2025)
Convergence Dynamics and Stabilization Strategies of Co-Evolving Generative Models
von: Gao, Weiguo, et al.
Veröffentlicht: (2025)
von: Gao, Weiguo, et al.
Veröffentlicht: (2025)
Weakly Supervised Learners for Correction of AI Errors with Provable Performance Guarantees
von: Tyukin, Ivan Y., et al.
Veröffentlicht: (2024)
von: Tyukin, Ivan Y., et al.
Veröffentlicht: (2024)
Learning Neural Network Classifiers with Low Model Complexity
von: Jayadeva, et al.
Veröffentlicht: (2017)
von: Jayadeva, et al.
Veröffentlicht: (2017)
DQN Performance with Epsilon Greedy Policies and Prioritized Experience Replay
von: Perkins, Daniel, et al.
Veröffentlicht: (2025)
von: Perkins, Daniel, et al.
Veröffentlicht: (2025)
Pushdown Reward Machines for Reinforcement Learning
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2025)
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2025)
Unsupervised Ensemble Learning Through Deep Energy-based Models
von: Maymon, Ariel, et al.
Veröffentlicht: (2026)
von: Maymon, Ariel, et al.
Veröffentlicht: (2026)
Maximally Permissive Reward Machines
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2024)
von: Varricchione, Giovanni, et al.
Veröffentlicht: (2024)
BatteryML:An Open-source platform for Machine Learning on Battery Degradation
von: Zhang, Han, et al.
Veröffentlicht: (2023)
von: Zhang, Han, et al.
Veröffentlicht: (2023)
SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Grouped Sequential Optimization Strategy -- the Application of Hyperparameter Importance Assessment in Deep Learning
von: Wang, Ruinan, et al.
Veröffentlicht: (2025)
von: Wang, Ruinan, et al.
Veröffentlicht: (2025)
Reciprocal Learning
von: Rodemann, Julian, et al.
Veröffentlicht: (2024)
von: Rodemann, Julian, et al.
Veröffentlicht: (2024)
Large Language Model Meets Graph Neural Network in Knowledge Distillation
von: Hu, Shengxiang, et al.
Veröffentlicht: (2024)
von: Hu, Shengxiang, et al.
Veröffentlicht: (2024)
Multi-State TD Target for Model-Free Reinforcement Learning
von: Wang, Wuhao, et al.
Veröffentlicht: (2024)
von: Wang, Wuhao, et al.
Veröffentlicht: (2024)
A Neural Affinity Framework for Abstract Reasoning: Diagnosing the Compositional Gap in Transformer Architectures via Procedural Task Taxonomy
von: Ingram, Miguel, et al.
Veröffentlicht: (2025)
von: Ingram, Miguel, et al.
Veröffentlicht: (2025)
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
von: Wang, Youkang, et al.
Veröffentlicht: (2025)
von: Wang, Youkang, et al.
Veröffentlicht: (2025)
What is the $\textit{intrinsic}$ dimension of your binary data? -- and how to compute it quickly
von: Hanika, Tom, et al.
Veröffentlicht: (2024)
von: Hanika, Tom, et al.
Veröffentlicht: (2024)
Quantifying First-Order Markov Violations in Noisy Reinforcement Learning: A Causal Discovery Approach
von: Mysore, Naveen
Veröffentlicht: (2025)
von: Mysore, Naveen
Veröffentlicht: (2025)
ASNN: Learning to Suggest Neural Architectures from Performance Distributions
von: Hong, Jinwook
Veröffentlicht: (2025)
von: Hong, Jinwook
Veröffentlicht: (2025)
EcoTransformer: Attention without Multiplication
von: Gao, Xin, et al.
Veröffentlicht: (2025)
von: Gao, Xin, et al.
Veröffentlicht: (2025)
Noradrenergic-inspired gain modulation attenuates the stability gap in joint training
von: Rodriguez-Garcia, Alejandro, et al.
Veröffentlicht: (2025)
von: Rodriguez-Garcia, Alejandro, et al.
Veröffentlicht: (2025)
Topological Foundations of Reinforcement Learning
von: Kadurha, David Krame
Veröffentlicht: (2024)
von: Kadurha, David Krame
Veröffentlicht: (2024)
Hybrid Imbalanced Regression Through Unified Data-Level and Algorithm-Level Balancing
von: Shahbazi, Shermin, et al.
Veröffentlicht: (2026)
von: Shahbazi, Shermin, et al.
Veröffentlicht: (2026)
Agent-Centric Personalized Multiple Clustering with Multi-Modal LLMs
von: Chen, Ziye, et al.
Veröffentlicht: (2025)
von: Chen, Ziye, et al.
Veröffentlicht: (2025)
Bayesian Power Steering: An Effective Approach for Domain Adaptation of Diffusion Models
von: Huang, Ding, et al.
Veröffentlicht: (2024)
von: Huang, Ding, et al.
Veröffentlicht: (2024)
Topological Perspectives on Optimal Multimodal Embedding Spaces
von: B, Abdul Aziz A., et al.
Veröffentlicht: (2024)
von: B, Abdul Aziz A., et al.
Veröffentlicht: (2024)
Cross-Domain Uncertainty Quantification for Selective Prediction: A Comprehensive Bound Ablation with Transfer-Informed Betting
von: Basu, Abhinaba
Veröffentlicht: (2026)
von: Basu, Abhinaba
Veröffentlicht: (2026)
Multi-Scale Graph Learning for Anti-Sparse Downscaling
von: Fan, Yingda, et al.
Veröffentlicht: (2025)
von: Fan, Yingda, et al.
Veröffentlicht: (2025)
Graph Transformers: A Survey
von: Shehzad, Ahsan, et al.
Veröffentlicht: (2024)
von: Shehzad, Ahsan, et al.
Veröffentlicht: (2024)
QGraphLIME - Explaining Quantum Graph Neural Networks
von: Jena, Haribandhu, et al.
Veröffentlicht: (2025)
von: Jena, Haribandhu, et al.
Veröffentlicht: (2025)
From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning
von: Park, Junseok, et al.
Veröffentlicht: (2025)
von: Park, Junseok, et al.
Veröffentlicht: (2025)
Continual Learning of Domain Knowledge from Human Feedback in Text-to-SQL
von: Cook, Thomas, et al.
Veröffentlicht: (2025)
von: Cook, Thomas, et al.
Veröffentlicht: (2025)
Order-Robust Class Incremental Learning: Graph-Driven Dynamic Similarity Grouping
von: Lai, Guannan, et al.
Veröffentlicht: (2025)
von: Lai, Guannan, et al.
Veröffentlicht: (2025)
Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests
von: Sáez, Arnau Igualde, et al.
Veröffentlicht: (2025)
von: Sáez, Arnau Igualde, et al.
Veröffentlicht: (2025)
PAC-MCTS: Bias-Aware Pruning for Robust LLM-Guided Search and Planning
von: Qian, Tianhao
Veröffentlicht: (2026)
von: Qian, Tianhao
Veröffentlicht: (2026)
A social path to human-like artificial intelligence
von: Duéñez-Guzmán, Edgar A., et al.
Veröffentlicht: (2024)
von: Duéñez-Guzmán, Edgar A., et al.
Veröffentlicht: (2024)
Memory-efficient Continual Learning with Prototypical Exemplar Condensation
von: Nguyen, Minh-Duong, et al.
Veröffentlicht: (2026)
von: Nguyen, Minh-Duong, et al.
Veröffentlicht: (2026)
Tracking Changing Probabilities via Dynamic Learners
von: Madani, Omid
Veröffentlicht: (2024)
von: Madani, Omid
Veröffentlicht: (2024)
Censored Sampling for Topology Design: Guiding Diffusion with Human Preferences
von: Kim, Euihyun, et al.
Veröffentlicht: (2025)
von: Kim, Euihyun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A LoRA-Based Approach to Fine-Tuning LLMs for Educational Guidance in Resource-Constrained Settings
von: Hosen, Md Millat
Veröffentlicht: (2025) -
Deep Reinforcement Learning Xiangqi Player with Monte Carlo Tree Search
von: Yilmaz, Berk, et al.
Veröffentlicht: (2025) -
Convergence Dynamics and Stabilization Strategies of Co-Evolving Generative Models
von: Gao, Weiguo, et al.
Veröffentlicht: (2025) -
Weakly Supervised Learners for Correction of AI Errors with Provable Performance Guarantees
von: Tyukin, Ivan Y., et al.
Veröffentlicht: (2024) -
Learning Neural Network Classifiers with Low Model Complexity
von: Jayadeva, et al.
Veröffentlicht: (2017)