Convex Dataset Valuation for Post-Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zeng, Siqi, Jung, Christopher, Li, Rui, Kang, Zhe, Li, Ming, Noorshams, Nima, Wang, Zhigang, Peng, Fuchun, Zhao, Han, Feng, Xue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Expectation Error Bounds for Transfer Learning in Linear Regression and Linear Neural Networks
von: Liu, Meitong, et al.
Veröffentlicht: (2026)
von: Liu, Meitong, et al.
Veröffentlicht: (2026)
Personalized Interpolation: Achieving Efficient Conversion Estimation with Flexible Optimization Windows
von: Zhang, Xin, et al.
Veröffentlicht: (2025)
von: Zhang, Xin, et al.
Veröffentlicht: (2025)
A Unified Knowledge-Distillation and Semi-Supervised Learning Framework to Improve Industrial Ads Delivery Systems
von: Eghbalzadeh, Hamid, et al.
Veröffentlicht: (2025)
von: Eghbalzadeh, Hamid, et al.
Veröffentlicht: (2025)
Proper Dataset Valuation by Pointwise Mutual Information
von: Zheng, Shuran, et al.
Veröffentlicht: (2024)
von: Zheng, Shuran, et al.
Veröffentlicht: (2024)
Algorithm-Relative Trajectory Valuation in Policy Gradient Control
von: Li, Shihao, et al.
Veröffentlicht: (2025)
von: Li, Shihao, et al.
Veröffentlicht: (2025)
Boost Post-Training Quantization via Null Space Optimization for Large Language Models
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025)
Fast-DataShapley: Neural Modeling for Training Data Valuation
von: Sun, Haifeng, et al.
Veröffentlicht: (2025)
von: Sun, Haifeng, et al.
Veröffentlicht: (2025)
MergeBench: A Benchmark for Merging Domain-Specialized LLMs
von: He, Yifei, et al.
Veröffentlicht: (2025)
von: He, Yifei, et al.
Veröffentlicht: (2025)
NoiseRater: Meta-Learned Noise Valuation for Diffusion Model Training
von: Wu, Fang, et al.
Veröffentlicht: (2026)
von: Wu, Fang, et al.
Veröffentlicht: (2026)
Learning Structured Representations by Embedding Class Hierarchy with Fast Optimal Transport
von: Zeng, Siqi, et al.
Veröffentlicht: (2024)
von: Zeng, Siqi, et al.
Veröffentlicht: (2024)
VISAGNN: Versatile Staleness-Aware Efficient Training on Large-Scale Graphs
von: Xue, Rui
Veröffentlicht: (2025)
von: Xue, Rui
Veröffentlicht: (2025)
Using Temperature Sampling to Effectively Train Robot Learning Policies on Imbalanced Datasets
von: Patil, Basavasagar, et al.
Veröffentlicht: (2025)
von: Patil, Basavasagar, et al.
Veröffentlicht: (2025)
Neural Dynamic Data Valuation: A Stochastic Optimal Control Approach
von: Liang, Zhangyong, et al.
Veröffentlicht: (2024)
von: Liang, Zhangyong, et al.
Veröffentlicht: (2024)
HOSL: Hybrid-Order Split Learning for Memory-Constrained Edge Training
von: Lnu, Aakriti, et al.
Veröffentlicht: (2026)
von: Lnu, Aakriti, et al.
Veröffentlicht: (2026)
Is Data Valuation Learnable and Interpretable?
von: Wu, Ou, et al.
Veröffentlicht: (2024)
von: Wu, Ou, et al.
Veröffentlicht: (2024)
Preference Discerning with LLM-Enhanced Generative Retrieval
von: Paischer, Fabian, et al.
Veröffentlicht: (2024)
von: Paischer, Fabian, et al.
Veröffentlicht: (2024)
Spend Wisely: Maximizing Post-Training Gains in Iterative Synthetic Data Bootstrapping
von: Yang, Pu, et al.
Veröffentlicht: (2025)
von: Yang, Pu, et al.
Veröffentlicht: (2025)
A Convex-optimization-based Layer-wise Post-training Pruner for Large Language Models
von: Zhao, Pengxiang, et al.
Veröffentlicht: (2024)
von: Zhao, Pengxiang, et al.
Veröffentlicht: (2024)
Who is In Charge? Dissecting Role Conflicts in Instruction Following
von: Zeng, Siqi
Veröffentlicht: (2025)
von: Zeng, Siqi
Veröffentlicht: (2025)
LRQ: Optimizing Post-Training Quantization for Large Language Models by Learning Low-Rank Weight-Scaling Matrices
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2024)
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2024)
Bridging On-Device and Cloud LLMs for Collaborative Reasoning: A Unified Methodology for Local Routing and Post-Training
von: Fang, Wenzhi, et al.
Veröffentlicht: (2025)
von: Fang, Wenzhi, et al.
Veröffentlicht: (2025)
Astro: Activation-guided Structured Regularization for Outlier-Robust LLM Post-Training Quantization
von: Chen, Xi, et al.
Veröffentlicht: (2026)
von: Chen, Xi, et al.
Veröffentlicht: (2026)
RobustVLA: Robustness-Aware Reinforcement Post-Training for Vision-Language-Action Models
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradients
von: Li, Ming, et al.
Veröffentlicht: (2025)
von: Li, Ming, et al.
Veröffentlicht: (2025)
FlexRound: Learnable Rounding based on Element-wise Division for Post-Training Quantization
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2023)
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2023)
Efficient Training of Neural Fractional-Order Differential Equation via Adjoint Backpropagation
von: Kang, Qiyu, et al.
Veröffentlicht: (2025)
von: Kang, Qiyu, et al.
Veröffentlicht: (2025)
Data Distribution Valuation
von: Xu, Xinyi, et al.
Veröffentlicht: (2024)
von: Xu, Xinyi, et al.
Veröffentlicht: (2024)
FaultDiffusion: Few-Shot Fault Time Series Generation with Diffusion Model
von: Xu, Yi, et al.
Veröffentlicht: (2025)
von: Xu, Yi, et al.
Veröffentlicht: (2025)
EcoVal: An Efficient Data Valuation Framework for Machine Learning
von: Tarun, Ayush K, et al.
Veröffentlicht: (2024)
von: Tarun, Ayush K, et al.
Veröffentlicht: (2024)
SOC-ICNN: From Polyhedral to Conic Geometry for Learning Convex Surrogate Functions
von: Liu, Kang, et al.
Veröffentlicht: (2026)
von: Liu, Kang, et al.
Veröffentlicht: (2026)
RiskPO: Risk-based Policy Optimization via Verifiable Reward for LLM Post-Training
von: Ren, Tao, et al.
Veröffentlicht: (2025)
von: Ren, Tao, et al.
Veröffentlicht: (2025)
Improved Dimension Dependence for Bandit Convex Optimization with Gradient Variations
von: Yu, Hang, et al.
Veröffentlicht: (2026)
von: Yu, Hang, et al.
Veröffentlicht: (2026)
Task-Stratified Knowledge Scaling Laws for Post-Training Quantized Large Language Models
von: Zhou, Chenxi, et al.
Veröffentlicht: (2025)
von: Zhou, Chenxi, et al.
Veröffentlicht: (2025)
Learning Structured Representations with Hyperbolic Embeddings
von: Sinha, Aditya, et al.
Veröffentlicht: (2024)
von: Sinha, Aditya, et al.
Veröffentlicht: (2024)
EXPRESS: An LLM-Generated Explainable Property Valuation System with Neighbor Imputation
von: Du, Wei-Wei, et al.
Veröffentlicht: (2025)
von: Du, Wei-Wei, et al.
Veröffentlicht: (2025)
Unlocking the Pre-Trained Model as a Dual-Alignment Calibrator for Post-Trained LLMs
von: Luo, Beier, et al.
Veröffentlicht: (2026)
von: Luo, Beier, et al.
Veröffentlicht: (2026)
GIO: Gradient Information Optimization for Training Dataset Selection
von: Everaert, Dante, et al.
Veröffentlicht: (2023)
von: Everaert, Dante, et al.
Veröffentlicht: (2023)
Erase then Rectify: A Training-Free Parameter Editing Approach for Cost-Effective Graph Unlearning
von: Yang, Zhe-Rui, et al.
Veröffentlicht: (2024)
von: Yang, Zhe-Rui, et al.
Veröffentlicht: (2024)
Byzantine-Robust Federated Learning Framework with Post-Quantum Secure Aggregation for Real-Time Threat Intelligence Sharing in Critical IoT Infrastructure
von: Rahmati, Milad, et al.
Veröffentlicht: (2026)
von: Rahmati, Milad, et al.
Veröffentlicht: (2026)
When LRP Diverges from Leave-One-Out in Transformers
von: You, Weiqiu, et al.
Veröffentlicht: (2025)
von: You, Weiqiu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Expectation Error Bounds for Transfer Learning in Linear Regression and Linear Neural Networks
von: Liu, Meitong, et al.
Veröffentlicht: (2026) -
Personalized Interpolation: Achieving Efficient Conversion Estimation with Flexible Optimization Windows
von: Zhang, Xin, et al.
Veröffentlicht: (2025) -
A Unified Knowledge-Distillation and Semi-Supervised Learning Framework to Improve Industrial Ads Delivery Systems
von: Eghbalzadeh, Hamid, et al.
Veröffentlicht: (2025) -
Proper Dataset Valuation by Pointwise Mutual Information
von: Zheng, Shuran, et al.
Veröffentlicht: (2024) -
Algorithm-Relative Trajectory Valuation in Policy Gradient Control
von: Li, Shihao, et al.
Veröffentlicht: (2025)