IT-OSE: Exploring Optimal Sample Size for Industrial Data Augmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Mingchun, Zhao, Rongqiang, Huang, Zhennan, Ding, Songyu, Liu, Jie |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DS-Diffusion: Data Style-Guided Diffusion Model for Time-Series Generation
by: Sun, Mingchun, et al.
Published: (2025)
by: Sun, Mingchun, et al.
Published: (2025)
IGADA-IoT: IoT Sensor Energy Optimization in Wireless Sensor Networks Driven by Automatic Data Augmentation
by: Sun, Mingchun, et al.
Published: (2026)
by: Sun, Mingchun, et al.
Published: (2026)
Feature Recalibration Based Olfactory-Visual Multimodal Model for Enhanced Rice Deterioration Detection
by: Zhao, Rongqiang, et al.
Published: (2026)
by: Zhao, Rongqiang, et al.
Published: (2026)
Synthetic Data Generation for Augmenting Small Samples
by: Liu, Dan, et al.
Published: (2025)
by: Liu, Dan, et al.
Published: (2025)
Exploring Time-Step Size in Reinforcement Learning for Sepsis Treatment
by: Sun, Yingchuan, et al.
Published: (2025)
by: Sun, Yingchuan, et al.
Published: (2025)
Time Transfer: On Optimal Learning Rate and Batch Size In The Infinite Data Limit
by: Filatov, Oleg, et al.
Published: (2024)
by: Filatov, Oleg, et al.
Published: (2024)
Enhancing Obsolescence Forecasting with Deep Generative Data Augmentation: A Semi-Supervised Framework for Low-Data Industrial Applications
by: Saad, Elie, et al.
Published: (2025)
by: Saad, Elie, et al.
Published: (2025)
Conflict-Aware Pseudo Labeling via Optimal Transport for Entity Alignment
by: Ding, Qijie, et al.
Published: (2022)
by: Ding, Qijie, et al.
Published: (2022)
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models
by: Luo, Yifan, et al.
Published: (2025)
by: Luo, Yifan, et al.
Published: (2025)
Gate Recurrent Unit for Efficient Industrial Gas Identification
by: Wang, Ding
Published: (2024)
by: Wang, Ding
Published: (2024)
RealGen: Retrieval Augmented Generation for Controllable Traffic Scenarios
by: Ding, Wenhao, et al.
Published: (2023)
by: Ding, Wenhao, et al.
Published: (2023)
Reward-Augmented Data Enhances Direct Preference Alignment of LLMs
by: Zhang, Shenao, et al.
Published: (2024)
by: Zhang, Shenao, et al.
Published: (2024)
Augmenting Limited and Biased RCTs through Pseudo-Sample Matching-Based Observational Data Fusion Method
by: Han, Kairong, et al.
Published: (2025)
by: Han, Kairong, et al.
Published: (2025)
Augmenting Offline RL with Unlabeled Data
by: Wang, Zhao, et al.
Published: (2024)
by: Wang, Zhao, et al.
Published: (2024)
Effective Sample Size and Generalization Bounds for Temporal Networks
by: Gahtan, Barak, et al.
Published: (2025)
by: Gahtan, Barak, et al.
Published: (2025)
Learning Probabilities of Causation with Mask-Augmented Data
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
Industrial Energy Disaggregation with Digital Twin-generated Dataset and Efficient Data Augmentation
by: Internò, Christian, et al.
Published: (2025)
by: Internò, Christian, et al.
Published: (2025)
Feature-to-Image Data Augmentation: Improving Model Feature Extraction with Cluster-Guided Synthetic Samples
by: Haghbin, Yasaman, et al.
Published: (2024)
by: Haghbin, Yasaman, et al.
Published: (2024)
MolMem: Memory-Augmented Agentic Reinforcement Learning for Sample-Efficient Molecular Optimization
by: Wang, Ziqing, et al.
Published: (2026)
by: Wang, Ziqing, et al.
Published: (2026)
Mixture of Diverse Size Experts
by: Sun, Manxi, et al.
Published: (2024)
by: Sun, Manxi, et al.
Published: (2024)
A Classical View on Benign Overfitting: The Role of Sample Size
by: Park, Junhyung, et al.
Published: (2025)
by: Park, Junhyung, et al.
Published: (2025)
Learning with Imbalanced Noisy Data by Preventing Bias in Sample Selection
by: Liu, Huafeng, et al.
Published: (2024)
by: Liu, Huafeng, et al.
Published: (2024)
GeoMAE: Masking Representation Learning for Spatio-Temporal Graph Forecasting with Missing Values
by: Ke, Songyu, et al.
Published: (2025)
by: Ke, Songyu, et al.
Published: (2025)
Optimal Embedding Learning Rate in LLMs: The Effect of Vocabulary Size
by: Hayou, Soufiane, et al.
Published: (2025)
by: Hayou, Soufiane, et al.
Published: (2025)
SAFLEX: Self-Adaptive Augmentation via Feature Label Extrapolation
by: Ding, Mucong, et al.
Published: (2024)
by: Ding, Mucong, et al.
Published: (2024)
Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport
by: Sun, Mingyang, et al.
Published: (2025)
by: Sun, Mingyang, et al.
Published: (2025)
RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation
by: Xue, Yuquan, et al.
Published: (2025)
by: Xue, Yuquan, et al.
Published: (2025)
Lossless Compression: A New Benchmark for Time Series Model Evaluation
by: Wan, Meng, et al.
Published: (2025)
by: Wan, Meng, et al.
Published: (2025)
A Comprehensive Survey on Data Augmentation
by: Wang, Zaitian, et al.
Published: (2024)
by: Wang, Zaitian, et al.
Published: (2024)
Exploring Multi-Modal Data with Tool-Augmented LLM Agents for Precise Causal Discovery
by: Shen, ChengAo, et al.
Published: (2024)
by: Shen, ChengAo, et al.
Published: (2024)
Real-Fake: Effective Training Data Synthesis Through Distribution Matching
by: Yuan, Jianhao, et al.
Published: (2023)
by: Yuan, Jianhao, et al.
Published: (2023)
Optimal Transport for Structure Learning Under Missing Data
by: Vo, Vy, et al.
Published: (2024)
by: Vo, Vy, et al.
Published: (2024)
Learning Optimal and Sample-Efficient Decision Policies with Guarantees
by: Shao, Daqian
Published: (2026)
by: Shao, Daqian
Published: (2026)
BEACON: Bayesian Optimal Stopping for Efficient LLM Sampling
by: Wan, Guangya, et al.
Published: (2025)
by: Wan, Guangya, et al.
Published: (2025)
Actor-Critics Can Achieve Optimal Sample Efficiency
by: Tan, Kevin, et al.
Published: (2025)
by: Tan, Kevin, et al.
Published: (2025)
Planning-Augmented Sampling with Early Guidance for High-Reward Discovery
by: Zhu, Rui, et al.
Published: (2025)
by: Zhu, Rui, et al.
Published: (2025)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
by: Huang, Xingshuai, et al.
Published: (2024)
by: Huang, Xingshuai, et al.
Published: (2024)
IsoCompute Playbook: Optimally Scaling Sampling Compute for LLM RL
by: Cheng, Zhoujun, et al.
Published: (2026)
by: Cheng, Zhoujun, et al.
Published: (2026)
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
by: Ji, Kaixuan, et al.
Published: (2026)
by: Ji, Kaixuan, et al.
Published: (2026)
Review of Data-centric Time Series Analysis from Sample, Feature, and Period
by: Sun, Chenxi, et al.
Published: (2024)
by: Sun, Chenxi, et al.
Published: (2024)
Similar Items
-
DS-Diffusion: Data Style-Guided Diffusion Model for Time-Series Generation
by: Sun, Mingchun, et al.
Published: (2025) -
IGADA-IoT: IoT Sensor Energy Optimization in Wireless Sensor Networks Driven by Automatic Data Augmentation
by: Sun, Mingchun, et al.
Published: (2026) -
Feature Recalibration Based Olfactory-Visual Multimodal Model for Enhanced Rice Deterioration Detection
by: Zhao, Rongqiang, et al.
Published: (2026) -
Synthetic Data Generation for Augmenting Small Samples
by: Liu, Dan, et al.
Published: (2025) -
Exploring Time-Step Size in Reinforcement Learning for Sepsis Treatment
by: Sun, Yingchuan, et al.
Published: (2025)