ENTP: Enhancing Low-Quality SFT Data via Neural-Symbolic Text Purge-Mix
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Zile, Li, Ling, Di, Na, Pang, Jinlong, Zhou, Yao, Cheng, Hao, Han, Bo, Wei, Jiaheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Token Cleaning: Fine-Grained Data Selection for LLM Supervised Fine-Tuning
by: Pang, Jinlong, et al.
Published: (2025)
by: Pang, Jinlong, et al.
Published: (2025)
ENTP: Encoder-only Next Token Prediction
by: Ewer, Ethan, et al.
Published: (2024)
by: Ewer, Ethan, et al.
Published: (2024)
SelectMix: Enhancing Label Noise Robustness through Targeted Sample Mixing
by: Liu, Qiuhao, et al.
Published: (2025)
by: Liu, Qiuhao, et al.
Published: (2025)
LM-mixup: Text Data Augmentation via Language Model based Mixup
by: Deng, Zhijie, et al.
Published: (2025)
by: Deng, Zhijie, et al.
Published: (2025)
LogPurge: Log Data Purification for Anomaly Detection via Rule-Enhanced Filtering
by: Zhang, Shenglin, et al.
Published: (2025)
by: Zhang, Shenglin, et al.
Published: (2025)
GAC: Noise-Aware Adaptive Mixing for Hybrid SFT-RL Post-Training
by: Hu, Yuelin, et al.
Published: (2026)
by: Hu, Yuelin, et al.
Published: (2026)
Incentivizing High-quality Participation From Federated Learning Agents
by: Pang, Jinlong, et al.
Published: (2025)
by: Pang, Jinlong, et al.
Published: (2025)
SpikeVoice: High-Quality Text-to-Speech Via Efficient Spiking Neural Network
by: Wang, Kexin, et al.
Published: (2024)
by: Wang, Kexin, et al.
Published: (2024)
Attentional Processing Biases in Young People With Binging and Purging Behavior
by: Aglaia Freccero, et al.
Published: (2025)
by: Aglaia Freccero, et al.
Published: (2025)
Purge-Gate: Backpropagation-Free Test-Time Adaptation for Point Clouds Classification via Token Purging
by: Yazdanpanah, Moslem, et al.
Published: (2025)
by: Yazdanpanah, Moslem, et al.
Published: (2025)
Improving Data Efficiency via Curating LLM-Driven Rating Systems
by: Pang, Jinlong, et al.
Published: (2024)
by: Pang, Jinlong, et al.
Published: (2024)
Label Smoothing Improves Gradient Ascent in LLM Unlearning
by: Pang, Zirui, et al.
Published: (2025)
by: Pang, Zirui, et al.
Published: (2025)
LLM Unlearning via Loss Adjustment with Only Forget Data
by: Wang, Yaxuan, et al.
Published: (2024)
by: Wang, Yaxuan, et al.
Published: (2024)
SFT-then-RL Outperforms Mixed-Policy Methods for LLM Reasoning
by: Limozin, Alexis, et al.
Published: (2026)
by: Limozin, Alexis, et al.
Published: (2026)
TMS: Trajectory-Mixed Supervision for Reward-Free, On-Policy SFT
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2026)
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2026)
Modeling Low-Resource Health Coaching Dialogues via Neuro-Symbolic Goal Summarization and Text-Units-Text Generation
by: Zhou, Yue, et al.
Published: (2024)
by: Zhou, Yue, et al.
Published: (2024)
Empirically Determining Binge/Purge Frequency Thresholds for Differentiating Anorexia Nervosa‐Restricting Subtype vs. Binge–Purge Subtype
by: Sophie R. Abber, et al.
Published: (2025)
by: Sophie R. Abber, et al.
Published: (2025)
Purging the Gray Zone: Latent-Geometric Denoising for Precise Knowledge Boundary Awareness
by: An, Hao, et al.
Published: (2026)
by: An, Hao, et al.
Published: (2026)
Recognition through Reasoning: Reinforcing Image Geo-localization with Large Vision-Language Models
by: Li, Ling, et al.
Published: (2025)
by: Li, Ling, et al.
Published: (2025)
C sequential optimization numbers
by: Hui, Zile
Published: (2024)
by: Hui, Zile
Published: (2024)
Implicit Reward as the Bridge: A Unified View of SFT and DPO Connections
by: Wang, Bo, et al.
Published: (2025)
by: Wang, Bo, et al.
Published: (2025)
PepThink-R1: LLM for Interpretable Cyclic Peptide Optimization with CoT SFT and Reinforcement Learning
by: Wang, Ruheng, et al.
Published: (2025)
by: Wang, Ruheng, et al.
Published: (2025)
Distinct Patterns of Dynamical Regulation in Passive Sensor Data Following Binge and Purge Behaviors
by: Jonathan E. Butner, et al.
Published: (2025)
by: Jonathan E. Butner, et al.
Published: (2025)
The Medical Complications of Purging Behaviours Associated With Eating Disorders
by: Dennis Gibson, et al.
Published: (2025)
by: Dennis Gibson, et al.
Published: (2025)
Symbolic Momentum Conservation and Curvature Entanglement in a Recursive Universe: A Force-Based Reality Framework (SFT-FBRF-8) Eighth Research Paper of the SFT-FBRF Series
by: JANAKARAJ, SIVARAM
Published: (2025)
by: JANAKARAJ, SIVARAM
Published: (2025)
AgenticMath: Enhancing LLM Reasoning via Agentic-based Math Data Generation
by: Liu, Xianyang, et al.
Published: (2025)
by: Liu, Xianyang, et al.
Published: (2025)
Harnessing Collective Structure Knowledge in Data Augmentation for Graph Neural Networks
by: Ma, Rongrong, et al.
Published: (2024)
by: Ma, Rongrong, et al.
Published: (2024)
Getting More Juice Out of the SFT Data: Reward Learning from Human Demonstration Improves SFT for LLM Alignment
by: Li, Jiaxiang, et al.
Published: (2024)
by: Li, Jiaxiang, et al.
Published: (2024)
Learning to Adapt SFT Data for Better Reasoning Generalization
by: Sun, Lisong, et al.
Published: (2026)
by: Sun, Lisong, et al.
Published: (2026)
Parametric Operator Inference to Simulate the Purging Process in Semiconductor Manufacturing
by: Kang, Seunghyon, et al.
Published: (2025)
by: Kang, Seunghyon, et al.
Published: (2025)
A landscape of contact manifolds via rational SFT
by: Moreno, Agustin, et al.
Published: (2020)
by: Moreno, Agustin, et al.
Published: (2020)
mSFT: Addressing Dataset Mixtures Overfitting Heterogeneously in Multi-task SFT
by: Koh, Woosung, et al.
Published: (2026)
by: Koh, Woosung, et al.
Published: (2026)
Small-Margin Preferences Still Matter-If You Train Them Right
by: Pang, Jinlong, et al.
Published: (2026)
by: Pang, Jinlong, et al.
Published: (2026)
RL makes MLLMs see better than SFT
by: Song, Junha, et al.
Published: (2025)
by: Song, Junha, et al.
Published: (2025)
KPTUltra : Dual‐Enhanced Knowledgeable Prompt Tuning for Few‐Shot Text Classification in Low‐Resource Scenarios
by: Wenlong Zha, et al.
Published: (2026)
by: Wenlong Zha, et al.
Published: (2026)
A High-Quality and Low-Complexity Streamable Neural Speech Codec with Knowledge Distillation
by: Zhang, En-Wei, et al.
Published: (2025)
by: Zhang, En-Wei, et al.
Published: (2025)
OFFSIDE: Benchmarking Unlearning Misinformation in Multimodal Large Language Models
by: Zheng, Hao, et al.
Published: (2025)
by: Zheng, Hao, et al.
Published: (2025)
Evaluating LLM-Contaminated Crowdsourcing Data Without Ground Truth
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
Cluster Purge Loss: Structuring Transformer Embeddings for Equivalent Mutants Detection
by: Danilov, Adelaide, et al.
Published: (2025)
by: Danilov, Adelaide, et al.
Published: (2025)
Suicidal Ideation in Adult Women: The Unique Roles of Binging, Purging, and Restricting
by: Holly K. Spinner, et al.
Published: (2025)
by: Holly K. Spinner, et al.
Published: (2025)
Similar Items
-
Token Cleaning: Fine-Grained Data Selection for LLM Supervised Fine-Tuning
by: Pang, Jinlong, et al.
Published: (2025) -
ENTP: Encoder-only Next Token Prediction
by: Ewer, Ethan, et al.
Published: (2024) -
SelectMix: Enhancing Label Noise Robustness through Targeted Sample Mixing
by: Liu, Qiuhao, et al.
Published: (2025) -
LM-mixup: Text Data Augmentation via Language Model based Mixup
by: Deng, Zhijie, et al.
Published: (2025) -
LogPurge: Log Data Purification for Anomaly Detection via Rule-Enhanced Filtering
by: Zhang, Shenglin, et al.
Published: (2025)