Shall Your Data Strategy Work? Perform a Swift Study
Fuente:
arXiv
Saved in:
| Main Authors: | Peng, Minlong, Yang, Jingyi, He, Zhongjun, Wu, Hua |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Not All Parameters Are Created Equal: Smart Isolation Boosts Fine-Tuning Performance
by: Wang, Yao, et al.
Published: (2025)
by: Wang, Yao, et al.
Published: (2025)
Towards Boosting Many-to-Many Multilingual Machine Translation with Large Language Models
by: Gao, Pengzhi, et al.
Published: (2024)
by: Gao, Pengzhi, et al.
Published: (2024)
AlignX: Advancing Multilingual Large Language Models with Multilingual Representation Alignment
by: Bu, Mengyu, et al.
Published: (2025)
by: Bu, Mengyu, et al.
Published: (2025)
DeCoRL: Decoupling Reasoning Chains via Parallel Sub-Step Generation and Cascaded Reinforcement for Interpretable and Scalable RLHF
by: Gao, Ziyuan, et al.
Published: (2025)
by: Gao, Ziyuan, et al.
Published: (2025)
Show Your Work with Confidence: Confidence Bands for Tuning Curves
by: Lourie, Nicholas, et al.
Published: (2023)
by: Lourie, Nicholas, et al.
Published: (2023)
Swift Cross-Dataset Pruning: Enhancing Fine-Tuning Efficiency in Natural Language Understanding
by: Nguyen, Binh-Nguyen, et al.
Published: (2025)
by: Nguyen, Binh-Nguyen, et al.
Published: (2025)
You Shall Know a Tool by the Traces it Leaves: The Predictability of Sentiment Analysis Tools
by: Baumartz, Daniel, et al.
Published: (2024)
by: Baumartz, Daniel, et al.
Published: (2024)
Shall We Team Up: Exploring Spontaneous Cooperation of Competing LLM Agents
by: Wu, Zengqing, et al.
Published: (2024)
by: Wu, Zengqing, et al.
Published: (2024)
Mask and You Shall Receive: Optimizing Masked Language Modeling For Pretraining BabyLMs
by: Edman, Lukas, et al.
Published: (2025)
by: Edman, Lukas, et al.
Published: (2025)
DocTER: Evaluating Document-based Knowledge Editing
by: Wu, Suhang, et al.
Published: (2023)
by: Wu, Suhang, et al.
Published: (2023)
TokenSwift: Lossless Acceleration of Ultra Long Sequence Generation
by: Wu, Tong, et al.
Published: (2025)
by: Wu, Tong, et al.
Published: (2025)
One2set + Large Language Model: Best Partners for Keyphrase Generation
by: Shao, Liangying, et al.
Published: (2024)
by: Shao, Liangying, et al.
Published: (2024)
How Much of Your Data Can Suck? Thresholds for Domain Performance and Emergent Misalignment in LLMs
by: Ouyang, Jian, et al.
Published: (2025)
by: Ouyang, Jian, et al.
Published: (2025)
Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models
by: Xue, Chao, et al.
Published: (2026)
by: Xue, Chao, et al.
Published: (2026)
Your Vision-Language Model Itself Is a Strong Filter: Towards High-Quality Instruction Tuning with Data Selection
by: Chen, Ruibo, et al.
Published: (2024)
by: Chen, Ruibo, et al.
Published: (2024)
Benchmarking Chinese Medical LLMs: A Medbench-based Analysis of Performance Gaps and Hierarchical Optimization Strategies
by: Jiang, Luyi, et al.
Published: (2025)
by: Jiang, Luyi, et al.
Published: (2025)
When or What? Understanding Consumer Engagement on Digital Platforms
by: Wu, Jingyi, et al.
Published: (2025)
by: Wu, Jingyi, et al.
Published: (2025)
AgentSwift: Efficient LLM Agent Design via Value-guided Hierarchical Search
by: Li, Yu, et al.
Published: (2025)
by: Li, Yu, et al.
Published: (2025)
TRACE: Discovering Task-Specific Parameter via Adaptation-Aware Probing for Continual Fine-Tuning
by: Han, Xiaosong, et al.
Published: (2026)
by: Han, Xiaosong, et al.
Published: (2026)
Swift-SVD: Theoretical Optimality Meets Practical Efficiency in Low-Rank LLM Compression
by: Qi, Ruoling, et al.
Published: (2026)
by: Qi, Ruoling, et al.
Published: (2026)
HiGMem: A Hierarchical and LLM-Guided Memory System for Long-Term Conversational Agents
by: Cao, Shuqi, et al.
Published: (2026)
by: Cao, Shuqi, et al.
Published: (2026)
Your Co-Workers Matter: Evaluating Collaborative Capabilities of Language Models in Blocks World
by: Wu, Guande, et al.
Published: (2024)
by: Wu, Guande, et al.
Published: (2024)
Parameter Importance is Not Static: Evolving Parameter Isolation for Supervised Fine-Tuning
by: Lin, Zekai, et al.
Published: (2026)
by: Lin, Zekai, et al.
Published: (2026)
What is in Your Safe Data? Identifying Benign Data that Breaks Safety
by: He, Luxi, et al.
Published: (2024)
by: He, Luxi, et al.
Published: (2024)
SwiftKV: Fast Prefill-Optimized Inference with Knowledge-Preserving Model Transformation
by: Qiao, Aurick, et al.
Published: (2024)
by: Qiao, Aurick, et al.
Published: (2024)
Investigating the Impact of Data Selection Strategies on Language Model Performance
by: Gu, Jiayao, et al.
Published: (2025)
by: Gu, Jiayao, et al.
Published: (2025)
LLMs are Single-threaded Reasoners: Demystifying the Working Mechanism of Soft Thinking
by: Wu, Junhong, et al.
Published: (2025)
by: Wu, Junhong, et al.
Published: (2025)
Do Data-based Curricula Work?
by: Surkov, Maxim K., et al.
Published: (2021)
by: Surkov, Maxim K., et al.
Published: (2021)
Talk to Your Slides: High-Efficiency Slide Editing via Language-Driven Structured Data Manipulation
by: Jung, Kyudan, et al.
Published: (2025)
by: Jung, Kyudan, et al.
Published: (2025)
Perturb Your Data: Paraphrase-Guided Training Data Watermarking
by: Shetty, Pranav, et al.
Published: (2025)
by: Shetty, Pranav, et al.
Published: (2025)
On Training Data Influence of GPT Models
by: Chai, Yekun, et al.
Published: (2024)
by: Chai, Yekun, et al.
Published: (2024)
AutoPR: Let's Automate Your Academic Promotion!
by: Chen, Qiguang, et al.
Published: (2025)
by: Chen, Qiguang, et al.
Published: (2025)
What's Wrong with Your Code Generated by Large Language Models? An Extensive Study
by: Dou, Shihan, et al.
Published: (2024)
by: Dou, Shihan, et al.
Published: (2024)
LLaMAX2: Your Translation-Enhanced Model also Performs Well in Reasoning
by: Gao, Changjiang, et al.
Published: (2025)
by: Gao, Changjiang, et al.
Published: (2025)
Reason Only When Needed: Efficient Generative Reward Modeling via Model-Internal Uncertainty
by: Xue, Chao, et al.
Published: (2026)
by: Xue, Chao, et al.
Published: (2026)
SwiftEval: Developing a Language-Specific Benchmark for LLM-generated Code Evaluation
by: Petrukha, Ivan, et al.
Published: (2025)
by: Petrukha, Ivan, et al.
Published: (2025)
$ρ$-$\texttt{EOS}$: Training-free Bidirectional Variable-Length Control for Masked Diffusion LLMs
by: Yang, Jingyi, et al.
Published: (2026)
by: Yang, Jingyi, et al.
Published: (2026)
CIBench: Evaluating Your LLMs with a Code Interpreter Plugin
by: Zhang, Chuyu, et al.
Published: (2024)
by: Zhang, Chuyu, et al.
Published: (2024)
Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen!
by: Zhang, Zhexin, et al.
Published: (2025)
by: Zhang, Zhexin, et al.
Published: (2025)
Self-training Strategies for Sentiment Analysis: An Empirical Study
by: Liu, Haochen, et al.
Published: (2023)
by: Liu, Haochen, et al.
Published: (2023)
Similar Items
-
Not All Parameters Are Created Equal: Smart Isolation Boosts Fine-Tuning Performance
by: Wang, Yao, et al.
Published: (2025) -
Towards Boosting Many-to-Many Multilingual Machine Translation with Large Language Models
by: Gao, Pengzhi, et al.
Published: (2024) -
AlignX: Advancing Multilingual Large Language Models with Multilingual Representation Alignment
by: Bu, Mengyu, et al.
Published: (2025) -
DeCoRL: Decoupling Reasoning Chains via Parallel Sub-Step Generation and Cascaded Reinforcement for Interpretable and Scalable RLHF
by: Gao, Ziyuan, et al.
Published: (2025) -
Show Your Work with Confidence: Confidence Bands for Tuning Curves
by: Lourie, Nicholas, et al.
Published: (2023)