Disentangling the Roles of Representation and Selection in Data Pruning
Fuente:
arXiv
Salvato in:
| Autori principali: | Du, Yupei, Song, Yingjin, Wong, Hugh Mee, Ignatev, Daniil, Gatt, Albert, Nguyen, Dong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FTFT: Efficient and Robust Fine-Tuning by Transferring Training Dynamics
di: Du, Yupei, et al.
Pubblicazione: (2023)
di: Du, Yupei, et al.
Pubblicazione: (2023)
DeMeVa at LeWiDi-2025: Modeling Perspectives with In-Context Learning and Label Distribution Learning
di: Ignatev, Daniil, et al.
Pubblicazione: (2025)
di: Ignatev, Daniil, et al.
Pubblicazione: (2025)
Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences?
di: Song, Yingjin, et al.
Pubblicazione: (2025)
di: Song, Yingjin, et al.
Pubblicazione: (2025)
VAQUUM: Are Vague Quantifiers Grounded in Visual Data?
di: Wong, Hugh Mee, et al.
Pubblicazione: (2025)
di: Wong, Hugh Mee, et al.
Pubblicazione: (2025)
When Models Decide and When They Bind: A Two-Stage Computation for Multiple-Choice Question-Answering
di: Wong, Hugh Mee, et al.
Pubblicazione: (2026)
di: Wong, Hugh Mee, et al.
Pubblicazione: (2026)
Don't Learn, Ground: A Case for Natural Language Inference with Visual Grounding
di: Ignatev, Daniil, et al.
Pubblicazione: (2025)
di: Ignatev, Daniil, et al.
Pubblicazione: (2025)
From Image Captioning to Visual Storytelling
di: Passadakis, Admitos, et al.
Pubblicazione: (2025)
di: Passadakis, Admitos, et al.
Pubblicazione: (2025)
Context-aware Visual Storytelling with Visual Prefix Tuning and Contrastive Learning
di: Song, Yingjin, et al.
Pubblicazione: (2024)
di: Song, Yingjin, et al.
Pubblicazione: (2024)
Structured Pruning for Diverse Best-of-N Reasoning Optimization
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2025)
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2025)
Demystifying When Pruning Works via Representation Hierarchies
di: He, Shwai, et al.
Pubblicazione: (2026)
di: He, Shwai, et al.
Pubblicazione: (2026)
Disentangling Language Roles in Multilingual LLM Task Execution
di: Zhan, Qishi, et al.
Pubblicazione: (2026)
di: Zhan, Qishi, et al.
Pubblicazione: (2026)
InnerThoughts: Disentangling Representations and Predictions in Large Language Models
di: Chételat, Didier, et al.
Pubblicazione: (2025)
di: Chételat, Didier, et al.
Pubblicazione: (2025)
Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference
di: Taniguchi, Rei, et al.
Pubblicazione: (2026)
di: Taniguchi, Rei, et al.
Pubblicazione: (2026)
RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations
di: Huang, Jing, et al.
Pubblicazione: (2024)
di: Huang, Jing, et al.
Pubblicazione: (2024)
You Do Not Fully Utilize Transformer's Representation Capacity
di: Gerasimov, Gleb, et al.
Pubblicazione: (2025)
di: Gerasimov, Gleb, et al.
Pubblicazione: (2025)
Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI
di: Man, Hieu, et al.
Pubblicazione: (2026)
di: Man, Hieu, et al.
Pubblicazione: (2026)
Dissecting Language Models: Machine Unlearning via Selective Pruning
di: Pochinkov, Nicholas, et al.
Pubblicazione: (2024)
di: Pochinkov, Nicholas, et al.
Pubblicazione: (2024)
Bypass Back-propagation: Optimization-based Structural Pruning for Large Language Models via Policy Gradient
di: Gao, Yuan, et al.
Pubblicazione: (2024)
di: Gao, Yuan, et al.
Pubblicazione: (2024)
Reason to Rote: Rethinking Memorization in Reasoning
di: Du, Yupei, et al.
Pubblicazione: (2025)
di: Du, Yupei, et al.
Pubblicazione: (2025)
Disentangled Representation Learning with Large Language Models for Text-Attributed Graphs
di: Qin, Yijian, et al.
Pubblicazione: (2023)
di: Qin, Yijian, et al.
Pubblicazione: (2023)
Analyze Feature Flow to Enhance Interpretation and Steering in Language Models
di: Laptev, Daniil, et al.
Pubblicazione: (2025)
di: Laptev, Daniil, et al.
Pubblicazione: (2025)
Disentangling Logic: The Role of Context in Large Language Model Reasoning Capabilities
di: Hua, Wenyue, et al.
Pubblicazione: (2024)
di: Hua, Wenyue, et al.
Pubblicazione: (2024)
Kronecker Factorization Improves Efficiency and Interpretability of Sparse Autoencoders
di: Kurochkin, Vadim, et al.
Pubblicazione: (2025)
di: Kurochkin, Vadim, et al.
Pubblicazione: (2025)
Quantifying Modality Contributions via Disentangling Multimodal Representations
di: Amit, Padegal, et al.
Pubblicazione: (2025)
di: Amit, Padegal, et al.
Pubblicazione: (2025)
Domain-Specific Pruning of Large Mixture-of-Experts Models with Few-shot Demonstrations
di: Dong, Zican, et al.
Pubblicazione: (2025)
di: Dong, Zican, et al.
Pubblicazione: (2025)
Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning
di: Bi, Jiaxi, et al.
Pubblicazione: (2026)
di: Bi, Jiaxi, et al.
Pubblicazione: (2026)
Self-Data Distillation for Recovering Quality in Pruned Large Language Models
di: Thangarasa, Vithursan, et al.
Pubblicazione: (2024)
di: Thangarasa, Vithursan, et al.
Pubblicazione: (2024)
Language Model-Driven Data Pruning Enables Efficient Active Learning
di: Azeemi, Abdul Hameed, et al.
Pubblicazione: (2024)
di: Azeemi, Abdul Hameed, et al.
Pubblicazione: (2024)
Perplexed by Perplexity: Perplexity-Based Data Pruning With Small Reference Models
di: Ankner, Zachary, et al.
Pubblicazione: (2024)
di: Ankner, Zachary, et al.
Pubblicazione: (2024)
Frustratingly Easy Task-aware Pruning for Large Language Models
di: Tian, Yuanhe, et al.
Pubblicazione: (2025)
di: Tian, Yuanhe, et al.
Pubblicazione: (2025)
Comparative Analysis of Demonstration Selection Algorithms for LLM In-Context Learning
di: Shu, Dong, et al.
Pubblicazione: (2024)
di: Shu, Dong, et al.
Pubblicazione: (2024)
UFO-RL: Uncertainty-Focused Optimization for Efficient Reinforcement Learning Data Selection
di: Zhao, Yang, et al.
Pubblicazione: (2025)
di: Zhao, Yang, et al.
Pubblicazione: (2025)
Pruning Weights but Not Truth: Safeguarding Truthfulness While Pruning LLMs
di: Fu, Yao, et al.
Pubblicazione: (2025)
di: Fu, Yao, et al.
Pubblicazione: (2025)
Is Random Attention Sufficient for Sequence Modeling? Disentangling Trainable Components in the Transformer
di: Dong, Yihe, et al.
Pubblicazione: (2025)
di: Dong, Yihe, et al.
Pubblicazione: (2025)
STUN: Structured-Then-Unstructured Pruning for Scalable MoE Pruning
di: Lee, Jaeseong, et al.
Pubblicazione: (2024)
di: Lee, Jaeseong, et al.
Pubblicazione: (2024)
Greedy Information Projection for LLM Data Selection
di: Dong, Victor Ye, et al.
Pubblicazione: (2026)
di: Dong, Victor Ye, et al.
Pubblicazione: (2026)
Transformers with Selective Access to Early Representations
di: Gunasekaran, Skye, et al.
Pubblicazione: (2026)
di: Gunasekaran, Skye, et al.
Pubblicazione: (2026)
Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning
di: Schaffelder, Max, et al.
Pubblicazione: (2025)
di: Schaffelder, Max, et al.
Pubblicazione: (2025)
IDEA Prune: An Integrated Enlarge-and-Prune Pipeline in Generative Language Model Pretraining
di: Li, Yixiao, et al.
Pubblicazione: (2025)
di: Li, Yixiao, et al.
Pubblicazione: (2025)
Interpretable Discriminative Text Representations via Agreement and Label Disentanglement
di: Wang, Tong, et al.
Pubblicazione: (2026)
di: Wang, Tong, et al.
Pubblicazione: (2026)
Documenti analoghi
-
FTFT: Efficient and Robust Fine-Tuning by Transferring Training Dynamics
di: Du, Yupei, et al.
Pubblicazione: (2023) -
DeMeVa at LeWiDi-2025: Modeling Perspectives with In-Context Learning and Label Distribution Learning
di: Ignatev, Daniil, et al.
Pubblicazione: (2025) -
Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences?
di: Song, Yingjin, et al.
Pubblicazione: (2025) -
VAQUUM: Are Vague Quantifiers Grounded in Visual Data?
di: Wong, Hugh Mee, et al.
Pubblicazione: (2025) -
When Models Decide and When They Bind: A Two-Stage Computation for Multiple-Choice Question-Answering
di: Wong, Hugh Mee, et al.
Pubblicazione: (2026)