SEED: Targeted Data Selection by Weighted Independent Set
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yuan, Guo, Lifeng, Pan, Junwen, Zheng, Wenzhao, Zhou, Wen, Cheng, Kuan, Keutzer, Kurt, Zhang, Shanghang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FactorLLM: Factorizing Knowledge via Mixture of Experts for Large Language Models
by: Zhao, Zhongyu, et al.
Published: (2024)
by: Zhao, Zhongyu, et al.
Published: (2024)
ODE$_t$(ODE$_l$): Shortcutting the Time and the Length in Diffusion and Flow Models for Faster Sampling
by: Gudovskiy, Denis, et al.
Published: (2025)
by: Gudovskiy, Denis, et al.
Published: (2025)
PixelGaussian: Generalizable 3D Gaussian Reconstruction from Arbitrary Views
by: Fei, Xin, et al.
Published: (2024)
by: Fei, Xin, et al.
Published: (2024)
Driv3R: Learning Dense 4D Reconstruction for Autonomous Driving
by: Fei, Xin, et al.
Published: (2024)
by: Fei, Xin, et al.
Published: (2024)
One Model is All You Need: ByT5-Sanskrit, a Unified Model for Sanskrit NLP Tasks
by: Nehrdich, Sebastian, et al.
Published: (2024)
by: Nehrdich, Sebastian, et al.
Published: (2024)
Simple and Effective Input Reformulations for Translation
by: Yu, Brian, et al.
Published: (2023)
by: Yu, Brian, et al.
Published: (2023)
Efficient RLVR Training via Weighted Mutual Information Data Selection
by: Zhou, Xinyu, et al.
Published: (2026)
by: Zhou, Xinyu, et al.
Published: (2026)
FGMP: Fine-Grained Mixed-Precision Weight and Activation Quantization for Hardware-Accelerated LLM Inference
by: Hooper, Coleman, et al.
Published: (2025)
by: Hooper, Coleman, et al.
Published: (2025)
GPD-1: Generative Pre-training for Driving
by: Xie, Zixun, et al.
Published: (2024)
by: Xie, Zixun, et al.
Published: (2024)
SEED: Domain-Specific Data Curation With Large Language Models
by: Chen, Zui, et al.
Published: (2023)
by: Chen, Zui, et al.
Published: (2023)
Learning to Select and Rank from Choice-Based Feedback: A Simple Nested Approach
by: Yang, Junwen, et al.
Published: (2023)
by: Yang, Junwen, et al.
Published: (2023)
Split-Ensemble: Efficient OOD-aware Ensemble via Task and Model Splitting
by: Chen, Anthony, et al.
Published: (2023)
by: Chen, Anthony, et al.
Published: (2023)
Explore and Establish Synergistic Effects Between Weight Pruning and Coreset Selection in Neural Network Training
by: Wan, Weilin, et al.
Published: (2025)
by: Wan, Weilin, et al.
Published: (2025)
Segment Any Motion in Videos
by: Huang, Nan, et al.
Published: (2025)
by: Huang, Nan, et al.
Published: (2025)
Path Choice Matters for Clear Attribution in Path Methods
by: Zhang, Borui, et al.
Published: (2024)
by: Zhang, Borui, et al.
Published: (2024)
Preventing Local Pitfalls in Vector Quantization via Optimal Transport
by: Zhang, Borui, et al.
Published: (2024)
by: Zhang, Borui, et al.
Published: (2024)
Uncovering Critical Sets of Deep Neural Networks via Sample-Independent Critical Lifting
by: Zhang, Leyang, et al.
Published: (2025)
by: Zhang, Leyang, et al.
Published: (2025)
Dobi-SVD: Differentiable SVD for LLM Compression and Some New Perspectives
by: Wang, Qinsi, et al.
Published: (2025)
by: Wang, Qinsi, et al.
Published: (2025)
R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation
by: Ljungbergh, William, et al.
Published: (2025)
by: Ljungbergh, William, et al.
Published: (2025)
On Representation Redundancy in Large-Scale Instruction Tuning Data Selection
by: Shu, Youwei, et al.
Published: (2026)
by: Shu, Youwei, et al.
Published: (2026)
SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference
by: Zhang, Yuan, et al.
Published: (2024)
by: Zhang, Yuan, et al.
Published: (2024)
Let the Target Select for Itself: Data Selection via Target-Aligned Paths
by: Yang, Huitao, et al.
Published: (2026)
by: Yang, Huitao, et al.
Published: (2026)
Characterizing Prompt Compression Methods for Long Context Inference
by: Jha, Siddharth, et al.
Published: (2024)
by: Jha, Siddharth, et al.
Published: (2024)
Computational Budget Should Be Considered in Data Selection
by: Wan, Weilin, et al.
Published: (2025)
by: Wan, Weilin, et al.
Published: (2025)
GeoDrive: 3D Geometry-Informed Driving World Model with Precise Action Control
by: Chen, Anthony, et al.
Published: (2025)
by: Chen, Anthony, et al.
Published: (2025)
EPIM: Efficient Processing-In-Memory Accelerators based on Epitome
by: Wang, Chenyu, et al.
Published: (2023)
by: Wang, Chenyu, et al.
Published: (2023)
D3: Diversity, Difficulty, and Dependability-Aware Data Selection for Sample-Efficient LLM Instruction Tuning
by: Zhang, Jia, et al.
Published: (2025)
by: Zhang, Jia, et al.
Published: (2025)
GaussianAD: Gaussian-Centric End-to-End Autonomous Driving
by: Zheng, Wenzhao, et al.
Published: (2024)
by: Zheng, Wenzhao, et al.
Published: (2024)
Graph as Point Set
by: Wang, Xiyuan, et al.
Published: (2024)
by: Wang, Xiyuan, et al.
Published: (2024)
SFTok: Bridging the Performance Gap in Discrete Tokenizers
by: Rao, Qihang, et al.
Published: (2025)
by: Rao, Qihang, et al.
Published: (2025)
Quantize-then-Rectify: Efficient VQ-VAE Training
by: Zhang, Borui, et al.
Published: (2025)
by: Zhang, Borui, et al.
Published: (2025)
Fast Shapley Value Estimation: A Unified Approach
by: Zhang, Borui, et al.
Published: (2023)
by: Zhang, Borui, et al.
Published: (2023)
Stag-1: Towards Realistic 4D Driving Simulation with Video Generation Model
by: Wang, Lening, et al.
Published: (2024)
by: Wang, Lening, et al.
Published: (2024)
QFT: Quantized Full-parameter Tuning of LLMs with Affordable Resources
by: Li, Zhikai, et al.
Published: (2023)
by: Li, Zhikai, et al.
Published: (2023)
UniDrive: Towards Universal Driving Perception Across Camera Configurations
by: Li, Ye, et al.
Published: (2024)
by: Li, Ye, et al.
Published: (2024)
Instruct Large Language Models to Drive like Humans
by: Zhang, Ruijun, et al.
Published: (2024)
by: Zhang, Ruijun, et al.
Published: (2024)
CDLM: Consistency Diffusion Language Models For Faster Sampling
by: Kim, Minseo, et al.
Published: (2025)
by: Kim, Minseo, et al.
Published: (2025)
One-Shot Federated Clustering of Non-Independent Completely Distributed Data
by: Zhang, Yiqun, et al.
Published: (2026)
by: Zhang, Yiqun, et al.
Published: (2026)
Sharpness-diversity tradeoff: improving flat ensembles with SharpBalance
by: Lu, Haiquan, et al.
Published: (2024)
by: Lu, Haiquan, et al.
Published: (2024)
SciML Agents: Write the Solver, Not the Solution
by: Gaonkar, Saarth, et al.
Published: (2025)
by: Gaonkar, Saarth, et al.
Published: (2025)
Similar Items
-
FactorLLM: Factorizing Knowledge via Mixture of Experts for Large Language Models
by: Zhao, Zhongyu, et al.
Published: (2024) -
ODE$_t$(ODE$_l$): Shortcutting the Time and the Length in Diffusion and Flow Models for Faster Sampling
by: Gudovskiy, Denis, et al.
Published: (2025) -
PixelGaussian: Generalizable 3D Gaussian Reconstruction from Arbitrary Views
by: Fei, Xin, et al.
Published: (2024) -
Driv3R: Learning Dense 4D Reconstruction for Autonomous Driving
by: Fei, Xin, et al.
Published: (2024) -
One Model is All You Need: ByT5-Sanskrit, a Unified Model for Sanskrit NLP Tasks
by: Nehrdich, Sebastian, et al.
Published: (2024)