Cut Less, Fold More: Model Compression through the Lens of Projection Geometry
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Saukh, Olga, Wang, Dong, Šikić, Haris, Cheng, Yun, Thiele, Lothar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Forget the Data and Fine-Tuning! Just Fold the Network to Compress
von: Wang, Dong, et al.
Veröffentlicht: (2025)
von: Wang, Dong, et al.
Veröffentlicht: (2025)
GRAIL: Post-hoc Compensation by Linear Reconstruction for Compressed Networks
von: Tang, Wenwu, et al.
Veröffentlicht: (2026)
von: Tang, Wenwu, et al.
Veröffentlicht: (2026)
PCDCNet: A Surrogate Model for Air Quality Forecasting with Physical-Chemical Dynamics and Constraints
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
Physics-Guided Inductive Spatiotemporal Kriging for PM2.5 with Satellite Gradient Constraints
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
Subspace-Configurable Networks
von: Wang, Dong, et al.
Veröffentlicht: (2023)
von: Wang, Dong, et al.
Veröffentlicht: (2023)
From LLMs to Edge: Parameter-Efficient Fine-Tuning on Edge Devices
von: Slamanig, Georg, et al.
Veröffentlicht: (2025)
von: Slamanig, Georg, et al.
Veröffentlicht: (2025)
Say Less, Mean More: Leveraging Pragmatics in Retrieval-Augmented Generation
von: Riaz, Haris, et al.
Veröffentlicht: (2025)
von: Riaz, Haris, et al.
Veröffentlicht: (2025)
Less is More: Undertraining Experts Improves Model Upcycling
von: Horoi, Stefan, et al.
Veröffentlicht: (2025)
von: Horoi, Stefan, et al.
Veröffentlicht: (2025)
Transformer Multivariate Forecasting: Less is More?
von: Xu, Jingjing, et al.
Veröffentlicht: (2023)
von: Xu, Jingjing, et al.
Veröffentlicht: (2023)
Less is More: on the Over-Globalizing Problem in Graph Transformers
von: Xing, Yujie, et al.
Veröffentlicht: (2024)
von: Xing, Yujie, et al.
Veröffentlicht: (2024)
Less is More: Recursive Reasoning with Tiny Networks
von: Jolicoeur-Martineau, Alexia
Veröffentlicht: (2025)
von: Jolicoeur-Martineau, Alexia
Veröffentlicht: (2025)
Less is More: Unlocking Specialization of Time Series Foundation Models via Structured Pruning
von: Zhao, Lifan, et al.
Veröffentlicht: (2025)
von: Zhao, Lifan, et al.
Veröffentlicht: (2025)
Forget Less, Generalize More: Unifying Temporal and Structural Adaptation for Dynamic Graphs
von: Chang, Qian, et al.
Veröffentlicht: (2026)
von: Chang, Qian, et al.
Veröffentlicht: (2026)
When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search
von: Robertson, John T., et al.
Veröffentlicht: (2026)
von: Robertson, John T., et al.
Veröffentlicht: (2026)
Less is More: Pseudo-Label Filtering for Continual Test-Time Adaptation
von: Tan, Jiayao, et al.
Veröffentlicht: (2024)
von: Tan, Jiayao, et al.
Veröffentlicht: (2024)
Draft Less, Retrieve More: Hybrid Tree Construction for Speculative Decoding
von: Shen, Yuhao, et al.
Veröffentlicht: (2026)
von: Shen, Yuhao, et al.
Veröffentlicht: (2026)
LIMR: Less is More for RL Scaling
von: Li, Xuefeng, et al.
Veröffentlicht: (2025)
von: Li, Xuefeng, et al.
Veröffentlicht: (2025)
Less is More for Improving Automatic Evaluation of Factual Consistency
von: Wang, Tong, et al.
Veröffentlicht: (2024)
von: Wang, Tong, et al.
Veröffentlicht: (2024)
Less is More: Local Intrinsic Dimensions of Contextual Language Models
von: Ruppik, Benjamin Matthias, et al.
Veröffentlicht: (2025)
von: Ruppik, Benjamin Matthias, et al.
Veröffentlicht: (2025)
When Less is More: 8-bit Quantization Improves Continual Learning in Large Language Models
von: Zhang, Michael S., et al.
Veröffentlicht: (2025)
von: Zhang, Michael S., et al.
Veröffentlicht: (2025)
Less Is More -- On the Importance of Sparsification for Transformers and Graph Neural Networks for TSP
von: Lischka, Attila, et al.
Veröffentlicht: (2024)
von: Lischka, Attila, et al.
Veröffentlicht: (2024)
Get More with LESS: Synthesizing Recurrence with KV Cache Compression for Efficient LLM Inference
von: Dong, Harry, et al.
Veröffentlicht: (2024)
von: Dong, Harry, et al.
Veröffentlicht: (2024)
When More is Less: Understanding Chain-of-Thought Length in LLMs
von: Wu, Yuyang, et al.
Veröffentlicht: (2025)
von: Wu, Yuyang, et al.
Veröffentlicht: (2025)
No More, No Less: Task Alignment in Terminal Agents
von: Mavali, Sina, et al.
Veröffentlicht: (2026)
von: Mavali, Sina, et al.
Veröffentlicht: (2026)
Supernova: Achieving More with Less in Transformer Architectures
von: Tanase, Andrei-Valentin, et al.
Veröffentlicht: (2025)
von: Tanase, Andrei-Valentin, et al.
Veröffentlicht: (2025)
Heterogeneous Graph Structure Learning through the Lens of Data-generating Processes
von: Jiang, Keyue, et al.
Veröffentlicht: (2025)
von: Jiang, Keyue, et al.
Veröffentlicht: (2025)
Train Less, Learn More: Adaptive Efficient Rollout Optimization for Group-Based Reinforcement Learning
von: Zhang, Zhi, et al.
Veröffentlicht: (2026)
von: Zhang, Zhi, et al.
Veröffentlicht: (2026)
Is Monotonic Sampling Necessary in Diffusion Models?
von: Khan, Muhammad Haris
Veröffentlicht: (2026)
von: Khan, Muhammad Haris
Veröffentlicht: (2026)
Less is More: Multimodal Region Representation via Pairwise Inter-view Learning
von: Namgung, Min, et al.
Veröffentlicht: (2025)
von: Namgung, Min, et al.
Veröffentlicht: (2025)
Reflect-then-Plan: Offline Model-Based Planning through a Doubly Bayesian Lens
von: Jeong, Jihwan, et al.
Veröffentlicht: (2025)
von: Jeong, Jihwan, et al.
Veröffentlicht: (2025)
Analyzing Memorization in Large Language Models through the Lens of Model Attribution
von: Menta, Tarun Ram, et al.
Veröffentlicht: (2025)
von: Menta, Tarun Ram, et al.
Veröffentlicht: (2025)
Less Is More, but Where? Dynamic Token Compression via LLM-Guided Keyframe Prior
von: Li, Yulin, et al.
Veröffentlicht: (2025)
von: Li, Yulin, et al.
Veröffentlicht: (2025)
Palu: Compressing KV-Cache with Low-Rank Projection
von: Chang, Chi-Chih, et al.
Veröffentlicht: (2024)
von: Chang, Chi-Chih, et al.
Veröffentlicht: (2024)
Provable Benefit of Cutout and CutMix for Feature Learning
von: Oh, Junsoo, et al.
Veröffentlicht: (2024)
von: Oh, Junsoo, et al.
Veröffentlicht: (2024)
Learning More with Less: A Dynamic Dual-Level Down-Sampling Framework for Efficient Policy Optimization
von: Wang, Chao, et al.
Veröffentlicht: (2025)
von: Wang, Chao, et al.
Veröffentlicht: (2025)
Quotient Geometry, Effective Curvature, and Implicit Bias in Simple Shallow Neural Networks
von: Dong, Hang-Cheng, et al.
Veröffentlicht: (2026)
von: Dong, Hang-Cheng, et al.
Veröffentlicht: (2026)
Less is More: Improving LLM Alignment via Preference Data Selection
von: Deng, Xun, et al.
Veröffentlicht: (2025)
von: Deng, Xun, et al.
Veröffentlicht: (2025)
Backward Lens: Projecting Language Model Gradients into the Vocabulary Space
von: Katz, Shahar, et al.
Veröffentlicht: (2024)
von: Katz, Shahar, et al.
Veröffentlicht: (2024)
Less is More: Denoising Knowledge Graphs For Retrieval Augmented Generation
von: Zheng, Yilun, et al.
Veröffentlicht: (2025)
von: Zheng, Yilun, et al.
Veröffentlicht: (2025)
Learn More, Forget Less: A Gradient-Aware Data Selection Approach for LLM
von: Liu, Yibai, et al.
Veröffentlicht: (2025)
von: Liu, Yibai, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Forget the Data and Fine-Tuning! Just Fold the Network to Compress
von: Wang, Dong, et al.
Veröffentlicht: (2025) -
GRAIL: Post-hoc Compensation by Linear Reconstruction for Compressed Networks
von: Tang, Wenwu, et al.
Veröffentlicht: (2026) -
PCDCNet: A Surrogate Model for Air Quality Forecasting with Physical-Chemical Dynamics and Constraints
von: Wang, Shuo, et al.
Veröffentlicht: (2025) -
Physics-Guided Inductive Spatiotemporal Kriging for PM2.5 with Satellite Gradient Constraints
von: Wang, Shuo, et al.
Veröffentlicht: (2025) -
Subspace-Configurable Networks
von: Wang, Dong, et al.
Veröffentlicht: (2023)