Language Models as Zero-shot Lossless Gradient Compressors: Towards General Neural Parameter Prior Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Hui-Po, Fritz, Mario |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PriorZero: Bridging Language Priors and World Models for Decision Making
von: Xiong, Junyu, et al.
Veröffentlicht: (2026)
von: Xiong, Junyu, et al.
Veröffentlicht: (2026)
Flora: Low-Rank Adapters Are Secretly Gradient Compressors
von: Hao, Yongchang, et al.
Veröffentlicht: (2024)
von: Hao, Yongchang, et al.
Veröffentlicht: (2024)
Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Models
von: Wu, Shutong, et al.
Veröffentlicht: (2025)
von: Wu, Shutong, et al.
Veröffentlicht: (2025)
EAC-MoE: Expert-Selection Aware Compressor for Mixture-of-Experts Large Language Models
von: Chen, Yuanteng, et al.
Veröffentlicht: (2025)
von: Chen, Yuanteng, et al.
Veröffentlicht: (2025)
Lossless Compression: A New Benchmark for Time Series Model Evaluation
von: Wan, Meng, et al.
Veröffentlicht: (2025)
von: Wan, Meng, et al.
Veröffentlicht: (2025)
MC#: Mixture Compressor for Mixture-of-Experts Large Models
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Towards Efficient Few-shot Graph Neural Architecture Search via Partitioning Gradient Contribution
von: Song, Wenhao, et al.
Veröffentlicht: (2025)
von: Song, Wenhao, et al.
Veröffentlicht: (2025)
Zero-shot Multivariate Time Series Forecasting Using Tabular Prior Fitted Networks
von: Jayawardhana, Mayuka, et al.
Veröffentlicht: (2026)
von: Jayawardhana, Mayuka, et al.
Veröffentlicht: (2026)
Low-Resource Crop Classification from Multi-Spectral Time Series Using Lossless Compressors
von: Cheng, Wei, et al.
Veröffentlicht: (2024)
von: Cheng, Wei, et al.
Veröffentlicht: (2024)
VLAD-Grasp: Zero-shot Grasp Detection via Vision-Language Models
von: Kulshrestha, Manav, et al.
Veröffentlicht: (2025)
von: Kulshrestha, Manav, et al.
Veröffentlicht: (2025)
Lookahead: An Inference Acceleration Framework for Large Language Model with Lossless Generation Accuracy
von: Zhao, Yao, et al.
Veröffentlicht: (2023)
von: Zhao, Yao, et al.
Veröffentlicht: (2023)
OTTER: Effortless Label Distribution Adaptation of Zero-shot Models
von: Shin, Changho, et al.
Veröffentlicht: (2024)
von: Shin, Changho, et al.
Veröffentlicht: (2024)
A Foundation Model for Zero-shot Logical Query Reasoning
von: Galkin, Mikhail, et al.
Veröffentlicht: (2024)
von: Galkin, Mikhail, et al.
Veröffentlicht: (2024)
Lossless Vocabulary Reduction for Auto-Regressive Language Models
von: Chijiwa, Daiki, et al.
Veröffentlicht: (2025)
von: Chijiwa, Daiki, et al.
Veröffentlicht: (2025)
ProgFed: Effective, Communication, and Computation Efficient Federated Learning by Progressive Training
von: Wang, Hui-Po, et al.
Veröffentlicht: (2021)
von: Wang, Hui-Po, et al.
Veröffentlicht: (2021)
Reverso: Efficient Time Series Foundation Models for Zero-shot Forecasting
von: Fu, Xinghong, et al.
Veröffentlicht: (2026)
von: Fu, Xinghong, et al.
Veröffentlicht: (2026)
NeuralPrefix: A Zero-shot Sensory Data Imputation Plugin
von: Khamis, Abdelwahed, et al.
Veröffentlicht: (2025)
von: Khamis, Abdelwahed, et al.
Veröffentlicht: (2025)
Small Models, Strong Priors: Architectural Inductive Bias for Parameter-Efficient Neural PDE Solvers
von: Sankaran, Shyam, et al.
Veröffentlicht: (2026)
von: Sankaran, Shyam, et al.
Veröffentlicht: (2026)
One-shot World Models Using a Transformer Trained on a Synthetic Prior
von: Ferreira, Fabio, et al.
Veröffentlicht: (2024)
von: Ferreira, Fabio, et al.
Veröffentlicht: (2024)
Shared Doubt: Zero-shot Cross-Lingual Confidence Estimation for Language Models
von: Kyriakou, Athina, et al.
Veröffentlicht: (2026)
von: Kyriakou, Athina, et al.
Veröffentlicht: (2026)
AnomalyGFM: Graph Foundation Model for Zero/Few-shot Anomaly Detection
von: Qiao, Hezhe, et al.
Veröffentlicht: (2025)
von: Qiao, Hezhe, et al.
Veröffentlicht: (2025)
A Zero-shot and Few-shot Study of Instruction-Finetuned Large Language Models Applied to Clinical and Biomedical Tasks
von: Labrak, Yanis, et al.
Veröffentlicht: (2023)
von: Labrak, Yanis, et al.
Veröffentlicht: (2023)
Towards Few-shot Self-explaining Graph Neural Networks
von: Peng, Jingyu, et al.
Veröffentlicht: (2024)
von: Peng, Jingyu, et al.
Veröffentlicht: (2024)
Zero-shot Concept Bottleneck Models
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2025)
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2025)
Zero-shot Load Forecasting for Integrated Energy Systems: A Large Language Model-based Framework with Multi-task Learning
von: Li, Jiaheng, et al.
Veröffentlicht: (2025)
von: Li, Jiaheng, et al.
Veröffentlicht: (2025)
Inductive Biases for Zero-shot Systematic Generalization in Language-informed Reinforcement Learning
von: Dijujin, Negin Hashemi, et al.
Veröffentlicht: (2025)
von: Dijujin, Negin Hashemi, et al.
Veröffentlicht: (2025)
Large Language Models are Few-shot Multivariate Time Series Classifiers
von: Chen, Yakun, et al.
Veröffentlicht: (2025)
von: Chen, Yakun, et al.
Veröffentlicht: (2025)
FoMo-0D: A Foundation Model for Zero-shot Tabular Outlier Detection
von: Shen, Yuchen, et al.
Veröffentlicht: (2024)
von: Shen, Yuchen, et al.
Veröffentlicht: (2024)
BiTA: Bi-Directional Tuning for Lossless Acceleration in Large Language Models
von: Lin, Feng, et al.
Veröffentlicht: (2024)
von: Lin, Feng, et al.
Veröffentlicht: (2024)
Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2024)
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2024)
Human-like Category Learning by Injecting Ecological Priors from Large Language Models into Neural Networks
von: Jagadish, Akshay K., et al.
Veröffentlicht: (2024)
von: Jagadish, Akshay K., et al.
Veröffentlicht: (2024)
Gradient Inversion Transcript: Leveraging Robust Generative Priors to Reconstruct Training Data from Gradient Leakage
von: Chen, Xinping, et al.
Veröffentlicht: (2025)
von: Chen, Xinping, et al.
Veröffentlicht: (2025)
Reimagining Parameter Space Exploration with Diffusion Models
von: Zhang, Lijun, et al.
Veröffentlicht: (2025)
von: Zhang, Lijun, et al.
Veröffentlicht: (2025)
Cooperative Open-ended Learning Framework for Zero-shot Coordination
von: Li, Yang, et al.
Veröffentlicht: (2023)
von: Li, Yang, et al.
Veröffentlicht: (2023)
STRCMP: Integrating Graph Structural Priors with Language Models for Combinatorial Optimization
von: Li, Xijun, et al.
Veröffentlicht: (2025)
von: Li, Xijun, et al.
Veröffentlicht: (2025)
FineZip : Pushing the Limits of Large Language Models for Practical Lossless Text Compression
von: Mittu, Fazal, et al.
Veröffentlicht: (2024)
von: Mittu, Fazal, et al.
Veröffentlicht: (2024)
Benchmarking Language Modeling for Lossless Compression of Full-Fidelity Audio
von: Long, Phillip, et al.
Veröffentlicht: (2026)
von: Long, Phillip, et al.
Veröffentlicht: (2026)
Masked Generative Priors Improve World Models Sequence Modelling Capabilities
von: Meo, Cristian, et al.
Veröffentlicht: (2024)
von: Meo, Cristian, et al.
Veröffentlicht: (2024)
Towards General Continuous Memory for Vision-Language Models
von: Wu, Wenyi, et al.
Veröffentlicht: (2025)
von: Wu, Wenyi, et al.
Veröffentlicht: (2025)
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization
von: Chaubard, Francois, et al.
Veröffentlicht: (2025)
von: Chaubard, Francois, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PriorZero: Bridging Language Priors and World Models for Decision Making
von: Xiong, Junyu, et al.
Veröffentlicht: (2026) -
Flora: Low-Rank Adapters Are Secretly Gradient Compressors
von: Hao, Yongchang, et al.
Veröffentlicht: (2024) -
Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Models
von: Wu, Shutong, et al.
Veröffentlicht: (2025) -
EAC-MoE: Expert-Selection Aware Compressor for Mixture-of-Experts Large Language Models
von: Chen, Yuanteng, et al.
Veröffentlicht: (2025) -
Lossless Compression: A New Benchmark for Time Series Model Evaluation
von: Wan, Meng, et al.
Veröffentlicht: (2025)