Proxy Compression for Language Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Lin, Li, Xinyu, Liu, Qian, Feng, Xiachong, Kong, Lingpeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Reparameterized Discrete Diffusion Model for Text Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
DynaAct: Large Language Model Reasoning with Dynamic Action Spaces
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)
Reasoning Does Not Necessarily Improve Role-Playing Ability
von: Feng, Xiachong, et al.
Veröffentlicht: (2025)
von: Feng, Xiachong, et al.
Veröffentlicht: (2025)
Self-Infilling Code Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
PromptCoT 2.0: Scaling Prompt Synthesis for Large Language Model Reasoning
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)
PromptCoT: Synthesizing Olympiad-level Problems for Mathematical Reasoning in Large Language Models
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)
Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
Jailbreaking as a Reward Misspecification Problem
von: Xie, Zhihui, et al.
Veröffentlicht: (2024)
von: Xie, Zhihui, et al.
Veröffentlicht: (2024)
Teaching Language Models to Critique via Reinforcement Learning
von: Xie, Zhihui, et al.
Veröffentlicht: (2025)
von: Xie, Zhihui, et al.
Veröffentlicht: (2025)
Scaling Reasoning without Attention
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)
Proxy-RLHF: Decoupling Generation and Alignment in Large Language Model with Proxy
von: Zhu, Yu, et al.
Veröffentlicht: (2024)
von: Zhu, Yu, et al.
Veröffentlicht: (2024)
The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models
von: Qin, Chonghan, et al.
Veröffentlicht: (2026)
von: Qin, Chonghan, et al.
Veröffentlicht: (2026)
DLM-Scope: Mechanistic Interpretability of Diffusion Language Models via Sparse Autoencoders
von: Wang, Xu, et al.
Veröffentlicht: (2026)
von: Wang, Xu, et al.
Veröffentlicht: (2026)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
Multimodal ArXiv: A Dataset for Improving Scientific Comprehension of Large Vision-Language Models
von: Li, Lei, et al.
Veröffentlicht: (2024)
von: Li, Lei, et al.
Veröffentlicht: (2024)
SubgoalXL: Subgoal-based Expert Learning for Theorem Proving
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
Semantic Consistency Regularization with Large Language Models for Semi-supervised Sentiment Analysis
von: Li, Kunrong, et al.
Veröffentlicht: (2025)
von: Li, Kunrong, et al.
Veröffentlicht: (2025)
Adaptive Feature-based Low-Rank Compression of Large Language Models via Bayesian Optimization
von: Ji, Yixin, et al.
Veröffentlicht: (2024)
von: Ji, Yixin, et al.
Veröffentlicht: (2024)
Basis Sharing: Cross-Layer Parameter Sharing for Large Language Model Compression
von: Wang, Jingcun, et al.
Veröffentlicht: (2024)
von: Wang, Jingcun, et al.
Veröffentlicht: (2024)
Focus-LIME: Surgical Interpretation of Long-Context Large Language Models via Proxy-Based Neighborhood Selection
von: Liu, Junhao, et al.
Veröffentlicht: (2026)
von: Liu, Junhao, et al.
Veröffentlicht: (2026)
Rethinking the Role of Proxy Rewards in Language Model Alignment
von: Kim, Sungdong, et al.
Veröffentlicht: (2024)
von: Kim, Sungdong, et al.
Veröffentlicht: (2024)
ILRe: Intermediate Layer Retrieval for Context Compression in Causal Language Models
von: Liang, Manlai, et al.
Veröffentlicht: (2025)
von: Liang, Manlai, et al.
Veröffentlicht: (2025)
Byte-token Enhanced Language Models for Temporal Point Processes Analysis
von: Kong, Quyu, et al.
Veröffentlicht: (2025)
von: Kong, Quyu, et al.
Veröffentlicht: (2025)
FoldGPT: Simple and Effective Large Language Model Compression Scheme
von: Liu, Songwei, et al.
Veröffentlicht: (2024)
von: Liu, Songwei, et al.
Veröffentlicht: (2024)
SVD-LLM: Truncation-aware Singular Value Decomposition for Large Language Model Compression
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Towards Unified Task Embeddings Across Multiple Models: Bridging the Gap for Prompt-Based Large Language Models and Beyond
von: Wang, Xinyu, et al.
Veröffentlicht: (2024)
von: Wang, Xinyu, et al.
Veröffentlicht: (2024)
GOFA: A Generative One-For-All Model for Joint Graph Language Modeling
von: Kong, Lecheng, et al.
Veröffentlicht: (2024)
von: Kong, Lecheng, et al.
Veröffentlicht: (2024)
Self-Refinement of Language Models from External Proxy Metrics Feedback
von: Ramji, Keshav, et al.
Veröffentlicht: (2024)
von: Ramji, Keshav, et al.
Veröffentlicht: (2024)
Diversity Measures: Domain-Independent Proxies for Failure in Language Model Queries
von: Ngu, Noel, et al.
Veröffentlicht: (2023)
von: Ngu, Noel, et al.
Veröffentlicht: (2023)
MiniDisc: Minimal Distillation Schedule for Language Model Compression
von: Zhang, Chen, et al.
Veröffentlicht: (2022)
von: Zhang, Chen, et al.
Veröffentlicht: (2022)
Are Compressed Language Models Less Subgroup Robust?
von: Gee, Leonidas, et al.
Veröffentlicht: (2024)
von: Gee, Leonidas, et al.
Veröffentlicht: (2024)
A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios
von: Feng, Xiachong, et al.
Veröffentlicht: (2024)
von: Feng, Xiachong, et al.
Veröffentlicht: (2024)
Saten: Sparse Augmented Tensor Networks for Post-Training Compression of Large Language Models
von: Solgi, Ryan, et al.
Veröffentlicht: (2025)
von: Solgi, Ryan, et al.
Veröffentlicht: (2025)
Improving Instruction Following in Language Models through Proxy-Based Uncertainty Estimation
von: Lee, JoonHo, et al.
Veröffentlicht: (2024)
von: Lee, JoonHo, et al.
Veröffentlicht: (2024)
Compressed Context Memory For Online Language Model Interaction
von: Kim, Jang-Hyun, et al.
Veröffentlicht: (2023)
von: Kim, Jang-Hyun, et al.
Veröffentlicht: (2023)
Found in the Middle: Permutation Self-Consistency Improves Listwise Ranking in Large Language Models
von: Tang, Raphael, et al.
Veröffentlicht: (2023)
von: Tang, Raphael, et al.
Veröffentlicht: (2023)
Bootstrapping Language Models with DPO Implicit Rewards
von: Chen, Changyu, et al.
Veröffentlicht: (2024)
von: Chen, Changyu, et al.
Veröffentlicht: (2024)
Efficient Parallel Samplers for Recurrent-Depth Models and Their Connection to Diffusion Language Models
von: Geiping, Jonas, et al.
Veröffentlicht: (2025)
von: Geiping, Jonas, et al.
Veröffentlicht: (2025)
Forecasting Downstream Performance of LLMs With Proxy Metrics
von: Patel, Arkil, et al.
Veröffentlicht: (2026)
von: Patel, Arkil, et al.
Veröffentlicht: (2026)
Teaching Large Language Models Number-Focused Headline Generation With Key Element Rationales
von: Qian, Zhen, et al.
Veröffentlicht: (2025)
von: Qian, Zhen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Reparameterized Discrete Diffusion Model for Text Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023) -
DynaAct: Large Language Model Reasoning with Dynamic Action Spaces
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025) -
Reasoning Does Not Necessarily Improve Role-Playing Ability
von: Feng, Xiachong, et al.
Veröffentlicht: (2025) -
Self-Infilling Code Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023) -
PromptCoT 2.0: Scaling Prompt Synthesis for Large Language Model Reasoning
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)