Top-$nσ$: Not All Logits Are You Need
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Chenxia, Liu, Jianchun, Xu, Hongli, Huang, Liusheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive KV-Cache Compression without Manually Setting Budget
by: Tang, Chenxia, et al.
Published: (2025)
by: Tang, Chenxia, et al.
Published: (2025)
Accelerating Mixture-of-Expert Inference with Adaptive Expert Split Mechanism
by: Yan, Jiaming, et al.
Published: (2025)
by: Yan, Jiaming, et al.
Published: (2025)
Enhancing Federated Graph Learning via Adaptive Fusion of Structural and Node Characteristics
by: Gao, Xianjun, et al.
Published: (2024)
by: Gao, Xianjun, et al.
Published: (2024)
Mitigating Catastrophic Forgetting with Adaptive Transformer Block Expansion in Federated Fine-Tuning
by: Huo, Yujia, et al.
Published: (2025)
by: Huo, Yujia, et al.
Published: (2025)
Collaborative Speculative Inference for Efficient LLM Inference Serving
by: Gao, Luyao, et al.
Published: (2025)
by: Gao, Luyao, et al.
Published: (2025)
Towards Communication-Efficient Decentralized Federated Graph Learning over Non-IID Data
by: Wang, Shilong, et al.
Published: (2025)
by: Wang, Shilong, et al.
Published: (2025)
Caesar: A Low-deviation Compression Approach for Efficient Federated Learning
by: Yan, Jiaming, et al.
Published: (2024)
by: Yan, Jiaming, et al.
Published: (2024)
Adaptive and Fine-grained Module-wise Expert Pruning for Efficient LoRA-MoE Fine-Tuning
by: Li, Weihang, et al.
Published: (2026)
by: Li, Weihang, et al.
Published: (2026)
Heterogeneous Learning Rate Scheduling for Neural Architecture Search on Long-Tailed Datasets
by: Tang, Chenxia
Published: (2024)
by: Tang, Chenxia
Published: (2024)
Accuracy is Not All You Need
by: Dutta, Abhinav, et al.
Published: (2024)
by: Dutta, Abhinav, et al.
Published: (2024)
Logits are All We Need to Adapt Closed Models
by: Hiranandani, Gaurush, et al.
Published: (2025)
by: Hiranandani, Gaurush, et al.
Published: (2025)
Improving LLM Reasoning via Dependency-Aware Query Decomposition and Logic-Parallel Content Expansion
by: Gao, Xianjun, et al.
Published: (2025)
by: Gao, Xianjun, et al.
Published: (2025)
FedQuad: Adaptive Layer-wise LoRA Deployment and Activation Quantization for Federated Fine-Tuning
by: Li, Rukuo, et al.
Published: (2025)
by: Li, Rukuo, et al.
Published: (2025)
Support is All You Need for Certified VAE Training
by: Xu, Changming, et al.
Published: (2025)
by: Xu, Changming, et al.
Published: (2025)
Attention is All You Need Until You Need Retention
by: Yaslioglu, M. Murat
Published: (2025)
by: Yaslioglu, M. Murat
Published: (2025)
SemiSFL: Split Federated Learning on Unlabeled and Non-IID Data
by: Xu, Yang, et al.
Published: (2023)
by: Xu, Yang, et al.
Published: (2023)
Context is All You Need
by: Delanois, Jean Erik, et al.
Published: (2026)
by: Delanois, Jean Erik, et al.
Published: (2026)
Attention Is All You Need But You Don't Need All Of It For Inference of Large Language Models
by: Tyukin, Georgy, et al.
Published: (2024)
by: Tyukin, Georgy, et al.
Published: (2024)
TransMLA: Multi-Head Latent Attention Is All You Need
by: Meng, Fanxu, et al.
Published: (2025)
by: Meng, Fanxu, et al.
Published: (2025)
Some Attention is All You Need for Retrieval
by: Michalak, Felix, et al.
Published: (2025)
by: Michalak, Felix, et al.
Published: (2025)
Half Search Space is All You Need
by: Rumiantsev, Pavel, et al.
Published: (2025)
by: Rumiantsev, Pavel, et al.
Published: (2025)
Multistep Inverse Is Not All You Need
by: Levine, Alexander, et al.
Published: (2024)
by: Levine, Alexander, et al.
Published: (2024)
Exploitation Is All You Need... for Exploration
by: Rentschler, Micah, et al.
Published: (2025)
by: Rentschler, Micah, et al.
Published: (2025)
Is Diversity All You Need for Scalable Robotic Manipulation?
by: Shi, Modi, et al.
Published: (2025)
by: Shi, Modi, et al.
Published: (2025)
CAMformer: Associative Memory is All You Need
by: Molom-Ochir, Tergel, et al.
Published: (2025)
by: Molom-Ochir, Tergel, et al.
Published: (2025)
FP64 is All You Need: Rethinking Failure Modes in Physics-Informed Neural Networks
by: Xu, Chenhui, et al.
Published: (2025)
by: Xu, Chenhui, et al.
Published: (2025)
No More Adam: Learning Rate Scaling at Initialization is All You Need
by: Xu, Minghao, et al.
Published: (2024)
by: Xu, Minghao, et al.
Published: (2024)
MoE Lens -- An Expert Is All You Need
by: Chaudhari, Marmik, et al.
Published: (2026)
by: Chaudhari, Marmik, et al.
Published: (2026)
Fusion or Confusion? Multimodal Complexity Is Not All You Need
by: Rheude, Tillmann, et al.
Published: (2025)
by: Rheude, Tillmann, et al.
Published: (2025)
Realizable Learning is All You Need
by: Hopkins, Max, et al.
Published: (2021)
by: Hopkins, Max, et al.
Published: (2021)
Cooperation Is All You Need
by: Adeel, Ahsan, et al.
Published: (2023)
by: Adeel, Ahsan, et al.
Published: (2023)
Attention Smoothing Is All You Need For Unlearning
by: Zade, Saleh Zare, et al.
Published: (2026)
by: Zade, Saleh Zare, et al.
Published: (2026)
Standard Gaussian Process is All You Need for High-Dimensional Bayesian Optimization
by: Xu, Zhitong, et al.
Published: (2024)
by: Xu, Zhitong, et al.
Published: (2024)
Block Rotation is All You Need for MXFP4 Quantization
by: Shao, Yuantian, et al.
Published: (2025)
by: Shao, Yuantian, et al.
Published: (2025)
All You Need Is Synthetic Task Augmentation
by: Godin, Guillaume
Published: (2025)
by: Godin, Guillaume
Published: (2025)
Element-wise Attention Is All You Need
by: Feng, Guoxin
Published: (2025)
by: Feng, Guoxin
Published: (2025)
Tensor Product Attention Is All You Need
by: Zhang, Yifan, et al.
Published: (2025)
by: Zhang, Yifan, et al.
Published: (2025)
More Agents Is All You Need
by: Li, Junyou, et al.
Published: (2024)
by: Li, Junyou, et al.
Published: (2024)
Alignment with Preference Optimization Is All You Need for LLM Safety
by: Alami, Reda, et al.
Published: (2024)
by: Alami, Reda, et al.
Published: (2024)
Uni-LoRA: One Vector is All You Need
by: Li, Kaiyang, et al.
Published: (2025)
by: Li, Kaiyang, et al.
Published: (2025)
Similar Items
-
Adaptive KV-Cache Compression without Manually Setting Budget
by: Tang, Chenxia, et al.
Published: (2025) -
Accelerating Mixture-of-Expert Inference with Adaptive Expert Split Mechanism
by: Yan, Jiaming, et al.
Published: (2025) -
Enhancing Federated Graph Learning via Adaptive Fusion of Structural and Node Characteristics
by: Gao, Xianjun, et al.
Published: (2024) -
Mitigating Catastrophic Forgetting with Adaptive Transformer Block Expansion in Federated Fine-Tuning
by: Huo, Yujia, et al.
Published: (2025) -
Collaborative Speculative Inference for Efficient LLM Inference Serving
by: Gao, Luyao, et al.
Published: (2025)