Top-$nσ$: Not All Logits Are You Need
Fuente:
arXiv
Guardado en:
| Autores principales: | Tang, Chenxia, Liu, Jianchun, Xu, Hongli, Huang, Liusheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Adaptive KV-Cache Compression without Manually Setting Budget
por: Tang, Chenxia, et al.
Publicado: (2025)
por: Tang, Chenxia, et al.
Publicado: (2025)
Accelerating Mixture-of-Expert Inference with Adaptive Expert Split Mechanism
por: Yan, Jiaming, et al.
Publicado: (2025)
por: Yan, Jiaming, et al.
Publicado: (2025)
Enhancing Federated Graph Learning via Adaptive Fusion of Structural and Node Characteristics
por: Gao, Xianjun, et al.
Publicado: (2024)
por: Gao, Xianjun, et al.
Publicado: (2024)
Mitigating Catastrophic Forgetting with Adaptive Transformer Block Expansion in Federated Fine-Tuning
por: Huo, Yujia, et al.
Publicado: (2025)
por: Huo, Yujia, et al.
Publicado: (2025)
Collaborative Speculative Inference for Efficient LLM Inference Serving
por: Gao, Luyao, et al.
Publicado: (2025)
por: Gao, Luyao, et al.
Publicado: (2025)
Towards Communication-Efficient Decentralized Federated Graph Learning over Non-IID Data
por: Wang, Shilong, et al.
Publicado: (2025)
por: Wang, Shilong, et al.
Publicado: (2025)
Caesar: A Low-deviation Compression Approach for Efficient Federated Learning
por: Yan, Jiaming, et al.
Publicado: (2024)
por: Yan, Jiaming, et al.
Publicado: (2024)
Adaptive and Fine-grained Module-wise Expert Pruning for Efficient LoRA-MoE Fine-Tuning
por: Li, Weihang, et al.
Publicado: (2026)
por: Li, Weihang, et al.
Publicado: (2026)
Heterogeneous Learning Rate Scheduling for Neural Architecture Search on Long-Tailed Datasets
por: Tang, Chenxia
Publicado: (2024)
por: Tang, Chenxia
Publicado: (2024)
Accuracy is Not All You Need
por: Dutta, Abhinav, et al.
Publicado: (2024)
por: Dutta, Abhinav, et al.
Publicado: (2024)
Logits are All We Need to Adapt Closed Models
por: Hiranandani, Gaurush, et al.
Publicado: (2025)
por: Hiranandani, Gaurush, et al.
Publicado: (2025)
Improving LLM Reasoning via Dependency-Aware Query Decomposition and Logic-Parallel Content Expansion
por: Gao, Xianjun, et al.
Publicado: (2025)
por: Gao, Xianjun, et al.
Publicado: (2025)
FedQuad: Adaptive Layer-wise LoRA Deployment and Activation Quantization for Federated Fine-Tuning
por: Li, Rukuo, et al.
Publicado: (2025)
por: Li, Rukuo, et al.
Publicado: (2025)
Support is All You Need for Certified VAE Training
por: Xu, Changming, et al.
Publicado: (2025)
por: Xu, Changming, et al.
Publicado: (2025)
Attention is All You Need Until You Need Retention
por: Yaslioglu, M. Murat
Publicado: (2025)
por: Yaslioglu, M. Murat
Publicado: (2025)
SemiSFL: Split Federated Learning on Unlabeled and Non-IID Data
por: Xu, Yang, et al.
Publicado: (2023)
por: Xu, Yang, et al.
Publicado: (2023)
Context is All You Need
por: Delanois, Jean Erik, et al.
Publicado: (2026)
por: Delanois, Jean Erik, et al.
Publicado: (2026)
Attention Is All You Need But You Don't Need All Of It For Inference of Large Language Models
por: Tyukin, Georgy, et al.
Publicado: (2024)
por: Tyukin, Georgy, et al.
Publicado: (2024)
TransMLA: Multi-Head Latent Attention Is All You Need
por: Meng, Fanxu, et al.
Publicado: (2025)
por: Meng, Fanxu, et al.
Publicado: (2025)
Some Attention is All You Need for Retrieval
por: Michalak, Felix, et al.
Publicado: (2025)
por: Michalak, Felix, et al.
Publicado: (2025)
Half Search Space is All You Need
por: Rumiantsev, Pavel, et al.
Publicado: (2025)
por: Rumiantsev, Pavel, et al.
Publicado: (2025)
Multistep Inverse Is Not All You Need
por: Levine, Alexander, et al.
Publicado: (2024)
por: Levine, Alexander, et al.
Publicado: (2024)
Exploitation Is All You Need... for Exploration
por: Rentschler, Micah, et al.
Publicado: (2025)
por: Rentschler, Micah, et al.
Publicado: (2025)
Is Diversity All You Need for Scalable Robotic Manipulation?
por: Shi, Modi, et al.
Publicado: (2025)
por: Shi, Modi, et al.
Publicado: (2025)
CAMformer: Associative Memory is All You Need
por: Molom-Ochir, Tergel, et al.
Publicado: (2025)
por: Molom-Ochir, Tergel, et al.
Publicado: (2025)
FP64 is All You Need: Rethinking Failure Modes in Physics-Informed Neural Networks
por: Xu, Chenhui, et al.
Publicado: (2025)
por: Xu, Chenhui, et al.
Publicado: (2025)
No More Adam: Learning Rate Scaling at Initialization is All You Need
por: Xu, Minghao, et al.
Publicado: (2024)
por: Xu, Minghao, et al.
Publicado: (2024)
MoE Lens -- An Expert Is All You Need
por: Chaudhari, Marmik, et al.
Publicado: (2026)
por: Chaudhari, Marmik, et al.
Publicado: (2026)
Fusion or Confusion? Multimodal Complexity Is Not All You Need
por: Rheude, Tillmann, et al.
Publicado: (2025)
por: Rheude, Tillmann, et al.
Publicado: (2025)
Realizable Learning is All You Need
por: Hopkins, Max, et al.
Publicado: (2021)
por: Hopkins, Max, et al.
Publicado: (2021)
Cooperation Is All You Need
por: Adeel, Ahsan, et al.
Publicado: (2023)
por: Adeel, Ahsan, et al.
Publicado: (2023)
Attention Smoothing Is All You Need For Unlearning
por: Zade, Saleh Zare, et al.
Publicado: (2026)
por: Zade, Saleh Zare, et al.
Publicado: (2026)
Standard Gaussian Process is All You Need for High-Dimensional Bayesian Optimization
por: Xu, Zhitong, et al.
Publicado: (2024)
por: Xu, Zhitong, et al.
Publicado: (2024)
Block Rotation is All You Need for MXFP4 Quantization
por: Shao, Yuantian, et al.
Publicado: (2025)
por: Shao, Yuantian, et al.
Publicado: (2025)
All You Need Is Synthetic Task Augmentation
por: Godin, Guillaume
Publicado: (2025)
por: Godin, Guillaume
Publicado: (2025)
Element-wise Attention Is All You Need
por: Feng, Guoxin
Publicado: (2025)
por: Feng, Guoxin
Publicado: (2025)
Tensor Product Attention Is All You Need
por: Zhang, Yifan, et al.
Publicado: (2025)
por: Zhang, Yifan, et al.
Publicado: (2025)
More Agents Is All You Need
por: Li, Junyou, et al.
Publicado: (2024)
por: Li, Junyou, et al.
Publicado: (2024)
Alignment with Preference Optimization Is All You Need for LLM Safety
por: Alami, Reda, et al.
Publicado: (2024)
por: Alami, Reda, et al.
Publicado: (2024)
Uni-LoRA: One Vector is All You Need
por: Li, Kaiyang, et al.
Publicado: (2025)
por: Li, Kaiyang, et al.
Publicado: (2025)
Ejemplares similares
-
Adaptive KV-Cache Compression without Manually Setting Budget
por: Tang, Chenxia, et al.
Publicado: (2025) -
Accelerating Mixture-of-Expert Inference with Adaptive Expert Split Mechanism
por: Yan, Jiaming, et al.
Publicado: (2025) -
Enhancing Federated Graph Learning via Adaptive Fusion of Structural and Node Characteristics
por: Gao, Xianjun, et al.
Publicado: (2024) -
Mitigating Catastrophic Forgetting with Adaptive Transformer Block Expansion in Federated Fine-Tuning
por: Huo, Yujia, et al.
Publicado: (2025) -
Collaborative Speculative Inference for Efficient LLM Inference Serving
por: Gao, Luyao, et al.
Publicado: (2025)