Rethinking the Outlier Distribution in Large Language Models: An In-depth Study
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Raman, Rahul, Sharma, Khushi, Zhang, Sai Qian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DLoRA: Distributed Parameter-Efficient Fine-Tuning Solution for Large Language Model
von: Gao, Chao, et al.
Veröffentlicht: (2024)
von: Gao, Chao, et al.
Veröffentlicht: (2024)
Systematic Outliers in Large Language Models
von: An, Yongqi, et al.
Veröffentlicht: (2025)
von: An, Yongqi, et al.
Veröffentlicht: (2025)
Evaluating the Role of Large Language Models in Legal Practice in India
von: Hemrajani, Rahul
Veröffentlicht: (2025)
von: Hemrajani, Rahul
Veröffentlicht: (2025)
Can Large Language Models Discern Evidence for Scientific Hypotheses? Case Studies in the Social Sciences
von: Koneru, Sai, et al.
Veröffentlicht: (2023)
von: Koneru, Sai, et al.
Veröffentlicht: (2023)
EDA Corpus: A Large Language Model Dataset for Enhanced Interaction with OpenROAD
von: Wu, Bing-Yue, et al.
Veröffentlicht: (2024)
von: Wu, Bing-Yue, et al.
Veröffentlicht: (2024)
The Labyrinth and the Thread: Rethinking Regularizations in Sequential Knowledge Editing for Large Language Models
von: Wang, Zheng, et al.
Veröffentlicht: (2026)
von: Wang, Zheng, et al.
Veröffentlicht: (2026)
Layer-Order Inversion: Rethinking Latent Multi-Hop Reasoning in Large Language Models
von: Liu, Xukai, et al.
Veröffentlicht: (2026)
von: Liu, Xukai, et al.
Veröffentlicht: (2026)
Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe
von: Li, Yaxuan, et al.
Veröffentlicht: (2026)
von: Li, Yaxuan, et al.
Veröffentlicht: (2026)
Refract ICL: Rethinking Example Selection in the Era of Million-Token Models
von: Akula, Arjun R., et al.
Veröffentlicht: (2025)
von: Akula, Arjun R., et al.
Veröffentlicht: (2025)
Rethinking Kullback-Leibler Divergence in Knowledge Distillation for Large Language Models
von: Wu, Taiqiang, et al.
Veröffentlicht: (2024)
von: Wu, Taiqiang, et al.
Veröffentlicht: (2024)
Improving the Language Understanding Capabilities of Large Language Models Using Reinforcement Learning
von: Hu, Bokai, et al.
Veröffentlicht: (2024)
von: Hu, Bokai, et al.
Veröffentlicht: (2024)
Outlier-Safe Pre-Training for Robust 4-Bit Quantization of Large Language Models
von: Park, Jungwoo, et al.
Veröffentlicht: (2025)
von: Park, Jungwoo, et al.
Veröffentlicht: (2025)
Rethinking Interpretability in the Era of Large Language Models
von: Singh, Chandan, et al.
Veröffentlicht: (2024)
von: Singh, Chandan, et al.
Veröffentlicht: (2024)
Rethinking Toxicity Evaluation in Large Language Models: A Multi-Label Perspective
von: Kou, Zhiqiang, et al.
Veröffentlicht: (2025)
von: Kou, Zhiqiang, et al.
Veröffentlicht: (2025)
Towards Universal Semantics With Large Language Models
von: Baartmans, Raymond, et al.
Veröffentlicht: (2025)
von: Baartmans, Raymond, et al.
Veröffentlicht: (2025)
GL-Fusion: Rethinking the Combination of Graph Neural Network and Large Language model
von: Yang, Haotong, et al.
Veröffentlicht: (2024)
von: Yang, Haotong, et al.
Veröffentlicht: (2024)
Rethinking Data Mixing from the Perspective of Large Language Models
von: Xu, Yuanjian, et al.
Veröffentlicht: (2026)
von: Xu, Yuanjian, et al.
Veröffentlicht: (2026)
From Symbolic to Natural-Language Relations: Rethinking Knowledge Graph Construction in the Era of Large Language Models
von: Han, Kanyao, et al.
Veröffentlicht: (2026)
von: Han, Kanyao, et al.
Veröffentlicht: (2026)
GenRES: Rethinking Evaluation for Generative Relation Extraction in the Era of Large Language Models
von: Jiang, Pengcheng, et al.
Veröffentlicht: (2024)
von: Jiang, Pengcheng, et al.
Veröffentlicht: (2024)
"Lost-in-the-Later": Framework for Quantifying Contextual Grounding in Large Language Models
von: Tao, Yufei, et al.
Veröffentlicht: (2025)
von: Tao, Yufei, et al.
Veröffentlicht: (2025)
Benchmarking Distributional Alignment of Large Language Models
von: Meister, Nicole, et al.
Veröffentlicht: (2024)
von: Meister, Nicole, et al.
Veröffentlicht: (2024)
Contextual Refinement of Translations: Large Language Models for Sentence and Document-Level Post-Editing
von: Koneru, Sai, et al.
Veröffentlicht: (2023)
von: Koneru, Sai, et al.
Veröffentlicht: (2023)
Pre-training LLMs using human-like development data corpus
von: Bhardwaj, Khushi, et al.
Veröffentlicht: (2023)
von: Bhardwaj, Khushi, et al.
Veröffentlicht: (2023)
Understanding the Dilemma of Unlearning for Large Language Models
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
Bengali Text Classification: An Evaluation of Large Language Model Approaches
von: Hoque, Md Mahmudul, et al.
Veröffentlicht: (2026)
von: Hoque, Md Mahmudul, et al.
Veröffentlicht: (2026)
Robust Prompt Optimization for Large Language Models Against Distribution Shifts
von: Li, Moxin, et al.
Veröffentlicht: (2023)
von: Li, Moxin, et al.
Veröffentlicht: (2023)
An Empirical Study on Prompt Compression for Large Language Models
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
BiSup: Bidirectional Quantization Error Suppression for Large Language Models
von: Zou, Minghui, et al.
Veröffentlicht: (2024)
von: Zou, Minghui, et al.
Veröffentlicht: (2024)
Faux Polyglot: A Study on Information Disparity in Multilingual Large Language Models
von: Sharma, Nikhil, et al.
Veröffentlicht: (2024)
von: Sharma, Nikhil, et al.
Veröffentlicht: (2024)
Harnessing the Power of Large Language Models for Empathetic Response Generation: Empirical Investigations and Improvements
von: Qian, Yushan, et al.
Veröffentlicht: (2023)
von: Qian, Yushan, et al.
Veröffentlicht: (2023)
ConSiDERS-The-Human Evaluation Framework: Rethinking Human Evaluation for Generative Large Language Models
von: Elangovan, Aparna, et al.
Veröffentlicht: (2024)
von: Elangovan, Aparna, et al.
Veröffentlicht: (2024)
Knowledge-based Consistency Testing of Large Language Models
von: Rajan, Sai Sathiesh, et al.
Veröffentlicht: (2024)
von: Rajan, Sai Sathiesh, et al.
Veröffentlicht: (2024)
An In-depth Evaluation of Large Language Models in Sentence Simplification with Error-based Human Assessment
von: Wu, Xuanxin, et al.
Veröffentlicht: (2024)
von: Wu, Xuanxin, et al.
Veröffentlicht: (2024)
Examining Inter-Consistency of Large Language Models Collaboration: An In-depth Analysis via Debate
von: Xiong, Kai, et al.
Veröffentlicht: (2023)
von: Xiong, Kai, et al.
Veröffentlicht: (2023)
On the Tip of the Tongue: Analyzing Conceptual Representation in Large Language Models with Reverse-Dictionary Probe
von: Xu, Ningyu, et al.
Veröffentlicht: (2024)
von: Xu, Ningyu, et al.
Veröffentlicht: (2024)
Let's Play Across Cultures: A Large Multilingual, Multicultural Benchmark for Assessing Language Models' Understanding of Sports
von: Singh, Punit Kumar, et al.
Veröffentlicht: (2025)
von: Singh, Punit Kumar, et al.
Veröffentlicht: (2025)
Rethinking the Role of Text Complexity in Language Model Pretraining
von: Velasco, Dan John, et al.
Veröffentlicht: (2025)
von: Velasco, Dan John, et al.
Veröffentlicht: (2025)
CASTILLO: Characterizing Response Length Distributions of Large Language Models
von: Perez-Ramirez, Daniel F., et al.
Veröffentlicht: (2025)
von: Perez-Ramirez, Daniel F., et al.
Veröffentlicht: (2025)
On Adversarial Robustness and Out-of-Distribution Robustness of Large Language Models
von: Yang, April, et al.
Veröffentlicht: (2024)
von: Yang, April, et al.
Veröffentlicht: (2024)
Disentangling Reasoning and Knowledge in Medical Large Language Models
von: Thapa, Rahul, et al.
Veröffentlicht: (2025)
von: Thapa, Rahul, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DLoRA: Distributed Parameter-Efficient Fine-Tuning Solution for Large Language Model
von: Gao, Chao, et al.
Veröffentlicht: (2024) -
Systematic Outliers in Large Language Models
von: An, Yongqi, et al.
Veröffentlicht: (2025) -
Evaluating the Role of Large Language Models in Legal Practice in India
von: Hemrajani, Rahul
Veröffentlicht: (2025) -
Can Large Language Models Discern Evidence for Scientific Hypotheses? Case Studies in the Social Sciences
von: Koneru, Sai, et al.
Veröffentlicht: (2023) -
EDA Corpus: A Large Language Model Dataset for Enhanced Interaction with OpenROAD
von: Wu, Bing-Yue, et al.
Veröffentlicht: (2024)