Combining Entropy and Matrix Nuclear Norm for Enhanced Evaluation of Language Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Vo, James |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Large Language Model Evaluation via Matrix Nuclear-Norm
von: Li, Yahan, et al.
Veröffentlicht: (2024)
von: Li, Yahan, et al.
Veröffentlicht: (2024)
Transformer Layer Injection: A Novel Approach for Efficient Upscaling of Large Language Models
von: Vo, James
Veröffentlicht: (2024)
von: Vo, James
Veröffentlicht: (2024)
Vi-Mistral-X: Building a Vietnamese Language Model with Advanced Continual Pre-training
von: Vo, James
Veröffentlicht: (2024)
von: Vo, James
Veröffentlicht: (2024)
SparseAccelerate: Efficient Long-Context Inference for Mid-Range GPUs
von: Vo, James
Veröffentlicht: (2024)
von: Vo, James
Veröffentlicht: (2024)
PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action
von: Shao, Yijia, et al.
Veröffentlicht: (2024)
von: Shao, Yijia, et al.
Veröffentlicht: (2024)
Efficient Second-Order Neural Network Optimization via Adaptive Trust Region Methods
von: Vo, James
Veröffentlicht: (2024)
von: Vo, James
Veröffentlicht: (2024)
Towards Robustness and Diversity: Continual Learning in Dialog Generation with Text-Mixup and Batch Nuclear-Norm Maximization
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
Entropy in Large Language Models
von: Scharringhausen, Marco
Veröffentlicht: (2026)
von: Scharringhausen, Marco
Veröffentlicht: (2026)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
von: Vo, Truong, et al.
Veröffentlicht: (2025)
von: Vo, Truong, et al.
Veröffentlicht: (2025)
Towards Reliable Medical Question Answering: Techniques and Challenges in Mitigating Hallucinations in Language Models
von: Pham, Duy Khoa, et al.
Veröffentlicht: (2024)
von: Pham, Duy Khoa, et al.
Veröffentlicht: (2024)
Entropy2Vec: Crosslingual Language Modeling Entropy as End-to-End Learnable Language Representations
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2025)
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2025)
NormAd: A Framework for Measuring the Cultural Adaptability of Large Language Models
von: Rao, Abhinav, et al.
Veröffentlicht: (2024)
von: Rao, Abhinav, et al.
Veröffentlicht: (2024)
DiNADO: Norm-Disentangled Neurally-Decomposed Oracles for Controlling Language Models
von: Lu, Sidi, et al.
Veröffentlicht: (2023)
von: Lu, Sidi, et al.
Veröffentlicht: (2023)
SocialGaze: Improving the Integration of Human Social Norms in Large Language Models
von: Vijjini, Anvesh Rao, et al.
Veröffentlicht: (2024)
von: Vijjini, Anvesh Rao, et al.
Veröffentlicht: (2024)
Execution-Based Evaluation of Natural Language to Bash and PowerShell for Incident Remediation
von: Vo, Ngoc Phuoc An, et al.
Veröffentlicht: (2024)
von: Vo, Ngoc Phuoc An, et al.
Veröffentlicht: (2024)
Measuring Social Norms of Large Language Models
von: Yuan, Ye, et al.
Veröffentlicht: (2024)
von: Yuan, Ye, et al.
Veröffentlicht: (2024)
Entropy-Based Data Selection for Language Models
von: Li, Hongming, et al.
Veröffentlicht: (2026)
von: Li, Hongming, et al.
Veröffentlicht: (2026)
SLMEval: Entropy-Based Calibration for Human-Aligned Evaluation of Large Language Models
von: Daynauth, Roland, et al.
Veröffentlicht: (2025)
von: Daynauth, Roland, et al.
Veröffentlicht: (2025)
EntropyCache: Decoded Token Entropy Guided KV Caching for Diffusion Language Models
von: Cheong, Minsoo, et al.
Veröffentlicht: (2026)
von: Cheong, Minsoo, et al.
Veröffentlicht: (2026)
NormGenesis: Multicultural Dialogue Generation via Exemplar-Guided Social Norm Modeling and Violation Recovery
von: Hong, Minki, et al.
Veröffentlicht: (2025)
von: Hong, Minki, et al.
Veröffentlicht: (2025)
Unlocking the Potential of Large Language Models in the Nuclear Industry with Synthetic Data
von: Anwar, Muhammad, et al.
Veröffentlicht: (2025)
von: Anwar, Muhammad, et al.
Veröffentlicht: (2025)
CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs
von: Ye, Yangfan, et al.
Veröffentlicht: (2026)
von: Ye, Yangfan, et al.
Veröffentlicht: (2026)
Systematic Generalization in Language Models Scales with Information Entropy
von: Wold, Sondre, et al.
Veröffentlicht: (2025)
von: Wold, Sondre, et al.
Veröffentlicht: (2025)
Sparse Matrix in Large Language Model Fine-tuning
von: He, Haoze, et al.
Veröffentlicht: (2024)
von: He, Haoze, et al.
Veröffentlicht: (2024)
VideoNorms: Benchmarking Cultural Awareness of Video Language Models
von: Varimalla, Nikhil Reddy, et al.
Veröffentlicht: (2025)
von: Varimalla, Nikhil Reddy, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models with Psychometrics
von: Li, Yuan, et al.
Veröffentlicht: (2024)
von: Li, Yuan, et al.
Veröffentlicht: (2024)
MultiLexNorm++: A Unified Benchmark and a Generative Model for Lexical Normalization for Asian Languages
von: Buaphet, Weerayut, et al.
Veröffentlicht: (2026)
von: Buaphet, Weerayut, et al.
Veröffentlicht: (2026)
To Words and Beyond: Probing Large Language Models for Sentence-Level Psycholinguistic Norms of Memorability and Reading Times
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
Aligning Large Language Models with Human Opinions through Persona Selection and Value--Belief--Norm Reasoning
von: Long, Do Xuan, et al.
Veröffentlicht: (2023)
von: Long, Do Xuan, et al.
Veröffentlicht: (2023)
GAIN: A Benchmark for Goal-Aligned Decision-Making of Large Language Models under Imperfect Norms
von: Kawarada, Masayuki, et al.
Veröffentlicht: (2026)
von: Kawarada, Masayuki, et al.
Veröffentlicht: (2026)
GeoNorm: Unify Pre-Norm and Post-Norm with Geodesic Optimization
von: Zheng, Chuanyang, et al.
Veröffentlicht: (2026)
von: Zheng, Chuanyang, et al.
Veröffentlicht: (2026)
Entropy-Based Decoding for Retrieval-Augmented Large Language Models
von: Qiu, Zexuan, et al.
Veröffentlicht: (2024)
von: Qiu, Zexuan, et al.
Veröffentlicht: (2024)
AWM: Accurate Weight-Matrix Fingerprint for Large Language Models
von: Zeng, Boyi, et al.
Veröffentlicht: (2025)
von: Zeng, Boyi, et al.
Veröffentlicht: (2025)
Combining Autoregressive and Autoencoder Language Models for Text Classification
von: Gonçalves, João
Veröffentlicht: (2024)
von: Gonçalves, João
Veröffentlicht: (2024)
Mixture-of-Agents Enhances Large Language Model Capabilities
von: Wang, Junlin, et al.
Veröffentlicht: (2024)
von: Wang, Junlin, et al.
Veröffentlicht: (2024)
Evaluating Vision-Language Models for Emotion Recognition
von: Bhattacharyya, Sree, et al.
Veröffentlicht: (2025)
von: Bhattacharyya, Sree, et al.
Veröffentlicht: (2025)
Enhancing Incremental Summarization with Structured Representations
von: Hwang, EunJeong, et al.
Veröffentlicht: (2024)
von: Hwang, EunJeong, et al.
Veröffentlicht: (2024)
SemiKong: Curating, Training, and Evaluating A Semiconductor Industry-Specific Large Language Model
von: Nguyen, Christopher, et al.
Veröffentlicht: (2024)
von: Nguyen, Christopher, et al.
Veröffentlicht: (2024)
Towards Secure and Private Language Models for Nuclear Power Plants
von: Anwar, Muhammad, et al.
Veröffentlicht: (2025)
von: Anwar, Muhammad, et al.
Veröffentlicht: (2025)
On the Entropy Calibration of Language Models
von: Cao, Steven, et al.
Veröffentlicht: (2025)
von: Cao, Steven, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Large Language Model Evaluation via Matrix Nuclear-Norm
von: Li, Yahan, et al.
Veröffentlicht: (2024) -
Transformer Layer Injection: A Novel Approach for Efficient Upscaling of Large Language Models
von: Vo, James
Veröffentlicht: (2024) -
Vi-Mistral-X: Building a Vietnamese Language Model with Advanced Continual Pre-training
von: Vo, James
Veröffentlicht: (2024) -
SparseAccelerate: Efficient Long-Context Inference for Mid-Range GPUs
von: Vo, James
Veröffentlicht: (2024) -
PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action
von: Shao, Yijia, et al.
Veröffentlicht: (2024)