Entropy, Thermodynamics and the Geometrization of the Language Model
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Yang, Wenzhe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Controlling Long-Horizon Behavior in Language Model Agents with Explicit State Dynamics
von: Subaharan, Sukesh
Veröffentlicht: (2026)
von: Subaharan, Sukesh
Veröffentlicht: (2026)
SATA-BENCH: Select All That Apply Benchmark for Multiple Choice Questions
von: Xu, Weijie, et al.
Veröffentlicht: (2025)
von: Xu, Weijie, et al.
Veröffentlicht: (2025)
Epidemic Information Extraction for Event-Based Surveillance using Large Language Models
von: Consoli, Sergio, et al.
Veröffentlicht: (2024)
von: Consoli, Sergio, et al.
Veröffentlicht: (2024)
BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models
von: Abhishek, Alok, et al.
Veröffentlicht: (2025)
von: Abhishek, Alok, et al.
Veröffentlicht: (2025)
Low-Resource Neural Machine Translation Using Recurrent Neural Networks and Transfer Learning: A Case Study on English-to-Igbo
von: Ekle, Ocheme Anthony, et al.
Veröffentlicht: (2025)
von: Ekle, Ocheme Anthony, et al.
Veröffentlicht: (2025)
SHARP: Social Harm Analysis via Risk Profiles for Measuring Inequities in Large Language Models
von: Abhishek, Alok, et al.
Veröffentlicht: (2026)
von: Abhishek, Alok, et al.
Veröffentlicht: (2026)
Large Language Models are Inconsistent and Biased Evaluators
von: Stureborg, Rickard, et al.
Veröffentlicht: (2024)
von: Stureborg, Rickard, et al.
Veröffentlicht: (2024)
Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems
von: Berman, Shmuel, et al.
Veröffentlicht: (2024)
von: Berman, Shmuel, et al.
Veröffentlicht: (2024)
From MTEB to MTOB: Retrieval-Augmented Classification for Descriptive Grammars
von: Kornilov, Albert, et al.
Veröffentlicht: (2024)
von: Kornilov, Albert, et al.
Veröffentlicht: (2024)
TSDS: Data Selection for Task-Specific Model Finetuning
von: Liu, Zifan, et al.
Veröffentlicht: (2024)
von: Liu, Zifan, et al.
Veröffentlicht: (2024)
An Automatic Text Classification Method Based on Hierarchical Taxonomies, Neural Networks and Document Embedding: The NETHIC Tool
von: Lomasto, Luigi, et al.
Veröffentlicht: (2026)
von: Lomasto, Luigi, et al.
Veröffentlicht: (2026)
Data and AI governance: Promoting equity, ethics, and fairness in large language models
von: Abhishek, Alok, et al.
Veröffentlicht: (2025)
von: Abhishek, Alok, et al.
Veröffentlicht: (2025)
Reasoning Promotes Robustness in Theory of Mind Tasks
von: de Haan, Ian B., et al.
Veröffentlicht: (2026)
von: de Haan, Ian B., et al.
Veröffentlicht: (2026)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
von: Yang, Yibo
Veröffentlicht: (2025)
von: Yang, Yibo
Veröffentlicht: (2025)
Evaluating Pixel Language Models on Non-Standardized Languages
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2024)
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2024)
NurValues: Real-World Nursing Values Evaluation for Large Language Models in Clinical Context
von: Yao, Ben, et al.
Veröffentlicht: (2025)
von: Yao, Ben, et al.
Veröffentlicht: (2025)
Tailoring Vaccine Messaging with Common-Ground Opinions
von: Stureborg, Rickard, et al.
Veröffentlicht: (2024)
von: Stureborg, Rickard, et al.
Veröffentlicht: (2024)
The Superalignment of Superhuman Intelligence with Large Language Models
von: Huang, Minlie, et al.
Veröffentlicht: (2024)
von: Huang, Minlie, et al.
Veröffentlicht: (2024)
An Unforgeable Publicly Verifiable Watermark for Large Language Models
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
Exploring State Tracking Capabilities of Large Language Models
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)
CUBO: Self-Contained Retrieval-Augmented Generation on Consumer Laptops 10 GB Corpora, 16 GB RAM, Single-Device Deployment
von: Astrino, Paolo
Veröffentlicht: (2026)
von: Astrino, Paolo
Veröffentlicht: (2026)
Morphological Synthesizer for Ge'ez Language: Addressing Morphological Complexity and Resource Limitations
von: Gebremariam, Gebrearegawi, et al.
Veröffentlicht: (2025)
von: Gebremariam, Gebrearegawi, et al.
Veröffentlicht: (2025)
Distilling Large Language Models for Efficient Clinical Information Extraction
von: Vedula, Karthik S., et al.
Veröffentlicht: (2024)
von: Vedula, Karthik S., et al.
Veröffentlicht: (2024)
A Survey of Text Watermarking in the Era of Large Language Models
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
Adaptive Steering and Remasking for Safe Generation in Diffusion Language Models
von: Lee, Yejin, et al.
Veröffentlicht: (2026)
von: Lee, Yejin, et al.
Veröffentlicht: (2026)
Triplètoile: Extraction of Knowledge from Microblogging Text
von: Zavarella, Vanni, et al.
Veröffentlicht: (2024)
von: Zavarella, Vanni, et al.
Veröffentlicht: (2024)
Towards Effective and Efficient Continual Pre-training of Large Language Models
von: Chen, Jie, et al.
Veröffentlicht: (2024)
von: Chen, Jie, et al.
Veröffentlicht: (2024)
Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models
von: Park, Seungcheol, et al.
Veröffentlicht: (2023)
von: Park, Seungcheol, et al.
Veröffentlicht: (2023)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
Arithmetic Without Algorithms: Language Models Solve Math With a Bag of Heuristics
von: Nikankin, Yaniv, et al.
Veröffentlicht: (2024)
von: Nikankin, Yaniv, et al.
Veröffentlicht: (2024)
Setting Standards in Turkish NLP: TR-MMLU for Large Language Model Evaluation
von: Bayram, M. Ali, et al.
Veröffentlicht: (2024)
von: Bayram, M. Ali, et al.
Veröffentlicht: (2024)
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
Accurate Sublayer Pruning for Large Language Models by Exploiting Latency and Tunability Information
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
von: Fang, Xi, et al.
Veröffentlicht: (2024)
von: Fang, Xi, et al.
Veröffentlicht: (2024)
d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
WeDLM: Reconciling Diffusion Language Models with Standard Causal Attention for Fast Inference
von: Liu, Aiwei, et al.
Veröffentlicht: (2025)
von: Liu, Aiwei, et al.
Veröffentlicht: (2025)
Benchmarking Energy Efficiency of Large Language Models Using vLLM
von: Pronk, K., et al.
Veröffentlicht: (2025)
von: Pronk, K., et al.
Veröffentlicht: (2025)
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
The Paradox of Poetic Intent in Back-Translation: Evaluating the Quality of Large Language Models in Chinese Translation
von: Weigang, Li, et al.
Veröffentlicht: (2025)
von: Weigang, Li, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Controlling Long-Horizon Behavior in Language Model Agents with Explicit State Dynamics
von: Subaharan, Sukesh
Veröffentlicht: (2026) -
SATA-BENCH: Select All That Apply Benchmark for Multiple Choice Questions
von: Xu, Weijie, et al.
Veröffentlicht: (2025) -
Epidemic Information Extraction for Event-Based Surveillance using Large Language Models
von: Consoli, Sergio, et al.
Veröffentlicht: (2024) -
BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models
von: Abhishek, Alok, et al.
Veröffentlicht: (2025) -
Low-Resource Neural Machine Translation Using Recurrent Neural Networks and Transfer Learning: A Case Study on English-to-Igbo
von: Ekle, Ocheme Anthony, et al.
Veröffentlicht: (2025)