Exploring the Impact of Temperature on Large Language Models:Hot or Cold?
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Lujun, Sleem, Lama, Gentile, Niccolo', Nichil, Geoffrey, State, Radu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Small Language Models in the Real World: Insights from Industrial Text Classification
by: Li, Lujun, et al.
Published: (2025)
by: Li, Lujun, et al.
Published: (2025)
Is Small Language Model the Silver Bullet to Low-Resource Languages Machine Translation?
by: Song, Yewei, et al.
Published: (2025)
by: Song, Yewei, et al.
Published: (2025)
Do Large Language Models Grasp The Grammar? Evidence from Grammar-Book-Guided Probing in Luxembourgish
by: Li, Lujun, et al.
Published: (2025)
by: Li, Lujun, et al.
Published: (2025)
The Necessity of Setting Temperature in LLM-as-a-Judge
by: Li, Lujun, et al.
Published: (2026)
by: Li, Lujun, et al.
Published: (2026)
Uncovering Zero-Shot Generalization Gaps in Time-Series Foundation Models Using Real-World Videos
by: Li, Lujun, et al.
Published: (2025)
by: Li, Lujun, et al.
Published: (2025)
NegBLEURT Forest: Leveraging Inconsistencies for Detecting Jailbreak Attacks
by: Sleem, Lama, et al.
Published: (2025)
by: Sleem, Lama, et al.
Published: (2025)
Agent Skill Framework: Perspectives on the Potential of Small Language Models in Industrial Environments
by: Xu, Yangjie, et al.
Published: (2026)
by: Xu, Yangjie, et al.
Published: (2026)
Vision Transformer-Based Time-Series Image Reconstruction for Cloud-Filling Applications
by: Li, Lujun, et al.
Published: (2025)
by: Li, Lujun, et al.
Published: (2025)
HalluGuard: Evidence-Grounded Small Reasoning Models to Mitigate Hallucinations in Retrieval-Augmented Generation
by: Bergeron, Loris, et al.
Published: (2025)
by: Bergeron, Loris, et al.
Published: (2025)
Temporal-Spatial Tubelet Embedding for Cloud-Robust MSI Reconstruction using MSI-SAR Fusion: A Multi-Head Self-Attention Video Vision Transformer Approach
by: Wang, Yiqun, et al.
Published: (2025)
by: Wang, Yiqun, et al.
Published: (2025)
How Much Does Persuasion Strategy Matter? LLM-Annotated Evidence from Charitable Donation Dialogues
by: Petrova, Tatiana, et al.
Published: (2026)
by: Petrova, Tatiana, et al.
Published: (2026)
On the Limitations of Large Language Models (LLMs): False Attribution
by: Adewumi, Tosin, et al.
Published: (2024)
by: Adewumi, Tosin, et al.
Published: (2024)
Piloting Copilot, Codex, and StarCoder2: Hot Temperature, Cold Prompts, or Black Magic?
by: Döderlein, Jean-Baptiste, et al.
Published: (2022)
by: Döderlein, Jean-Baptiste, et al.
Published: (2022)
Limitations of Normalization in Attention Mechanism
by: Mudarisov, Timur, et al.
Published: (2025)
by: Mudarisov, Timur, et al.
Published: (2025)
Language Model Uncertainty Quantification with Attention Chain
by: Li, Yinghao, et al.
Published: (2025)
by: Li, Yinghao, et al.
Published: (2025)
Shaping Explanations: Semantic Reward Modeling with Encoder-Only Transformers for GRPO
by: Pappone, Francesco, et al.
Published: (2025)
by: Pappone, Francesco, et al.
Published: (2025)
Pruner-Zero: Evolving Symbolic Pruning Metric from scratch for Large Language Models
by: Dong, Peijie, et al.
Published: (2024)
by: Dong, Peijie, et al.
Published: (2024)
Exploring the Translation Mechanism of Large Language Models
by: Zhang, Hongbin, et al.
Published: (2025)
by: Zhang, Hongbin, et al.
Published: (2025)
Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?
by: Geigle, Gregor, et al.
Published: (2024)
by: Geigle, Gregor, et al.
Published: (2024)
SaudiCulture: A Benchmark for Evaluating Large Language Models Cultural Competence within Saudi Arabia
by: Ayash, Lama, et al.
Published: (2025)
by: Ayash, Lama, et al.
Published: (2025)
Exploring Large Language Models for Translating Romanian Computational Problems into English
by: Dumitran, Adrian Marius, et al.
Published: (2025)
by: Dumitran, Adrian Marius, et al.
Published: (2025)
LPZero: Language Model Zero-cost Proxy Search from Zero
by: Dong, Peijie, et al.
Published: (2024)
by: Dong, Peijie, et al.
Published: (2024)
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
by: Geigle, Gregor, et al.
Published: (2024)
by: Geigle, Gregor, et al.
Published: (2024)
From Text to Multimodality: Exploring the Evolution and Impact of Large Language Models in Medical Practice
by: Niu, Qian, et al.
Published: (2024)
by: Niu, Qian, et al.
Published: (2024)
Exploring Mathematical Extrapolation of Large Language Models with Synthetic Data
by: Li, Haolong, et al.
Published: (2024)
by: Li, Haolong, et al.
Published: (2024)
Exploring the Impact of Personality Traits on Conversational Recommender Systems: A Simulation with Large Language Models
by: Zhao, Xiaoyan, et al.
Published: (2025)
by: Zhao, Xiaoyan, et al.
Published: (2025)
Exploring the Impact of Corpus Diversity on Financial Pretrained Language Models
by: Choe, Jaeyoung, et al.
Published: (2023)
by: Choe, Jaeyoung, et al.
Published: (2023)
Blockly2Hooks: Smart Contracts for Everyone with the XRP Ledger and Google Blockly
by: Trestioreanu, Lucian, et al.
Published: (2025)
by: Trestioreanu, Lucian, et al.
Published: (2025)
Exploring and Mitigating Fawning Hallucinations in Large Language Models
by: Shangguan, Zixuan, et al.
Published: (2025)
by: Shangguan, Zixuan, et al.
Published: (2025)
Exploring Forgetting in Large Language Model Pre-Training
by: Liao, Chonghua, et al.
Published: (2024)
by: Liao, Chonghua, et al.
Published: (2024)
Exploring Group and Symmetry Principles in Large Language Models
by: Imani, Shima, et al.
Published: (2024)
by: Imani, Shima, et al.
Published: (2024)
Exploring the Potential of Large Language Models in Computational Argumentation
by: Chen, Guizhen, et al.
Published: (2023)
by: Chen, Guizhen, et al.
Published: (2023)
ProCoT: Stimulating Critical Thinking and Writing of Students through Engagement with Large Language Models (LLMs)
by: Adewumi, Tosin, et al.
Published: (2023)
by: Adewumi, Tosin, et al.
Published: (2023)
Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf
by: Xu, Yuzhuang, et al.
Published: (2023)
by: Xu, Yuzhuang, et al.
Published: (2023)
Exploring the Reliability of Large Language Models as Customized Evaluators for Diverse NLP Tasks
by: Li, Qintong, et al.
Published: (2023)
by: Li, Qintong, et al.
Published: (2023)
Is Temperature the Creativity Parameter of Large Language Models?
by: Peeperkorn, Max, et al.
Published: (2024)
by: Peeperkorn, Max, et al.
Published: (2024)
Paraphrase and Solve: Exploring and Exploiting the Impact of Surface Form on Mathematical Reasoning in Large Language Models
by: Zhou, Yue, et al.
Published: (2024)
by: Zhou, Yue, et al.
Published: (2024)
Large Language Models Explore by Latent Distilling
by: Zeng, Yuanhao, et al.
Published: (2026)
by: Zeng, Yuanhao, et al.
Published: (2026)
EmoVerse: Exploring Multimodal Large Language Models for Sentiment and Emotion Understanding
by: Li, Ao, et al.
Published: (2024)
by: Li, Ao, et al.
Published: (2024)
Unveiling Imitation Learning: Exploring the Impact of Data Falsity to Large Language Model
by: Cho, Hyunsoo
Published: (2024)
by: Cho, Hyunsoo
Published: (2024)
Similar Items
-
Small Language Models in the Real World: Insights from Industrial Text Classification
by: Li, Lujun, et al.
Published: (2025) -
Is Small Language Model the Silver Bullet to Low-Resource Languages Machine Translation?
by: Song, Yewei, et al.
Published: (2025) -
Do Large Language Models Grasp The Grammar? Evidence from Grammar-Book-Guided Probing in Luxembourgish
by: Li, Lujun, et al.
Published: (2025) -
The Necessity of Setting Temperature in LLM-as-a-Judge
by: Li, Lujun, et al.
Published: (2026) -
Uncovering Zero-Shot Generalization Gaps in Time-Series Foundation Models Using Real-World Videos
by: Li, Lujun, et al.
Published: (2025)