Do LLMs Benefit From Their Own Words?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Jenny Y., Choshen, Leshem, Astudillo, Ramon, Broderick, Tamara, Andreas, Jacob |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Hitchhiker's Guide to Scaling Law Estimation
von: Choshen, Leshem, et al.
Veröffentlicht: (2024)
von: Choshen, Leshem, et al.
Veröffentlicht: (2024)
Beneath the Surface of Consistency: Exploring Cross-lingual Knowledge Representation Sharing in LLMs
von: Ifergan, Maxim, et al.
Veröffentlicht: (2024)
von: Ifergan, Maxim, et al.
Veröffentlicht: (2024)
Fuse to Forget: Bias Reduction and Selective Memorization through Model Fusion
von: Zaman, Kerem, et al.
Veröffentlicht: (2023)
von: Zaman, Kerem, et al.
Veröffentlicht: (2023)
Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
von: Damani, Mehul, et al.
Veröffentlicht: (2025)
von: Damani, Mehul, et al.
Veröffentlicht: (2025)
ComPEFT: Compression for Communicating Parameter Efficient Updates via Sparsification and Quantization
von: Yadav, Prateek, et al.
Veröffentlicht: (2023)
von: Yadav, Prateek, et al.
Veröffentlicht: (2023)
tinyBenchmarks: evaluating LLMs with fewer examples
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
ErrorMap and ErrorAtlas: Charting the Failure Landscape of Large Language Models
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2026)
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2026)
Training Language Models to Explain Their Own Computations
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
CRISP: Complex Reasoning with Interpretable Step-based Plans
von: Vetzler, Matan, et al.
Veröffentlicht: (2025)
von: Vetzler, Matan, et al.
Veröffentlicht: (2025)
Caught in the Web of Words: Do LLMs Fall for Spin in Medical Literature?
von: Yun, Hye Sun, et al.
Veröffentlicht: (2025)
von: Yun, Hye Sun, et al.
Veröffentlicht: (2025)
Can Gradient Descent Simulate Prompting?
von: Zhang, Eric, et al.
Veröffentlicht: (2025)
von: Zhang, Eric, et al.
Veröffentlicht: (2025)
Latency and Token-Aware Test-Time Compute
von: Huang, Jenny Y., et al.
Veröffentlicht: (2025)
von: Huang, Jenny Y., et al.
Veröffentlicht: (2025)
From Long to Short: LLMs Excel at Trimming Own Reasoning Chains
von: Han, Wei, et al.
Veröffentlicht: (2025)
von: Han, Wei, et al.
Veröffentlicht: (2025)
Robustness as an Emergent Property of Task Performance
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2026)
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2026)
Resolving Interference (RI): Disentangling Models for Improved Model Merging
von: Ramesh, Pratik, et al.
Veröffentlicht: (2026)
von: Ramesh, Pratik, et al.
Veröffentlicht: (2026)
Efficient multi-prompt evaluation of LLMs
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
From Tokens to Words: On the Inner Lexicon of LLMs
von: Kaplan, Guy, et al.
Veröffentlicht: (2024)
von: Kaplan, Guy, et al.
Veröffentlicht: (2024)
The ShareLM Collection and Plugin: Contributing Human-Model Chats for the Benefit of the Community
von: Don-Yehiya, Shachar, et al.
Veröffentlicht: (2024)
von: Don-Yehiya, Shachar, et al.
Veröffentlicht: (2024)
Visual Grounding Helps Learn Word Meanings in Low-Data Regimes
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2023)
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2023)
From Word to World: Evaluate and Mitigate Culture Bias in LLMs via Word Association Test
von: Dai, Xunlian, et al.
Veröffentlicht: (2025)
von: Dai, Xunlian, et al.
Veröffentlicht: (2025)
TextArena
von: Guertler, Leon, et al.
Veröffentlicht: (2025)
von: Guertler, Leon, et al.
Veröffentlicht: (2025)
Do LLMs Follow Their Own Rules? A Reflexive Audit of Self-Stated Safety Policies
von: Mittal, Avni
Veröffentlicht: (2026)
von: Mittal, Avni
Veröffentlicht: (2026)
Genie: Achieving Human Parity in Content-Grounded Datasets Generation
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs
von: Fan, Siqi, et al.
Veröffentlicht: (2025)
von: Fan, Siqi, et al.
Veröffentlicht: (2025)
From Words to Proverbs: Evaluating LLMs Linguistic and Cultural Competence in Saudi Dialects with Absher
von: Al-Monef, Renad, et al.
Veröffentlicht: (2025)
von: Al-Monef, Renad, et al.
Veröffentlicht: (2025)
Bring Your Own Prompts: Use-Case-Specific Bias and Fairness Evaluation for LLMs
von: Bouchard, Dylan
Veröffentlicht: (2024)
von: Bouchard, Dylan
Veröffentlicht: (2024)
A Survey on Model MoErging: Recycling and Routing Among Specialized Experts for Collaborative Learning
von: Yadav, Prateek, et al.
Veröffentlicht: (2024)
von: Yadav, Prateek, et al.
Veröffentlicht: (2024)
A Tale of Two Structures: Do LLMs Capture the Fractal Complexity of Language?
von: Alabdulmohsin, Ibrahim, et al.
Veröffentlicht: (2025)
von: Alabdulmohsin, Ibrahim, et al.
Veröffentlicht: (2025)
Do Large Language Models Understand Word Senses?
von: Meconi, Domenico, et al.
Veröffentlicht: (2025)
von: Meconi, Domenico, et al.
Veröffentlicht: (2025)
Creating a digital poet
von: Tohar, Vered, et al.
Veröffentlicht: (2026)
von: Tohar, Vered, et al.
Veröffentlicht: (2026)
Explain in Your Own Words: Improving Reasoning via Token-Selective Dual Knowledge Distillation
von: Kim, Minsang, et al.
Veröffentlicht: (2026)
von: Kim, Minsang, et al.
Veröffentlicht: (2026)
How Well do LLMs Compress Their Own Chain-of-Thought? A Token Complexity Approach
von: Lee, Ayeong, et al.
Veröffentlicht: (2025)
von: Lee, Ayeong, et al.
Veröffentlicht: (2025)
Compress then Serve: Serving Thousands of LoRA Adapters with Little Overhead
von: Brüel-Gabrielsson, Rickard, et al.
Veröffentlicht: (2024)
von: Brüel-Gabrielsson, Rickard, et al.
Veröffentlicht: (2024)
Word Form Matters: LLMs' Semantic Reconstruction under Typoglycemia
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
Unitxt: Flexible, Shareable and Reusable Data Preparation and Evaluation for Generative AI
von: Bandel, Elron, et al.
Veröffentlicht: (2024)
von: Bandel, Elron, et al.
Veröffentlicht: (2024)
Do LLMs Dream of Ontologies?
von: Bombieri, Marco, et al.
Veröffentlicht: (2024)
von: Bombieri, Marco, et al.
Veröffentlicht: (2024)
Latent Principle Discovery for Language Model Self-Improvement
von: Ramji, Keshav, et al.
Veröffentlicht: (2025)
von: Ramji, Keshav, et al.
Veröffentlicht: (2025)
Can LLMs Understand the Impact of Trauma? Costs and Benefits of LLMs Coding the Interviews of Firearm Violence Survivors
von: Zhu, Jessica H., et al.
Veröffentlicht: (2026)
von: Zhu, Jessica H., et al.
Veröffentlicht: (2026)
Beyond Words: A Latent Memory Approach to Internal Reasoning in LLMs
von: Orlicki, José I.
Veröffentlicht: (2025)
von: Orlicki, José I.
Veröffentlicht: (2025)
Deductive Closure Training of Language Models for Coherence, Accuracy, and Updatability
von: Akyürek, Afra Feyza, et al.
Veröffentlicht: (2024)
von: Akyürek, Afra Feyza, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Hitchhiker's Guide to Scaling Law Estimation
von: Choshen, Leshem, et al.
Veröffentlicht: (2024) -
Beneath the Surface of Consistency: Exploring Cross-lingual Knowledge Representation Sharing in LLMs
von: Ifergan, Maxim, et al.
Veröffentlicht: (2024) -
Fuse to Forget: Bias Reduction and Selective Memorization through Model Fusion
von: Zaman, Kerem, et al.
Veröffentlicht: (2023) -
Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
von: Damani, Mehul, et al.
Veröffentlicht: (2025) -
ComPEFT: Compression for Communicating Parameter Efficient Updates via Sparsification and Quantization
von: Yadav, Prateek, et al.
Veröffentlicht: (2023)