Can LLMs Compute with Reasons?
Fuente:
arXiv
Salvato in:
| Autori principali: | Sandilya, Harshit, Raj, Peehu, Bafna, Jainit Sushil, Mukhopadhyay, Srija, Sharma, Shivansh, Sharma, Ellwil, Sharma, Arastu, Trivedi, Neeta, Shrivastava, Manish, Kumar, Rajesh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Eradicating Negative Transfer in Multi-Physics Foundation Models via Sparse Mixture-of-Experts Routing
di: Sharma, Ellwil, et al.
Pubblicazione: (2026)
di: Sharma, Ellwil, et al.
Pubblicazione: (2026)
Bug In the Code Stack: Can LLMs Find Bugs in Large Python Code Stacks
di: Lee, Hokyung, et al.
Pubblicazione: (2024)
di: Lee, Hokyung, et al.
Pubblicazione: (2024)
Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs
di: Balter, Samuel G., et al.
Pubblicazione: (2026)
di: Balter, Samuel G., et al.
Pubblicazione: (2026)
A Case Study of Balanced Query Recommendation on Wikipedia
di: Mishra, Harshit, et al.
Pubblicazione: (2025)
di: Mishra, Harshit, et al.
Pubblicazione: (2025)
Evaluation of Table Representations to Answer Questions from Tables in Documents : A Case Study using 3GPP Specifications
di: Roychowdhury, Sujoy, et al.
Pubblicazione: (2024)
di: Roychowdhury, Sujoy, et al.
Pubblicazione: (2024)
Can Watermarked LLMs be Identified by Users via Crafted Prompts?
di: Liu, Aiwei, et al.
Pubblicazione: (2024)
di: Liu, Aiwei, et al.
Pubblicazione: (2024)
Prompt Engineering and the Effectiveness of Large Language Models in Enhancing Human Productivity
di: Anam, Rizal Khoirul
Pubblicazione: (2025)
di: Anam, Rizal Khoirul
Pubblicazione: (2025)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
The Personalization Trap: How User Memory Alters Emotional Reasoning in LLMs
di: Fang, Xi, et al.
Pubblicazione: (2025)
di: Fang, Xi, et al.
Pubblicazione: (2025)
Task Complexity Matters: An Empirical Study of Reasoning in LLMs for Sentiment Analysis
di: Huang, Donghao, et al.
Pubblicazione: (2026)
di: Huang, Donghao, et al.
Pubblicazione: (2026)
MORABLES: A Benchmark for Assessing Abstract Moral Reasoning in LLMs with Fables
di: Marcuzzo, Matteo, et al.
Pubblicazione: (2025)
di: Marcuzzo, Matteo, et al.
Pubblicazione: (2025)
LLMs for Legal Subsumption in German Employment Contracts
di: Wardas, Oliver, et al.
Pubblicazione: (2025)
di: Wardas, Oliver, et al.
Pubblicazione: (2025)
PropXplain: Can LLMs Enable Explainable Propaganda Detection?
di: Hasanain, Maram, et al.
Pubblicazione: (2025)
di: Hasanain, Maram, et al.
Pubblicazione: (2025)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
di: Yang, Yibo
Pubblicazione: (2025)
di: Yang, Yibo
Pubblicazione: (2025)
When Retrieval Succeeds and Fails: Rethinking Retrieval-Augmented Generation for LLMs
di: Wang, Yongjie, et al.
Pubblicazione: (2025)
di: Wang, Yongjie, et al.
Pubblicazione: (2025)
Distractor Injection Attacks on Large Reasoning Models: Characterization and Defense
di: Zhang, Zhehao, et al.
Pubblicazione: (2025)
di: Zhang, Zhehao, et al.
Pubblicazione: (2025)
The Unlikely Duel: Evaluating Creative Writing in LLMs through a Unique Scenario
di: Gómez-Rodríguez, Carlos, et al.
Pubblicazione: (2024)
di: Gómez-Rodríguez, Carlos, et al.
Pubblicazione: (2024)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
di: Fang, Xi, et al.
Pubblicazione: (2024)
di: Fang, Xi, et al.
Pubblicazione: (2024)
New Skills or Sharper Primitives? A Probabilistic Perspective on the Emergence of Reasoning in RLVR
di: Wang, Zhilin, et al.
Pubblicazione: (2026)
di: Wang, Zhilin, et al.
Pubblicazione: (2026)
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
di: Pranida, Salsabila Zahirah, et al.
Pubblicazione: (2025)
di: Pranida, Salsabila Zahirah, et al.
Pubblicazione: (2025)
Charting a Decade of Computational Linguistics in Italy: The CLiC-it Corpus
di: Alzetta, Chiara, et al.
Pubblicazione: (2025)
di: Alzetta, Chiara, et al.
Pubblicazione: (2025)
Challenges and Applications of Large Language Models: A Comparison of GPT and DeepSeek family of models
di: Sharma, Shubham, et al.
Pubblicazione: (2025)
di: Sharma, Shubham, et al.
Pubblicazione: (2025)
Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs
di: Hill, Brennen, et al.
Pubblicazione: (2025)
di: Hill, Brennen, et al.
Pubblicazione: (2025)
Do LLMs have a Gender (Entropy) Bias?
di: Prabhune, Sonal, et al.
Pubblicazione: (2025)
di: Prabhune, Sonal, et al.
Pubblicazione: (2025)
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
di: Miliani, Martina, et al.
Pubblicazione: (2025)
di: Miliani, Martina, et al.
Pubblicazione: (2025)
Rethinking the Multilingual Reasoning Gap with Layer Swap
di: Lasbordes, Maxence, et al.
Pubblicazione: (2026)
di: Lasbordes, Maxence, et al.
Pubblicazione: (2026)
How much do LLMs learn from negative examples?
di: Hamdan, Shadi, et al.
Pubblicazione: (2025)
di: Hamdan, Shadi, et al.
Pubblicazione: (2025)
Communicative Agents for Slideshow Storytelling Video Generation based on LLMs
di: Fan, Jingxing, et al.
Pubblicazione: (2025)
di: Fan, Jingxing, et al.
Pubblicazione: (2025)
LLM-Assisted Crisis Management: Building Advanced LLM Platforms for Effective Emergency Response and Public Collaboration
di: Otal, Hakan T., et al.
Pubblicazione: (2024)
di: Otal, Hakan T., et al.
Pubblicazione: (2024)
A Study into Investigating Temporal Robustness of LLMs
di: Wallat, Jonas, et al.
Pubblicazione: (2025)
di: Wallat, Jonas, et al.
Pubblicazione: (2025)
MicroRemed: Benchmarking LLMs in Microservices Remediation
di: Zhang, Lingzhe, et al.
Pubblicazione: (2025)
di: Zhang, Lingzhe, et al.
Pubblicazione: (2025)
Pun Unintended: LLMs and the Illusion of Humor Understanding
di: Zangari, Alessandro, et al.
Pubblicazione: (2025)
di: Zangari, Alessandro, et al.
Pubblicazione: (2025)
Tethered Reasoning: Decoupling Entropy from Hallucination in Quantized LLMs via Manifold Steering
di: Atkinson, Craig
Pubblicazione: (2026)
di: Atkinson, Craig
Pubblicazione: (2026)
From Noise to Diversity: Random Embedding Injection in LLM Reasoning
di: Kim, Heejun, et al.
Pubblicazione: (2026)
di: Kim, Heejun, et al.
Pubblicazione: (2026)
RHealthTwin: Towards Responsible and Multimodal Digital Twins for Personalized Well-being
di: Ferdousi, Rahatara, et al.
Pubblicazione: (2025)
di: Ferdousi, Rahatara, et al.
Pubblicazione: (2025)
Towards Probabilistic Question Answering Over Tabular Data
di: Shen, Chen, et al.
Pubblicazione: (2025)
di: Shen, Chen, et al.
Pubblicazione: (2025)
SpokenNativQA: Multilingual Everyday Spoken Queries for LLMs
di: Alam, Firoj, et al.
Pubblicazione: (2025)
di: Alam, Firoj, et al.
Pubblicazione: (2025)
AVEC: Bootstrapping Privacy for Local LLMs
di: Gaikwad, Madhava
Pubblicazione: (2025)
di: Gaikwad, Madhava
Pubblicazione: (2025)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
di: Orgad, Hadas, et al.
Pubblicazione: (2024)
di: Orgad, Hadas, et al.
Pubblicazione: (2024)
HR-MultiWOZ: A Task Oriented Dialogue (TOD) Dataset for HR LLM Agent
di: Xu, Weijie, et al.
Pubblicazione: (2024)
di: Xu, Weijie, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Eradicating Negative Transfer in Multi-Physics Foundation Models via Sparse Mixture-of-Experts Routing
di: Sharma, Ellwil, et al.
Pubblicazione: (2026) -
Bug In the Code Stack: Can LLMs Find Bugs in Large Python Code Stacks
di: Lee, Hokyung, et al.
Pubblicazione: (2024) -
Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs
di: Balter, Samuel G., et al.
Pubblicazione: (2026) -
A Case Study of Balanced Query Recommendation on Wikipedia
di: Mishra, Harshit, et al.
Pubblicazione: (2025) -
Evaluation of Table Representations to Answer Questions from Tables in Documents : A Case Study using 3GPP Specifications
di: Roychowdhury, Sujoy, et al.
Pubblicazione: (2024)