Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Puerto, Haritz, Tutek, Martin, Aditya, Somak, Zhu, Xiaodan, Gurevych, Iryna |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fine-Tuning on Diverse Reasoning Chains Drives Within-Inference CoT Refinement in LLMs
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
CATfOOD: Counterfactual Augmented Training for Improving Out-of-Domain Performance and Calibration
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2023)
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2023)
From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves
von: Puerto, Haritz, et al.
Veröffentlicht: (2026)
von: Puerto, Haritz, et al.
Veröffentlicht: (2026)
Robust Utility-Preserving Text Anonymization Based on Large Language Models
von: Yang, Tianyu, et al.
Veröffentlicht: (2024)
von: Yang, Tianyu, et al.
Veröffentlicht: (2024)
DARA: Decomposition-Alignment-Reasoning Autonomous Language Agent for Question Answering over Knowledge Graphs
von: Fang, Haishuo, et al.
Veröffentlicht: (2024)
von: Fang, Haishuo, et al.
Veröffentlicht: (2024)
PragWorld: A Benchmark Evaluating LLMs' Local World Model under Minimal Linguistic Alterations and Conversational Dynamics
von: Vashistha, Sachin, et al.
Veröffentlicht: (2025)
von: Vashistha, Sachin, et al.
Veröffentlicht: (2025)
SciCoQA: Quality Assurance for Scientific Paper--Code Alignment
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2026)
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2026)
SpaRC and SpaRP: Spatial Reasoning Characterization and Path Generation for Understanding Spatial Reasoning Capability of Large Language Models
von: Rizvi, Md Imbesat Hassan, et al.
Veröffentlicht: (2024)
von: Rizvi, Md Imbesat Hassan, et al.
Veröffentlicht: (2024)
Preemptive Detection and Correction of Misaligned Actions in LLM Agents
von: Fang, Haishuo, et al.
Veröffentlicht: (2024)
von: Fang, Haishuo, et al.
Veröffentlicht: (2024)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
Socratic Reasoning Improves Positive Text Rewriting
von: Goel, Anmol, et al.
Veröffentlicht: (2024)
von: Goel, Anmol, et al.
Veröffentlicht: (2024)
Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions
von: Hong, Pengfei, et al.
Veröffentlicht: (2024)
von: Hong, Pengfei, et al.
Veröffentlicht: (2024)
Like a Good Nearest Neighbor: Practical Content Moderation and Text Classification
von: Bates, Luke, et al.
Veröffentlicht: (2023)
von: Bates, Luke, et al.
Veröffentlicht: (2023)
Systematic Task Exploration with LLMs: A Study in Citation Text Generation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2024)
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2024)
Automatic Reviewers Fail to Detect Faulty Reasoning in Research Papers: A New Counterfactual Evaluation Framework
von: Dycke, Nils, et al.
Veröffentlicht: (2025)
von: Dycke, Nils, et al.
Veröffentlicht: (2025)
How are Prompts Different in Terms of Sensitivity?
von: Lu, Sheng, et al.
Veröffentlicht: (2023)
von: Lu, Sheng, et al.
Veröffentlicht: (2023)
ChartAttack: Testing the Vulnerability of LLMs to Malicious Prompting in Chart Generation
von: Ortiz-Barajas, Jesus-German, et al.
Veröffentlicht: (2026)
von: Ortiz-Barajas, Jesus-German, et al.
Veröffentlicht: (2026)
SPARE: Single-Pass Annotation with Reference-Guided Evaluation for Automatic Process Supervision and Reward Modelling
von: Rizvi, Md Imbesat Hassan, et al.
Veröffentlicht: (2025)
von: Rizvi, Md Imbesat Hassan, et al.
Veröffentlicht: (2025)
Are Multilingual LLMs Culturally-Diverse Reasoners? An Investigation into Multicultural Proverbs and Sayings
von: Liu, Chen Cecilia, et al.
Veröffentlicht: (2023)
von: Liu, Chen Cecilia, et al.
Veröffentlicht: (2023)
DOCE: Finding the Sweet Spot for Execution-Based Code Generation
von: Li, Haau-Sing, et al.
Veröffentlicht: (2024)
von: Li, Haau-Sing, et al.
Veröffentlicht: (2024)
MuCRASP: Multimodal Chain-of-thought Reasoning aware Structured Pruning
von: Dutta, Aritra, et al.
Veröffentlicht: (2026)
von: Dutta, Aritra, et al.
Veröffentlicht: (2026)
[WIP] Jailbreak Paradox: The Achilles' Heel of LLMs
von: Rao, Abhinav, et al.
Veröffentlicht: (2024)
von: Rao, Abhinav, et al.
Veröffentlicht: (2024)
Eliciting Better Multilingual Structured Reasoning from LLMs through Code
von: Li, Bryan, et al.
Veröffentlicht: (2024)
von: Li, Bryan, et al.
Veröffentlicht: (2024)
Are Emergent Abilities in Large Language Models just In-Context Learning?
von: Lu, Sheng, et al.
Veröffentlicht: (2023)
von: Lu, Sheng, et al.
Veröffentlicht: (2023)
STRICTA: Structured Reasoning in Critical Text Assessment for Peer Review and Beyond
von: Dycke, Nils, et al.
Veröffentlicht: (2024)
von: Dycke, Nils, et al.
Veröffentlicht: (2024)
Conditioning LLMs to Generate Code-Switched Text
von: Heredia, Maite, et al.
Veröffentlicht: (2025)
von: Heredia, Maite, et al.
Veröffentlicht: (2025)
Citation Failure: Definition, Analysis and Efficient Mitigation
von: Buchmann, Jan, et al.
Veröffentlicht: (2025)
von: Buchmann, Jan, et al.
Veröffentlicht: (2025)
The Inherent Limits of Pretrained LLMs: The Unexpected Convergence of Instruction Tuning and In-Context Learning Capabilities
von: Bigoulaeva, Irina, et al.
Veröffentlicht: (2025)
von: Bigoulaeva, Irina, et al.
Veröffentlicht: (2025)
Sensitivity, Performance, Robustness: Deconstructing the Effect of Sociodemographic Prompting
von: Beck, Tilman, et al.
Veröffentlicht: (2023)
von: Beck, Tilman, et al.
Veröffentlicht: (2023)
The Lou Dataset -- Exploring the Impact of Gender-Fair Language in German Text Classification
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
Enabling Natural Zero-Shot Prompting on Encoder Models via Statement-Tuning
von: Elshabrawy, Ahmed, et al.
Veröffentlicht: (2024)
von: Elshabrawy, Ahmed, et al.
Veröffentlicht: (2024)
ObscuraCoder: Powering Efficient Code LM Pre-Training Via Obfuscation Grounding
von: Paul, Indraneil, et al.
Veröffentlicht: (2025)
von: Paul, Indraneil, et al.
Veröffentlicht: (2025)
Chain-of-Code Collapse: Reasoning Failures in LLMs via Adversarial Prompting in Code Generation
von: Roh, Jaechul, et al.
Veröffentlicht: (2025)
von: Roh, Jaechul, et al.
Veröffentlicht: (2025)
Leaky Thoughts: Large Reasoning Models Are Not Private Thinkers
von: Green, Tommaso, et al.
Veröffentlicht: (2025)
von: Green, Tommaso, et al.
Veröffentlicht: (2025)
On Code-Induced Reasoning in LLMs
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning
von: Das, Debrup, et al.
Veröffentlicht: (2024)
von: Das, Debrup, et al.
Veröffentlicht: (2024)
TEXT2AFFORD: Probing Object Affordance Prediction abilities of Language Models solely from Text
von: Adak, Sayantan, et al.
Veröffentlicht: (2024)
von: Adak, Sayantan, et al.
Veröffentlicht: (2024)
LLMs as Cultural Archives: Cultural Commonsense Knowledge Graph Extraction
von: Tonga, Junior Cedric, et al.
Veröffentlicht: (2026)
von: Tonga, Junior Cedric, et al.
Veröffentlicht: (2026)
CORE-T: COherent REtrieval of Tables for Text-to-SQL
von: Soliman, Hassan, et al.
Veröffentlicht: (2026)
von: Soliman, Hassan, et al.
Veröffentlicht: (2026)
Commitment Checklist: Auditing Author Commitments in Peer Review
von: Chen, Chung-Chi, et al.
Veröffentlicht: (2026)
von: Chen, Chung-Chi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Fine-Tuning on Diverse Reasoning Chains Drives Within-Inference CoT Refinement in LLMs
von: Puerto, Haritz, et al.
Veröffentlicht: (2024) -
CATfOOD: Counterfactual Augmented Training for Improving Out-of-Domain Performance and Calibration
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2023) -
From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves
von: Puerto, Haritz, et al.
Veröffentlicht: (2026) -
Robust Utility-Preserving Text Anonymization Based on Large Language Models
von: Yang, Tianyu, et al.
Veröffentlicht: (2024) -
DARA: Decomposition-Alignment-Reasoning Autonomous Language Agent for Question Answering over Knowledge Graphs
von: Fang, Haishuo, et al.
Veröffentlicht: (2024)