PEA: Enhancing LLM Performance on Computational-Reasoning Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Zi, Weng, Shiwei, Alhanahnah, Mohannad, Jha, Somesh, Reps, Tom |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rethinking Diversity in Deep Neural Network Testing
von: Wang, Zi, et al.
Veröffentlicht: (2023)
von: Wang, Zi, et al.
Veröffentlicht: (2023)
Functional Homotopy: Smoothing Discrete Optimization via Continuous Parameters for LLM Jailbreak Attacks
von: Wang, Zi, et al.
Veröffentlicht: (2024)
von: Wang, Zi, et al.
Veröffentlicht: (2024)
DepsRAG: Towards Agentic Reasoning and Planning for Software Dependency Management
von: Alhanahnah, Mohannad, et al.
Veröffentlicht: (2024)
von: Alhanahnah, Mohannad, et al.
Veröffentlicht: (2024)
How Not to Detect Prompt Injections with an LLM
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
PolicyBank: Evolving Policy Understanding for LLM Agents
von: Choi, Jihye, et al.
Veröffentlicht: (2026)
von: Choi, Jihye, et al.
Veröffentlicht: (2026)
Data-to-Dashboard: Multi-Agent LLM Framework for Insightful Visualization in Enterprise Analytics
von: Zhang, Ran, et al.
Veröffentlicht: (2025)
von: Zhang, Ran, et al.
Veröffentlicht: (2025)
A New Era in LLM Security: Exploring Security Concerns in Real-World LLM-based Systems
von: Wu, Fangzhou, et al.
Veröffentlicht: (2024)
von: Wu, Fangzhou, et al.
Veröffentlicht: (2024)
Demystifying the Roles of LLM Layers in Retrieval, Knowledge, and Reasoning
von: Song, Xinyuan, et al.
Veröffentlicht: (2025)
von: Song, Xinyuan, et al.
Veröffentlicht: (2025)
ATLAS: Constraints-Aware Multi-Agent Collaboration for Real-World Travel Planning
von: Choi, Jihye, et al.
Veröffentlicht: (2025)
von: Choi, Jihye, et al.
Veröffentlicht: (2025)
Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning
von: Bi, Zhenni, et al.
Veröffentlicht: (2024)
von: Bi, Zhenni, et al.
Veröffentlicht: (2024)
Adaptive Concept Bottleneck for Foundation Models Under Distribution Shifts
von: Choi, Jihye, et al.
Veröffentlicht: (2024)
von: Choi, Jihye, et al.
Veröffentlicht: (2024)
SACTOR: LLM-Driven Correct and Idiomatic C to Rust Translation with Static Analysis and FFI-Based Verification
von: Zhou, Tianyang, et al.
Veröffentlicht: (2025)
von: Zhou, Tianyang, et al.
Veröffentlicht: (2025)
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
SoK: Software Debloating Landscape and Future Directions
von: Alhanahnah, Mohannad, et al.
Veröffentlicht: (2024)
von: Alhanahnah, Mohannad, et al.
Veröffentlicht: (2024)
Impact of Noise on LLM-Models Performance in Abstraction and Reasoning Corpus (ARC) Tasks with Model Temperature Considerations
von: Khandalkar, Nikhil, et al.
Veröffentlicht: (2025)
von: Khandalkar, Nikhil, et al.
Veröffentlicht: (2025)
MHA-RAG: Improving Efficiency, Accuracy, and Consistency by Encoding Exemplars as Soft Prompts
von: Jain, Abhinav, et al.
Veröffentlicht: (2025)
von: Jain, Abhinav, et al.
Veröffentlicht: (2025)
LLM & HPC:Benchmarking DeepSeek's Performance in High-Performance Computing Tasks
von: Nader, Noujoud, et al.
Veröffentlicht: (2025)
von: Nader, Noujoud, et al.
Veröffentlicht: (2025)
Auto-SPT: Automating Semantic Preserving Transformations for Code
von: Hooda, Ashish, et al.
Veröffentlicht: (2025)
von: Hooda, Ashish, et al.
Veröffentlicht: (2025)
Multi-Task GRPO: Reliable LLM Reasoning Across Tasks
von: Ramesh, Shyam Sundhar, et al.
Veröffentlicht: (2026)
von: Ramesh, Shyam Sundhar, et al.
Veröffentlicht: (2026)
Beyond Memorization: Testing LLM Reasoning on Unseen Theory of Computation Tasks
von: Shelat, Shlok, et al.
Veröffentlicht: (2026)
von: Shelat, Shlok, et al.
Veröffentlicht: (2026)
Bag of Tricks for Inference-time Computation of LLM Reasoning
von: Liu, Fan, et al.
Veröffentlicht: (2025)
von: Liu, Fan, et al.
Veröffentlicht: (2025)
Extending Token Computation for LLM Reasoning
von: Liao, Bingli, et al.
Veröffentlicht: (2024)
von: Liao, Bingli, et al.
Veröffentlicht: (2024)
Enhance Reasoning for Large Language Models in the Game Werewolf
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
Prompt Tuning Strikes Back: Customizing Foundation Models with Low-Rank Prompt Adaptation
von: Jain, Abhinav, et al.
Veröffentlicht: (2024)
von: Jain, Abhinav, et al.
Veröffentlicht: (2024)
Reason from Future: Reverse Thought Chain Enhances LLM Reasoning
von: Xu, Yinlong, et al.
Veröffentlicht: (2025)
von: Xu, Yinlong, et al.
Veröffentlicht: (2025)
Effective Learning for Small Reasoning Models: An Empirical Study on 0.5B Reasoning LLMs
von: Zhuang, Xialie, et al.
Veröffentlicht: (2025)
von: Zhuang, Xialie, et al.
Veröffentlicht: (2025)
HyperTree Planning: Enhancing LLM Reasoning via Hierarchical Thinking
von: Gui, Runquan, et al.
Veröffentlicht: (2025)
von: Gui, Runquan, et al.
Veröffentlicht: (2025)
Formal Policy Enforcement for Real-World Agentic Systems
von: Palumbo, Nils, et al.
Veröffentlicht: (2026)
von: Palumbo, Nils, et al.
Veröffentlicht: (2026)
Debate Only When Necessary: Adaptive Multiagent Collaboration for Efficient LLM Reasoning
von: Eo, Sugyeong, et al.
Veröffentlicht: (2025)
von: Eo, Sugyeong, et al.
Veröffentlicht: (2025)
CPL: Critical Plan Step Learning Boosts LLM Generalization in Reasoning Tasks
von: Wang, Tianlong, et al.
Veröffentlicht: (2024)
von: Wang, Tianlong, et al.
Veröffentlicht: (2024)
Lost-in-Distance: Impact of Contextual Proximity on LLM Performance in Graph Tasks
von: Firooz, Hamed, et al.
Veröffentlicht: (2024)
von: Firooz, Hamed, et al.
Veröffentlicht: (2024)
LLM-guided Task and Motion Planning using Knowledge-based Reasoning
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2024)
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2024)
DNN Task Assignment in UAV Networks: A Generative AI Enhanced Multi-Agent Reinforcement Learning Approach
von: Tang, Xin, et al.
Veröffentlicht: (2024)
von: Tang, Xin, et al.
Veröffentlicht: (2024)
Prolonged Reasoning Is Not All You Need: Certainty-Based Adaptive Routing for Efficient LLM/MLLM Reasoning
von: Lu, Jinghui, et al.
Veröffentlicht: (2025)
von: Lu, Jinghui, et al.
Veröffentlicht: (2025)
MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for Pharmacovigilance
von: Choi, Jihye, et al.
Veröffentlicht: (2024)
von: Choi, Jihye, et al.
Veröffentlicht: (2024)
Undetectable Backdoors in Model Parameters: Hiding Sparse Secrets in High Dimensions
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2026)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2026)
Enhancing LLM Reasoning with Reward-guided Tree Search
von: Jiang, Jinhao, et al.
Veröffentlicht: (2024)
von: Jiang, Jinhao, et al.
Veröffentlicht: (2024)
Batch-of-Thought: Cross-Instance Learning for Enhanced LLM Reasoning
von: Yang, Xuan, et al.
Veröffentlicht: (2026)
von: Yang, Xuan, et al.
Veröffentlicht: (2026)
Let's Reason Formally: Natural-Formal Hybrid Reasoning Enhances LLM's Math Capability
von: Wang, Ruida, et al.
Veröffentlicht: (2025)
von: Wang, Ruida, et al.
Veröffentlicht: (2025)
Exploring Spatial Representation to Enhance LLM Reasoning in Aerial Vision-Language Navigation
von: Gao, Yunpeng, et al.
Veröffentlicht: (2024)
von: Gao, Yunpeng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Rethinking Diversity in Deep Neural Network Testing
von: Wang, Zi, et al.
Veröffentlicht: (2023) -
Functional Homotopy: Smoothing Discrete Optimization via Continuous Parameters for LLM Jailbreak Attacks
von: Wang, Zi, et al.
Veröffentlicht: (2024) -
DepsRAG: Towards Agentic Reasoning and Planning for Software Dependency Management
von: Alhanahnah, Mohannad, et al.
Veröffentlicht: (2024) -
How Not to Detect Prompt Injections with an LLM
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025) -
PolicyBank: Evolving Policy Understanding for LLM Agents
von: Choi, Jihye, et al.
Veröffentlicht: (2026)