Quantum-Audit: Evaluating the Reasoning Limits of LLMs on Quantum Computing
Fuente:
arXiv
Salvato in:
| Autori principali: | Afane, Mohamed, Laufer, Kayla, Wei, Wenqi, Mao, Ying, Farooq, Junaid, Wang, Ying, Chen, Juntao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Differentiable Architecture Search for Adversarially Robust Quantum Computer Vision
di: Afane, Mohamed, et al.
Pubblicazione: (2026)
di: Afane, Mohamed, et al.
Pubblicazione: (2026)
Next-Generation Phishing: How LLM Agents Empower Cyber Attackers
di: Afane, Khalifa, et al.
Pubblicazione: (2024)
di: Afane, Khalifa, et al.
Pubblicazione: (2024)
ATP: Adaptive Threshold Pruning for Efficient Data Encoding in Quantum Neural Networks
di: Afane, Mohamed, et al.
Pubblicazione: (2025)
di: Afane, Mohamed, et al.
Pubblicazione: (2025)
SCOUT: A Defense Against Data Poisoning Attacks in Fine-Tuned Language Models
di: Afane, Mohamed, et al.
Pubblicazione: (2025)
di: Afane, Mohamed, et al.
Pubblicazione: (2025)
Can LLMs Help Allocate Public Health Resources? A Case Study on Childhood Lead Testing
di: Afane, Mohamed, et al.
Pubblicazione: (2025)
di: Afane, Mohamed, et al.
Pubblicazione: (2025)
SQUASH: A SWAP-Based Quantum Attack to Sabotage Hybrid Quantum Neural Networks
di: Kumar, Rahul, et al.
Pubblicazione: (2025)
di: Kumar, Rahul, et al.
Pubblicazione: (2025)
Analyzing and Optimizing the Distribution of Blood Lead Level Testing for Children in New York City: A Data-Driven Approach
di: Afane, Mohamed, et al.
Pubblicazione: (2025)
di: Afane, Mohamed, et al.
Pubblicazione: (2025)
Benchmarking Legal RAG: The Promise and Limits of AI Statutory Surveys
di: Afane, Mohamed, et al.
Pubblicazione: (2026)
di: Afane, Mohamed, et al.
Pubblicazione: (2026)
Symbolic Specification and Reasoning for Quantum Data and Operations
di: Ying, Mingsheng
Pubblicazione: (2025)
di: Ying, Mingsheng
Pubblicazione: (2025)
Reasoning Beyond Limits: Advances and Open Problems for LLMs
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025)
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025)
Understanding and Patching Compositional Reasoning in LLMs
di: Li, Zhaoyi, et al.
Pubblicazione: (2024)
di: Li, Zhaoyi, et al.
Pubblicazione: (2024)
Efficient Reasoning for LLMs through Speculative Chain-of-Thought
di: Wang, Jikai, et al.
Pubblicazione: (2025)
di: Wang, Jikai, et al.
Pubblicazione: (2025)
Hardware-aware Circuit Cutting and Distributed Qubit Mapping for Connected Quantum Systems
di: Du, Zefan, et al.
Pubblicazione: (2024)
di: Du, Zefan, et al.
Pubblicazione: (2024)
Efficient Circuit Cutting and Scheduling in a Multi-Node Quantum System with Dynamic EPR Pairs
di: Du, Zefan, et al.
Pubblicazione: (2024)
di: Du, Zefan, et al.
Pubblicazione: (2024)
Toward Inclusive Educational AI: Auditing Frontier LLMs through a Multiplexity Lens
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
Are LLMs Rigorous Logical Reasoners? Empowering Natural Language Proof Generation by Stepwise Decoding with Contrastive Learning
di: Su, Ying, et al.
Pubblicazione: (2023)
di: Su, Ying, et al.
Pubblicazione: (2023)
Quantum Büchi Automata
di: Wang, Qisheng, et al.
Pubblicazione: (2018)
di: Wang, Qisheng, et al.
Pubblicazione: (2018)
IslamicLegalBench: Evaluating LLMs Knowledge and Reasoning of Islamic Law Across 1,200 Years of Islamic Pluralist Legal Traditions
di: Elmahjub, Ezieddin, et al.
Pubblicazione: (2026)
di: Elmahjub, Ezieddin, et al.
Pubblicazione: (2026)
Metaphors We Compute By: A Computational Audit of Cultural Translation vs. Thinking in LLMs
di: Chang, Yuan, et al.
Pubblicazione: (2026)
di: Chang, Yuan, et al.
Pubblicazione: (2026)
Explore the Reasoning Capability of LLMs in the Chess Testbed
di: Wang, Shu, et al.
Pubblicazione: (2024)
di: Wang, Shu, et al.
Pubblicazione: (2024)
GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
di: Duan, Jinhao, et al.
Pubblicazione: (2024)
di: Duan, Jinhao, et al.
Pubblicazione: (2024)
EvoWiki: Evaluating LLMs on Evolving Knowledge
di: Tang, Wei, et al.
Pubblicazione: (2024)
di: Tang, Wei, et al.
Pubblicazione: (2024)
Hardware-aware and Resource-efficient Circuit Packing and Scheduling on Trapped-Ion Quantum Computers
di: Palma, Miguel, et al.
Pubblicazione: (2025)
di: Palma, Miguel, et al.
Pubblicazione: (2025)
AuditBench: Evaluating Alignment Auditing Techniques on Models with Hidden Behaviors
di: Sheshadri, Abhay, et al.
Pubblicazione: (2026)
di: Sheshadri, Abhay, et al.
Pubblicazione: (2026)
Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability
di: Liang, Xiao, et al.
Pubblicazione: (2026)
di: Liang, Xiao, et al.
Pubblicazione: (2026)
Tokenization Constraints in LLMs: A Study of Symbolic and Arithmetic Reasoning Limits
di: Zhang, Xiang, et al.
Pubblicazione: (2025)
di: Zhang, Xiang, et al.
Pubblicazione: (2025)
Grading the Unspoken: Evaluating Tacit Reasoning in Quantum Field Theory and String Theory with LLMs
di: Yu, Xingyang, et al.
Pubblicazione: (2026)
di: Yu, Xingyang, et al.
Pubblicazione: (2026)
A Practical Quantum Hoare Logic with Classical Variables, I
di: Ying, Mingsheng
Pubblicazione: (2024)
di: Ying, Mingsheng
Pubblicazione: (2024)
FinAuditing: A Financial Taxonomy-Structured Multi-Document Benchmark for Evaluating LLMs
di: Wang, Yan, et al.
Pubblicazione: (2025)
di: Wang, Yan, et al.
Pubblicazione: (2025)
Test-Time Policy Adaptation for Enhanced Multi-Turn Interactions with LLMs
di: Wei, Chenxing, et al.
Pubblicazione: (2025)
di: Wei, Chenxing, et al.
Pubblicazione: (2025)
Model Utility Law: Evaluating LLMs beyond Performance through Mechanism Interpretable Metric
di: Cao, Yixin, et al.
Pubblicazione: (2025)
di: Cao, Yixin, et al.
Pubblicazione: (2025)
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
di: Huang, Wei, et al.
Pubblicazione: (2024)
di: Huang, Wei, et al.
Pubblicazione: (2024)
AraReasoner: Evaluating Reasoning-Based LLMs for Arabic NLP
di: Hasanaath, Ahmed, et al.
Pubblicazione: (2025)
di: Hasanaath, Ahmed, et al.
Pubblicazione: (2025)
STAIR: Spatial-Temporal Reasoning with Auditable Intermediate Results for Video Question Answering
di: Wang, Yueqian, et al.
Pubblicazione: (2024)
di: Wang, Yueqian, et al.
Pubblicazione: (2024)
Double-Calibration: Towards Reliable LLMs via Calibrating Knowledge and Reasoning Confidence
di: Lu, Yuyin, et al.
Pubblicazione: (2026)
di: Lu, Yuyin, et al.
Pubblicazione: (2026)
Verification of Recursively Defined Quantum Circuits
di: Ying, Mingsheng, et al.
Pubblicazione: (2024)
di: Ying, Mingsheng, et al.
Pubblicazione: (2024)
Enabling Discriminative Reasoning in LLMs for Legal Judgment Prediction
di: Deng, Chenlong, et al.
Pubblicazione: (2024)
di: Deng, Chenlong, et al.
Pubblicazione: (2024)
Unmasking Reasoning Processes: A Process-aware Benchmark for Evaluating Structural Mathematical Reasoning in LLMs
di: Zheng, Xiang, et al.
Pubblicazione: (2026)
di: Zheng, Xiang, et al.
Pubblicazione: (2026)
Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs
di: Alkaeed, Mahdi, et al.
Pubblicazione: (2026)
di: Alkaeed, Mahdi, et al.
Pubblicazione: (2026)
Can LLMs Write Faithfully? An Agent-Based Evaluation of LLM-generated Islamic Content
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Differentiable Architecture Search for Adversarially Robust Quantum Computer Vision
di: Afane, Mohamed, et al.
Pubblicazione: (2026) -
Next-Generation Phishing: How LLM Agents Empower Cyber Attackers
di: Afane, Khalifa, et al.
Pubblicazione: (2024) -
ATP: Adaptive Threshold Pruning for Efficient Data Encoding in Quantum Neural Networks
di: Afane, Mohamed, et al.
Pubblicazione: (2025) -
SCOUT: A Defense Against Data Poisoning Attacks in Fine-Tuned Language Models
di: Afane, Mohamed, et al.
Pubblicazione: (2025) -
Can LLMs Help Allocate Public Health Resources? A Case Study on Childhood Lead Testing
di: Afane, Mohamed, et al.
Pubblicazione: (2025)