Understanding LLM Scientific Reasoning through Promptings and Model's Explanation on the Answers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rueda, Alice, Hassan, Mohammed S., Perivolaris, Argyrios, Teferra, Bazen G., Samavi, Reza, Rambhatla, Sirisha, Wu, Yuqi, Zhang, Yanbo, Cao, Bo, Sharma, Divya, Krishnan, Sridhar, Bhat, Venkat |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Human vs. LLM-Based Thematic Analysis for Digital Mental Health Research: Proof-of-Concept Comparative Study
von: Parkington, Karisa, et al.
Veröffentlicht: (2025)
von: Parkington, Karisa, et al.
Veröffentlicht: (2025)
Estimating Quality in Therapeutic Conversations: A Multi-Dimensional Natural Language Processing Framework
von: Rueda, Alice, et al.
Veröffentlicht: (2025)
von: Rueda, Alice, et al.
Veröffentlicht: (2025)
Medical Misinformation in AI-Assisted Self-Diagnosis: Development of a Method (EvalPrompt) for Analyzing Large Language Models
von: Zada, Troy, et al.
Veröffentlicht: (2023)
von: Zada, Troy, et al.
Veröffentlicht: (2023)
LOCATEdit: Graph Laplacian Optimized Cross Attention for Localized Text-Guided Image Editing
von: Soni, Achint, et al.
Veröffentlicht: (2025)
von: Soni, Achint, et al.
Veröffentlicht: (2025)
SafeTuneBed: A Toolkit for Benchmarking LLM Safety Alignment in Fine-Tuning
von: Hossain, Saad, et al.
Veröffentlicht: (2025)
von: Hossain, Saad, et al.
Veröffentlicht: (2025)
SubTrack++ : Gradient Subspace Tracking for Scalable LLM Training
von: Rajabi, Sahar, et al.
Veröffentlicht: (2025)
von: Rajabi, Sahar, et al.
Veröffentlicht: (2025)
Randomized Gradient Subspaces for Efficient Large Language Model Training
von: Rajabi, Sahar, et al.
Veröffentlicht: (2025)
von: Rajabi, Sahar, et al.
Veröffentlicht: (2025)
Zero-Shot Object Re-Identification in Egocentric Kitchen Videos via Multi-Stage SAM3 Feature Fusion
von: Klepachevskyi, Dmytro, et al.
Veröffentlicht: (2026)
von: Klepachevskyi, Dmytro, et al.
Veröffentlicht: (2026)
PinPoint: Prompting with Informative Interior Points
von: Sadeghi, Pouya, et al.
Veröffentlicht: (2026)
von: Sadeghi, Pouya, et al.
Veröffentlicht: (2026)
CycleGAN Models for MRI Image Translation
von: Czobit, Cassandra, et al.
Veröffentlicht: (2023)
von: Czobit, Cassandra, et al.
Veröffentlicht: (2023)
GNN's Uncertainty Quantification using Self-Distillation
von: Daneshvar, Hirad, et al.
Veröffentlicht: (2025)
von: Daneshvar, Hirad, et al.
Veröffentlicht: (2025)
Evidential Uncertainty Sets in Deep Classifiers Using Conformal Prediction
von: Karimi, Hamed, et al.
Veröffentlicht: (2024)
von: Karimi, Hamed, et al.
Veröffentlicht: (2024)
Quantifying Deep Learning Model Uncertainty in Conformal Prediction
von: Karimi, Hamed, et al.
Veröffentlicht: (2023)
von: Karimi, Hamed, et al.
Veröffentlicht: (2023)
Chain of Unit-Physics: A Primitive-Centric Approach to Scientific Code Synthesis
von: Sharma, Vansh, et al.
Veröffentlicht: (2025)
von: Sharma, Vansh, et al.
Veröffentlicht: (2025)
PromptExp: Multi-granularity Prompt Explanation of Large Language Models
von: Dong, Ximing, et al.
Veröffentlicht: (2024)
von: Dong, Ximing, et al.
Veröffentlicht: (2024)
Gen4D: Synthesizing Humans and Scenes in the Wild
von: Bright, Jerrin, et al.
Veröffentlicht: (2025)
von: Bright, Jerrin, et al.
Veröffentlicht: (2025)
LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy
von: Karimi, Hamed, et al.
Veröffentlicht: (2026)
von: Karimi, Hamed, et al.
Veröffentlicht: (2026)
Color‐Coding Dietary Proteins by Source Based on Impacts of Their Production and Consumption on the Planet and the People—Redefining Proteins in the Future Food Systems
von: Tadesse Fikre Teferra
Veröffentlicht: (2025)
von: Tadesse Fikre Teferra
Veröffentlicht: (2025)
Performance Evaluation of a Radial Distribution Network Under Emerging Load Prediction Modeling Approach and DG Integration Using a Particle Swarm Optimization Algorithm
von: Demsew Mitiku Teferra
Veröffentlicht: (2025)
von: Demsew Mitiku Teferra
Veröffentlicht: (2025)
Bvac-AI DataSet
von: Nasir, Samavi
Veröffentlicht: (2024)
von: Nasir, Samavi
Veröffentlicht: (2024)
Avatar4D: Synthesizing Domain-Specific 4D Humans for Real-World Pose Estimation
von: Bright, Jerrin, et al.
Veröffentlicht: (2025)
von: Bright, Jerrin, et al.
Veröffentlicht: (2025)
Large Vocabulary Spontaneous Speech Recognition for Tigrigna
von: Kahsu, Ataklti, et al.
Veröffentlicht: (2023)
von: Kahsu, Ataklti, et al.
Veröffentlicht: (2023)
A numerical study into neural network surrogate model performance for uncertainty propagation
von: Wade, Noah, et al.
Veröffentlicht: (2026)
von: Wade, Noah, et al.
Veröffentlicht: (2026)
Domain-Guided Masked Autoencoders for Unique Player Identification
von: Balaji, Bavesh, et al.
Veröffentlicht: (2024)
von: Balaji, Bavesh, et al.
Veröffentlicht: (2024)
CUE-R: Beyond the Final Answer in Retrieval-Augmented Generation
von: Jain, Siddharth, et al.
Veröffentlicht: (2026)
von: Jain, Siddharth, et al.
Veröffentlicht: (2026)
CEAR: Certified Ensemble Adversarial Robustness in DNNs
von: Sadig, Daniel, et al.
Veröffentlicht: (2026)
von: Sadig, Daniel, et al.
Veröffentlicht: (2026)
Cascading Robustness Verification: Toward Efficient Model-Agnostic Certification
von: Maleki, Mohammadreza, et al.
Veröffentlicht: (2026)
von: Maleki, Mohammadreza, et al.
Veröffentlicht: (2026)
Prompt Informed Reinforcement Learning for Visual Coverage Path Planning
von: Margapuri, Venkat
Veröffentlicht: (2025)
von: Margapuri, Venkat
Veröffentlicht: (2025)
Superposition as Lossy Compression: Measure with Sparse Autoencoders and Connect to Adversarial Vulnerability
von: Bereska, Leonard, et al.
Veröffentlicht: (2025)
von: Bereska, Leonard, et al.
Veröffentlicht: (2025)
LangDA: Building Context-Awareness via Language for Domain Adaptive Semantic Segmentation
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Adversarial Pressure
von: Li, Yubo, et al.
Veröffentlicht: (2026)
von: Li, Yubo, et al.
Veröffentlicht: (2026)
Geometric conditions for matrix domination in two dimensions
von: Christodoulou, Argyrios
Veröffentlicht: (2019)
von: Christodoulou, Argyrios
Veröffentlicht: (2019)
A note on the boundary dynamics of holomorphic iterated function systems
von: Christodoulou, Argyrios
Veröffentlicht: (2025)
von: Christodoulou, Argyrios
Veröffentlicht: (2025)
Parameter spaces of locally constant cocycles
von: Christodoulou, Argyrios
Veröffentlicht: (2020)
von: Christodoulou, Argyrios
Veröffentlicht: (2020)
Scientific QA System with Verifiable Answers
von: Ljajić, Adela, et al.
Veröffentlicht: (2024)
von: Ljajić, Adela, et al.
Veröffentlicht: (2024)
Decoding Financial Behaviour: An Analysis of urbanised households in India using AIDIS 77th round
von: Sharma, Divya
Veröffentlicht: (2024)
von: Sharma, Divya
Veröffentlicht: (2024)
Unveiling Saving and Credit Dynamics: Insights from Financial Diaries and Surveys among Low-Income Households in Unauthorized Colonies in Delhi
von: Sharma, Divya
Veröffentlicht: (2024)
von: Sharma, Divya
Veröffentlicht: (2024)
SelfEval: Leveraging the discriminative nature of generative models for evaluation
von: Rambhatla, Sai Saketh, et al.
Veröffentlicht: (2023)
von: Rambhatla, Sai Saketh, et al.
Veröffentlicht: (2023)
Missingness Bias Calibration in Feature Attribution Explanations
von: Sridhar, Shailesh, et al.
Veröffentlicht: (2026)
von: Sridhar, Shailesh, et al.
Veröffentlicht: (2026)
A Reliable Knowledge Processing Framework for Combustion Science using Foundation Models
von: Sharma, Vansh, et al.
Veröffentlicht: (2023)
von: Sharma, Vansh, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Human vs. LLM-Based Thematic Analysis for Digital Mental Health Research: Proof-of-Concept Comparative Study
von: Parkington, Karisa, et al.
Veröffentlicht: (2025) -
Estimating Quality in Therapeutic Conversations: A Multi-Dimensional Natural Language Processing Framework
von: Rueda, Alice, et al.
Veröffentlicht: (2025) -
Medical Misinformation in AI-Assisted Self-Diagnosis: Development of a Method (EvalPrompt) for Analyzing Large Language Models
von: Zada, Troy, et al.
Veröffentlicht: (2023) -
LOCATEdit: Graph Laplacian Optimized Cross Attention for Localized Text-Guided Image Editing
von: Soni, Achint, et al.
Veröffentlicht: (2025) -
SafeTuneBed: A Toolkit for Benchmarking LLM Safety Alignment in Fine-Tuning
von: Hossain, Saad, et al.
Veröffentlicht: (2025)