MM-PhyQA: Multimodal Physics Question-Answering With Multi-Image CoT Prompting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Anand, Avinash, Kapuriya, Janak, Singh, Apoorv, Saraf, Jay, Lal, Naman, Verma, Astha, Gupta, Rushali, Shah, Rajiv |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MM-PhyRLHF: Reinforcement Learning Framework for Multimodal Physics Question-Answering
von: Kapuriya, Janak, et al.
Veröffentlicht: (2024)
von: Kapuriya, Janak, et al.
Veröffentlicht: (2024)
Enhancing Scientific Visual Question Answering via Vision-Caption aware Supervised Fine-Tuning
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025)
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025)
KG-CTG: Citation Generation through Knowledge Graph-guided Large Language Models
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
Context-Enhanced Language Models for Generating Multi-Paper Citations
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
Keystroke Dynamics Against Academic Dishonesty in the Age of LLMs
von: Kundu, Debnath, et al.
Veröffentlicht: (2024)
von: Kundu, Debnath, et al.
Veröffentlicht: (2024)
Spiritual-LLM : Gita Inspired Mental Health Therapy In the Era of LLMs
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025)
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025)
Certified Zeroth-order Black-Box Defense with Robust UNet Denoiser
von: Verma, Astha, et al.
Veröffentlicht: (2023)
von: Verma, Astha, et al.
Veröffentlicht: (2023)
LittiChoQA: Literary Texts in Indic Languages Chosen for Question Answering
von: Khandelwal, Aarya, et al.
Veröffentlicht: (2026)
von: Khandelwal, Aarya, et al.
Veröffentlicht: (2026)
A Progressive Evaluation Framework for Multicultural Analysis of Story Visualization
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025)
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025)
Knowledge Graphs are all you need: Leveraging KGs in Physics Question Answering
von: Addala, Krishnasai, et al.
Veröffentlicht: (2024)
von: Addala, Krishnasai, et al.
Veröffentlicht: (2024)
$π$-CoT: Prolog-Initialized Chain-of-Thought Prompting for Multi-Hop Question-Answering
von: Wan, Chao, et al.
Veröffentlicht: (2025)
von: Wan, Chao, et al.
Veröffentlicht: (2025)
Instruction Tuning and CoT Prompting for Contextual Medical QA with LLMs
von: Le, Chenqian, et al.
Veröffentlicht: (2025)
von: Le, Chenqian, et al.
Veröffentlicht: (2025)
Answering Questions in Stages: Prompt Chaining for Contract QA
von: Roegiest, Adam, et al.
Veröffentlicht: (2024)
von: Roegiest, Adam, et al.
Veröffentlicht: (2024)
MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering
von: Li, Xu, et al.
Veröffentlicht: (2025)
von: Li, Xu, et al.
Veröffentlicht: (2025)
Exploring the Role of Diversity in Example Selection for In-Context Learning
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025)
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025)
RanLayNet: A Dataset for Document Layout Detection used for Domain Adaptation and Generalization
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
Long-context Non-factoid Question Answering in Indic Languages
von: Mishra, Ritwik, et al.
Veröffentlicht: (2025)
von: Mishra, Ritwik, et al.
Veröffentlicht: (2025)
MMToM-QA: Multimodal Theory of Mind Question Answering
von: Jin, Chuanyang, et al.
Veröffentlicht: (2024)
von: Jin, Chuanyang, et al.
Veröffentlicht: (2024)
Memory-QA: Answering Recall Questions Based on Multimodal Memories
von: Jiang, Hongda, et al.
Veröffentlicht: (2025)
von: Jiang, Hongda, et al.
Veröffentlicht: (2025)
Step-CoT: Stepwise Visual Chain-of-Thought for Medical Visual Question Answering
von: Fan, Lin, et al.
Veröffentlicht: (2026)
von: Fan, Lin, et al.
Veröffentlicht: (2026)
HistoryBankQA: Multilingual Temporal Question Answering on Historical Events
von: Mandal, Biswadip, et al.
Veröffentlicht: (2025)
von: Mandal, Biswadip, et al.
Veröffentlicht: (2025)
Multilingual Non-Factoid Question Answering with Answer Paragraph Selection
von: Mishra, Ritwik, et al.
Veröffentlicht: (2024)
von: Mishra, Ritwik, et al.
Veröffentlicht: (2024)
DashboardQA: Benchmarking Multimodal Agents for Question Answering on Interactive Dashboards
von: Kartha, Aaryaman, et al.
Veröffentlicht: (2025)
von: Kartha, Aaryaman, et al.
Veröffentlicht: (2025)
Steps are all you need: Rethinking STEM Education with Prompt Engineering
von: Addala, Krishnasai, et al.
Veröffentlicht: (2024)
von: Addala, Krishnasai, et al.
Veröffentlicht: (2024)
Semantic Frame Aggregation-based Transformer for Live Video Comment Generation
von: Fatima, Anam, et al.
Veröffentlicht: (2025)
von: Fatima, Anam, et al.
Veröffentlicht: (2025)
Overview of TREC 2024 Medical Video Question Answering (MedVidQA) Track
von: Gupta, Deepak, et al.
Veröffentlicht: (2024)
von: Gupta, Deepak, et al.
Veröffentlicht: (2024)
D-SCoRE: Document-Centric Segmentation and CoT Reasoning with Structured Export for QA-CoT Data Generation
von: Zhou, Weibo, et al.
Veröffentlicht: (2025)
von: Zhou, Weibo, et al.
Veröffentlicht: (2025)
BLaVe-CoT: Consistency-Aware Visual Question Answering for Blind and Low Vision Users
von: Cheng, Wanyin, et al.
Veröffentlicht: (2025)
von: Cheng, Wanyin, et al.
Veröffentlicht: (2025)
InfoChartQA: A Benchmark for Multimodal Question Answering on Infographic Charts
von: Xie, Tianchi, et al.
Veröffentlicht: (2025)
von: Xie, Tianchi, et al.
Veröffentlicht: (2025)
PolQA: Polish Question Answering Dataset
von: Rybak, Piotr, et al.
Veröffentlicht: (2022)
von: Rybak, Piotr, et al.
Veröffentlicht: (2022)
VoQA: Visual-only Question Answering
von: An, Jianing, et al.
Veröffentlicht: (2025)
von: An, Jianing, et al.
Veröffentlicht: (2025)
P-RAG: Prompt-Enhanced Parametric RAG with LoRA and Selective CoT for Biomedical and Multi-Hop QA
von: Lyu, Xingda, et al.
Veröffentlicht: (2026)
von: Lyu, Xingda, et al.
Veröffentlicht: (2026)
RoadscapesQA: A Multitask, Multimodal Dataset for Visual Question Answering on Indian Roads
von: Iyer, Vijayasri, et al.
Veröffentlicht: (2026)
von: Iyer, Vijayasri, et al.
Veröffentlicht: (2026)
MovieRecapsQA: A Multimodal Open-Ended Video Question-Answering Benchmark
von: Shaar, Shaden, et al.
Veröffentlicht: (2026)
von: Shaar, Shaden, et al.
Veröffentlicht: (2026)
WikiMixQA: A Multimodal Benchmark for Question Answering over Tables and Charts
von: Foroutan, Negar, et al.
Veröffentlicht: (2025)
von: Foroutan, Negar, et al.
Veröffentlicht: (2025)
CoReQA: Uncovering Potentials of Language Models in Code Repository Question Answering
von: Chen, Jialiang, et al.
Veröffentlicht: (2025)
von: Chen, Jialiang, et al.
Veröffentlicht: (2025)
Efficient and Interpretable Information Retrieval for Product Question Answering with Heterogeneous Data
von: Biswas, Biplob, et al.
Veröffentlicht: (2024)
von: Biswas, Biplob, et al.
Veröffentlicht: (2024)
Advancements in Scientific Controllable Text Generation Methods
von: Goel, Arnav, et al.
Veröffentlicht: (2023)
von: Goel, Arnav, et al.
Veröffentlicht: (2023)
LingoQA: Visual Question Answering for Autonomous Driving
von: Marcu, Ana-Maria, et al.
Veröffentlicht: (2023)
von: Marcu, Ana-Maria, et al.
Veröffentlicht: (2023)
DebateQA: Evaluating Question Answering on Debatable Knowledge
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MM-PhyRLHF: Reinforcement Learning Framework for Multimodal Physics Question-Answering
von: Kapuriya, Janak, et al.
Veröffentlicht: (2024) -
Enhancing Scientific Visual Question Answering via Vision-Caption aware Supervised Fine-Tuning
von: Kapuriya, Janak, et al.
Veröffentlicht: (2025) -
KG-CTG: Citation Generation through Knowledge Graph-guided Large Language Models
von: Anand, Avinash, et al.
Veröffentlicht: (2024) -
Context-Enhanced Language Models for Generating Multi-Paper Citations
von: Anand, Avinash, et al.
Veröffentlicht: (2024) -
Keystroke Dynamics Against Academic Dishonesty in the Age of LLMs
von: Kundu, Debnath, et al.
Veröffentlicht: (2024)