MM-PhyRLHF: Reinforcement Learning Framework for Multimodal Physics Question-Answering
Fuente:
arXiv
Saved in:
| Main Authors: | Kapuriya, Janak, Kirtani, Chhavi, Singh, Apoorv, Saraf, Jay, Lal, Naman, Kumar, Jatin, Shivam, Adarsh Raj, Verma, Astha, Anand, Avinash, Shah, Rajiv Ratn |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MM-PhyQA: Multimodal Physics Question-Answering With Multi-Image CoT Prompting
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Enhancing Scientific Visual Question Answering via Vision-Caption aware Supervised Fine-Tuning
by: Kapuriya, Janak, et al.
Published: (2025)
by: Kapuriya, Janak, et al.
Published: (2025)
Knowledge Graphs are all you need: Leveraging KGs in Physics Question Answering
by: Addala, Krishnasai, et al.
Published: (2024)
by: Addala, Krishnasai, et al.
Published: (2024)
KG-CTG: Citation Generation through Knowledge Graph-guided Large Language Models
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Context-Enhanced Language Models for Generating Multi-Paper Citations
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Spiritual-LLM : Gita Inspired Mental Health Therapy In the Era of LLMs
by: Kapuriya, Janak, et al.
Published: (2025)
by: Kapuriya, Janak, et al.
Published: (2025)
Keystroke Dynamics Against Academic Dishonesty in the Age of LLMs
by: Kundu, Debnath, et al.
Published: (2024)
by: Kundu, Debnath, et al.
Published: (2024)
Mathify: Evaluating Large Language Models on Mathematical Problem Solving Tasks
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Steps are all you need: Rethinking STEM Education with Prompt Engineering
by: Addala, Krishnasai, et al.
Published: (2024)
by: Addala, Krishnasai, et al.
Published: (2024)
Certified Zeroth-order Black-Box Defense with Robust UNet Denoiser
by: Verma, Astha, et al.
Published: (2023)
by: Verma, Astha, et al.
Published: (2023)
Enhancing LLMs for Physics Problem-Solving using Reinforcement Learning with Human-AI Feedback
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Long-context Non-factoid Question Answering in Indic Languages
by: Mishra, Ritwik, et al.
Published: (2025)
by: Mishra, Ritwik, et al.
Published: (2025)
LittiChoQA: Literary Texts in Indic Languages Chosen for Question Answering
by: Khandelwal, Aarya, et al.
Published: (2026)
by: Khandelwal, Aarya, et al.
Published: (2026)
IRIS: Interleaved Reinforcement with Incremental Staged Curriculum for Cross-Lingual Mathematical Reasoning
by: Gupta, Navya, et al.
Published: (2026)
by: Gupta, Navya, et al.
Published: (2026)
Multilingual Non-Factoid Question Answering with Answer Paragraph Selection
by: Mishra, Ritwik, et al.
Published: (2024)
by: Mishra, Ritwik, et al.
Published: (2024)
RanLayNet: A Dataset for Document Layout Detection used for Domain Adaptation and Generalization
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
A Progressive Evaluation Framework for Multicultural Analysis of Story Visualization
by: Kapuriya, Janak, et al.
Published: (2025)
by: Kapuriya, Janak, et al.
Published: (2025)
Advancements in Scientific Controllable Text Generation Methods
by: Goel, Arnav, et al.
Published: (2023)
by: Goel, Arnav, et al.
Published: (2023)
Su-RoBERTa: A Semi-supervised Approach to Predicting Suicide Risk through Social Media using Base Language Models
by: Tank, Chayan, et al.
Published: (2024)
by: Tank, Chayan, et al.
Published: (2024)
Improving Physics Reasoning in Large Language Models Using Mixture of Refinement Agents
by: Jaiswal, Raj, et al.
Published: (2024)
by: Jaiswal, Raj, et al.
Published: (2024)
Exploring the Role of Diversity in Example Selection for In-Context Learning
by: Kapuriya, Janak, et al.
Published: (2025)
by: Kapuriya, Janak, et al.
Published: (2025)
RConE: Rough Cone Embedding for Multi-Hop Logical Query Answering on Multi-Modal Knowledge Graphs
by: Kharbanda, Mayank, et al.
Published: (2024)
by: Kharbanda, Mayank, et al.
Published: (2024)
Depression Detection and Analysis using Large Language Models on Textual and Audio-Visual Modalities
by: Tank, Chayan, et al.
Published: (2024)
by: Tank, Chayan, et al.
Published: (2024)
TC-OCR: TableCraft OCR for Efficient Detection & Recognition of Table Structure & Content
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Semantic Frame Aggregation-based Transformer for Live Video Comment Generation
by: Fatima, Anam, et al.
Published: (2025)
by: Fatima, Anam, et al.
Published: (2025)
Improving Multimodal LLMs Ability In Geometry Problem Solving, Reasoning, And Multistep Scoring
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
ReviewEval: An Evaluation Framework for AI-Generated Reviews
by: Garg, Madhav Krishan, et al.
Published: (2025)
by: Garg, Madhav Krishan, et al.
Published: (2025)
Efficient and Interpretable Information Retrieval for Product Question Answering with Heterogeneous Data
by: Biswas, Biplob, et al.
Published: (2024)
by: Biswas, Biplob, et al.
Published: (2024)
Better and Worse with Scale: How Contextual Entrainment Diverges with Model Size
by: Kukreja, Dikshant, et al.
Published: (2026)
by: Kukreja, Dikshant, et al.
Published: (2026)
Peering into the Mind of Language Models: An Approach for Attribution in Contextual Question Answering
by: Phukan, Anirudh, et al.
Published: (2024)
by: Phukan, Anirudh, et al.
Published: (2024)
Med-CoDE: Medical Critique based Disagreement Evaluation Framework
by: Gupta, Mohit, et al.
Published: (2025)
by: Gupta, Mohit, et al.
Published: (2025)
Speech Representation Learning Revisited: The Necessity of Separate Learnable Parameters and Robust Data Augmentation
by: Yadav, Hemant, et al.
Published: (2024)
by: Yadav, Hemant, et al.
Published: (2024)
Analysing the Masked predictive coding training criterion for pre-training a Speech Representation Model
by: Yadav, Hemant, et al.
Published: (2023)
by: Yadav, Hemant, et al.
Published: (2023)
MS-HuBERT: Mitigating Pre-training and Inference Mismatch in Masked Language Modelling methods for learning Speech Representations
by: Yadav, Hemant, et al.
Published: (2024)
by: Yadav, Hemant, et al.
Published: (2024)
JOOCI: a Framework for Learning Comprehensive Speech Representations
by: Yadav, Hemant, et al.
Published: (2024)
by: Yadav, Hemant, et al.
Published: (2024)
Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy
by: Kumarappan, Adarsh, et al.
Published: (2026)
by: Kumarappan, Adarsh, et al.
Published: (2026)
MM-RLHF: The Next Step Forward in Multimodal LLM Alignment
by: Zhang, Yi-Fan, et al.
Published: (2025)
by: Zhang, Yi-Fan, et al.
Published: (2025)
Prevalence of Gastrointestinal Parasites in Blackbuck ( Antelope cervicapra Linnaeus, 1758) of Blackbuck Conservation Area, Khairapur, Bardia, Nepal
by: Muna Thapa, et al.
Published: (2026)
by: Muna Thapa, et al.
Published: (2026)
MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering
by: Li, Xu, et al.
Published: (2025)
by: Li, Xu, et al.
Published: (2025)
Similar Items
-
MM-PhyQA: Multimodal Physics Question-Answering With Multi-Image CoT Prompting
by: Anand, Avinash, et al.
Published: (2024) -
Enhancing Scientific Visual Question Answering via Vision-Caption aware Supervised Fine-Tuning
by: Kapuriya, Janak, et al.
Published: (2025) -
Knowledge Graphs are all you need: Leveraging KGs in Physics Question Answering
by: Addala, Krishnasai, et al.
Published: (2024) -
KG-CTG: Citation Generation through Knowledge Graph-guided Large Language Models
by: Anand, Avinash, et al.
Published: (2024) -
Context-Enhanced Language Models for Generating Multi-Paper Citations
by: Anand, Avinash, et al.
Published: (2024)