Enhancing LLMs for Physics Problem-Solving using Reinforcement Learning with Human-AI Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Anand, Avinash, Prasad, Kritarth, Kirtani, Chhavi, Nair, Ashwin R, Gupta, Mohit, Garg, Saloni, Gautam, Anurag, Buldeo, Snehal, Shah, Rajiv Ratn |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Mathify: Evaluating Large Language Models on Mathematical Problem Solving Tasks
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
KG-CTG: Citation Generation through Knowledge Graph-guided Large Language Models
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Context-Enhanced Language Models for Generating Multi-Paper Citations
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Improving Multimodal LLMs Ability In Geometry Problem Solving, Reasoning, And Multistep Scoring
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Steps are all you need: Rethinking STEM Education with Prompt Engineering
by: Addala, Krishnasai, et al.
Published: (2024)
by: Addala, Krishnasai, et al.
Published: (2024)
Knowledge Graphs are all you need: Leveraging KGs in Physics Question Answering
by: Addala, Krishnasai, et al.
Published: (2024)
by: Addala, Krishnasai, et al.
Published: (2024)
MM-PhyRLHF: Reinforcement Learning Framework for Multimodal Physics Question-Answering
by: Kapuriya, Janak, et al.
Published: (2024)
by: Kapuriya, Janak, et al.
Published: (2024)
IRIS: Interleaved Reinforcement with Incremental Staged Curriculum for Cross-Lingual Mathematical Reasoning
by: Gupta, Navya, et al.
Published: (2026)
by: Gupta, Navya, et al.
Published: (2026)
Med-CoDE: Medical Critique based Disagreement Evaluation Framework
by: Gupta, Mohit, et al.
Published: (2025)
by: Gupta, Mohit, et al.
Published: (2025)
Advancements in Scientific Controllable Text Generation Methods
by: Goel, Arnav, et al.
Published: (2023)
by: Goel, Arnav, et al.
Published: (2023)
Keystroke Dynamics Against Academic Dishonesty in the Age of LLMs
by: Kundu, Debnath, et al.
Published: (2024)
by: Kundu, Debnath, et al.
Published: (2024)
Better and Worse with Scale: How Contextual Entrainment Diverges with Model Size
by: Kukreja, Dikshant, et al.
Published: (2026)
by: Kukreja, Dikshant, et al.
Published: (2026)
ReviewEval: An Evaluation Framework for AI-Generated Reviews
by: Garg, Madhav Krishan, et al.
Published: (2025)
by: Garg, Madhav Krishan, et al.
Published: (2025)
TC-OCR: TableCraft OCR for Efficient Detection & Recognition of Table Structure & Content
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Depression Detection and Analysis using Large Language Models on Textual and Audio-Visual Modalities
by: Tank, Chayan, et al.
Published: (2024)
by: Tank, Chayan, et al.
Published: (2024)
On Optimal Steering to Achieve Exact Fairness
by: Sharma, Mohit, et al.
Published: (2025)
by: Sharma, Mohit, et al.
Published: (2025)
RanLayNet: A Dataset for Document Layout Detection used for Domain Adaptation and Generalization
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Spiritual-LLM : Gita Inspired Mental Health Therapy In the Era of LLMs
by: Kapuriya, Janak, et al.
Published: (2025)
by: Kapuriya, Janak, et al.
Published: (2025)
Su-RoBERTa: A Semi-supervised Approach to Predicting Suicide Risk through Social Media using Base Language Models
by: Tank, Chayan, et al.
Published: (2024)
by: Tank, Chayan, et al.
Published: (2024)
Structured Definitions and Segmentations for Legal Reasoning in LLMs: A Study on Indian Legal Data
by: Khatri, Mann, et al.
Published: (2025)
by: Khatri, Mann, et al.
Published: (2025)
Improving Physics Reasoning in Large Language Models Using Mixture of Refinement Agents
by: Jaiswal, Raj, et al.
Published: (2024)
by: Jaiswal, Raj, et al.
Published: (2024)
Speech Representation Learning Revisited: The Necessity of Separate Learnable Parameters and Robust Data Augmentation
by: Yadav, Hemant, et al.
Published: (2024)
by: Yadav, Hemant, et al.
Published: (2024)
LittiChoQA: Literary Texts in Indic Languages Chosen for Question Answering
by: Khandelwal, Aarya, et al.
Published: (2026)
by: Khandelwal, Aarya, et al.
Published: (2026)
RConE: Rough Cone Embedding for Multi-Hop Logical Query Answering on Multi-Modal Knowledge Graphs
by: Kharbanda, Mayank, et al.
Published: (2024)
by: Kharbanda, Mayank, et al.
Published: (2024)
Analysing the Masked predictive coding training criterion for pre-training a Speech Representation Model
by: Yadav, Hemant, et al.
Published: (2023)
by: Yadav, Hemant, et al.
Published: (2023)
Long-context Non-factoid Question Answering in Indic Languages
by: Mishra, Ritwik, et al.
Published: (2025)
by: Mishra, Ritwik, et al.
Published: (2025)
MS-HuBERT: Mitigating Pre-training and Inference Mismatch in Masked Language Modelling methods for learning Speech Representations
by: Yadav, Hemant, et al.
Published: (2024)
by: Yadav, Hemant, et al.
Published: (2024)
JOOCI: a Framework for Learning Comprehensive Speech Representations
by: Yadav, Hemant, et al.
Published: (2024)
by: Yadav, Hemant, et al.
Published: (2024)
Teaching Human Behavior Improves Content Understanding Abilities Of LLMs
by: Singh, Somesh, et al.
Published: (2024)
by: Singh, Somesh, et al.
Published: (2024)
Designing an Intelligent Parcel Management System using IoT & Machine Learning
by: Gupta, Mohit, et al.
Published: (2024)
by: Gupta, Mohit, et al.
Published: (2024)
Faster Machine Translation Ensembling with Reinforcement Learning and Competitive Correction
by: Prasad, Kritarth, et al.
Published: (2025)
by: Prasad, Kritarth, et al.
Published: (2025)
Psychologically-Grounded Graph Modeling for Interpretable Depression Detection
by: Vyalla, Rishitej Reddy, et al.
Published: (2026)
by: Vyalla, Rishitej Reddy, et al.
Published: (2026)
Multilingual Coreference Resolution in Low-resource South Asian Languages
by: Mishra, Ritwik, et al.
Published: (2024)
by: Mishra, Ritwik, et al.
Published: (2024)
Multilingual Non-Factoid Question Answering with Answer Paragraph Selection
by: Mishra, Ritwik, et al.
Published: (2024)
by: Mishra, Ritwik, et al.
Published: (2024)
In-Domain African Languages Translation Using LLMs and Multi-armed Bandits
by: Singh, Pratik Rakesh, et al.
Published: (2025)
by: Singh, Pratik Rakesh, et al.
Published: (2025)
DubWise: Video-Guided Speech Duration Control in Multimodal LLM-based Text-to-Speech for Dubbing
by: Sahipjohn, Neha, et al.
Published: (2024)
by: Sahipjohn, Neha, et al.
Published: (2024)
VECL-TTS: Voice identity and Emotional style controllable Cross-Lingual Text-to-Speech
by: Gudmalwar, Ashishkumar, et al.
Published: (2024)
by: Gudmalwar, Ashishkumar, et al.
Published: (2024)
Isometric Neural Machine Translation using Phoneme Count Ratio Reward-based Reinforcement Learning
by: Mhaskar, Shivam Ratnakant, et al.
Published: (2024)
by: Mhaskar, Shivam Ratnakant, et al.
Published: (2024)
EmoReg: Directional Latent Vector Modeling for Emotional Intensity Regularization in Diffusion-based Voice Conversion
by: Gudmalwar, Ashishkumar, et al.
Published: (2024)
by: Gudmalwar, Ashishkumar, et al.
Published: (2024)
Similar Items
-
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English
by: Anand, Avinash, et al.
Published: (2024) -
Mathify: Evaluating Large Language Models on Mathematical Problem Solving Tasks
by: Anand, Avinash, et al.
Published: (2024) -
KG-CTG: Citation Generation through Knowledge Graph-guided Large Language Models
by: Anand, Avinash, et al.
Published: (2024) -
Context-Enhanced Language Models for Generating Multi-Paper Citations
by: Anand, Avinash, et al.
Published: (2024) -
Improving Multimodal LLMs Ability In Geometry Problem Solving, Reasoning, And Multistep Scoring
by: Anand, Avinash, et al.
Published: (2024)