Multimodal Mathematical Reasoning with Diverse Solving Perspective
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shi, Wenhao, Hu, Zhiqiang, Bin, Yi, Yang, Yang, Ng, See-Kiong, Shen, Heng Tao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models
von: Shi, Wenhao, et al.
Veröffentlicht: (2024)
von: Shi, Wenhao, et al.
Veröffentlicht: (2024)
GalleryGPT: Analyzing Paintings with Large Multimodal Models
von: Bin, Yi, et al.
Veröffentlicht: (2024)
von: Bin, Yi, et al.
Veröffentlicht: (2024)
Hint-before-Solving Prompting: Guiding LLMs to Effectively Utilize Encoded Knowledge
von: Fu, Jinlan, et al.
Veröffentlicht: (2024)
von: Fu, Jinlan, et al.
Veröffentlicht: (2024)
Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet
von: Zhao, James Xu, et al.
Veröffentlicht: (2025)
von: Zhao, James Xu, et al.
Veröffentlicht: (2025)
Dipper: Diversity in Prompts for Producing Large Language Model Ensembles in Reasoning tasks
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2024)
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2024)
FutureOmni: Evaluating Future Forecasting from Omni-Modal Context for Multimodal LLMs
von: Chen, Qian, et al.
Veröffentlicht: (2026)
von: Chen, Qian, et al.
Veröffentlicht: (2026)
SemCORE: A Semantic-Enhanced Generative Cross-Modal Retrieval Framework with MLLMs
von: Li, Haoxuan, et al.
Veröffentlicht: (2025)
von: Li, Haoxuan, et al.
Veröffentlicht: (2025)
Aligning Large Language Models for Faithful Integrity Against Opposing Argument
von: Zhao, Yong, et al.
Veröffentlicht: (2025)
von: Zhao, Yong, et al.
Veröffentlicht: (2025)
CHiP: Cross-modal Hierarchical Direct Preference Optimization for Multimodal LLMs
von: Fu, Jinlan, et al.
Veröffentlicht: (2025)
von: Fu, Jinlan, et al.
Veröffentlicht: (2025)
CET2: Modelling Topic Transitions for Coherent and Engaging Knowledge-Grounded Conversations
von: Xu, Lin, et al.
Veröffentlicht: (2024)
von: Xu, Lin, et al.
Veröffentlicht: (2024)
Analyzing Reasoning Consistency in Large Multimodal Models under Cross-Modal Conflicts
von: Zhu, Zhihao, et al.
Veröffentlicht: (2026)
von: Zhu, Zhihao, et al.
Veröffentlicht: (2026)
LETGAMES: An LLM-Powered Gamified Approach to Cognitive Training for Patients with Cognitive Impairment
von: Shi, Jingwei, et al.
Veröffentlicht: (2026)
von: Shi, Jingwei, et al.
Veröffentlicht: (2026)
ThinkTank-ME: A Multi-Expert Framework for Middle East Event Forecasting
von: Li, Haoxuan, et al.
Veröffentlicht: (2026)
von: Li, Haoxuan, et al.
Veröffentlicht: (2026)
HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding
von: Zhang, Haowei, et al.
Veröffentlicht: (2026)
von: Zhang, Haowei, et al.
Veröffentlicht: (2026)
Confidence Elicitation: A New Attack Vector for Large Language Models
von: Formento, Brian, et al.
Veröffentlicht: (2025)
von: Formento, Brian, et al.
Veröffentlicht: (2025)
Mercury: A Code Efficiency Benchmark for Code Large Language Models
von: Du, Mingzhe, et al.
Veröffentlicht: (2024)
von: Du, Mingzhe, et al.
Veröffentlicht: (2024)
Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)
Don't Just Say "I don't know"! Self-aligning Large Language Models for Responding to Unknown Questions with Explanations
von: Deng, Yang, et al.
Veröffentlicht: (2024)
von: Deng, Yang, et al.
Veröffentlicht: (2024)
Plug-and-Play Policy Planner for Large Language Model Powered Dialogue Agents
von: Deng, Yang, et al.
Veröffentlicht: (2023)
von: Deng, Yang, et al.
Veröffentlicht: (2023)
Ask-before-Plan: Proactive Language Agents for Real-World Planning
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
Describe-then-Reason: Improving Multimodal Mathematical Reasoning through Visual Comprehension Training
von: Jia, Mengzhao, et al.
Veröffentlicht: (2024)
von: Jia, Mengzhao, et al.
Veröffentlicht: (2024)
Chain of Thought Explanation for Dialogue State Tracking
von: Xu, Lin, et al.
Veröffentlicht: (2024)
von: Xu, Lin, et al.
Veröffentlicht: (2024)
MASim: Multilingual Agent-Based Simulation for Social Science
von: Zhang, Xuan, et al.
Veröffentlicht: (2025)
von: Zhang, Xuan, et al.
Veröffentlicht: (2025)
METRO: Towards Strategy Induction from Expert Dialogue Transcripts for Non-collaborative Dialogues
von: Yang, Haofu, et al.
Veröffentlicht: (2026)
von: Yang, Haofu, et al.
Veröffentlicht: (2026)
Beyond Prompt: Fine-grained Simulation of Cognitively Impaired Standardized Patients via Stochastic Steering
von: Zhang, Weikang, et al.
Veröffentlicht: (2026)
von: Zhang, Weikang, et al.
Veröffentlicht: (2026)
MCM-DPO: Multifaceted Cross-Modal Direct Preference Optimization for Alt-text Generation
von: Fu, Jinlan, et al.
Veröffentlicht: (2025)
von: Fu, Jinlan, et al.
Veröffentlicht: (2025)
PHAnToM: Persona-based Prompting Has An Effect on Theory-of-Mind Reasoning in Large Language Models
von: Tan, Fiona Anting, et al.
Veröffentlicht: (2024)
von: Tan, Fiona Anting, et al.
Veröffentlicht: (2024)
Rewarding the Rare: Uniqueness-Aware RL for Creative Problem Solving in LLMs
von: Hu, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Hu, Zhiyuan, et al.
Veröffentlicht: (2026)
METER: Evaluating Multi-Level Contextual Causal Reasoning in Large Language Models
von: Li, Pengfeng, et al.
Veröffentlicht: (2026)
von: Li, Pengfeng, et al.
Veröffentlicht: (2026)
On the Multi-turn Instruction Following for Conversational Web Agents
von: Deng, Yang, et al.
Veröffentlicht: (2024)
von: Deng, Yang, et al.
Veröffentlicht: (2024)
TRACE: TRansformer-based Attribution using Contrastive Embeddings in LLMs
von: Wang, Cheng, et al.
Veröffentlicht: (2024)
von: Wang, Cheng, et al.
Veröffentlicht: (2024)
Towards Proactive Information Probing: Customer Service Chatbots Harvesting Value from Conversation
von: Huang, Chen, et al.
Veröffentlicht: (2026)
von: Huang, Chen, et al.
Veröffentlicht: (2026)
FACT-AUDIT: An Adaptive Multi-Agent Framework for Dynamic Fact-Checking Evaluation of Large Language Models
von: Lin, Hongzhan, et al.
Veröffentlicht: (2025)
von: Lin, Hongzhan, et al.
Veröffentlicht: (2025)
ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving
von: Gou, Zhibin, et al.
Veröffentlicht: (2023)
von: Gou, Zhibin, et al.
Veröffentlicht: (2023)
Vision-and-Language Pretraining
von: Nguyen, Thong, et al.
Veröffentlicht: (2022)
von: Nguyen, Thong, et al.
Veröffentlicht: (2022)
MAMA: Meta-optimized Angular Margin Contrastive Framework for Video-Language Representation Learning
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)
Guiding VLM Agents with Process Rewards at Inference Time for GUI Navigation
von: Hu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Hu, Zhiyuan, et al.
Veröffentlicht: (2025)
MAgIC: Investigation of Large Language Model Powered Multi-Agent in Cognition, Adaptability, Rationality and Collaboration
von: Xu, Lin, et al.
Veröffentlicht: (2023)
von: Xu, Lin, et al.
Veröffentlicht: (2023)
BPP-Search: Enhancing Tree of Thought Reasoning for Mathematical Modeling Problem Solving
von: Wang, Teng, et al.
Veröffentlicht: (2024)
von: Wang, Teng, et al.
Veröffentlicht: (2024)
Knowledge Boundary of Large Language Models: A Survey
von: Li, Moxin, et al.
Veröffentlicht: (2024)
von: Li, Moxin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models
von: Shi, Wenhao, et al.
Veröffentlicht: (2024) -
GalleryGPT: Analyzing Paintings with Large Multimodal Models
von: Bin, Yi, et al.
Veröffentlicht: (2024) -
Hint-before-Solving Prompting: Guiding LLMs to Effectively Utilize Encoded Knowledge
von: Fu, Jinlan, et al.
Veröffentlicht: (2024) -
Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet
von: Zhao, James Xu, et al.
Veröffentlicht: (2025) -
Dipper: Diversity in Prompts for Producing Large Language Model Ensembles in Reasoning tasks
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2024)