Plane Geometry Problem Solving with Multi-modal Reasoning: A Survey
Fuente:
arXiv
Saved in:
| Main Authors: | Cho, Seunghyuk, Qin, Zhenyue, Liu, Yang, Choi, Youngbin, Lee, Seungbeom, Kim, Dongwoo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GeoDANO: Geometric VLM with Domain Agnostic Vision Encoder
by: Cho, Seunghyuk, et al.
Published: (2025)
by: Cho, Seunghyuk, et al.
Published: (2025)
Feature Unlearning for Pre-trained GANs and VAEs
by: Moon, Saemi, et al.
Published: (2023)
by: Moon, Saemi, et al.
Published: (2023)
Routing by Reaching: Composition of Pre-trained GFlowNets for Multi-Objective Generation
by: Yoon, Seokwon, et al.
Published: (2026)
by: Yoon, Seokwon, et al.
Published: (2026)
A Survey of Deep Learning for Geometry Problem Solving
by: Ma, Jianzhe, et al.
Published: (2025)
by: Ma, Jianzhe, et al.
Published: (2025)
Holistic Unlearning Benchmark: A Multi-Faceted Evaluation for Text-to-Image Diffusion Model Unlearning
by: Moon, Saemi, et al.
Published: (2024)
by: Moon, Saemi, et al.
Published: (2024)
In-Place Feedback: Reliable Refinement for Multi-Turn Expert-LLM Collaboration
by: Choi, Youngbin, et al.
Published: (2025)
by: Choi, Youngbin, et al.
Published: (2025)
Multi-modal Data Spectrum: Multi-modal Datasets are Multi-dimensional
by: Madaan, Divyam, et al.
Published: (2025)
by: Madaan, Divyam, et al.
Published: (2025)
NavFormer: IGRF Forecasting in Moving Coordinate Frames
by: Hwang, Yoontae, et al.
Published: (2026)
by: Hwang, Yoontae, et al.
Published: (2026)
Concepts or Skills? Rethinking Instruction Selection for Multi-modal Models
by: Bai, Andrew, et al.
Published: (2025)
by: Bai, Andrew, et al.
Published: (2025)
Multi-level and Multi-modal Action Anticipation
by: Kim, Seulgi, et al.
Published: (2025)
by: Kim, Seulgi, et al.
Published: (2025)
TK-Planes: Tiered K-Planes with High Dimensional Feature Vectors for Dynamic UAV-based Scenes
by: Maxey, Christopher, et al.
Published: (2024)
by: Maxey, Christopher, et al.
Published: (2024)
Multi-level Cross-modal Alignment for Image Clustering
by: Qiu, Liping, et al.
Published: (2024)
by: Qiu, Liping, et al.
Published: (2024)
Diffusion-based Data Augmentation and Knowledge Distillation with Generated Soft Labels Solving Data Scarcity Problems of SAR Oil Spill Segmentation
by: Moon, Jaeho, et al.
Published: (2024)
by: Moon, Jaeho, et al.
Published: (2024)
VideoHallu: Evaluating and Mitigating Multi-modal Hallucinations on Synthetic Video Understanding
by: Li, Zongxia, et al.
Published: (2025)
by: Li, Zongxia, et al.
Published: (2025)
Concept Unlearning via Cross-Attention Activation Projection for Diffusion Models
by: Moon, Saemi, et al.
Published: (2026)
by: Moon, Saemi, et al.
Published: (2026)
Integrating Intermediate Layer Optimization and Projected Gradient Descent for Solving Inverse Problems with Diffusion Models
by: Zheng, Yang, et al.
Published: (2025)
by: Zheng, Yang, et al.
Published: (2025)
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation
by: Kang, Wonjun, et al.
Published: (2025)
by: Kang, Wonjun, et al.
Published: (2025)
Simultaneous Long-tailed Recognition and Multi-modal Fusion for Highly Imbalanced Multi-modal Data
by: Yoon, Heegeon, et al.
Published: (2026)
by: Yoon, Heegeon, et al.
Published: (2026)
On the Multi-modal Vulnerability of Diffusion Models
by: Yang, Dingcheng, et al.
Published: (2024)
by: Yang, Dingcheng, et al.
Published: (2024)
Neural Predictor-Corrector: Solving Homotopy Problems with Reinforcement Learning
by: Mai, Jiayao, et al.
Published: (2026)
by: Mai, Jiayao, et al.
Published: (2026)
Countering Multi-modal Representation Collapse through Rank-targeted Fusion
by: Kim, Seulgi, et al.
Published: (2025)
by: Kim, Seulgi, et al.
Published: (2025)
Jointly Modeling Inter- & Intra-Modality Dependencies for Multi-modal Learning
by: Madaan, Divyam, et al.
Published: (2024)
by: Madaan, Divyam, et al.
Published: (2024)
CoPL: Collaborative Preference Learning for Personalizing LLMs
by: Choi, Youngbin, et al.
Published: (2025)
by: Choi, Youngbin, et al.
Published: (2025)
One Self-Configurable Model to Solve Many Abstract Visual Reasoning Problems
by: Małkiński, Mikołaj, et al.
Published: (2023)
by: Małkiński, Mikołaj, et al.
Published: (2023)
Repulsive Latent Score Distillation for Solving Inverse Problems
by: Zilberstein, Nicolas, et al.
Published: (2024)
by: Zilberstein, Nicolas, et al.
Published: (2024)
ReLayout: Integrating Relation Reasoning for Content-aware Layout Generation with Multi-modal Large Language Models
by: Tian, Jiaxu, et al.
Published: (2025)
by: Tian, Jiaxu, et al.
Published: (2025)
Learnable Cross-modal Knowledge Distillation for Multi-modal Learning with Missing Modality
by: Wang, Hu, et al.
Published: (2023)
by: Wang, Hu, et al.
Published: (2023)
Harnessing Input-Adaptive Inference for Efficient VLN
by: Kang, Dongwoo, et al.
Published: (2025)
by: Kang, Dongwoo, et al.
Published: (2025)
Cross-modal Active Complementary Learning with Self-refining Correspondence
by: Qin, Yang, et al.
Published: (2023)
by: Qin, Yang, et al.
Published: (2023)
Measurement Geometry and Design for Trustworthy Generative Inverse Problems
by: Jin, Pengfei, et al.
Published: (2026)
by: Jin, Pengfei, et al.
Published: (2026)
Representation-Centric Survey of Supervised Skeletal Action Recognition and the New Benchmark
by: Liu, Yang, et al.
Published: (2022)
by: Liu, Yang, et al.
Published: (2022)
Towards Multi-modal Transformers in Federated Learning
by: Sun, Guangyu, et al.
Published: (2024)
by: Sun, Guangyu, et al.
Published: (2024)
Multi-modal learning for geospatial vegetation forecasting
by: Benson, Vitus, et al.
Published: (2023)
by: Benson, Vitus, et al.
Published: (2023)
Incorporating Pre-trained Diffusion Models in Solving the Schrödinger Bridge Problem
by: Tang, Zhicong, et al.
Published: (2025)
by: Tang, Zhicong, et al.
Published: (2025)
DMPlug: A Plug-in Method for Solving Inverse Problems with Diffusion Models
by: Wang, Hengkang, et al.
Published: (2024)
by: Wang, Hengkang, et al.
Published: (2024)
HFI: A unified framework for training-free detection and implicit watermarking of latent diffusion model generated images
by: Choi, Sungik, et al.
Published: (2024)
by: Choi, Sungik, et al.
Published: (2024)
Delving into Multi-modal Multi-task Foundation Models for Road Scene Understanding: From Learning Paradigm Perspectives
by: Luo, Sheng, et al.
Published: (2024)
by: Luo, Sheng, et al.
Published: (2024)
Transductive Generalization via Optimal Transport and Its Application to Graph Node Classification
by: Park, MoonJeong, et al.
Published: (2026)
by: Park, MoonJeong, et al.
Published: (2026)
cadrille: Multi-modal CAD Reconstruction with Reinforcement Learning
by: Kolodiazhnyi, Maksim, et al.
Published: (2025)
by: Kolodiazhnyi, Maksim, et al.
Published: (2025)
Diffusion Models for Solving Inverse Problems via Posterior Sampling with Piecewise Guidance
by: Mohseni-Sehdeh, Saeed, et al.
Published: (2025)
by: Mohseni-Sehdeh, Saeed, et al.
Published: (2025)
Similar Items
-
GeoDANO: Geometric VLM with Domain Agnostic Vision Encoder
by: Cho, Seunghyuk, et al.
Published: (2025) -
Feature Unlearning for Pre-trained GANs and VAEs
by: Moon, Saemi, et al.
Published: (2023) -
Routing by Reaching: Composition of Pre-trained GFlowNets for Multi-Objective Generation
by: Yoon, Seokwon, et al.
Published: (2026) -
A Survey of Deep Learning for Geometry Problem Solving
by: Ma, Jianzhe, et al.
Published: (2025) -
Holistic Unlearning Benchmark: A Multi-Faceted Evaluation for Text-to-Image Diffusion Model Unlearning
by: Moon, Saemi, et al.
Published: (2024)