Explain with Visual Keypoints Like a Real Mentor! A Benchmark for Multimodal Solution Explanation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Park, Jaewoo, Park, Jungyang, Jang, Dongju, Chung, Jiwan, Yoo, Byungwoo, Shin, Jaewoo, Park, Seonjoon, Kim, Taehyeong, Yu, Youngjae |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tracing Mathematical Proficiency Through Problem-Solving Processes
von: Park, Jungyang, et al.
Veröffentlicht: (2025)
von: Park, Jungyang, et al.
Veröffentlicht: (2025)
GuideDog: A Real-World Egocentric Multimodal Dataset for Blind and Low-Vision Accessibility-Aware Guidance
von: Kim, Junhyeok, et al.
Veröffentlicht: (2025)
von: Kim, Junhyeok, et al.
Veröffentlicht: (2025)
Teaching Metric Distance to Discrete Autoregressive Language Models
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway
von: Park, Jueon, et al.
Veröffentlicht: (2026)
von: Park, Jueon, et al.
Veröffentlicht: (2026)
Zero-shot Multimodal Document Retrieval via Cross-modal Question Generation
von: Choi, Yejin, et al.
Veröffentlicht: (2025)
von: Choi, Yejin, et al.
Veröffentlicht: (2025)
Background-Aware Defect Generation for Robust Industrial Anomaly Detection
von: Cho, Youngjae, et al.
Veröffentlicht: (2024)
von: Cho, Youngjae, et al.
Veröffentlicht: (2024)
Snell meets Fagnano. Path optimization through an imperfect mirror
von: Arnold, Maxim, et al.
Veröffentlicht: (2025)
von: Arnold, Maxim, et al.
Veröffentlicht: (2025)
Keep Away From It! Examining the Contamination Effect of Insect‐Based Foods in Retail Environments
von: Zining Wang, et al.
Veröffentlicht: (2025)
von: Zining Wang, et al.
Veröffentlicht: (2025)
Challenges and Lessons from MIDOG 2025: A Two-Stage Approach to Domain-Robust Mitotic Figure Detection
von: Song, Euiseop, et al.
Veröffentlicht: (2025)
von: Song, Euiseop, et al.
Veröffentlicht: (2025)
ASGuard: Activation-Scaling Guard to Mitigate Targeted Jailbreaking Attack
von: Park, Yein, et al.
Veröffentlicht: (2025)
von: Park, Yein, et al.
Veröffentlicht: (2025)
Ensuring Functional Correctness of Large Code Models with Selective Generation
von: Jeong, Jaewoo, et al.
Veröffentlicht: (2025)
von: Jeong, Jaewoo, et al.
Veröffentlicht: (2025)
MolDeTox: Evaluating Language Model's Stepwise Fragment Editing for Molecular Detoxification
von: Park, Jueon, et al.
Veröffentlicht: (2026)
von: Park, Jueon, et al.
Veröffentlicht: (2026)
On a lower bound of Hausdorff dimension of weighted singular vectors
von: Kim, Taehyeong, et al.
Veröffentlicht: (2022)
von: Kim, Taehyeong, et al.
Veröffentlicht: (2022)
Leveraging Positional Encoding for Robust Multi-Reference-Based Object 6D Pose Estimation
von: Park, Jaewoo, et al.
Veröffentlicht: (2024)
von: Park, Jaewoo, et al.
Veröffentlicht: (2024)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models
von: Park, Junyoung, et al.
Veröffentlicht: (2026)
von: Park, Junyoung, et al.
Veröffentlicht: (2026)
Beyond Attack Success Rate: Temporal Logit Observability for LLM Safety Failures
von: Park, Junyoung, et al.
Veröffentlicht: (2026)
von: Park, Junyoung, et al.
Veröffentlicht: (2026)
Are Any-to-Any Models More Consistent Across Modality Transfers Than Specialists?
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
On the rank index of projective curves of almost minimal degree
von: Jung, Jaewoo, et al.
Veröffentlicht: (2024)
von: Jung, Jaewoo, et al.
Veröffentlicht: (2024)
Algorithms for Models with Intractable Normalizing Functions
von: Haran, Murali, et al.
Veröffentlicht: (2026)
von: Haran, Murali, et al.
Veröffentlicht: (2026)
EP-HDC: Hyperdimensional Computing with Encrypted Parameters for High-Throughput Privacy-Preserving Inference
von: Park, Jaewoo, et al.
Veröffentlicht: (2025)
von: Park, Jaewoo, et al.
Veröffentlicht: (2025)
On quadratic persistence and Pythagoras numbers of totally real projective varieties
von: Han, Jong In, et al.
Veröffentlicht: (2025)
von: Han, Jong In, et al.
Veröffentlicht: (2025)
Adaptive Replay Buffer for Offline-to-Online Reinforcement Learning
von: Song, Chihyeon, et al.
Veröffentlicht: (2025)
von: Song, Chihyeon, et al.
Veröffentlicht: (2025)
Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training
von: Park, Yein, et al.
Veröffentlicht: (2025)
von: Park, Yein, et al.
Veröffentlicht: (2025)
Selective Vision is the Challenge for Visual Reasoning: A Benchmark for Visual Argument Understanding
von: Chung, Jiwan, et al.
Veröffentlicht: (2024)
von: Chung, Jiwan, et al.
Veröffentlicht: (2024)
A Stein Gradient Descent Approach for Doubly Intractable Distributions
von: Lee, Heesang, et al.
Veröffentlicht: (2024)
von: Lee, Heesang, et al.
Veröffentlicht: (2024)
Subtle Risks, Critical Failures: A Framework for Diagnosing Physical Safety of LLMs for Embodied Decision Making
von: Son, Yejin, et al.
Veröffentlicht: (2025)
von: Son, Yejin, et al.
Veröffentlicht: (2025)
A More Accurate Approximation of Activation Function with Few Spikes Neurons
von: Jeong, Dayena, et al.
Veröffentlicht: (2024)
von: Jeong, Dayena, et al.
Veröffentlicht: (2024)
Benchmarking Foundation Models on Exceptional Cases: Dataset Creation and Validation
von: Kang, Suho, et al.
Veröffentlicht: (2024)
von: Kang, Suho, et al.
Veröffentlicht: (2024)
Multi-agent Long-term 3D Human Pose Forecasting via Interaction-aware Trajectory Conditioning
von: Jeong, Jaewoo, et al.
Veröffentlicht: (2024)
von: Jeong, Jaewoo, et al.
Veröffentlicht: (2024)
Fast Computer Model Calibration using Annealed and Transformed Variational Inference
von: Cho, Dongkyu Derek, et al.
Veröffentlicht: (2022)
von: Cho, Dongkyu Derek, et al.
Veröffentlicht: (2022)
Understanding Open-Set Recognition by Jacobian Norm and Inter-Class Separation
von: Park, Jaewoo, et al.
Veröffentlicht: (2022)
von: Park, Jaewoo, et al.
Veröffentlicht: (2022)
Does Time Have Its Place? Temporal Heads: Where Language Models Recall Time-specific Information
von: Park, Yein, et al.
Veröffentlicht: (2025)
von: Park, Yein, et al.
Veröffentlicht: (2025)
Automated Kernel Discovery Towards Understanding High-dimensional Bayesian Optimization
von: Yun, Taeyoung, et al.
Veröffentlicht: (2026)
von: Yun, Taeyoung, et al.
Veröffentlicht: (2026)
Diffusion Fine-Tuning via Reparameterized Policy Gradient of the Soft Q-Function
von: Kang, Hyeongyu, et al.
Veröffentlicht: (2025)
von: Kang, Hyeongyu, et al.
Veröffentlicht: (2025)
Adversarial Feature Alignment: Balancing Robustness and Accuracy in Deep Learning via Adversarial Training
von: Park, Leo Hyun, et al.
Veröffentlicht: (2024)
von: Park, Leo Hyun, et al.
Veröffentlicht: (2024)
RefPose: Leveraging Reference Geometric Correspondences for Accurate 6D Pose Estimation of Unseen Objects
von: Kim, Jaeguk, et al.
Veröffentlicht: (2025)
von: Kim, Jaeguk, et al.
Veröffentlicht: (2025)
Families of Two-Impulse Optimal Rendezvous Transfers Between Elliptic Orbits
von: Park, Beom, et al.
Veröffentlicht: (2026)
von: Park, Beom, et al.
Veröffentlicht: (2026)
A Delayed Acceptance Auxiliary Variable MCMC for Spatial Models with Intractable Likelihood Function
von: Lee, Jong Hyeon, et al.
Veröffentlicht: (2025)
von: Lee, Jong Hyeon, et al.
Veröffentlicht: (2025)
SpatialBoost: Enhancing Visual Representation through Language-Guided Reasoning
von: Jeon, Byungwoo, et al.
Veröffentlicht: (2026)
von: Jeon, Byungwoo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Tracing Mathematical Proficiency Through Problem-Solving Processes
von: Park, Jungyang, et al.
Veröffentlicht: (2025) -
GuideDog: A Real-World Egocentric Multimodal Dataset for Blind and Low-Vision Accessibility-Aware Guidance
von: Kim, Junhyeok, et al.
Veröffentlicht: (2025) -
Teaching Metric Distance to Discrete Autoregressive Language Models
von: Chung, Jiwan, et al.
Veröffentlicht: (2025) -
ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway
von: Park, Jueon, et al.
Veröffentlicht: (2026) -
Zero-shot Multimodal Document Retrieval via Cross-modal Question Generation
von: Choi, Yejin, et al.
Veröffentlicht: (2025)