Nodes Are Early, Edges Are Late: Probing Diagram Representations in Large Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yoshida, Haruto, Kudo, Keito, Aoki, Yoichi, Tanaka, Ryota, Saito, Itsumi, Sakaguchi, Keisuke, Inui, Kentaro |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
First Heuristic Then Rational: Dynamic Use of Heuristics in Language Model Reasoning
by: Aoki, Yoichi, et al.
Published: (2024)
by: Aoki, Yoichi, et al.
Published: (2024)
ACORN: Aspect-wise Commonsense Reasoning Explanation Evaluation
by: Brassard, Ana, et al.
Published: (2024)
by: Brassard, Ana, et al.
Published: (2024)
LLMs Faithfully and Iteratively Compute Answers During CoT: A Systematic Analysis With Multi-step Arithmetics
by: Kudo, Keito, et al.
Published: (2024)
by: Kudo, Keito, et al.
Published: (2024)
Weight-based Analysis of Detokenization in Language Models: Understanding the First Stage of Inference Without Inference
by: Kamoda, Go, et al.
Published: (2025)
by: Kamoda, Go, et al.
Published: (2025)
J-UniMorph: Japanese Morphological Annotation through the Universal Feature Schema
by: Matsuzaki, Kosuke, et al.
Published: (2024)
by: Matsuzaki, Kosuke, et al.
Published: (2024)
Empirical Analysis of Large Vision-Language Models against Goal Hijacking via Visual Prompt Injection
by: Kimura, Subaru, et al.
Published: (2024)
by: Kimura, Subaru, et al.
Published: (2024)
SLVMEval: Synthetic Meta Evaluation Benchmark for Text-to-Long Video Generation
by: Matsuda, Ryosuke, et al.
Published: (2026)
by: Matsuda, Ryosuke, et al.
Published: (2026)
The Curse of Popularity: Popular Entities have Catastrophic Side Effects when Deleting Knowledge from Language Models
by: Takahashi, Ryosuke, et al.
Published: (2024)
by: Takahashi, Ryosuke, et al.
Published: (2024)
FinchGPT: a Transformer based language model for birdsong analysis
by: Kobayashi, Kosei, et al.
Published: (2025)
by: Kobayashi, Kosei, et al.
Published: (2025)
Monotonic Representation of Numeric Properties in Language Models
by: Heinzerling, Benjamin, et al.
Published: (2024)
by: Heinzerling, Benjamin, et al.
Published: (2024)
Representational Analysis of Binding in Language Models
by: Dai, Qin, et al.
Published: (2024)
by: Dai, Qin, et al.
Published: (2024)
Can Language Models Handle a Non-Gregorian Calendar? The Case of the Japanese wareki
by: Sasaki, Mutsumi, et al.
Published: (2025)
by: Sasaki, Mutsumi, et al.
Published: (2025)
Evaluation of the Automated Labeling Method for Taxonomic Nomenclature Through Prompt-Optimized Large Language Model
by: Inoshita, Keito, et al.
Published: (2025)
by: Inoshita, Keito, et al.
Published: (2025)
Cell-Based Representation of Relational Binding in Language Models
by: Dai, Qin, et al.
Published: (2026)
by: Dai, Qin, et al.
Published: (2026)
Annotating Errors in English Learners' Written Language Production: Advancing Automated Written Feedback Systems
by: Coyne, Steven, et al.
Published: (2025)
by: Coyne, Steven, et al.
Published: (2025)
Repetition Neurons: How Do Language Models Produce Repetitions?
by: Hiraoka, Tatsuya, et al.
Published: (2024)
by: Hiraoka, Tatsuya, et al.
Published: (2024)
Skeletonization-Based Adversarial Perturbations on Large Vision Language Model's Mathematical Text Recognition
by: Yoshida, Masatomo, et al.
Published: (2026)
by: Yoshida, Masatomo, et al.
Published: (2026)
VIFSS: View-Invariant and Figure Skating-Specific Pose Representation Learning for Temporal Action Segmentation
by: Tanaka, Ryota, et al.
Published: (2025)
by: Tanaka, Ryota, et al.
Published: (2025)
Instruction-Following Evaluation of Large Vision-Language Models
by: Shiono, Daiki, et al.
Published: (2025)
by: Shiono, Daiki, et al.
Published: (2025)
RealTime QA: What's the Answer Right Now?
by: Kasai, Jungo, et al.
Published: (2022)
by: Kasai, Jungo, et al.
Published: (2022)
Linear Representations of Hierarchical Concepts in Language Models
by: Sakata, Masaki, et al.
Published: (2026)
by: Sakata, Masaki, et al.
Published: (2026)
Unlocking Prompt Infilling Capability for Diffusion Language Models
by: Fujinuma, Yoshinari, et al.
Published: (2026)
by: Fujinuma, Yoshinari, et al.
Published: (2026)
Spelling-out is not Straightforward: LLMs' Capability of Tokenization from Token to Characters
by: Hiraoka, Tatsuya, et al.
Published: (2025)
by: Hiraoka, Tatsuya, et al.
Published: (2025)
Reconsidering Positional Supervision in Masked Diffusion Language Model Training
by: Ye, Mengyu, et al.
Published: (2026)
by: Ye, Mengyu, et al.
Published: (2026)
Large Language Models Are Human-Like Internally
by: Kuribayashi, Tatsuki, et al.
Published: (2025)
by: Kuribayashi, Tatsuki, et al.
Published: (2025)
Who Does This Name Remind You of ? Nationality Prediction via Large Language Model Associative Memory
by: Inoshita, Keito
Published: (2026)
by: Inoshita, Keito
Published: (2026)
Nationality and Region Prediction from Names: A Comparative Study of Neural Models and Large Language Models
by: Inoshita, Keito
Published: (2026)
by: Inoshita, Keito
Published: (2026)
LLMs Can Compensate for Deficiencies in Visual Representations
by: Takishita, Sho, et al.
Published: (2025)
by: Takishita, Sho, et al.
Published: (2025)
Hypergraph Vision Transformers: Images are More than Nodes, More than Edges
by: Fixelle, Joshua
Published: (2025)
by: Fixelle, Joshua
Published: (2025)
TopK Language Models
by: Takahashi, Ryosuke, et al.
Published: (2025)
by: Takahashi, Ryosuke, et al.
Published: (2025)
XDR-LVLM: An Explainable Vision-Language Large Model for Diabetic Retinopathy Diagnosis
by: Ito, Masato, et al.
Published: (2025)
by: Ito, Masato, et al.
Published: (2025)
Syntactic Learnability of Echo State Neural Language Models at Scale
by: Ueda, Ryo, et al.
Published: (2025)
by: Ueda, Ryo, et al.
Published: (2025)
A Large Collection of Model-generated Contradictory Responses for Consistency-aware Dialogue Systems
by: Sato, Shiki, et al.
Published: (2024)
by: Sato, Shiki, et al.
Published: (2024)
From Edges to Depth: Probing the Spatial Hierarchy in Vision Transformers
by: Sanghavi, Jainum
Published: (2026)
by: Sanghavi, Jainum
Published: (2026)
3D Pose-Based Temporal Action Segmentation for Figure Skating: A Fine-Grained and Jump Procedure-Aware Annotation Approach
by: Tanaka, Ryota, et al.
Published: (2024)
by: Tanaka, Ryota, et al.
Published: (2024)
Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge
by: Zhang, Ying, et al.
Published: (2025)
by: Zhang, Ying, et al.
Published: (2025)
Rectifying Belief Space via Unlearning to Harness LLMs' Reasoning
by: Niwa, Ayana, et al.
Published: (2025)
by: Niwa, Ayana, et al.
Published: (2025)
Evaluating Multimodal Large Language Models on Vertically Written Japanese Text
by: Sasagawa, Keito, et al.
Published: (2025)
by: Sasagawa, Keito, et al.
Published: (2025)
On Entity Identification in Language Models
by: Sakata, Masaki, et al.
Published: (2025)
by: Sakata, Masaki, et al.
Published: (2025)
TADA: Making Node-link Diagrams Accessible to Blind and Low-Vision People
by: Zhao, Yichun, et al.
Published: (2023)
by: Zhao, Yichun, et al.
Published: (2023)
Similar Items
-
First Heuristic Then Rational: Dynamic Use of Heuristics in Language Model Reasoning
by: Aoki, Yoichi, et al.
Published: (2024) -
ACORN: Aspect-wise Commonsense Reasoning Explanation Evaluation
by: Brassard, Ana, et al.
Published: (2024) -
LLMs Faithfully and Iteratively Compute Answers During CoT: A Systematic Analysis With Multi-step Arithmetics
by: Kudo, Keito, et al.
Published: (2024) -
Weight-based Analysis of Detokenization in Language Models: Understanding the First Stage of Inference Without Inference
by: Kamoda, Go, et al.
Published: (2025) -
J-UniMorph: Japanese Morphological Annotation through the Universal Feature Schema
by: Matsuzaki, Kosuke, et al.
Published: (2024)