TopoBench: Benchmarking LLMs on Hard Topological Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Maniparambil, Mayug, Hoehing, Nils, Kapuriya, Janak, Karuvally, Arjun, Rushe, Ellen, Ventresque, Anthony, O'Connor, Noel, Reid, Fergal |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks
by: Hoehing, Nils, et al.
Published: (2025)
by: Hoehing, Nils, et al.
Published: (2025)
TopoBench: A Framework for Benchmarking Topological Deep Learning
by: Telyatnikov, Lev, et al.
Published: (2024)
by: Telyatnikov, Lev, et al.
Published: (2024)
Metamorphic Testing for Pose Estimation Systems
by: Duran, Matias, et al.
Published: (2025)
by: Duran, Matias, et al.
Published: (2025)
Hold-One-Shot-Out (HOSO) for Validation-Free Few-Shot CLIP Adapters
by: Vorster, Chris, et al.
Published: (2026)
by: Vorster, Chris, et al.
Published: (2026)
Underrepresented in Foundation Model Pretraining Data? A One-Shot Probe
by: Vorster, Chris, et al.
Published: (2026)
by: Vorster, Chris, et al.
Published: (2026)
Are Natural-Domain Foundation Models Effective for Accelerated Cardiac MRI Reconstruction?
by: Hashmi, Anam, et al.
Published: (2026)
by: Hashmi, Anam, et al.
Published: (2026)
Ensemble Learning with Sparse Hypercolumns
by: Dietlmeier, Julia, et al.
Published: (2026)
by: Dietlmeier, Julia, et al.
Published: (2026)
Pinpoint Counterfactuals: Reducing social bias in foundation models via localized counterfactual generation
by: Sirotkin, Kirill, et al.
Published: (2024)
by: Sirotkin, Kirill, et al.
Published: (2024)
Harnessing Frozen Unimodal Encoders for Flexible Multimodal Alignment
by: Maniparambil, Mayug, et al.
Published: (2024)
by: Maniparambil, Mayug, et al.
Published: (2024)
A Progressive Evaluation Framework for Multicultural Analysis of Story Visualization
by: Kapuriya, Janak, et al.
Published: (2025)
by: Kapuriya, Janak, et al.
Published: (2025)
Do Vision and Language Encoders Represent the World Similarly?
by: Maniparambil, Mayug, et al.
Published: (2024)
by: Maniparambil, Mayug, et al.
Published: (2024)
Test-Time Adaptation with SaLIP: A Cascade of SAM and CLIP for Zero shot Medical Image Segmentation
by: Aleem, Sidra, et al.
Published: (2024)
by: Aleem, Sidra, et al.
Published: (2024)
Spiritual-LLM : Gita Inspired Mental Health Therapy In the Era of LLMs
by: Kapuriya, Janak, et al.
Published: (2025)
by: Kapuriya, Janak, et al.
Published: (2025)
Tiny Recursive Reasoning with Mamba-2 Attention Hybrid
by: Wang, Wenlong, et al.
Published: (2026)
by: Wang, Wenlong, et al.
Published: (2026)
Exploring the Role of Diversity in Example Selection for In-Context Learning
by: Kapuriya, Janak, et al.
Published: (2025)
by: Kapuriya, Janak, et al.
Published: (2025)
Hidden Traveling Waves bind Working Memory Variables in Recurrent Neural Networks
by: Karuvally, Arjun, et al.
Published: (2024)
by: Karuvally, Arjun, et al.
Published: (2024)
Semantic Frame Aggregation-based Transformer for Live Video Comment Generation
by: Fatima, Anam, et al.
Published: (2025)
by: Fatima, Anam, et al.
Published: (2025)
DW-Bench: Benchmarking LLMs on Data Warehouse Graph Topology Reasoning
by: Ahmed, Ahmed G. A. H, et al.
Published: (2026)
by: Ahmed, Ahmed G. A. H, et al.
Published: (2026)
Exponential Dynamic Energy Network for High Capacity Sequence Memory
by: Karuvally, Arjun, et al.
Published: (2025)
by: Karuvally, Arjun, et al.
Published: (2025)
Metformin and exercise prescription: Time for evidence‐based guidance
by: Kellie Hoehing, et al.
Published: (2024)
by: Kellie Hoehing, et al.
Published: (2024)
TopoPrune: Robust Data Pruning via Unified Latent Space Topology
by: Roy, Arjun, et al.
Published: (2026)
by: Roy, Arjun, et al.
Published: (2026)
DNR Bench: Benchmarking Over-Reasoning in Reasoning LLMs
by: Hashemi, Masoud, et al.
Published: (2025)
by: Hashemi, Masoud, et al.
Published: (2025)
ChaosBench-Logic: A Benchmark for Logical and Symbolic Reasoning on Chaotic Dynamical Systems
by: Thomas, Noel
Published: (2026)
by: Thomas, Noel
Published: (2026)
Topo2Seq: Enhanced Topology Reasoning via Topology Sequence Learning
by: Yang, Yiming, et al.
Published: (2025)
by: Yang, Yiming, et al.
Published: (2025)
HardSecBench: Benchmarking the Security Awareness of LLMs for Hardware Code Generation
by: Chen, Qirui, et al.
Published: (2026)
by: Chen, Qirui, et al.
Published: (2026)
Transient Dynamics in Lattices of Differentiating Ring Oscillators
by: DelMastro, Peter, et al.
Published: (2025)
by: DelMastro, Peter, et al.
Published: (2025)
CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
by: Lin, Zicheng, et al.
Published: (2024)
by: Lin, Zicheng, et al.
Published: (2024)
DPrivBench: Benchmarking LLMs' Reasoning for Differential Privacy
by: Wang, Erchi, et al.
Published: (2026)
by: Wang, Erchi, et al.
Published: (2026)
Coordinates from Context: Using LLMs to Ground Complex Location References
by: Masis, Tessa, et al.
Published: (2025)
by: Masis, Tessa, et al.
Published: (2025)
TopoStreamer: Temporal Lane Segment Topology Reasoning in Autonomous Driving
by: Yang, Yiming, et al.
Published: (2025)
by: Yang, Yiming, et al.
Published: (2025)
TopoLogic: An Interpretable Pipeline for Lane Topology Reasoning on Driving Scenes
by: Fu, Yanping, et al.
Published: (2024)
by: Fu, Yanping, et al.
Published: (2024)
RiddleBench: A New Generative Reasoning Benchmark for LLMs
by: Halder, Deepon, et al.
Published: (2025)
by: Halder, Deepon, et al.
Published: (2025)
FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning
by: Wang, Zeyu, et al.
Published: (2026)
by: Wang, Zeyu, et al.
Published: (2026)
HiBench: Benchmarking LLMs Capability on Hierarchical Structure Reasoning
by: Jiang, Zhuohang, et al.
Published: (2025)
by: Jiang, Zhuohang, et al.
Published: (2025)
FinTradeBench: A Financial Reasoning Benchmark for LLMs
by: Agrawal, Yogesh, et al.
Published: (2026)
by: Agrawal, Yogesh, et al.
Published: (2026)
DeonticBench: A Benchmark for Reasoning over Rules
by: Dou, Guangyao, et al.
Published: (2026)
by: Dou, Guangyao, et al.
Published: (2026)
TopoPoint: Enhance Topology Reasoning via Endpoint Detection in Autonomous Driving
by: Fu, Yanping, et al.
Published: (2025)
by: Fu, Yanping, et al.
Published: (2025)
RelTopo: Multi-Level Relational Modeling for Driving Scene Topology Reasoning
by: Luo, Yueru, et al.
Published: (2025)
by: Luo, Yueru, et al.
Published: (2025)
MM-PhyQA: Multimodal Physics Question-Answering With Multi-Image CoT Prompting
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
GeoGramBench: Benchmarking the Geometric Program Reasoning in Modern LLMs
by: Luo, Shixian, et al.
Published: (2025)
by: Luo, Shixian, et al.
Published: (2025)
Similar Items
-
Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks
by: Hoehing, Nils, et al.
Published: (2025) -
TopoBench: A Framework for Benchmarking Topological Deep Learning
by: Telyatnikov, Lev, et al.
Published: (2024) -
Metamorphic Testing for Pose Estimation Systems
by: Duran, Matias, et al.
Published: (2025) -
Hold-One-Shot-Out (HOSO) for Validation-Free Few-Shot CLIP Adapters
by: Vorster, Chris, et al.
Published: (2026) -
Underrepresented in Foundation Model Pretraining Data? A One-Shot Probe
by: Vorster, Chris, et al.
Published: (2026)