SPAN: Learning Similarity between Scene Graphs and Images with Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Cong, Yuren, Liao, Wentong, Rosenhahn, Bodo, Yang, Michael Ying |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FDSG: Forecasting Dynamic Scene Graphs
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
Robust Shape Fitting for 3D Scene Abstraction
by: Kluger, Florian, et al.
Published: (2024)
by: Kluger, Florian, et al.
Published: (2024)
Pruning by Block Benefit: Exploring the Properties of Vision Transformer Blocks during Domain Adaptation
by: Glandorf, Patrick, et al.
Published: (2025)
by: Glandorf, Patrick, et al.
Published: (2025)
BUSSARD: Normalizing Flows for Bijective Universal Scene-Specific Anomalous Relationship Detection
by: Schween, Melissa, et al.
Published: (2026)
by: Schween, Melissa, et al.
Published: (2026)
HydraMix: Multi-Image Feature Mixing for Small Data Image Classification
by: Reinders, Christoph, et al.
Published: (2025)
by: Reinders, Christoph, et al.
Published: (2025)
Improving 3D Foot Motion Reconstruction in Markerless Monocular Human Motion Capture
by: Wehrbein, Tom, et al.
Published: (2026)
by: Wehrbein, Tom, et al.
Published: (2026)
PARSAC: Accelerating Robust Multi-Model Fitting with Parallel Sample Consensus
by: Kluger, Florian, et al.
Published: (2024)
by: Kluger, Florian, et al.
Published: (2024)
Segment Any Object Model (SAOM): Real-to-Simulation Fine-Tuning Strategy for Multi-Class Multi-Instance Segmentation
by: Khan, Mariia, et al.
Published: (2024)
by: Khan, Mariia, et al.
Published: (2024)
Multi-Flow: Multi-View-Enriched Normalizing Flows for Industrial Anomaly Detection
by: Kruse, Mathis, et al.
Published: (2025)
by: Kruse, Mathis, et al.
Published: (2025)
Video Patch Pruning: Efficient Video Instance Segmentation via Early Token Reduction
by: Glandorf, Patrick, et al.
Published: (2026)
by: Glandorf, Patrick, et al.
Published: (2026)
UncertainSAM: Fast and Efficient Uncertainty Quantification of the Segment Anything Model
by: Kaiser, Timo, et al.
Published: (2025)
by: Kaiser, Timo, et al.
Published: (2025)
Cell Tracking according to Biological Needs -- Strong Mitosis-aware Multi-Hypothesis Tracker with Aleatoric Uncertainty
by: Kaiser, Timo, et al.
Published: (2024)
by: Kaiser, Timo, et al.
Published: (2024)
CHOTA: A Higher Order Accuracy Metric for Cell Tracking
by: Kaiser, Timo, et al.
Published: (2024)
by: Kaiser, Timo, et al.
Published: (2024)
Multi-Object Tracking Retrieval with LLaVA-Video: A Training-Free Solution to MOT25-StAG Challenge
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
DINO-QPM: Adapting Visual Foundation Models for Globally Interpretable Image Classification
by: Zimmermann, Robert, et al.
Published: (2026)
by: Zimmermann, Robert, et al.
Published: (2026)
Interpretable Decision-Making for End-to-End Autonomous Driving
by: Mirzaie, Mona, et al.
Published: (2025)
by: Mirzaie, Mona, et al.
Published: (2025)
Q-SENN: Quantized Self-Explaining Neural Networks
by: Norrenbrock, Thomas, et al.
Published: (2023)
by: Norrenbrock, Thomas, et al.
Published: (2023)
Utilizing Uncertainty in 2D Pose Detectors for Probabilistic 3D Human Mesh Recovery
by: Wehrbein, Tom, et al.
Published: (2024)
by: Wehrbein, Tom, et al.
Published: (2024)
Personalized 3D Human Pose and Shape Refinement
by: Wehrbein, Tom, et al.
Published: (2024)
by: Wehrbein, Tom, et al.
Published: (2024)
FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing
by: Cong, Yuren, et al.
Published: (2023)
by: Cong, Yuren, et al.
Published: (2023)
From Gameplay Traces to Game Mechanics: Causal Induction with Large Language Models
by: Jiwatode, Mohit, et al.
Published: (2026)
by: Jiwatode, Mohit, et al.
Published: (2026)
SplatPose & Detect: Pose-Agnostic 3D Anomaly Detection
by: Kruse, Mathis, et al.
Published: (2024)
by: Kruse, Mathis, et al.
Published: (2024)
Explore Internal and External Similarity for Single Image Deraining with Graph Neural Networks
by: Wang, Cong, et al.
Published: (2024)
by: Wang, Cong, et al.
Published: (2024)
SPAN: Spatial-Projection Alignment for Monocular 3D Object Detection
by: Wang, Yifan, et al.
Published: (2025)
by: Wang, Yifan, et al.
Published: (2025)
Improved Convex Decomposition with Ensembling and Negative Primitives
by: Vavilala, Vaibhav, et al.
Published: (2024)
by: Vavilala, Vaibhav, et al.
Published: (2024)
QPM: Discrete Optimization for Globally Interpretable Image Classification
by: Norrenbrock, Thomas, et al.
Published: (2025)
by: Norrenbrock, Thomas, et al.
Published: (2025)
SG-Adapter: Enhancing Text-to-Image Generation with Scene Graph Guidance
by: Shen, Guibao, et al.
Published: (2024)
by: Shen, Guibao, et al.
Published: (2024)
CHiQPM: Calibrated Hierarchical Interpretable Image Classification
by: Norrenbrock, Thomas, et al.
Published: (2025)
by: Norrenbrock, Thomas, et al.
Published: (2025)
GenTron: Diffusion Transformers for Image and Video Generation
by: Chen, Shoufa, et al.
Published: (2023)
by: Chen, Shoufa, et al.
Published: (2023)
SGEdit: Bridging LLM with Text2Image Generative Model for Scene Graph-based Image Editing
by: Zhang, Zhiyuan, et al.
Published: (2024)
by: Zhang, Zhiyuan, et al.
Published: (2024)
Image Similarity using An Ensemble of Context-Sensitive Models
by: Liao, Zukang, et al.
Published: (2024)
by: Liao, Zukang, et al.
Published: (2024)
Joint Generative Modeling of Grounded Scene Graphs and Images via Diffusion Models
by: Xu, Bicheng, et al.
Published: (2024)
by: Xu, Bicheng, et al.
Published: (2024)
SPAN: Continuous Modeling of Suspicion Progression for Temporal Intention Localization
by: Hu, Xinyi, et al.
Published: (2025)
by: Hu, Xinyi, et al.
Published: (2025)
Scene Graph Generation via Conditional Random Fields
by: Cong, Weilin, et al.
Published: (2018)
by: Cong, Weilin, et al.
Published: (2018)
Image Patch-Matching with Graph-Based Learning in Street Scenes
by: She, Rui, et al.
Published: (2023)
by: She, Rui, et al.
Published: (2023)
WorldAfford: Affordance Grounding based on Natural Language Instructions
by: Chen, Changmao, et al.
Published: (2024)
by: Chen, Changmao, et al.
Published: (2024)
Motion-aware Contrastive Learning for Temporal Panoptic Scene Graph Generation
by: Nguyen, Thong Thanh, et al.
Published: (2024)
by: Nguyen, Thong Thanh, et al.
Published: (2024)
Inst3D-LMM: Instance-Aware 3D Scene Understanding with Multi-modal Instruction Tuning
by: Yu, Hanxun, et al.
Published: (2025)
by: Yu, Hanxun, et al.
Published: (2025)
Pairwise Similarity Regularization for Semi-supervised Graph Medical Image Segmentation
by: Zhou, Jialu, et al.
Published: (2025)
by: Zhou, Jialu, et al.
Published: (2025)
Graph-Guided Dual-Level Augmentation for 3D Scene Segmentation
by: Lin, Hongbin, et al.
Published: (2025)
by: Lin, Hongbin, et al.
Published: (2025)
Similar Items
-
FDSG: Forecasting Dynamic Scene Graphs
by: Yang, Yi, et al.
Published: (2025) -
Robust Shape Fitting for 3D Scene Abstraction
by: Kluger, Florian, et al.
Published: (2024) -
Pruning by Block Benefit: Exploring the Properties of Vision Transformer Blocks during Domain Adaptation
by: Glandorf, Patrick, et al.
Published: (2025) -
BUSSARD: Normalizing Flows for Bijective Universal Scene-Specific Anomalous Relationship Detection
by: Schween, Melissa, et al.
Published: (2026) -
HydraMix: Multi-Image Feature Mixing for Small Data Image Classification
by: Reinders, Christoph, et al.
Published: (2025)