ART: Adaptive Relation Tuning for Generalized Relation Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Sudhakaran, Gopika, Shindo, Hikaru, Schramowski, Patrick, Schaub-Meyer, Simone, Kersting, Kristian, Roth, Stefan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DeiSAM: Segment Anything with Deictic Prompting
by: Shindo, Hikaru, et al.
Published: (2024)
by: Shindo, Hikaru, et al.
Published: (2024)
DIAGen: Semantically Diverse Image Augmentation with Generative Models for Few-Shot Learning
by: Lingenberg, Tobias, et al.
Published: (2024)
by: Lingenberg, Tobias, et al.
Published: (2024)
Learning Differentiable Logic Programs for Abstract Visual Reasoning
by: Shindo, Hikaru, et al.
Published: (2023)
by: Shindo, Hikaru, et al.
Published: (2023)
Core Tokensets for Data-efficient Sequential Training of Transformers
by: Paul, Subarnaduti, et al.
Published: (2024)
by: Paul, Subarnaduti, et al.
Published: (2024)
Efficient Masked Attention Transformer for Few-Shot Classification and Segmentation
by: Carrión-Ojeda, Dustin, et al.
Published: (2025)
by: Carrión-Ojeda, Dustin, et al.
Published: (2025)
V-LoL: A Diagnostic Dataset for Visual Logical Learning
by: Helff, Lukas, et al.
Published: (2023)
by: Helff, Lukas, et al.
Published: (2023)
ART: Adaptive Relational Transformer for Pedestrian Trajectory Prediction with Temporal-Aware Relations
by: Li, Ruochen, et al.
Published: (2026)
by: Li, Ruochen, et al.
Published: (2026)
LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models
by: Helff, Lukas, et al.
Published: (2024)
by: Helff, Lukas, et al.
Published: (2024)
Exploiting Cultural Biases via Homoglyphs in Text-to-Image Synthesis
by: Struppek, Lukas, et al.
Published: (2022)
by: Struppek, Lukas, et al.
Published: (2022)
How to Train your Text-to-Image Model: Evaluating Design Choices for Synthetic Training Captions
by: Brack, Manuel, et al.
Published: (2025)
by: Brack, Manuel, et al.
Published: (2025)
LEDITS++: Limitless Image Editing using Text-to-Image Models
by: Brack, Manuel, et al.
Published: (2023)
by: Brack, Manuel, et al.
Published: (2023)
Boosting Omnidirectional Stereo Matching with a Pre-trained Depth Foundation Model
by: Endres, Jannik, et al.
Published: (2025)
by: Endres, Jannik, et al.
Published: (2025)
No Safe Dose: How Training Data Drives Unsafe Image Generation
by: Friedrich, Felix, et al.
Published: (2026)
by: Friedrich, Felix, et al.
Published: (2026)
STORM: Segment, Track, and Object Re-Localization from a Single Image
by: Deng, Yu, et al.
Published: (2025)
by: Deng, Yu, et al.
Published: (2025)
SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems
by: Shindo, Hikaru, et al.
Published: (2026)
by: Shindo, Hikaru, et al.
Published: (2026)
GLASS: Guided Latent Slot Diffusion for Object-Centric Learning
by: Singh, Krishnakant, et al.
Published: (2024)
by: Singh, Krishnakant, et al.
Published: (2024)
Evaluating Object-Centric Models beyond Object Discovery
by: Singh, Krishnakant, et al.
Published: (2026)
by: Singh, Krishnakant, et al.
Published: (2026)
Benchmarking the Attribution Quality of Vision Models
by: Hesse, Robin, et al.
Published: (2024)
by: Hesse, Robin, et al.
Published: (2024)
Markers Identification for Relative Pose Estimation of an Uncooperative Target
by: Candan, Batu, et al.
Published: (2024)
by: Candan, Batu, et al.
Published: (2024)
Boosting Unsupervised Semantic Segmentation with Principal Mask Proposals
by: Hahn, Oliver, et al.
Published: (2024)
by: Hahn, Oliver, et al.
Published: (2024)
Removing Cost Volumes from Optical Flow Estimators
by: Kiefhaber, Simon, et al.
Published: (2025)
by: Kiefhaber, Simon, et al.
Published: (2025)
Adaptive Layer Selection for Efficient Vision Transformer Fine-Tuning
by: Devoto, Alessio, et al.
Published: (2024)
by: Devoto, Alessio, et al.
Published: (2024)
ViRED: Prediction of Visual Relations in Engineering Drawings
by: Gu, Chao, et al.
Published: (2024)
by: Gu, Chao, et al.
Published: (2024)
Few-shot Open Relation Extraction with Gaussian Prototype and Adaptive Margin
by: Guo, Tianlin, et al.
Published: (2024)
by: Guo, Tianlin, et al.
Published: (2024)
Relation Learning and Aggregate-attention for Multi-person Motion Prediction
by: Qu, Kehua, et al.
Published: (2024)
by: Qu, Kehua, et al.
Published: (2024)
OCAtari: Object-Centric Atari 2600 Reinforcement Learning Environments
by: Delfosse, Quentin, et al.
Published: (2023)
by: Delfosse, Quentin, et al.
Published: (2023)
Semantic Relation-Enhanced CLIP Adapter for Domain Adaptive Zero-Shot Learning
by: Yu, Jiaao, et al.
Published: (2025)
by: Yu, Jiaao, et al.
Published: (2025)
Predictive Reasoning with Augmented Anomaly Contrastive Learning for Compositional Visual Relations
by: Li, Chengtai, et al.
Published: (2026)
by: Li, Chengtai, et al.
Published: (2026)
Image-Text Relation Prediction for Multilingual Tweets
by: Rikters, Matīss, et al.
Published: (2025)
by: Rikters, Matīss, et al.
Published: (2025)
Disentangling Polysemantic Channels in Convolutional Neural Networks
by: Hesse, Robin, et al.
Published: (2025)
by: Hesse, Robin, et al.
Published: (2025)
Pix2Code: Learning to Compose Neural Visual Concepts as Programs
by: Wüst, Antonia, et al.
Published: (2024)
by: Wüst, Antonia, et al.
Published: (2024)
Is Synthetic Data all We Need? Benchmarking the Robustness of Models Trained with Synthetic Images
by: Singh, Krishnakant, et al.
Published: (2024)
by: Singh, Krishnakant, et al.
Published: (2024)
MUFASA: A Multi-Layer Framework for Slot Attention
by: Bock, Sebastian, et al.
Published: (2026)
by: Bock, Sebastian, et al.
Published: (2026)
Pair then Relation: Pair-Net for Panoptic Scene Graph Generation
by: Wang, Jinghao, et al.
Published: (2023)
by: Wang, Jinghao, et al.
Published: (2023)
LumosX: Relate Any Identities with Their Attributes for Personalized Video Generation
by: Xing, Jiazheng, et al.
Published: (2026)
by: Xing, Jiazheng, et al.
Published: (2026)
Relation-aware Hierarchical Prompt for Open-vocabulary Scene Graph Generation
by: Liu, Tao, et al.
Published: (2024)
by: Liu, Tao, et al.
Published: (2024)
Finding DoRI: Discovery of Retained Images in Diffusion Models
by: Kowalczuk, Antoni, et al.
Published: (2025)
by: Kowalczuk, Antoni, et al.
Published: (2025)
RS-Net: Context-Aware Relation Scoring for Dynamic Scene Graph Generation
by: Jo, Hae-Won, et al.
Published: (2025)
by: Jo, Hae-Won, et al.
Published: (2025)
Cameras as Relative Positional Encoding
by: Li, Ruilong, et al.
Published: (2025)
by: Li, Ruilong, et al.
Published: (2025)
EmoNet-Face: An Expert-Annotated Benchmark for Synthetic Emotion Recognition
by: Schuhmann, Christoph, et al.
Published: (2025)
by: Schuhmann, Christoph, et al.
Published: (2025)
Similar Items
-
DeiSAM: Segment Anything with Deictic Prompting
by: Shindo, Hikaru, et al.
Published: (2024) -
DIAGen: Semantically Diverse Image Augmentation with Generative Models for Few-Shot Learning
by: Lingenberg, Tobias, et al.
Published: (2024) -
Learning Differentiable Logic Programs for Abstract Visual Reasoning
by: Shindo, Hikaru, et al.
Published: (2023) -
Core Tokensets for Data-efficient Sequential Training of Transformers
by: Paul, Subarnaduti, et al.
Published: (2024) -
Efficient Masked Attention Transformer for Few-Shot Classification and Segmentation
by: Carrión-Ojeda, Dustin, et al.
Published: (2025)