INTRA: Interaction Relationship-aware Weakly Supervised Affordance Grounding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jang, Ji Ha, Seo, Hoigi, Chun, Se Young |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models
von: Jeong, Wongi, et al.
Veröffentlicht: (2026)
von: Jeong, Wongi, et al.
Veröffentlicht: (2026)
Efficient Personalization of Quantized Diffusion Model without Backpropagation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Training-free Mixed-Resolution Latent Upsampling for Spatially Accelerated Diffusion Transformers
von: Jeong, Wongi, et al.
Veröffentlicht: (2025)
von: Jeong, Wongi, et al.
Veröffentlicht: (2025)
Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Triadic Dynamics Aware Diffusion Posterior Sampling for Inverse Problems: Optimizing Guidance and Stochasticity Schedules
von: Bang, Junseo, et al.
Veröffentlicht: (2026)
von: Bang, Junseo, et al.
Veröffentlicht: (2026)
BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion
von: Kim, Gwanghyun, et al.
Veröffentlicht: (2024)
von: Kim, Gwanghyun, et al.
Veröffentlicht: (2024)
Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models
von: Seo, Hoigi, et al.
Veröffentlicht: (2026)
von: Seo, Hoigi, et al.
Veröffentlicht: (2026)
Robust 3D-Masked Part-level Editing in 3D Gaussian Splatting with Regularized Score Distillation Sampling
von: Kim, Hayeon, et al.
Veröffentlicht: (2025)
von: Kim, Hayeon, et al.
Veröffentlicht: (2025)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Weakly-Supervised Affordance Grounding Guided by Part-Level Semantic Priors
von: Xu, Peiran, et al.
Veröffentlicht: (2025)
von: Xu, Peiran, et al.
Veröffentlicht: (2025)
Selective Contrastive Learning for Weakly Supervised Affordance Grounding
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
Geometrical Properties of Text Token Embeddings for Strong Semantic Binding in Text-to-Image Generation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Uncertainty-guided Compositional Alignment with Part-to-Whole Semantic Representativeness in Hyperbolic Vision-Language Models
von: Kim, Hayeon, et al.
Veröffentlicht: (2026)
von: Kim, Hayeon, et al.
Veröffentlicht: (2026)
Closed-Loop Transfer for Weakly-supervised Affordance Grounding
von: Tang, Jiajin, et al.
Veröffentlicht: (2025)
von: Tang, Jiajin, et al.
Veröffentlicht: (2025)
AffordanceSAM: Segment Anything Once More in Affordance Grounding
von: Jiang, Dengyang, et al.
Veröffentlicht: (2025)
von: Jiang, Dengyang, et al.
Veröffentlicht: (2025)
FCL-COD: Weakly Supervised Camouflaged Object Detection with Frequency-aware and Contrastive Learning
von: Ni, Jingchen, et al.
Veröffentlicht: (2026)
von: Ni, Jingchen, et al.
Veröffentlicht: (2026)
VAGNet: Grounding 3D Affordance from Human-Object Interactions in Videos
von: Mao, Aihua, et al.
Veröffentlicht: (2026)
von: Mao, Aihua, et al.
Veröffentlicht: (2026)
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
von: Gao, Xianqiang, et al.
Veröffentlicht: (2024)
von: Gao, Xianqiang, et al.
Veröffentlicht: (2024)
AffordanceLLM: Grounding Affordance from Vision Language Models
von: Qian, Shengyi, et al.
Veröffentlicht: (2024)
von: Qian, Shengyi, et al.
Veröffentlicht: (2024)
RegFormer: Transferable Relational Grounding for Efficient Weakly-Supervised Human-Object Interaction Detection
von: Park, Jihwan, et al.
Veröffentlicht: (2026)
von: Park, Jihwan, et al.
Veröffentlicht: (2026)
Short-term Object Interaction Anticipation with Disentangled Object Detection @ Ego4D Short Term Object Interaction Anticipation Challenge
von: Cho, Hyunjin, et al.
Veröffentlicht: (2024)
von: Cho, Hyunjin, et al.
Veröffentlicht: (2024)
Grounding 3D Scene Affordance From Egocentric Interactions
von: Liu, Cuiyu, et al.
Veröffentlicht: (2024)
von: Liu, Cuiyu, et al.
Veröffentlicht: (2024)
Aerial View River Landform Video segmentation: A Weakly Supervised Context-aware Temporal Consistency Distillation Approach
von: Chen, Chi-Han, et al.
Veröffentlicht: (2025)
von: Chen, Chi-Han, et al.
Veröffentlicht: (2025)
Beyond the Ground Truth: Enhanced Supervision for Image Restoration
von: Ryou, Donghun, et al.
Veröffentlicht: (2025)
von: Ryou, Donghun, et al.
Veröffentlicht: (2025)
General Geometry-aware Weakly Supervised 3D Object Detection
von: Zhang, Guowen, et al.
Veröffentlicht: (2024)
von: Zhang, Guowen, et al.
Veröffentlicht: (2024)
UL-VIO: Ultra-lightweight Visual-Inertial Odometry with Noise Robust Test-time Adaptation
von: Park, Jinho, et al.
Veröffentlicht: (2024)
von: Park, Jinho, et al.
Veröffentlicht: (2024)
AdaRadar: Rate Adaptive Spectral Compression for Radar-based Perception
von: Park, Jinho, et al.
Veröffentlicht: (2026)
von: Park, Jinho, et al.
Veröffentlicht: (2026)
Affostruction: 3D Affordance Grounding with Generative Reconstruction
von: Park, Chunghyun, et al.
Veröffentlicht: (2026)
von: Park, Chunghyun, et al.
Veröffentlicht: (2026)
Weakly Supervised Temporal Sentence Grounding via Positive Sample Mining
von: Dong, Lu, et al.
Veröffentlicht: (2025)
von: Dong, Lu, et al.
Veröffentlicht: (2025)
DAG: Unleash the Potential of Diffusion Model for Open-Vocabulary 3D Affordance Grounding
von: Wang, Hanqing, et al.
Veröffentlicht: (2025)
von: Wang, Hanqing, et al.
Veröffentlicht: (2025)
Gaussian Mixture Proposals with Pull-Push Learning Scheme to Capture Diverse Events for Weakly Supervised Temporal Video Grounding
von: Kim, Sunoh, et al.
Veröffentlicht: (2023)
von: Kim, Sunoh, et al.
Veröffentlicht: (2023)
Text2Place: Affordance-aware Text Guided Human Placement
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
DIAL: Dense Image-text ALignment for Weakly Supervised Semantic Segmentation
von: Jang, Soojin, et al.
Veröffentlicht: (2024)
von: Jang, Soojin, et al.
Veröffentlicht: (2024)
Siamese Learning with Joint Alignment and Regression for Weakly-Supervised Video Paragraph Grounding
von: Tan, Chaolei, et al.
Veröffentlicht: (2024)
von: Tan, Chaolei, et al.
Veröffentlicht: (2024)
STPro: Spatial and Temporal Progressive Learning for Weakly Supervised Spatio-Temporal Grounding
von: Garg, Aaryan, et al.
Veröffentlicht: (2025)
von: Garg, Aaryan, et al.
Veröffentlicht: (2025)
Contextual Self-paced Learning for Weakly Supervised Spatio-Temporal Video Grounding
von: Kumar, Akash, et al.
Veröffentlicht: (2025)
von: Kumar, Akash, et al.
Veröffentlicht: (2025)
CRAFT: A Neuro-Symbolic Framework for Visual Functional Affordance Grounding
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
Shared Coupling-bridge for Weakly Supervised Local Feature Learning
von: Sun, Jiayuan, et al.
Veröffentlicht: (2022)
von: Sun, Jiayuan, et al.
Veröffentlicht: (2022)
Grounding 3D Object Affordance with Language Instructions, Visual Observations and Interactions
von: Zhu, He, et al.
Veröffentlicht: (2025)
von: Zhu, He, et al.
Veröffentlicht: (2025)
AlignCAT: Visual-Linguistic Alignment of Category and Attribute for Weakly Supervised Visual Grounding
von: Wang, Yidan, et al.
Veröffentlicht: (2025)
von: Wang, Yidan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models
von: Jeong, Wongi, et al.
Veröffentlicht: (2026) -
Efficient Personalization of Quantized Diffusion Model without Backpropagation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025) -
Training-free Mixed-Resolution Latent Upsampling for Spatially Accelerated Diffusion Transformers
von: Jeong, Wongi, et al.
Veröffentlicht: (2025) -
Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025) -
Triadic Dynamics Aware Diffusion Posterior Sampling for Inverse Problems: Optimizing Guidance and Stochasticity Schedules
von: Bang, Junseo, et al.
Veröffentlicht: (2026)