Segmenting Object Affordances: Reproducibility and Sensitivity to Scale
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Apicella, Tommaso, Xompero, Alessio, Gastaldo, Paolo, Cavallaro, Andrea |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Visual Affordance Prediction: Survey and Reproducibility
von: Apicella, Tommaso, et al.
Veröffentlicht: (2025)
von: Apicella, Tommaso, et al.
Veröffentlicht: (2025)
Affordance segmentation of hand-occluded containers from exocentric images
von: Apicella, Tommaso, et al.
Veröffentlicht: (2023)
von: Apicella, Tommaso, et al.
Veröffentlicht: (2023)
Learning Privacy from Visual Entities
von: Xompero, Alessio, et al.
Veröffentlicht: (2025)
von: Xompero, Alessio, et al.
Veröffentlicht: (2025)
Stereo Hand-Object Reconstruction for Human-to-Robot Handover
von: Pang, Yik Lung, et al.
Veröffentlicht: (2024)
von: Pang, Yik Lung, et al.
Veröffentlicht: (2024)
Zero-shot image privacy classification with Vision-Language Models
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
Explaining models relating objects and privacy
von: Xompero, Alessio, et al.
Veröffentlicht: (2024)
von: Xompero, Alessio, et al.
Veröffentlicht: (2024)
Container Localisation and Mass Estimation with an RGB-D Camera
von: Apicella, Tommaso, et al.
Veröffentlicht: (2022)
von: Apicella, Tommaso, et al.
Veröffentlicht: (2022)
Memory-Augmented Vision-Language Agents for Persistent and Semantically Consistent Object Captioning
von: Galliena, Tommaso, et al.
Veröffentlicht: (2026)
von: Galliena, Tommaso, et al.
Veröffentlicht: (2026)
Embodied Image Captioning: Self-supervised Learning Agents for Spatially Coherent Image Descriptions
von: Galliena, Tommaso, et al.
Veröffentlicht: (2025)
von: Galliena, Tommaso, et al.
Veröffentlicht: (2025)
AffordanceSAM: Segment Anything Once More in Affordance Grounding
von: Jiang, Dengyang, et al.
Veröffentlicht: (2025)
von: Jiang, Dengyang, et al.
Veröffentlicht: (2025)
The impact of abstract and object tags on image privacy classification
von: Baranouskaya, Darya, et al.
Veröffentlicht: (2025)
von: Baranouskaya, Darya, et al.
Veröffentlicht: (2025)
PrivLEX: Detecting legal concepts in images through Vision-Language Models
von: Baranouskaya, Darya, et al.
Veröffentlicht: (2026)
von: Baranouskaya, Darya, et al.
Veröffentlicht: (2026)
Leverage Task Context for Object Affordance Ranking
von: Huang, Haojie, et al.
Veröffentlicht: (2024)
von: Huang, Haojie, et al.
Veröffentlicht: (2024)
Lifelong Imitation Learning with Multimodal Latent Replay and Incremental Adjustment
von: Yu, Fanqi, et al.
Veröffentlicht: (2026)
von: Yu, Fanqi, et al.
Veröffentlicht: (2026)
ManipGPT: Is Affordance Segmentation by Large Vision Models Enough for Articulated Object Manipulation?
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
Towards Affordance-Aware Articulation Synthesis for Rigged Objects
von: Yu, Yu-Chu, et al.
Veröffentlicht: (2025)
von: Yu, Yu-Chu, et al.
Veröffentlicht: (2025)
Image-guided topic modeling for interpretable privacy classification
von: Baia, Alina Elena, et al.
Veröffentlicht: (2024)
von: Baia, Alina Elena, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Domain Adaptation for Vitiligo Segmentation in Clinical Photographs
von: Jiang, Wentao, et al.
Veröffentlicht: (2025)
von: Jiang, Wentao, et al.
Veröffentlicht: (2025)
AffordanceGrasp-R1:Leveraging Reasoning-Based Affordance Segmentation with Reinforcement Learning for Robotic Grasping
von: Zhou, Dingyi, et al.
Veröffentlicht: (2026)
von: Zhou, Dingyi, et al.
Veröffentlicht: (2026)
FlowOVD: Learning Generative Latent Flows for Zero-shot Open-vocabulary Detection
von: Wei, Yao, et al.
Veröffentlicht: (2026)
von: Wei, Yao, et al.
Veröffentlicht: (2026)
Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion
von: He, Jixuan, et al.
Veröffentlicht: (2024)
von: He, Jixuan, et al.
Veröffentlicht: (2024)
Integrating Affordances and Attention models for Short-Term Object Interaction Anticipation
von: Labadia, Lorenzo Mur, et al.
Veröffentlicht: (2026)
von: Labadia, Lorenzo Mur, et al.
Veröffentlicht: (2026)
IAAO: Interactive Affordance Learning for Articulated Objects in 3D Environments
von: Zhang, Can, et al.
Veröffentlicht: (2025)
von: Zhang, Can, et al.
Veröffentlicht: (2025)
Cross-modal Counterfactual Explanations: Uncovering Decision Factors and Dataset Biases in Subjective Classification
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
Unlocking 3D Affordance Segmentation with 2D Semantic Knowledge
von: Huang, Yu, et al.
Veröffentlicht: (2025)
von: Huang, Yu, et al.
Veröffentlicht: (2025)
Uncertainty Estimation in Instance Segmentation of Affordances via Bayesian Visual Transformers
von: Mur-Labadia, Lorenzo, et al.
Veröffentlicht: (2026)
von: Mur-Labadia, Lorenzo, et al.
Veröffentlicht: (2026)
AFF-ttention! Affordances and Attention models for Short-Term Object Interaction Anticipation
von: Mur-Labadia, Lorenzo, et al.
Veröffentlicht: (2024)
von: Mur-Labadia, Lorenzo, et al.
Veröffentlicht: (2024)
VAGNet: Grounding 3D Affordance from Human-Object Interactions in Videos
von: Mao, Aihua, et al.
Veröffentlicht: (2026)
von: Mao, Aihua, et al.
Veröffentlicht: (2026)
Sparse multi-view hand-object reconstruction for unseen environments
von: Pang, Yik Lung, et al.
Veröffentlicht: (2024)
von: Pang, Yik Lung, et al.
Veröffentlicht: (2024)
Black-box Attacks on Image Activity Prediction and its Natural Language Explanations
von: Baia, Alina Elena, et al.
Veröffentlicht: (2023)
von: Baia, Alina Elena, et al.
Veröffentlicht: (2023)
InteractAnything: Zero-shot Human Object Interaction Synthesis via LLM Feedback and Object Affordance Parsing
von: Zhang, Jinlu, et al.
Veröffentlicht: (2025)
von: Zhang, Jinlu, et al.
Veröffentlicht: (2025)
Potential Field as Scene Affordance for Behavior Change-Based Visual Risk Object Identification
von: Pao, Pang-Yuan, et al.
Veröffentlicht: (2024)
von: Pao, Pang-Yuan, et al.
Veröffentlicht: (2024)
Gaussian Heritage: 3D Digitization of Cultural Heritage with Integrated Object Segmentation
von: Dahaghin, Mahtab, et al.
Veröffentlicht: (2024)
von: Dahaghin, Mahtab, et al.
Veröffentlicht: (2024)
Object Affordance Recognition and Grounding via Multi-scale Cross-modal Representation Learning
von: Wan, Xinhang, et al.
Veröffentlicht: (2025)
von: Wan, Xinhang, et al.
Veröffentlicht: (2025)
TRACER: Texture-Robust Affordance Chain-of-Thought for Deformable-Object Refinement
von: Jia, Wanjun, et al.
Veröffentlicht: (2026)
von: Jia, Wanjun, et al.
Veröffentlicht: (2026)
Global Motion Understanding in Large-Scale Video Object Segmentation
von: Fedynyak, Volodymyr, et al.
Veröffentlicht: (2024)
von: Fedynyak, Volodymyr, et al.
Veröffentlicht: (2024)
Tsanet: Temporal and Scale Alignment for Unsupervised Video Object Segmentation
von: Lee, Seunghoon, et al.
Veröffentlicht: (2023)
von: Lee, Seunghoon, et al.
Veröffentlicht: (2023)
Affogato: Learning Open-Vocabulary Affordance Grounding with Automated Data Generation at Scale
von: Lee, Junha, et al.
Veröffentlicht: (2025)
von: Lee, Junha, et al.
Veröffentlicht: (2025)
Learning to Evaluate Autonomous Behaviour in Human-Robot Interaction
von: Tiezzi, Matteo, et al.
Veröffentlicht: (2025)
von: Tiezzi, Matteo, et al.
Veröffentlicht: (2025)
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
von: Gao, Xianqiang, et al.
Veröffentlicht: (2024)
von: Gao, Xianqiang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Visual Affordance Prediction: Survey and Reproducibility
von: Apicella, Tommaso, et al.
Veröffentlicht: (2025) -
Affordance segmentation of hand-occluded containers from exocentric images
von: Apicella, Tommaso, et al.
Veröffentlicht: (2023) -
Learning Privacy from Visual Entities
von: Xompero, Alessio, et al.
Veröffentlicht: (2025) -
Stereo Hand-Object Reconstruction for Human-to-Robot Handover
von: Pang, Yik Lung, et al.
Veröffentlicht: (2024) -
Zero-shot image privacy classification with Vision-Language Models
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)