Top-Down Guidance for Learning Object-Centric Representations
Fuente:
arXiv
Saved in:
| Main Authors: | Zou, Junhong, Zhu, Xiangyu, Zhang, Zhaoxiang, Lei, Zhen |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Large Vision-Language Models' Understanding for Flow Field Data
by: Zhang, Xiaomei, et al.
Published: (2025)
by: Zhang, Xiaomei, et al.
Published: (2025)
Revisiting Marr in Face: The Building of 2D--2.5D--3D Representations in Deep Neural Networks
by: Zhu, Xiangyu, et al.
Published: (2024)
by: Zhu, Xiangyu, et al.
Published: (2024)
3D Face Reconstruction with the Geometric Guidance of Facial Part Segmentation
by: Wang, Zidu, et al.
Published: (2023)
by: Wang, Zidu, et al.
Published: (2023)
Seek for Incantations: Towards Accurate Text-to-Image Diffusion Synthesis through Prompt Engineering
by: Yu, Chang, et al.
Published: (2024)
by: Yu, Chang, et al.
Published: (2024)
General Geometry-aware Weakly Supervised 3D Object Detection
by: Zhang, Guowen, et al.
Published: (2024)
by: Zhang, Guowen, et al.
Published: (2024)
On-Road Object Importance Estimation: A New Dataset and A Model with Multi-Fold Top-Down Guidance
by: Nan, Zhixiong, et al.
Published: (2024)
by: Nan, Zhixiong, et al.
Published: (2024)
Top2Pano: Learning to Generate Indoor Panoramas from Top-Down View
by: Zhang, Zitong, et al.
Published: (2025)
by: Zhang, Zitong, et al.
Published: (2025)
Organized Grouped Discrete Representation for Object-Centric Learning
by: Zhao, Rongzhen, et al.
Published: (2024)
by: Zhao, Rongzhen, et al.
Published: (2024)
TextOCVP: Object-Centric Video Prediction with Language Guidance
by: Villar-Corrales, Angel, et al.
Published: (2025)
by: Villar-Corrales, Angel, et al.
Published: (2025)
Adaptive Guidance Learning for Camouflaged Object Detection
by: Chen, Zhennan, et al.
Published: (2024)
by: Chen, Zhennan, et al.
Published: (2024)
Voxel Mamba: Group-Free State Space Models for Point Cloud based 3D Object Detection
by: Zhang, Guowen, et al.
Published: (2024)
by: Zhang, Guowen, et al.
Published: (2024)
Leveraging Bottom-Up and Top-Down Attention for Few-Shot Object Detection
by: Chen, Xianyu, et al.
Published: (2020)
by: Chen, Xianyu, et al.
Published: (2020)
DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer
by: Ma, Zhiyuan, et al.
Published: (2024)
by: Ma, Zhiyuan, et al.
Published: (2024)
Grouped Discrete Representation for Object-Centric Learning
by: Zhao, Rongzhen, et al.
Published: (2024)
by: Zhao, Rongzhen, et al.
Published: (2024)
Zero-Shot Object-Centric Representation Learning
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
Learning Global Object-Centric Representations via Disentangled Slot Attention
by: Chen, Tonglin, et al.
Published: (2024)
by: Chen, Tonglin, et al.
Published: (2024)
Context Matters: Learning Global Semantics via Object-Centric Representation
by: Zhong, Jike, et al.
Published: (2025)
by: Zhong, Jike, et al.
Published: (2025)
Learning Object-Centric Representations Based on Slots in Real World Scenarios
by: Akan, Adil Kaan
Published: (2025)
by: Akan, Adil Kaan
Published: (2025)
Learning Object-Centric Representations in SAR Images with Multi-Level Feature Fusion
by: Jang, Oh-Tae, et al.
Published: (2025)
by: Jang, Oh-Tae, et al.
Published: (2025)
Scene-Agnostic Object-Centric Representation Learning for 3D Gaussian Splatting
by: Hsu, Tsuheng, et al.
Published: (2026)
by: Hsu, Tsuheng, et al.
Published: (2026)
ORGAN: Object-Centric Representation Learning using Cycle Consistent Generative Adversarial Networks
by: Küchler, Joël, et al.
Published: (2026)
by: Küchler, Joël, et al.
Published: (2026)
Object-Centric Representation Learning for Enhanced 3D Semantic Scene Graph Prediction
by: Heo, KunHo, et al.
Published: (2025)
by: Heo, KunHo, et al.
Published: (2025)
Towards Realistic Hand-Object Interaction with Gravity-Field Based Diffusion Bridge
by: Xu, Miao, et al.
Published: (2025)
by: Xu, Miao, et al.
Published: (2025)
Modeling Spoof Noise by De-spoofing Diffusion and its Application in Face Anti-spoofing
by: Zhang, Bin, et al.
Published: (2024)
by: Zhang, Bin, et al.
Published: (2024)
DevFD: Developmental Face Forgery Detection by Learning Shared and Orthogonal LoRA Subspaces
by: Zhang, Tianshuo, et al.
Published: (2025)
by: Zhang, Tianshuo, et al.
Published: (2025)
Grouped Discrete Representation Guides Object-Centric Learning
by: Zhao, Rongzhen, et al.
Published: (2024)
by: Zhao, Rongzhen, et al.
Published: (2024)
Explicitly Disentangled Representations in Object-Centric Learning
by: Majellaro, Riccardo, et al.
Published: (2024)
by: Majellaro, Riccardo, et al.
Published: (2024)
Object-Centric Action-Enhanced Representations for Robot Visuo-Motor Policy Learning
by: Giannakakis, Nikos, et al.
Published: (2025)
by: Giannakakis, Nikos, et al.
Published: (2025)
Disentangled Object-Centric Image Representation for Robotic Manipulation
by: Emukpere, David, et al.
Published: (2025)
by: Emukpere, David, et al.
Published: (2025)
A Hyperbolic Perspective on Hierarchical Structure in Object-Centric Scene Representations
by: Madan, Neelu, et al.
Published: (2026)
by: Madan, Neelu, et al.
Published: (2026)
Are Object-Centric Representations Better At Compositional Generalization?
by: Kapl, Ferdinand, et al.
Published: (2026)
by: Kapl, Ferdinand, et al.
Published: (2026)
Object-Centric Relational Representations for Image Generation
by: Butera, Luca, et al.
Published: (2023)
by: Butera, Luca, et al.
Published: (2023)
Open Vocabulary 3D Scene Understanding via Geometry Guided Self-Distillation
by: Wang, Pengfei, et al.
Published: (2024)
by: Wang, Pengfei, et al.
Published: (2024)
SlotMemory: Object-Centric KV Memory for Streaming Long-Video Generation
by: Dou, Weijia, et al.
Published: (2026)
by: Dou, Weijia, et al.
Published: (2026)
Rethinking Layered Graphic Design Generation with a Top-Down Approach
by: Chen, Jingye, et al.
Published: (2025)
by: Chen, Jingye, et al.
Published: (2025)
Cycle Consistency in Video Object-Centric Learning
by: Zhao, Rongzhen, et al.
Published: (2026)
by: Zhao, Rongzhen, et al.
Published: (2026)
Neurosymbolic Object-Centric Learning with Distant Supervision
by: Colamonaco, Stefano, et al.
Published: (2025)
by: Colamonaco, Stefano, et al.
Published: (2025)
Representation Engineering: A Top-Down Approach to AI Transparency
by: Zou, Andy, et al.
Published: (2023)
by: Zou, Andy, et al.
Published: (2023)
Generative Active Learning for Image Synthesis Personalization
by: Zhang, Xulu, et al.
Published: (2024)
by: Zhang, Xulu, et al.
Published: (2024)
CarFormer: Self-Driving with Learned Object-Centric Representations
by: Hamdan, Shadi, et al.
Published: (2024)
by: Hamdan, Shadi, et al.
Published: (2024)
Similar Items
-
Improving Large Vision-Language Models' Understanding for Flow Field Data
by: Zhang, Xiaomei, et al.
Published: (2025) -
Revisiting Marr in Face: The Building of 2D--2.5D--3D Representations in Deep Neural Networks
by: Zhu, Xiangyu, et al.
Published: (2024) -
3D Face Reconstruction with the Geometric Guidance of Facial Part Segmentation
by: Wang, Zidu, et al.
Published: (2023) -
Seek for Incantations: Towards Accurate Text-to-Image Diffusion Synthesis through Prompt Engineering
by: Yu, Chang, et al.
Published: (2024) -
General Geometry-aware Weakly Supervised 3D Object Detection
by: Zhang, Guowen, et al.
Published: (2024)