From Vicious to Virtuous Cycles: Synergistic Representation Learning for Unsupervised Video Object-Centric Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Seong, Hyun Seok, Moon, WonJun, Heo, Jae-Pil |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric Learning
von: Moon, WonJun, et al.
Veröffentlicht: (2026)
von: Moon, WonJun, et al.
Veröffentlicht: (2026)
Selective Contrastive Learning for Weakly Supervised Affordance Grounding
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
Progressive Proxy Anchor Propagation for Unsupervised Semantic Segmentation
von: Seong, Hyun Seok, et al.
Veröffentlicht: (2024)
von: Seong, Hyun Seok, et al.
Veröffentlicht: (2024)
Temporal Alignment-Free Video Matching for Few-shot Action Recognition
von: Lee, SuBeen, et al.
Veröffentlicht: (2025)
von: Lee, SuBeen, et al.
Veröffentlicht: (2025)
VLCounter: Text-aware Visual Representation for Zero-Shot Object Counting
von: Kang, Seunggu, et al.
Veröffentlicht: (2023)
von: Kang, Seunggu, et al.
Veröffentlicht: (2023)
Correlation-Guided Query-Dependency Calibration for Video Temporal Grounding
von: Moon, WonJun, et al.
Veröffentlicht: (2023)
von: Moon, WonJun, et al.
Veröffentlicht: (2023)
Auxiliary Descriptive Knowledge for Few-Shot Adaptation of Vision-Language Model
von: Lee, SuBeen, et al.
Veröffentlicht: (2025)
von: Lee, SuBeen, et al.
Veröffentlicht: (2025)
Looking Beyond the Window: Global-Local Aligned CLIP for Training-free Open-Vocabulary Semantic Segmentation
von: Lee, ByeongCheol, et al.
Veröffentlicht: (2026)
von: Lee, ByeongCheol, et al.
Veröffentlicht: (2026)
Masking Matters: Unlocking the Spatial Reasoning Capabilities of LLMs for 3D Scene-Language Understanding
von: Jeon, Yerim, et al.
Veröffentlicht: (2025)
von: Jeon, Yerim, et al.
Veröffentlicht: (2025)
Mitigating Semantic Collapse in Partially Relevant Video Retrieval
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
Mitigating Background Shift in Class-Incremental Semantic Segmentation
von: Park, Gilhan, et al.
Veröffentlicht: (2024)
von: Park, Gilhan, et al.
Veröffentlicht: (2024)
GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats
von: Hyun, Sangeek, et al.
Veröffentlicht: (2024)
von: Hyun, Sangeek, et al.
Veröffentlicht: (2024)
Prototypes are Balanced Units for Efficient and Effective Partially Relevant Video Retrieval
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
EAGLE: Eigen Aggregation Learning for Object-Centric Unsupervised Semantic Segmentation
von: Kim, Chanyoung, et al.
Veröffentlicht: (2024)
von: Kim, Chanyoung, et al.
Veröffentlicht: (2024)
Foreground-Covering Prototype Generation and Matching for SAM-Aided Few-Shot Segmentation
von: Park, Suho, et al.
Veröffentlicht: (2025)
von: Park, Suho, et al.
Veröffentlicht: (2025)
Scalable GANs with Transformers
von: Hyun, Sangeek, et al.
Veröffentlicht: (2025)
von: Hyun, Sangeek, et al.
Veröffentlicht: (2025)
GTA: Guided Transfer of Spatial Attention from Object-Centric Representations
von: Seo, SeokHyun, et al.
Veröffentlicht: (2024)
von: Seo, SeokHyun, et al.
Veröffentlicht: (2024)
Cycle Consistency in Video Object-Centric Learning
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2026)
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2026)
Style Injection in Diffusion: A Training-free Approach for Adapting Large-scale Diffusion Models for Style Transfer
von: Chung, Jiwoo, et al.
Veröffentlicht: (2023)
von: Chung, Jiwoo, et al.
Veröffentlicht: (2023)
Object-Centric Representation Learning for Enhanced 3D Semantic Scene Graph Prediction
von: Heo, KunHo, et al.
Veröffentlicht: (2025)
von: Heo, KunHo, et al.
Veröffentlicht: (2025)
Mutually-Aware Feature Learning for Few-Shot Object Counting
von: Jeon, Yerim, et al.
Veröffentlicht: (2024)
von: Jeon, Yerim, et al.
Veröffentlicht: (2024)
Unsupervised Object-Centric Learning from Multiple Unspecified Viewpoints
von: Yuan, Jinyang, et al.
Veröffentlicht: (2024)
von: Yuan, Jinyang, et al.
Veröffentlicht: (2024)
Grouped Discrete Representation for Object-Centric Learning
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2024)
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2024)
Zero-Shot Object-Centric Representation Learning
von: Didolkar, Aniket, et al.
Veröffentlicht: (2024)
von: Didolkar, Aniket, et al.
Veröffentlicht: (2024)
Are We Done with Object-Centric Learning?
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2025)
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2025)
Hierarchical Compact Clustering Attention (COCA) for Unsupervised Object-Centric Learning
von: Küçüksözen, Can, et al.
Veröffentlicht: (2025)
von: Küçüksözen, Can, et al.
Veröffentlicht: (2025)
WorldComp2D: Spatio-semantic Representations of Object Identity and Location from Local Views
von: Jin, SeongMin, et al.
Veröffentlicht: (2026)
von: Jin, SeongMin, et al.
Veröffentlicht: (2026)
DynaVol: Unsupervised Learning for Dynamic Scenes through Object-Centric Voxelization
von: Zhao, Yanpeng, et al.
Veröffentlicht: (2023)
von: Zhao, Yanpeng, et al.
Veröffentlicht: (2023)
DDLP: Unsupervised Object-Centric Video Prediction with Deep Dynamic Latent Particles
von: Daniel, Tal, et al.
Veröffentlicht: (2023)
von: Daniel, Tal, et al.
Veröffentlicht: (2023)
Explicitly Disentangled Representations in Object-Centric Learning
von: Majellaro, Riccardo, et al.
Veröffentlicht: (2024)
von: Majellaro, Riccardo, et al.
Veröffentlicht: (2024)
Disambiguating 2D-3D Correspondences in Gaussian Splatting-based Feature Fields for Visual Localization
von: Lee, Miso, et al.
Veröffentlicht: (2026)
von: Lee, Miso, et al.
Veröffentlicht: (2026)
Ambiguity-Restrained Text-Video Representation Learning for Partially Relevant Video Retrieval
von: Cho, CH, et al.
Veröffentlicht: (2025)
von: Cho, CH, et al.
Veröffentlicht: (2025)
Reasoning-Enhanced Object-Centric Learning for Videos
von: Li, Jian, et al.
Veröffentlicht: (2024)
von: Li, Jian, et al.
Veröffentlicht: (2024)
Cross-scale Aligned Supervision for Training GANs
von: Hyun, Sangeek, et al.
Veröffentlicht: (2026)
von: Hyun, Sangeek, et al.
Veröffentlicht: (2026)
Rethinking Temporal Consistency in Video Object-Centric Learning: From Prediction to Correspondence
von: Li, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Li, Zhiyuan, et al.
Veröffentlicht: (2026)
Noise-free Optimization in Early Training Steps for Image Super-Resolution
von: Lee, MinKyu, et al.
Veröffentlicht: (2023)
von: Lee, MinKyu, et al.
Veröffentlicht: (2023)
Multiple Object Stitching for Unsupervised Representation Learning
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
Diversity-aware Channel Pruning for StyleGAN Compression
von: Chung, Jiwoo, et al.
Veröffentlicht: (2024)
von: Chung, Jiwoo, et al.
Veröffentlicht: (2024)
Auto-Encoded Supervision for Perceptual Image Super-Resolution
von: Lee, MinKyu, et al.
Veröffentlicht: (2024)
von: Lee, MinKyu, et al.
Veröffentlicht: (2024)
Unsupervised Learning of Disentangled Representations from Video
von: Denton, Remi, et al.
Veröffentlicht: (2017)
von: Denton, Remi, et al.
Veröffentlicht: (2017)
Ähnliche Einträge
-
Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric Learning
von: Moon, WonJun, et al.
Veröffentlicht: (2026) -
Selective Contrastive Learning for Weakly Supervised Affordance Grounding
von: Moon, WonJun, et al.
Veröffentlicht: (2025) -
Progressive Proxy Anchor Propagation for Unsupervised Semantic Segmentation
von: Seong, Hyun Seok, et al.
Veröffentlicht: (2024) -
Temporal Alignment-Free Video Matching for Few-shot Action Recognition
von: Lee, SuBeen, et al.
Veröffentlicht: (2025) -
VLCounter: Text-aware Visual Representation for Zero-Shot Object Counting
von: Kang, Seunggu, et al.
Veröffentlicht: (2023)