GenView: Enhancing View Quality with Pretrained Generative Model for Self-Supervised Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Xiaojie, Yang, Yibo, Li, Xiangtai, Wu, Jianlong, Yu, Yue, Ghanem, Bernard, Zhang, Min |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GenView++: Unifying Adaptive Generative Augmentation and Quality-Driven Supervision for Contrastive Representation Learning
por: Li, Xiaojie, et al.
Publicado: (2025)
por: Li, Xiaojie, et al.
Publicado: (2025)
Enhancing Online Continual Learning with Plug-and-Play State Space Model and Class-Conditional Mixture of Discretization
por: Liu, Sihao, et al.
Publicado: (2024)
por: Liu, Sihao, et al.
Publicado: (2024)
Mamba-FSCIL: Dynamic Adaptation with Selective State Space Model for Few-Shot Class-Incremental Learning
por: Li, Xiaojie, et al.
Publicado: (2024)
por: Li, Xiaojie, et al.
Publicado: (2024)
Continuous Knowledge-Preserving Decomposition with Adaptive Layer Selection for Few-Shot Class-Incremental Learning
por: Li, Xiaojie, et al.
Publicado: (2025)
por: Li, Xiaojie, et al.
Publicado: (2025)
On Pretraining Data Diversity for Self-Supervised Learning
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)
Vivid-ZOO: Multi-View Video Generation with Diffusion Model
por: Li, Bing, et al.
Publicado: (2024)
por: Li, Bing, et al.
Publicado: (2024)
LipGen: Viseme-Guided Lip Video Generation for Enhancing Visual Speech Recognition
por: Hao, Bowen, et al.
Publicado: (2025)
por: Hao, Bowen, et al.
Publicado: (2025)
ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models
por: Fang, Zixun, et al.
Publicado: (2025)
por: Fang, Zixun, et al.
Publicado: (2025)
Multi-View Crowd Counting With Self-Supervised Learning
por: Mo, Hong, et al.
Publicado: (2025)
por: Mo, Hong, et al.
Publicado: (2025)
MVTN: Learning Multi-View Transformations for 3D Understanding
por: Hamdi, Abdullah, et al.
Publicado: (2022)
por: Hamdi, Abdullah, et al.
Publicado: (2022)
PanopticPartFormer++: A Unified and Decoupled View for Panoptic Part Segmentation
por: Li, Xiangtai, et al.
Publicado: (2023)
por: Li, Xiangtai, et al.
Publicado: (2023)
RobuRCDet: Enhancing Robustness of Radar-Camera Fusion in Bird's Eye View for 3D Object Detection
por: Yue, Jingtong, et al.
Publicado: (2025)
por: Yue, Jingtong, et al.
Publicado: (2025)
Self-Supervised Discriminative Feature Learning for Deep Multi-View Clustering
por: Xu, Jie, et al.
Publicado: (2021)
por: Xu, Jie, et al.
Publicado: (2021)
A Multi-View Consistency Framework with Semi-Supervised Domain Adaptation
por: Hong, Yuting, et al.
Publicado: (2026)
por: Hong, Yuting, et al.
Publicado: (2026)
ConsID-Gen: View-Consistent and Identity-Preserving Image-to-Video Generation
por: Wu, Mingyang, et al.
Publicado: (2026)
por: Wu, Mingyang, et al.
Publicado: (2026)
UniRecGen: Unifying Multi-View 3D Reconstruction and Generation
por: Huang, Zhisheng, et al.
Publicado: (2026)
por: Huang, Zhisheng, et al.
Publicado: (2026)
WonderFree: Enhancing Novel View Quality and Cross-View Consistency for 3D Scene Exploration
por: Ni, Chaojun, et al.
Publicado: (2025)
por: Ni, Chaojun, et al.
Publicado: (2025)
ViewFusion: Learning Composable Diffusion Models for Novel View Synthesis
por: Spiegl, Bernard, et al.
Publicado: (2024)
por: Spiegl, Bernard, et al.
Publicado: (2024)
Uncertainty and Self-Supervision in Single-View Depth
por: Rodriguez-Puigvert, Javier
Publicado: (2024)
por: Rodriguez-Puigvert, Javier
Publicado: (2024)
Tuning-Free Visual Customization via View Iterative Self-Attention Control
por: Li, Xiaojie, et al.
Publicado: (2024)
por: Li, Xiaojie, et al.
Publicado: (2024)
Towards Open Vocabulary Learning: A Survey
por: Wu, Jianzong, et al.
Publicado: (2023)
por: Wu, Jianzong, et al.
Publicado: (2023)
USP: Unified Self-Supervised Pretraining for Image Generation and Understanding
por: Chu, Xiangxiang, et al.
Publicado: (2025)
por: Chu, Xiangxiang, et al.
Publicado: (2025)
Towards Cross-View-Consistent Self-Supervised Surround Depth Estimation
por: Ding, Laiyan, et al.
Publicado: (2024)
por: Ding, Laiyan, et al.
Publicado: (2024)
Semi-Supervised Learning for Visual Bird's Eye View Semantic Segmentation
por: Zhu, Junyu, et al.
Publicado: (2023)
por: Zhu, Junyu, et al.
Publicado: (2023)
MVEB: Self-Supervised Learning with Multi-View Entropy Bottleneck
por: Wen, Liangjian, et al.
Publicado: (2024)
por: Wen, Liangjian, et al.
Publicado: (2024)
RendBEV: Semantic Novel View Synthesis for Self-Supervised Bird's Eye View Segmentation
por: Monteagudo, Henrique Piñeiro, et al.
Publicado: (2025)
por: Monteagudo, Henrique Piñeiro, et al.
Publicado: (2025)
Enhancing Semi-Supervised Multi-View Graph Convolutional Networks via Supervised Contrastive Learning and Self-Training
por: Xiao, Huaiyuan, et al.
Publicado: (2025)
por: Xiao, Huaiyuan, et al.
Publicado: (2025)
Unposed Sparse Views Room Layout Reconstruction in the Age of Pretrain Model
por: Huang, Yaxuan, et al.
Publicado: (2025)
por: Huang, Yaxuan, et al.
Publicado: (2025)
ARVideo: Autoregressive Pretraining for Self-Supervised Video Representation Learning
por: Ren, Sucheng, et al.
Publicado: (2024)
por: Ren, Sucheng, et al.
Publicado: (2024)
VARS: Video Assistant Referee System for Automated Soccer Decision Making from Multiple Views
por: Held, Jan, et al.
Publicado: (2023)
por: Held, Jan, et al.
Publicado: (2023)
TESPEC: Temporally-Enhanced Self-Supervised Pretraining for Event Cameras
por: Mohammadi, Mohammad, et al.
Publicado: (2025)
por: Mohammadi, Mohammad, et al.
Publicado: (2025)
SUDO: Enhancing Text-to-Image Diffusion Models with Self-Supervised Direct Preference Optimization
por: Peng, Liang, et al.
Publicado: (2025)
por: Peng, Liang, et al.
Publicado: (2025)
Pix4Point: Image Pretrained Standard Transformers for 3D Point Cloud Understanding
por: Qian, Guocheng, et al.
Publicado: (2022)
por: Qian, Guocheng, et al.
Publicado: (2022)
CharacterGen: Efficient 3D Character Generation from Single Images with Multi-View Pose Canonicalization
por: Peng, Hao-Yang, et al.
Publicado: (2024)
por: Peng, Hao-Yang, et al.
Publicado: (2024)
Self-Supervised Partial Cycle-Consistency for Multi-View Matching
por: Taggenbrock, Fedor, et al.
Publicado: (2025)
por: Taggenbrock, Fedor, et al.
Publicado: (2025)
Semi-Supervised Multi-View Crowd Counting by Ranking Multi-View Fusion Models
por: Zhang, Qi, et al.
Publicado: (2025)
por: Zhang, Qi, et al.
Publicado: (2025)
DiffPlace: Street View Generation via Place-Controllable Diffusion Model Enhancing Place Recognition
por: Li, Ji, et al.
Publicado: (2026)
por: Li, Ji, et al.
Publicado: (2026)
D$^2$GS: Depth-and-Density Guided Gaussian Splatting for Stable and Accurate Sparse-View Reconstruction
por: Song, Meixi, et al.
Publicado: (2025)
por: Song, Meixi, et al.
Publicado: (2025)
LoLep: Single-View View Synthesis with Locally-Learned Planes and Self-Attention Occlusion Inference
por: Wang, Cong, et al.
Publicado: (2023)
por: Wang, Cong, et al.
Publicado: (2023)
RMP-SAM: Towards Real-Time Multi-Purpose Segment Anything
por: Xu, Shilin, et al.
Publicado: (2024)
por: Xu, Shilin, et al.
Publicado: (2024)
Ejemplares similares
-
GenView++: Unifying Adaptive Generative Augmentation and Quality-Driven Supervision for Contrastive Representation Learning
por: Li, Xiaojie, et al.
Publicado: (2025) -
Enhancing Online Continual Learning with Plug-and-Play State Space Model and Class-Conditional Mixture of Discretization
por: Liu, Sihao, et al.
Publicado: (2024) -
Mamba-FSCIL: Dynamic Adaptation with Selective State Space Model for Few-Shot Class-Incremental Learning
por: Li, Xiaojie, et al.
Publicado: (2024) -
Continuous Knowledge-Preserving Decomposition with Adaptive Layer Selection for Few-Shot Class-Incremental Learning
por: Li, Xiaojie, et al.
Publicado: (2025) -
On Pretraining Data Diversity for Self-Supervised Learning
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)