Coevolving Representations in Joint Image-Feature Diffusion
Fuente:
arXiv
Salvato in:
| Autori principali: | Kouzelis, Theodoros, Gidaris, Spyros, Komodakis, Nikos |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Boosting Generative Image Modeling via Joint Image-Feature Synthesis
di: Kouzelis, Theodoros, et al.
Pubblicazione: (2025)
di: Kouzelis, Theodoros, et al.
Pubblicazione: (2025)
Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction
di: Karypidis, Efstathios, et al.
Pubblicazione: (2026)
di: Karypidis, Efstathios, et al.
Pubblicazione: (2026)
Multi-Token Prediction Needs Registers
di: Gerontopoulos, Anastasios, et al.
Pubblicazione: (2025)
di: Gerontopoulos, Anastasios, et al.
Pubblicazione: (2025)
Advancing Semantic Future Prediction through Multimodal Visual Sequence Transformers
di: Karypidis, Efstathios, et al.
Pubblicazione: (2025)
di: Karypidis, Efstathios, et al.
Pubblicazione: (2025)
SPOT: Self-Training with Patch-Order Permutation for Object-Centric Learning with Autoregressive Transformers
di: Kakogeorgiou, Ioannis, et al.
Pubblicazione: (2023)
di: Kakogeorgiou, Ioannis, et al.
Pubblicazione: (2023)
DINO-Foresight: Looking into the Future with DINO
di: Karypidis, Efstathios, et al.
Pubblicazione: (2024)
di: Karypidis, Efstathios, et al.
Pubblicazione: (2024)
EQ-VAE: Equivariance Regularized Latent Space for Improved Generative Image Modeling
di: Kouzelis, Theodoros, et al.
Pubblicazione: (2025)
di: Kouzelis, Theodoros, et al.
Pubblicazione: (2025)
MOCA: Self-supervised Representation Learning by Predicting Masked Online Codebook Assignments
di: Gidaris, Spyros, et al.
Pubblicazione: (2023)
di: Gidaris, Spyros, et al.
Pubblicazione: (2023)
Enabling Local Editing in Diffusion Models by Joint and Individual Component Analysis
di: Kouzelis, Theodoros, et al.
Pubblicazione: (2024)
di: Kouzelis, Theodoros, et al.
Pubblicazione: (2024)
DIP: Unsupervised Dense In-Context Post-training of Visual Representations
di: Sirko-Galouchenko, Sophia, et al.
Pubblicazione: (2025)
di: Sirko-Galouchenko, Sophia, et al.
Pubblicazione: (2025)
No Train, all Gain: Self-Supervised Gradients Improve Deep Frozen Representations
di: Simoncini, Walter, et al.
Pubblicazione: (2024)
di: Simoncini, Walter, et al.
Pubblicazione: (2024)
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
di: Siméoni, Oriane, et al.
Pubblicazione: (2023)
di: Siméoni, Oriane, et al.
Pubblicazione: (2023)
Retrieve and Segment: Are a Few Examples Enough to Bridge the Supervision Gap in Open-Vocabulary Segmentation?
di: Aravanis, Tilemachos, et al.
Pubblicazione: (2026)
di: Aravanis, Tilemachos, et al.
Pubblicazione: (2026)
POP-3D: Open-Vocabulary 3D Occupancy Prediction from Images
di: Vobecky, Antonin, et al.
Pubblicazione: (2024)
di: Vobecky, Antonin, et al.
Pubblicazione: (2024)
Franca: Nested Matryoshka Clustering for Scalable Visual Representation Learning
di: Venkataramanan, Shashanka, et al.
Pubblicazione: (2025)
di: Venkataramanan, Shashanka, et al.
Pubblicazione: (2025)
Boosting Visual Instruction Tuning with Self-Supervised Guidance
di: Sirko-Galouchenko, Sophia, et al.
Pubblicazione: (2026)
di: Sirko-Galouchenko, Sophia, et al.
Pubblicazione: (2026)
ToNNO: Tomographic Reconstruction of a Neural Network's Output for Weakly Supervised Segmentation of 3D Medical Images
di: Schmidt-Mengin, Marius, et al.
Pubblicazione: (2024)
di: Schmidt-Mengin, Marius, et al.
Pubblicazione: (2024)
OccFeat: Self-supervised Occupancy Feature Prediction for Pretraining BEV Segmentation Networks
di: Sirko-Galouchenko, Sophia, et al.
Pubblicazione: (2024)
di: Sirko-Galouchenko, Sophia, et al.
Pubblicazione: (2024)
Vanilla ViT for Automotive Point Cloud Semantic Segmentation
di: Puy, Gilles, et al.
Pubblicazione: (2026)
di: Puy, Gilles, et al.
Pubblicazione: (2026)
Drive&Segment: Unsupervised Semantic Segmentation of Urban Scenes via Cross-modal Distillation
di: Vobecky, Antonin, et al.
Pubblicazione: (2022)
di: Vobecky, Antonin, et al.
Pubblicazione: (2022)
Three Pillars improving Vision Foundation Model Distillation for Lidar
di: Puy, Gilles, et al.
Pubblicazione: (2023)
di: Puy, Gilles, et al.
Pubblicazione: (2023)
A High-Level Survey of Optical Remote Sensing
di: Koletsis, Panagiotis, et al.
Pubblicazione: (2026)
di: Koletsis, Panagiotis, et al.
Pubblicazione: (2026)
Attention, Please! Revisiting Attentive Probing Through the Lens of Efficiency
di: Psomas, Bill, et al.
Pubblicazione: (2025)
di: Psomas, Bill, et al.
Pubblicazione: (2025)
Hierarchical Fusion and Joint Aggregation: A Multi-Level Feature Representation Method for AIGC Image Quality Assessment
di: Meng, Linghe, et al.
Pubblicazione: (2025)
di: Meng, Linghe, et al.
Pubblicazione: (2025)
HyPER-GAN: Hybrid Patch-Based Image-to-Image Translation for Real-Time Photorealism Enhancement
di: Pasios, Stefanos, et al.
Pubblicazione: (2026)
di: Pasios, Stefanos, et al.
Pubblicazione: (2026)
Teacher-Feature Drifting: One-Step Diffusion Distillation with Pretrained Diffusion Representations
di: Zhang, Yuan, et al.
Pubblicazione: (2026)
di: Zhang, Yuan, et al.
Pubblicazione: (2026)
Instant 3D Human Avatar Generation using Image Diffusion Models
di: Kolotouros, Nikos, et al.
Pubblicazione: (2024)
di: Kolotouros, Nikos, et al.
Pubblicazione: (2024)
SafeWork-R1: Coevolving Safety and Intelligence under the AI-45$^{\circ}$ Law
di: Lab, Shanghai AI, et al.
Pubblicazione: (2025)
di: Lab, Shanghai AI, et al.
Pubblicazione: (2025)
AstroSpy: On detecting Fake Images in Astronomy via Joint Image-Spectral Representations
di: Alam, Mohammed Talha, et al.
Pubblicazione: (2024)
di: Alam, Mohammed Talha, et al.
Pubblicazione: (2024)
Joint Superpixel and Self-Representation Learning for Scalable Hyperspectral Image Clustering
di: Li, Xianlu, et al.
Pubblicazione: (2025)
di: Li, Xianlu, et al.
Pubblicazione: (2025)
PRIOR: Prototype Representation Joint Learning from Medical Images and Reports
di: Cheng, Pujin, et al.
Pubblicazione: (2023)
di: Cheng, Pujin, et al.
Pubblicazione: (2023)
Joint Conditional Diffusion Model for Image Restoration with Mixed Degradations
di: Yue, Yufeng, et al.
Pubblicazione: (2024)
di: Yue, Yufeng, et al.
Pubblicazione: (2024)
Interpreting and Analysing CLIP's Zero-Shot Image Classification via Mutual Knowledge
di: Sammani, Fawaz, et al.
Pubblicazione: (2024)
di: Sammani, Fawaz, et al.
Pubblicazione: (2024)
JoDiffusion: Jointly Diffusing Image with Pixel-Level Annotations for Semantic Segmentation Promotion
di: Wang, Haoyu, et al.
Pubblicazione: (2025)
di: Wang, Haoyu, et al.
Pubblicazione: (2025)
Environment-Aware Satellite Image Generation with Diffusion Models
di: Kostagiolas, Nikos, et al.
Pubblicazione: (2025)
di: Kostagiolas, Nikos, et al.
Pubblicazione: (2025)
Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders
di: Tong, Shengbang, et al.
Pubblicazione: (2026)
di: Tong, Shengbang, et al.
Pubblicazione: (2026)
Underwater Diffusion Attention Network with Contrastive Language-Image Joint Learning for Underwater Image Enhancement
di: Shaahid, Afrah, et al.
Pubblicazione: (2025)
di: Shaahid, Afrah, et al.
Pubblicazione: (2025)
Joint Generative Modeling of Grounded Scene Graphs and Images via Diffusion Models
di: Xu, Bicheng, et al.
Pubblicazione: (2024)
di: Xu, Bicheng, et al.
Pubblicazione: (2024)
Feature Denoising Diffusion Model for Blind Image Quality Assessment
di: Li, Xudong, et al.
Pubblicazione: (2024)
di: Li, Xudong, et al.
Pubblicazione: (2024)
Diffusion Noise Feature: Accurate and Fast Generated Image Detection
di: Zhang, Yichi, et al.
Pubblicazione: (2023)
di: Zhang, Yichi, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Boosting Generative Image Modeling via Joint Image-Feature Synthesis
di: Kouzelis, Theodoros, et al.
Pubblicazione: (2025) -
Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction
di: Karypidis, Efstathios, et al.
Pubblicazione: (2026) -
Multi-Token Prediction Needs Registers
di: Gerontopoulos, Anastasios, et al.
Pubblicazione: (2025) -
Advancing Semantic Future Prediction through Multimodal Visual Sequence Transformers
di: Karypidis, Efstathios, et al.
Pubblicazione: (2025) -
SPOT: Self-Training with Patch-Order Permutation for Object-Centric Learning with Autoregressive Transformers
di: Kakogeorgiou, Ioannis, et al.
Pubblicazione: (2023)