Boosting Generative Image Modeling via Joint Image-Feature Synthesis
Fuente:
arXiv
Guardado en:
| Autores principales: | Kouzelis, Theodoros, Karypidis, Efstathios, Kakogeorgiou, Ioannis, Gidaris, Spyros, Komodakis, Nikos |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Coevolving Representations in Joint Image-Feature Diffusion
por: Kouzelis, Theodoros, et al.
Publicado: (2026)
por: Kouzelis, Theodoros, et al.
Publicado: (2026)
Advancing Semantic Future Prediction through Multimodal Visual Sequence Transformers
por: Karypidis, Efstathios, et al.
Publicado: (2025)
por: Karypidis, Efstathios, et al.
Publicado: (2025)
DINO-Foresight: Looking into the Future with DINO
por: Karypidis, Efstathios, et al.
Publicado: (2024)
por: Karypidis, Efstathios, et al.
Publicado: (2024)
Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction
por: Karypidis, Efstathios, et al.
Publicado: (2026)
por: Karypidis, Efstathios, et al.
Publicado: (2026)
EQ-VAE: Equivariance Regularized Latent Space for Improved Generative Image Modeling
por: Kouzelis, Theodoros, et al.
Publicado: (2025)
por: Kouzelis, Theodoros, et al.
Publicado: (2025)
SPOT: Self-Training with Patch-Order Permutation for Object-Centric Learning with Autoregressive Transformers
por: Kakogeorgiou, Ioannis, et al.
Publicado: (2023)
por: Kakogeorgiou, Ioannis, et al.
Publicado: (2023)
Multi-Token Prediction Needs Registers
por: Gerontopoulos, Anastasios, et al.
Publicado: (2025)
por: Gerontopoulos, Anastasios, et al.
Publicado: (2025)
Technical Report for the 5th CLVision Challenge at CVPR: Addressing the Class-Incremental with Repetition using Unlabeled Data -- 4th Place Solution
por: Moraiti, Panagiota, et al.
Publicado: (2025)
por: Moraiti, Panagiota, et al.
Publicado: (2025)
MOCA: Self-supervised Representation Learning by Predicting Masked Online Codebook Assignments
por: Gidaris, Spyros, et al.
Publicado: (2023)
por: Gidaris, Spyros, et al.
Publicado: (2023)
Composed Image Retrieval for Remote Sensing
por: Psomas, Bill, et al.
Publicado: (2024)
por: Psomas, Bill, et al.
Publicado: (2024)
Enabling Local Editing in Diffusion Models by Joint and Individual Component Analysis
por: Kouzelis, Theodoros, et al.
Publicado: (2024)
por: Kouzelis, Theodoros, et al.
Publicado: (2024)
Attention, Please! Revisiting Attentive Probing Through the Lens of Efficiency
por: Psomas, Bill, et al.
Publicado: (2025)
por: Psomas, Bill, et al.
Publicado: (2025)
Boosting Visual Instruction Tuning with Self-Supervised Guidance
por: Sirko-Galouchenko, Sophia, et al.
Publicado: (2026)
por: Sirko-Galouchenko, Sophia, et al.
Publicado: (2026)
Benchmarking Composed Image Retrieval for Applied Earth Observation
por: Psomas, Bill, et al.
Publicado: (2026)
por: Psomas, Bill, et al.
Publicado: (2026)
REGLUE Your Latents with Global and Local Semantics for Entangled Diffusion
por: Petsangourakis, Giorgos, et al.
Publicado: (2025)
por: Petsangourakis, Giorgos, et al.
Publicado: (2025)
POP-3D: Open-Vocabulary 3D Occupancy Prediction from Images
por: Vobecky, Antonin, et al.
Publicado: (2024)
por: Vobecky, Antonin, et al.
Publicado: (2024)
ToNNO: Tomographic Reconstruction of a Neural Network's Output for Weakly Supervised Segmentation of 3D Medical Images
por: Schmidt-Mengin, Marius, et al.
Publicado: (2024)
por: Schmidt-Mengin, Marius, et al.
Publicado: (2024)
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
por: Siméoni, Oriane, et al.
Publicado: (2023)
por: Siméoni, Oriane, et al.
Publicado: (2023)
Retrieve and Segment: Are a Few Examples Enough to Bridge the Supervision Gap in Open-Vocabulary Segmentation?
por: Aravanis, Tilemachos, et al.
Publicado: (2026)
por: Aravanis, Tilemachos, et al.
Publicado: (2026)
DIP: Unsupervised Dense In-Context Post-training of Visual Representations
por: Sirko-Galouchenko, Sophia, et al.
Publicado: (2025)
por: Sirko-Galouchenko, Sophia, et al.
Publicado: (2025)
TryOffAnyone: Tiled Cloth Generation from a Dressed Person
por: Xarchakos, Ioannis, et al.
Publicado: (2024)
por: Xarchakos, Ioannis, et al.
Publicado: (2024)
Drive&Segment: Unsupervised Semantic Segmentation of Urban Scenes via Cross-modal Distillation
por: Vobecky, Antonin, et al.
Publicado: (2022)
por: Vobecky, Antonin, et al.
Publicado: (2022)
No Train, all Gain: Self-Supervised Gradients Improve Deep Frozen Representations
por: Simoncini, Walter, et al.
Publicado: (2024)
por: Simoncini, Walter, et al.
Publicado: (2024)
OccFeat: Self-supervised Occupancy Feature Prediction for Pretraining BEV Segmentation Networks
por: Sirko-Galouchenko, Sophia, et al.
Publicado: (2024)
por: Sirko-Galouchenko, Sophia, et al.
Publicado: (2024)
Vanilla ViT for Automotive Point Cloud Semantic Segmentation
por: Puy, Gilles, et al.
Publicado: (2026)
por: Puy, Gilles, et al.
Publicado: (2026)
Three Pillars improving Vision Foundation Model Distillation for Lidar
por: Puy, Gilles, et al.
Publicado: (2023)
por: Puy, Gilles, et al.
Publicado: (2023)
Joint Generative Modeling of Grounded Scene Graphs and Images via Diffusion Models
por: Xu, Bicheng, et al.
Publicado: (2024)
por: Xu, Bicheng, et al.
Publicado: (2024)
Interpreting and Analysing CLIP's Zero-Shot Image Classification via Mutual Knowledge
por: Sammani, Fawaz, et al.
Publicado: (2024)
por: Sammani, Fawaz, et al.
Publicado: (2024)
Comparison Analysis of Traditional Machine Learning and Deep Learning Techniques for Data and Image Classification
por: Karypidis, Efstathios, et al.
Publicado: (2022)
por: Karypidis, Efstathios, et al.
Publicado: (2022)
NVComposer: Boosting Generative Novel View Synthesis with Multiple Sparse and Unposed Images
por: Li, Lingen, et al.
Publicado: (2024)
por: Li, Lingen, et al.
Publicado: (2024)
HyPER-GAN: Hybrid Patch-Based Image-to-Image Translation for Real-Time Photorealism Enhancement
por: Pasios, Stefanos, et al.
Publicado: (2026)
por: Pasios, Stefanos, et al.
Publicado: (2026)
Boosting Few-Shot Detection with Large Language Models and Layout-to-Image Synthesis
por: Abdullah, Ahmed, et al.
Publicado: (2024)
por: Abdullah, Ahmed, et al.
Publicado: (2024)
Interpretable Vision Transformers in Image Classification via SVDA
por: Arampatzakis, Vasileios, et al.
Publicado: (2026)
por: Arampatzakis, Vasileios, et al.
Publicado: (2026)
Instant 3D Human Avatar Generation using Image Diffusion Models
por: Kolotouros, Nikos, et al.
Publicado: (2024)
por: Kolotouros, Nikos, et al.
Publicado: (2024)
TextBoost: Boosting Text Encoder for Personalized Text-to-Image Generation
por: Park, NaHyeon, et al.
Publicado: (2024)
por: Park, NaHyeon, et al.
Publicado: (2024)
Environment-Aware Satellite Image Generation with Diffusion Models
por: Kostagiolas, Nikos, et al.
Publicado: (2025)
por: Kostagiolas, Nikos, et al.
Publicado: (2025)
Generative Adversarial Synthesis and Deep Feature Discrimination of Brain Tumor MRI Images
por: Ali, Md Sumon, et al.
Publicado: (2025)
por: Ali, Md Sumon, et al.
Publicado: (2025)
Guided Image Synthesis via Initial Image Editing in Diffusion Model
por: Mao, Jiafeng, et al.
Publicado: (2023)
por: Mao, Jiafeng, et al.
Publicado: (2023)
Masked Image Modelling for retinal OCT understanding
por: Pissas, Theodoros, et al.
Publicado: (2024)
por: Pissas, Theodoros, et al.
Publicado: (2024)
Boosting Image Restoration via Priors from Pre-trained Models
por: Xu, Xiaogang, et al.
Publicado: (2024)
por: Xu, Xiaogang, et al.
Publicado: (2024)
Ejemplares similares
-
Coevolving Representations in Joint Image-Feature Diffusion
por: Kouzelis, Theodoros, et al.
Publicado: (2026) -
Advancing Semantic Future Prediction through Multimodal Visual Sequence Transformers
por: Karypidis, Efstathios, et al.
Publicado: (2025) -
DINO-Foresight: Looking into the Future with DINO
por: Karypidis, Efstathios, et al.
Publicado: (2024) -
Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction
por: Karypidis, Efstathios, et al.
Publicado: (2026) -
EQ-VAE: Equivariance Regularized Latent Space for Improved Generative Image Modeling
por: Kouzelis, Theodoros, et al.
Publicado: (2025)