Diffusion Models as Data Mining Tools
Fuente:
arXiv
Salvato in:
| Autori principali: | Siglidis, Ioannis, Holynski, Aleksander, Efros, Alexei A., Aubry, Mathieu, Ginosar, Shiry |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale
di: Koepke, A. Sophia, et al.
Pubblicazione: (2026)
di: Koepke, A. Sophia, et al.
Pubblicazione: (2026)
Gaussian Masked Autoencoders
di: Rajasegaran, Jathushan, et al.
Pubblicazione: (2025)
di: Rajasegaran, Jathushan, et al.
Pubblicazione: (2025)
Continuous 3D Perception Model with Persistent State
di: Wang, Qianqian, et al.
Pubblicazione: (2025)
di: Wang, Qianqian, et al.
Pubblicazione: (2025)
An Interpretable Deep Learning Approach for Morphological Script Type Analysis
di: Vlachou-Efstathiou, Malamatenia, et al.
Pubblicazione: (2024)
di: Vlachou-Efstathiou, Malamatenia, et al.
Pubblicazione: (2024)
GPS as a Control Signal for Image Generation
di: Feng, Chao, et al.
Pubblicazione: (2025)
di: Feng, Chao, et al.
Pubblicazione: (2025)
Interpreting CLIP's Image Representation via Text-Based Decomposition
di: Gandelsman, Yossi, et al.
Pubblicazione: (2023)
di: Gandelsman, Yossi, et al.
Pubblicazione: (2023)
Disentangled 3D Scene Generation with Layout Learning
di: Epstein, Dave, et al.
Pubblicazione: (2024)
di: Epstein, Dave, et al.
Pubblicazione: (2024)
Vision Transformers Don't Need Trained Registers
di: Jiang, Nick, et al.
Pubblicazione: (2025)
di: Jiang, Nick, et al.
Pubblicazione: (2025)
KiVA: Kid-inspired Visual Analogies for Testing Large Multimodal Models
di: Yiu, Eunice, et al.
Pubblicazione: (2024)
di: Yiu, Eunice, et al.
Pubblicazione: (2024)
Synthesizing Moving People with 3D Control
di: Li, Boyi, et al.
Pubblicazione: (2024)
di: Li, Boyi, et al.
Pubblicazione: (2024)
Poly-Autoregressive Prediction for Modeling Interactions
di: Thakkar, Neerja, et al.
Pubblicazione: (2025)
di: Thakkar, Neerja, et al.
Pubblicazione: (2025)
Frozen Forecasting: A Unified Evaluation
di: Walker, Jacob C, et al.
Pubblicazione: (2025)
di: Walker, Jacob C, et al.
Pubblicazione: (2025)
Neural USD: An object-centric framework for iterative editing and control
di: Escontrela, Alejandro, et al.
Pubblicazione: (2025)
di: Escontrela, Alejandro, et al.
Pubblicazione: (2025)
ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training
di: Jin, Haian, et al.
Pubblicazione: (2026)
di: Jin, Haian, et al.
Pubblicazione: (2026)
HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement
di: Schouten, Marco, et al.
Pubblicazione: (2026)
di: Schouten, Marco, et al.
Pubblicazione: (2026)
Pose Priors from Language Models
di: Subramanian, Sanjay, et al.
Pubblicazione: (2024)
di: Subramanian, Sanjay, et al.
Pubblicazione: (2024)
Advances in Diffusion Models for Image Data Augmentation: A Review of Methods, Models, Evaluation Metrics and Future Research Directions
di: Alimisis, Panagiotis, et al.
Pubblicazione: (2024)
di: Alimisis, Panagiotis, et al.
Pubblicazione: (2024)
Mining Your Own Secrets: Diffusion Classifier Scores for Continual Personalization of Text-to-Image Diffusion Models
di: Jha, Saurav, et al.
Pubblicazione: (2024)
di: Jha, Saurav, et al.
Pubblicazione: (2024)
VLMine: Long-Tail Data Mining with Vision Language Models
di: Ye, Mao, et al.
Pubblicazione: (2024)
di: Ye, Mao, et al.
Pubblicazione: (2024)
OpenStreetView-5M: The Many Roads to Global Visual Geolocation
di: Astruc, Guillaume, et al.
Pubblicazione: (2024)
di: Astruc, Guillaume, et al.
Pubblicazione: (2024)
Forecasting Motion in the Wild
di: Thakkar, Neerja, et al.
Pubblicazione: (2026)
di: Thakkar, Neerja, et al.
Pubblicazione: (2026)
Video Interpolation with Diffusion Models
di: Jain, Siddhant, et al.
Pubblicazione: (2024)
di: Jain, Siddhant, et al.
Pubblicazione: (2024)
A SAM based Tool for Semi-Automatic Food Annotation
di: Rahman, Lubnaa Abdur, et al.
Pubblicazione: (2024)
di: Rahman, Lubnaa Abdur, et al.
Pubblicazione: (2024)
Generative Powers of Ten
di: Wang, Xiaojuan, et al.
Pubblicazione: (2023)
di: Wang, Xiaojuan, et al.
Pubblicazione: (2023)
FashionSD-X: Multimodal Fashion Garment Synthesis using Latent Diffusion
di: Singh, Abhishek Kumar, et al.
Pubblicazione: (2024)
di: Singh, Abhishek Kumar, et al.
Pubblicazione: (2024)
Rethinking Score Distillation as a Bridge Between Image Distributions
di: McAllister, David, et al.
Pubblicazione: (2024)
di: McAllister, David, et al.
Pubblicazione: (2024)
DiffusionAct: Controllable Diffusion Autoencoder for One-shot Face Reenactment
di: Bounareli, Stella, et al.
Pubblicazione: (2024)
di: Bounareli, Stella, et al.
Pubblicazione: (2024)
Effective Data Augmentation With Diffusion Models
di: Trabucco, Brandon, et al.
Pubblicazione: (2023)
di: Trabucco, Brandon, et al.
Pubblicazione: (2023)
Mobile Video Diffusion
di: Yahia, Haitam Ben, et al.
Pubblicazione: (2024)
di: Yahia, Haitam Ben, et al.
Pubblicazione: (2024)
Latent Drifting in Diffusion Models for Counterfactual Medical Image Synthesis
di: Yeganeh, Yousef, et al.
Pubblicazione: (2024)
di: Yeganeh, Yousef, et al.
Pubblicazione: (2024)
Bootstrapping Diffusion: Diffusion Model Training Leveraging Partial and Corrupted Data
di: Ma, Xudong
Pubblicazione: (2025)
di: Ma, Xudong
Pubblicazione: (2025)
Synergy and Synchrony in Couple Dances
di: Maluleke, Vongani, et al.
Pubblicazione: (2024)
di: Maluleke, Vongani, et al.
Pubblicazione: (2024)
Infinite Texture: Text-guided High Resolution Diffusion Texture Synthesis
di: Wang, Yifan, et al.
Pubblicazione: (2024)
di: Wang, Yifan, et al.
Pubblicazione: (2024)
Readout Guidance: Learning Control from Diffusion Features
di: Luo, Grace, et al.
Pubblicazione: (2023)
di: Luo, Grace, et al.
Pubblicazione: (2023)
ShapeShifter: 3D Variations Using Multiscale and Sparse Point-Voxel Diffusion
di: Maruani, Nissim, et al.
Pubblicazione: (2025)
di: Maruani, Nissim, et al.
Pubblicazione: (2025)
Diffusion Image Generation with Explicit Modeling of Data Manifold Geometry
di: Xue, Duoduo, et al.
Pubblicazione: (2026)
di: Xue, Duoduo, et al.
Pubblicazione: (2026)
Diffusion Hyperfeatures: Searching Through Time and Space for Semantic Correspondence
di: Luo, Grace, et al.
Pubblicazione: (2023)
di: Luo, Grace, et al.
Pubblicazione: (2023)
M$^2$-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining
di: Lv, Rui, et al.
Pubblicazione: (2026)
di: Lv, Rui, et al.
Pubblicazione: (2026)
Interpreting the Second-Order Effects of Neurons in CLIP
di: Gandelsman, Yossi, et al.
Pubblicazione: (2024)
di: Gandelsman, Yossi, et al.
Pubblicazione: (2024)
Visual Jenga: Discovering Object Dependencies via Counterfactual Inpainting
di: Bhattad, Anand, et al.
Pubblicazione: (2025)
di: Bhattad, Anand, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale
di: Koepke, A. Sophia, et al.
Pubblicazione: (2026) -
Gaussian Masked Autoencoders
di: Rajasegaran, Jathushan, et al.
Pubblicazione: (2025) -
Continuous 3D Perception Model with Persistent State
di: Wang, Qianqian, et al.
Pubblicazione: (2025) -
An Interpretable Deep Learning Approach for Morphological Script Type Analysis
di: Vlachou-Efstathiou, Malamatenia, et al.
Pubblicazione: (2024) -
GPS as a Control Signal for Image Generation
di: Feng, Chao, et al.
Pubblicazione: (2025)