Data Augmentation via Latent Diffusion for Saliency Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Aydemir, Bahar, Bhattacharjee, Deblina, Zhang, Tong, Salzmann, Mathieu, Süsstrunk, Sabine |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TempSAL -- Uncovering Temporal Information for Deep Saliency Prediction
by: Aydemir, Bahar, et al.
Published: (2023)
by: Aydemir, Bahar, et al.
Published: (2023)
Unlocking Comics: The AI4VA Dataset for Visual Understanding
by: Grönquist, Peter, et al.
Published: (2024)
by: Grönquist, Peter, et al.
Published: (2024)
OMH: Structured Sparsity via Optimally Matched Hierarchy for Unsupervised Semantic Segmentation
by: Ozaydin, Baran, et al.
Published: (2024)
by: Ozaydin, Baran, et al.
Published: (2024)
DSI2I: Dense Style for Unpaired Image-to-Image Translation
by: Ozaydin, Baran, et al.
Published: (2022)
by: Ozaydin, Baran, et al.
Published: (2022)
Coherent and Multi-modality Image Inpainting via Latent Space Optimization
by: Pan, Lingzhi, et al.
Published: (2024)
by: Pan, Lingzhi, et al.
Published: (2024)
Controlling the Fidelity and Diversity of Deep Generative Models via Pseudo Density
by: Li, Shuangqi, et al.
Published: (2024)
by: Li, Shuangqi, et al.
Published: (2024)
Canonical Latent Representations in Conditional Diffusion Models
by: Xu, Yitao, et al.
Published: (2025)
by: Xu, Yitao, et al.
Published: (2025)
Mitigating Object Dependencies: Improving Point Cloud Self-Supervised Learning through Object Exchange
by: Wu, Yanhao, et al.
Published: (2024)
by: Wu, Yanhao, et al.
Published: (2024)
FDS: Frequency-Aware Denoising Score for Text-Guided Latent Diffusion Image Editing
by: Ren, Yufan, et al.
Published: (2025)
by: Ren, Yufan, et al.
Published: (2025)
Adaptive Multi-step Refinement Network for Robust Point Cloud Registration
by: Chen, Zhi, et al.
Published: (2023)
by: Chen, Zhi, et al.
Published: (2023)
Subtractive Modulative Network with Learnable Periodic Activations
by: Wang, Tiou, et al.
Published: (2026)
by: Wang, Tiou, et al.
Published: (2026)
AdaNCA: Neural Cellular Automata As Adaptors For More Robust Vision Transformer
by: Xu, Yitao, et al.
Published: (2024)
by: Xu, Yitao, et al.
Published: (2024)
Enhancing Frequency Forgery Clues for Diffusion-Generated Image Detection
by: Zhang, Daichi, et al.
Published: (2025)
by: Zhang, Daichi, et al.
Published: (2025)
NEMTO: Neural Environment Matting for Novel View and Relighting Synthesis of Transparent Objects
by: Wang, Dongqing, et al.
Published: (2023)
by: Wang, Dongqing, et al.
Published: (2023)
InNeRF360: Text-Guided 3D-Consistent Object Inpainting on 360-degree Neural Radiance Fields
by: Wang, Dongqing, et al.
Published: (2023)
by: Wang, Dongqing, et al.
Published: (2023)
SINDER: Repairing the Singular Defects of DINOv2
by: Wang, Haoqi, et al.
Published: (2024)
by: Wang, Haoqi, et al.
Published: (2024)
2-Shots in the Dark: Low-Light Denoising with Minimal Data Acquisition
by: Lu, Liying, et al.
Published: (2025)
by: Lu, Liying, et al.
Published: (2025)
Saliency Guided Optimization of Diffusion Latents
by: Wang, Xiwen, et al.
Published: (2024)
by: Wang, Xiwen, et al.
Published: (2024)
Color Constancy in Hyperspectral Imaging via Reduced Spectral Spaces
by: Vidarsson, G. Dofri, et al.
Published: (2026)
by: Vidarsson, G. Dofri, et al.
Published: (2026)
Dark Noise Diffusion: Noise Synthesis for Low-Light Image Denoising
by: Lu, Liying, et al.
Published: (2025)
by: Lu, Liying, et al.
Published: (2025)
Emergent Dynamics in Neural Cellular Automata
by: Xu, Yitao, et al.
Published: (2024)
by: Xu, Yitao, et al.
Published: (2024)
DVMNet++: Rethinking Relative Pose Estimation for Unseen Objects
by: Zhao, Chen, et al.
Published: (2024)
by: Zhao, Chen, et al.
Published: (2024)
Leveraging Hierarchical Image-Text Misalignment for Universal Fake Image Detection
by: Zhang, Daichi, et al.
Published: (2025)
by: Zhang, Daichi, et al.
Published: (2025)
Q-Drift: Quantization-Aware Drift Correction for Diffusion Model Sampling
by: Ryu, Sooyoung, et al.
Published: (2026)
by: Ryu, Sooyoung, et al.
Published: (2026)
VibrantLeaves: A principled parametric image generator for training deep restoration models
by: Achddou, Raphael, et al.
Published: (2025)
by: Achddou, Raphael, et al.
Published: (2025)
Zooming into Comics: Region-Aware RL Improves Fine-Grained Comic Understanding in Vision-Language Models
by: Chen, Yule, et al.
Published: (2025)
by: Chen, Yule, et al.
Published: (2025)
FaceSaliencyAug: Mitigating Geographic, Gender and Stereotypical Biases via Saliency-Based Data Augmentation
by: Kumar, Teerath, et al.
Published: (2024)
by: Kumar, Teerath, et al.
Published: (2024)
DiffSal: Joint Audio and Video Learning for Diffusion Saliency Prediction
by: Xiong, Junwen, et al.
Published: (2024)
by: Xiong, Junwen, et al.
Published: (2024)
Source-Free Domain-Invariant Performance Prediction
by: Khramtsova, Ekaterina, et al.
Published: (2024)
by: Khramtsova, Ekaterina, et al.
Published: (2024)
VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models
by: Ren, Yufan, et al.
Published: (2025)
by: Ren, Yufan, et al.
Published: (2025)
Self-Ensembling Gaussian Splatting for Few-Shot Novel View Synthesis
by: Zhao, Chen, et al.
Published: (2024)
by: Zhao, Chen, et al.
Published: (2024)
Data Augmentation via Latent Diffusion Models for Detecting Smell-Related Objects in Historical Artworks
by: Sheta, Ahmed, et al.
Published: (2025)
by: Sheta, Ahmed, et al.
Published: (2025)
GigaPose: Fast and Robust Novel Object Pose Estimation via One Correspondence
by: Nguyen, Van Nguyen, et al.
Published: (2023)
by: Nguyen, Van Nguyen, et al.
Published: (2023)
Text-Audio-Visual-conditioned Diffusion Model for Video Saliency Prediction
by: Yu, Li, et al.
Published: (2025)
by: Yu, Li, et al.
Published: (2025)
Learning to Weight Parameters for Training Data Attribution
by: Li, Shuangqi, et al.
Published: (2025)
by: Li, Shuangqi, et al.
Published: (2025)
LocPoseNet: Robust Location Prior for Unseen Object Pose Estimation
by: Zhao, Chen, et al.
Published: (2022)
by: Zhao, Chen, et al.
Published: (2022)
Explore In-Context Segmentation via Latent Diffusion Models
by: Wang, Chaoyang, et al.
Published: (2024)
by: Wang, Chaoyang, et al.
Published: (2024)
Mesh Neural Cellular Automata
by: Pajouheshgar, Ehsan, et al.
Published: (2023)
by: Pajouheshgar, Ehsan, et al.
Published: (2023)
Online Long-term Point Tracking in the Foundation Model Era
by: Aydemir, Görkay
Published: (2025)
by: Aydemir, Görkay
Published: (2025)
Generalize or Detect? Towards Robust Semantic Segmentation Under Multiple Distribution Shifts
by: Gao, Zhitong, et al.
Published: (2024)
by: Gao, Zhitong, et al.
Published: (2024)
Similar Items
-
TempSAL -- Uncovering Temporal Information for Deep Saliency Prediction
by: Aydemir, Bahar, et al.
Published: (2023) -
Unlocking Comics: The AI4VA Dataset for Visual Understanding
by: Grönquist, Peter, et al.
Published: (2024) -
OMH: Structured Sparsity via Optimally Matched Hierarchy for Unsupervised Semantic Segmentation
by: Ozaydin, Baran, et al.
Published: (2024) -
DSI2I: Dense Style for Unpaired Image-to-Image Translation
by: Ozaydin, Baran, et al.
Published: (2022) -
Coherent and Multi-modality Image Inpainting via Latent Space Optimization
by: Pan, Lingzhi, et al.
Published: (2024)