Cora: Correspondence-aware image editing using few step diffusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Alimohammadi, Amirhossein, Mikaeili, Aryan, Nag, Sauradip, Hassanpour, Negar, Tagliasacchi, Andrea, Mahdavi-Amiri, Ali |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DUDF: Differentiable Unsigned Distance Fields with Hyperbolic Scaling
von: Fainstein, Miguel, et al.
Veröffentlicht: (2024)
von: Fainstein, Miguel, et al.
Veröffentlicht: (2024)
Consistent Zero-shot 3D Texture Synthesis Using Geometry-aware Diffusion and Temporal Video Models
von: Kang, Donggoo, et al.
Veröffentlicht: (2025)
von: Kang, Donggoo, et al.
Veröffentlicht: (2025)
Quantized Vision-Language Models for Damage Assessment: A Comparative Study of LLaVA-1.5-7B Quantization Levels
von: Yasuno, Takato
Veröffentlicht: (2026)
von: Yasuno, Takato
Veröffentlicht: (2026)
Pointing-Based Object Recognition
von: Hajdúch, Lukáš, et al.
Veröffentlicht: (2026)
von: Hajdúch, Lukáš, et al.
Veröffentlicht: (2026)
Gaussian Splatting: 3D Reconstruction and Novel View Synthesis, a Review
von: Dalal, Anurag, et al.
Veröffentlicht: (2024)
von: Dalal, Anurag, et al.
Veröffentlicht: (2024)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
FAME: Feature Activation Map Explanation on Image Classification and Face Recognition
von: Zhang, Xinyi, et al.
Veröffentlicht: (2026)
von: Zhang, Xinyi, et al.
Veröffentlicht: (2026)
STimage-1K4M: A histopathology image-gene expression dataset for spatial transcriptomics
von: Chen, Jiawen, et al.
Veröffentlicht: (2024)
von: Chen, Jiawen, et al.
Veröffentlicht: (2024)
Supervised Contrastive Learning for Few-Shot AI-Generated Image Detection and Attribution
von: Urueña, Jaime Álvarez, et al.
Veröffentlicht: (2025)
von: Urueña, Jaime Álvarez, et al.
Veröffentlicht: (2025)
Scalable Face Security Vision Foundation Model for Deepfake, Diffusion, and Spoofing Detection
von: Wang, Gaojian, et al.
Veröffentlicht: (2025)
von: Wang, Gaojian, et al.
Veröffentlicht: (2025)
When in Doubt, Think Slow: Iterative Reasoning with Latent Imagination
von: Benfeghoul, Martin, et al.
Veröffentlicht: (2024)
von: Benfeghoul, Martin, et al.
Veröffentlicht: (2024)
Symmetry Awareness Encoded Deep Learning Framework for Brain Imaging Analysis
von: Ma, Yang, et al.
Veröffentlicht: (2024)
von: Ma, Yang, et al.
Veröffentlicht: (2024)
Differentiable Hierarchical Visual Tokenization
von: Aasan, Marius, et al.
Veröffentlicht: (2025)
von: Aasan, Marius, et al.
Veröffentlicht: (2025)
Analyzing Quality, Bias, and Performance in Text-to-Image Generative Models
von: Masrourisaadat, Nila, et al.
Veröffentlicht: (2024)
von: Masrourisaadat, Nila, et al.
Veröffentlicht: (2024)
Feature-Augmented Deep Networks for Multiscale Building Segmentation in High-Resolution UAV and Satellite Imagery
von: Maniyar, Chintan B., et al.
Veröffentlicht: (2025)
von: Maniyar, Chintan B., et al.
Veröffentlicht: (2025)
Towards Onboard Continuous Change Detection for Floods
von: Kyselica, Daniel, et al.
Veröffentlicht: (2026)
von: Kyselica, Daniel, et al.
Veröffentlicht: (2026)
Non-Robust Features are Not Always Useful in One-Class Classification
von: Lau, Matthew, et al.
Veröffentlicht: (2024)
von: Lau, Matthew, et al.
Veröffentlicht: (2024)
Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models
von: Li, Jinhao, et al.
Veröffentlicht: (2024)
von: Li, Jinhao, et al.
Veröffentlicht: (2024)
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
von: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Veröffentlicht: (2025)
von: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Veröffentlicht: (2025)
EDSNet: Efficient-DSNet for Video Summarization
von: Prasad, Ashish, et al.
Veröffentlicht: (2024)
von: Prasad, Ashish, et al.
Veröffentlicht: (2024)
See-through: Single-image Layer Decomposition for Anime Characters
von: Lin, Jian, et al.
Veröffentlicht: (2026)
von: Lin, Jian, et al.
Veröffentlicht: (2026)
GeoPos: A Minimal Positional Encoding for Enhanced Fine-Grained Details in Image Synthesis Using Convolutional Neural Networks
von: Hosseini, Mehran, et al.
Veröffentlicht: (2024)
von: Hosseini, Mehran, et al.
Veröffentlicht: (2024)
HOSC: A Periodic Activation Function for Preserving Sharp Features in Implicit Neural Representations
von: Serrano, Danzel, et al.
Veröffentlicht: (2024)
von: Serrano, Danzel, et al.
Veröffentlicht: (2024)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
von: Shahin, Nada, et al.
Veröffentlicht: (2025)
von: Shahin, Nada, et al.
Veröffentlicht: (2025)
Learning 3D object-centric representation through prediction
von: Day, John, et al.
Veröffentlicht: (2024)
von: Day, John, et al.
Veröffentlicht: (2024)
S-HR-VQVAE: Sequential Hierarchical Residual Learning Vector Quantized Variational Autoencoder for Video Prediction
von: Adiban, Mohammad, et al.
Veröffentlicht: (2023)
von: Adiban, Mohammad, et al.
Veröffentlicht: (2023)
U-Net-Like Spiking Neural Networks for Single Image Dehazing
von: Li, Huibin, et al.
Veröffentlicht: (2025)
von: Li, Huibin, et al.
Veröffentlicht: (2025)
Fusing Structure from Motion and Simulation-Augmented Pose Regression from Optical Flow for Challenging Indoor Environments
von: Ott, Felix, et al.
Veröffentlicht: (2023)
von: Ott, Felix, et al.
Veröffentlicht: (2023)
3DGEER: 3D Gaussian Rendering Made Exact and Efficient for Generic Cameras
von: Huang, Zixun, et al.
Veröffentlicht: (2025)
von: Huang, Zixun, et al.
Veröffentlicht: (2025)
Appearance-Invariant Detection of Suggestive Motion via Laban Movement Descriptors on SMPL Skeletons
von: Ahn, Jaehoon, et al.
Veröffentlicht: (2026)
von: Ahn, Jaehoon, et al.
Veröffentlicht: (2026)
A Spitting Image: Modular Superpixel Tokenization in Vision Transformers
von: Aasan, Marius, et al.
Veröffentlicht: (2024)
von: Aasan, Marius, et al.
Veröffentlicht: (2024)
Lost in Latent Space: Disentangled Models and the Challenge of Combinatorial Generalisation
von: Montero, Milton L., et al.
Veröffentlicht: (2022)
von: Montero, Milton L., et al.
Veröffentlicht: (2022)
ADAT: Time-Series-Aware Adaptive Transformer Architecture for Sign Language Translation
von: Shahin, Nada, et al.
Veröffentlicht: (2025)
von: Shahin, Nada, et al.
Veröffentlicht: (2025)
GLEaN: A Text-to-image Bias Detection Approach for Public Comprehension
von: Ding, Bochu, et al.
Veröffentlicht: (2026)
von: Ding, Bochu, et al.
Veröffentlicht: (2026)
Sat-JEPA-Diff: Bridging Self-Supervised Learning and Generative Diffusion for Remote Sensing
von: Komurcu, Kursat, et al.
Veröffentlicht: (2026)
von: Komurcu, Kursat, et al.
Veröffentlicht: (2026)
SSD-GS: Scattering and Shadow Decomposition for Relightable 3D Gaussian Splatting
von: Zheng, Iris, et al.
Veröffentlicht: (2026)
von: Zheng, Iris, et al.
Veröffentlicht: (2026)
MRD: Using Physically Based Differentiable Rendering to Probe Vision Models for 3D Scene Understanding
von: Beilharz, Benjamin, et al.
Veröffentlicht: (2025)
von: Beilharz, Benjamin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DUDF: Differentiable Unsigned Distance Fields with Hyperbolic Scaling
von: Fainstein, Miguel, et al.
Veröffentlicht: (2024) -
Consistent Zero-shot 3D Texture Synthesis Using Geometry-aware Diffusion and Temporal Video Models
von: Kang, Donggoo, et al.
Veröffentlicht: (2025) -
Quantized Vision-Language Models for Damage Assessment: A Comparative Study of LLaVA-1.5-7B Quantization Levels
von: Yasuno, Takato
Veröffentlicht: (2026) -
Pointing-Based Object Recognition
von: Hajdúch, Lukáš, et al.
Veröffentlicht: (2026) -
Gaussian Splatting: 3D Reconstruction and Novel View Synthesis, a Review
von: Dalal, Anurag, et al.
Veröffentlicht: (2024)