d-Sketch: Improving Visual Fidelity of Sketch-to-Image Translation with Pretrained Latent Diffusion Models without Retraining
Fuente:
arXiv
Saved in:
| Main Authors: | Roy, Prasun, Bhattacharya, Saumik, Ghosh, Subhankar, Pal, Umapada, Blumenstein, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Semantically Consistent Person Image Generation
by: Roy, Prasun, et al.
Published: (2023)
by: Roy, Prasun, et al.
Published: (2023)
Scene Aware Person Image Generation through Global Contextual Conditioning
by: Roy, Prasun, et al.
Published: (2022)
by: Roy, Prasun, et al.
Published: (2022)
Effects of Degradations on Deep Neural Network Architectures
by: Roy, Prasun, et al.
Published: (2018)
by: Roy, Prasun, et al.
Published: (2018)
Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation
by: Roy, Prasun, et al.
Published: (2025)
by: Roy, Prasun, et al.
Published: (2025)
TIPS: Text-Induced Pose Synthesis
by: Roy, Prasun, et al.
Published: (2022)
by: Roy, Prasun, et al.
Published: (2022)
Sketch&Patch++: Efficient Structure-Aware 3D Gaussian Representation
by: Shi, Yuang, et al.
Published: (2026)
by: Shi, Yuang, et al.
Published: (2026)
STEFANN: Scene Text Editor using Font Adaptive Neural Network
by: Roy, Prasun, et al.
Published: (2019)
by: Roy, Prasun, et al.
Published: (2019)
Multi-scale Attention Guided Pose Transfer
by: Roy, Prasun, et al.
Published: (2022)
by: Roy, Prasun, et al.
Published: (2022)
DrawVideo: Generating Long Video from Storyboard Keyframe Sketches
by: Xu, Chuanzhi, et al.
Published: (2026)
by: Xu, Chuanzhi, et al.
Published: (2026)
Streaming Real-Time Rendered Scenes as 3D Gaussians
by: Siekkinen, Matti, et al.
Published: (2026)
by: Siekkinen, Matti, et al.
Published: (2026)
A Single Atlas is All You Need: Decoder-Side Gaussian Splatting for Immersive Video
by: Mieloch, Dawid, et al.
Published: (2026)
by: Mieloch, Dawid, et al.
Published: (2026)
Rip Current Detection in Nearshore Areas through UAV Video Analysis with Almost Local-Isometric Embedding Techniques on Sphere
by: Sun, Anchen, et al.
Published: (2023)
by: Sun, Anchen, et al.
Published: (2023)
BASICS: Broad quality Assessment of Static point clouds In Compression Scenarios
by: Ak, Ali, et al.
Published: (2023)
by: Ak, Ali, et al.
Published: (2023)
FASTER: A Font-Agnostic Scene Text Editing and Rendering Framework
by: Das, Alloy, et al.
Published: (2023)
by: Das, Alloy, et al.
Published: (2023)
Relightable Gaussian Splatting for Virtual Production Using Image-Based Illumination
by: Azzarelli, Adrian, et al.
Published: (2026)
by: Azzarelli, Adrian, et al.
Published: (2026)
Resolution limit of the eye: how many pixels can we see?
by: Ashraf, Maliha, et al.
Published: (2024)
by: Ashraf, Maliha, et al.
Published: (2024)
Streaming of rendered content with adaptive frame rate and resolution
by: Liu, Yaru, et al.
Published: (2026)
by: Liu, Yaru, et al.
Published: (2026)
AIM 2024 Challenge on Efficient Video Super-Resolution for AV1 Compressed Content
by: Conde, Marcos V, et al.
Published: (2024)
by: Conde, Marcos V, et al.
Published: (2024)
Contrastive Multi-Modal Hypergraph Reasoning for 3D Crowd Mesh Recovery
by: Sun, Minghao, et al.
Published: (2026)
by: Sun, Minghao, et al.
Published: (2026)
AV1 Motion Vector Fidelity and Application for Efficient Optical Flow
by: Zouein, Julien, et al.
Published: (2025)
by: Zouein, Julien, et al.
Published: (2025)
FreqPrior: Improving Video Diffusion Models with Frequency Filtering Gaussian Noise
by: Yuan, Yunlong, et al.
Published: (2025)
by: Yuan, Yunlong, et al.
Published: (2025)
CAGS: Color-Adaptive Volumetric Video Streaming with Dynamic 3D Gaussian Splatting
by: Yin, Daheng, et al.
Published: (2026)
by: Yin, Daheng, et al.
Published: (2026)
Compact Visual Data Representation for Green Multimedia -- A Human Visual System Perspective
by: Chen, Peilin, et al.
Published: (2024)
by: Chen, Peilin, et al.
Published: (2024)
NeurOp-Diff:Continuous Remote Sensing Image Super-Resolution via Neural Operator Diffusion
by: Xu, Zihao, et al.
Published: (2025)
by: Xu, Zihao, et al.
Published: (2025)
Foveated Compression for Immersive Telepresence Visualization
by: Schwarz, Max, et al.
Published: (2025)
by: Schwarz, Max, et al.
Published: (2025)
Rethinking Security of Diffusion-based Generative Steganography
by: Zhu, Jihao, et al.
Published: (2026)
by: Zhu, Jihao, et al.
Published: (2026)
Audio-Visual Cross-Modal Compression for Generative Face Video Coding
by: Xu, Youmin, et al.
Published: (2025)
by: Xu, Youmin, et al.
Published: (2025)
Improved Screen Content Coding in VVC Using Soft Context Formation
by: Och, Hannah, et al.
Published: (2023)
by: Och, Hannah, et al.
Published: (2023)
DiV-INR: Extreme Low-Bitrate Diffusion Video Compression with INR Conditioning
by: Çetin, Eren, et al.
Published: (2026)
by: Çetin, Eren, et al.
Published: (2026)
Enhanced Quality Aware-Scalable Underwater Image Compression
by: Zhu, Linwei, et al.
Published: (2025)
by: Zhu, Linwei, et al.
Published: (2025)
Adaptive Wireless Image Semantic Transmission and Over-The-Air Testing
by: Ding, Jiarun, et al.
Published: (2024)
by: Ding, Jiarun, et al.
Published: (2024)
Prompt-based Multimodal Semantic Communication for Multi-spectral Image Segmentation
by: Zhang, Haoshuo, et al.
Published: (2025)
by: Zhang, Haoshuo, et al.
Published: (2025)
Unified ROI-based Image Compression Paradigm with Generalized Gaussian Model
by: Hu, Kai, et al.
Published: (2026)
by: Hu, Kai, et al.
Published: (2026)
Combined Channel and Spatial Attention-based Stereo Endoscopic Image Super-Resolution
by: Hayat, Mansoor, et al.
Published: (2023)
by: Hayat, Mansoor, et al.
Published: (2023)
Real-time 3D Visualization of Radiance Fields on Light Field Displays
by: Kim, Jonghyun, et al.
Published: (2025)
by: Kim, Jonghyun, et al.
Published: (2025)
Freehand Sketch Generation from Mechanical Components
by: Liao, Zhichao, et al.
Published: (2024)
by: Liao, Zhichao, et al.
Published: (2024)
ABC: Adaptive BayesNet Structure Learning for Computational Scalable Multi-task Image Compression
by: Zhang, Yufeng, et al.
Published: (2025)
by: Zhang, Yufeng, et al.
Published: (2025)
MarsSQE: Stereo Quality Enhancement for Martian Images Using Bi-level Cross-view Attention
by: Xu, Mai, et al.
Published: (2024)
by: Xu, Mai, et al.
Published: (2024)
AirGS: Real-Time 4D Gaussian Streaming for Free-Viewpoint Video Experiences
by: Wang, Zhe, et al.
Published: (2025)
by: Wang, Zhe, et al.
Published: (2025)
Synthetic Video Enhances Physical Fidelity in Video Synthesis
by: Zhao, Qi, et al.
Published: (2025)
by: Zhao, Qi, et al.
Published: (2025)
Similar Items
-
Semantically Consistent Person Image Generation
by: Roy, Prasun, et al.
Published: (2023) -
Scene Aware Person Image Generation through Global Contextual Conditioning
by: Roy, Prasun, et al.
Published: (2022) -
Effects of Degradations on Deep Neural Network Architectures
by: Roy, Prasun, et al.
Published: (2018) -
Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation
by: Roy, Prasun, et al.
Published: (2025) -
TIPS: Text-Induced Pose Synthesis
by: Roy, Prasun, et al.
Published: (2022)