EDT: An Efficient Diffusion Transformer Framework Inspired by Human-like Sketching
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Xinwang, Liu, Ning, Zhu, Yichen, Feng, Feifei, Tang, Jian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CoProSketch: Controllable and Progressive Sketch Generation with Diffusion Model
by: Zhan, Ruohao, et al.
Published: (2025)
by: Zhan, Ruohao, et al.
Published: (2025)
SketchGraphNet: A Memory-Efficient Hybrid Graph Transformer for Large-Scale Sketch Corpora Recognition
by: Chen, Shilong, et al.
Published: (2026)
by: Chen, Shilong, et al.
Published: (2026)
VidSketch: Hand-drawn Sketch-Driven Video Generation with Diffusion Control
by: Jiang, Lifan, et al.
Published: (2025)
by: Jiang, Lifan, et al.
Published: (2025)
U-Sketch: An Efficient Approach for Sketch to Image Diffusion Models
by: Mitsouras, Ilias, et al.
Published: (2024)
by: Mitsouras, Ilias, et al.
Published: (2024)
Content-Conditioned Generation of Stylized Free hand Sketches
by: Liu, Jiajun, et al.
Published: (2024)
by: Liu, Jiajun, et al.
Published: (2024)
ImgCoT: Compressing Long Chain of Thought into Compact Visual Tokens for Efficient Reasoning of Large Language Model
by: Chen, Xiaoshu, et al.
Published: (2026)
by: Chen, Xiaoshu, et al.
Published: (2026)
EPSD: Early Pruning with Self-Distillation for Efficient Model Compression
by: Chen, Dong, et al.
Published: (2024)
by: Chen, Dong, et al.
Published: (2024)
Neural Network Optimization Reimagined: Decoupled Techniques for Scratch and Fine-Tuning
by: Ning, Xin, et al.
Published: (2026)
by: Ning, Xin, et al.
Published: (2026)
Efficient Prompt Tuning of Large Vision-Language Model for Fine-Grained Ship Classification
by: Lan, Long, et al.
Published: (2024)
by: Lan, Long, et al.
Published: (2024)
RefSAM: Efficiently Adapting Segmenting Anything Model for Referring Video Object Segmentation
by: Li, Yonglin, et al.
Published: (2023)
by: Li, Yonglin, et al.
Published: (2023)
Gen-AI Police Sketches with Stable Diffusion
by: Fidalgo, Nicholas, et al.
Published: (2025)
by: Fidalgo, Nicholas, et al.
Published: (2025)
SketchRef: a Multi-Task Evaluation Benchmark for Sketch Synthesis
by: Lin, Xingyue, et al.
Published: (2024)
by: Lin, Xingyue, et al.
Published: (2024)
Diffusion-Based Restoration for Multi-Modal 3D Object Detection in Adverse Weather
by: He, Zhijian, et al.
Published: (2025)
by: He, Zhijian, et al.
Published: (2025)
InterSketch: An Interleaved Reasoning Model with Self-correcting Visual Sketch and Stepwise Reward
by: Ning, Zhiwei, et al.
Published: (2026)
by: Ning, Zhiwei, et al.
Published: (2026)
SRDiffusion: Accelerate Video Diffusion Inference via Sketching-Rendering Cooperation
by: Cheng, Shenggan, et al.
Published: (2025)
by: Cheng, Shenggan, et al.
Published: (2025)
Q-DiT: Accurate Post-Training Quantization for Diffusion Transformers
by: Chen, Lei, et al.
Published: (2024)
by: Chen, Lei, et al.
Published: (2024)
Do Generalised Classifiers really work on Human Drawn Sketches?
by: Bandyopadhyay, Hmrishav, et al.
Published: (2024)
by: Bandyopadhyay, Hmrishav, et al.
Published: (2024)
OminiControl2: Efficient Conditioning for Diffusion Transformers
by: Tan, Zhenxiong, et al.
Published: (2025)
by: Tan, Zhenxiong, et al.
Published: (2025)
Can Vision-Language Models Replace Human Annotators: A Case Study with CelebA Dataset
by: Lu, Haoming, et al.
Published: (2024)
by: Lu, Haoming, et al.
Published: (2024)
Seeing Through Deepfakes: A Human-Inspired Framework for Multi-Face Detection
by: Hu, Juan, et al.
Published: (2025)
by: Hu, Juan, et al.
Published: (2025)
VMonarch: Efficient Video Diffusion Transformers with Structured Attention
by: Liang, Cheng, et al.
Published: (2026)
by: Liang, Cheng, et al.
Published: (2026)
DGAE: Diffusion-Guided Autoencoder for Efficient Latent Representation Learning
by: Liu, Dongxu, et al.
Published: (2025)
by: Liu, Dongxu, et al.
Published: (2025)
KnobGen: Controlling the Sophistication of Artwork in Sketch-Based Diffusion Models
by: Navard, Pouyan, et al.
Published: (2024)
by: Navard, Pouyan, et al.
Published: (2024)
SANA-Video: Efficient Video Generation with Block Linear Diffusion Transformer
by: Chen, Junsong, et al.
Published: (2025)
by: Chen, Junsong, et al.
Published: (2025)
HQ-DiT: Efficient Diffusion Transformer with FP4 Hybrid Quantization
by: Liu, Wenxuan, et al.
Published: (2024)
by: Liu, Wenxuan, et al.
Published: (2024)
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
by: Brioschi, Riccardo, et al.
Published: (2025)
by: Brioschi, Riccardo, et al.
Published: (2025)
HARMamba: Efficient and Lightweight Wearable Sensor Human Activity Recognition Based on Bidirectional Mamba
by: Li, Shuangjian, et al.
Published: (2024)
by: Li, Shuangjian, et al.
Published: (2024)
Towards Self-Improvement of Diffusion Models via Group Preference Optimization
by: Chen, Renjie, et al.
Published: (2025)
by: Chen, Renjie, et al.
Published: (2025)
CLIP is All You Need for Human-like Semantic Representations in Stable Diffusion
by: Braunstein, Cameron, et al.
Published: (2025)
by: Braunstein, Cameron, et al.
Published: (2025)
Prompt-Consistency Image Generation (PCIG): A Unified Framework Integrating LLMs, Knowledge Graphs, and Controllable Diffusion Models
by: Sun, Yichen, et al.
Published: (2024)
by: Sun, Yichen, et al.
Published: (2024)
OmniDiT: Extending Diffusion Transformer to Omni-VTON Framework
by: Zeng, Weixuan, et al.
Published: (2026)
by: Zeng, Weixuan, et al.
Published: (2026)
Sketch2NeRF: Multi-view Sketch-guided Text-to-3D Generation
by: Chen, Minglin, et al.
Published: (2024)
by: Chen, Minglin, et al.
Published: (2024)
NuWa: Deriving Lightweight Task-Specific Vision Transformers for Edge Devices
by: Wei, Ziteng, et al.
Published: (2025)
by: Wei, Ziteng, et al.
Published: (2025)
When Training-Free NAS Meets Vision Transformer: A Neural Tangent Kernel Perspective
by: Zhou, Qiqi, et al.
Published: (2024)
by: Zhou, Qiqi, et al.
Published: (2024)
Balanced Multi-view Clustering
by: Li, Zhenglai, et al.
Published: (2025)
by: Li, Zhenglai, et al.
Published: (2025)
SketchINR: A First Look into Sketches as Implicit Neural Representations
by: Bandyopadhyay, Hmrishav, et al.
Published: (2024)
by: Bandyopadhyay, Hmrishav, et al.
Published: (2024)
RealDex: Towards Human-like Grasping for Robotic Dexterous Hand
by: Liu, Yumeng, et al.
Published: (2024)
by: Liu, Yumeng, et al.
Published: (2024)
DiffSketcher: Text Guided Vector Sketch Synthesis through Latent Diffusion Models
by: Xing, Ximing, et al.
Published: (2023)
by: Xing, Ximing, et al.
Published: (2023)
DDiT: Dynamic Patch Scheduling for Efficient Diffusion Transformers
by: Kim, Dahye, et al.
Published: (2026)
by: Kim, Dahye, et al.
Published: (2026)
Reflection Removal through Efficient Adaptation of Diffusion Transformers
by: Zakarin, Daniyar, et al.
Published: (2025)
by: Zakarin, Daniyar, et al.
Published: (2025)
Similar Items
-
CoProSketch: Controllable and Progressive Sketch Generation with Diffusion Model
by: Zhan, Ruohao, et al.
Published: (2025) -
SketchGraphNet: A Memory-Efficient Hybrid Graph Transformer for Large-Scale Sketch Corpora Recognition
by: Chen, Shilong, et al.
Published: (2026) -
VidSketch: Hand-drawn Sketch-Driven Video Generation with Diffusion Control
by: Jiang, Lifan, et al.
Published: (2025) -
U-Sketch: An Efficient Approach for Sketch to Image Diffusion Models
by: Mitsouras, Ilias, et al.
Published: (2024) -
Content-Conditioned Generation of Stylized Free hand Sketches
by: Liu, Jiajun, et al.
Published: (2024)