LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Yiren, Chen, Danze, Shou, Mike Zheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OmniPSD: Layered PSD Generation with Diffusion Transformer
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
MakeAnything: Harnessing Diffusion Transformers for Multi-Domain Procedural Sequence Generation
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
VISTA: Triplet-Supervised Video Style Transfer with Diffusion Transformers
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
DiffSim: Taming Diffusion Models for Evaluating Visual Similarity
von: Song, Yiren, et al.
Veröffentlicht: (2024)
von: Song, Yiren, et al.
Veröffentlicht: (2024)
Edit2Perceive: Image Editing Diffusion Models Are Strong Dense Perceivers
von: Shi, Yiqing, et al.
Veröffentlicht: (2025)
von: Shi, Yiqing, et al.
Veröffentlicht: (2025)
Mitty: Diffusion-based Human-to-Robot Video Generation
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
TransAnimate: Taming Layer Diffusion to Generate RGBA Video
von: Chen, Xuewei, et al.
Veröffentlicht: (2025)
von: Chen, Xuewei, et al.
Veröffentlicht: (2025)
OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
AmodalSVG: Amodal Image Vectorization via Semantic Layer Peeling
von: Hu, Juncheng, et al.
Veröffentlicht: (2026)
von: Hu, Juncheng, et al.
Veröffentlicht: (2026)
X-Humanoid: Robotize Human Videos to Generate Humanoid Videos at Scale
von: Yang, Pei, et al.
Veröffentlicht: (2025)
von: Yang, Pei, et al.
Veröffentlicht: (2025)
Steganalysis on Digital Watermarking: Is Your Defense Truly Impervious?
von: Yang, Pei, et al.
Veröffentlicht: (2024)
von: Yang, Pei, et al.
Veröffentlicht: (2024)
StreamingEffect: Real-Time Human-Centric Video Effect Generation
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
IDProtector: An Adversarial Noise Encoder to Protect Against ID-Preserving Image Generation
von: Song, Yiren, et al.
Veröffentlicht: (2024)
von: Song, Yiren, et al.
Veröffentlicht: (2024)
RingID: Rethinking Tree-Ring Watermarking for Enhanced Multi-Key Identification
von: Ci, Hai, et al.
Veröffentlicht: (2024)
von: Ci, Hai, et al.
Veröffentlicht: (2024)
DiffDecompose: Layer-Wise Decomposition of Alpha-Composited Images via Diffusion Transformers
von: Wang, Zitong, et al.
Veröffentlicht: (2025)
von: Wang, Zitong, et al.
Veröffentlicht: (2025)
D-AR: Diffusion via Autoregressive Models
von: Gao, Ziteng, et al.
Veröffentlicht: (2025)
von: Gao, Ziteng, et al.
Veröffentlicht: (2025)
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
WorldWander: Bridging Egocentric and Exocentric Worlds in Video Generation
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
WMAdapter: Adding WaterMark Control to Latent Diffusion Models
von: Ci, Hai, et al.
Veröffentlicht: (2024)
von: Ci, Hai, et al.
Veröffentlicht: (2024)
Personalized Vision via Visual In-Context Learning
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
SWEET: Sparse World Modeling with Image Editing for Embodied Task Execution
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
OmniHumanoid: Streaming Cross-Embodiment Video Generation with Paired-Free Adaptation
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
UENR-600K: A Large-Scale Physically Grounded Dataset for Nighttime Video Deraining
von: Yang, Pei, et al.
Veröffentlicht: (2026)
von: Yang, Pei, et al.
Veröffentlicht: (2026)
SVGThinker: Instruction-Aligned and Reasoning-Driven Text-to-SVG Generation
von: Chen, Hanqi, et al.
Veröffentlicht: (2025)
von: Chen, Hanqi, et al.
Veröffentlicht: (2025)
Backward-Compatible Aligned Representations via an Orthogonal Transformation Layer
von: Ricci, Simone, et al.
Veröffentlicht: (2024)
von: Ricci, Simone, et al.
Veröffentlicht: (2024)
TPDiff: Temporal Pyramid Video Diffusion Model
von: Ran, Lingmin, et al.
Veröffentlicht: (2025)
von: Ran, Lingmin, et al.
Veröffentlicht: (2025)
Loom: Diffusion-Transformer for Interleaved Generation
von: Ye, Mingcheng, et al.
Veröffentlicht: (2025)
von: Ye, Mingcheng, et al.
Veröffentlicht: (2025)
The Consistency Critic: Correcting Inconsistencies in Generated Images via Reference-Guided Attentive Alignment
von: Ouyang, Ziheng, et al.
Veröffentlicht: (2025)
von: Ouyang, Ziheng, et al.
Veröffentlicht: (2025)
LiveSVG: Zero-Shot SVG Animation via Video Generation
von: Levy, Matan, et al.
Veröffentlicht: (2026)
von: Levy, Matan, et al.
Veröffentlicht: (2026)
Pluggable Pruning with Contiguous Layer Distillation for Diffusion Transformers
von: Ma, Jian, et al.
Veröffentlicht: (2025)
von: Ma, Jian, et al.
Veröffentlicht: (2025)
H2R-Grounder: A Paired-Data-Free Paradigm for Translating Human Interaction Videos into Physically Grounded Robot Videos
von: Ci, Hai, et al.
Veröffentlicht: (2025)
von: Ci, Hai, et al.
Veröffentlicht: (2025)
UniMoD: Efficient Unified Multimodal Transformers with Mixture-of-Depths
von: Mao, Weijia, et al.
Veröffentlicht: (2025)
von: Mao, Weijia, et al.
Veröffentlicht: (2025)
PhotoDoodle: Learning Artistic Image Editing from Few-Shot Pairwise Data
von: Huang, Shijie, et al.
Veröffentlicht: (2025)
von: Huang, Shijie, et al.
Veröffentlicht: (2025)
PAI-Studio: Cinematic Video Background Replacement with Camera-Aware Motion
von: Gao, Heyuan, et al.
Veröffentlicht: (2026)
von: Gao, Heyuan, et al.
Veröffentlicht: (2026)
Anti-Reference: Universal and Immediate Defense Against Reference-Based Generation
von: Song, Yiren, et al.
Veröffentlicht: (2024)
von: Song, Yiren, et al.
Veröffentlicht: (2024)
PANDA: Towards Generalist Video Anomaly Detection via Agentic AI Engineer
von: Yang, Zhiwei, et al.
Veröffentlicht: (2025)
von: Yang, Zhiwei, et al.
Veröffentlicht: (2025)
HALO: Human-Aligned End-to-end Image Retargeting with Layered Transformations
von: Xu, Yiran, et al.
Veröffentlicht: (2025)
von: Xu, Yiran, et al.
Veröffentlicht: (2025)
Tuning-Free Image Editing with Fidelity and Editability via Unified Latent Diffusion Model
von: Mao, Qi, et al.
Veröffentlicht: (2025)
von: Mao, Qi, et al.
Veröffentlicht: (2025)
InternSVG: Towards Unified SVG Tasks with Multimodal Large Language Models
von: Wang, Haomin, et al.
Veröffentlicht: (2025)
von: Wang, Haomin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
OmniPSD: Layered PSD Generation with Diffusion Transformer
von: Liu, Cheng, et al.
Veröffentlicht: (2025) -
MakeAnything: Harnessing Diffusion Transformers for Multi-Domain Procedural Sequence Generation
von: Song, Yiren, et al.
Veröffentlicht: (2025) -
VISTA: Triplet-Supervised Video Style Transfer with Diffusion Transformers
von: Song, Yiren, et al.
Veröffentlicht: (2026) -
DiffSim: Taming Diffusion Models for Evaluating Visual Similarity
von: Song, Yiren, et al.
Veröffentlicht: (2024) -
Edit2Perceive: Image Editing Diffusion Models Are Strong Dense Perceivers
von: Shi, Yiqing, et al.
Veröffentlicht: (2025)