Teaching an Agent to Sketch One Part at a Time
Fuente:
arXiv
Saved in:
| Main Authors: | Du, Xiaodan, Xu, Ruize, Yunis, David, Vinker, Yael, Shakhnarovich, Greg |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
by: Du, Xiaodan, et al.
Published: (2023)
by: Du, Xiaodan, et al.
Published: (2023)
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
by: Avrahami, Omri, et al.
Published: (2023)
by: Avrahami, Omri, et al.
Published: (2023)
Instance Segmentation of Scene Sketches Using Natural Image Priors
by: Tang, Mia, et al.
Published: (2025)
by: Tang, Mia, et al.
Published: (2025)
SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
by: Gueuwou, Shester, et al.
Published: (2024)
by: Gueuwou, Shester, et al.
Published: (2024)
ReCapture: Generative Video Camera Controls for User-Provided Videos using Masked Video Fine-Tuning
by: Zhang, David Junhao, et al.
Published: (2024)
by: Zhang, David Junhao, et al.
Published: (2024)
One Trajectory, One Token: Grounded Video Tokenization via Panoptic Sub-object Trajectory
by: Zheng, Chenhao, et al.
Published: (2025)
by: Zheng, Chenhao, et al.
Published: (2025)
Sketch2Colab: Sketch-Conditioned Multi-Human Animation via Controllable Flow Distillation
by: Daiya, Divyanshu, et al.
Published: (2026)
by: Daiya, Divyanshu, et al.
Published: (2026)
RealFill: Reference-Driven Generation for Authentic Image Completion
by: Tang, Luming, et al.
Published: (2023)
by: Tang, Luming, et al.
Published: (2023)
PASTA: Controllable Part-Aware Shape Generation with Autoregressive Transformers
by: Li, Songlin, et al.
Published: (2024)
by: Li, Songlin, et al.
Published: (2024)
HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models
by: Ruiz, Nataniel, et al.
Published: (2023)
by: Ruiz, Nataniel, et al.
Published: (2023)
STEP-Parts: Geometric Partitioning of Boundary Representations for Large-Scale CAD Processing
by: Fan, Shen, et al.
Published: (2026)
by: Fan, Shen, et al.
Published: (2026)
SplArt: Articulation Estimation and Part-Level Reconstruction with 3D Gaussian Splatting
by: Lin, Shengjie, et al.
Published: (2025)
by: Lin, Shengjie, et al.
Published: (2025)
AirSketch: Generative Motion to Sketch
by: Lim, Hui Xian Grace, et al.
Published: (2024)
by: Lim, Hui Xian Grace, et al.
Published: (2024)
SENS: Part-Aware Sketch-based Implicit Neural Shape Modeling
by: Binninger, Alexandre, et al.
Published: (2023)
by: Binninger, Alexandre, et al.
Published: (2023)
DGS-LRM: Real-Time Deformable 3D Gaussian Reconstruction From Monocular Videos
by: Lin, Chieh Hubert, et al.
Published: (2025)
by: Lin, Chieh Hubert, et al.
Published: (2025)
Unbounded: A Generative Infinite Game of Character Life Simulation
by: Li, Jialu, et al.
Published: (2024)
by: Li, Jialu, et al.
Published: (2024)
COLLAGE: Collaborative Human-Agent Interaction Generation using Hierarchical Latent Diffusion and Language Models
by: Daiya, Divyanshu, et al.
Published: (2024)
by: Daiya, Divyanshu, et al.
Published: (2024)
PyPotteryInk: One-Step Diffusion Model for Sketch to Publication-ready Archaeological Drawings
by: Cardarelli, Lorenzo
Published: (2025)
by: Cardarelli, Lorenzo
Published: (2025)
BulletGen: Improving 4D Reconstruction with Bullet-Time Generation
by: Rozumny, Denis, et al.
Published: (2025)
by: Rozumny, Denis, et al.
Published: (2025)
Generative Visual Communication in the Era of Vision-Language Models
by: Vinker, Yael
Published: (2024)
by: Vinker, Yael
Published: (2024)
SRDiffusion: Accelerate Video Diffusion Inference via Sketching-Rendering Cooperation
by: Cheng, Shenggan, et al.
Published: (2025)
by: Cheng, Shenggan, et al.
Published: (2025)
Magic Insert: Style-Aware Drag-and-Drop
by: Ruiz, Nataniel, et al.
Published: (2024)
by: Ruiz, Nataniel, et al.
Published: (2024)
Boosting 3D Object Generation through PBR Materials
by: Wang, Yitong, et al.
Published: (2024)
by: Wang, Yitong, et al.
Published: (2024)
SynBench: A Synthetic Benchmark for Non-rigid 3D Point Cloud Registration
by: Monji-Azad, Sara, et al.
Published: (2024)
by: Monji-Azad, Sara, et al.
Published: (2024)
ComboStoc: Combinatorial Stochasticity for Diffusion Generative Models
by: Xu, Rui, et al.
Published: (2024)
by: Xu, Rui, et al.
Published: (2024)
Uncertainty-Informed Volume Visualization using Implicit Neural Representation
by: Saklani, Shanu, et al.
Published: (2024)
by: Saklani, Shanu, et al.
Published: (2024)
BloomScene: Lightweight Structured 3D Gaussian Splatting for Crossmodal Scene Generation
by: Hou, Xiaolu, et al.
Published: (2025)
by: Hou, Xiaolu, et al.
Published: (2025)
Meta 3D Gen
by: Bensadoun, Raphael, et al.
Published: (2024)
by: Bensadoun, Raphael, et al.
Published: (2024)
Time-to-Move: Training-Free Motion Controlled Video Generation via Dual-Clock Denoising
by: Singer, Assaf, et al.
Published: (2025)
by: Singer, Assaf, et al.
Published: (2025)
VectorGym: A Multitask Benchmark for SVG Code Generation, Sketching, and Editing
by: Rodriguez, Juan, et al.
Published: (2026)
by: Rodriguez, Juan, et al.
Published: (2026)
LoopDraw: a Loop-Based Autoregressive Model for Shape Synthesis and Editing
by: Dinh, Nam Anh, et al.
Published: (2022)
by: Dinh, Nam Anh, et al.
Published: (2022)
Navigating with Annealing Guidance Scale in Diffusion Space
by: Yehezkel, Shai, et al.
Published: (2025)
by: Yehezkel, Shai, et al.
Published: (2025)
Image Generation from Contextually-Contradictory Prompts
by: Huberman, Saar, et al.
Published: (2025)
by: Huberman, Saar, et al.
Published: (2025)
Edge-preserving noise for diffusion models
by: Vandersanden, Jente, et al.
Published: (2024)
by: Vandersanden, Jente, et al.
Published: (2024)
LoMOE: Localized Multi-Object Editing via Multi-Diffusion
by: Chakrabarty, Goirik, et al.
Published: (2024)
by: Chakrabarty, Goirik, et al.
Published: (2024)
Few-Shot Unsupervised Implicit Neural Shape Representation Learning with Spatial Adversaries
by: Ouasfi, Amine, et al.
Published: (2024)
by: Ouasfi, Amine, et al.
Published: (2024)
Towards Practical Single-shot Motion Synthesis
by: Roditakis, Konstantinos, et al.
Published: (2024)
by: Roditakis, Konstantinos, et al.
Published: (2024)
Improving Video Generation with Human Feedback
by: Liu, Jie, et al.
Published: (2025)
by: Liu, Jie, et al.
Published: (2025)
Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and Try-Off
by: Lee, Seungyong, et al.
Published: (2025)
by: Lee, Seungyong, et al.
Published: (2025)
On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers
by: Dahary, Omer, et al.
Published: (2026)
by: Dahary, Omer, et al.
Published: (2026)
Similar Items
-
Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
by: Du, Xiaodan, et al.
Published: (2023) -
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
by: Avrahami, Omri, et al.
Published: (2023) -
Instance Segmentation of Scene Sketches Using Natural Image Priors
by: Tang, Mia, et al.
Published: (2025) -
SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
by: Gueuwou, Shester, et al.
Published: (2024) -
ReCapture: Generative Video Camera Controls for User-Provided Videos using Masked Video Fine-Tuning
by: Zhang, David Junhao, et al.
Published: (2024)