Learning to Edit Visual Programs with Self-Supervision
Fuente:
arXiv
Saved in:
| Main Authors: | Jones, R. Kenny, Zhang, Renhao, Ganeshan, Aditya, Ritchie, Daniel |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Infer Generative Template Programs for Visual Concepts
by: Jones, R. Kenny, et al.
Published: (2024)
by: Jones, R. Kenny, et al.
Published: (2024)
Pattern Analogies: Learning to Perform Programmatic Image Edits by Analogy
by: Ganeshan, Aditya, et al.
Published: (2024)
by: Ganeshan, Aditya, et al.
Published: (2024)
ParSEL: Parameterized Shape Editing with Language
by: Ganeshan, Aditya, et al.
Published: (2024)
by: Ganeshan, Aditya, et al.
Published: (2024)
PartComposer: Learning and Composing Part-Level Concepts from Single-Image Examples
by: Liu, Junyu, et al.
Published: (2025)
by: Liu, Junyu, et al.
Published: (2025)
Machine Learning for Scientific Visualization: Ensemble Data Analysis
by: Gadirov, Hamid
Published: (2025)
by: Gadirov, Hamid
Published: (2025)
LoMOE: Localized Multi-Object Editing via Multi-Diffusion
by: Chakrabarty, Goirik, et al.
Published: (2024)
by: Chakrabarty, Goirik, et al.
Published: (2024)
Open-Universe Indoor Scene Generation using LLM Program Synthesis and Uncurated Object Databases
by: Aguina-Kang, Rio, et al.
Published: (2024)
by: Aguina-Kang, Rio, et al.
Published: (2024)
Uncertainty-Informed Volume Visualization using Implicit Neural Representation
by: Saklani, Shanu, et al.
Published: (2024)
by: Saklani, Shanu, et al.
Published: (2024)
Accurate Differential Operators for Hybrid Neural Fields
by: Chetan, Aditya, et al.
Published: (2023)
by: Chetan, Aditya, et al.
Published: (2023)
Diffusion Self-Distillation for Zero-Shot Customized Image Generation
by: Cai, Shengqu, et al.
Published: (2024)
by: Cai, Shengqu, et al.
Published: (2024)
MARVEL-40M+: Multi-Level Visual Elaboration for High-Fidelity Text-to-3D Content Creation
by: Sinha, Sankalp, et al.
Published: (2024)
by: Sinha, Sankalp, et al.
Published: (2024)
Learning to Place Objects with Programs and Iterative Self Training
by: Chang, Adrian, et al.
Published: (2025)
by: Chang, Adrian, et al.
Published: (2025)
On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers
by: Dahary, Omer, et al.
Published: (2026)
by: Dahary, Omer, et al.
Published: (2026)
DesignAsCode: Bridging Structural Editability and Visual Fidelity in Graphic Design Generation
by: Liu, Ziyuan, et al.
Published: (2026)
by: Liu, Ziyuan, et al.
Published: (2026)
Unsupervised Occupancy Learning from Sparse Point Cloud
by: Ouasfi, Amine, et al.
Published: (2024)
by: Ouasfi, Amine, et al.
Published: (2024)
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
by: Dahary, Omer, et al.
Published: (2024)
by: Dahary, Omer, et al.
Published: (2024)
Revisiting Deepfake Detection: Chronological Continual Learning and the Limits of Generalization
by: Fontana, Federico, et al.
Published: (2025)
by: Fontana, Federico, et al.
Published: (2025)
Navigating with Annealing Guidance Scale in Diffusion Space
by: Yehezkel, Shai, et al.
Published: (2025)
by: Yehezkel, Shai, et al.
Published: (2025)
Image Generation from Contextually-Contradictory Prompts
by: Huberman, Saar, et al.
Published: (2025)
by: Huberman, Saar, et al.
Published: (2025)
Be Decisive: Noise-Induced Layouts for Multi-Subject Generation
by: Dahary, Omer, et al.
Published: (2025)
by: Dahary, Omer, et al.
Published: (2025)
Few-Shot Unsupervised Implicit Neural Shape Representation Learning with Spatial Adversaries
by: Ouasfi, Amine, et al.
Published: (2024)
by: Ouasfi, Amine, et al.
Published: (2024)
SCULPT: Shape-Conditioned Unpaired Learning of Pose-dependent Clothed and Textured Human Meshes
by: Sanyal, Soubhik, et al.
Published: (2023)
by: Sanyal, Soubhik, et al.
Published: (2023)
InfiniteDiffusion: Bridging Learned Fidelity and Procedural Utility for Open-World Terrain Generation
by: Goslin, Alexander
Published: (2025)
by: Goslin, Alexander
Published: (2025)
GaussianVAE: Adaptive Learning Dynamics of 3D Gaussians for High-Fidelity Super-Resolution
by: Khalid, Shuja, et al.
Published: (2025)
by: Khalid, Shuja, et al.
Published: (2025)
PAPR in Motion: Seamless Point-level 3D Scene Interpolation
by: Peng, Shichong, et al.
Published: (2024)
by: Peng, Shichong, et al.
Published: (2024)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
by: Zheng, Shuhong, et al.
Published: (2026)
by: Zheng, Shuhong, et al.
Published: (2026)
Asset Harvester: Extracting 3D Assets from Autonomous Driving Logs for Simulation
by: Cao, Tianshi, et al.
Published: (2026)
by: Cao, Tianshi, et al.
Published: (2026)
DiffusionBrowser: Interactive Diffusion Previews via Multi-Branch Decoders
by: Hong, Susung, et al.
Published: (2025)
by: Hong, Susung, et al.
Published: (2025)
Improved Baselines with Representation Autoencoders
by: Singh, Jaskirat, et al.
Published: (2026)
by: Singh, Jaskirat, et al.
Published: (2026)
What matters for Representation Alignment: Global Information or Spatial Structure?
by: Singh, Jaskirat, et al.
Published: (2025)
by: Singh, Jaskirat, et al.
Published: (2025)
End-to-End Training for Unified Tokenization and Latent Denoising
by: Duggal, Shivam, et al.
Published: (2026)
by: Duggal, Shivam, et al.
Published: (2026)
One Trajectory, One Token: Grounded Video Tokenization via Panoptic Sub-object Trajectory
by: Zheng, Chenhao, et al.
Published: (2025)
by: Zheng, Chenhao, et al.
Published: (2025)
MeshSplat: Generalizable Sparse-View Surface Reconstruction via Gaussian Splatting
by: Chang, Hanzhi, et al.
Published: (2025)
by: Chang, Hanzhi, et al.
Published: (2025)
RAGDiffusion: Faithful Cloth Generation via External Knowledge Assimilation
by: Tan, Xianfeng, et al.
Published: (2024)
by: Tan, Xianfeng, et al.
Published: (2024)
Controlling Text-to-Image Diffusion by Orthogonal Finetuning
by: Qiu, Zeju, et al.
Published: (2023)
by: Qiu, Zeju, et al.
Published: (2023)
An objective comparison of methods for augmented reality in laparoscopic liver resection by preoperative-to-intraoperative image fusion
by: Ali, Sharib, et al.
Published: (2024)
by: Ali, Sharib, et al.
Published: (2024)
BloomScene: Lightweight Structured 3D Gaussian Splatting for Crossmodal Scene Generation
by: Hou, Xiaolu, et al.
Published: (2025)
by: Hou, Xiaolu, et al.
Published: (2025)
LRM: Large Reconstruction Model for Single Image to 3D
by: Hong, Yicong, et al.
Published: (2023)
by: Hong, Yicong, et al.
Published: (2023)
ReCapture: Generative Video Camera Controls for User-Provided Videos using Masked Video Fine-Tuning
by: Zhang, David Junhao, et al.
Published: (2024)
by: Zhang, David Junhao, et al.
Published: (2024)
ArtiFixer: Enhancing and Extending 3D Reconstruction with Auto-Regressive Diffusion Models
by: de Lutio, Riccardo, et al.
Published: (2026)
by: de Lutio, Riccardo, et al.
Published: (2026)
Similar Items
-
Learning to Infer Generative Template Programs for Visual Concepts
by: Jones, R. Kenny, et al.
Published: (2024) -
Pattern Analogies: Learning to Perform Programmatic Image Edits by Analogy
by: Ganeshan, Aditya, et al.
Published: (2024) -
ParSEL: Parameterized Shape Editing with Language
by: Ganeshan, Aditya, et al.
Published: (2024) -
PartComposer: Learning and Composing Part-Level Concepts from Single-Image Examples
by: Liu, Junyu, et al.
Published: (2025) -
Machine Learning for Scientific Visualization: Ensemble Data Analysis
by: Gadirov, Hamid
Published: (2025)