PSDesigner: Automated Graphic Design with a Human-Like Creative Workflow
Fuente:
arXiv
Saved in:
| Main Authors: | Shuai, Xincheng, Tang, Song, Huang, Yutong, Ding, Henghui, Tao, Dacheng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Free-Form Scene Editor: Enabling Multi-Round Object Manipulation like in a 3D Engine
by: Shuai, Xincheng, et al.
Published: (2025)
by: Shuai, Xincheng, et al.
Published: (2025)
GlyphPrinter: Region-Grouped Direct Preference Optimization for Glyph-Accurate Visual Text Rendering
by: Shuai, Xincheng, et al.
Published: (2026)
by: Shuai, Xincheng, et al.
Published: (2026)
SceneDesigner: Controllable Multi-Object Image Generation with 9-DoF Pose Manipulation
by: Qin, Zhenyuan, et al.
Published: (2025)
by: Qin, Zhenyuan, et al.
Published: (2025)
Free-Form Motion Control: Controlling the 6D Poses of Camera and Objects in Video Generation
by: Shuai, Xincheng, et al.
Published: (2025)
by: Shuai, Xincheng, et al.
Published: (2025)
A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models
by: Shuai, Xincheng, et al.
Published: (2024)
by: Shuai, Xincheng, et al.
Published: (2024)
AnyI2V: Animating Any Conditional Image with Motion Control
by: Li, Ziye, et al.
Published: (2025)
by: Li, Ziye, et al.
Published: (2025)
ROSE: Retrieval-Oriented Segmentation Enhancement
by: Tang, Song, et al.
Published: (2026)
by: Tang, Song, et al.
Published: (2026)
CreatiDesign: A Unified Multi-Conditional Diffusion Transformer for Creative Graphic Design
by: Zhang, Hui, et al.
Published: (2025)
by: Zhang, Hui, et al.
Published: (2025)
Decoupling Static and Hierarchical Motion Perception for Referring Video Segmentation
by: He, Shuting, et al.
Published: (2024)
by: He, Shuting, et al.
Published: (2024)
RefMask3D: Language-Guided Transformer for 3D Referring Segmentation
by: He, Shuting, et al.
Published: (2024)
by: He, Shuting, et al.
Published: (2024)
Cross-Domain Knowledge Distillation for Low-Resolution Human Pose Estimation
by: Gu, Zejun, et al.
Published: (2024)
by: Gu, Zejun, et al.
Published: (2024)
SAM3-DMS: Decoupled Memory Selection for Multi-target Video Segmentation of SAM3
by: Shen, Ruiqi, et al.
Published: (2026)
by: Shen, Ruiqi, et al.
Published: (2026)
MOVE: Motion-Guided Few-Shot Video Object Segmentation
by: Ying, Kaining, et al.
Published: (2025)
by: Ying, Kaining, et al.
Published: (2025)
Segment Anything Across Shots: A Method and Benchmark
by: Hu, Hengrui, et al.
Published: (2025)
by: Hu, Hengrui, et al.
Published: (2025)
On Geometry-Enhanced Parameter-Efficient Fine-Tuning for 3D Scene Segmentation
by: Tang, Liyao, et al.
Published: (2025)
by: Tang, Liyao, et al.
Published: (2025)
Multimodal Referring Segmentation: A Survey
by: Ding, Henghui, et al.
Published: (2025)
by: Ding, Henghui, et al.
Published: (2025)
Towards Motion Turing Test: Evaluating Human-Likeness in Humanoid Robots
by: Li, Mingzhe, et al.
Published: (2026)
by: Li, Mingzhe, et al.
Published: (2026)
Geolocation with Real Human Gameplay Data: A Large-Scale Dataset and Human-Like Reasoning Framework
by: Song, Zirui, et al.
Published: (2025)
by: Song, Zirui, et al.
Published: (2025)
Mitigating the Curse of Dimensionality for Certified Robustness via Dual Randomized Smoothing
by: Xia, Song, et al.
Published: (2024)
by: Xia, Song, et al.
Published: (2024)
EffectErase: Joint Video Object Removal and Insertion for High-Quality Effect Erasing
by: Fu, Yang, et al.
Published: (2026)
by: Fu, Yang, et al.
Published: (2026)
SegPoint: Segment Any Point Cloud via Large Language Model
by: He, Shuting, et al.
Published: (2024)
by: He, Shuting, et al.
Published: (2024)
DesignProbe: A Graphic Design Benchmark for Multimodal Large Language Models
by: Lin, Jieru, et al.
Published: (2024)
by: Lin, Jieru, et al.
Published: (2024)
From Elements to Design: A Layered Approach for Automatic Graphic Design Composition
by: Lin, Jiawei, et al.
Published: (2024)
by: Lin, Jiawei, et al.
Published: (2024)
Neuron: Learning Context-Aware Evolving Representations for Zero-Shot Skeleton Action Recognition
by: Chen, Yang, et al.
Published: (2024)
by: Chen, Yang, et al.
Published: (2024)
Automated Workflow for the Detection of Vugs
by: Nasim, M. Quamer, et al.
Published: (2025)
by: Nasim, M. Quamer, et al.
Published: (2025)
PartCraft: Crafting Creative Objects by Parts
by: Ng, Kam Woh, et al.
Published: (2024)
by: Ng, Kam Woh, et al.
Published: (2024)
Towards Omnimodal Expressions and Reasoning in Referring Audio-Visual Segmentation
by: Ying, Kaining, et al.
Published: (2025)
by: Ying, Kaining, et al.
Published: (2025)
ASSISTGUI: Task-Oriented Desktop Graphical User Interface Automation
by: Gao, Difei, et al.
Published: (2023)
by: Gao, Difei, et al.
Published: (2023)
Can Large Vision Language Models Read Maps Like a Human?
by: Xing, Shuo, et al.
Published: (2025)
by: Xing, Shuo, et al.
Published: (2025)
Local-consistent Transformation Learning for Rotation-invariant Point Cloud Analysis
by: Chen, Yiyang, et al.
Published: (2024)
by: Chen, Yiyang, et al.
Published: (2024)
Harnessing Text-to-Image Diffusion Models for Point Cloud Self-Supervised Learning
by: Chen, Yiyang, et al.
Published: (2025)
by: Chen, Yiyang, et al.
Published: (2025)
Vibe Spaces for Creatively Connecting and Expressing Visual Concepts
by: Yang, Huzheng, et al.
Published: (2025)
by: Yang, Huzheng, et al.
Published: (2025)
Predicting Visual Attention in Graphic Design Documents
by: Chakraborty, Souradeep, et al.
Published: (2024)
by: Chakraborty, Souradeep, et al.
Published: (2024)
Ref-SAM3D: Bridging SAM3D with Text for Reference 3D Reconstruction
by: Zhou, Yun, et al.
Published: (2025)
by: Zhou, Yun, et al.
Published: (2025)
Towards Modality-agnostic Label-efficient Segmentation with Entropy-Regularized Distribution Alignment
by: Tang, Liyao, et al.
Published: (2024)
by: Tang, Liyao, et al.
Published: (2024)
Automated Neural Architecture Design for Industrial Defect Detection
by: Liu, Yuxi, et al.
Published: (2025)
by: Liu, Yuxi, et al.
Published: (2025)
CRAFT: Designing Creative and Functional 3D Objects
by: Guo, Michelle, et al.
Published: (2024)
by: Guo, Michelle, et al.
Published: (2024)
Contact-aware Human Motion Generation from Textual Descriptions
by: Ma, Sihan, et al.
Published: (2024)
by: Ma, Sihan, et al.
Published: (2024)
IGD: Instructional Graphic Design with Multimodal Layer Generation
by: Qu, Yadong, et al.
Published: (2025)
by: Qu, Yadong, et al.
Published: (2025)
Multimodal Structure Learning: Disentangling Shared and Specific Topology via Cross-Modal Graphical Lasso
by: Wang, Fei, et al.
Published: (2026)
by: Wang, Fei, et al.
Published: (2026)
Similar Items
-
Free-Form Scene Editor: Enabling Multi-Round Object Manipulation like in a 3D Engine
by: Shuai, Xincheng, et al.
Published: (2025) -
GlyphPrinter: Region-Grouped Direct Preference Optimization for Glyph-Accurate Visual Text Rendering
by: Shuai, Xincheng, et al.
Published: (2026) -
SceneDesigner: Controllable Multi-Object Image Generation with 9-DoF Pose Manipulation
by: Qin, Zhenyuan, et al.
Published: (2025) -
Free-Form Motion Control: Controlling the 6D Poses of Camera and Objects in Video Generation
by: Shuai, Xincheng, et al.
Published: (2025) -
A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models
by: Shuai, Xincheng, et al.
Published: (2024)