Saved in:
| Main Authors: | Zeng, Bohan, Li, Shanglin, Feng, Yutang, Yang, Ling, Li, Hong, Gao, Sicheng, Liu, Jiaming, He, Conghui, Zhang, Wentao, Liu, Jianzhuang, Zhang, Baochang, Yan, Shuicheng |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2310.05375 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ZONE: Zero-Shot Instruction-Guided Local Editing
by: Li, Shanglin, et al.
Published: (2023)
by: Li, Shanglin, et al.
Published: (2023)
EditWorld: Simulating World Dynamics for Instruction-Following Image Editing
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
Dual Diffusion Models for Multi-modal Guided 3D Avatar Generation
by: Li, Hong, et al.
Published: (2026)
by: Li, Hong, et al.
Published: (2026)
Any2AnyTryon: Leveraging Adaptive Position Embeddings for Versatile Virtual Clothing Tasks
by: Guo, Hailong, et al.
Published: (2025)
by: Guo, Hailong, et al.
Published: (2025)
WideRange4D: Enabling High-Quality 4D Reconstruction with Wide-Range Movements and Scenes
by: Yang, Ling, et al.
Published: (2025)
by: Yang, Ling, et al.
Published: (2025)
Motion Manipulation via Unsupervised Keypoint Positioning in Face Animation
by: Li, Hong, et al.
Published: (2026)
by: Li, Hong, et al.
Published: (2026)
MagicEraser: Erasing Any Objects via Semantics-Aware Control
by: Li, Fan, et al.
Published: (2024)
by: Li, Fan, et al.
Published: (2024)
EA‐YOLO: An Efficient and Accurate UAV Image Object Detection Algorithm
by: Dehao Dong, et al.
Published: (2024)
by: Dehao Dong, et al.
Published: (2024)
Decoupling Appearance Variations with 3D Consistent Features in Gaussian Splatting
by: Lin, Jiaqi, et al.
Published: (2025)
by: Lin, Jiaqi, et al.
Published: (2025)
Ctrl-X: Controlling Structure and Appearance for Text-To-Image Generation Without Guidance
by: Lin, Kuan Heng, et al.
Published: (2024)
by: Lin, Kuan Heng, et al.
Published: (2024)
Learning Background Prompts to Discover Implicit Knowledge for Open Vocabulary Object Detection
by: Li, Jiaming, et al.
Published: (2024)
by: Li, Jiaming, et al.
Published: (2024)
SynthVLM: Towards High-Quality and Efficient Synthesis of Image-Caption Datasets for Vision-Language Models
by: Liu, Zheng, et al.
Published: (2024)
by: Liu, Zheng, et al.
Published: (2024)
Semantic Score Distillation Sampling for Compositional Text-to-3D Generation
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
AttriCLIP: A Non-Incremental Learner for Incremental Knowledge Learning
by: Wang, Runqi, et al.
Published: (2023)
by: Wang, Runqi, et al.
Published: (2023)
Scale Contrastive Learning with Selective Attentions for Blind Image Quality Assessment
by: Hu, Runze, et al.
Published: (2024)
by: Hu, Runze, et al.
Published: (2024)
Data Augmentation in Human-Centric Vision
by: Jiang, Wentao, et al.
Published: (2024)
by: Jiang, Wentao, et al.
Published: (2024)
DiffuX2CT: Diffusion Learning to Reconstruct CT Images from Biplanar X-Rays
by: Liu, Xuhui, et al.
Published: (2024)
by: Liu, Xuhui, et al.
Published: (2024)
Trans4D: Realistic Geometry-Aware Transition for Compositional Text-to-4D Synthesis
by: Zeng, Bohan, et al.
Published: (2024)
by: Zeng, Bohan, et al.
Published: (2024)
T-Rex2: Towards Generic Object Detection via Text-Visual Prompt Synergy
by: Jiang, Qing, et al.
Published: (2024)
by: Jiang, Qing, et al.
Published: (2024)
Uncertainty-Aware Gradient Stabilization for Small Object Detection
by: Sun, Huixin, et al.
Published: (2023)
by: Sun, Huixin, et al.
Published: (2023)
Causally Guided Gaussian Perturbations for Out-Of-Distribution Generalization in Medical Imaging
by: Pei, Haoran, et al.
Published: (2025)
by: Pei, Haoran, et al.
Published: (2025)
Learning from Mistakes: Iterative Prompt Relabeling for Text-to-Image Diffusion Model Training
by: Chen, Xinyan, et al.
Published: (2023)
by: Chen, Xinyan, et al.
Published: (2023)
Chapter 13 Pulmonary Diseases and Traditional Chinese Medicine
by: Zhao, Baochang, et al.
Published: (2024)
by: Zhao, Baochang, et al.
Published: (2024)
CamoSAM2: Motion-Appearance Induced Auto-Refining Prompts for Video Camouflaged Object Detection
by: Zhang, Xin, et al.
Published: (2025)
by: Zhang, Xin, et al.
Published: (2025)
Why Training-Free Token Reduction Collapses: The Inherent Instability of Pairwise Scoring Signals
by: Shanglin, Yang
Published: (2026)
by: Shanglin, Yang
Published: (2026)
PAM: A Pose-Appearance-Motion Engine for Sim-to-Real HOI Video Generation
by: Gao, Mingju, et al.
Published: (2026)
by: Gao, Mingju, et al.
Published: (2026)
HumanEdit: A High-Quality Human-Rewarded Dataset for Instruction-based Image Editing
by: Bai, Jinbin, et al.
Published: (2024)
by: Bai, Jinbin, et al.
Published: (2024)
Direction-Aware Hybrid Representation Learning for 3D Hand Pose and Shape Estimation
by: Liu, Shiyong, et al.
Published: (2025)
by: Liu, Shiyong, et al.
Published: (2025)
Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling
by: Wang, Yuran, et al.
Published: (2025)
by: Wang, Yuran, et al.
Published: (2025)
Multi-Step Visual Reasoning with Visual Tokens Scaling and Verification
by: Bai, Tianyi, et al.
Published: (2025)
by: Bai, Tianyi, et al.
Published: (2025)
La vida en China / Lin Yutang ; ilustraciones de Howard Simon ; traducción de Carlos Villegas
by: Yutang, Lin
Published: (1972)
by: Yutang, Lin
Published: (1972)
Confucianism and Democratic Constitutionalism in East Asia: Evaluating Confucian Democratic Perfectionism
by: Yutang Jin
Published: (2025)
by: Yutang Jin
Published: (2025)
Zero-P-to-3: Zero-Shot Partial-View Images to 3D Object
by: Lin, Yuxuan, et al.
Published: (2025)
by: Lin, Yuxuan, et al.
Published: (2025)
ControlLight: Towards Controllable, Consistent, and Generalizable Low-Light Enhancement
by: Yang, Yufeng, et al.
Published: (2026)
by: Yang, Yufeng, et al.
Published: (2026)
SyncMV4D: Synchronized Multi-view Joint Diffusion of Appearance and Motion for Hand-Object Interaction Synthesis
by: Dang, Lingwei, et al.
Published: (2025)
by: Dang, Lingwei, et al.
Published: (2025)
MinerU-Diffusion: Rethinking Document OCR as Inverse Rendering via Diffusion Decoding
by: Dong, Hejun, et al.
Published: (2026)
by: Dong, Hejun, et al.
Published: (2026)
Enhancing Object Discovery for Unsupervised Instance Segmentation and Object Detection
by: Feng, Xingyu, et al.
Published: (2025)
by: Feng, Xingyu, et al.
Published: (2025)
TAPTRv2: Attention-based Position Update Improves Tracking Any Point
by: Li, Hongyang, et al.
Published: (2024)
by: Li, Hongyang, et al.
Published: (2024)
Geometry-Editable and Appearance-Preserving Object Compositon
by: Lin, Jianman, et al.
Published: (2025)
by: Lin, Jianman, et al.
Published: (2025)
CrayonRobo: Object-Centric Prompt-Driven Vision-Language-Action Model for Robotic Manipulation
by: Li, Xiaoqi, et al.
Published: (2025)
by: Li, Xiaoqi, et al.
Published: (2025)
Similar Items
-
ZONE: Zero-Shot Instruction-Guided Local Editing
by: Li, Shanglin, et al.
Published: (2023) -
EditWorld: Simulating World Dynamics for Instruction-Following Image Editing
by: Yang, Ling, et al.
Published: (2024) -
Dual Diffusion Models for Multi-modal Guided 3D Avatar Generation
by: Li, Hong, et al.
Published: (2026) -
Any2AnyTryon: Leveraging Adaptive Position Embeddings for Versatile Virtual Clothing Tasks
by: Guo, Hailong, et al.
Published: (2025) -
WideRange4D: Enabling High-Quality 4D Reconstruction with Wide-Range Movements and Scenes
by: Yang, Ling, et al.
Published: (2025)