Saved in:
| Main Authors: | Bai, Chen, Shao, Zeman, Zhang, Guoxiang, Liang, Di, Yang, Jie, Zhang, Zhuorui, Guo, Yujian, Zhong, Chengzhang, Qiu, Yiqiao, Wang, Zhendong, Guan, Yichen, Zheng, Xiaoyin, Wang, Tao, Lu, Cheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2401.17509 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Photorealistic Object Insertion with Diffusion-Guided Inverse Rendering
by: Liang, Ruofan, et al.
Published: (2024)
by: Liang, Ruofan, et al.
Published: (2024)
ObjectDrop: Bootstrapping Counterfactuals for Photorealistic Object Removal and Insertion
by: Winter, Daniel, et al.
Published: (2024)
by: Winter, Daniel, et al.
Published: (2024)
Controllable Video Object Insertion via Multiview Priors
by: Qi, Xia, et al.
Published: (2026)
by: Qi, Xia, et al.
Published: (2026)
NavigScene: Bridging Local Perception and Global Navigation for Beyond-Visual-Range Autonomous Driving
by: Peng, Qucheng, et al.
Published: (2025)
by: Peng, Qucheng, et al.
Published: (2025)
Place Anything into Any Video
by: Liu, Ziling, et al.
Published: (2024)
by: Liu, Ziling, et al.
Published: (2024)
Depth Anything with Any Prior
by: Wang, Zehan, et al.
Published: (2025)
by: Wang, Zehan, et al.
Published: (2025)
Teleportraits: Training-Free People Insertion into Any Scene
by: Gao, Jialu, et al.
Published: (2025)
by: Gao, Jialu, et al.
Published: (2025)
Material Anything: Generating Materials for Any 3D Object via Diffusion
by: Huang, Xin, et al.
Published: (2024)
by: Huang, Xin, et al.
Published: (2024)
Drive-JEPA: Video JEPA Meets Multimodal Trajectory Distillation for End-to-End Driving
by: Wang, Linhan, et al.
Published: (2026)
by: Wang, Linhan, et al.
Published: (2026)
Tracking and Segmenting Anything in Any Modality
by: Zhang, Tianlu, et al.
Published: (2025)
by: Zhang, Tianlu, et al.
Published: (2025)
AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation
by: Wei, Huawei, et al.
Published: (2024)
by: Wei, Huawei, et al.
Published: (2024)
Structurally Prune Anything: Any Architecture, Any Framework, Any Time
by: Wang, Xun, et al.
Published: (2024)
by: Wang, Xun, et al.
Published: (2024)
Motion Anything: Any to Motion Generation
by: Zhang, Zeyu, et al.
Published: (2025)
by: Zhang, Zeyu, et al.
Published: (2025)
RealMaster: Lifting Rendered Scenes into Photorealistic Video
by: Cohen-Bar, Dana, et al.
Published: (2026)
by: Cohen-Bar, Dana, et al.
Published: (2026)
SAM 2++: Tracking Anything at Any Granularity
by: Zhang, Jiaming, et al.
Published: (2025)
by: Zhang, Jiaming, et al.
Published: (2025)
Mirage: One-Step Video Diffusion for Photorealistic and Coherent Asset Editing in Driving Scenes
by: Wang, Shuyun, et al.
Published: (2025)
by: Wang, Shuyun, et al.
Published: (2025)
OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models
by: Chen, Jinshu, et al.
Published: (2025)
by: Chen, Jinshu, et al.
Published: (2025)
DetAny4D: Detect Anything 4D Temporally in a Streaming RGB Video
by: Hou, Jiawei, et al.
Published: (2025)
by: Hou, Jiawei, et al.
Published: (2025)
X-SAM: From Segment Anything to Any Segmentation
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
Generative Object Insertion in Gaussian Splatting with a Multi-View Diffusion Model
by: Zhong, Hongliang, et al.
Published: (2024)
by: Zhong, Hongliang, et al.
Published: (2024)
Judge Anything: MLLM as a Judge Across Any Modality
by: Pu, Shu, et al.
Published: (2025)
by: Pu, Shu, et al.
Published: (2025)
Food Portion Estimation via 3D Object Scaling
by: Vinod, Gautham, et al.
Published: (2024)
by: Vinod, Gautham, et al.
Published: (2024)
VideoAnydoor: High-fidelity Video Object Insertion with Precise Motion Control
by: Tu, Yuanpeng, et al.
Published: (2025)
by: Tu, Yuanpeng, et al.
Published: (2025)
SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion
by: Chen, Xinyu, et al.
Published: (2026)
by: Chen, Xinyu, et al.
Published: (2026)
$NavA^3$: Understanding Any Instruction, Navigating Anywhere, Finding Anything
by: Zhang, Lingfeng, et al.
Published: (2025)
by: Zhang, Lingfeng, et al.
Published: (2025)
Count Anything at Any Granularity
by: Liu, Chang, et al.
Published: (2026)
by: Liu, Chang, et al.
Published: (2026)
Say Anything with Any Style
by: Tan, Shuai, et al.
Published: (2024)
by: Tan, Shuai, et al.
Published: (2024)
Depth Anything at Any Condition
by: Sun, Boyuan, et al.
Published: (2025)
by: Sun, Boyuan, et al.
Published: (2025)
Relighting Scenes with Object Insertions in Neural Radiance Fields
by: Zhu, Xuening, et al.
Published: (2024)
by: Zhu, Xuening, et al.
Published: (2024)
Compress Any Segment Anything Model (SAM)
by: Fan, Juntong, et al.
Published: (2025)
by: Fan, Juntong, et al.
Published: (2025)
Segment Anything for Video: A Comprehensive Review of Video Object Segmentation and Tracking from Past to Future
by: Xu, Guoping, et al.
Published: (2025)
by: Xu, Guoping, et al.
Published: (2025)
Any to Full: Prompting Depth Anything for Depth Completion in One Stage
by: Zhou, Zhiyuan, et al.
Published: (2026)
by: Zhou, Zhiyuan, et al.
Published: (2026)
ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model
by: Liu, Fangfu, et al.
Published: (2024)
by: Liu, Fangfu, et al.
Published: (2024)
4Real: Towards Photorealistic 4D Scene Generation via Video Diffusion Models
by: Yu, Heng, et al.
Published: (2024)
by: Yu, Heng, et al.
Published: (2024)
CustAny: Customizing Anything from A Single Example
by: Kong, Lingjie, et al.
Published: (2024)
by: Kong, Lingjie, et al.
Published: (2024)
Detect in Any Scene: An Agentic Framework for Object Detection with Experience-Aware Reasoning
by: Zhang, Wenlun, et al.
Published: (2026)
by: Zhang, Wenlun, et al.
Published: (2026)
InteractAnything: Zero-shot Human Object Interaction Synthesis via LLM Feedback and Object Affordance Parsing
by: Zhang, Jinlu, et al.
Published: (2025)
by: Zhang, Jinlu, et al.
Published: (2025)
Object Placement for Anything
by: Gao, Bingjie, et al.
Published: (2025)
by: Gao, Bingjie, et al.
Published: (2025)
SNAP: Towards Segmenting Anything in Any Point Cloud
by: Gupta, Aniket, et al.
Published: (2025)
by: Gupta, Aniket, et al.
Published: (2025)
MDReID: Modality-Decoupled Learning for Any-to-Any Multi-Modal Object Re-Identification
by: Feng, Yingying, et al.
Published: (2025)
by: Feng, Yingying, et al.
Published: (2025)
Similar Items
-
Photorealistic Object Insertion with Diffusion-Guided Inverse Rendering
by: Liang, Ruofan, et al.
Published: (2024) -
ObjectDrop: Bootstrapping Counterfactuals for Photorealistic Object Removal and Insertion
by: Winter, Daniel, et al.
Published: (2024) -
Controllable Video Object Insertion via Multiview Priors
by: Qi, Xia, et al.
Published: (2026) -
NavigScene: Bridging Local Perception and Global Navigation for Beyond-Visual-Range Autonomous Driving
by: Peng, Qucheng, et al.
Published: (2025) -
Place Anything into Any Video
by: Liu, Ziling, et al.
Published: (2024)