OmniEraser: Remove Objects and Their Effects in Images with Paired Video-Frame Data
Fuente:
arXiv
Salvato in:
| Autori principali: | Wei, Runpu, Yin, Zijin, Zhang, Shuo, Zhou, Lanxiang, Wang, Xueyi, Ban, Chao, Cao, Tianwei, Sun, Hao, He, Zhongjiang, Liang, Kongming, Ma, Zhanyu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FairHuman: Boosting Hand and Face Quality in Human Image Generation with Minimum Potential Delay Fairness in Diffusion Models
di: Wang, Yuxuan, et al.
Pubblicazione: (2025)
di: Wang, Yuxuan, et al.
Pubblicazione: (2025)
Geometric Image Editing via Effects-Sensitive In-Context Inpainting with Diffusion Transformers
di: Zhang, Shuo, et al.
Pubblicazione: (2026)
di: Zhang, Shuo, et al.
Pubblicazione: (2026)
Benchmarking Semantic Segmentation Models via Appearance and Geometry Attribute Editing
di: Yin, Zijin, et al.
Pubblicazione: (2026)
di: Yin, Zijin, et al.
Pubblicazione: (2026)
Polyp-E: Benchmarking the Robustness of Deep Segmentation Models via Polyp Editing
di: Wei, Runpu, et al.
Pubblicazione: (2024)
di: Wei, Runpu, et al.
Pubblicazione: (2024)
Detailed Object Description with Controllable Dimensions
di: Wang, Xinran, et al.
Pubblicazione: (2024)
di: Wang, Xinran, et al.
Pubblicazione: (2024)
Curriculum Group Policy Optimization: Adaptive Sampling for Unleashing the Potential of Text-to-Image Generation
di: Li, Baoteng, et al.
Pubblicazione: (2026)
di: Li, Baoteng, et al.
Pubblicazione: (2026)
Benchmarking Segmentation Models with Mask-Preserved Attribute Editing
di: Yin, Zijin, et al.
Pubblicazione: (2024)
di: Yin, Zijin, et al.
Pubblicazione: (2024)
PGP-SAM: Prototype-Guided Prompt Learning for Efficient Few-Shot Medical Image Segmentation
di: Yan, Zhonghao, et al.
Pubblicazione: (2025)
di: Yan, Zhonghao, et al.
Pubblicazione: (2025)
ProTA: Probabilistic Token Aggregation for Text-Video Retrieval
di: Fang, Han, et al.
Pubblicazione: (2024)
di: Fang, Han, et al.
Pubblicazione: (2024)
ConMo: Controllable Motion Disentanglement and Recomposition for Zero-Shot Motion Transfer
di: Gao, Jiayi, et al.
Pubblicazione: (2025)
di: Gao, Jiayi, et al.
Pubblicazione: (2025)
DataEvolver: Let Your Data Build and Improve Itself via Goal-Driven Loop Agents
di: Zhang, Qisong, et al.
Pubblicazione: (2026)
di: Zhang, Qisong, et al.
Pubblicazione: (2026)
Generative Visual Chain-of-Thought for Image Editing
di: Yin, Zijin, et al.
Pubblicazione: (2026)
di: Yin, Zijin, et al.
Pubblicazione: (2026)
AdaEraser: Training-Free Object Removal via Adaptive Attention Suppression
di: Liu, Dingming
Pubblicazione: (2026)
di: Liu, Dingming
Pubblicazione: (2026)
GenEraser: Generalizable Video Object Removal via Balanced Text-Mask Guidance and Decoupled Locator-Preserver
di: Chen, Yuqing, et al.
Pubblicazione: (2026)
di: Chen, Yuqing, et al.
Pubblicazione: (2026)
Disentangle and denoise: Tackling context misalignment for video moment retrieval
di: Ma, Kaijing, et al.
Pubblicazione: (2024)
di: Ma, Kaijing, et al.
Pubblicazione: (2024)
From Simple to Professional: A Combinatorial Controllable Image Captioning Agent
di: Wang, Xinran, et al.
Pubblicazione: (2024)
di: Wang, Xinran, et al.
Pubblicazione: (2024)
SmartEraser: Remove Anything from Images using Masked-Region Guidance
di: Jiang, Longtao, et al.
Pubblicazione: (2025)
di: Jiang, Longtao, et al.
Pubblicazione: (2025)
Attentive Eraser: Unleashing Diffusion Model's Object Removal Potential via Self-Attention Redirection Guidance
di: Sun, Wenhao, et al.
Pubblicazione: (2024)
di: Sun, Wenhao, et al.
Pubblicazione: (2024)
DeepEraser: Deep Iterative Context Mining for Generic Text Eraser
di: Feng, Hao, et al.
Pubblicazione: (2024)
di: Feng, Hao, et al.
Pubblicazione: (2024)
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation
di: Wang, Xinran, et al.
Pubblicazione: (2025)
di: Wang, Xinran, et al.
Pubblicazione: (2025)
VideoEraser: Concept Erasure in Text-to-Video Diffusion Models
di: Xu, Naen, et al.
Pubblicazione: (2025)
di: Xu, Naen, et al.
Pubblicazione: (2025)
Evaluating Attribute Comprehension in Large Vision-Language Models
di: Zhang, Haiwen, et al.
Pubblicazione: (2024)
di: Zhang, Haiwen, et al.
Pubblicazione: (2024)
DiffuEraser: A Diffusion Model for Video Inpainting
di: Li, Xiaowen, et al.
Pubblicazione: (2025)
di: Li, Xiaowen, et al.
Pubblicazione: (2025)
ErrorEraser: Unlearning Data Bias for Improved Continual Learning
di: Cao, Xuemei, et al.
Pubblicazione: (2025)
di: Cao, Xuemei, et al.
Pubblicazione: (2025)
Inference-Time Rule Eraser: Fair Recognition via Distilling and Removing Biased Rules
di: Zhang, Yi, et al.
Pubblicazione: (2024)
di: Zhang, Yi, et al.
Pubblicazione: (2024)
DriveRX: A Vision-Language Reasoning Model for Cross-Task Autonomous Driving
di: Diao, Muxi, et al.
Pubblicazione: (2025)
di: Diao, Muxi, et al.
Pubblicazione: (2025)
EraserDiT: Fast Video Inpainting with Diffusion Transformer Model
di: Liu, Jie, et al.
Pubblicazione: (2025)
di: Liu, Jie, et al.
Pubblicazione: (2025)
OmniSTVG: Toward Spatio-Temporal Omni-Object Video Grounding
di: Yao, Jiali, et al.
Pubblicazione: (2025)
di: Yao, Jiali, et al.
Pubblicazione: (2025)
MagicEraser: Erasing Any Objects via Semantics-Aware Control
di: Li, Fan, et al.
Pubblicazione: (2024)
di: Li, Fan, et al.
Pubblicazione: (2024)
Efficient Face Super-Resolution via Wavelet-based Feature Enhancement Network
di: Li, Wenjie, et al.
Pubblicazione: (2024)
di: Li, Wenjie, et al.
Pubblicazione: (2024)
Omni-Router: Sharing Routing Decisions in Sparse Mixture-of-Experts for Speech Recognition
di: Gu, Zijin, et al.
Pubblicazione: (2025)
di: Gu, Zijin, et al.
Pubblicazione: (2025)
Hepato-LLaVA: An Expert MLLM with Sparse Topo-Pack Attention for Hepatocellular Pathology Analysis on Whole Slide Images
di: Yang, Yuxuan, et al.
Pubblicazione: (2026)
di: Yang, Yuxuan, et al.
Pubblicazione: (2026)
Ternary Quantum Eraser Cryptography
di: Halawani, Ahmed, et al.
Pubblicazione: (2026)
di: Halawani, Ahmed, et al.
Pubblicazione: (2026)
Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models
di: Tong, Yujun, et al.
Pubblicazione: (2026)
di: Tong, Yujun, et al.
Pubblicazione: (2026)
Trusted Unified Feature-Neighborhood Dynamics for Multi-View Classification
di: Huang, Haojian, et al.
Pubblicazione: (2024)
di: Huang, Haojian, et al.
Pubblicazione: (2024)
Combinatorial Laplacians and Relative Homology of Complex Pairs
di: Zhan, Xiongfeng, et al.
Pubblicazione: (2025)
di: Zhan, Xiongfeng, et al.
Pubblicazione: (2025)
DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions
di: Wang, Xinran, et al.
Pubblicazione: (2026)
di: Wang, Xinran, et al.
Pubblicazione: (2026)
AutoRed: A Free-form Adversarial Prompt Generation Framework for Automated Red Teaming
di: Diao, Muxi, et al.
Pubblicazione: (2025)
di: Diao, Muxi, et al.
Pubblicazione: (2025)
OmniPaint: Mastering Object-Oriented Editing via Disentangled Insertion-Removal Inpainting
di: Yu, Yongsheng, et al.
Pubblicazione: (2025)
di: Yu, Yongsheng, et al.
Pubblicazione: (2025)
OmniHumanoid: Streaming Cross-Embodiment Video Generation with Paired-Free Adaptation
di: Song, Yiren, et al.
Pubblicazione: (2026)
di: Song, Yiren, et al.
Pubblicazione: (2026)
Documenti analoghi
-
FairHuman: Boosting Hand and Face Quality in Human Image Generation with Minimum Potential Delay Fairness in Diffusion Models
di: Wang, Yuxuan, et al.
Pubblicazione: (2025) -
Geometric Image Editing via Effects-Sensitive In-Context Inpainting with Diffusion Transformers
di: Zhang, Shuo, et al.
Pubblicazione: (2026) -
Benchmarking Semantic Segmentation Models via Appearance and Geometry Attribute Editing
di: Yin, Zijin, et al.
Pubblicazione: (2026) -
Polyp-E: Benchmarking the Robustness of Deep Segmentation Models via Polyp Editing
di: Wei, Runpu, et al.
Pubblicazione: (2024) -
Detailed Object Description with Controllable Dimensions
di: Wang, Xinran, et al.
Pubblicazione: (2024)