UniSER: A Foundation Model for Unified Soft Effects Removal
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Jingdong, Zhang, Lingzhi, Liu, Qing, Chiu, Mang Tik, Barnes, Connelly, Wang, Yizhou, You, Haoran, Liu, Xiaoyang, Zhou, Yuqian, Lin, Zhe, Shechtman, Eli, Amirghodsi, Sohrab, Li, Xin, Wang, Wenping, Zhan, Xiaohang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fine-grained Defocus Blur Control for Generative Image Models
by: Shrivastava, Ayush, et al.
Published: (2025)
by: Shrivastava, Ayush, et al.
Published: (2025)
Layer- and Timestep-Adaptive Differentiable Token Compression Ratios for Efficient Diffusion Transformers
by: You, Haoran, et al.
Published: (2024)
by: You, Haoran, et al.
Published: (2024)
Cautious Next Token Prediction
by: Wang, Yizhou, et al.
Published: (2025)
by: Wang, Yizhou, et al.
Published: (2025)
Structure-Guided Image Completion with Image-level and Object-level Semantic Discriminators
by: Zheng, Haitian, et al.
Published: (2022)
by: Zheng, Haitian, et al.
Published: (2022)
Inline Critic Steers Image Editing
by: Kang, Weitai, et al.
Published: (2026)
by: Kang, Weitai, et al.
Published: (2026)
MTPano: Multi-Task Panoramic Scene Understanding via Label-Free Integration of Dense Prediction Priors
by: Zhang, Jingdong, et al.
Published: (2026)
by: Zhang, Jingdong, et al.
Published: (2026)
Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search
by: Zhang, Jingdong, et al.
Published: (2026)
by: Zhang, Jingdong, et al.
Published: (2026)
Distilling Diffusion Models into Conditional GANs
by: Kang, Minguk, et al.
Published: (2024)
by: Kang, Minguk, et al.
Published: (2024)
3R-GS: Best Practice in Optimizing Camera Poses Along with 3DGS
by: Huang, Zhisheng, et al.
Published: (2025)
by: Huang, Zhisheng, et al.
Published: (2025)
Jump Cut Smoothing for Talking Heads
by: Wang, Xiaojuan, et al.
Published: (2024)
by: Wang, Xiaojuan, et al.
Published: (2024)
Removing Distributional Discrepancies in Captions Improves Image-Text Alignment
by: Li, Yuheng, et al.
Published: (2024)
by: Li, Yuheng, et al.
Published: (2024)
Image Neural Field Diffusion Models
by: Chen, Yinbo, et al.
Published: (2024)
by: Chen, Yinbo, et al.
Published: (2024)
Multi-Task Label Discovery via Hierarchical Task Tokens for Partially Annotated Dense Predictions
by: Zhang, Jingdong, et al.
Published: (2024)
by: Zhang, Jingdong, et al.
Published: (2024)
ZipIR: Latent Pyramid Diffusion Transformer for High-Resolution Image Restoration
by: Yu, Yongsheng, et al.
Published: (2025)
by: Yu, Yongsheng, et al.
Published: (2025)
UniLION: Towards Unified Autonomous Driving Model with Linear Group RNNs
by: Liu, Zhe, et al.
Published: (2025)
by: Liu, Zhe, et al.
Published: (2025)
TurboEdit: Instant text-based image editing
by: Wu, Zongze, et al.
Published: (2024)
by: Wu, Zongze, et al.
Published: (2024)
UniAlignment: Semantic Alignment for Unified Image Generation, Understanding, Manipulation and Perception
by: Song, Xinyang, et al.
Published: (2025)
by: Song, Xinyang, et al.
Published: (2025)
A Novel Robot Hand with Hoeckens Linkages and Soft Phalanges for Scooping and Self-Adaptive Grasping in Environmental Constraints
by: Guo, Wentao, et al.
Published: (2025)
by: Guo, Wentao, et al.
Published: (2025)
Reprogrammable Multi‐Responsiveness of Regenerated Silk for Versatile Soft Actuators
by: Jianliang Xiao, et al.
Published: (2024)
by: Jianliang Xiao, et al.
Published: (2024)
Remove Symmetries to Control Model Expressivity and Improve Optimization
by: Ziyin, Liu, et al.
Published: (2024)
by: Ziyin, Liu, et al.
Published: (2024)
Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks
by: Zhang, Jiazhao, et al.
Published: (2024)
by: Zhang, Jiazhao, et al.
Published: (2024)
UniChange: Unifying Change Detection with Multimodal Large Language Model
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
Identifying Prompted Artist Names from Generated Images
by: Su, Grace, et al.
Published: (2025)
by: Su, Grace, et al.
Published: (2025)
Editable Image Elements for Controllable Synthesis
by: Mu, Jiteng, et al.
Published: (2024)
by: Mu, Jiteng, et al.
Published: (2024)
Soft Coulomb Gap Limits the Performance of Organic Thermoelectrics
by: Liu, Yuqian, et al.
Published: (2025)
by: Liu, Yuqian, et al.
Published: (2025)
NewMove: Customizing text-to-video models with novel motions
by: Materzynska, Joanna, et al.
Published: (2023)
by: Materzynska, Joanna, et al.
Published: (2023)
SliderSpace: Decomposing the Visual Capabilities of Diffusion Models
by: Gandikota, Rohit, et al.
Published: (2025)
by: Gandikota, Rohit, et al.
Published: (2025)
Improved Baselines with Representation Autoencoders
by: Singh, Jaskirat, et al.
Published: (2026)
by: Singh, Jaskirat, et al.
Published: (2026)
SoftShadow: Leveraging Soft Masks for Penumbra-Aware Shadow Removal
by: Wang, Xinrui, et al.
Published: (2024)
by: Wang, Xinrui, et al.
Published: (2024)
VideoGigaGAN: Towards Detail-rich Video Super-Resolution
by: Xu, Yiran, et al.
Published: (2024)
by: Xu, Yiran, et al.
Published: (2024)
Connellys' Classroom Cutaway
by: Connelly, John, et al.
Published: (2008)
by: Connelly, John, et al.
Published: (2008)
SPGen: Spherical Projection as Consistent and Flexible Representation for Single Image 3D Shape Generation
by: Zhang, Jingdong, et al.
Published: (2025)
by: Zhang, Jingdong, et al.
Published: (2025)
Dissolution of Primary Carbides and Formation and Healing of Kirkendall Voids in Bearing Steel under Pulsed Electric Current
by: Zhongxue Wang, et al.
Published: (2024)
by: Zhongxue Wang, et al.
Published: (2024)
From Slow Bidirectional to Fast Autoregressive Video Diffusion Models
by: Yin, Tianwei, et al.
Published: (2024)
by: Yin, Tianwei, et al.
Published: (2024)
Detecting Human Artifacts from Text-to-Image Models
by: Wang, Kaihong, et al.
Published: (2024)
by: Wang, Kaihong, et al.
Published: (2024)
Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion
by: Huang, Xun, et al.
Published: (2025)
by: Huang, Xun, et al.
Published: (2025)
RIFLE: Removal of Image Flicker-Banding via Latent Diffusion Enhancement
by: Zhu, Libo, et al.
Published: (2025)
by: Zhu, Libo, et al.
Published: (2025)
Magic Fixup: Streamlining Photo Editing by Watching Dynamic Videos
by: Alzayer, Hadi, et al.
Published: (2024)
by: Alzayer, Hadi, et al.
Published: (2024)
Customizing Text-to-Image Diffusion with Object Viewpoint Control
by: Kumari, Nupur, et al.
Published: (2024)
by: Kumari, Nupur, et al.
Published: (2024)
Spiderwebs on the Sphere and an Isoperimetric Theorem
by: Connelly, Robert, et al.
Published: (2025)
by: Connelly, Robert, et al.
Published: (2025)
Similar Items
-
Fine-grained Defocus Blur Control for Generative Image Models
by: Shrivastava, Ayush, et al.
Published: (2025) -
Layer- and Timestep-Adaptive Differentiable Token Compression Ratios for Efficient Diffusion Transformers
by: You, Haoran, et al.
Published: (2024) -
Cautious Next Token Prediction
by: Wang, Yizhou, et al.
Published: (2025) -
Structure-Guided Image Completion with Image-level and Object-level Semantic Discriminators
by: Zheng, Haitian, et al.
Published: (2022) -
Inline Critic Steers Image Editing
by: Kang, Weitai, et al.
Published: (2026)