Benchmarking Segmentation Models with Mask-Preserved Attribute Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Yin, Zijin, Liang, Kongming, Li, Bing, Ma, Zhanyu, Guo, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking Semantic Segmentation Models via Appearance and Geometry Attribute Editing
by: Yin, Zijin, et al.
Published: (2026)
by: Yin, Zijin, et al.
Published: (2026)
Polyp-E: Benchmarking the Robustness of Deep Segmentation Models via Polyp Editing
by: Wei, Runpu, et al.
Published: (2024)
by: Wei, Runpu, et al.
Published: (2024)
PGP-SAM: Prototype-Guided Prompt Learning for Efficient Few-Shot Medical Image Segmentation
by: Yan, Zhonghao, et al.
Published: (2025)
by: Yan, Zhonghao, et al.
Published: (2025)
ConMo: Controllable Motion Disentanglement and Recomposition for Zero-Shot Motion Transfer
by: Gao, Jiayi, et al.
Published: (2025)
by: Gao, Jiayi, et al.
Published: (2025)
Generative Visual Chain-of-Thought for Image Editing
by: Yin, Zijin, et al.
Published: (2026)
by: Yin, Zijin, et al.
Published: (2026)
Evaluating Attribute Comprehension in Large Vision-Language Models
by: Zhang, Haiwen, et al.
Published: (2024)
by: Zhang, Haiwen, et al.
Published: (2024)
Efficient Face Super-Resolution via Wavelet-based Feature Enhancement Network
by: Li, Wenjie, et al.
Published: (2024)
by: Li, Wenjie, et al.
Published: (2024)
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation
by: Wang, Xinran, et al.
Published: (2025)
by: Wang, Xinran, et al.
Published: (2025)
Detailed Object Description with Controllable Dimensions
by: Wang, Xinran, et al.
Published: (2024)
by: Wang, Xinran, et al.
Published: (2024)
DriveRX: A Vision-Language Reasoning Model for Cross-Task Autonomous Driving
by: Diao, Muxi, et al.
Published: (2025)
by: Diao, Muxi, et al.
Published: (2025)
OmniEraser: Remove Objects and Their Effects in Images with Paired Video-Frame Data
by: Wei, Runpu, et al.
Published: (2025)
by: Wei, Runpu, et al.
Published: (2025)
Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models
by: Tong, Yujun, et al.
Published: (2026)
by: Tong, Yujun, et al.
Published: (2026)
CineTechBench: A Benchmark for Cinematographic Technique Understanding and Generation
by: Wang, Xinran, et al.
Published: (2025)
by: Wang, Xinran, et al.
Published: (2025)
FairHuman: Boosting Hand and Face Quality in Human Image Generation with Minimum Potential Delay Fairness in Diffusion Models
by: Wang, Yuxuan, et al.
Published: (2025)
by: Wang, Yuxuan, et al.
Published: (2025)
Geometric Image Editing via Effects-Sensitive In-Context Inpainting with Diffusion Transformers
by: Zhang, Shuo, et al.
Published: (2026)
by: Zhang, Shuo, et al.
Published: (2026)
From Simple to Professional: A Combinatorial Controllable Image Captioning Agent
by: Wang, Xinran, et al.
Published: (2024)
by: Wang, Xinran, et al.
Published: (2024)
Controllable-Continuous Color Editing in Diffusion Model via Color Mapping
by: Yang, Yuqi, et al.
Published: (2025)
by: Yang, Yuqi, et al.
Published: (2025)
DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions
by: Wang, Xinran, et al.
Published: (2026)
by: Wang, Xinran, et al.
Published: (2026)
Toward Generalizable Forgery Detection and Reasoning
by: Gao, Yueying, et al.
Published: (2025)
by: Gao, Yueying, et al.
Published: (2025)
Hepato-LLaVA: An Expert MLLM with Sparse Topo-Pack Attention for Hepatocellular Pathology Analysis on Whole Slide Images
by: Yang, Yuxuan, et al.
Published: (2026)
by: Yang, Yuxuan, et al.
Published: (2026)
IncreFA: Breaking the Static Wall of Generative Model Attribution
by: Qin, Haotian, et al.
Published: (2026)
by: Qin, Haotian, et al.
Published: (2026)
SegTalker: Segmentation-based Talking Face Generation with Mask-guided Local Editing
by: Xiong, Lingyu, et al.
Published: (2024)
by: Xiong, Lingyu, et al.
Published: (2024)
Recolour What Matters: Region-Aware Colour Editing via Token-Level Diffusion
by: Yang, Yuqi, et al.
Published: (2026)
by: Yang, Yuqi, et al.
Published: (2026)
TimeMachine: Fine-Grained Facial Age Editing with Identity Preservation
by: Mi, Yilin, et al.
Published: (2025)
by: Mi, Yilin, et al.
Published: (2025)
Rethinking Video Segmentation with Masked Video Consistency: Did the Model Learn as Intended?
by: Liang, Chen, et al.
Published: (2024)
by: Liang, Chen, et al.
Published: (2024)
Pedestrian Attribute Editing for Gait Recognition and Anonymization
by: Ma, Jingzhe, et al.
Published: (2023)
by: Ma, Jingzhe, et al.
Published: (2023)
SEED: A Benchmark Dataset for Sequential Facial Attribute Editing with Diffusion Models
by: Zhu, Yule, et al.
Published: (2025)
by: Zhu, Yule, et al.
Published: (2025)
MAKIMA: Tuning-free Multi-Attribute Open-domain Video Editing via Mask-Guided Attention Modulation
by: Zheng, Haoyu, et al.
Published: (2024)
by: Zheng, Haoyu, et al.
Published: (2024)
Self-Supervised Selective-Guided Diffusion Model for Old-Photo Face Restoration
by: Li, Wenjie, et al.
Published: (2025)
by: Li, Wenjie, et al.
Published: (2025)
FourierSR: A Fourier Token-based Plugin for Efficient Image Super-Resolution
by: Li, Wenjie, et al.
Published: (2025)
by: Li, Wenjie, et al.
Published: (2025)
SpecGen: Neural Spectral BRDF Generation via Spectral-Spatial Tri-plane Aggregation
by: Jin, Zhenyu, et al.
Published: (2025)
by: Jin, Zhenyu, et al.
Published: (2025)
Masked Representation Modeling for Domain-Adaptive Segmentation
by: Zhou, Wenlve, et al.
Published: (2025)
by: Zhou, Wenlve, et al.
Published: (2025)
Towards Privacy-Preserving Fine-Grained Visual Classification via Hierarchical Learning from Label Proportions
by: Chang, Jinyi, et al.
Published: (2025)
by: Chang, Jinyi, et al.
Published: (2025)
MaskINT: Video Editing via Interpolative Non-autoregressive Masked Transformers
by: Ma, Haoyu, et al.
Published: (2023)
by: Ma, Haoyu, et al.
Published: (2023)
SAMRefiner: Taming Segment Anything Model for Universal Mask Refinement
by: Lin, Yuqi, et al.
Published: (2025)
by: Lin, Yuqi, et al.
Published: (2025)
Semantic Segmentation on VSPW Dataset through Masked Video Consistency
by: Liang, Chen, et al.
Published: (2024)
by: Liang, Chen, et al.
Published: (2024)
From Pixel to Mask: A Survey of Out-of-Distribution Segmentation
by: Zhao, Wenjie, et al.
Published: (2025)
by: Zhao, Wenjie, et al.
Published: (2025)
PolarVLM: Bridging the Semantic-Physical Gap in Vision-Language Models
by: Li, Yuliang, et al.
Published: (2026)
by: Li, Yuliang, et al.
Published: (2026)
Fast and Efficient: Mask Neural Fields for 3D Scene Segmentation
by: Gao, Zihan, et al.
Published: (2024)
by: Gao, Zihan, et al.
Published: (2024)
BA-SAM: Scalable Bias-Mode Attention Mask for Segment Anything Model
by: Song, Yiran, et al.
Published: (2024)
by: Song, Yiran, et al.
Published: (2024)
Similar Items
-
Benchmarking Semantic Segmentation Models via Appearance and Geometry Attribute Editing
by: Yin, Zijin, et al.
Published: (2026) -
Polyp-E: Benchmarking the Robustness of Deep Segmentation Models via Polyp Editing
by: Wei, Runpu, et al.
Published: (2024) -
PGP-SAM: Prototype-Guided Prompt Learning for Efficient Few-Shot Medical Image Segmentation
by: Yan, Zhonghao, et al.
Published: (2025) -
ConMo: Controllable Motion Disentanglement and Recomposition for Zero-Shot Motion Transfer
by: Gao, Jiayi, et al.
Published: (2025) -
Generative Visual Chain-of-Thought for Image Editing
by: Yin, Zijin, et al.
Published: (2026)