Benchmarking Semantic Segmentation Models via Appearance and Geometry Attribute Editing
Fuente:
arXiv
Salvato in:
| Autori principali: | Yin, Zijin, Li, Bing, Liang, Kongming, Sun, Hao, He, Zhongjiang, Ma, Zhanyu, Guo, Jun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Benchmarking Segmentation Models with Mask-Preserved Attribute Editing
di: Yin, Zijin, et al.
Pubblicazione: (2024)
di: Yin, Zijin, et al.
Pubblicazione: (2024)
Polyp-E: Benchmarking the Robustness of Deep Segmentation Models via Polyp Editing
di: Wei, Runpu, et al.
Pubblicazione: (2024)
di: Wei, Runpu, et al.
Pubblicazione: (2024)
Detailed Object Description with Controllable Dimensions
di: Wang, Xinran, et al.
Pubblicazione: (2024)
di: Wang, Xinran, et al.
Pubblicazione: (2024)
PGP-SAM: Prototype-Guided Prompt Learning for Efficient Few-Shot Medical Image Segmentation
di: Yan, Zhonghao, et al.
Pubblicazione: (2025)
di: Yan, Zhonghao, et al.
Pubblicazione: (2025)
Geometric Image Editing via Effects-Sensitive In-Context Inpainting with Diffusion Transformers
di: Zhang, Shuo, et al.
Pubblicazione: (2026)
di: Zhang, Shuo, et al.
Pubblicazione: (2026)
FairHuman: Boosting Hand and Face Quality in Human Image Generation with Minimum Potential Delay Fairness in Diffusion Models
di: Wang, Yuxuan, et al.
Pubblicazione: (2025)
di: Wang, Yuxuan, et al.
Pubblicazione: (2025)
OmniEraser: Remove Objects and Their Effects in Images with Paired Video-Frame Data
di: Wei, Runpu, et al.
Pubblicazione: (2025)
di: Wei, Runpu, et al.
Pubblicazione: (2025)
Generative Visual Chain-of-Thought for Image Editing
di: Yin, Zijin, et al.
Pubblicazione: (2026)
di: Yin, Zijin, et al.
Pubblicazione: (2026)
ConMo: Controllable Motion Disentanglement and Recomposition for Zero-Shot Motion Transfer
di: Gao, Jiayi, et al.
Pubblicazione: (2025)
di: Gao, Jiayi, et al.
Pubblicazione: (2025)
Evaluating Attribute Comprehension in Large Vision-Language Models
di: Zhang, Haiwen, et al.
Pubblicazione: (2024)
di: Zhang, Haiwen, et al.
Pubblicazione: (2024)
Efficient Face Super-Resolution via Wavelet-based Feature Enhancement Network
di: Li, Wenjie, et al.
Pubblicazione: (2024)
di: Li, Wenjie, et al.
Pubblicazione: (2024)
Curriculum Group Policy Optimization: Adaptive Sampling for Unleashing the Potential of Text-to-Image Generation
di: Li, Baoteng, et al.
Pubblicazione: (2026)
di: Li, Baoteng, et al.
Pubblicazione: (2026)
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation
di: Wang, Xinran, et al.
Pubblicazione: (2025)
di: Wang, Xinran, et al.
Pubblicazione: (2025)
Controllable-Continuous Color Editing in Diffusion Model via Color Mapping
di: Yang, Yuqi, et al.
Pubblicazione: (2025)
di: Yang, Yuqi, et al.
Pubblicazione: (2025)
DriveRX: A Vision-Language Reasoning Model for Cross-Task Autonomous Driving
di: Diao, Muxi, et al.
Pubblicazione: (2025)
di: Diao, Muxi, et al.
Pubblicazione: (2025)
Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models
di: Tong, Yujun, et al.
Pubblicazione: (2026)
di: Tong, Yujun, et al.
Pubblicazione: (2026)
CineTechBench: A Benchmark for Cinematographic Technique Understanding and Generation
di: Wang, Xinran, et al.
Pubblicazione: (2025)
di: Wang, Xinran, et al.
Pubblicazione: (2025)
From Simple to Professional: A Combinatorial Controllable Image Captioning Agent
di: Wang, Xinran, et al.
Pubblicazione: (2024)
di: Wang, Xinran, et al.
Pubblicazione: (2024)
Toward Generalizable Forgery Detection and Reasoning
di: Gao, Yueying, et al.
Pubblicazione: (2025)
di: Gao, Yueying, et al.
Pubblicazione: (2025)
Layout2Scene: 3D Semantic Layout Guided Scene Generation via Geometry and Appearance Diffusion Priors
di: Chen, Minglin, et al.
Pubblicazione: (2025)
di: Chen, Minglin, et al.
Pubblicazione: (2025)
Recolour What Matters: Region-Aware Colour Editing via Token-Level Diffusion
di: Yang, Yuqi, et al.
Pubblicazione: (2026)
di: Yang, Yuqi, et al.
Pubblicazione: (2026)
Geometry-Editable and Appearance-Preserving Object Compositon
di: Lin, Jianman, et al.
Pubblicazione: (2025)
di: Lin, Jianman, et al.
Pubblicazione: (2025)
PGAHum: Prior-Guided Geometry and Appearance Learning for High-Fidelity Animatable Human Reconstruction
di: Wang, Hao, et al.
Pubblicazione: (2024)
di: Wang, Hao, et al.
Pubblicazione: (2024)
IncreFA: Breaking the Static Wall of Generative Model Attribution
di: Qin, Haotian, et al.
Pubblicazione: (2026)
di: Qin, Haotian, et al.
Pubblicazione: (2026)
PolarVLM: Bridging the Semantic-Physical Gap in Vision-Language Models
di: Li, Yuliang, et al.
Pubblicazione: (2026)
di: Li, Yuliang, et al.
Pubblicazione: (2026)
The GAN that Warped: Semantic Attribute Editing with Unpaired Data
di: Dorta, Gara, et al.
Pubblicazione: (2018)
di: Dorta, Gara, et al.
Pubblicazione: (2018)
UniEdit: A Unified Tuning-Free Framework for Video Motion and Appearance Editing
di: Bai, Jianhong, et al.
Pubblicazione: (2024)
di: Bai, Jianhong, et al.
Pubblicazione: (2024)
Hepato-LLaVA: An Expert MLLM with Sparse Topo-Pack Attention for Hepatocellular Pathology Analysis on Whole Slide Images
di: Yang, Yuxuan, et al.
Pubblicazione: (2026)
di: Yang, Yuxuan, et al.
Pubblicazione: (2026)
DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions
di: Wang, Xinran, et al.
Pubblicazione: (2026)
di: Wang, Xinran, et al.
Pubblicazione: (2026)
Velox: Learning Representations of 4D Geometry and Appearance
di: Malik, Anagh, et al.
Pubblicazione: (2026)
di: Malik, Anagh, et al.
Pubblicazione: (2026)
Learning Naturally Aggregated Appearance for Efficient 3D Editing
di: Cheng, Ka Leong, et al.
Pubblicazione: (2023)
di: Cheng, Ka Leong, et al.
Pubblicazione: (2023)
Cityscape-Adverse: Benchmarking Robustness of Semantic Segmentation with Realistic Scene Modifications via Diffusion-Based Image Editing
di: Suryanto, Naufal, et al.
Pubblicazione: (2024)
di: Suryanto, Naufal, et al.
Pubblicazione: (2024)
TAPESTRY: From Geometry to Appearance via Consistent Turntable Videos
di: Zeng, Yan, et al.
Pubblicazione: (2026)
di: Zeng, Yan, et al.
Pubblicazione: (2026)
Early Semantic Grounding in Image Editing Models for Zero-Shot Referring Image Segmentation
di: He, Jingxuan, et al.
Pubblicazione: (2026)
di: He, Jingxuan, et al.
Pubblicazione: (2026)
DFormerv2: Geometry Self-Attention for RGBD Semantic Segmentation
di: Yin, Bo-Wen, et al.
Pubblicazione: (2025)
di: Yin, Bo-Wen, et al.
Pubblicazione: (2025)
Adaptive Evidential Learning for Temporal-Semantic Robustness in Moment Retrieval
di: Huang, Haojian, et al.
Pubblicazione: (2025)
di: Huang, Haojian, et al.
Pubblicazione: (2025)
Leveraging Contrastive Learning for Semantic Segmentation with Consistent Labels Across Varying Appearances
di: Montalvo, Javier, et al.
Pubblicazione: (2024)
di: Montalvo, Javier, et al.
Pubblicazione: (2024)
GA-Drive: Geometry-Appearance Decoupled Modeling for Free-viewpoint Driving Scene Generation
di: Zhang, Hao, et al.
Pubblicazione: (2026)
di: Zhang, Hao, et al.
Pubblicazione: (2026)
Relative Difficulty Distillation for Semantic Segmentation
di: Liang, Dong, et al.
Pubblicazione: (2024)
di: Liang, Dong, et al.
Pubblicazione: (2024)
GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion
di: Zhu, Hanxin, et al.
Pubblicazione: (2026)
di: Zhu, Hanxin, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Benchmarking Segmentation Models with Mask-Preserved Attribute Editing
di: Yin, Zijin, et al.
Pubblicazione: (2024) -
Polyp-E: Benchmarking the Robustness of Deep Segmentation Models via Polyp Editing
di: Wei, Runpu, et al.
Pubblicazione: (2024) -
Detailed Object Description with Controllable Dimensions
di: Wang, Xinran, et al.
Pubblicazione: (2024) -
PGP-SAM: Prototype-Guided Prompt Learning for Efficient Few-Shot Medical Image Segmentation
di: Yan, Zhonghao, et al.
Pubblicazione: (2025) -
Geometric Image Editing via Effects-Sensitive In-Context Inpainting with Diffusion Transformers
di: Zhang, Shuo, et al.
Pubblicazione: (2026)