ImageAttributionBench: How Far Are We from Generalizable Attribution?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mou, Tingshu, Wei, Zhipeng, Gong, Chao, Chen, Jingjing, Ma, Xingjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ViSRA: A Video-based Spatial Reasoning Agent for Multi-modal Large Language Models
von: Mou, Tingshu, et al.
Veröffentlicht: (2026)
von: Mou, Tingshu, et al.
Veröffentlicht: (2026)
ReToMe-VA: Recursive Token Merging for Video Diffusion-based Unrestricted Adversarial Attack
von: Gao, Ziyi, et al.
Veröffentlicht: (2024)
von: Gao, Ziyi, et al.
Veröffentlicht: (2024)
Reliable and Efficient Concept Erasure of Text-to-Image Diffusion Models
von: Gong, Chao, et al.
Veröffentlicht: (2024)
von: Gong, Chao, et al.
Veröffentlicht: (2024)
PixelWorld: How Far Are We from Perceiving Everything as Pixels?
von: Lyu, Zhiheng, et al.
Veröffentlicht: (2025)
von: Lyu, Zhiheng, et al.
Veröffentlicht: (2025)
VCE: Safe Autoregressive Image Generation via Visual Contrast Exploitation
von: Han, Feng, et al.
Veröffentlicht: (2025)
von: Han, Feng, et al.
Veröffentlicht: (2025)
UniREditBench: A Unified Reasoning-based Image Editing Benchmark
von: Han, Feng, et al.
Veröffentlicht: (2025)
von: Han, Feng, et al.
Veröffentlicht: (2025)
ATTIQA: Generalizable Image Quality Feature Extractor using Attribute-aware Pretraining
von: Kwon, Daekyu, et al.
Veröffentlicht: (2024)
von: Kwon, Daekyu, et al.
Veröffentlicht: (2024)
IDEA-Bench: How Far are Generative Models from Professional Designing?
von: Liang, Chen, et al.
Veröffentlicht: (2024)
von: Liang, Chen, et al.
Veröffentlicht: (2024)
DynamicEarth: How Far are We from Open-Vocabulary Change Detection?
von: Li, Kaiyu, et al.
Veröffentlicht: (2025)
von: Li, Kaiyu, et al.
Veröffentlicht: (2025)
PICABench: How Far Are We from Physically Realistic Image Editing?
von: Pu, Yuandong, et al.
Veröffentlicht: (2025)
von: Pu, Yuandong, et al.
Veröffentlicht: (2025)
How Far Are We from Generating Missing Modalities with Foundation Models?
von: Ke, Guanzhou, et al.
Veröffentlicht: (2025)
von: Ke, Guanzhou, et al.
Veröffentlicht: (2025)
Attribution as Retrieval: Model-Agnostic AI-Generated Image Attribution
von: Wang, Hongsong, et al.
Veröffentlicht: (2026)
von: Wang, Hongsong, et al.
Veröffentlicht: (2026)
Pedestrian Attribute Editing for Gait Recognition and Anonymization
von: Ma, Jingzhe, et al.
Veröffentlicht: (2023)
von: Ma, Jingzhe, et al.
Veröffentlicht: (2023)
Copyright Infringement Detection in Text-to-Image Diffusion Models via Differential Privacy
von: Man, Xiafeng, et al.
Veröffentlicht: (2025)
von: Man, Xiafeng, et al.
Veröffentlicht: (2025)
Learning to Infer Unseen Single-/Multi-Attribute-Object Compositions with Graph Networks
von: Chen, Hui, et al.
Veröffentlicht: (2020)
von: Chen, Hui, et al.
Veröffentlicht: (2020)
How Far Can We Compress Instant-NGP-Based NeRF?
von: Chen, Yihang, et al.
Veröffentlicht: (2024)
von: Chen, Yihang, et al.
Veröffentlicht: (2024)
Rethinking Generative Image Pretraining: How Far Are We From Scaling Up Next-Pixel Prediction?
von: Yan, Xinchen, et al.
Veröffentlicht: (2025)
von: Yan, Xinchen, et al.
Veröffentlicht: (2025)
EchoingPixels: Cross-Modal Adaptive Token Reduction for Efficient Audio-Visual LLMs
von: Gong, Chao, et al.
Veröffentlicht: (2025)
von: Gong, Chao, et al.
Veröffentlicht: (2025)
DuMo: Dual Encoder Modulation Network for Precise Concept Erasure
von: Han, Feng, et al.
Veröffentlicht: (2025)
von: Han, Feng, et al.
Veröffentlicht: (2025)
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
von: Chen, Zhe, et al.
Veröffentlicht: (2024)
von: Chen, Zhe, et al.
Veröffentlicht: (2024)
FaceBench: A Multi-View Multi-Level Facial Attribute VQA Dataset for Benchmarking Face Perception MLLMs
von: Wang, Xiaoqin, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoqin, et al.
Veröffentlicht: (2025)
High-Fidelity GAN Inversion for Image Attribute Editing
von: Wang, Tengfei, et al.
Veröffentlicht: (2021)
von: Wang, Tengfei, et al.
Veröffentlicht: (2021)
Insight-A: Attribution-aware for Multimodal Misinformation Detection
von: Wu, Junjie, et al.
Veröffentlicht: (2025)
von: Wu, Junjie, et al.
Veröffentlicht: (2025)
TAPT: Test-Time Adversarial Prompt Tuning for Robust Inference in Vision-Language Models
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
VideoAVE: A Multi-Attribute Video-to-Text Attribute Value Extraction Dataset and Benchmark Models
von: Cheng, Ming, et al.
Veröffentlicht: (2025)
von: Cheng, Ming, et al.
Veröffentlicht: (2025)
Compositional Attribute Imbalance in Vision Datasets
von: Chen, Jiayi, et al.
Veröffentlicht: (2025)
von: Chen, Jiayi, et al.
Veröffentlicht: (2025)
Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
InstructAttribute: Fine-grained Object Attributes editing with Instruction
von: Yin, Xingxi, et al.
Veröffentlicht: (2025)
von: Yin, Xingxi, et al.
Veröffentlicht: (2025)
WildDoc: How Far Are We from Achieving Comprehensive and Robust Document Understanding in the Wild?
von: Wang, An-Lan, et al.
Veröffentlicht: (2025)
von: Wang, An-Lan, et al.
Veröffentlicht: (2025)
GIM: Learning Generalizable Image Matcher From Internet Videos
von: Shen, Xuelun, et al.
Veröffentlicht: (2024)
von: Shen, Xuelun, et al.
Veröffentlicht: (2024)
ACMo: Attribute Controllable Motion Generation
von: Wei, Mingjie, et al.
Veröffentlicht: (2025)
von: Wei, Mingjie, et al.
Veröffentlicht: (2025)
How Far Are We from Intelligent Visual Deductive Reasoning?
von: Zhang, Yizhe, et al.
Veröffentlicht: (2024)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2024)
DO-Bench: An Attributable Benchmark for Diagnosing Object Hallucination in Vision-Language Models
von: Wang, JiYang, et al.
Veröffentlicht: (2026)
von: Wang, JiYang, et al.
Veröffentlicht: (2026)
Composing Object Relations and Attributes for Image-Text Matching
von: Pham, Khoi, et al.
Veröffentlicht: (2024)
von: Pham, Khoi, et al.
Veröffentlicht: (2024)
Evaluating Attribute Confusion in Fashion Text-to-Image Generation
von: Liu, Ziyue, et al.
Veröffentlicht: (2025)
von: Liu, Ziyue, et al.
Veröffentlicht: (2025)
Detecting Origin Attribution for Text-to-Image Diffusion Models
von: Xu, Katherine, et al.
Veröffentlicht: (2024)
von: Xu, Katherine, et al.
Veröffentlicht: (2024)
U-Face: An Efficient and Generalizable Framework for Unsupervised Facial Attribute Editing via Subspace Learning
von: Liu, Bo, et al.
Veröffentlicht: (2026)
von: Liu, Bo, et al.
Veröffentlicht: (2026)
Unsupervised Synthetic Image Attribution: Alignment and Disentanglement
von: Liu, Zongfang, et al.
Veröffentlicht: (2026)
von: Liu, Zongfang, et al.
Veröffentlicht: (2026)
DreamMix: Decoupling Object Attributes for Enhanced Editability in Customized Image Inpainting
von: Yang, Yicheng, et al.
Veröffentlicht: (2024)
von: Yang, Yicheng, et al.
Veröffentlicht: (2024)
FaceGemma: Enhancing Image Captioning with Facial Attributes for Portrait Images
von: Haque, Naimul, et al.
Veröffentlicht: (2023)
von: Haque, Naimul, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
ViSRA: A Video-based Spatial Reasoning Agent for Multi-modal Large Language Models
von: Mou, Tingshu, et al.
Veröffentlicht: (2026) -
ReToMe-VA: Recursive Token Merging for Video Diffusion-based Unrestricted Adversarial Attack
von: Gao, Ziyi, et al.
Veröffentlicht: (2024) -
Reliable and Efficient Concept Erasure of Text-to-Image Diffusion Models
von: Gong, Chao, et al.
Veröffentlicht: (2024) -
PixelWorld: How Far Are We from Perceiving Everything as Pixels?
von: Lyu, Zhiheng, et al.
Veröffentlicht: (2025) -
VCE: Safe Autoregressive Image Generation via Visual Contrast Exploitation
von: Han, Feng, et al.
Veröffentlicht: (2025)