Language-based Image Colorization: A Benchmark and Beyond
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yifan, Yang, Shuai, Liu, Jiaying |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-Supervised Skeleton-Based Action Representation Learning: A Benchmark and Beyond
von: Zhang, Jiahang, et al.
Veröffentlicht: (2024)
von: Zhang, Jiahang, et al.
Veröffentlicht: (2024)
VibeFlow: Versatile Video Chroma-Lux Editing through Self-Supervised Learning
von: Li, Yifan, et al.
Veröffentlicht: (2026)
von: Li, Yifan, et al.
Veröffentlicht: (2026)
PTDiffusion: Free Lunch for Generating Optical Illusion Hidden Pictures with Phase-Transferred Diffusion Model
von: Gao, Xiang, et al.
Veröffentlicht: (2025)
von: Gao, Xiang, et al.
Veröffentlicht: (2025)
Benchmarking Endoscopic Surgical Image Restoration and Beyond
von: Pei, Jialun, et al.
Veröffentlicht: (2025)
von: Pei, Jialun, et al.
Veröffentlicht: (2025)
Can MLLMs Reason Beyond Language? VisReason: A Comprehensive Benchmark for Vision-Centric Reasoning
von: Guo, Longteng, et al.
Veröffentlicht: (2026)
von: Guo, Longteng, et al.
Veröffentlicht: (2026)
Beyond Walking: A Large-Scale Image-Text Benchmark for Text-based Person Anomaly Search
von: Yang, Shuyu, et al.
Veröffentlicht: (2024)
von: Yang, Shuyu, et al.
Veröffentlicht: (2024)
Beyond Open Vocabulary: Multimodal Prompting for Object Detection in Remote Sensing Images
von: Yang, Shuai, et al.
Veröffentlicht: (2026)
von: Yang, Shuai, et al.
Veröffentlicht: (2026)
Control Color: Multimodal Diffusion-based Interactive Image Colorization
von: Liang, Zhexin, et al.
Veröffentlicht: (2024)
von: Liang, Zhexin, et al.
Veröffentlicht: (2024)
Adaptive Context Matters: Towards Provable Multi-Modality Guidance for Super-Resolution
von: Luo, Jinyi, et al.
Veröffentlicht: (2026)
von: Luo, Jinyi, et al.
Veröffentlicht: (2026)
GenColorBench: A Color Evaluation Benchmark for Text-to-Image Generation Models
von: Butt, Muhammad Atif, et al.
Veröffentlicht: (2025)
von: Butt, Muhammad Atif, et al.
Veröffentlicht: (2025)
Beyond Model Design: Data-Centric Training and Self-Ensemble for Gaussian Color Image Denoising
von: Chang, Gengjia, et al.
Veröffentlicht: (2026)
von: Chang, Gengjia, et al.
Veröffentlicht: (2026)
Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
OmniFM: Toward Modality-Robust and Task-Agnostic Federated Learning for Heterogeneous Medical Imaging
von: Liu, Meilin, et al.
Veröffentlicht: (2026)
von: Liu, Meilin, et al.
Veröffentlicht: (2026)
Beyond the Visible: Benchmarking Occlusion Perception in Multimodal Large Language Models
von: Liu, Zhaochen, et al.
Veröffentlicht: (2025)
von: Liu, Zhaochen, et al.
Veröffentlicht: (2025)
ColorConceptBench: A Benchmark for Probabilistic Color-Concept Understanding in Text-to-Image Models
von: Ruan, Chenxi, et al.
Veröffentlicht: (2026)
von: Ruan, Chenxi, et al.
Veröffentlicht: (2026)
DFBench: Benchmarking Deepfake Image Detection Capability of Large Multimodal Models
von: Wang, Jiarui, et al.
Veröffentlicht: (2025)
von: Wang, Jiarui, et al.
Veröffentlicht: (2025)
Beyond the Destination: A Novel Benchmark for Exploration-Aware Embodied Question Answering
von: Jiang, Kaixuan, et al.
Veröffentlicht: (2025)
von: Jiang, Kaixuan, et al.
Veröffentlicht: (2025)
Contrast-X: A Multi-Modal Contrast Image Synthesis Benchmark and Universal Modality Flow Matching
von: Chen, Yifan, et al.
Veröffentlicht: (2026)
von: Chen, Yifan, et al.
Veröffentlicht: (2026)
Beyond Memorization: A Multi-Modal Ordinal Regression Benchmark to Expose Popularity Bias in Vision-Language Models
von: Szu-Tu, Li-Zhong, et al.
Veröffentlicht: (2025)
von: Szu-Tu, Li-Zhong, et al.
Veröffentlicht: (2025)
Beyond Forgetting in Continual Medical Image Segmentation: A Comprehensive Benchmark Study
von: Wang, Bomin, et al.
Veröffentlicht: (2026)
von: Wang, Bomin, et al.
Veröffentlicht: (2026)
Advancing Visual Reliability: Color-Accurate Underwater Image Enhancement for Real-Time Underwater Missions
von: Zhou, Yiqiang, et al.
Veröffentlicht: (2026)
von: Zhou, Yiqiang, et al.
Veröffentlicht: (2026)
From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models
von: Li, Lingyao, et al.
Veröffentlicht: (2025)
von: Li, Lingyao, et al.
Veröffentlicht: (2025)
Boosting Weakly-Supervised Referring Image Segmentation via Progressive Comprehension
von: Yang, Zaiquan, et al.
Veröffentlicht: (2024)
von: Yang, Zaiquan, et al.
Veröffentlicht: (2024)
Beyond the Last Frame: Process-aware Evaluation for Generative Video Reasoning
von: Li, Yifan, et al.
Veröffentlicht: (2025)
von: Li, Yifan, et al.
Veröffentlicht: (2025)
Palette-based Color Transfer between Images
von: Lv, Chenlei, et al.
Veröffentlicht: (2024)
von: Lv, Chenlei, et al.
Veröffentlicht: (2024)
VL-RewardBench: A Challenging Benchmark for Vision-Language Generative Reward Models
von: Li, Lei, et al.
Veröffentlicht: (2024)
von: Li, Lei, et al.
Veröffentlicht: (2024)
FBSDiff: Plug-and-Play Frequency Band Substitution of Diffusion Features for Highly Controllable Text-Driven Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
UniREditBench: A Unified Reasoning-based Image Editing Benchmark
von: Han, Feng, et al.
Veröffentlicht: (2025)
von: Han, Feng, et al.
Veröffentlicht: (2025)
FRESCO: Spatial-Temporal Correspondence for Zero-Shot Video Translation
von: Yang, Shuai, et al.
Veröffentlicht: (2024)
von: Yang, Shuai, et al.
Veröffentlicht: (2024)
BioBench: A Blueprint to Move Beyond ImageNet for Scientific ML Benchmarks
von: Stevens, Samuel
Veröffentlicht: (2025)
von: Stevens, Samuel
Veröffentlicht: (2025)
MotionBank: A Large-scale Video Motion Benchmark with Disentangled Rule-based Annotations
von: Xu, Liang, et al.
Veröffentlicht: (2024)
von: Xu, Liang, et al.
Veröffentlicht: (2024)
ColorFlow: Retrieval-Augmented Image Sequence Colorization
von: Zhuang, Junhao, et al.
Veröffentlicht: (2024)
von: Zhuang, Junhao, et al.
Veröffentlicht: (2024)
Intelligent Artistic Typography: A Comprehensive Review of Artistic Text Design and Generation
von: Bai, Yuhang, et al.
Veröffentlicht: (2024)
von: Bai, Yuhang, et al.
Veröffentlicht: (2024)
VRSBench: A Versatile Vision-Language Benchmark Dataset for Remote Sensing Image Understanding
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
RecipeGen: A Benchmark for Real-World Recipe Image Generation
von: Zhang, Ruoxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Ruoxuan, et al.
Veröffentlicht: (2025)
Rethinking Facial Expression Recognition in the Era of Multimodal Large Language Models: Benchmark, Datasets, and Beyond
von: Zhang, Fan, et al.
Veröffentlicht: (2025)
von: Zhang, Fan, et al.
Veröffentlicht: (2025)
A Benchmark for Multi-Lingual Vision-Language Learning in Remote Sensing Image Captioning
von: Zhou, Qing, et al.
Veröffentlicht: (2025)
von: Zhou, Qing, et al.
Veröffentlicht: (2025)
Animate-X++: Universal Character Image Animation with Dynamic Backgrounds
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
Learning Spatially Decoupled Color Representations for Facial Image Colorization
von: Zhu, Hangyan, et al.
Veröffentlicht: (2024)
von: Zhu, Hangyan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Self-Supervised Skeleton-Based Action Representation Learning: A Benchmark and Beyond
von: Zhang, Jiahang, et al.
Veröffentlicht: (2024) -
VibeFlow: Versatile Video Chroma-Lux Editing through Self-Supervised Learning
von: Li, Yifan, et al.
Veröffentlicht: (2026) -
PTDiffusion: Free Lunch for Generating Optical Illusion Hidden Pictures with Phase-Transferred Diffusion Model
von: Gao, Xiang, et al.
Veröffentlicht: (2025) -
Benchmarking Endoscopic Surgical Image Restoration and Beyond
von: Pei, Jialun, et al.
Veröffentlicht: (2025) -
Can MLLMs Reason Beyond Language? VisReason: A Comprehensive Benchmark for Vision-Centric Reasoning
von: Guo, Longteng, et al.
Veröffentlicht: (2026)