Too Vivid to Be Real? Benchmarking and Calibrating Generative Color Fidelity
Fuente:
arXiv
Saved in:
| Main Authors: | Fang, Zhengyao, Jia, Zexi, Zhong, Yijia, Luo, Pengcheng, Zhang, Jinchao, Lu, Guangming, Yu, Jun, Pei, Wenjie |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Manifold-Optimal Guidance: A Unified Riemannian Control View of Diffusion Guidance
by: Jia, Zexi, et al.
Published: (2026)
by: Jia, Zexi, et al.
Published: (2026)
Evaluating Generative Models via One-Dimensional Code Distributions
by: Jia, Zexi, et al.
Published: (2026)
by: Jia, Zexi, et al.
Published: (2026)
Prune Redundancy, Preserve Essence: Vision Token Compression in VLMs via Synergistic Importance-Diversity
by: Fang, Zhengyao, et al.
Published: (2026)
by: Fang, Zhengyao, et al.
Published: (2026)
Recognition-Synergistic Scene Text Editing
by: Fang, Zhengyao, et al.
Published: (2025)
by: Fang, Zhengyao, et al.
Published: (2025)
StyleDecoupler: Generalizable Artistic Style Disentanglement
by: Jia, Zexi, et al.
Published: (2026)
by: Jia, Zexi, et al.
Published: (2026)
WeCromCL: Weakly Supervised Cross-Modality Contrastive Learning for Transcription-only Supervised Text Spotting
by: Wu, Jingjing, et al.
Published: (2024)
by: Wu, Jingjing, et al.
Published: (2024)
EditInfinity: Image Editing with Binary-Quantized Generative Models
by: Wang, Jiahuan, et al.
Published: (2025)
by: Wang, Jiahuan, et al.
Published: (2025)
CoDA: Color Distribution Probing for Efficient and Generalizable AI-Generated Image Detection
by: Jia, Zexi, et al.
Published: (2026)
by: Jia, Zexi, et al.
Published: (2026)
D$^2$ST-Adapter: Disentangled-and-Deformable Spatio-Temporal Adapter for Few-shot Action Recognition
by: Pei, Wenjie, et al.
Published: (2023)
by: Pei, Wenjie, et al.
Published: (2023)
Video-ToC: Video Tree-of-Cue Reasoning
by: Tan, Qizhong, et al.
Published: (2026)
by: Tan, Qizhong, et al.
Published: (2026)
VividDreamer: Towards High-Fidelity and Efficient Text-to-3D Generation
by: Chen, Zixuan, et al.
Published: (2024)
by: Chen, Zixuan, et al.
Published: (2024)
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
by: Hu, Teng, et al.
Published: (2025)
by: Hu, Teng, et al.
Published: (2025)
VividDreamer: Invariant Score Distillation For Hyper-Realistic Text-to-3D Generation
by: Zhuo, Wenjie, et al.
Published: (2024)
by: Zhuo, Wenjie, et al.
Published: (2024)
Global-Local Stepwise Generative Network for Ultra High-Resolution Image Restoration
by: Feng, Xin, et al.
Published: (2022)
by: Feng, Xin, et al.
Published: (2022)
Beyond Heuristics: Learnable Density Control for 3D Gaussian Splatting
by: Ning, Zhenhua, et al.
Published: (2026)
by: Ning, Zhenhua, et al.
Published: (2026)
Exploring Specular Reflection Inconsistency for Generalizable Face Forgery Detection
by: Fei, Hongyan, et al.
Published: (2026)
by: Fei, Hongyan, et al.
Published: (2026)
DiffTrans: Differentiable Geometry-Materials Decomposition for Reconstructing Transparent Objects
by: Li, Changpu, et al.
Published: (2026)
by: Li, Changpu, et al.
Published: (2026)
HairDiffusion: Vivid Multi-Colored Hair Editing via Latent Diffusion
by: Zeng, Yu, et al.
Published: (2024)
by: Zeng, Yu, et al.
Published: (2024)
Control-CLIP: Decoupling Category and Style Guidance in CLIP for Specific-Domain Generation
by: Jia, Zexi, et al.
Published: (2025)
by: Jia, Zexi, et al.
Published: (2025)
RDTF: Resource-efficient Dual-mask Training Framework for Multi-frame Animated Sticker Generation
by: Yuan, Zhiqiang, et al.
Published: (2025)
by: Yuan, Zhiqiang, et al.
Published: (2025)
Learning Compatible Multi-Prize Subnetworks for Asymmetric Retrieval
by: Sun, Yushuai, et al.
Published: (2025)
by: Sun, Yushuai, et al.
Published: (2025)
Domain-Rectifying Adapter for Cross-Domain Few-Shot Segmentation
by: Su, Jiapeng, et al.
Published: (2024)
by: Su, Jiapeng, et al.
Published: (2024)
RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation
by: Wen, Junwei, et al.
Published: (2026)
by: Wen, Junwei, et al.
Published: (2026)
UniVoxel: Fast Inverse Rendering by Unified Voxelization of Scene Representation
by: Wu, Shuang, et al.
Published: (2024)
by: Wu, Shuang, et al.
Published: (2024)
VividDream: Generating 3D Scene with Ambient Dynamics
by: Lee, Yao-Chih, et al.
Published: (2024)
by: Lee, Yao-Chih, et al.
Published: (2024)
VGG-Tex: A Vivid Geometry-Guided Facial Texture Estimation Model for High Fidelity Monocular 3D Face Reconstruction
by: Wu, Haoyu, et al.
Published: (2024)
by: Wu, Haoyu, et al.
Published: (2024)
VividFace: A Diffusion-Based Hybrid Framework for High-Fidelity Video Face Swapping
by: Shao, Hao, et al.
Published: (2024)
by: Shao, Hao, et al.
Published: (2024)
Let Storytelling Tell Vivid Stories: An Expressive and Fluent Multimodal Storyteller
by: Zang, Chuanqi, et al.
Published: (2024)
by: Zang, Chuanqi, et al.
Published: (2024)
Semantic to Structure: Learning Structural Representations for Infringement Detection
by: Huang, Chuanwei, et al.
Published: (2025)
by: Huang, Chuanwei, et al.
Published: (2025)
Vivid-ZOO: Multi-View Video Generation with Diffusion Model
by: Li, Bing, et al.
Published: (2024)
by: Li, Bing, et al.
Published: (2024)
A Visual Leap in CLIP Compositionality Reasoning through Generation of Counterfactual Sets
by: Jia, Zexi, et al.
Published: (2025)
by: Jia, Zexi, et al.
Published: (2025)
F2RVLM: Boosting Fine-grained Fragment Retrieval for Multi-Modal Long-form Dialogue with Vision Language Model
by: Bi, Hanbo, et al.
Published: (2025)
by: Bi, Hanbo, et al.
Published: (2025)
Dif-Fusion: Towards High Color Fidelity in Infrared and Visible Image Fusion with Diffusion Models
by: Yue, Jun, et al.
Published: (2023)
by: Yue, Jun, et al.
Published: (2023)
Textoon: Generating Vivid 2D Cartoon Characters from Text Descriptions
by: He, Chao, et al.
Published: (2025)
by: He, Chao, et al.
Published: (2025)
Enhancing Spatial Reasoning in Multimodal Large Language Models through Reasoning-based Segmentation
by: Ning, Zhenhua, et al.
Published: (2025)
by: Ning, Zhenhua, et al.
Published: (2025)
Guided Real Image Dehazing using YCbCr Color Space
by: Fang, Wenxuan, et al.
Published: (2024)
by: Fang, Wenxuan, et al.
Published: (2024)
VividMed: Vision Language Model with Versatile Visual Grounding for Medicine
by: Luo, Lingxiao, et al.
Published: (2024)
by: Luo, Lingxiao, et al.
Published: (2024)
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model
by: Li, Yan, et al.
Published: (2024)
by: Li, Yan, et al.
Published: (2024)
Architect: Generating Vivid and Interactive 3D Scenes with Hierarchical 2D Inpainting
by: Wang, Yian, et al.
Published: (2024)
by: Wang, Yian, et al.
Published: (2024)
Phenotype-Guided Generative Model for High-Fidelity Cardiac MRI Synthesis: Advancing Pretraining and Clinical Applications
by: Li, Ziyu, et al.
Published: (2025)
by: Li, Ziyu, et al.
Published: (2025)
Similar Items
-
Manifold-Optimal Guidance: A Unified Riemannian Control View of Diffusion Guidance
by: Jia, Zexi, et al.
Published: (2026) -
Evaluating Generative Models via One-Dimensional Code Distributions
by: Jia, Zexi, et al.
Published: (2026) -
Prune Redundancy, Preserve Essence: Vision Token Compression in VLMs via Synergistic Importance-Diversity
by: Fang, Zhengyao, et al.
Published: (2026) -
Recognition-Synergistic Scene Text Editing
by: Fang, Zhengyao, et al.
Published: (2025) -
StyleDecoupler: Generalizable Artistic Style Disentanglement
by: Jia, Zexi, et al.
Published: (2026)