AceTone: Bridging Words and Colors for Conditional Image Grading
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Tianren, Liao, Mingxiang, Zhang, Xijin, Ye, Qixiang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ReDDiT: Rehashing Noise for Discrete Visual Generation
by: Ma, Tianren, et al.
Published: (2025)
by: Ma, Tianren, et al.
Published: (2025)
ClawMachine: Learning to Fetch Visual Tokens for Referential Comprehension
by: Ma, Tianren, et al.
Published: (2024)
by: Ma, Tianren, et al.
Published: (2024)
DynRefer: Delving into Region-level Multimodal Tasks via Dynamic Resolution
by: Zhao, Yuzhong, et al.
Published: (2024)
by: Zhao, Yuzhong, et al.
Published: (2024)
Evaluation of Text-to-Video Generation Models: A Dynamics Perspective
by: Liao, Mingxiang, et al.
Published: (2024)
by: Liao, Mingxiang, et al.
Published: (2024)
ChatterBox: Multi-round Multimodal Referring and Grounding
by: Tian, Yunjie, et al.
Published: (2024)
by: Tian, Yunjie, et al.
Published: (2024)
Artemis: Towards Referential Understanding in Complex Videos
by: Qiu, Jihao, et al.
Published: (2024)
by: Qiu, Jihao, et al.
Published: (2024)
Thinking with Images via Self-Calling Agent
by: Yang, Wenxi, et al.
Published: (2025)
by: Yang, Wenxi, et al.
Published: (2025)
CC-Diff: Enhancing Contextual Coherence in Remote Sensing Image Synthesis
by: Zhang, Mu, et al.
Published: (2024)
by: Zhang, Mu, et al.
Published: (2024)
AceVFI: A Comprehensive Survey of Advances in Video Frame Interpolation
by: Kye, Dahyeon, et al.
Published: (2025)
by: Kye, Dahyeon, et al.
Published: (2025)
S&D Messenger: Exchanging Semantic and Domain Knowledge for Generic Semi-Supervised Medical Image Segmentation
by: Zhang, Qixiang, et al.
Published: (2024)
by: Zhang, Qixiang, et al.
Published: (2024)
Delving Deep into Semantic Relation Distillation
by: Yan, Zhaoyi, et al.
Published: (2025)
by: Yan, Zhaoyi, et al.
Published: (2025)
Probabilistic Prior Driven Attention Mechanism Based on Diffusion Model for Imaging Through Atmospheric Turbulence
by: Sun, Guodong, et al.
Published: (2024)
by: Sun, Guodong, et al.
Published: (2024)
Mixture of Physical Priors Adapter for Parameter-Efficient Fine-Tuning
by: Wang, Zhaozhi, et al.
Published: (2024)
by: Wang, Zhaozhi, et al.
Published: (2024)
Delta-SVD: Efficient Compression for Personalized Text-to-Image Models
by: Zhang, Tangyuan, et al.
Published: (2025)
by: Zhang, Tangyuan, et al.
Published: (2025)
RoadTones: Tone Controllable Text Generation from Road Event Videos
by: Parikh, Chirag, et al.
Published: (2026)
by: Parikh, Chirag, et al.
Published: (2026)
MOC: Meta-Optimized Classifier for Few-Shot Whole Slide Image Classification
by: Xiang, Tianqi, et al.
Published: (2025)
by: Xiang, Tianqi, et al.
Published: (2025)
SGEdit: Bridging LLM with Text2Image Generative Model for Scene Graph-based Image Editing
by: Zhang, Zhiyuan, et al.
Published: (2024)
by: Zhang, Zhiyuan, et al.
Published: (2024)
True to Tone? Quantifying Skin Tone Fidelity and Bias in Photographic-to-Virtual Human Pipelines
by: Schneider, Gabriel Ferri, et al.
Published: (2026)
by: Schneider, Gabriel Ferri, et al.
Published: (2026)
Toward High-Fidelity Visual Reconstruction: From EEG-Based Conditioned Generation to Joint-Modal Guided Rebuilding
by: Gong, Zhijian, et al.
Published: (2026)
by: Gong, Zhijian, et al.
Published: (2026)
Burst Image Super-Resolution via Multi-Cross Attention Encoding and Multi-Scan State-Space Decoding
by: Huang, Tengda, et al.
Published: (2025)
by: Huang, Tengda, et al.
Published: (2025)
Colorimeter-Supervised Skin Tone Estimation from Dermatoscopic Images for Fairness Auditing
by: Benčević, Marin, et al.
Published: (2026)
by: Benčević, Marin, et al.
Published: (2026)
Expandable Residual Approximation for Knowledge Distillation
by: Yan, Zhaoyi, et al.
Published: (2025)
by: Yan, Zhaoyi, et al.
Published: (2025)
Spatial Transform Decoupling for Oriented Object Detection
by: Yu, Hongtian, et al.
Published: (2023)
by: Yu, Hongtian, et al.
Published: (2023)
Self-supervised Feature-Gate Coupling for Dynamic Network Pruning
by: Shi, Mengnan, et al.
Published: (2021)
by: Shi, Mengnan, et al.
Published: (2021)
VideoAnchor: Reinforcing Subspace-Structured Visual Cues for Coherent Visual-Spatial Reasoning
by: Wang, Zhaozhi, et al.
Published: (2025)
by: Wang, Zhaozhi, et al.
Published: (2025)
Depth-guided Texture Diffusion for Image Semantic Segmentation
by: Sun, Wei, et al.
Published: (2024)
by: Sun, Wei, et al.
Published: (2024)
AsyncDSB: Schedule-Asynchronous Diffusion Schrödinger Bridge for Image Inpainting
by: Han, Zihao, et al.
Published: (2024)
by: Han, Zihao, et al.
Published: (2024)
Painting with Words: Elevating Detailed Image Captioning with Benchmark and Alignment Learning
by: Ye, Qinghao, et al.
Published: (2025)
by: Ye, Qinghao, et al.
Published: (2025)
Distribution-aware Interactive Attention Network and Large-scale Cloud Recognition Benchmark on FY-4A Satellite Image
by: Zhang, Jiaqing, et al.
Published: (2024)
by: Zhang, Jiaqing, et al.
Published: (2024)
Preserving Silent Features for Domain Generalization
by: Zhao, Chujie, et al.
Published: (2024)
by: Zhao, Chujie, et al.
Published: (2024)
Integrating View Conditions for Image Synthesis
by: Bai, Jinbin, et al.
Published: (2023)
by: Bai, Jinbin, et al.
Published: (2023)
ColorSwap: A Color and Word Order Dataset for Multimodal Evaluation
by: Burapacheep, Jirayu, et al.
Published: (2024)
by: Burapacheep, Jirayu, et al.
Published: (2024)
3D Shape Augmentation with Content-Aware Shape Resizing
by: Chen, Mingxiang, et al.
Published: (2024)
by: Chen, Mingxiang, et al.
Published: (2024)
NeuroPump: Simultaneous Geometric and Color Rectification for Underwater Images
by: Guo, Yue, et al.
Published: (2024)
by: Guo, Yue, et al.
Published: (2024)
YOLOv12: Attention-Centric Real-Time Object Detectors
by: Tian, Yunjie, et al.
Published: (2025)
by: Tian, Yunjie, et al.
Published: (2025)
OpenView: Empowering MLLMs with Out-of-view VQA
by: Chen, Qixiang, et al.
Published: (2025)
by: Chen, Qixiang, et al.
Published: (2025)
Modeling Image Tone Dichotomy with the Power Function
by: Martinez, Axel, et al.
Published: (2024)
by: Martinez, Axel, et al.
Published: (2024)
Tone Matters: The Impact of Linguistic Tone on Hallucination in VLMs
by: Hong, Weihao, et al.
Published: (2026)
by: Hong, Weihao, et al.
Published: (2026)
WordVIS: A Color Worth A Thousand Words
by: Khan, Umar, et al.
Published: (2024)
by: Khan, Umar, et al.
Published: (2024)
Video Color Grading via Look-Up Table Generation
by: Shin, Seunghyun, et al.
Published: (2025)
by: Shin, Seunghyun, et al.
Published: (2025)
Similar Items
-
ReDDiT: Rehashing Noise for Discrete Visual Generation
by: Ma, Tianren, et al.
Published: (2025) -
ClawMachine: Learning to Fetch Visual Tokens for Referential Comprehension
by: Ma, Tianren, et al.
Published: (2024) -
DynRefer: Delving into Region-level Multimodal Tasks via Dynamic Resolution
by: Zhao, Yuzhong, et al.
Published: (2024) -
Evaluation of Text-to-Video Generation Models: A Dynamics Perspective
by: Liao, Mingxiang, et al.
Published: (2024) -
ChatterBox: Multi-round Multimodal Referring and Grounding
by: Tian, Yunjie, et al.
Published: (2024)