HanMoVLM: Large Vision-Language Models for Professional Artistic Painting Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Hongji, Zhou, Yucheng, Han, Wencheng, Li, Songlian, Zhao, Xiaotong, Shen, Jianbing |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Self-Rewarding Large Vision-Language Models for Optimizing Prompts in Text-to-Image Generation
by: Yang, Hongji, et al.
Published: (2025)
by: Yang, Hongji, et al.
Published: (2025)
DC-ControlNet: Decoupling Inter- and Intra-Element Conditions in Image Generation with Diffusion Models
by: Yang, Hongji, et al.
Published: (2025)
by: Yang, Hongji, et al.
Published: (2025)
CogOmniControl: Reasoning-Driven Controllable Video Generation via Creative Intent Cognition
by: Yang, Hongji, et al.
Published: (2026)
by: Yang, Hongji, et al.
Published: (2026)
HiCoGen: Hierarchical Compositional Text-to-Image Generation in Diffusion Models via Reinforcement Learning
by: Yang, Hongji, et al.
Published: (2025)
by: Yang, Hongji, et al.
Published: (2025)
Visual In-Context Learning for Large Vision-Language Models
by: Zhou, Yucheng, et al.
Published: (2024)
by: Zhou, Yucheng, et al.
Published: (2024)
Improving Medical Large Vision-Language Models with Abnormal-Aware Feedback
by: Zhou, Yucheng, et al.
Published: (2025)
by: Zhou, Yucheng, et al.
Published: (2025)
High-Precision Self-Supervised Monocular Depth Estimation with Rich-Resource Prior
by: Han, Wencheng, et al.
Published: (2024)
by: Han, Wencheng, et al.
Published: (2024)
Rethinking Visual Dependency in Long-Context Reasoning for Large Vision-Language Models
by: Zhou, Yucheng, et al.
Published: (2024)
by: Zhou, Yucheng, et al.
Published: (2024)
PaintCopilot: Modeling Painting as Autonomous Artistic Continuation
by: Wen, Yunge, et al.
Published: (2026)
by: Wen, Yunge, et al.
Published: (2026)
Sim4Seg: Boosting Multimodal Multi-disease Medical Diagnosis Segmentation with Region-Aware Vision-Language Similarity Masks
by: Song, Lingran, et al.
Published: (2025)
by: Song, Lingran, et al.
Published: (2025)
Decoupling Fine Detail and Global Geometry for Compressed Depth Map Super-Resolution
by: Zheng, Huan, et al.
Published: (2024)
by: Zheng, Huan, et al.
Published: (2024)
RAWMamba: Unified sRGB-to-RAW De-rendering With State Space Model
by: Chen, Hongjun, et al.
Published: (2024)
by: Chen, Hongjun, et al.
Published: (2024)
AdaOcc: Adaptive Forward View Transformation and Flow Modeling for 3D Occupancy and Flow Prediction
by: Chen, Dubing, et al.
Published: (2024)
by: Chen, Dubing, et al.
Published: (2024)
Towards Geometry-Aware and Motion-Guided Video Human Mesh Recovery
by: Chen, Hongjun, et al.
Published: (2026)
by: Chen, Hongjun, et al.
Published: (2026)
Towards Better Cephalometric Landmark Detection with Diffusion Data Generation
by: Guo, Dongqian, et al.
Published: (2025)
by: Guo, Dongqian, et al.
Published: (2025)
RepVF: A Unified Vector Fields Representation for Multi-task 3D Perception
by: Li, Chunliang, et al.
Published: (2024)
by: Li, Chunliang, et al.
Published: (2024)
Artistic Intelligence: A Diffusion-Based Framework for High-Fidelity Landscape Painting Synthesis
by: Yang, Wanggong, et al.
Published: (2024)
by: Yang, Wanggong, et al.
Published: (2024)
Towards High-Fidelity 3D Portrait Generation with Rich Details by Cross-View Prior-Aware Diffusion
by: Wei, Haoran, et al.
Published: (2024)
by: Wei, Haoran, et al.
Published: (2024)
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models
by: Tian, Xiaoyu, et al.
Published: (2024)
by: Tian, Xiaoyu, et al.
Published: (2024)
DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving
by: Han, Wencheng, et al.
Published: (2024)
by: Han, Wencheng, et al.
Published: (2024)
From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation
by: Song, Han, et al.
Published: (2026)
by: Song, Han, et al.
Published: (2026)
Bridging Scene Generation and Planning: Driving with World Model via Unifying Vision and Motion Representation
by: Gui, Xingtai, et al.
Published: (2026)
by: Gui, Xingtai, et al.
Published: (2026)
PPJudge: Towards Human-Aligned Assessment of Artistic Painting Process
by: Jiang, Shiqi, et al.
Published: (2025)
by: Jiang, Shiqi, et al.
Published: (2025)
Q-VLM: Post-training Quantization for Large Vision-Language Models
by: Wang, Changyuan, et al.
Published: (2024)
by: Wang, Changyuan, et al.
Published: (2024)
VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models
by: Zhang, Jianke, et al.
Published: (2026)
by: Zhang, Jianke, et al.
Published: (2026)
Breaking Down Monocular Ambiguity: Exploiting Temporal Evolution for 3D Lane Detection
by: Zheng, Huan, et al.
Published: (2025)
by: Zheng, Huan, et al.
Published: (2025)
Language Prompt for Autonomous Driving
by: Wu, Dongming, et al.
Published: (2023)
by: Wu, Dongming, et al.
Published: (2023)
Condition Errors Refinement in Autoregressive Image Generation with Diffusion Loss
by: Zhou, Yucheng, et al.
Published: (2026)
by: Zhou, Yucheng, et al.
Published: (2026)
Paintings and Drawings Aesthetics Assessment with Rich Attributes for Various Artistic Categories
by: Jin, Xin, et al.
Published: (2024)
by: Jin, Xin, et al.
Published: (2024)
VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
by: Shen, Haozhan, et al.
Published: (2025)
by: Shen, Haozhan, et al.
Published: (2025)
Reducing CT Metal Artifacts by Learning Latent Space Alignment with Gemstone Spectral Imaging Data
by: Han, Wencheng, et al.
Published: (2025)
by: Han, Wencheng, et al.
Published: (2025)
FloorplanVLM: A Vision-Language Model for Floorplan Vectorization
by: Liu, Yuanqing, et al.
Published: (2026)
by: Liu, Yuanqing, et al.
Published: (2026)
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
by: Zeng, Zhitao, et al.
Published: (2025)
by: Zeng, Zhitao, et al.
Published: (2025)
RLGF: Reinforcement Learning with Geometric Feedback for Autonomous Driving Video Generation
by: Yan, Tianyi, et al.
Published: (2025)
by: Yan, Tianyi, et al.
Published: (2025)
RelationVLM: Making Large Vision-Language Models Understand Visual Relations
by: Huang, Zhipeng, et al.
Published: (2024)
by: Huang, Zhipeng, et al.
Published: (2024)
IIR-VLM: In-Context Instance-level Recognition for Large Vision-Language Models
by: Shi, Liang, et al.
Published: (2026)
by: Shi, Liang, et al.
Published: (2026)
PolarVLM: Bridging the Semantic-Physical Gap in Vision-Language Models
by: Li, Yuliang, et al.
Published: (2026)
by: Li, Yuliang, et al.
Published: (2026)
ArtGPT-4: Towards Artistic-understanding Large Vision-Language Models with Enhanced Adapter
by: Yuan, Zhengqing, et al.
Published: (2023)
by: Yuan, Zhengqing, et al.
Published: (2023)
ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models
by: Ye, Wencheng, et al.
Published: (2025)
by: Ye, Wencheng, et al.
Published: (2025)
SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information
by: Sun, Jiashuo, et al.
Published: (2024)
by: Sun, Jiashuo, et al.
Published: (2024)
Similar Items
-
Self-Rewarding Large Vision-Language Models for Optimizing Prompts in Text-to-Image Generation
by: Yang, Hongji, et al.
Published: (2025) -
DC-ControlNet: Decoupling Inter- and Intra-Element Conditions in Image Generation with Diffusion Models
by: Yang, Hongji, et al.
Published: (2025) -
CogOmniControl: Reasoning-Driven Controllable Video Generation via Creative Intent Cognition
by: Yang, Hongji, et al.
Published: (2026) -
HiCoGen: Hierarchical Compositional Text-to-Image Generation in Diffusion Models via Reinforcement Learning
by: Yang, Hongji, et al.
Published: (2025) -
Visual In-Context Learning for Large Vision-Language Models
by: Zhou, Yucheng, et al.
Published: (2024)