Salvato in:
| Autori principali: | Zhao, Yuanpei, Lin, Jie, Zhang, Chao, Wang, Yilin, Li, Mao, Li, Chenhui, Hou, Jie, Lv, Tangjie |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2605.19776 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CROP: Expert-Aligned Image Cropping via Compositional Reasoning and Optimizing Preference
di: Dong, Zhitong, et al.
Pubblicazione: (2026)
di: Dong, Zhitong, et al.
Pubblicazione: (2026)
FGM-HD: Boosting Generation Diversity of Fractal Generative Models through Hausdorff Dimension Induction
di: Zhang, Haowei, et al.
Pubblicazione: (2025)
di: Zhang, Haowei, et al.
Pubblicazione: (2025)
DGA-Net: Enhancing SAM with Depth Prompting and Graph-Anchor Guidance for Camouflaged Object Detection
di: Li, Yuetong, et al.
Pubblicazione: (2026)
di: Li, Yuetong, et al.
Pubblicazione: (2026)
AACP: Aesthetics assessment of children's paintings based on self-supervised learning
di: Jiang, Shiqi, et al.
Pubblicazione: (2024)
di: Jiang, Shiqi, et al.
Pubblicazione: (2024)
DreamFuse: Adaptive Image Fusion with Diffusion Transformer
di: Huang, Junjia, et al.
Pubblicazione: (2025)
di: Huang, Junjia, et al.
Pubblicazione: (2025)
StoryWeaver: A Unified World Model for Knowledge-Enhanced Story Character Customization
di: Zhang, Jinlu, et al.
Pubblicazione: (2024)
di: Zhang, Jinlu, et al.
Pubblicazione: (2024)
AesExpert: Towards Multi-modality Foundation Model for Image Aesthetics Perception
di: Huang, Yipo, et al.
Pubblicazione: (2024)
di: Huang, Yipo, et al.
Pubblicazione: (2024)
Aesthetic Image Captioning with Saliency Enhanced MLLMs
di: Tao, Yilin, et al.
Pubblicazione: (2025)
di: Tao, Yilin, et al.
Pubblicazione: (2025)
An Order-Complexity Aesthetic Assessment Model for Aesthetic-aware Music Recommendation
di: Jin, Xin, et al.
Pubblicazione: (2024)
di: Jin, Xin, et al.
Pubblicazione: (2024)
Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization
di: Liang, Zhanhao, et al.
Pubblicazione: (2024)
di: Liang, Zhanhao, et al.
Pubblicazione: (2024)
AesRM: Improving Video Aesthetics with Expert-Level Feedback
di: Han, Yujin, et al.
Pubblicazione: (2026)
di: Han, Yujin, et al.
Pubblicazione: (2026)
MAProtoNet: A Multi-scale Attentive Interpretable Prototypical Part Network for 3D Magnetic Resonance Imaging Brain Tumor Classification
di: Li, Binghua, et al.
Pubblicazione: (2024)
di: Li, Binghua, et al.
Pubblicazione: (2024)
TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering
di: Mao, Dongxing, et al.
Pubblicazione: (2026)
di: Mao, Dongxing, et al.
Pubblicazione: (2026)
Let Storytelling Tell Vivid Stories: An Expressive and Fluent Multimodal Storyteller
di: Zang, Chuanqi, et al.
Pubblicazione: (2024)
di: Zang, Chuanqi, et al.
Pubblicazione: (2024)
AesBench: An Expert Benchmark for Multimodal Large Language Models on Image Aesthetics Perception
di: Huang, Yipo, et al.
Pubblicazione: (2024)
di: Huang, Yipo, et al.
Pubblicazione: (2024)
Fixed Anchors Are Not Enough: Dynamic Retrieval and Persistent Homology for Dataset Distillation
di: Li, Muquan, et al.
Pubblicazione: (2026)
di: Li, Muquan, et al.
Pubblicazione: (2026)
Dual Expert Distillation Network for Generalized Zero-Shot Learning
di: Rao, Zhijie, et al.
Pubblicazione: (2024)
di: Rao, Zhijie, et al.
Pubblicazione: (2024)
AuG-KD: Anchor-Based Mixup Generation for Out-of-Domain Knowledge Distillation
di: Tang, Zihao, et al.
Pubblicazione: (2024)
di: Tang, Zihao, et al.
Pubblicazione: (2024)
AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation
di: Qian, Rui, et al.
Pubblicazione: (2026)
di: Qian, Rui, et al.
Pubblicazione: (2026)
Look Ma, No Ground Truth! Ground-Truth-Free Tuning of Structure from Motion and Visual SLAM
di: Fontan, Alejandro, et al.
Pubblicazione: (2024)
di: Fontan, Alejandro, et al.
Pubblicazione: (2024)
Amodal Ground Truth and Completion in the Wild
di: Zhan, Guanqi, et al.
Pubblicazione: (2023)
di: Zhan, Guanqi, et al.
Pubblicazione: (2023)
Distill, Diffuse, and Semanticize (DDS): Annotation-Free 3D Scene Understanding Based on Multi-Granularity Distillation and Graph-Diffusion-Based Segmentation
di: Wang, Yijing, et al.
Pubblicazione: (2026)
di: Wang, Yijing, et al.
Pubblicazione: (2026)
DebGCD: Debiased Learning with Distribution Guidance for Generalized Category Discovery
di: Liu, Yuanpei, et al.
Pubblicazione: (2025)
di: Liu, Yuanpei, et al.
Pubblicazione: (2025)
Fuse Before Transfer: Knowledge Fusion for Heterogeneous Distillation
di: Li, Guopeng, et al.
Pubblicazione: (2024)
di: Li, Guopeng, et al.
Pubblicazione: (2024)
Learning Cross-view Visual Geo-localization without Ground Truth
di: Li, Haoyuan, et al.
Pubblicazione: (2024)
di: Li, Haoyuan, et al.
Pubblicazione: (2024)
Robust Loss Functions for Object Grasping under Limited Ground Truth
di: Deng, Yangfan, et al.
Pubblicazione: (2024)
di: Deng, Yangfan, et al.
Pubblicazione: (2024)
Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding
di: Wang, Wanfu, et al.
Pubblicazione: (2025)
di: Wang, Wanfu, et al.
Pubblicazione: (2025)
FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing
di: Chen, Ze, et al.
Pubblicazione: (2026)
di: Chen, Ze, et al.
Pubblicazione: (2026)
From Priors to Perception: Grounding Video-LLMs in Physical Reality
di: Zhao, Zicheng, et al.
Pubblicazione: (2026)
di: Zhao, Zicheng, et al.
Pubblicazione: (2026)
Robust Mesh Saliency Ground Truth Acquisition in VR via View Cone Sampling and Manifold Diffusion
di: Zheng, Guoquan, et al.
Pubblicazione: (2026)
di: Zheng, Guoquan, et al.
Pubblicazione: (2026)
WaterMono: Teacher-Guided Anomaly Masking and Enhancement Boosting for Robust Underwater Self-Supervised Monocular Depth Estimation
di: Ding, Yilin, et al.
Pubblicazione: (2024)
di: Ding, Yilin, et al.
Pubblicazione: (2024)
TAR-TVG: Enhancing VLMs with Timestamp Anchor-Constrained Reasoning for Temporal Video Grounding
di: Guo, Chaohong, et al.
Pubblicazione: (2025)
di: Guo, Chaohong, et al.
Pubblicazione: (2025)
ArtiMuse: Fine-Grained Image Aesthetics Assessment with Joint Scoring and Expert-Level Understanding
di: Cao, Shuo, et al.
Pubblicazione: (2025)
di: Cao, Shuo, et al.
Pubblicazione: (2025)
Lightweight Contrastive Distilled Hashing for Online Cross-modal Retrieval
di: Li, Jiaxing, et al.
Pubblicazione: (2025)
di: Li, Jiaxing, et al.
Pubblicazione: (2025)
Diffusion-based Facial Aesthetics Enhancement with 3D Structure Guidance
di: Li, Lisha, et al.
Pubblicazione: (2025)
di: Li, Lisha, et al.
Pubblicazione: (2025)
PPJudge: Towards Human-Aligned Assessment of Artistic Painting Process
di: Jiang, Shiqi, et al.
Pubblicazione: (2025)
di: Jiang, Shiqi, et al.
Pubblicazione: (2025)
Adapting Fine-Grained Cross-View Localization to Areas without Fine Ground Truth
di: Xia, Zimin, et al.
Pubblicazione: (2024)
di: Xia, Zimin, et al.
Pubblicazione: (2024)
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding
di: Yang, Zuhao, et al.
Pubblicazione: (2025)
di: Yang, Zuhao, et al.
Pubblicazione: (2025)
Beyond the Ground Truth: Enhanced Supervision for Image Restoration
di: Ryou, Donghun, et al.
Pubblicazione: (2025)
di: Ryou, Donghun, et al.
Pubblicazione: (2025)
From Images to Detection: Machine Learning for Blood Pattern Classification
di: Li, Yilin, et al.
Pubblicazione: (2025)
di: Li, Yilin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CROP: Expert-Aligned Image Cropping via Compositional Reasoning and Optimizing Preference
di: Dong, Zhitong, et al.
Pubblicazione: (2026) -
FGM-HD: Boosting Generation Diversity of Fractal Generative Models through Hausdorff Dimension Induction
di: Zhang, Haowei, et al.
Pubblicazione: (2025) -
DGA-Net: Enhancing SAM with Depth Prompting and Graph-Anchor Guidance for Camouflaged Object Detection
di: Li, Yuetong, et al.
Pubblicazione: (2026) -
AACP: Aesthetics assessment of children's paintings based on self-supervised learning
di: Jiang, Shiqi, et al.
Pubblicazione: (2024) -
DreamFuse: Adaptive Image Fusion with Diffusion Transformer
di: Huang, Junjia, et al.
Pubblicazione: (2025)