Charm: The Missing Piece in ViT fine-tuning for Image Aesthetic Assessment
Fuente:
arXiv
Saved in:
| Main Authors: | Behrad, Fatemeh, Tuytelaars, Tinne, Wagemans, Johan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Role of Individual Differences in Current Approaches to Computational Image Aesthetics
by: Chen, Li-Wei, et al.
Published: (2025)
by: Chen, Li-Wei, et al.
Published: (2025)
From Concepts to Judgments: Interpretable Image Aesthetic Assessment
by: Liu, Xiao-Chang, et al.
Published: (2026)
by: Liu, Xiao-Chang, et al.
Published: (2026)
BackFlip: The Impact of Local and Global Data Augmentations on Artistic Image Aesthetic Assessment
by: Strafforello, Ombretta, et al.
Published: (2024)
by: Strafforello, Ombretta, et al.
Published: (2024)
PEO: Training-Free Aesthetic Quality Enhancement in Pre-Trained Text-to-Image Diffusion Models with Prompt Embedding Optimization
by: Margaryan, Hovhannes, et al.
Published: (2025)
by: Margaryan, Hovhannes, et al.
Published: (2025)
Finding Closure: A Closer Look at the Gestalt Law of Closure in Convolutional Neural Networks
by: Zhang, Yuyan, et al.
Published: (2024)
by: Zhang, Yuyan, et al.
Published: (2024)
Investigating the Gestalt Principle of Closure in Deep Convolutional Neural Networks
by: Zhang, Yuyan, et al.
Published: (2024)
by: Zhang, Yuyan, et al.
Published: (2024)
Implicit Gaussian Splatting with Efficient Multi-Level Tri-Plane Representation
by: Wu, Minye, et al.
Published: (2024)
by: Wu, Minye, et al.
Published: (2024)
Analysis of Spatial augmentation in Self-supervised models in the purview of training and test distributions
by: Jha, Abhishek, et al.
Published: (2024)
by: Jha, Abhishek, et al.
Published: (2024)
Visually-Aware Context Modeling for News Image Captioning
by: Qu, Tingyu, et al.
Published: (2023)
by: Qu, Tingyu, et al.
Published: (2023)
Remembering by Reconstructing: Domain Incremental Learning With Test-Time Training on Video Streams
by: Swinnen, Jonathan, et al.
Published: (2026)
by: Swinnen, Jonathan, et al.
Published: (2026)
Pretrained ViTs Yield Versatile Representations For Medical Images
by: Matsoukas, Christos, et al.
Published: (2023)
by: Matsoukas, Christos, et al.
Published: (2023)
DM-Align: Leveraging the Power of Natural Language Instructions to Make Changes to Images
by: Trusca, Maria Mihaela, et al.
Published: (2024)
by: Trusca, Maria Mihaela, et al.
Published: (2024)
RGS-DR: Deferred Reflections and Residual Shading in 2D Gaussian Splatting
by: Kouros, Georgios, et al.
Published: (2025)
by: Kouros, Georgios, et al.
Published: (2025)
Object-Centric Pretraining via Target Encoder Bootstrapping
by: Đukić, Nikola, et al.
Published: (2025)
by: Đukić, Nikola, et al.
Published: (2025)
Animate Your Motion: Turning Still Images into Dynamic Videos
by: Li, Mingxiao, et al.
Published: (2024)
by: Li, Mingxiao, et al.
Published: (2024)
Towards More Accurate Personalized Image Generation: Addressing Overfitting and Evaluation Bias
by: Li, Mingxiao, et al.
Published: (2025)
by: Li, Mingxiao, et al.
Published: (2025)
Spec-Gloss Surfels and Normal-Diffuse Priors for Relightable Glossy Objects
by: Kouros, Georgios, et al.
Published: (2025)
by: Kouros, Georgios, et al.
Published: (2025)
Introducing Routing Functions to Vision-Language Parameter-Efficient Fine-Tuning with Low-Rank Bottlenecks
by: Qu, Tingyu, et al.
Published: (2024)
by: Qu, Tingyu, et al.
Published: (2024)
Your ViT is Secretly an Image Segmentation Model
by: Kerssies, Tommie, et al.
Published: (2025)
by: Kerssies, Tommie, et al.
Published: (2025)
Same accuracy, twice as fast: continuous training surpasses retraining from scratch
by: Verwimp, Eli, et al.
Published: (2025)
by: Verwimp, Eli, et al.
Published: (2025)
Eff-GRot: Efficient and Generalizable Rotation Estimation with Transformers
by: Mathioulakis, Fanis, et al.
Published: (2025)
by: Mathioulakis, Fanis, et al.
Published: (2025)
Parameter Efficient Fine-tuning of Self-supervised ViTs without Catastrophic Forgetting
by: Bafghi, Reza Akbarian, et al.
Published: (2024)
by: Bafghi, Reza Akbarian, et al.
Published: (2024)
Deeper Inside Deep ViT
by: Hong, Sungrae
Published: (2025)
by: Hong, Sungrae
Published: (2025)
Multi-task convolutional neural network for image aesthetic assessment
by: Soydaner, Derya, et al.
Published: (2023)
by: Soydaner, Derya, et al.
Published: (2023)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
by: Zhong, Yunshan, et al.
Published: (2023)
by: Zhong, Yunshan, et al.
Published: (2023)
Unsupervised Parameter Efficient Source-free Post-pretraining
by: Jha, Abhishek, et al.
Published: (2025)
by: Jha, Abhishek, et al.
Published: (2025)
IML-ViT: Benchmarking Image Manipulation Localization by Vision Transformer
by: Ma, Xiaochen, et al.
Published: (2023)
by: Ma, Xiaochen, et al.
Published: (2023)
Let ViT Speak: Generative Language-Image Pre-training
by: Fang, Yan, et al.
Published: (2026)
by: Fang, Yan, et al.
Published: (2026)
Forgetting of task-specific knowledge in model merging-based continual learning
by: Hess, Timm, et al.
Published: (2025)
by: Hess, Timm, et al.
Published: (2025)
TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models
by: Qu, Tingyu, et al.
Published: (2024)
by: Qu, Tingyu, et al.
Published: (2024)
DFQ-ViT: Data-Free Quantization for Vision Transformers without Fine-tuning
by: Tong, Yujia, et al.
Published: (2025)
by: Tong, Yujia, et al.
Published: (2025)
RepViT: Revisiting Mobile CNN From ViT Perspective
by: Wang, Ao, et al.
Published: (2023)
by: Wang, Ao, et al.
Published: (2023)
ViT-FIQA: Assessing Face Image Quality using Vision Transformers
by: Atzori, Andrea, et al.
Published: (2025)
by: Atzori, Andrea, et al.
Published: (2025)
DC-ViT: Modulating Spatial and Channel Interactions for Multi-Channel Images
by: Marikkar, Umar, et al.
Published: (2026)
by: Marikkar, Umar, et al.
Published: (2026)
Is this chart lying to me? Automating the detection of misleading visualizations
by: Tonglet, Jonathan, et al.
Published: (2025)
by: Tonglet, Jonathan, et al.
Published: (2025)
From Missing Pieces to Masterpieces: Image Completion with Context-Adaptive Diffusion
by: Shamsolmoali, Pourya, et al.
Published: (2025)
by: Shamsolmoali, Pourya, et al.
Published: (2025)
OASIS: Online Sample Selection for Continual Visual Instruction Tuning
by: Lee, Minjae, et al.
Published: (2025)
by: Lee, Minjae, et al.
Published: (2025)
Unveiling the Ambiguity in Neural Inverse Rendering: A Parameter Compensation Analysis
by: Kouros, Georgios, et al.
Published: (2024)
by: Kouros, Georgios, et al.
Published: (2024)
Implicit to Explicit Entropy Regularization: Benchmarking ViT Fine-tuning under Noisy Labels
by: Marrium, Maria, et al.
Published: (2024)
by: Marrium, Maria, et al.
Published: (2024)
Rethinking Random Masking in Self-Distillation on ViT
by: Seong, Jihyeon, et al.
Published: (2025)
by: Seong, Jihyeon, et al.
Published: (2025)
Similar Items
-
On the Role of Individual Differences in Current Approaches to Computational Image Aesthetics
by: Chen, Li-Wei, et al.
Published: (2025) -
From Concepts to Judgments: Interpretable Image Aesthetic Assessment
by: Liu, Xiao-Chang, et al.
Published: (2026) -
BackFlip: The Impact of Local and Global Data Augmentations on Artistic Image Aesthetic Assessment
by: Strafforello, Ombretta, et al.
Published: (2024) -
PEO: Training-Free Aesthetic Quality Enhancement in Pre-Trained Text-to-Image Diffusion Models with Prompt Embedding Optimization
by: Margaryan, Hovhannes, et al.
Published: (2025) -
Finding Closure: A Closer Look at the Gestalt Law of Closure in Convolutional Neural Networks
by: Zhang, Yuyan, et al.
Published: (2024)