Online Anchor-based Training for Image Classification Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Tzelepi, Maria, Mezaris, Vasileios |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LMM-Regularized CLIP Embeddings for Image Classification
by: Tzelepi, Maria, et al.
Published: (2024)
by: Tzelepi, Maria, et al.
Published: (2024)
Disturbing Image Detection Using LMM-Elicited Emotion Embeddings
by: Tzelepi, Maria, et al.
Published: (2024)
by: Tzelepi, Maria, et al.
Published: (2024)
Improving Multimodal Hateful Meme Detection Exploiting LMM-Generated Knowledge
by: Tzelepi, Maria, et al.
Published: (2025)
by: Tzelepi, Maria, et al.
Published: (2025)
Exploiting LMM-based knowledge for image classification tasks
by: Tzelepi, Maria, et al.
Published: (2024)
by: Tzelepi, Maria, et al.
Published: (2024)
P-TAME: Explain Any Image Classifier with Trained Perturbations
by: Ntrougkas, Mariano V., et al.
Published: (2025)
by: Ntrougkas, Mariano V., et al.
Published: (2025)
B-FPGM: Lightweight Face Detection via Bayesian-Optimized Soft FPGM Pruning
by: Kaparinos, Nikolaos, et al.
Published: (2025)
by: Kaparinos, Nikolaos, et al.
Published: (2025)
SDAKD: Student Discriminator Assisted Knowledge Distillation for Super-Resolution Generative Adversarial Networks
by: Kaparinos, Nikolaos, et al.
Published: (2025)
by: Kaparinos, Nikolaos, et al.
Published: (2025)
MMFusion: Combining Image Forensic Filters for Visual Manipulation Detection and Localization
by: Triaridis, Kostas, et al.
Published: (2023)
by: Triaridis, Kostas, et al.
Published: (2023)
A Human-Annotated Video Dataset for Training and Evaluation of 360-Degree Video Summarization Methods
by: Kontostathis, Ioannis, et al.
Published: (2024)
by: Kontostathis, Ioannis, et al.
Published: (2024)
LLaVA-CKD: Bottom-Up Cascaded Knowledge Distillation for Vision-Language Models
by: Gkalelis, Nikolaos, et al.
Published: (2026)
by: Gkalelis, Nikolaos, et al.
Published: (2026)
Sens-VisualNews: A Benchmark Dataset for Sensational Image Detection
by: Goulas, Andreas, et al.
Published: (2026)
by: Goulas, Andreas, et al.
Published: (2026)
TAME: Attention Mechanism Based Feature Fusion for Generating Explanation Maps of Convolutional Neural Networks
by: Ntrougkas, Mariano, et al.
Published: (2023)
by: Ntrougkas, Mariano, et al.
Published: (2023)
TSalV360: A Method and Dataset for Text-driven Saliency Detection in 360-Degrees Videos
by: Kontostathis, Ioannis, et al.
Published: (2025)
by: Kontostathis, Ioannis, et al.
Published: (2025)
VidCtx: Context-aware Video Question Answering with Image Models
by: Goulas, Andreas, et al.
Published: (2024)
by: Goulas, Andreas, et al.
Published: (2024)
An Integrated Framework for Multi-Granular Explanation of Video Summarization
by: Tsigos, Konstantinos, et al.
Published: (2024)
by: Tsigos, Konstantinos, et al.
Published: (2024)
An Experimental Study on Generating Plausible Textual Explanations for Video Summarization
by: Eleftheriadis, Thomas, et al.
Published: (2025)
by: Eleftheriadis, Thomas, et al.
Published: (2025)
SD-MVSum: Script-Driven Multimodal Video Summarization Method and Datasets
by: Mylonas, Manolis, et al.
Published: (2025)
by: Mylonas, Manolis, et al.
Published: (2025)
SD-VSum: A Method and Dataset for Script-Driven Video Summarization
by: Mylonas, Manolis, et al.
Published: (2025)
by: Mylonas, Manolis, et al.
Published: (2025)
Improving the Perturbation-Based Explanation of Deepfake Detectors Through the Use of Adversarially-Generated Samples
by: Tsigos, Konstantinos, et al.
Published: (2025)
by: Tsigos, Konstantinos, et al.
Published: (2025)
T-TAME: Trainable Attention Mechanism for Explaining Convolutional Networks and Vision Transformers
by: Ntrougkas, Mariano V., et al.
Published: (2024)
by: Ntrougkas, Mariano V., et al.
Published: (2024)
Towards Quantitative Evaluation of Explainable AI Methods for Deepfake Detection
by: Tsigos, Konstantinos, et al.
Published: (2024)
by: Tsigos, Konstantinos, et al.
Published: (2024)
Visual and audio scene classification for detecting discrepancies in video: a baseline method and experimental protocol
by: Apostolidis, Konstantinos, et al.
Published: (2024)
by: Apostolidis, Konstantinos, et al.
Published: (2024)
Interpretable Vision Transformers in Image Classification via SVDA
by: Arampatzakis, Vasileios, et al.
Published: (2026)
by: Arampatzakis, Vasileios, et al.
Published: (2026)
Anchor Token Matching: Implicit Structure Locking for Training-free AR Image Editing
by: Hu, Taihang, et al.
Published: (2025)
by: Hu, Taihang, et al.
Published: (2025)
AnchorFlow: Training-Free 3D Editing via Latent Anchor-Aligned Flows
by: Zhou, Zhenglin, et al.
Published: (2025)
by: Zhou, Zhenglin, et al.
Published: (2025)
HAT: History-Augmented Anchor Transformer for Online Temporal Action Localization
by: Reza, Sakib, et al.
Published: (2024)
by: Reza, Sakib, et al.
Published: (2024)
VALA: Learning Latent Anchors for Training-Free and Temporally Consistent
by: Wu, Zhangkai, et al.
Published: (2025)
by: Wu, Zhangkai, et al.
Published: (2025)
TEMA: Anchor the Image, Follow the Text for Multi-Modification Composed Image Retrieval
by: Li, Zixu, et al.
Published: (2026)
by: Li, Zixu, et al.
Published: (2026)
Zero-Training Task-Specific Model Synthesis for Few-Shot Medical Image Classification
by: Qin, Yao, et al.
Published: (2025)
by: Qin, Yao, et al.
Published: (2025)
Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting
by: Feng, Hao, et al.
Published: (2025)
by: Feng, Hao, et al.
Published: (2025)
Many-Task Federated Fine-Tuning via Unified Task Vectors
by: Tsouvalas, Vasileios, et al.
Published: (2025)
by: Tsouvalas, Vasileios, et al.
Published: (2025)
Anchor-based Robust Finetuning of Vision-Language Models
by: Han, Jinwei, et al.
Published: (2024)
by: Han, Jinwei, et al.
Published: (2024)
AnchorHOI: Zero-shot Generation of 4D Human-Object Interaction via Anchor-based Prior Distillation
by: Dai, Sisi, et al.
Published: (2025)
by: Dai, Sisi, et al.
Published: (2025)
Liberating Seen Classes: Boosting Few-Shot and Zero-Shot Text Classification via Anchor Generation and Classification Reframing
by: Liu, Han, et al.
Published: (2024)
by: Liu, Han, et al.
Published: (2024)
Iterative Online Image Synthesis via Diffusion Model for Imbalanced Classification
by: Li, Shuhan, et al.
Published: (2024)
by: Li, Shuhan, et al.
Published: (2024)
Online Class-Incremental Learning For Real-World Food Image Classification
by: Raghavan, Siddeshwar, et al.
Published: (2023)
by: Raghavan, Siddeshwar, et al.
Published: (2023)
SlideGCD: Slide-based Graph Collaborative Training with Knowledge Distillation for Whole Slide Image Classification
by: Shu, Tong, et al.
Published: (2024)
by: Shu, Tong, et al.
Published: (2024)
LDTR: Transformer-based Lane Detection with Anchor-chain Representation
by: Yang, Zhongyu, et al.
Published: (2024)
by: Yang, Zhongyu, et al.
Published: (2024)
AnchorDiff: Training-Free Concept Grounding for MM-DiTs via Anchor-Based Graph Propagation
by: Zhang, Jian, et al.
Published: (2026)
by: Zhang, Jian, et al.
Published: (2026)
Envision3D: One Image to 3D with Anchor Views Interpolation
by: Pang, Yatian, et al.
Published: (2024)
by: Pang, Yatian, et al.
Published: (2024)
Similar Items
-
LMM-Regularized CLIP Embeddings for Image Classification
by: Tzelepi, Maria, et al.
Published: (2024) -
Disturbing Image Detection Using LMM-Elicited Emotion Embeddings
by: Tzelepi, Maria, et al.
Published: (2024) -
Improving Multimodal Hateful Meme Detection Exploiting LMM-Generated Knowledge
by: Tzelepi, Maria, et al.
Published: (2025) -
Exploiting LMM-based knowledge for image classification tasks
by: Tzelepi, Maria, et al.
Published: (2024) -
P-TAME: Explain Any Image Classifier with Trained Perturbations
by: Ntrougkas, Mariano V., et al.
Published: (2025)