KNN Transformer with Pyramid Prompts for Few-Shot Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Wenhao, Wang, Qiangchang, Zhao, Peng, Yin, Yilong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DVLA-RL: Dual-Level Vision-Language Alignment with Reinforcement Learning Gating for Few-Shot Learning
by: Li, Wenhao, et al.
Published: (2026)
by: Li, Wenhao, et al.
Published: (2026)
VT-FSL: Bridging Vision and Text with LLMs for Few-Shot Learning
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
Doodle Your Keypoints: Sketch-Based Few-Shot Keypoint Detection
by: Maity, Subhajit, et al.
Published: (2025)
by: Maity, Subhajit, et al.
Published: (2025)
Automated Discontinuity Set Characterisation in Enclosed Rock Face Point Clouds Using Single-Shot Filtering and Cyclic Orientation Transformation
by: Patra, Dibyayan, et al.
Published: (2026)
by: Patra, Dibyayan, et al.
Published: (2026)
SETR: A Two-Stage Semantic-Enhanced Framework for Zero-Shot Composed Image Retrieval
by: Xiao, Yuqi, et al.
Published: (2025)
by: Xiao, Yuqi, et al.
Published: (2025)
AVadCLIP: Audio-Visual Collaboration for Robust Video Anomaly Detection
by: Wu, Peng, et al.
Published: (2025)
by: Wu, Peng, et al.
Published: (2025)
DVGBench: Implicit-to-Explicit Visual Grounding Benchmark in UAV Imagery with Large Vision-Language Models
by: Zhou, Yue, et al.
Published: (2026)
by: Zhou, Yue, et al.
Published: (2026)
CMAB: A First National-Scale Multi-Attribute Building Dataset in China Derived from Open Source Data and GeoAI
by: Zhang, Yecheng, et al.
Published: (2024)
by: Zhang, Yecheng, et al.
Published: (2024)
Supersampling of Data from Structured-light Scanner with Deep Learning
by: Melicherčík, Martin, et al.
Published: (2023)
by: Melicherčík, Martin, et al.
Published: (2023)
Efficient Diffusion Models: A Comprehensive Survey from Principles to Practices
by: Ma, Zhiyuan, et al.
Published: (2024)
by: Ma, Zhiyuan, et al.
Published: (2024)
Consistency-Driven Calibration and Matching for Few-Shot Class-Incremental Learning
by: Wang, Qinzhe, et al.
Published: (2025)
by: Wang, Qinzhe, et al.
Published: (2025)
Processing and Segmentation of Human Teeth from 2D Images using Weakly Supervised Learning
by: Kunzo, Tomáš, et al.
Published: (2023)
by: Kunzo, Tomáš, et al.
Published: (2023)
TimeCausality: Evaluating the Causal Ability in Time Dimension for Vision Language Models
by: Wang, Zeqing, et al.
Published: (2025)
by: Wang, Zeqing, et al.
Published: (2025)
A Deep Learning Approach to Identify Rock Bolts in Complex 3D Point Clouds of Underground Mines Captured Using Mobile Laser Scanners
by: Patra, Dibyayan, et al.
Published: (2025)
by: Patra, Dibyayan, et al.
Published: (2025)
SkeletonX: Data-Efficient Skeleton-based Action Recognition via Cross-sample Feature Aggregation
by: Zhang, Zongye, et al.
Published: (2025)
by: Zhang, Zongye, et al.
Published: (2025)
Learning to count small and clustered objects with application to bacterial colonies
by: Zheng, Minghua, et al.
Published: (2026)
by: Zheng, Minghua, et al.
Published: (2026)
SAM Encoder Breach by Adversarial Simplicial Complex Triggers Downstream Model Failures
by: Qin, Yi, et al.
Published: (2025)
by: Qin, Yi, et al.
Published: (2025)
MCA-Bench: A Multimodal Benchmark for Evaluating CAPTCHA Robustness Against VLM-based Attacks
by: Wu, Zonglin, et al.
Published: (2025)
by: Wu, Zonglin, et al.
Published: (2025)
Decoupled Sensitivity-Consistency Learning for Weakly Supervised Video Anomaly Detection
by: Zheng, Hantao, et al.
Published: (2026)
by: Zheng, Hantao, et al.
Published: (2026)
EarthVL: A Progressive Earth Vision-Language Understanding and Generation Framework
by: Wang, Junjue, et al.
Published: (2026)
by: Wang, Junjue, et al.
Published: (2026)
Robust Multi-Source Covid-19 Detection in CT Images
by: Pritha, Asmita Yuki, et al.
Published: (2026)
by: Pritha, Asmita Yuki, et al.
Published: (2026)
DisasterM3: A Remote Sensing Vision-Language Dataset for Disaster Damage Assessment and Response
by: Wang, Junjue, et al.
Published: (2025)
by: Wang, Junjue, et al.
Published: (2025)
DCT-HistoTransformer: Efficient Lightweight Vision Transformer with DCT Integration for histopathological image analysis
by: Ranjbar, Mahtab, et al.
Published: (2024)
by: Ranjbar, Mahtab, et al.
Published: (2024)
Group Activity Recognition using Unreliable Tracked Pose
by: Thilakarathne, Haritha, et al.
Published: (2024)
by: Thilakarathne, Haritha, et al.
Published: (2024)
Video-Based Human Pose Regression via Decoupled Space-Time Aggregation
by: He, Jijie, et al.
Published: (2024)
by: He, Jijie, et al.
Published: (2024)
EUFCC-340K: A Faceted Hierarchical Dataset for Metadata Annotation in GLAM Collections
by: Net, Francesc, et al.
Published: (2024)
by: Net, Francesc, et al.
Published: (2024)
Orientation-conditioned Facial Texture Mapping for Video-based Facial Remote Photoplethysmography Estimation
by: Cantrill, Sam, et al.
Published: (2024)
by: Cantrill, Sam, et al.
Published: (2024)
TDIP: Tunable Deep Image Processing, a Real Time Melt Pool Monitoring Solution
by: Akhavan, Javid, et al.
Published: (2024)
by: Akhavan, Javid, et al.
Published: (2024)
Dynamic Brightness Adaptation for Robust Multi-modal Image Fusion
by: Sun, Yiming, et al.
Published: (2024)
by: Sun, Yiming, et al.
Published: (2024)
Scalable and Realistic Virtual Try-on Application for Foundation Makeup with Kubelka-Munk Theory
by: Pang, Hui, et al.
Published: (2025)
by: Pang, Hui, et al.
Published: (2025)
Cost Savings from Automatic Quality Assessment of Generated Images
by: Giro-i-Nieto, Xavier, et al.
Published: (2025)
by: Giro-i-Nieto, Xavier, et al.
Published: (2025)
Deepfake Detection Generalization with Diffusion Noise
by: Qi, Hongyuan, et al.
Published: (2026)
by: Qi, Hongyuan, et al.
Published: (2026)
Towards Integrated Rock Support Visualisation in 3D Point Cloud of Underground Mines
by: Patra, Dibyayan, et al.
Published: (2026)
by: Patra, Dibyayan, et al.
Published: (2026)
HyperFM: An Efficient Hyperspectral Foundation Model with Spectral Grouping
by: Tushar, Zahid Hassan, et al.
Published: (2026)
by: Tushar, Zahid Hassan, et al.
Published: (2026)
Cross-View-Prediction: Exploring Contrastive Feature for Hyperspectral Image Classification
by: Zhang, Anyu, et al.
Published: (2022)
by: Zhang, Anyu, et al.
Published: (2022)
Facial Spatiotemporal Graphs: Leveraging the 3D Facial Surface for Remote Physiological Measurement
by: Cantrill, Sam, et al.
Published: (2026)
by: Cantrill, Sam, et al.
Published: (2026)
Exploring Diffusion with Test-Time Training on Efficient Image Restoration
by: Lu, Rongchang, et al.
Published: (2025)
by: Lu, Rongchang, et al.
Published: (2025)
Tiny-YOLOSAM: Fast Hybrid Image Segmentation
by: Xu, Kenneth, et al.
Published: (2025)
by: Xu, Kenneth, et al.
Published: (2025)
Skeletonization-Based Adversarial Perturbations on Large Vision Language Model's Mathematical Text Recognition
by: Yoshida, Masatomo, et al.
Published: (2026)
by: Yoshida, Masatomo, et al.
Published: (2026)
Evaluating the Significance of Outdoor Advertising from Driver's Perspective Using Computer Vision
by: Černeková, Zuzana, et al.
Published: (2023)
by: Černeková, Zuzana, et al.
Published: (2023)
Similar Items
-
DVLA-RL: Dual-Level Vision-Language Alignment with Reinforcement Learning Gating for Few-Shot Learning
by: Li, Wenhao, et al.
Published: (2026) -
VT-FSL: Bridging Vision and Text with LLMs for Few-Shot Learning
by: Li, Wenhao, et al.
Published: (2025) -
Doodle Your Keypoints: Sketch-Based Few-Shot Keypoint Detection
by: Maity, Subhajit, et al.
Published: (2025) -
Automated Discontinuity Set Characterisation in Enclosed Rock Face Point Clouds Using Single-Shot Filtering and Cyclic Orientation Transformation
by: Patra, Dibyayan, et al.
Published: (2026) -
SETR: A Two-Stage Semantic-Enhanced Framework for Zero-Shot Composed Image Retrieval
by: Xiao, Yuqi, et al.
Published: (2025)