Proto-CLIP: Vision-Language Prototypical Network for Few-Shot Learning
Fuente:
arXiv
Saved in:
| Main Authors: | P, Jishnu Jaykumar, Palanisamy, Kamalesh, Chao, Yu-Wei, Du, Xinya, Xiang, Yu |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adapting Pre-Trained Vision Models for Novel Instance Detection and Segmentation
by: Lu, Yangxiao, et al.
Published: (2024)
by: Lu, Yangxiao, et al.
Published: (2024)
SCENEREPLICA: Benchmarking Real-World Robot Manipulation by Creating Replicable Scenes
by: Khargonkar, Ninad, et al.
Published: (2023)
by: Khargonkar, Ninad, et al.
Published: (2023)
Generalizable Vision-Language Few-Shot Adaptation with Predictive Prompts and Negative Learning
by: Mandalika, Sriram
Published: (2025)
by: Mandalika, Sriram
Published: (2025)
Few-Shot Fruit Segmentation via Transfer Learning
by: James, Jordan A., et al.
Published: (2024)
by: James, Jordan A., et al.
Published: (2024)
Mechanistic Finetuning of Vision-Language-Action Models via Few-Shot Demonstrations
by: Mitra, Chancharik, et al.
Published: (2025)
by: Mitra, Chancharik, et al.
Published: (2025)
DOPE: Dual Object Perception-Enhancement Network for Vision-and-Language Navigation
by: Yu, Yinfeng, et al.
Published: (2025)
by: Yu, Yinfeng, et al.
Published: (2025)
Few-Shot Panoptic Segmentation With Foundation Models
by: Käppeler, Markus, et al.
Published: (2023)
by: Käppeler, Markus, et al.
Published: (2023)
RVT-2: Learning Precise Manipulation from Few Demonstrations
by: Goyal, Ankit, et al.
Published: (2024)
by: Goyal, Ankit, et al.
Published: (2024)
Automated Robotic Moisture Monitoring in Agricultural Fields
by: Palanisamy, Senthil, et al.
Published: (2026)
by: Palanisamy, Senthil, et al.
Published: (2026)
ProtoConNet: Prototypical Augmentation and Alignment for Open-Set Few-Shot Image Classification
by: Shi, Kexuan, et al.
Published: (2025)
by: Shi, Kexuan, et al.
Published: (2025)
IPNET:Influential Prototypical Networks for Few Shot Learning
by: Chowdhury, Ranjana Roy, et al.
Published: (2022)
by: Chowdhury, Ranjana Roy, et al.
Published: (2022)
UniPrototype: Humn-Robot Skill Learning with Uniform Prototypes
by: Hu, Xiao, et al.
Published: (2025)
by: Hu, Xiao, et al.
Published: (2025)
Constraint-Aware Zero-Shot Vision-Language Navigation in Continuous Environments
by: Chen, Kehan, et al.
Published: (2024)
by: Chen, Kehan, et al.
Published: (2024)
CLIP-Clique: Graph-based Correspondence Matching Augmented by Vision Language Models for Object-based Global Localization
by: Matsuzaki, Shigemichi, et al.
Published: (2024)
by: Matsuzaki, Shigemichi, et al.
Published: (2024)
ProMi: An Efficient Prototype-Mixture Baseline for Few-Shot Segmentation with Bounding-Box Annotations
by: Chiaroni, Florent, et al.
Published: (2025)
by: Chiaroni, Florent, et al.
Published: (2025)
Zero-Shot 3D Visual Grounding from Vision-Language Models
by: Li, Rong, et al.
Published: (2025)
by: Li, Rong, et al.
Published: (2025)
Unify the Views: View-Consistent Prototype Learning for Few-Shot Segmentation
by: Liu, Hongli, et al.
Published: (2026)
by: Liu, Hongli, et al.
Published: (2026)
Detecting Endangered Marine Species in Autonomous Underwater Vehicle Imagery Using Point Annotations and Few-Shot Learning
by: Doig, Heather, et al.
Published: (2024)
by: Doig, Heather, et al.
Published: (2024)
SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation
by: Shi, Xiangyu, et al.
Published: (2025)
by: Shi, Xiangyu, et al.
Published: (2025)
CLIP-guided Prototype Modulating for Few-shot Action Recognition
by: Wang, Xiang, et al.
Published: (2023)
by: Wang, Xiang, et al.
Published: (2023)
Calibrated Cache Model for Few-Shot Vision-Language Model Adaptation
by: Ding, Kun, et al.
Published: (2024)
by: Ding, Kun, et al.
Published: (2024)
Fast-SmartWay: Panoramic-Free End-to-End Zero-Shot Vision-and-Language Navigation
by: Shi, Xiangyu, et al.
Published: (2025)
by: Shi, Xiangyu, et al.
Published: (2025)
StreamingVLA: Streaming Vision-Language-Action Model with Action Flow Matching and Adaptive Early Observation
by: Shi, Yiran, et al.
Published: (2026)
by: Shi, Yiran, et al.
Published: (2026)
MPA: Multimodal Prototype Augmentation for Few-Shot Learning
by: Wu, Liwen, et al.
Published: (2026)
by: Wu, Liwen, et al.
Published: (2026)
Query-guided Prototype Evolution Network for Few-Shot Segmentation
by: Cong, Runmin, et al.
Published: (2024)
by: Cong, Runmin, et al.
Published: (2024)
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs
by: Fang, Yu, et al.
Published: (2026)
by: Fang, Yu, et al.
Published: (2026)
GSMem: 3D Gaussian Splatting as Persistent Spatial Memory for Zero-Shot Embodied Exploration and Reasoning
by: Lu, Yiren, et al.
Published: (2026)
by: Lu, Yiren, et al.
Published: (2026)
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
by: Dai, Tingjun, et al.
Published: (2026)
by: Dai, Tingjun, et al.
Published: (2026)
Robotic-CLIP: Fine-tuning CLIP on Action Data for Robotic Applications
by: Nguyen, Nghia, et al.
Published: (2024)
by: Nguyen, Nghia, et al.
Published: (2024)
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
by: Qiao, Yanyuan, et al.
Published: (2024)
by: Qiao, Yanyuan, et al.
Published: (2024)
Three-Step Nav: A Hierarchical Global-Local Planner for Zero-Shot Vision-and-Language Navigation
by: Zheng, Wanrong, et al.
Published: (2026)
by: Zheng, Wanrong, et al.
Published: (2026)
SpatialNav: Leveraging Spatial Scene Graphs for Zero-Shot Vision-and-Language Navigation
by: Zhang, Jiwen, et al.
Published: (2026)
by: Zhang, Jiwen, et al.
Published: (2026)
ArtiBench and ArtiBrain: Benchmarking Generalizable Vision-Language Articulated Object Manipulation
by: Wu, Yuhan, et al.
Published: (2025)
by: Wu, Yuhan, et al.
Published: (2025)
VITAL: Interactive Few-Shot Imitation Learning via Visual Human-in-the-Loop Corrections
by: Kasaei, Hamidreza, et al.
Published: (2024)
by: Kasaei, Hamidreza, et al.
Published: (2024)
Distilled Feature Fields Enable Few-Shot Language-Guided Manipulation
by: Shen, William, et al.
Published: (2023)
by: Shen, William, et al.
Published: (2023)
Fully Fine-tuned CLIP Models are Efficient Few-Shot Learners
by: Liu, Mushui, et al.
Published: (2024)
by: Liu, Mushui, et al.
Published: (2024)
FUNCTO: Function-Centric One-Shot Imitation Learning for Tool Manipulation
by: Tang, Chao, et al.
Published: (2025)
by: Tang, Chao, et al.
Published: (2025)
Logits DeConfusion with CLIP for Few-Shot Learning
by: Li, Shuo, et al.
Published: (2025)
by: Li, Shuo, et al.
Published: (2025)
ProtoCLIP: Prototype-Aligned Latent Refinement for Robust Zero-Shot Chest X-Ray Classification
by: Kittler, Florian, et al.
Published: (2026)
by: Kittler, Florian, et al.
Published: (2026)
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
by: Lyu, Kailin, et al.
Published: (2026)
by: Lyu, Kailin, et al.
Published: (2026)
Similar Items
-
Adapting Pre-Trained Vision Models for Novel Instance Detection and Segmentation
by: Lu, Yangxiao, et al.
Published: (2024) -
SCENEREPLICA: Benchmarking Real-World Robot Manipulation by Creating Replicable Scenes
by: Khargonkar, Ninad, et al.
Published: (2023) -
Generalizable Vision-Language Few-Shot Adaptation with Predictive Prompts and Negative Learning
by: Mandalika, Sriram
Published: (2025) -
Few-Shot Fruit Segmentation via Transfer Learning
by: James, Jordan A., et al.
Published: (2024) -
Mechanistic Finetuning of Vision-Language-Action Models via Few-Shot Demonstrations
by: Mitra, Chancharik, et al.
Published: (2025)