Generate, Transduct, Adapt: Iterative Transduction with VLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Saha, Oindrila, Lawrence, Logan, Van Horn, Grant, Maji, Subhransu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improved Zero-Shot Classification by Adapting VLMs with Text Descriptions
by: Saha, Oindrila, et al.
Published: (2024)
by: Saha, Oindrila, et al.
Published: (2024)
You May Speak Freely: Improving the Fine-Grained Visual Recognition Capabilities of Multimodal Large Language Models with Answer Extraction
by: Lawrence, Logan, et al.
Published: (2025)
by: Lawrence, Logan, et al.
Published: (2025)
Not All Birds Look The Same: Identity-Preserving Generation For Birds
by: Sun, Aaron, et al.
Published: (2025)
by: Sun, Aaron, et al.
Published: (2025)
Merlin L48 Spectrogram Dataset
by: Sun, Aaron, et al.
Published: (2025)
by: Sun, Aaron, et al.
Published: (2025)
RealBirdID: Benchmarking Bird Species Identification in the Era of MLLMs
by: Lawrence, Logan, et al.
Published: (2026)
by: Lawrence, Logan, et al.
Published: (2026)
Human-in-the-Loop Visual Re-ID for Population Size Estimation
by: Perez, Gustavo, et al.
Published: (2023)
by: Perez, Gustavo, et al.
Published: (2023)
3D Space as a Scratchpad for Editable Text-to-Image Generation
by: Saha, Oindrila, et al.
Published: (2026)
by: Saha, Oindrila, et al.
Published: (2026)
SIGMA-GEN: Structure and Identity Guided Multi-subject Assembly for Image Generation
by: Saha, Oindrila, et al.
Published: (2025)
by: Saha, Oindrila, et al.
Published: (2025)
Masked Autoencoders with Limited Data: Does It Work? A Fine-Grained Bioacoustics Case Study
by: Liu, Wuao, et al.
Published: (2026)
by: Liu, Wuao, et al.
Published: (2026)
LATA: Laplacian-Assisted Transductive Adaptation for Conformal Uncertainty in Medical VLMs
by: Bozorgtabar, Behzad, et al.
Published: (2026)
by: Bozorgtabar, Behzad, et al.
Published: (2026)
Consensus-Driven Active Model Selection
by: Kay, Justin, et al.
Published: (2025)
by: Kay, Justin, et al.
Published: (2025)
Moment Sampling in Video LLMs for Long-Form Video QA
by: Chasmai, Mustafa, et al.
Published: (2025)
by: Chasmai, Mustafa, et al.
Published: (2025)
Boosting Vision-Language Models with Transduction
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
WildSAT: Learning Satellite Image Representations from Wildlife Observations
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
Improved Feature Generating Framework for Transductive Zero-shot Learning
by: Ye, Zihan, et al.
Published: (2024)
by: Ye, Zihan, et al.
Published: (2024)
UNEM: UNrolled Generalized EM for Transductive Few-Shot Learning
by: Zhou, Long, et al.
Published: (2024)
by: Zhou, Long, et al.
Published: (2024)
Transductive One-Shot Learning Meet Subspace Decomposition
by: Stein, Kyle, et al.
Published: (2025)
by: Stein, Kyle, et al.
Published: (2025)
Learning for Transductive Threshold Calibration in Open-World Recognition
by: Zhang, Qin, et al.
Published: (2023)
by: Zhang, Qin, et al.
Published: (2023)
Adapting Large VLMs with Iterative and Manual Instructions for Generative Low-light Enhancement
by: Sun, Xiaoran, et al.
Published: (2025)
by: Sun, Xiaoran, et al.
Published: (2025)
Task2Box: Box Embeddings for Modeling Asymmetric Task Relationships
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
Active Measurement of Two-Point Correlations
by: Hamilton, Max, et al.
Published: (2026)
by: Hamilton, Max, et al.
Published: (2026)
Language-Aware Information Maximization for Transductive Few-Shot CLIP
by: Baklouti, Ghassen, et al.
Published: (2025)
by: Baklouti, Ghassen, et al.
Published: (2025)
Transductive Zero-Shot and Few-Shot CLIP
by: Martin, Ségolène, et al.
Published: (2024)
by: Martin, Ségolène, et al.
Published: (2024)
Transductive Learning for Near-Duplicate Image Detection in Scanned Photo Collections
by: Net, Francesc, et al.
Published: (2024)
by: Net, Francesc, et al.
Published: (2024)
YouDream: Generating Anatomically Controllable Consistent Text-to-3D Animals
by: Mishra, Sandeep, et al.
Published: (2024)
by: Mishra, Sandeep, et al.
Published: (2024)
C3DAG: Controlled 3D Animal Generation using 3D pose guidance
by: Mishra, Sandeep, et al.
Published: (2024)
by: Mishra, Sandeep, et al.
Published: (2024)
SCAP: Transductive Test-Time Adaptation via Supportive Clique-based Attribute Prompting
by: Zhang, Chenyu, et al.
Published: (2025)
by: Zhang, Chenyu, et al.
Published: (2025)
Feedforward Few-shot Species Range Estimation
by: Lange, Christian, et al.
Published: (2025)
by: Lange, Christian, et al.
Published: (2025)
VIDMP3: Video Editing by Representing Motion with Pose and Position Priors
by: Mishra, Sandeep, et al.
Published: (2025)
by: Mishra, Sandeep, et al.
Published: (2025)
EOL: Transductive Few-Shot Open-Set Recognition by Enhancing Outlier Logits
by: Ochal, Mateusz, et al.
Published: (2024)
by: Ochal, Mateusz, et al.
Published: (2024)
Unbiased Max-Min Embedding Classification for Transductive Few-Shot Learning: Clustering and Classification Are All You Need
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
Active Measurement: Efficient Estimation at Scale
by: Hamilton, Max, et al.
Published: (2025)
by: Hamilton, Max, et al.
Published: (2025)
Transductive Visual Programming: Evolving Tool Libraries from Experience for Spatial Reasoning
by: Wu, Shengguang, et al.
Published: (2025)
by: Wu, Shengguang, et al.
Published: (2025)
AdaptPrompt: Parameter-Efficient Adaptation of VLMs for Generalizable Deepfake Detection
by: Jiang, Yichen, et al.
Published: (2025)
by: Jiang, Yichen, et al.
Published: (2025)
CATRF: Codec-Adaptive TriPlane Radiance Fields for Volumetric Content Delivery
by: Chen, Tung-I, et al.
Published: (2026)
by: Chen, Tung-I, et al.
Published: (2026)
Unmasking Deep Fakes: Leveraging Deep Learning for Video Authenticity Detection
by: Hasan, Mahmudul, et al.
Published: (2025)
by: Hasan, Mahmudul, et al.
Published: (2025)
Contrastive Learning and Cycle Consistency-based Transductive Transfer Learning for Target Annotation
by: Sami, Shoaib Meraj, et al.
Published: (2024)
by: Sami, Shoaib Meraj, et al.
Published: (2024)
Improving Satellite Imagery Masking using Multi-task and Transfer Learning
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
Counting to Four is still a Chore for VLMs
by: Anh, Duy Le Dinh, et al.
Published: (2026)
by: Anh, Duy Le Dinh, et al.
Published: (2026)
Counting Fish with Temporal Representations of Sonar Video
by: Van Brunt, Kai, et al.
Published: (2025)
by: Van Brunt, Kai, et al.
Published: (2025)
Similar Items
-
Improved Zero-Shot Classification by Adapting VLMs with Text Descriptions
by: Saha, Oindrila, et al.
Published: (2024) -
You May Speak Freely: Improving the Fine-Grained Visual Recognition Capabilities of Multimodal Large Language Models with Answer Extraction
by: Lawrence, Logan, et al.
Published: (2025) -
Not All Birds Look The Same: Identity-Preserving Generation For Birds
by: Sun, Aaron, et al.
Published: (2025) -
Merlin L48 Spectrogram Dataset
by: Sun, Aaron, et al.
Published: (2025) -
RealBirdID: Benchmarking Bird Species Identification in the Era of MLLMs
by: Lawrence, Logan, et al.
Published: (2026)