Saved in:
| Main Authors: | Ranario, Earl, Earles, Mason J. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.15977 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AgRowStitch: A High-fidelity Image Stitching Pipeline for Ground-based Agricultural Images
by: Uyehara, Isaac Kazuo, et al.
Published: (2025)
by: Uyehara, Isaac Kazuo, et al.
Published: (2025)
Enabling Plant Phenotyping in Weedy Environments using Multi-Modal Imagery via Synthetic and Generated Training Data
by: Ranario, Earl, et al.
Published: (2025)
by: Ranario, Earl, et al.
Published: (2025)
AGILE: A Diffusion-Based Attention-Guided Image and Label Translation for Efficient Cross-Domain Plant Trait Identification
by: Ranario, Earl, et al.
Published: (2025)
by: Ranario, Earl, et al.
Published: (2025)
Does Your VFM Speak Plant? The Botanical Grammar of Vision Foundation Models for Object Detection
by: Lundqvist, Lars, et al.
Published: (2026)
by: Lundqvist, Lars, et al.
Published: (2026)
A Vision Language Model for Generating Procedural Plant Architecture Representations from Simulated Images
by: Yun, Heesup, et al.
Published: (2026)
by: Yun, Heesup, et al.
Published: (2026)
Using Vision Language Foundation Models to Generate Plant Simulation Configurations via In-Context Learning
by: Yun, Heesup, et al.
Published: (2026)
by: Yun, Heesup, et al.
Published: (2026)
Initialization matters in few-shot adaptation of vision-language models for histopathological image classification
by: Meseguer, Pablo, et al.
Published: (2026)
by: Meseguer, Pablo, et al.
Published: (2026)
California Crop Yield Benchmark: Combining Satellite Image, Climate, Evapotranspiration, and Soil Data Layers for County-Level Yield Forecasting of Over 70 Crops
by: Kamangir, Hamid, et al.
Published: (2025)
by: Kamangir, Hamid, et al.
Published: (2025)
iNatAg: Multi-Class Classification Models Enabled by a Large-Scale Benchmark Dataset with 4.7M Images of 2,959 Crop and Weed Species
by: Jain, Naitik, et al.
Published: (2025)
by: Jain, Naitik, et al.
Published: (2025)
On the test-time zero-shot generalization of vision-language models: Do we really need prompt learning?
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
MI-VisionShot: Few-shot adaptation of vision-language models for slide-level classification of histopathological images
by: Meseguer, Pablo, et al.
Published: (2024)
by: Meseguer, Pablo, et al.
Published: (2024)
Zero-shot segmentation of skin tumors in whole-slide images with vision-language foundation models
by: Moreno, Santiago, et al.
Published: (2025)
by: Moreno, Santiago, et al.
Published: (2025)
FreeZe: Training-free zero-shot 6D pose estimation with geometric and vision foundation models
by: Caraffa, Andrea, et al.
Published: (2023)
by: Caraffa, Andrea, et al.
Published: (2023)
Prompt as Knowledge Bank: Boost Vision-language model via Structural Representation for zero-shot medical detection
by: Yang, Yuguang, et al.
Published: (2025)
by: Yang, Yuguang, et al.
Published: (2025)
Enhancing AI microscopy for foodborne bacterial classification via adversarial domain adaptation across optical and biological variability
by: Bhattacharya, Siddhartha, et al.
Published: (2024)
by: Bhattacharya, Siddhartha, et al.
Published: (2024)
Interpreting vision transformers via residual replacement model
by: Kim, Jinyeong, et al.
Published: (2025)
by: Kim, Jinyeong, et al.
Published: (2025)
Retrieval-enriched zero-shot image classification in low-resource domains
by: Dall'Asen, Nicola, et al.
Published: (2024)
by: Dall'Asen, Nicola, et al.
Published: (2024)
An analysis of vision-language models for fabric retrieval
by: Giuliari, Francesco, et al.
Published: (2025)
by: Giuliari, Francesco, et al.
Published: (2025)
CMAViT: Integrating Climate, Managment, and Remote Sensing Data for Crop Yield Estimation with Multimodel Vision Transformers
by: Kamangir, Hamid, et al.
Published: (2024)
by: Kamangir, Hamid, et al.
Published: (2024)
Taming generative video models for zero-shot optical flow extraction
by: Kim, Seungwoo, et al.
Published: (2025)
by: Kim, Seungwoo, et al.
Published: (2025)
Zero-shot large vision-language model prompting for automated bone identification in paleoradiology x-ray archives
by: Dong, Owen, et al.
Published: (2026)
by: Dong, Owen, et al.
Published: (2026)
ESP-Zero: Unsupervised enhancement of zero-shot classification for Extremely Sparse Point cloud
by: Han, Jiayi, et al.
Published: (2024)
by: Han, Jiayi, et al.
Published: (2024)
Are vision language models robust to uncertain inputs?
by: Wang, Xi, et al.
Published: (2025)
by: Wang, Xi, et al.
Published: (2025)
VisTA-SR: Improving the Accuracy and Resolution of Low-Cost Thermal Imaging Cameras for Agriculture
by: Yun, Heesup, et al.
Published: (2024)
by: Yun, Heesup, et al.
Published: (2024)
Cascading multi-agent anomaly detection in surveillance systems via vision-language models and embedding-based classification
by: Rehman, Tayyab, et al.
Published: (2026)
by: Rehman, Tayyab, et al.
Published: (2026)
Sleep-stage efficient classification using a lightweight self-supervised model
by: Durães, Eldiane Borges dos Santos, et al.
Published: (2026)
by: Durães, Eldiane Borges dos Santos, et al.
Published: (2026)
Accurate and efficient zero-shot 6D pose estimation with frozen foundation models
by: Caraffa, Andrea, et al.
Published: (2025)
by: Caraffa, Andrea, et al.
Published: (2025)
VCP-CLIP: A visual context prompting model for zero-shot anomaly segmentation
by: Qu, Zhen, et al.
Published: (2024)
by: Qu, Zhen, et al.
Published: (2024)
Visual symbolic mechanisms: Emergent symbol processing in vision language models
by: Assouel, Rim, et al.
Published: (2025)
by: Assouel, Rim, et al.
Published: (2025)
Do large language vision models understand 3D shapes?
by: Eppel, Sagi
Published: (2024)
by: Eppel, Sagi
Published: (2024)
Video models are zero-shot learners and reasoners
by: Wiedemer, Thaddäus, et al.
Published: (2025)
by: Wiedemer, Thaddäus, et al.
Published: (2025)
Quantifying the human visual exposome with vision language models
by: Rominger, Christian, et al.
Published: (2026)
by: Rominger, Christian, et al.
Published: (2026)
What matters when building vision-language models?
by: Laurençon, Hugo, et al.
Published: (2024)
by: Laurençon, Hugo, et al.
Published: (2024)
bi-modal textual prompt learning for vision-language models in remote sensing
by: Kashyap, Pankhi, et al.
Published: (2026)
by: Kashyap, Pankhi, et al.
Published: (2026)
Interpreting the linear structure of vision-language model embedding spaces
by: Papadimitriou, Isabel, et al.
Published: (2025)
by: Papadimitriou, Isabel, et al.
Published: (2025)
Amortizing intractable inference in diffusion models for vision, language, and control
by: Venkatraman, Siddarth, et al.
Published: (2024)
by: Venkatraman, Siddarth, et al.
Published: (2024)
Thinker: A vision-language foundation model for embodied intelligence
by: Pan, Baiyu, et al.
Published: (2026)
by: Pan, Baiyu, et al.
Published: (2026)
Enhancing medical vision-language contrastive learning via inter-matching relation modelling
by: Li, Mingjian, et al.
Published: (2024)
by: Li, Mingjian, et al.
Published: (2024)
A multi-modal vision-language model for generalizable annotation-free pathology localization
by: Yang, Hao, et al.
Published: (2024)
by: Yang, Hao, et al.
Published: (2024)
Enhancing the vision-language foundation model with key semantic knowledge-emphasized report refinement
by: Huang, Weijian, et al.
Published: (2024)
by: Huang, Weijian, et al.
Published: (2024)
Similar Items
-
AgRowStitch: A High-fidelity Image Stitching Pipeline for Ground-based Agricultural Images
by: Uyehara, Isaac Kazuo, et al.
Published: (2025) -
Enabling Plant Phenotyping in Weedy Environments using Multi-Modal Imagery via Synthetic and Generated Training Data
by: Ranario, Earl, et al.
Published: (2025) -
AGILE: A Diffusion-Based Attention-Guided Image and Label Translation for Efficient Cross-Domain Plant Trait Identification
by: Ranario, Earl, et al.
Published: (2025) -
Does Your VFM Speak Plant? The Botanical Grammar of Vision Foundation Models for Object Detection
by: Lundqvist, Lars, et al.
Published: (2026) -
A Vision Language Model for Generating Procedural Plant Architecture Representations from Simulated Images
by: Yun, Heesup, et al.
Published: (2026)