Agro-Consensus: Semantic Self-Consistency in Vision-Language Models for Crop Disease Management in Developing Countries
Fuente:
arXiv
Salvato in:
| Autori principali: | Gupta, Mihir, Desai, Pratik, Greer, Ross |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Self-Consistency in Vision-Language Models for Precision Agriculture: Multi-Response Consensus for Crop Disease Management
di: Gupta, Mihir, et al.
Pubblicazione: (2025)
di: Gupta, Mihir, et al.
Pubblicazione: (2025)
Evaluating Cascaded Methods of Vision-Language Models for Zero-Shot Detection and Association of Hardhats for Increased Construction Safety
di: Choi, Lucas, et al.
Pubblicazione: (2024)
di: Choi, Lucas, et al.
Pubblicazione: (2024)
Beyond General Prompts: Automated Prompt Refinement using Contrastive Class Alignment Scores for Disambiguating Objects in Vision-Language Models
di: Choi, Lucas, et al.
Pubblicazione: (2025)
di: Choi, Lucas, et al.
Pubblicazione: (2025)
Evaluating Vision-Language Models for Zero-Shot Detection, Classification, and Association of Motorcycles, Passengers, and Helmets
di: Choi, Lucas, et al.
Pubblicazione: (2024)
di: Choi, Lucas, et al.
Pubblicazione: (2024)
Multi-Frame, Lightweight & Efficient Vision-Language Models for Question Answering in Autonomous Driving
di: Gopalkrishnan, Akshay, et al.
Pubblicazione: (2024)
di: Gopalkrishnan, Akshay, et al.
Pubblicazione: (2024)
AgroBench: Vision-Language Model Benchmark in Agriculture
di: Shinoda, Risa, et al.
Pubblicazione: (2025)
di: Shinoda, Risa, et al.
Pubblicazione: (2025)
Vision-Language Semantic Grounding for Multi-Domain Crop-Weed Segmentation
di: Hossain, Nazia, et al.
Pubblicazione: (2026)
di: Hossain, Nazia, et al.
Pubblicazione: (2026)
DepthVision: Enabling Robust Vision-Language Models with GAN-Based LiDAR-to-RGB Synthesis for Autonomous Driving
di: Kirchner, Sven, et al.
Pubblicazione: (2025)
di: Kirchner, Sven, et al.
Pubblicazione: (2025)
Finding the Reflection Point: Unpadding Images to Remove Data Augmentation Artifacts in Large Open Source Image Datasets for Machine Learning
di: Choi, Lucas, et al.
Pubblicazione: (2025)
di: Choi, Lucas, et al.
Pubblicazione: (2025)
SMc2f: Robust Scenario Mining for Robotic Autonomy from Coarse to Fine
di: Chen, Yifei, et al.
Pubblicazione: (2026)
di: Chen, Yifei, et al.
Pubblicazione: (2026)
Natural Language Instructions for Scene-Responsive Human-in-the-Loop Motion Planning in Autonomous Driving using Vision-Language-Action Models
di: Martinez-Sanchez, Angel, et al.
Pubblicazione: (2026)
di: Martinez-Sanchez, Angel, et al.
Pubblicazione: (2026)
CropVLM: A Domain-Adapted Vision-Language Model for Open-Set Crop Analysis
di: Boudiaf, Abderrahmene, et al.
Pubblicazione: (2026)
di: Boudiaf, Abderrahmene, et al.
Pubblicazione: (2026)
Driver Activity Classification Using Generalizable Representations from Vision-Language Models
di: Greer, Ross, et al.
Pubblicazione: (2024)
di: Greer, Ross, et al.
Pubblicazione: (2024)
Understanding Virality: A Rubric based Vision-Language Model Framework for Short-Form Edutainment Evaluation
di: Gupta, Arnav, et al.
Pubblicazione: (2025)
di: Gupta, Arnav, et al.
Pubblicazione: (2025)
AgroGPT: Efficient Agricultural Vision-Language Model with Expert Tuning
di: Awais, Muhammad, et al.
Pubblicazione: (2024)
di: Awais, Muhammad, et al.
Pubblicazione: (2024)
Cropper: Vision-Language Model for Image Cropping through In-Context Learning
di: Lee, Seung Hyun, et al.
Pubblicazione: (2024)
di: Lee, Seung Hyun, et al.
Pubblicazione: (2024)
Semantic Graph Consistency: Going Beyond Patches for Regularizing Self-Supervised Vision Transformers
di: Devaguptapu, Chaitanya, et al.
Pubblicazione: (2024)
di: Devaguptapu, Chaitanya, et al.
Pubblicazione: (2024)
How Reasoning Influences Intersectional Biases in Vision Language Models
di: Desai, Adit, et al.
Pubblicazione: (2025)
di: Desai, Adit, et al.
Pubblicazione: (2025)
Can Vision-Language Models Understand and Interpret Dynamic Gestures from Pedestrians? Pilot Datasets and Exploration Towards Instructive Nonverbal Commands for Cooperative Autonomous Vehicles
di: Bossen, Tonko E. W., et al.
Pubblicazione: (2025)
di: Bossen, Tonko E. W., et al.
Pubblicazione: (2025)
Self-Calibrated Consistency can Fight Back for Adversarial Robustness in Vision-Language Models
di: Liu, Jiaxiang, et al.
Pubblicazione: (2025)
di: Liu, Jiaxiang, et al.
Pubblicazione: (2025)
Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model
di: Wang, Chenfeng, et al.
Pubblicazione: (2026)
di: Wang, Chenfeng, et al.
Pubblicazione: (2026)
SC-Tune: Unleashing Self-Consistent Referential Comprehension in Large Vision Language Models
di: Yue, Tongtian, et al.
Pubblicazione: (2024)
di: Yue, Tongtian, et al.
Pubblicazione: (2024)
Perception Without Vision for Trajectory Prediction: Ego Vehicle Dynamics as Scene Representation for Efficient Active Learning in Autonomous Driving
di: Greer, Ross, et al.
Pubblicazione: (2024)
di: Greer, Ross, et al.
Pubblicazione: (2024)
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation
di: Roy, Parthib, et al.
Pubblicazione: (2024)
di: Roy, Parthib, et al.
Pubblicazione: (2024)
Memory-Augmented Vision-Language Agents for Persistent and Semantically Consistent Object Captioning
di: Galliena, Tommaso, et al.
Pubblicazione: (2026)
di: Galliena, Tommaso, et al.
Pubblicazione: (2026)
Towards Self-Refinement of Vision-Language Models with Triangular Consistency
di: Deng, Yunlong, et al.
Pubblicazione: (2025)
di: Deng, Yunlong, et al.
Pubblicazione: (2025)
Test-Time Consistency in Vision Language Models
di: Chou, Shih-Han, et al.
Pubblicazione: (2025)
di: Chou, Shih-Han, et al.
Pubblicazione: (2025)
Technical Report for Argoverse2 Scenario Mining Challenges on Iterative Error Correction and Spatially-Aware Prompting
di: Chen, Yifei, et al.
Pubblicazione: (2025)
di: Chen, Yifei, et al.
Pubblicazione: (2025)
Towards Explainable, Safe Autonomous Driving with Language Embeddings for Novelty Identification and Active Learning: Framework and Experimental Analysis with Real-World Data Sets
di: Greer, Ross, et al.
Pubblicazione: (2024)
di: Greer, Ross, et al.
Pubblicazione: (2024)
Investigating Spatial Attention Bias in Vision-Language Models
di: Chaudhary, Aryan, et al.
Pubblicazione: (2025)
di: Chaudhary, Aryan, et al.
Pubblicazione: (2025)
Design and Implementation of FourCropNet: A CNN-Based System for Efficient Multi-Crop Disease Detection and Management
di: Khandagale, H. P., et al.
Pubblicazione: (2025)
di: Khandagale, H. P., et al.
Pubblicazione: (2025)
Consistency-guided Prompt Learning for Vision-Language Models
di: Roy, Shuvendu, et al.
Pubblicazione: (2023)
di: Roy, Shuvendu, et al.
Pubblicazione: (2023)
Unveiling the Tapestry of Consistency in Large Vision-Language Models
di: Zhang, Yuan, et al.
Pubblicazione: (2024)
di: Zhang, Yuan, et al.
Pubblicazione: (2024)
A Two-Stage Multitask Vision-Language Framework for Explainable Crop Disease Visual Question Answering
di: Hossain, Md. Zahid, et al.
Pubblicazione: (2026)
di: Hossain, Md. Zahid, et al.
Pubblicazione: (2026)
Downstream Analysis of Foundational Medical Vision Models for Disease Progression
di: Demir, Basar, et al.
Pubblicazione: (2025)
di: Demir, Basar, et al.
Pubblicazione: (2025)
Spatio-Temporal Attention for Consistent Video Semantic Segmentation in Automated Driving
di: Varghese, Serin, et al.
Pubblicazione: (2026)
di: Varghese, Serin, et al.
Pubblicazione: (2026)
Automated Data Curation Using GPS & NLP to Generate Instruction-Action Pairs for Autonomous Vehicle Vision-Language Navigation Datasets
di: Roque, Guillermo, et al.
Pubblicazione: (2025)
di: Roque, Guillermo, et al.
Pubblicazione: (2025)
MM-R$^3$: On (In-)Consistency of Vision-Language Models (VLMs)
di: Chou, Shih-Han, et al.
Pubblicazione: (2024)
di: Chou, Shih-Han, et al.
Pubblicazione: (2024)
Latent Semantic Consensus For Deterministic Geometric Model Fitting
di: Xiao, Guobao, et al.
Pubblicazione: (2024)
di: Xiao, Guobao, et al.
Pubblicazione: (2024)
ConsensusDrop: Fusing Visual and Cross-Modal Saliency for Efficient Vision Language Models
di: Parikh, Dhruv, et al.
Pubblicazione: (2026)
di: Parikh, Dhruv, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Self-Consistency in Vision-Language Models for Precision Agriculture: Multi-Response Consensus for Crop Disease Management
di: Gupta, Mihir, et al.
Pubblicazione: (2025) -
Evaluating Cascaded Methods of Vision-Language Models for Zero-Shot Detection and Association of Hardhats for Increased Construction Safety
di: Choi, Lucas, et al.
Pubblicazione: (2024) -
Beyond General Prompts: Automated Prompt Refinement using Contrastive Class Alignment Scores for Disambiguating Objects in Vision-Language Models
di: Choi, Lucas, et al.
Pubblicazione: (2025) -
Evaluating Vision-Language Models for Zero-Shot Detection, Classification, and Association of Motorcycles, Passengers, and Helmets
di: Choi, Lucas, et al.
Pubblicazione: (2024) -
Multi-Frame, Lightweight & Efficient Vision-Language Models for Question Answering in Autonomous Driving
di: Gopalkrishnan, Akshay, et al.
Pubblicazione: (2024)