CLIP-Guided Multi-Task Regression for Multi-View Plant Phenotyping
Fuente:
arXiv
Saved in:
| Main Authors: | Warmers, Simon, Zawish, Muhammad, Dharejo, Fayaz Ali, Davy, Steven, Timofte, Radu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
WaveHiT-SR: Hierarchical Wavelet Network for Efficient Image Super-Resolution
by: Ali, Fayaz, et al.
Published: (2025)
by: Ali, Fayaz, et al.
Published: (2025)
Degradation-Aware All-in-One Image Restoration via Latent Prior Encoding
by: Sharif, S M A, et al.
Published: (2025)
by: Sharif, S M A, et al.
Published: (2025)
TriFusion-SR: Joint Tri-Modal Medical Image Fusion and SR
by: Dharejo, Fayaz Ali, et al.
Published: (2026)
by: Dharejo, Fayaz Ali, et al.
Published: (2026)
Illuminating Darkness: Learning to Enhance Low-light Images In-the-Wild
by: Sharif, S M A, et al.
Published: (2025)
by: Sharif, S M A, et al.
Published: (2025)
NeuCODEX: Edge-Cloud Co-Inference with Spike-Driven Compression and Dynamic Early-Exit
by: Hassan, Maurf, et al.
Published: (2025)
by: Hassan, Maurf, et al.
Published: (2025)
Energy-Efficient Uncertainty-Aware Biomass Composition Prediction at the Edge
by: Zawish, Muhammad, et al.
Published: (2024)
by: Zawish, Muhammad, et al.
Published: (2024)
ViewSparsifier: Killing Redundancy in Multi-View Plant Phenotyping
by: Kampa, Robin-Nico, et al.
Published: (2025)
by: Kampa, Robin-Nico, et al.
Published: (2025)
Cat: Post-Training Quantization Error Reduction via Cluster-based Affine Transformation
by: Zoljodi, Ali, et al.
Published: (2025)
by: Zoljodi, Ali, et al.
Published: (2025)
Practical Manipulation Model for Robust Deepfake Detection
by: Hopf, Benedikt, et al.
Published: (2025)
by: Hopf, Benedikt, et al.
Published: (2025)
Experts-Guided Unbalanced Optimal Transport for ISP Learning from Unpaired and/or Paired Data
by: Perevozchikov, Georgy, et al.
Published: (2025)
by: Perevozchikov, Georgy, et al.
Published: (2025)
From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs
by: Shrestha, Usha, et al.
Published: (2026)
by: Shrestha, Usha, et al.
Published: (2026)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
by: Song, Dan, et al.
Published: (2023)
by: Song, Dan, et al.
Published: (2023)
A Light Weight Multi-Features-View Convolution Neural Network For Plant Disease Identification
by: Khan, Muhammad Kaleem Ullah
Published: (2026)
by: Khan, Muhammad Kaleem Ullah
Published: (2026)
Layover or Direct Flight: Rethinking Audio-Guided Image Segmentation
by: Santos, Joel Alberto, et al.
Published: (2025)
by: Santos, Joel Alberto, et al.
Published: (2025)
MuDreamer: Learning Predictive World Models without Reconstruction
by: Burchi, Maxime, et al.
Published: (2024)
by: Burchi, Maxime, et al.
Published: (2024)
Learned Lightweight Smartphone ISP with Unpaired Data
by: Arhire, Andrei, et al.
Published: (2025)
by: Arhire, Andrei, et al.
Published: (2025)
Bridging the Task Gap: Multi-Task Adversarial Transferability in CLIP and Its Derivatives
by: Liu, Kuanrong, et al.
Published: (2025)
by: Liu, Kuanrong, et al.
Published: (2025)
Duoduo CLIP: Efficient 3D Understanding with Multi-View Images
by: Lee, Han-Hung, et al.
Published: (2024)
by: Lee, Han-Hung, et al.
Published: (2024)
The Regularizing Power of Language-Training Deepfake Detectors
by: Hopf, Benedikt, et al.
Published: (2026)
by: Hopf, Benedikt, et al.
Published: (2026)
The Return of Structural Handwritten Mathematical Expression Recognition
by: Seitz, Jakob, et al.
Published: (2025)
by: Seitz, Jakob, et al.
Published: (2025)
AugmentGest: Can Random Data Cropping Augmentation Boost Gesture Recognition Performance?
by: Aboudeshish, Nada, et al.
Published: (2025)
by: Aboudeshish, Nada, et al.
Published: (2025)
CLIP-Decoder : ZeroShot Multilabel Classification using Multimodal CLIP Aligned Representation
by: Ali, Muhammad, et al.
Published: (2024)
by: Ali, Muhammad, et al.
Published: (2024)
LeMoRe: Learn More Details for Lightweight Semantic Segmentation
by: Abid, Mian Muhammad Naeem, et al.
Published: (2025)
by: Abid, Mian Muhammad Naeem, et al.
Published: (2025)
ContextFormer: Redefining Efficiency in Semantic Segmentation
by: Abid, Mian Muhammad Naeem, et al.
Published: (2025)
by: Abid, Mian Muhammad Naeem, et al.
Published: (2025)
Video CLIP Model for Multi-View Echocardiography Interpretation
by: Takizawa, Ryo, et al.
Published: (2025)
by: Takizawa, Ryo, et al.
Published: (2025)
Modulate and Reconstruct: Learning Hyperspectral Imaging from Misaligned Smartphone Views
by: Reutsky, Daniil, et al.
Published: (2025)
by: Reutsky, Daniil, et al.
Published: (2025)
Higher fidelity perceptual image and video compression with a latent conditioned residual denoising diffusion model
by: Brenig, Jonas, et al.
Published: (2025)
by: Brenig, Jonas, et al.
Published: (2025)
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
by: Monsefi, Amin Karimi, et al.
Published: (2024)
by: Monsefi, Amin Karimi, et al.
Published: (2024)
Auto-Regressively Generating Multi-View Consistent Images
by: Hu, JiaKui, et al.
Published: (2025)
by: Hu, JiaKui, et al.
Published: (2025)
Babel-ImageNet: Massively Multilingual Evaluation of Vision-and-Language Representations
by: Geigle, Gregor, et al.
Published: (2023)
by: Geigle, Gregor, et al.
Published: (2023)
Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?
by: Geigle, Gregor, et al.
Published: (2024)
by: Geigle, Gregor, et al.
Published: (2024)
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
by: Geigle, Gregor, et al.
Published: (2024)
by: Geigle, Gregor, et al.
Published: (2024)
Closed-Loop LLM Discovery of Non-Standard Channel Priors in Vision Models
by: Uzun, Tolgay Atinc, et al.
Published: (2026)
by: Uzun, Tolgay Atinc, et al.
Published: (2026)
CLIP3D-AD: Extending CLIP for 3D Few-Shot Anomaly Detection with Multi-View Images Generation
by: Zuo, Zuo, et al.
Published: (2024)
by: Zuo, Zuo, et al.
Published: (2024)
Multi Camera Connected Vision System with Multi View Analytics: A Comprehensive Survey
by: Munsif, Muhammad, et al.
Published: (2025)
by: Munsif, Muhammad, et al.
Published: (2025)
CPLIP: Zero-Shot Learning for Histopathology with Comprehensive Vision-Language Alignment
by: Javed, Sajid, et al.
Published: (2024)
by: Javed, Sajid, et al.
Published: (2024)
Equivariant Multi-Modality Image Fusion
by: Zhao, Zixiang, et al.
Published: (2023)
by: Zhao, Zixiang, et al.
Published: (2023)
Learning Transformer-based World Models with Contrastive Predictive Coding
by: Burchi, Maxime, et al.
Published: (2025)
by: Burchi, Maxime, et al.
Published: (2025)
Accurate and Efficient World Modeling with Masked Latent Transformers
by: Burchi, Maxime, et al.
Published: (2025)
by: Burchi, Maxime, et al.
Published: (2025)
Task-Augmented Cross-View Imputation Network for Partial Multi-View Incomplete Multi-Label Classification
by: Zhao, Lian, et al.
Published: (2024)
by: Zhao, Lian, et al.
Published: (2024)
Similar Items
-
WaveHiT-SR: Hierarchical Wavelet Network for Efficient Image Super-Resolution
by: Ali, Fayaz, et al.
Published: (2025) -
Degradation-Aware All-in-One Image Restoration via Latent Prior Encoding
by: Sharif, S M A, et al.
Published: (2025) -
TriFusion-SR: Joint Tri-Modal Medical Image Fusion and SR
by: Dharejo, Fayaz Ali, et al.
Published: (2026) -
Illuminating Darkness: Learning to Enhance Low-light Images In-the-Wild
by: Sharif, S M A, et al.
Published: (2025) -
NeuCODEX: Edge-Cloud Co-Inference with Spike-Driven Compression and Dynamic Early-Exit
by: Hassan, Maurf, et al.
Published: (2025)