Global and Local Entailment Learning for Natural World Imagery
Fuente:
arXiv
Saved in:
| Main Authors: | Sastry, Srikumar, Dhakal, Aayush, Xing, Eric, Khanal, Subash, Jacobs, Nathan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GeoSynth: Contextually-Aware High-Resolution Satellite Image Synthesis
by: Sastry, Srikumar, et al.
Published: (2024)
by: Sastry, Srikumar, et al.
Published: (2024)
RANGE: Retrieval Augmented Neural Fields for Multi-Resolution Geo-Embeddings
by: Dhakal, Aayush, et al.
Published: (2025)
by: Dhakal, Aayush, et al.
Published: (2025)
LD-SDM: Language-Driven Hierarchical Species Distribution Modeling
by: Sastry, Srikumar, et al.
Published: (2023)
by: Sastry, Srikumar, et al.
Published: (2023)
TaxaBind: A Unified Embedding Space for Ecological Applications
by: Sastry, Srikumar, et al.
Published: (2024)
by: Sastry, Srikumar, et al.
Published: (2024)
Sat2Cap: Mapping Fine-Grained Textual Descriptions from Satellite Images
by: Dhakal, Aayush, et al.
Published: (2023)
by: Dhakal, Aayush, et al.
Published: (2023)
PSM: Learning Probabilistic Embeddings for Multi-scale Zero-Shot Soundscape Mapping
by: Khanal, Subash, et al.
Published: (2024)
by: Khanal, Subash, et al.
Published: (2024)
Sat2Sound: A Unified Framework for Zero-Shot Soundscape Mapping
by: Khanal, Subash, et al.
Published: (2025)
by: Khanal, Subash, et al.
Published: (2025)
SimLBR: Learning to Detect Fake Images by Learning to Detect Real Images
by: Dhakal, Aayush, et al.
Published: (2026)
by: Dhakal, Aayush, et al.
Published: (2026)
ProM3E: Probabilistic Masked MultiModal Embedding Model for Ecology
by: Sastry, Srikumar, et al.
Published: (2025)
by: Sastry, Srikumar, et al.
Published: (2025)
GeoDiT: Point-Conditioned Diffusion Transformer for Satellite Image Synthesis
by: Sastry, Srikumar, et al.
Published: (2026)
by: Sastry, Srikumar, et al.
Published: (2026)
VectorSynth: Fine-Grained Satellite Image Synthesis with Structured Semantics
by: Cher, Daniel, et al.
Published: (2025)
by: Cher, Daniel, et al.
Published: (2025)
Mixed-View Panorama Synthesis using Geospatially Guided Diffusion
by: Xiong, Zhexiao, et al.
Published: (2024)
by: Xiong, Zhexiao, et al.
Published: (2024)
GEOBIND: Binding Text, Image, and Audio through Satellite Images
by: Dhakal, Aayush, et al.
Published: (2024)
by: Dhakal, Aayush, et al.
Published: (2024)
DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments
by: Sarkar, Anindya, et al.
Published: (2026)
by: Sarkar, Anindya, et al.
Published: (2026)
GOMAA-Geo: GOal Modality Agnostic Active Geo-localization
by: Sarkar, Anindya, et al.
Published: (2024)
by: Sarkar, Anindya, et al.
Published: (2024)
Towards Open-World Generation of Stereo Images and Unsupervised Matching
by: Qiao, Feng, et al.
Published: (2025)
by: Qiao, Feng, et al.
Published: (2025)
TuneVLSeg: Prompt Tuning Benchmark for Vision-Language Segmentation Models
by: Adhikari, Rabin, et al.
Published: (2024)
by: Adhikari, Rabin, et al.
Published: (2024)
ConText-CIR: Learning from Concepts in Text for Composed Image Retrieval
by: Xing, Eric, et al.
Published: (2025)
by: Xing, Eric, et al.
Published: (2025)
QuARI: Query Adaptive Retrieval Improvement
by: Xing, Eric, et al.
Published: (2025)
by: Xing, Eric, et al.
Published: (2025)
Severe Domain Shift in Skeleton-Based Action Recognition:A Study of Uncertainty Failure in Real-World Gym Environments
by: Khanal, Aaditya, et al.
Published: (2026)
by: Khanal, Aaditya, et al.
Published: (2026)
good4cir: Generating Detailed Synthetic Captions for Composed Image Retrieval
by: Kolouju, Pranavi, et al.
Published: (2025)
by: Kolouju, Pranavi, et al.
Published: (2025)
VLSM-Adapter: Finetuning Vision-Language Segmentation Efficiently with Lightweight Blocks
by: Dhakal, Manish, et al.
Published: (2024)
by: Dhakal, Manish, et al.
Published: (2024)
HYPE: Hyperbolic Entailment Filtering for Underspecified Images and Texts
by: Kim, Wonjae, et al.
Published: (2024)
by: Kim, Wonjae, et al.
Published: (2024)
On the Status of Foundation Models for SAR Imagery
by: Inkawhich, Nathan
Published: (2025)
by: Inkawhich, Nathan
Published: (2025)
Exploring Transfer Learning in Medical Image Segmentation using Vision-Language Models
by: Poudel, Kanchan, et al.
Published: (2023)
by: Poudel, Kanchan, et al.
Published: (2023)
PRUE: A Practical Recipe for Field Boundary Segmentation at Scale
by: Muhawenayo, Gedeon, et al.
Published: (2026)
by: Muhawenayo, Gedeon, et al.
Published: (2026)
StereoGenBench: A Synthetic Multi-Camera Benchmark for Stereo Generation under Controlled Baseline Regimes
by: Cui, Yangzhi, et al.
Published: (2026)
by: Cui, Yangzhi, et al.
Published: (2026)
Adversarial Robustness Analysis of Vision-Language Models in Medical Image Segmentation
by: Budathoki, Anjila, et al.
Published: (2025)
by: Budathoki, Anjila, et al.
Published: (2025)
The first global agricultural field boundary map at 10m resolution
by: Robinson, Caleb, et al.
Published: (2026)
by: Robinson, Caleb, et al.
Published: (2026)
Detecting Visual Triggers in Cannabis Imagery: A CLIP-Based Multi-Labeling Framework with Local-Global Aggregation
by: Lu, Linqi, et al.
Published: (2024)
by: Lu, Linqi, et al.
Published: (2024)
HyperAlign: Hyperbolic Entailment Cones for Adaptive Text-to-Image Alignment Assessment
by: Chen, Wenzhi, et al.
Published: (2026)
by: Chen, Wenzhi, et al.
Published: (2026)
GenOpticalFlow: A Generative Approach to Unsupervised Optical Flow Learning
by: Luo, Yixuan, et al.
Published: (2026)
by: Luo, Yixuan, et al.
Published: (2026)
DELST: Dual Entailment Learning for Hyperbolic Image-Gene Pretraining in Spatial Transcriptomics
by: Chen, Xulin, et al.
Published: (2025)
by: Chen, Xulin, et al.
Published: (2025)
Compositional Entailment Learning for Hyperbolic Vision-Language Models
by: Pal, Avik, et al.
Published: (2024)
by: Pal, Avik, et al.
Published: (2024)
UniDrive-WM: Unified Understanding, Planning and Generation World Model For Autonomous Driving
by: Xiong, Zhexiao, et al.
Published: (2026)
by: Xiong, Zhexiao, et al.
Published: (2026)
Benchmarking Object Detectors under Real-World Distribution Shifts in Satellite Imagery
by: Al-Emadi, Sara, et al.
Published: (2025)
by: Al-Emadi, Sara, et al.
Published: (2025)
Mixture of Global and Local Experts with Diffusion Transformer for Controllable Face Generation
by: Zou, Xuechao, et al.
Published: (2025)
by: Zou, Xuechao, et al.
Published: (2025)
DeclutterNeRF: Generative-Free 3D Scene Recovery for Occlusion Removal
by: Liu, Wanzhou, et al.
Published: (2025)
by: Liu, Wanzhou, et al.
Published: (2025)
From Orthomosaics to Raw UAV Imagery: Enhancing Palm Detection and Crown-Center Localization
by: Zhu, Rongkun, et al.
Published: (2025)
by: Zhu, Rongkun, et al.
Published: (2025)
Cross-View Geo-Localization with Street-View and VHR Satellite Imagery in Decentrality Settings
by: Xia, Panwang, et al.
Published: (2024)
by: Xia, Panwang, et al.
Published: (2024)
Similar Items
-
GeoSynth: Contextually-Aware High-Resolution Satellite Image Synthesis
by: Sastry, Srikumar, et al.
Published: (2024) -
RANGE: Retrieval Augmented Neural Fields for Multi-Resolution Geo-Embeddings
by: Dhakal, Aayush, et al.
Published: (2025) -
LD-SDM: Language-Driven Hierarchical Species Distribution Modeling
by: Sastry, Srikumar, et al.
Published: (2023) -
TaxaBind: A Unified Embedding Space for Ecological Applications
by: Sastry, Srikumar, et al.
Published: (2024) -
Sat2Cap: Mapping Fine-Grained Textual Descriptions from Satellite Images
by: Dhakal, Aayush, et al.
Published: (2023)