GEOBIND: Binding Text, Image, and Audio through Satellite Images
Fuente:
arXiv
Saved in:
| Main Authors: | Dhakal, Aayush, Khanal, Subash, Sastry, Srikumar, Ahmad, Adeel, Jacobs, Nathan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TaxaBind: A Unified Embedding Space for Ecological Applications
by: Sastry, Srikumar, et al.
Published: (2024)
by: Sastry, Srikumar, et al.
Published: (2024)
Sat2Sound: A Unified Framework for Zero-Shot Soundscape Mapping
by: Khanal, Subash, et al.
Published: (2025)
by: Khanal, Subash, et al.
Published: (2025)
GeoSynth: Contextually-Aware High-Resolution Satellite Image Synthesis
by: Sastry, Srikumar, et al.
Published: (2024)
by: Sastry, Srikumar, et al.
Published: (2024)
Sat2Cap: Mapping Fine-Grained Textual Descriptions from Satellite Images
by: Dhakal, Aayush, et al.
Published: (2023)
by: Dhakal, Aayush, et al.
Published: (2023)
LD-SDM: Language-Driven Hierarchical Species Distribution Modeling
by: Sastry, Srikumar, et al.
Published: (2023)
by: Sastry, Srikumar, et al.
Published: (2023)
RANGE: Retrieval Augmented Neural Fields for Multi-Resolution Geo-Embeddings
by: Dhakal, Aayush, et al.
Published: (2025)
by: Dhakal, Aayush, et al.
Published: (2025)
GeoDiT: Point-Conditioned Diffusion Transformer for Satellite Image Synthesis
by: Sastry, Srikumar, et al.
Published: (2026)
by: Sastry, Srikumar, et al.
Published: (2026)
Global and Local Entailment Learning for Natural World Imagery
by: Sastry, Srikumar, et al.
Published: (2025)
by: Sastry, Srikumar, et al.
Published: (2025)
PSM: Learning Probabilistic Embeddings for Multi-scale Zero-Shot Soundscape Mapping
by: Khanal, Subash, et al.
Published: (2024)
by: Khanal, Subash, et al.
Published: (2024)
SimLBR: Learning to Detect Fake Images by Learning to Detect Real Images
by: Dhakal, Aayush, et al.
Published: (2026)
by: Dhakal, Aayush, et al.
Published: (2026)
ProM3E: Probabilistic Masked MultiModal Embedding Model for Ecology
by: Sastry, Srikumar, et al.
Published: (2025)
by: Sastry, Srikumar, et al.
Published: (2025)
DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments
by: Sarkar, Anindya, et al.
Published: (2026)
by: Sarkar, Anindya, et al.
Published: (2026)
VectorSynth: Fine-Grained Satellite Image Synthesis with Structured Semantics
by: Cher, Daniel, et al.
Published: (2025)
by: Cher, Daniel, et al.
Published: (2025)
GOMAA-Geo: GOal Modality Agnostic Active Geo-localization
by: Sarkar, Anindya, et al.
Published: (2024)
by: Sarkar, Anindya, et al.
Published: (2024)
Exploring Transfer Learning in Medical Image Segmentation using Vision-Language Models
by: Poudel, Kanchan, et al.
Published: (2023)
by: Poudel, Kanchan, et al.
Published: (2023)
ThreatFormer-IDS: Robust Transformer Intrusion Detection with Zero-Day Generalization and Explainable Attribution
by: Nayak, Srikumar
Published: (2026)
by: Nayak, Srikumar
Published: (2026)
Geometrical Properties of Text Token Embeddings for Strong Semantic Binding in Text-to-Image Generation
by: Seo, Hoigi, et al.
Published: (2025)
by: Seo, Hoigi, et al.
Published: (2025)
Token Merging for Training-Free Semantic Binding in Text-to-Image Synthesis
by: Hu, Taihang, et al.
Published: (2024)
by: Hu, Taihang, et al.
Published: (2024)
Grounding Text-to-Image Diffusion Models for Controlled High-Quality Image Generation
by: Süleyman, Ahmad, et al.
Published: (2025)
by: Süleyman, Ahmad, et al.
Published: (2025)
VLSM-Adapter: Finetuning Vision-Language Segmentation Efficiently with Lightweight Blocks
by: Dhakal, Manish, et al.
Published: (2024)
by: Dhakal, Manish, et al.
Published: (2024)
Multimodal Medical Image Binding via Shared Text Embeddings
by: Liu, Yunhao, et al.
Published: (2025)
by: Liu, Yunhao, et al.
Published: (2025)
HQFS: Hybrid Quantum Classical Financial Security with VQC Forecasting, QUBO Annealing, and Audit-Ready Post-Quantum Signing
by: Nayak, Srikumar
Published: (2026)
by: Nayak, Srikumar
Published: (2026)
Named Entity Recognition for Payment Data Using NLP
by: Nayak, Srikumar
Published: (2026)
by: Nayak, Srikumar
Published: (2026)
Calibrated Credit Intelligence: Shift-Robust and Fair Risk Scoring with Bayesian Uncertainty and Gradient Boosting
by: Nayak, Srikumar
Published: (2026)
by: Nayak, Srikumar
Published: (2026)
Addressing Image Hallucination in Text-to-Image Generation through Factual Image Retrieval
by: Lim, Youngsun, et al.
Published: (2024)
by: Lim, Youngsun, et al.
Published: (2024)
StyleForge: Enhancing Text-to-Image Synthesis for Any Artistic Styles with Dual Binding
by: Park, Junseo, et al.
Published: (2024)
by: Park, Junseo, et al.
Published: (2024)
How Bias Binds: Measuring Hidden Associations for Bias Control in Text-to-Image Compositions
by: Li, Jeng-Lin, et al.
Published: (2025)
by: Li, Jeng-Lin, et al.
Published: (2025)
Node-Based Editing for Multimodal Generation of Text, Audio, Image, and Video
by: Kyaw, Alexander Htet, et al.
Published: (2025)
by: Kyaw, Alexander Htet, et al.
Published: (2025)
A Cellular Doctrine of Morality: Intrinsic Active Precision and the Mind-Reality Overload Dilemma
by: Adeel, Ahsan
Published: (2026)
by: Adeel, Ahsan
Published: (2026)
UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain Adaptation
by: Zhong, Siru, et al.
Published: (2024)
by: Zhong, Siru, et al.
Published: (2024)
good4cir: Generating Detailed Synthetic Captions for Composed Image Retrieval
by: Kolouju, Pranavi, et al.
Published: (2025)
by: Kolouju, Pranavi, et al.
Published: (2025)
RLShield: Practical Multi-Agent RL for Financial Cyber Defense with Attack-Surface MDPs and Real-Time Response Orchestration
by: Nayak, Srikumar
Published: (2026)
by: Nayak, Srikumar
Published: (2026)
InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis
by: Bentham, Oliver, et al.
Published: (2026)
by: Bentham, Oliver, et al.
Published: (2026)
Multi-view Image Diffusion via Coordinate Noise and Fourier Attention
by: Theiss, Justin, et al.
Published: (2024)
by: Theiss, Justin, et al.
Published: (2024)
Accelerating Discrete Facility Layout Optimization: A Hybrid CDCL and CP-SAT Architecture
by: Gibson, Joshua, et al.
Published: (2025)
by: Gibson, Joshua, et al.
Published: (2025)
Self-Supervision in Time for Satellite Images(S3-TSS): A novel method of SSL technique in Satellite images
by: Maurya, Akansh, et al.
Published: (2024)
by: Maurya, Akansh, et al.
Published: (2024)
MindCraft: Revolutionizing Education through AI-Powered Personalized Learning and Mentorship for Rural India
by: Bardia, Arihant, et al.
Published: (2025)
by: Bardia, Arihant, et al.
Published: (2025)
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
by: Lanier, Michael, et al.
Published: (2024)
by: Lanier, Michael, et al.
Published: (2024)
Source-Free and Image-Only Unsupervised Domain Adaptation for Category Level Object Pose Estimation
by: Kaushik, Prakhar, et al.
Published: (2024)
by: Kaushik, Prakhar, et al.
Published: (2024)
Object-centric Binding in Contrastive Language-Image Pretraining
by: Assouel, Rim, et al.
Published: (2025)
by: Assouel, Rim, et al.
Published: (2025)
Similar Items
-
TaxaBind: A Unified Embedding Space for Ecological Applications
by: Sastry, Srikumar, et al.
Published: (2024) -
Sat2Sound: A Unified Framework for Zero-Shot Soundscape Mapping
by: Khanal, Subash, et al.
Published: (2025) -
GeoSynth: Contextually-Aware High-Resolution Satellite Image Synthesis
by: Sastry, Srikumar, et al.
Published: (2024) -
Sat2Cap: Mapping Fine-Grained Textual Descriptions from Satellite Images
by: Dhakal, Aayush, et al.
Published: (2023) -
LD-SDM: Language-Driven Hierarchical Species Distribution Modeling
by: Sastry, Srikumar, et al.
Published: (2023)