Embedding-Only Uplink for Onboard Retrieval Under Shift in Remote Sensing
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Sim, Sangcheol |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
RSTeller: Scaling Up Visual Language Modeling in Remote Sensing with Rich Linguistic Semantics from Openly Available Data and Large Language Models
von: Ge, Junyao, et al.
Veröffentlicht: (2024)
von: Ge, Junyao, et al.
Veröffentlicht: (2024)
LATTE: Latent Trajectory Embedding for Diffusion-Generated Image Detection
von: Vasilcoiu, Ana, et al.
Veröffentlicht: (2025)
von: Vasilcoiu, Ana, et al.
Veröffentlicht: (2025)
DeltaVLM: Interactive Remote Sensing Image Change Analysis via Instruction-guided Difference Perception
von: Deng, Pei, et al.
Veröffentlicht: (2025)
von: Deng, Pei, et al.
Veröffentlicht: (2025)
Interpretable Tau-PET Synthesis from Multimodal T1-Weighted and FLAIR MRI Using Partial Information Decomposition Guided Disentangled Quantized Half-UNet
von: Chopra, Agamdeep S., et al.
Veröffentlicht: (2026)
von: Chopra, Agamdeep S., et al.
Veröffentlicht: (2026)
A Two-Stage, Object-Centric Deep Learning Framework for Robust Exam Cheating Detection
von: Le, Van-Truong, et al.
Veröffentlicht: (2026)
von: Le, Van-Truong, et al.
Veröffentlicht: (2026)
Fashion Florence: Fine-Tuning Florence-2 for Structured Fashion Attribute Extraction
von: Berlia, Anushree
Veröffentlicht: (2026)
von: Berlia, Anushree
Veröffentlicht: (2026)
ChartComplete: A Taxonomy-based Inclusive Chart Dataset
von: Mustapha, Ahmad, et al.
Veröffentlicht: (2026)
von: Mustapha, Ahmad, et al.
Veröffentlicht: (2026)
GenMatter: Perceiving Physical Objects with Generative Matter Models
von: Li, Eric, et al.
Veröffentlicht: (2026)
von: Li, Eric, et al.
Veröffentlicht: (2026)
Automated Plant Disease and Pest Detection System Using Hybrid Lightweight CNN-MobileViT Models for Diagnosis of Indigenous Crops
von: Gebremedhin, Tekleab G., et al.
Veröffentlicht: (2025)
von: Gebremedhin, Tekleab G., et al.
Veröffentlicht: (2025)
CausalVQA: A Physically Grounded Causal Reasoning Benchmark for Video Models
von: Foss, Aaron, et al.
Veröffentlicht: (2025)
von: Foss, Aaron, et al.
Veröffentlicht: (2025)
Application of YOLOv8 in monocular downward multiple Car Target detection
von: Lyu, Shijie
Veröffentlicht: (2025)
von: Lyu, Shijie
Veröffentlicht: (2025)
THIRDEYE: Cue-Aware Monocular Depth Estimation via Brain-Inspired Multi-Stage Fusion
von: Ioan, Calin Teodor
Veröffentlicht: (2025)
von: Ioan, Calin Teodor
Veröffentlicht: (2025)
From Dead Pixels to Editable Slides: Infographic Reconstruction into Native Google Slides via Vision-Language Region Understanding
von: Gonzalez, Leonardo
Veröffentlicht: (2026)
von: Gonzalez, Leonardo
Veröffentlicht: (2026)
Beyond Few-shot Object Detection: A Detailed Survey
von: Chudasama, Vishal, et al.
Veröffentlicht: (2024)
von: Chudasama, Vishal, et al.
Veröffentlicht: (2024)
HY-Himmel Technical Report: Hierarchical Interleaved Multi-stream Motion Encoding for Long Video Understanding
von: Jin, Haopeng, et al.
Veröffentlicht: (2026)
von: Jin, Haopeng, et al.
Veröffentlicht: (2026)
Intrinsic Image Fusion for Multi-View 3D Material Reconstruction
von: Kocsis, Peter, et al.
Veröffentlicht: (2025)
von: Kocsis, Peter, et al.
Veröffentlicht: (2025)
IntrinsiX: High-Quality PBR Generation using Image Priors
von: Kocsis, Peter, et al.
Veröffentlicht: (2025)
von: Kocsis, Peter, et al.
Veröffentlicht: (2025)
A Hybrid Deterministic Framework for Named Entity Extraction in Broadcast News Video
von: Lucas, Andrea Filiberto, et al.
Veröffentlicht: (2026)
von: Lucas, Andrea Filiberto, et al.
Veröffentlicht: (2026)
4D Synchronized Fields: Motion-Language Gaussian Splatting for Temporal Scene Understanding
von: Barhdadi, Mohamed Rayan, et al.
Veröffentlicht: (2026)
von: Barhdadi, Mohamed Rayan, et al.
Veröffentlicht: (2026)
Intrinsic Image Diffusion for Indoor Single-view Material Estimation
von: Kocsis, Peter, et al.
Veröffentlicht: (2023)
von: Kocsis, Peter, et al.
Veröffentlicht: (2023)
StoryMovie: A Dataset for Semantic Alignment of Visual Stories with Movie Scripts and Subtitles
von: Oliveira, Daniel, et al.
Veröffentlicht: (2026)
von: Oliveira, Daniel, et al.
Veröffentlicht: (2026)
Perceptual Flow Network for Visually Grounded Reasoning
von: Li, Yangfu, et al.
Veröffentlicht: (2026)
von: Li, Yangfu, et al.
Veröffentlicht: (2026)
Mask-Conditioned Voxel Diffusion for Joint Geometry and Color Inpainting
von: Sumuk, Aarya
Veröffentlicht: (2026)
von: Sumuk, Aarya
Veröffentlicht: (2026)
PhysVideoGenerator: Towards Physically Aware Video Generation via Latent Physics Guidance
von: Satish, Siddarth Nilol Kundur, et al.
Veröffentlicht: (2026)
von: Satish, Siddarth Nilol Kundur, et al.
Veröffentlicht: (2026)
Reducing Object Hallucination in LVLMs via Emphasizing Image-negative Tokens
von: Shen, Meng, et al.
Veröffentlicht: (2026)
von: Shen, Meng, et al.
Veröffentlicht: (2026)
A Simple Baseline for Streaming Video Understanding
von: Shen, Yujiao, et al.
Veröffentlicht: (2026)
von: Shen, Yujiao, et al.
Veröffentlicht: (2026)
SCA-Net: Spatial-Contextual Aggregation Network for Enhanced Small Building and Road Change Detection
von: Gholibeigi, Emad, et al.
Veröffentlicht: (2026)
von: Gholibeigi, Emad, et al.
Veröffentlicht: (2026)
Synthetic-Child: An AIGC-Based Synthetic Data Pipeline for Privacy-Preserving Child Posture Estimation
von: Zeng, Taowen
Veröffentlicht: (2026)
von: Zeng, Taowen
Veröffentlicht: (2026)
Pixel-Level Pavement Distress Assessment Using Instance Segmentation
von: Dewick, Logan, et al.
Veröffentlicht: (2026)
von: Dewick, Logan, et al.
Veröffentlicht: (2026)
Context in object detection: a systematic literature review
von: Jamali, Mahtab, et al.
Veröffentlicht: (2025)
von: Jamali, Mahtab, et al.
Veröffentlicht: (2025)
Pedestrian Detection in Low-Light Conditions: A Comprehensive Survey
von: Ghari, Bahareh, et al.
Veröffentlicht: (2024)
von: Ghari, Bahareh, et al.
Veröffentlicht: (2024)
FlowIBR: Leveraging Pre-Training for Efficient Neural Image-Based Rendering of Dynamic Scenes
von: Büsching, Marcel, et al.
Veröffentlicht: (2023)
von: Büsching, Marcel, et al.
Veröffentlicht: (2023)
IMASHRIMP: Automatic White Shrimp (Penaeus vannamei) Biometrical Analysis from Laboratory Images Using Computer Vision and Deep Learning
von: González, Abiam Remache, et al.
Veröffentlicht: (2025)
von: González, Abiam Remache, et al.
Veröffentlicht: (2025)
OCC-MLLM-CoT-Alpha: Towards Multi-stage Occlusion Recognition Based on Large Language Models via 3D-Aware Supervision and Chain-of-Thoughts Guidance
von: Wang, Chaoyi, et al.
Veröffentlicht: (2025)
von: Wang, Chaoyi, et al.
Veröffentlicht: (2025)
NOAH: Benchmarking Narrative Prior driven Hallucination and Omission in Video Large Language Models
von: Lee, Kyuho, et al.
Veröffentlicht: (2025)
von: Lee, Kyuho, et al.
Veröffentlicht: (2025)
Towards a Generalizable Fusion Architecture for Multimodal Object Detection
von: Berjawi, Jad, et al.
Veröffentlicht: (2025)
von: Berjawi, Jad, et al.
Veröffentlicht: (2025)
EmoVerse: A MLLMs-Driven Emotion Representation Dataset for Interpretable Visual Emotion Analysis
von: Guo, Yijie, et al.
Veröffentlicht: (2025)
von: Guo, Yijie, et al.
Veröffentlicht: (2025)
SimWorld: A Unified Benchmark for Simulator-Conditioned Scene Generation via World Model
von: Li, Xinqing, et al.
Veröffentlicht: (2025)
von: Li, Xinqing, et al.
Veröffentlicht: (2025)
Exploring Surround-View Fisheye Camera 3D Object Detection
von: Li, Changcai, et al.
Veröffentlicht: (2025)
von: Li, Changcai, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025) -
RSTeller: Scaling Up Visual Language Modeling in Remote Sensing with Rich Linguistic Semantics from Openly Available Data and Large Language Models
von: Ge, Junyao, et al.
Veröffentlicht: (2024) -
LATTE: Latent Trajectory Embedding for Diffusion-Generated Image Detection
von: Vasilcoiu, Ana, et al.
Veröffentlicht: (2025) -
DeltaVLM: Interactive Remote Sensing Image Change Analysis via Instruction-guided Difference Perception
von: Deng, Pei, et al.
Veröffentlicht: (2025) -
Interpretable Tau-PET Synthesis from Multimodal T1-Weighted and FLAIR MRI Using Partial Information Decomposition Guided Disentangled Quantized Half-UNet
von: Chopra, Agamdeep S., et al.
Veröffentlicht: (2026)