RemoteZero: Geospatial Reasoning with Zero Human Annotations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yao, Liang, Liu, Fan, Xu, Shengxiang, Zhang, Chuanyi, Min, Rui, Di, Shimin, Zheng, Yuhui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RemoteReasoner: Towards Unifying Geospatial Reasoning Workflow
von: Yao, Liang, et al.
Veröffentlicht: (2025)
von: Yao, Liang, et al.
Veröffentlicht: (2025)
RemoteShield: Enable Robust Multimodal Large Language Models for Earth Observation
von: Min, Rui, et al.
Veröffentlicht: (2026)
von: Min, Rui, et al.
Veröffentlicht: (2026)
RemoteAgent: Bridging Vague Human Intents and Earth Observation with RL-based Agentic MLLMs
von: Yao, Liang, et al.
Veröffentlicht: (2026)
von: Yao, Liang, et al.
Veröffentlicht: (2026)
RemoteSAM: Towards Segment Anything for Earth Observation
von: Yao, Liang, et al.
Veröffentlicht: (2025)
von: Yao, Liang, et al.
Veröffentlicht: (2025)
Chain-of-Talkers (CoTalk): Fast Human Annotation of Dense Image Captions
von: Shen, Yijun, et al.
Veröffentlicht: (2025)
von: Shen, Yijun, et al.
Veröffentlicht: (2025)
RemoteTrimmer: Adaptive Structural Pruning for Remote Sensing Image Classification
von: Zou, Guangwenjie, et al.
Veröffentlicht: (2024)
von: Zou, Guangwenjie, et al.
Veröffentlicht: (2024)
UEMM-Air: Make Unmanned Aerial Vehicles Perform More Multi-modal Tasks
von: Yao, Liang, et al.
Veröffentlicht: (2024)
von: Yao, Liang, et al.
Veröffentlicht: (2024)
GeoZero: Incentivizing Reasoning from Scratch on Geospatial Scenes
von: Wang, Di, et al.
Veröffentlicht: (2025)
von: Wang, Di, et al.
Veröffentlicht: (2025)
Disentangle Object and Non-object Infrared Features via Language Guidance
von: Liu, Fan, et al.
Veröffentlicht: (2026)
von: Liu, Fan, et al.
Veröffentlicht: (2026)
Evaluating Remote Sensing Image Captions Beyond Metric Biases
von: Chen, Ziyun, et al.
Veröffentlicht: (2026)
von: Chen, Ziyun, et al.
Veröffentlicht: (2026)
Combating Noisy Labels via Dynamic Connection Masking
von: Zhang, Xinlei, et al.
Veröffentlicht: (2025)
von: Zhang, Xinlei, et al.
Veröffentlicht: (2025)
Geospatial-Reasoning-Driven Vocabulary-Agnostic Remote Sensing Semantic Segmentation
von: Zhou, Chufeng, et al.
Veröffentlicht: (2026)
von: Zhou, Chufeng, et al.
Veröffentlicht: (2026)
Unlocking Zero-Shot Geospatial Reasoning via Indirect Rewards
von: Xu, Chenhui, et al.
Veröffentlicht: (2025)
von: Xu, Chenhui, et al.
Veröffentlicht: (2025)
Domain-invariant Progressive Knowledge Distillation for UAV-based Object Detection
von: Yao, Liang, et al.
Veröffentlicht: (2024)
von: Yao, Liang, et al.
Veröffentlicht: (2024)
BVINet: Unlocking Blind Video Inpainting with Zero Annotations
von: Wu, Zhiliang, et al.
Veröffentlicht: (2025)
von: Wu, Zhiliang, et al.
Veröffentlicht: (2025)
V-Zero: Self-Improving Multimodal Reasoning with Zero Annotation
von: Wang, Han, et al.
Veröffentlicht: (2026)
von: Wang, Han, et al.
Veröffentlicht: (2026)
Simultaneous Image-to-Zero and Zero-to-Noise: Diffusion Models with Analytical Image Attenuation
von: Huang, Yuhang, et al.
Veröffentlicht: (2023)
von: Huang, Yuhang, et al.
Veröffentlicht: (2023)
Towards Real Zero-Shot Camouflaged Object Segmentation without Camouflaged Annotations
von: Lei, Cheng, et al.
Veröffentlicht: (2024)
von: Lei, Cheng, et al.
Veröffentlicht: (2024)
Prompting DirectSAM for Semantic Contour Extraction in Remote Sensing Images
von: Miao, Shiyu, et al.
Veröffentlicht: (2024)
von: Miao, Shiyu, et al.
Veröffentlicht: (2024)
GeoViS: Geospatially Rewarded Visual Search for Remote Sensing Visual Grounding
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
Boost UAV-based Ojbect Detection via Scale-Invariant Feature Disentanglement and Adversarial Learning
von: Liu, Fan, et al.
Veröffentlicht: (2024)
von: Liu, Fan, et al.
Veröffentlicht: (2024)
GaussianBody: Clothed Human Reconstruction via 3d Gaussian Splatting
von: Li, Mengtian, et al.
Veröffentlicht: (2024)
von: Li, Mengtian, et al.
Veröffentlicht: (2024)
Agent3D-Zero: An Agent for Zero-shot 3D Understanding
von: Zhang, Sha, et al.
Veröffentlicht: (2024)
von: Zhang, Sha, et al.
Veröffentlicht: (2024)
FG-MDM: Towards Zero-Shot Human Motion Generation via ChatGPT-Refined Descriptions
von: Shi, Xu, et al.
Veröffentlicht: (2023)
von: Shi, Xu, et al.
Veröffentlicht: (2023)
RSVG-ZeroOV: Exploring a Training-Free Framework for Zero-Shot Open-Vocabulary Visual Grounding in Remote Sensing Images
von: Li, Ke, et al.
Veröffentlicht: (2025)
von: Li, Ke, et al.
Veröffentlicht: (2025)
ZeroStereo: Zero-shot Stereo Matching from Single Images
von: Wang, Xianqi, et al.
Veröffentlicht: (2025)
von: Wang, Xianqi, et al.
Veröffentlicht: (2025)
Multi-Scale Gaussian-Language Map for Zero-shot Embodied Navigation and Reasoning
von: Zhang, Sixian, et al.
Veröffentlicht: (2026)
von: Zhang, Sixian, et al.
Veröffentlicht: (2026)
Zero-shot Composed Text-Image Retrieval
von: Liu, Yikun, et al.
Veröffentlicht: (2023)
von: Liu, Yikun, et al.
Veröffentlicht: (2023)
ReasonNavi: Human-Inspired Global Map Reasoning for Zero-Shot Embodied Navigation
von: Ao, Yuzhuo, et al.
Veröffentlicht: (2026)
von: Ao, Yuzhuo, et al.
Veröffentlicht: (2026)
CrossEarth: Geospatial Vision Foundation Model for Domain Generalizable Remote Sensing Semantic Segmentation
von: Gong, Ziyang, et al.
Veröffentlicht: (2024)
von: Gong, Ziyang, et al.
Veröffentlicht: (2024)
CFG-Zero*: Improved Classifier-Free Guidance for Flow Matching Models
von: Fan, Weichen, et al.
Veröffentlicht: (2025)
von: Fan, Weichen, et al.
Veröffentlicht: (2025)
Bridging Semantics and Geometry: A Decoupled LVLM-SAM Framework for Reasoning Segmentation in Optical Remote Sensing
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
ZeroSCD: Zero-Shot Street Scene Change Detection
von: Kannan, Shyam Sundar, et al.
Veröffentlicht: (2024)
von: Kannan, Shyam Sundar, et al.
Veröffentlicht: (2024)
Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion
von: Zhou, Zhenglin, et al.
Veröffentlicht: (2025)
von: Zhou, Zhenglin, et al.
Veröffentlicht: (2025)
ZONE: Zero-Shot Instruction-Guided Local Editing
von: Li, Shanglin, et al.
Veröffentlicht: (2023)
von: Li, Shanglin, et al.
Veröffentlicht: (2023)
AvatarGO: Zero-shot 4D Human-Object Interaction Generation and Animation
von: Cao, Yukang, et al.
Veröffentlicht: (2024)
von: Cao, Yukang, et al.
Veröffentlicht: (2024)
Performance of Human Annotators in Object Detection and Segmentation of Remotely Sensed Data
von: Blushtein-Livnon, Roni, et al.
Veröffentlicht: (2024)
von: Blushtein-Livnon, Roni, et al.
Veröffentlicht: (2024)
MSNav: Zero-Shot Vision-and-Language Navigation with Dynamic Memory and LLM Spatial Reasoning
von: Liu, Chenghao, et al.
Veröffentlicht: (2025)
von: Liu, Chenghao, et al.
Veröffentlicht: (2025)
ID-Animator: Zero-Shot Identity-Preserving Human Video Generation
von: He, Xuanhua, et al.
Veröffentlicht: (2024)
von: He, Xuanhua, et al.
Veröffentlicht: (2024)
ZeroMamba: Exploring Visual State Space Model for Zero-Shot Learning
von: Hou, Wenjin, et al.
Veröffentlicht: (2024)
von: Hou, Wenjin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
RemoteReasoner: Towards Unifying Geospatial Reasoning Workflow
von: Yao, Liang, et al.
Veröffentlicht: (2025) -
RemoteShield: Enable Robust Multimodal Large Language Models for Earth Observation
von: Min, Rui, et al.
Veröffentlicht: (2026) -
RemoteAgent: Bridging Vague Human Intents and Earth Observation with RL-based Agentic MLLMs
von: Yao, Liang, et al.
Veröffentlicht: (2026) -
RemoteSAM: Towards Segment Anything for Earth Observation
von: Yao, Liang, et al.
Veröffentlicht: (2025) -
Chain-of-Talkers (CoTalk): Fast Human Annotation of Dense Image Captions
von: Shen, Yijun, et al.
Veröffentlicht: (2025)