Text Before Vision: Staged Knowledge Injection Matters for Agentic RLVR in Ultra-High-Resolution Remote Sensing Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Fengxiang, Chen, Mingshuo, Li, Yueying, Yang, Yajie, Zhou, Yuhao, Wang, Di, Zhang, Yifan, Wang, Haoyu, Zhao, Haiyan, Sun, Hongda, Lan, Long, Song, Jun, Wang, Yulin, Zhang, Jing, Zhang, Wenlong, Du, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GeoEyes: On-Demand Visual Focusing for Evidence-Grounded Understanding of Ultra-High-Resolution Remote Sensing Imagery
von: Wang, Fengxiang, et al.
Veröffentlicht: (2026)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2026)
Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding
von: Li, Yueying, et al.
Veröffentlicht: (2026)
von: Li, Yueying, et al.
Veröffentlicht: (2026)
XLRS-Bench: Could Your Multimodal LLMs Understand Extremely Large Ultra-High-Resolution Remote Sensing Imagery?
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
RoMA: Scaling up Mamba-based Foundation Models for Remote Sensing
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
GeoLLaVA-8K: Scaling Remote-Sensing Multimodal Large Language Models to 8K Resolution
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
ImageRAG: Enhancing Ultra High Resolution Remote Sensing Imagery Analysis with ImageRAG
von: Zhang, Zilun, et al.
Veröffentlicht: (2024)
von: Zhang, Zilun, et al.
Veröffentlicht: (2024)
A Benchmark for Ultra-High-Resolution Remote Sensing MLLMs
von: Dang, Yunkai, et al.
Veröffentlicht: (2025)
von: Dang, Yunkai, et al.
Veröffentlicht: (2025)
GeoViS: Geospatially Rewarded Visual Search for Remote Sensing Visual Grounding
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
Demystifying and Detecting Agentic Workflow Injection Vulnerabilities in GitHub Actions
von: Wang, Shenao, et al.
Veröffentlicht: (2026)
von: Wang, Shenao, et al.
Veröffentlicht: (2026)
Multi-Perspective Subimage CLIP with Keyword Guidance for Remote Sensing Image-Text Retrieval
von: Li, Yifan, et al.
Veröffentlicht: (2026)
von: Li, Yifan, et al.
Veröffentlicht: (2026)
Development and Validation of a Prediction Model for Respiratory Failure in Patients With Sepsis‐Associated Acute Kidney Injury (SA‐AKI) Within 48 Hours of Admission
von: Bin Wang, et al.
Veröffentlicht: (2025)
von: Bin Wang, et al.
Veröffentlicht: (2025)
Diffusion Enhancement for Cloud Removal in Ultra-Resolution Remote Sensing Imagery
von: Sui, Jialu, et al.
Veröffentlicht: (2024)
von: Sui, Jialu, et al.
Veröffentlicht: (2024)
"Your AI, My Shell": Demystifying Prompt Injection Attacks on Agentic AI Coding Editors
von: Liu, Yue, et al.
Veröffentlicht: (2025)
von: Liu, Yue, et al.
Veröffentlicht: (2025)
GeoLink: Empowering Remote Sensing Foundation Model with OpenStreetMap Data
von: Bai, Lubian, et al.
Veröffentlicht: (2025)
von: Bai, Lubian, et al.
Veröffentlicht: (2025)
LMFNet: An Efficient Multimodal Fusion Approach for Semantic Segmentation in High-Resolution Remote Sensing
von: Wang, Tong, et al.
Veröffentlicht: (2024)
von: Wang, Tong, et al.
Veröffentlicht: (2024)
Dual Consensus: Escaping from Spurious Majority in Unsupervised RLVR via Two-Stage Vote Mechanism
von: Du, Kaixuan, et al.
Veröffentlicht: (2026)
von: Du, Kaixuan, et al.
Veröffentlicht: (2026)
RSEdit: Text-Guided Image Editing for Remote Sensing
von: Zhenyuan, Chen, et al.
Veröffentlicht: (2026)
von: Zhenyuan, Chen, et al.
Veröffentlicht: (2026)
MSSDF: Modality-Shared Self-supervised Distillation for High-Resolution Multi-modal Remote Sensing Image Learning
von: Wang, Tong, et al.
Veröffentlicht: (2025)
von: Wang, Tong, et al.
Veröffentlicht: (2025)
Dual-Stage Safe Herding Framework for Adversarial Attacker in Dynamic Environment
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)
Not only where, But when: Temporal Scheduling for RLVR
von: Zhang, Jinghao, et al.
Veröffentlicht: (2026)
von: Zhang, Jinghao, et al.
Veröffentlicht: (2026)
Repair Before Veto: Repair-Augmented Constraint Learning for Contextual Decisions
von: Wang, Yifan
Veröffentlicht: (2026)
von: Wang, Yifan
Veröffentlicht: (2026)
GeoVista: Visually Grounded Active Perception for Ultra-High-Resolution Remote Sensing Understanding
von: Zhu, Jiashun, et al.
Veröffentlicht: (2026)
von: Zhu, Jiashun, et al.
Veröffentlicht: (2026)
iEBAKER: Improved Remote Sensing Image-Text Retrieval Framework via Eliminate Before Align and Keyword Explicit Reasoning
von: Zhang, Yan, et al.
Veröffentlicht: (2025)
von: Zhang, Yan, et al.
Veröffentlicht: (2025)
UHR-DETR: Efficient End-to-End Small Object Detection for Ultra-High-Resolution Remote Sensing Imagery
von: Li, Jingfang, et al.
Veröffentlicht: (2026)
von: Li, Jingfang, et al.
Veröffentlicht: (2026)
Part II: ROLL Flash -- Accelerating RLVR and Agentic Training with Asynchrony
von: Lu, Han, et al.
Veröffentlicht: (2025)
von: Lu, Han, et al.
Veröffentlicht: (2025)
Before the Model Learns the Bug:Fuzzing RLVR Verifiers
von: Ray, Jaideep
Veröffentlicht: (2026)
von: Ray, Jaideep
Veröffentlicht: (2026)
The Unlearnability Phenomenon in RLVR for Language Models
von: Chen, Yulin, et al.
Veröffentlicht: (2026)
von: Chen, Yulin, et al.
Veröffentlicht: (2026)
AdapTools: Adaptive Tool-based Indirect Prompt Injection Attacks on Agentic LLMs
von: Wang, Che, et al.
Veröffentlicht: (2026)
von: Wang, Che, et al.
Veröffentlicht: (2026)
S5: Scalable Semi-Supervised Semantic Segmentation in Remote Sensing
von: Lv, Liang, et al.
Veröffentlicht: (2025)
von: Lv, Liang, et al.
Veröffentlicht: (2025)
Ultra-Low Complexity On-Orbit Compression for Remote Sensing Imagery via Block Modulated Imaging
von: Wang, Zhibin, et al.
Veröffentlicht: (2024)
von: Wang, Zhibin, et al.
Veröffentlicht: (2024)
FreeText: Training-Free Text Rendering in Diffusion Transformers via Attention Localization and Spectral Glyph Injection
von: Zhang, Ruiqiang, et al.
Veröffentlicht: (2026)
von: Zhang, Ruiqiang, et al.
Veröffentlicht: (2026)
Comprehensive Analysis of the Correlation Between Immune‐Inflammatory Biomarkers and Intestinal Necrosis in Patients With Acute Intestinal Ischemia
von: Yu Tian, et al.
Veröffentlicht: (2025)
von: Yu Tian, et al.
Veröffentlicht: (2025)
SkeletonAgent: An Agentic Interaction Framework for Skeleton-based Action Recognition
von: Liu, Hongda, et al.
Veröffentlicht: (2025)
von: Liu, Hongda, et al.
Veröffentlicht: (2025)
Moduli spaces and the algebra of conformal blocks
von: Zhang, Yanglong, et al.
Veröffentlicht: (2026)
von: Zhang, Yanglong, et al.
Veröffentlicht: (2026)
Fault‐Tolerant Control for Path Following of Underactuated Autonomous Underwater Vehicle With Input Delays and Output Constraints
von: Hao Wang, et al.
Veröffentlicht: (2025)
von: Hao Wang, et al.
Veröffentlicht: (2025)
New Skills or Sharper Primitives? A Probabilistic Perspective on the Emergence of Reasoning in RLVR
von: Wang, Zhilin, et al.
Veröffentlicht: (2026)
von: Wang, Zhilin, et al.
Veröffentlicht: (2026)
Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing
von: Chen, Xi, et al.
Veröffentlicht: (2026)
von: Chen, Xi, et al.
Veröffentlicht: (2026)
Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers
von: Wang, Haochuan Kevin, et al.
Veröffentlicht: (2026)
von: Wang, Haochuan Kevin, et al.
Veröffentlicht: (2026)
Co-Training Vision Language Models for Remote Sensing Multi-task Learning
von: Li, Qingyun, et al.
Veröffentlicht: (2025)
von: Li, Qingyun, et al.
Veröffentlicht: (2025)
GRASP: Guided Region-Aware Sparse Prompting for Adapting MLLMs to Remote Sensing
von: Sun, Qigan, et al.
Veröffentlicht: (2026)
von: Sun, Qigan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
GeoEyes: On-Demand Visual Focusing for Evidence-Grounded Understanding of Ultra-High-Resolution Remote Sensing Imagery
von: Wang, Fengxiang, et al.
Veröffentlicht: (2026) -
Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding
von: Li, Yueying, et al.
Veröffentlicht: (2026) -
XLRS-Bench: Could Your Multimodal LLMs Understand Extremely Large Ultra-High-Resolution Remote Sensing Imagery?
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025) -
RoMA: Scaling up Mamba-based Foundation Models for Remote Sensing
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025) -
GeoLLaVA-8K: Scaling Remote-Sensing Multimodal Large Language Models to 8K Resolution
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)