Visual and Text Prompt Segmentation: A Novel Multi-Model Framework for Remote Sensing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zi, Xing, Jin, Kairui, Tao, Xian, Li, Jun, Braytee, Ali, Shah, Rajiv Ratn, Prasad, Mukesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Prompt-based Multimodal Semantic Communication for Multi-spectral Image Segmentation
von: Zhang, Haoshuo, et al.
Veröffentlicht: (2025)
von: Zhang, Haoshuo, et al.
Veröffentlicht: (2025)
A H.265/HEVC Fine-Grained ROI Video Encryption Algorithm Based on Coding Unit and Prompt Segmentation
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
Understanding What Is Not Said:Referring Remote Sensing Image Segmentation with Scarce Expressions
von: Ye, Kai, et al.
Veröffentlicht: (2025)
von: Ye, Kai, et al.
Veröffentlicht: (2025)
Generative AI for Multimedia Communication: Recent Advances, An Information-Theoretic Framework, and Future Opportunities
von: Jin, Yili, et al.
Veröffentlicht: (2025)
von: Jin, Yili, et al.
Veröffentlicht: (2025)
HybridPrompt: Bridging Generative Priors and Traditional Codecs for Mobile Streaming
von: Liu, Liming, et al.
Veröffentlicht: (2026)
von: Liu, Liming, et al.
Veröffentlicht: (2026)
Compact Visual Data Representation for Green Multimedia -- A Human Visual System Perspective
von: Chen, Peilin, et al.
Veröffentlicht: (2024)
von: Chen, Peilin, et al.
Veröffentlicht: (2024)
Foveated Compression for Immersive Telepresence Visualization
von: Schwarz, Max, et al.
Veröffentlicht: (2025)
von: Schwarz, Max, et al.
Veröffentlicht: (2025)
Audio-Visual Cross-Modal Compression for Generative Face Video Coding
von: Xu, Youmin, et al.
Veröffentlicht: (2025)
von: Xu, Youmin, et al.
Veröffentlicht: (2025)
MarsSQE: Stereo Quality Enhancement for Martian Images Using Bi-level Cross-view Attention
von: Xu, Mai, et al.
Veröffentlicht: (2024)
von: Xu, Mai, et al.
Veröffentlicht: (2024)
NiMark: A Non-intrusive Watermarking Framework against Screen-shooting Attacks
von: Wu, Yufeng, et al.
Veröffentlicht: (2026)
von: Wu, Yufeng, et al.
Veröffentlicht: (2026)
Optimizing Mobile-Friendly Viewport Prediction for Live 360-Degree Video Streaming
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
SatFusion: A Unified Framework for Enhancing Remote Sensing Images via Multi-Frame and Multi-Source Images Fusion
von: Tong, Yufei, et al.
Veröffentlicht: (2025)
von: Tong, Yufei, et al.
Veröffentlicht: (2025)
Raster Domain Text Steganography: A Unified Framework for Multimodal Secure Embedding
von: Kandala, A V Uday Kiran
Veröffentlicht: (2025)
von: Kandala, A V Uday Kiran
Veröffentlicht: (2025)
Adaptive Wireless Image Semantic Transmission and Over-The-Air Testing
von: Ding, Jiarun, et al.
Veröffentlicht: (2024)
von: Ding, Jiarun, et al.
Veröffentlicht: (2024)
Generative Flow Networks for Personalized Multimedia Systems: A Case Study on Short Video Feeds
von: Jin, Yili, et al.
Veröffentlicht: (2025)
von: Jin, Yili, et al.
Veröffentlicht: (2025)
Symmetric Entropy-Constrained Video Coding for Machines
von: Sun, Yuxiao, et al.
Veröffentlicht: (2025)
von: Sun, Yuxiao, et al.
Veröffentlicht: (2025)
A Visual Perception-Based Tunable Framework and Evaluation Benchmark for H.265/HEVC ROI Encryption
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
AeroLite: Tag-Guided Lightweight Generation of Aerial Image Captions
von: Zi, Xing, et al.
Veröffentlicht: (2025)
von: Zi, Xing, et al.
Veröffentlicht: (2025)
Change Detection Between Optical Remote Sensing Imagery and Map Data via Segment Anything Model (SAM)
von: Chen, Hongruixuan, et al.
Veröffentlicht: (2024)
von: Chen, Hongruixuan, et al.
Veröffentlicht: (2024)
PAL: Prompting Analytic Learning with Missing Modality for Multi-Modal Class-Incremental Learning
von: Yue, Xianghu, et al.
Veröffentlicht: (2025)
von: Yue, Xianghu, et al.
Veröffentlicht: (2025)
Robust Live Streaming over LEO Satellite Constellations: Measurement, Analysis, and Handover-Aware Adaptation
von: Fang, Hao, et al.
Veröffentlicht: (2025)
von: Fang, Hao, et al.
Veröffentlicht: (2025)
Enhanced Template-based Intra Mode Derivation with Adaptive Block Vector Replacement
von: Zhang, Jiaqi, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaqi, et al.
Veröffentlicht: (2025)
Transform and Entropy Coding in AV2
von: Nalci, Alican, et al.
Veröffentlicht: (2026)
von: Nalci, Alican, et al.
Veröffentlicht: (2026)
Tube-Structured Incremental Semantic HARQ for Generative Video Receivers
von: Wang, Xuesong, et al.
Veröffentlicht: (2026)
von: Wang, Xuesong, et al.
Veröffentlicht: (2026)
Rate-Quality or Energy-Quality Pareto Fronts for Adaptive Video Streaming?
von: Katsenou, Angeliki, et al.
Veröffentlicht: (2024)
von: Katsenou, Angeliki, et al.
Veröffentlicht: (2024)
Smaller is Better: Generative Models Can Power Short Video Preloading
von: Liu, Liming, et al.
Veröffentlicht: (2026)
von: Liu, Liming, et al.
Veröffentlicht: (2026)
Camel: Frame-Level Bandwidth Estimation for Low-Latency Live Streaming under Video Bitrate Undershooting
von: Liu, Liming, et al.
Veröffentlicht: (2026)
von: Liu, Liming, et al.
Veröffentlicht: (2026)
Fast Multirate Encoding for 360° Video in OMAF Streaming Workflows
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
Decoding Complexity-Rate-Quality Pareto-Front for Adaptive VVC Streaming
von: Katsenou, Angeliki, et al.
Veröffentlicht: (2024)
von: Katsenou, Angeliki, et al.
Veröffentlicht: (2024)
Memory-Anchored Multimodal Reasoning for Explainable Video Forensics
von: Chen, Chen, et al.
Veröffentlicht: (2025)
von: Chen, Chen, et al.
Veröffentlicht: (2025)
GScomp-QA: A Subjective Dataset for Quality Assessment of Compressed Gaussian Splatting
von: Martin, Pedro, et al.
Veröffentlicht: (2026)
von: Martin, Pedro, et al.
Veröffentlicht: (2026)
Unravelling the Power of Single-Pass Look-Ahead in Modern Codecs for Optimized Transcoding Deployment
von: Vibhoothi, Vibhoothi, et al.
Veröffentlicht: (2024)
von: Vibhoothi, Vibhoothi, et al.
Veröffentlicht: (2024)
Perception-Aware Video Semantic Communication
von: Huang, Yinhuan, et al.
Veröffentlicht: (2026)
von: Huang, Yinhuan, et al.
Veröffentlicht: (2026)
Video Compression Beyond VVC: Quantitative Analysis of Intra Coding Tools in Enhanced Compression Model (ECM)
von: Abdoli, Mohsen, et al.
Veröffentlicht: (2024)
von: Abdoli, Mohsen, et al.
Veröffentlicht: (2024)
Dynamic resolution switching for live streaming
von: Xiong, Xin, et al.
Veröffentlicht: (2026)
von: Xiong, Xin, et al.
Veröffentlicht: (2026)
Enhanced Quality Aware-Scalable Underwater Image Compression
von: Zhu, Linwei, et al.
Veröffentlicht: (2025)
von: Zhu, Linwei, et al.
Veröffentlicht: (2025)
Partition Tree Search Acceleration for VVC: Survey and Evaluation with VTM Evolution
von: Kherchouche, M. E. A., et al.
Veröffentlicht: (2026)
von: Kherchouche, M. E. A., et al.
Veröffentlicht: (2026)
H.265/HEVC Video Steganalysis Based on CU Block Structure Gradients and IPM Mapping
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
Sandwiched Compression: Repurposing Standard Codecs with Neural Network Wrappers
von: Guleryuz, Onur G., et al.
Veröffentlicht: (2024)
von: Guleryuz, Onur G., et al.
Veröffentlicht: (2024)
Encoding Time and Energy Model for SVT-AV1 based on Video Complexity
von: Eichermüller, Lena, et al.
Veröffentlicht: (2024)
von: Eichermüller, Lena, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Prompt-based Multimodal Semantic Communication for Multi-spectral Image Segmentation
von: Zhang, Haoshuo, et al.
Veröffentlicht: (2025) -
A H.265/HEVC Fine-Grained ROI Video Encryption Algorithm Based on Coding Unit and Prompt Segmentation
von: Zhang, Xiang, et al.
Veröffentlicht: (2026) -
Understanding What Is Not Said:Referring Remote Sensing Image Segmentation with Scarce Expressions
von: Ye, Kai, et al.
Veröffentlicht: (2025) -
Generative AI for Multimedia Communication: Recent Advances, An Information-Theoretic Framework, and Future Opportunities
von: Jin, Yili, et al.
Veröffentlicht: (2025) -
HybridPrompt: Bridging Generative Priors and Traditional Codecs for Mobile Streaming
von: Liu, Liming, et al.
Veröffentlicht: (2026)