Any2Any: Unified Arbitrary Modality Translation for Remote Sensing
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Haoyang, Zhang, Jing, Wang, Hebaixu, Wang, Shiqin, Huang, Pohsun, Li, Jiayuan, Guo, Haonan, Wang, Di, Wang, Zheng, Du, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Residual Diffusion Bridge Model for Image Restoration
by: Wang, Hebaixu, et al.
Published: (2025)
by: Wang, Hebaixu, et al.
Published: (2025)
DGSolver: Diffusion Generalist Solver with Universal Posterior Sampling for Image Restoration
by: Wang, Hebaixu, et al.
Published: (2025)
by: Wang, Hebaixu, et al.
Published: (2025)
Universal Pansharpening Foundation Model
by: Wang, Hebaixu, et al.
Published: (2026)
by: Wang, Hebaixu, et al.
Published: (2026)
VLRS-Bench: A Vision-Language Reasoning Benchmark for Remote Sensing
by: Luo, Zhiming, et al.
Published: (2026)
by: Luo, Zhiming, et al.
Published: (2026)
AnyAD: Unified Any-Modality Anomaly Detection in Incomplete Multi-Sequence MRI
by: Wu, Changwei, et al.
Published: (2025)
by: Wu, Changwei, et al.
Published: (2025)
Seeing Clearly without Training: Mitigating Hallucinations in Multimodal LLMs for Remote Sensing
by: Liu, Yi, et al.
Published: (2026)
by: Liu, Yi, et al.
Published: (2026)
Learning Modality-agnostic Representation for Semantic Segmentation from Any Modalities
by: Zheng, Xu, et al.
Published: (2024)
by: Zheng, Xu, et al.
Published: (2024)
FysicsWorld: A Unified Full-Modality Benchmark for Any-to-Any Understanding, Generation, and Reasoning
by: Jiang, Yue, et al.
Published: (2025)
by: Jiang, Yue, et al.
Published: (2025)
Animate Any Character in Any World
by: Wang, Yitong, et al.
Published: (2025)
by: Wang, Yitong, et al.
Published: (2025)
SARMAE: Masked Autoencoder for SAR Representation Learning
by: Liu, Danxu, et al.
Published: (2025)
by: Liu, Danxu, et al.
Published: (2025)
AnyCrowd: Instance-Isolated Identity-Pose Binding for Arbitrary Multi-Character Animation
by: Xie, Zhenyu, et al.
Published: (2026)
by: Xie, Zhenyu, et al.
Published: (2026)
UniM: A Unified Any-to-Any Interleaved Multimodal Benchmark
by: Li, Yanlin, et al.
Published: (2026)
by: Li, Yanlin, et al.
Published: (2026)
Remote Sensing ChatGPT: Solving Remote Sensing Tasks with ChatGPT and Visual Models
by: Guo, Haonan, et al.
Published: (2024)
by: Guo, Haonan, et al.
Published: (2024)
Any-to-Any Learning in Computational Pathology via Triplet Multimodal Pretraining
by: Sun, Qichen, et al.
Published: (2025)
by: Sun, Qichen, et al.
Published: (2025)
MTP: Advancing Remote Sensing Foundation Model via Multi-Task Pretraining
by: Wang, Di, et al.
Published: (2024)
by: Wang, Di, et al.
Published: (2024)
Segment Any Motion in Videos
by: Huang, Nan, et al.
Published: (2025)
by: Huang, Nan, et al.
Published: (2025)
Structurally Prune Anything: Any Architecture, Any Framework, Any Time
by: Wang, Xun, et al.
Published: (2024)
by: Wang, Xun, et al.
Published: (2024)
FreeTuner: Any Subject in Any Style with Training-free Diffusion
by: Xu, Youcan, et al.
Published: (2024)
by: Xu, Youcan, et al.
Published: (2024)
Insight Any Instance: Promptable Instance Segmentation for Remote Sensing Images
by: Li, Xuexue
Published: (2024)
by: Li, Xuexue
Published: (2024)
Any Model, Any Place, Any Time: Get Remote Sensing Foundation Model Embeddings On Demand
by: Ye, Dingqi, et al.
Published: (2026)
by: Ye, Dingqi, et al.
Published: (2026)
Depth Anything with Any Prior
by: Wang, Zehan, et al.
Published: (2025)
by: Wang, Zehan, et al.
Published: (2025)
Any-to-Bokeh: Arbitrary-Subject Video Refocusing with Video Diffusion Model
by: Yang, Yang, et al.
Published: (2025)
by: Yang, Yang, et al.
Published: (2025)
MDReID: Modality-Decoupled Learning for Any-to-Any Multi-Modal Object Re-Identification
by: Feng, Yingying, et al.
Published: (2025)
by: Feng, Yingying, et al.
Published: (2025)
AnyBand-Diff: A Unified Remote Sensing Image Generation and Band Repair Framework with Spectral Priors
by: Zhao, Zuopeng, et al.
Published: (2026)
by: Zhao, Zuopeng, et al.
Published: (2026)
Anything in Any Scene: Photorealistic Video Object Insertion
by: Bai, Chen, et al.
Published: (2024)
by: Bai, Chen, et al.
Published: (2024)
Any-Optical-Model: A Universal Foundation Model for Optical Remote Sensing
by: Li, Xuyang, et al.
Published: (2025)
by: Li, Xuyang, et al.
Published: (2025)
ClipGStream: Clip-Stream Gaussian Splatting for Any Length and Any Motion Multi-View Dynamic Scene Reconstruction
by: Liang, Jie, et al.
Published: (2026)
by: Liang, Jie, et al.
Published: (2026)
LeMeViT: Efficient Vision Transformer with Learnable Meta Tokens for Remote Sensing Image Interpretation
by: Jiang, Wentao, et al.
Published: (2024)
by: Jiang, Wentao, et al.
Published: (2024)
TrackAny3D: Transferring Pretrained 3D Models for Category-unified 3D Point Cloud Tracking
by: Wang, Mengmeng, et al.
Published: (2025)
by: Wang, Mengmeng, et al.
Published: (2025)
AnyEdit: Mastering Unified High-Quality Image Editing for Any Idea
by: Yu, Qifan, et al.
Published: (2024)
by: Yu, Qifan, et al.
Published: (2024)
Judge Anything: MLLM as a Judge Across Any Modality
by: Pu, Shu, et al.
Published: (2025)
by: Pu, Shu, et al.
Published: (2025)
HorizonForge: Driving Scene Editing with Any Trajectories and Any Vehicles
by: Wang, Yifan, et al.
Published: (2026)
by: Wang, Yifan, et al.
Published: (2026)
AURORA-KITTI: Any-Weather Depth Completion and Denoising in the Wild
by: Wang, Yiting, et al.
Published: (2026)
by: Wang, Yiting, et al.
Published: (2026)
AnyRecon: Arbitrary-View 3D Reconstruction with Video Diffusion Model
by: Chen, Yutian, et al.
Published: (2026)
by: Chen, Yutian, et al.
Published: (2026)
AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling
by: Li, Yiheng, et al.
Published: (2026)
by: Li, Yiheng, et al.
Published: (2026)
AnyPcc: Compressing Any Point Cloud with a Single Universal Model
by: Wang, Kangli, et al.
Published: (2025)
by: Wang, Kangli, et al.
Published: (2025)
Any2Caption:Interpreting Any Condition to Caption for Controllable Video Generation
by: Wu, Shengqiong, et al.
Published: (2025)
by: Wu, Shengqiong, et al.
Published: (2025)
OutfitAnyone: Ultra-high Quality Virtual Try-On for Any Clothing and Any Person
by: Sun, Ke, et al.
Published: (2024)
by: Sun, Ke, et al.
Published: (2024)
AnyTrans: Translate AnyText in the Image with Large Scale Models
by: Qian, Zhipeng, et al.
Published: (2024)
by: Qian, Zhipeng, et al.
Published: (2024)
Single-Model and Any-Modality for Video Object Tracking
by: Wu, Zongwei, et al.
Published: (2023)
by: Wu, Zongwei, et al.
Published: (2023)
Similar Items
-
Residual Diffusion Bridge Model for Image Restoration
by: Wang, Hebaixu, et al.
Published: (2025) -
DGSolver: Diffusion Generalist Solver with Universal Posterior Sampling for Image Restoration
by: Wang, Hebaixu, et al.
Published: (2025) -
Universal Pansharpening Foundation Model
by: Wang, Hebaixu, et al.
Published: (2026) -
VLRS-Bench: A Vision-Language Reasoning Benchmark for Remote Sensing
by: Luo, Zhiming, et al.
Published: (2026) -
AnyAD: Unified Any-Modality Anomaly Detection in Incomplete Multi-Sequence MRI
by: Wu, Changwei, et al.
Published: (2025)