Saved in:
| Main Authors: | Wang, Xinghao, Liu, Feng, Su, Rui, Wang, Zhihui, Fang, Lihua, Zhou, Lianqing, Bai, Lei, Ouyang, Wanli |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2502.19960 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast Information Streaming Handler (FisH): A Unified Seismic Neural Network for Single Station Real-Time Earthquake Early Warning
by: Zhang, Tianning, et al.
Published: (2024)
by: Zhang, Tianning, et al.
Published: (2024)
Geometry-enhanced Pre-training on Interatomic Potentials
by: Cui, Taoyong, et al.
Published: (2023)
by: Cui, Taoyong, et al.
Published: (2023)
DispFormer: A Pretrained Transformer Incorporating Physical Constraints for Dispersion Curve Inversion
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
Progressive Cross-Stream Cooperation in Spatial and Temporal Domain for Action Localization
by: Su, Rui, et al.
Published: (2019)
by: Su, Rui, et al.
Published: (2019)
OpenSWI: A Massive-Scale Benchmark Dataset for Surface Wave Dispersion Curve Inversion
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
A Mutil-conditional Diffusion Transformer for Versatile Seismic Wave Generation
by: Duan, Longfei, et al.
Published: (2025)
by: Duan, Longfei, et al.
Published: (2025)
Multi-modal Vision Pre-training for Medical Image Analysis
by: Rui, Shaohao, et al.
Published: (2024)
by: Rui, Shaohao, et al.
Published: (2024)
ComfyBench: Benchmarking LLM-based Agents in ComfyUI for Autonomously Designing Collaborative AI Systems
by: Xue, Xiangyuan, et al.
Published: (2024)
by: Xue, Xiangyuan, et al.
Published: (2024)
HVDistill: Transferring Knowledge from Images to Point Clouds via Unsupervised Hybrid-View Distillation
by: Zhang, Sha, et al.
Published: (2024)
by: Zhang, Sha, et al.
Published: (2024)
Native-Resolution Image Synthesis
by: Wang, Zidong, et al.
Published: (2025)
by: Wang, Zidong, et al.
Published: (2025)
Med-Banana-50K: A Cross-modality Large-Scale Dataset for Text-guided Medical Image Editing
by: Chen, Zhihui, et al.
Published: (2025)
by: Chen, Zhihui, et al.
Published: (2025)
Improving Weakly Supervised Temporal Action Localization by Exploiting Multi-resolution Information in Temporal Domain
by: Su, Rui, et al.
Published: (2025)
by: Su, Rui, et al.
Published: (2025)
SeisLM: a Foundation Model for Seismic Waveforms
by: Liu, Tianlin, et al.
Published: (2024)
by: Liu, Tianlin, et al.
Published: (2024)
Multi-modal Mood Reader: Pre-trained Model Empowers Cross-Subject Emotion Recognition
by: Dong, Yihang, et al.
Published: (2024)
by: Dong, Yihang, et al.
Published: (2024)
Uni3D-LLM: Unifying Point Cloud Perception, Generation and Editing with Large Language Models
by: Liu, Dingning, et al.
Published: (2024)
by: Liu, Dingning, et al.
Published: (2024)
Seis2Rock: A Data-Driven Approach to Direct Petrophysical Inversion of Pre-Stack Seismic Data
by: Corrales, Miguel, et al.
Published: (2023)
by: Corrales, Miguel, et al.
Published: (2023)
SeisRDT: Latent Diffusion Model Based On Representation Learning For Seismic Data Interpolation And Reconstruction
by: Wang, Shuang, et al.
Published: (2025)
by: Wang, Shuang, et al.
Published: (2025)
RASST: Fast Cross-modal Retrieval-Augmented Simultaneous Speech Translation
by: Luo, Jiaxuan, et al.
Published: (2026)
by: Luo, Jiaxuan, et al.
Published: (2026)
InvCoSS: Inversion-driven Continual Self-supervised Learning in Medical Multi-modal Image Pre-training
by: Luo, Zihao, et al.
Published: (2025)
by: Luo, Zihao, et al.
Published: (2025)
Towards Efficient and Intelligent Laser Weeding: Method and Dataset for Weed Stem Detection
by: Liu, Dingning, et al.
Published: (2025)
by: Liu, Dingning, et al.
Published: (2025)
T3D: Advancing 3D Medical Vision-Language Pre-training by Learning Multi-View Visual Consistency
by: Liu, Che, et al.
Published: (2023)
by: Liu, Che, et al.
Published: (2023)
Iterative Pretraining Framework for Interatomic Potentials
by: Cui, Taoyong, et al.
Published: (2025)
by: Cui, Taoyong, et al.
Published: (2025)
TransGPT: Multi-modal Generative Pre-trained Transformer for Transportation
by: Wang, Peng, et al.
Published: (2024)
by: Wang, Peng, et al.
Published: (2024)
SlimQwen: Exploring the Pruning and Distillation in Large MoE Model Pre-training
by: Tang, Shengkun, et al.
Published: (2026)
by: Tang, Shengkun, et al.
Published: (2026)
Semantics-enhanced Cross-modal Masked Image Modeling for Vision-Language Pre-training
by: Liu, Haowei, et al.
Published: (2024)
by: Liu, Haowei, et al.
Published: (2024)
SeisFusion: Constrained Diffusion Model with Input Guidance for 3D Seismic Data Interpolation and Reconstruction
by: Wang, Shuang, et al.
Published: (2024)
by: Wang, Shuang, et al.
Published: (2024)
MoDA: Multi-modal Diffusion Architecture for Talking Head Generation
by: Li, Xinyang, et al.
Published: (2025)
by: Li, Xinyang, et al.
Published: (2025)
Multimodal LLM Enhanced Cross-lingual Cross-modal Retrieval
by: Wang, Yabing, et al.
Published: (2024)
by: Wang, Yabing, et al.
Published: (2024)
Boosting Instance Awareness via Cross-View Correlation with 4D Radar and Camera for 3D Object Detection
by: Bai, Xiaokai, et al.
Published: (2026)
by: Bai, Xiaokai, et al.
Published: (2026)
FiTv2: Scalable and Improved Flexible Vision Transformer for Diffusion Model
by: Wang, ZiDong, et al.
Published: (2024)
by: Wang, ZiDong, et al.
Published: (2024)
Efficient Transferability Assessment for Selection of Pre-trained Detectors
by: Wang, Zhao, et al.
Published: (2024)
by: Wang, Zhao, et al.
Published: (2024)
Multi-modal Cross-domain Self-supervised Pre-training for fMRI and EEG Fusion
by: Wei, Xinxu, et al.
Published: (2024)
by: Wei, Xinxu, et al.
Published: (2024)
High‐Performance Photodetection in MoSe 2 /Bi 2 Se 3 /MoTe 2 Dual Heterostructure Enabled by Topological Interlayer Design
by: Huijin Xiao, et al.
Published: (2025)
by: Huijin Xiao, et al.
Published: (2025)
Learning Geometry-Guided Depth via Projective Modeling for Monocular 3D Object Detection
by: Zhang, Yinmin, et al.
Published: (2021)
by: Zhang, Yinmin, et al.
Published: (2021)
Superpixel Semantics Representation and Pre-training for Vision-Language Task
by: Zhang, Siyu, et al.
Published: (2023)
by: Zhang, Siyu, et al.
Published: (2023)
Semantic Aware Linear Transfer by Recycling Pre-trained Language Models for Cross-lingual Transfer
by: Lee, Seungyoon, et al.
Published: (2025)
by: Lee, Seungyoon, et al.
Published: (2025)
UniPAD: A Universal Pre-training Paradigm for Autonomous Driving
by: Yang, Honghui, et al.
Published: (2023)
by: Yang, Honghui, et al.
Published: (2023)
Exploring Transferability of Multimodal Adversarial Samples for Vision-Language Pre-training Models with Contrastive Learning
by: Wang, Youze, et al.
Published: (2023)
by: Wang, Youze, et al.
Published: (2023)
Cross-Domain Pre-training with Language Models for Transferable Time Series Representations
by: Cheng, Mingyue, et al.
Published: (2024)
by: Cheng, Mingyue, et al.
Published: (2024)
Large-scale Multiple Testing of Cross-covariance Functions with Applications to Functional Network Models
by: Fang, Qin, et al.
Published: (2024)
by: Fang, Qin, et al.
Published: (2024)
Similar Items
-
Fast Information Streaming Handler (FisH): A Unified Seismic Neural Network for Single Station Real-Time Earthquake Early Warning
by: Zhang, Tianning, et al.
Published: (2024) -
Geometry-enhanced Pre-training on Interatomic Potentials
by: Cui, Taoyong, et al.
Published: (2023) -
DispFormer: A Pretrained Transformer Incorporating Physical Constraints for Dispersion Curve Inversion
by: Liu, Feng, et al.
Published: (2025) -
Progressive Cross-Stream Cooperation in Spatial and Temporal Domain for Action Localization
by: Su, Rui, et al.
Published: (2019) -
OpenSWI: A Massive-Scale Benchmark Dataset for Surface Wave Dispersion Curve Inversion
by: Liu, Feng, et al.
Published: (2025)