MERLIN: Building Low-SNR Robust Multimodal LLMs for Electromagnetic Signals
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shen, Junyu, She, Zhendong, Zhang, Chenghanyu, Sun, Yuchuang, Luo, Luqing, Tan, Dingwei, Guo, Zonghao, Guo, Bo, Han, Zehua, Xie, Wupeng, Mu, Yaxin, Zhang, Peng, Li, Peipei, Wang, Fengxiang, Sun, Yangang, Sun, Maosong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PReD: An LLM-based Foundation Multimodal Model for Electromagnetic Perception, Recognition, and Decision
von: Han, Zehua, et al.
Veröffentlicht: (2026)
von: Han, Zehua, et al.
Veröffentlicht: (2026)
EMind: A Foundation Model for Multi-task Electromagnetic Signals Understanding
von: Luo, Luqing, et al.
Veröffentlicht: (2025)
von: Luo, Luqing, et al.
Veröffentlicht: (2025)
XLRS-Bench: Could Your Multimodal LLMs Understand Extremely Large Ultra-High-Resolution Remote Sensing Imagery?
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
RIS-MAE: A Self-Supervised Modulation Classification Method Based on Raw IQ Signals and Masked Autoencoder
von: Liu, Yunfei, et al.
Veröffentlicht: (2025)
von: Liu, Yunfei, et al.
Veröffentlicht: (2025)
LLaVA-UHD v3: Progressive Visual Compression for Efficient Native-Resolution Encoding in MLLMs
von: Sun, Shichu, et al.
Veröffentlicht: (2025)
von: Sun, Shichu, et al.
Veröffentlicht: (2025)
SpineBench: Benchmarking Multimodal LLMs for Spinal Pathology Analysis
von: Zhang, Chenghanyu, et al.
Veröffentlicht: (2025)
von: Zhang, Chenghanyu, et al.
Veröffentlicht: (2025)
Classification Method for Fatigue Driving Signals Based on Multiple Classifier Analysis
von: Zhendong Mu
Veröffentlicht: (2025)
von: Zhendong Mu
Veröffentlicht: (2025)
Cost Information, Insider Trading, and Product Market Equilibrium
von: Dingwei Gu, et al.
Veröffentlicht: (2026)
von: Dingwei Gu, et al.
Veröffentlicht: (2026)
Scaling Decentralized Learning with FLock
von: Cheng, Zehua, et al.
Veröffentlicht: (2025)
von: Cheng, Zehua, et al.
Veröffentlicht: (2025)
More transparency, more trust: Evidence from anti‐corruption campaign in China
von: Yaxin Zhang, et al.
Veröffentlicht: (2024)
von: Yaxin Zhang, et al.
Veröffentlicht: (2024)
DTRT: Enhancing Human Intent Estimation and Role Allocation for Physical Human-Robot Collaboration
von: Liu, Haotian, et al.
Veröffentlicht: (2025)
von: Liu, Haotian, et al.
Veröffentlicht: (2025)
FAR-Dex: Few-shot Data Augmentation and Adaptive Residual Policy Refinement for Dexterous Manipulation
von: Bai, Yushan, et al.
Veröffentlicht: (2026)
von: Bai, Yushan, et al.
Veröffentlicht: (2026)
MM-UAVBench: How Well Do Multimodal Large Language Models See, Think, and Plan in Low-Altitude UAV Scenarios?
von: Dai, Shiqi, et al.
Veröffentlicht: (2025)
von: Dai, Shiqi, et al.
Veröffentlicht: (2025)
Will Pre-Training Ever End? A First Step Toward Next-Generation Foundation MLLMs via Self-Improving Systematic Cognition
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2025)
Cheers: Decoupling Patch Details from Semantic Representations Enables Unified Multimodal Comprehension and Generation
von: Zhang, Yichen, et al.
Veröffentlicht: (2026)
von: Zhang, Yichen, et al.
Veröffentlicht: (2026)
Magnetic field properties of the SNR HB 9
von: Xiao, Li, et al.
Veröffentlicht: (2025)
von: Xiao, Li, et al.
Veröffentlicht: (2025)
States Hidden in Hidden States: LLMs Emerge Discrete State Representations Implicitly
von: Chen, Junhao, et al.
Veröffentlicht: (2024)
von: Chen, Junhao, et al.
Veröffentlicht: (2024)
Survey on AI-Generated Media Detection: From Non-MLLM to MLLM
von: Zou, Yueying, et al.
Veröffentlicht: (2025)
von: Zou, Yueying, et al.
Veröffentlicht: (2025)
Large-field CO (J=1-0) observations toward SNR G150.3+4.5
von: Feng, Jian-Cheng, et al.
Veröffentlicht: (2024)
von: Feng, Jian-Cheng, et al.
Veröffentlicht: (2024)
Harnessing Massive Satellite Imagery with Efficient Masked Image Modeling
von: Wang, Fengxiang, et al.
Veröffentlicht: (2024)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2024)
Fantastic Semantics and Where to Find Them: Investigating Which Layers of Generative LLMs Reflect Lexical Semantics
von: Liu, Zhu, et al.
Veröffentlicht: (2024)
von: Liu, Zhu, et al.
Veröffentlicht: (2024)
KARL: Knowledge-Aware Reasoning and Reinforcement Learning for Knowledge-Intensive Visual Grounding
von: Ma, Xinyu, et al.
Veröffentlicht: (2025)
von: Ma, Xinyu, et al.
Veröffentlicht: (2025)
The CodeInverter Suite: Control-Flow and Data-Mapping Augmented Binary Decompilation with LLMs
von: Liu, Peipei, et al.
Veröffentlicht: (2025)
von: Liu, Peipei, et al.
Veröffentlicht: (2025)
GeoViS: Geospatially Rewarded Visual Search for Remote Sensing Visual Grounding
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
Fabrication of Uniform Anionic Polymeric Nanoplatelets as Building Blocks for Constructing Conductive Hydrogels with Enhancing Conductive and Mechanical Properties
von: Zehua Li, et al.
Veröffentlicht: (2024)
von: Zehua Li, et al.
Veröffentlicht: (2024)
Hypercontractivity for a family of quantum Ornstein-Uhlenbeck semigroups
von: Sun, Longfa, et al.
Veröffentlicht: (2026)
von: Sun, Longfa, et al.
Veröffentlicht: (2026)
Probabilistic Modeling of Jailbreak on Multimodal LLMs: From Quantification to Application
von: Xu, Wenzhuo, et al.
Veröffentlicht: (2025)
von: Xu, Wenzhuo, et al.
Veröffentlicht: (2025)
Self-Supervised Enhancement of Forward-Looking Sonar Images: Bridging Cross-Modal Degradation Gaps through Feature Space Transformation and Multi-Frame Fusion
von: Zhang, Zhisheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zhisheng, et al.
Veröffentlicht: (2025)
A Proximal Gradient Method With Probabilistic Multi-Gossip Communications for Decentralized Composite Optimization
von: Guo, Luyao, et al.
Veröffentlicht: (2023)
von: Guo, Luyao, et al.
Veröffentlicht: (2023)
Differentially Private Decentralized Optimization with Relay Communication
von: Wang, Luqing, et al.
Veröffentlicht: (2022)
von: Wang, Luqing, et al.
Veröffentlicht: (2022)
Elucidating the SNR-t Bias of Diffusion Probabilistic Models
von: Yu, Meng, et al.
Veröffentlicht: (2026)
von: Yu, Meng, et al.
Veröffentlicht: (2026)
Improve Temporal Awareness of LLMs for Sequential Recommendation
von: Chu, Zhendong, et al.
Veröffentlicht: (2024)
von: Chu, Zhendong, et al.
Veröffentlicht: (2024)
High-dimensional Statistics Applications to Batch Effects in Metabolomics
von: Guo, Zhendong
Veröffentlicht: (2024)
von: Guo, Zhendong
Veröffentlicht: (2024)
Implicit Guidance and Explicit Representation of Semantic Information in Points Cloud: A Survey
von: Tang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Tang, Jingyuan, et al.
Veröffentlicht: (2025)
LLaVA-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images
von: Xu, Ruyi, et al.
Veröffentlicht: (2024)
von: Xu, Ruyi, et al.
Veröffentlicht: (2024)
A Lightweight Hybrid Dual Channel Speech Enhancement System under Low-SNR Conditions
von: Wang, Zheng, et al.
Veröffentlicht: (2025)
von: Wang, Zheng, et al.
Veröffentlicht: (2025)
Measurement and Analysis of Scattering From Building Surfaces at Millimeter-Wave Frequency
von: Guo, Yulu, et al.
Veröffentlicht: (2025)
von: Guo, Yulu, et al.
Veröffentlicht: (2025)
High-Accuracy Schottky Diagnostics for Low-SNR Betatron Tune Measurement in Ramping Synchrotrons
von: Sun, Peihan, et al.
Veröffentlicht: (2024)
von: Sun, Peihan, et al.
Veröffentlicht: (2024)
LLaVA-UHD v2: an MLLM Integrating High-Resolution Semantic Pyramid via Hierarchical Window Transformer
von: Zhang, Yipeng, et al.
Veröffentlicht: (2024)
von: Zhang, Yipeng, et al.
Veröffentlicht: (2024)
RoMA: Scaling up Mamba-based Foundation Models for Remote Sensing
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PReD: An LLM-based Foundation Multimodal Model for Electromagnetic Perception, Recognition, and Decision
von: Han, Zehua, et al.
Veröffentlicht: (2026) -
EMind: A Foundation Model for Multi-task Electromagnetic Signals Understanding
von: Luo, Luqing, et al.
Veröffentlicht: (2025) -
XLRS-Bench: Could Your Multimodal LLMs Understand Extremely Large Ultra-High-Resolution Remote Sensing Imagery?
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025) -
RIS-MAE: A Self-Supervised Modulation Classification Method Based on Raw IQ Signals and Masked Autoencoder
von: Liu, Yunfei, et al.
Veröffentlicht: (2025) -
LLaVA-UHD v3: Progressive Visual Compression for Efficient Native-Resolution Encoding in MLLMs
von: Sun, Shichu, et al.
Veröffentlicht: (2025)