Advancing Multi-Robot Networks via MLLM-Driven Sensing, Communication, and Computation: A Comprehensive Survey
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Hyun Jong, Lee, Howon, Shim, Kyuhong, Kwak, Jeongho, Kim, Hyunsoo, Kim, Donghoon, Ngo, Khoa Anh, Ryu, Sehyun, Choi, Jaehyun, Kim, Youbin, Moon, Chanjun, Ryoo, Michael, Shim, Byonghyo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Preserving Pre-trained Representation Space: On Effectiveness of Prefix-tuning for Large Multi-modal Models
by: Kim, Donghoon, et al.
Published: (2024)
by: Kim, Donghoon, et al.
Published: (2024)
Visually Guided Decoding: Gradient-Free Hard Prompt Inversion with Language Models
by: Kim, Donghoon, et al.
Published: (2025)
by: Kim, Donghoon, et al.
Published: (2025)
Role of Sensing and Computer Vision in 6G Wireless Communications
by: Kim, Seungnyun, et al.
Published: (2024)
by: Kim, Seungnyun, et al.
Published: (2024)
Learning Primitive Relations for Compositional Zero-Shot Learning
by: Lee, Insu, et al.
Published: (2025)
by: Lee, Insu, et al.
Published: (2025)
Adaptive Capacity Allocation for Vision Language Action Fine-tuning
by: Kim, Donghoon, et al.
Published: (2026)
by: Kim, Donghoon, et al.
Published: (2026)
Large Multimodal Models-Empowered Task-Oriented Autonomous Communications: Design Methodology and Implementation Challenges
by: Yang, Hyun Jong, et al.
Published: (2025)
by: Yang, Hyun Jong, et al.
Published: (2025)
Large Multimodal Model-Aided Scheduling for 6G Autonomous Communications
by: Kim, Sunwoo, et al.
Published: (2026)
by: Kim, Sunwoo, et al.
Published: (2026)
Towards Comprehensive Scene Understanding: Integrating First and Third-Person Views for LVLMs
by: Lee, Insu, et al.
Published: (2025)
by: Lee, Insu, et al.
Published: (2025)
Revealing Multi-View Hallucination in Large Vision-Language Models
by: Park, Wooje, et al.
Published: (2026)
by: Park, Wooje, et al.
Published: (2026)
Mask2Flow-TSE: Two-Stage Target Speaker Extraction with Masking and Flow Matching
by: Moon, Junwon, et al.
Published: (2026)
by: Moon, Junwon, et al.
Published: (2026)
Adaptive Resource Allocation Optimization Using Large Language Models in Dynamic Wireless Environments
by: Noh, Hyeonho, et al.
Published: (2025)
by: Noh, Hyeonho, et al.
Published: (2025)
Massive Data Generation for Deep Learning-aided Wireless Systems Using Meta Learning and Generative Adversarial Network
by: Kim, Jinhong, et al.
Published: (2022)
by: Kim, Jinhong, et al.
Published: (2022)
InfiniPot: Infinite Context Processing on Memory-Constrained LLMs
by: Kim, Minsoo, et al.
Published: (2024)
by: Kim, Minsoo, et al.
Published: (2024)
InfiniPot-V: Memory-Constrained KV Cache Compression for Streaming Video Understanding
by: Kim, Minsoo, et al.
Published: (2025)
by: Kim, Minsoo, et al.
Published: (2025)
Deep Learning-aided Parametric Sparse Channel Estimation for Terahertz Massive MIMO Systems
by: Kim, Jinhong, et al.
Published: (2024)
by: Kim, Jinhong, et al.
Published: (2024)
VOMTC: Vision Objects for Millimeter and Terahertz Communications
by: Kim, Sunwoo, et al.
Published: (2024)
by: Kim, Sunwoo, et al.
Published: (2024)
Semantic Token Reweighting for Interpretable and Controllable Text Embeddings in CLIP
by: Kim, Eunji, et al.
Published: (2024)
by: Kim, Eunji, et al.
Published: (2024)
Unlocking Transfer Learning for Open-World Few-Shot Recognition
by: Kim, Byeonggeun, et al.
Published: (2024)
by: Kim, Byeonggeun, et al.
Published: (2024)
Universal Spin Screening Clouds in Local Moment Phases
by: Kim, Minsoo L., et al.
Published: (2024)
by: Kim, Minsoo L., et al.
Published: (2024)
Blockage-Aware Multi-RIS WSR Maximization via Per-RIS Indexed Synchronization Sequences and Closed-Form Riemannian Updates
by: Ryu, Sehyun, et al.
Published: (2025)
by: Ryu, Sehyun, et al.
Published: (2025)
Standards-Compliant DM-RS Allocation via Temporal Channel Prediction for Massive MIMO Systems
by: Ryu, Sehyun, et al.
Published: (2025)
by: Ryu, Sehyun, et al.
Published: (2025)
Transformer-assisted Parametric CSI Feedback for mmWave Massive MIMO Systems
by: Ju, Hyungyu, et al.
Published: (2024)
by: Ju, Hyungyu, et al.
Published: (2024)
P2VA: Converting Persona Descriptions into Voice Attributes for Fair and Controllable Text-to-Speech
by: Lee, Yejin, et al.
Published: (2025)
by: Lee, Yejin, et al.
Published: (2025)
Benchmark Profiling: Mechanistic Diagnosis of LLM Benchmarks
by: Kim, Dongjun, et al.
Published: (2025)
by: Kim, Dongjun, et al.
Published: (2025)
LANGSAE EDITING: Improving Multilingual Information Retrieval via Post-hoc Language Identity Removal
by: Kim, Dongjun, et al.
Published: (2026)
by: Kim, Dongjun, et al.
Published: (2026)
Quantitative Hydrodynamic Limit of the Chern--Simons--Higgs System
by: Kim, Jeongho, et al.
Published: (2026)
by: Kim, Jeongho, et al.
Published: (2026)
Evaluating Hallucinations in Audio-Visual Multimodal LLMs with Spoken Queries under Diverse Acoustic Conditions
by: Park, Hansol, et al.
Published: (2025)
by: Park, Hansol, et al.
Published: (2025)
Blind Channel Estimation for RIS-Assisted Millimeter Wave Communication Systems
by: Jia, Dianhao, et al.
Published: (2025)
by: Jia, Dianhao, et al.
Published: (2025)
Noise Variance Optimization in Differential Privacy: A Game-Theoretic Approach Through Per-Instance Differential Privacy
by: Ryu, Sehyun, et al.
Published: (2024)
by: Ryu, Sehyun, et al.
Published: (2024)
Do MLLMs Capture How Interfaces Guide User Behavior? A Benchmark for Multimodal UI/UX Design Understanding
by: Jeon, Jaehyun, et al.
Published: (2025)
by: Jeon, Jaehyun, et al.
Published: (2025)
Learning Contextual Retrieval for Robust Conversational Search
by: Yang, Seunghan, et al.
Published: (2025)
by: Yang, Seunghan, et al.
Published: (2025)
SUPER-AD: Semantic Uncertainty-aware Planning for End-to-End Robust Autonomous Driving
by: Ryu, Wonjeong, et al.
Published: (2025)
by: Ryu, Wonjeong, et al.
Published: (2025)
Self-Calibration DOA Estimation for Movable Antenna Systems with Antenna Position Errors
by: Ye, Chengzhi, et al.
Published: (2026)
by: Ye, Chengzhi, et al.
Published: (2026)
Determination of Bandwidth of Q-filter in Disturbance Observers to Guarantee Transient and Steady State Performance under Measurement Noise
by: Kim, Gaeun, et al.
Published: (2025)
by: Kim, Gaeun, et al.
Published: (2025)
Asymptotic stability of the high-dimensional Kuramoto model on Stiefel manifolds
by: Kim, Dohyun, et al.
Published: (2024)
by: Kim, Dohyun, et al.
Published: (2024)
Classifier-guided CLIP Distillation for Unsupervised Multi-label Classification
by: Kim, Dongseob, et al.
Published: (2025)
by: Kim, Dongseob, et al.
Published: (2025)
Unlocking Fine-Grained and Within-Utterance Speaking Style Control in Prompt-Based Text-to-Speech Models
by: Kang, Jaehoon, et al.
Published: (2026)
by: Kang, Jaehoon, et al.
Published: (2026)
Whisper-CD: Accurate Long-Form Speech Recognition using Multi-Negative Contrastive Decoding
by: Ahn, Hoseong, et al.
Published: (2026)
by: Ahn, Hoseong, et al.
Published: (2026)
Hypotaxy‐Enabled Selective Growth of MoS 2 for Laterally‐Connected Edge and van der Waals Bottom Contacts (Adv. Funct. Mater. 41/2026)
by: Donghoon Moon, et al.
Published: (2026)
by: Donghoon Moon, et al.
Published: (2026)
Stationary solutions to the spherically symmetric compressible fluid with capillarity effect
by: Kim, Jeongho
Published: (2026)
by: Kim, Jeongho
Published: (2026)
Similar Items
-
Preserving Pre-trained Representation Space: On Effectiveness of Prefix-tuning for Large Multi-modal Models
by: Kim, Donghoon, et al.
Published: (2024) -
Visually Guided Decoding: Gradient-Free Hard Prompt Inversion with Language Models
by: Kim, Donghoon, et al.
Published: (2025) -
Role of Sensing and Computer Vision in 6G Wireless Communications
by: Kim, Seungnyun, et al.
Published: (2024) -
Learning Primitive Relations for Compositional Zero-Shot Learning
by: Lee, Insu, et al.
Published: (2025) -
Adaptive Capacity Allocation for Vision Language Action Fine-tuning
by: Kim, Donghoon, et al.
Published: (2026)