Gespeichert in:
| Hauptverfasser: | Shao, Run, Yang, Cheng, Li, Qiujun, Zhu, Qing, Zhang, Yongjun, Li, YanSheng, Liu, Yu, Tang, Yong, Liu, Dapeng, Yang, Shizhong, Li, Haifeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2401.00546 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AllSpark: Reborn Labeled Features from Unlabeled in Transformer for Semi-Supervised Semantic Segmentation
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
LocationAgent: A Hierarchical Agent for Image Geolocation via Decoupling Strategy and Evidence from Parametric Knowledge
von: Li, Qiujun, et al.
Veröffentlicht: (2026)
von: Li, Qiujun, et al.
Veröffentlicht: (2026)
RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding
von: Xu, Linrui, et al.
Veröffentlicht: (2024)
von: Xu, Linrui, et al.
Veröffentlicht: (2024)
Fabrication and Analysis of the Wear Properties of High‐Vanadium High‐Speed Steel through Spark Plasma Sintering
von: Shuaiwu Tong, et al.
Veröffentlicht: (2024)
von: Shuaiwu Tong, et al.
Veröffentlicht: (2024)
Adaptive Channel Estimation and Hybrid Beamforming for RIS aided Vehicular Communication
von: Li, Tianyou, et al.
Veröffentlicht: (2026)
von: Li, Tianyou, et al.
Veröffentlicht: (2026)
STA-GANN: A Valid and Generalizable Spatio-Temporal Kriging Approach
von: Li, Yujie, et al.
Veröffentlicht: (2025)
von: Li, Yujie, et al.
Veröffentlicht: (2025)
PixelRefer: A Unified Framework for Spatio-Temporal Object Referring with Arbitrary Granularity
von: Yuan, Yuqian, et al.
Veröffentlicht: (2025)
von: Yuan, Yuqian, et al.
Veröffentlicht: (2025)
The Wittgensteinian Representation Hypothesis: Is Language the Attractor of Multimodal Convergence?
von: Zhang, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Zhang, Zhaoyang, et al.
Veröffentlicht: (2026)
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
von: Guo, Xiangyu, et al.
Veröffentlicht: (2025)
von: Guo, Xiangyu, et al.
Veröffentlicht: (2025)
SpaceVLLM: Endowing Multimodal Large Language Model with Spatio-Temporal Video Grounding Capability
von: Wang, Jiankang, et al.
Veröffentlicht: (2025)
von: Wang, Jiankang, et al.
Veröffentlicht: (2025)
Spatio-Temporal Few-Shot Learning via Diffusive Neural Network Generation
von: Yuan, Yuan, et al.
Veröffentlicht: (2024)
von: Yuan, Yuan, et al.
Veröffentlicht: (2024)
Value-Decomposed Reinforcement Learning Framework for Taxiway Routing with Hierarchical Conflict-Aware Observations
von: Zhou, Shizhong, et al.
Veröffentlicht: (2026)
von: Zhou, Shizhong, et al.
Veröffentlicht: (2026)
A Refer-and-Ground Multimodal Large Language Model for Biomedicine
von: Huang, Xiaoshuang, et al.
Veröffentlicht: (2024)
von: Huang, Xiaoshuang, et al.
Veröffentlicht: (2024)
LiDAR Prompted Spatio-Temporal Multi-View Stereo for Autonomous Driving
von: Sun, Qihao, et al.
Veröffentlicht: (2026)
von: Sun, Qihao, et al.
Veröffentlicht: (2026)
Unleashing the Potential of Multimodal LLMs for Zero-Shot Spatio-Temporal Video Grounding
von: Yang, Zaiquan, et al.
Veröffentlicht: (2025)
von: Yang, Zaiquan, et al.
Veröffentlicht: (2025)
Jointly Modeling Spatio-Temporal Features of Tactile Signals for Action Classification
von: Lin, Jimmy, et al.
Veröffentlicht: (2024)
von: Lin, Jimmy, et al.
Veröffentlicht: (2024)
Crip Spacetime: Access, Failure, and Accountability in Academic Life. By MargaretPrice, Durham: Duke University Press, 2024. 240 pp. $26.95 (paper). ISBN: 978‐1‐47‐803037‐9; $102.95 (hardcover). ISBN: 978‐1‐47‐802613‐6
von: Leyan Zheng, et al.
Veröffentlicht: (2025)
von: Leyan Zheng, et al.
Veröffentlicht: (2025)
STaR-Attack: A Spatio-Temporal and Narrative Reasoning Attack Framework for Unified Multimodal Understanding and Generation Models
von: Guo, Shaoxiong, et al.
Veröffentlicht: (2025)
von: Guo, Shaoxiong, et al.
Veröffentlicht: (2025)
Air Quality Prediction with A Meteorology-Guided Modality-Decoupled Spatio-Temporal Network
von: Yin, Hang, et al.
Veröffentlicht: (2025)
von: Yin, Hang, et al.
Veröffentlicht: (2025)
Transformer RGBT Tracking with Spatio-Temporal Multimodal Tokens
von: Sun, Dengdi, et al.
Veröffentlicht: (2024)
von: Sun, Dengdi, et al.
Veröffentlicht: (2024)
Robust Multimodal Semantic Segmentation with Balanced Modality Contributions
von: Tan, Jiaqi, et al.
Veröffentlicht: (2025)
von: Tan, Jiaqi, et al.
Veröffentlicht: (2025)
Multimodal Contrastive Learning via Uni-Modal Coding and Cross-Modal Prediction for Multimodal Sentiment Analysis
von: Lin, Ronghao, et al.
Veröffentlicht: (2022)
von: Lin, Ronghao, et al.
Veröffentlicht: (2022)
UniFlow: A Foundation Model for Unified Urban Spatio-Temporal Flow Prediction
von: Yuan, Yuan, et al.
Veröffentlicht: (2024)
von: Yuan, Yuan, et al.
Veröffentlicht: (2024)
GLIDE: Graph-guided Leap Inference for Diffusion Estimation of Spatio-Temporal Point Processes
von: Zhou, Guanyu, et al.
Veröffentlicht: (2026)
von: Zhou, Guanyu, et al.
Veröffentlicht: (2026)
Mining Multi-Modality Spatio-Temporal Cues for Video Important Person Identification
von: Wang, Xiao, et al.
Veröffentlicht: (2026)
von: Wang, Xiao, et al.
Veröffentlicht: (2026)
UrbanGPT: Spatio-Temporal Large Language Models
von: Li, Zhonghang, et al.
Veröffentlicht: (2024)
von: Li, Zhonghang, et al.
Veröffentlicht: (2024)
AsyReC: A Multimodal Graph-based Framework for Spatio-Temporal Asymmetric Dyadic Relationship Classification
von: Tang, Wang, et al.
Veröffentlicht: (2025)
von: Tang, Wang, et al.
Veröffentlicht: (2025)
TreeMIL: A Multi-instance Learning Framework for Time Series Anomaly Detection with Inexact Supervision
von: Liu, Chen, et al.
Veröffentlicht: (2024)
von: Liu, Chen, et al.
Veröffentlicht: (2024)
Learnability in Online Kernel Selection with Memory Constraint via Data-dependent Regret Analysis
von: Li, Junfan, et al.
Veröffentlicht: (2024)
von: Li, Junfan, et al.
Veröffentlicht: (2024)
Improved Kernel Alignment Regret Bound for Online Kernel Learning
von: Li, Junfan, et al.
Veröffentlicht: (2022)
von: Li, Junfan, et al.
Veröffentlicht: (2022)
Graph Learning-Driven Multi-Vessel Association: Fusing Multimodal Data for Maritime Intelligence
von: Lu, Yuxu, et al.
Veröffentlicht: (2025)
von: Lu, Yuxu, et al.
Veröffentlicht: (2025)
Interacted Object Grounding in Spatio-Temporal Human-Object Interactions
von: Liu, Xiaoyang, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoyang, et al.
Veröffentlicht: (2024)
Decoupled Diffusion Sparks Adaptive Scene Generation
von: Zhou, Yunsong, et al.
Veröffentlicht: (2025)
von: Zhou, Yunsong, et al.
Veröffentlicht: (2025)
DMTrack: Spatio-Temporal Multimodal Tracking via Dual-Adapter
von: Li, Weihong, et al.
Veröffentlicht: (2025)
von: Li, Weihong, et al.
Veröffentlicht: (2025)
STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training
von: Liu, Minglu, et al.
Veröffentlicht: (2026)
von: Liu, Minglu, et al.
Veröffentlicht: (2026)
Enhancing SNN-based Spatio-Temporal Learning: A Benchmark Dataset and Cross-Modality Attention Model
von: Zhou, Shibo, et al.
Veröffentlicht: (2024)
von: Zhou, Shibo, et al.
Veröffentlicht: (2024)
Long Context is Not Long at All: A Prospector of Long-Dependency Data for Large Language Models
von: Chen, Longze, et al.
Veröffentlicht: (2024)
von: Chen, Longze, et al.
Veröffentlicht: (2024)
Modality-Guided Dynamic Graph Fusion and Temporal Diffusion for Self-Supervised RGB-T Tracking
von: Li, Shenglan, et al.
Veröffentlicht: (2025)
von: Li, Shenglan, et al.
Veröffentlicht: (2025)
Multimodal Classification via Modal-Aware Interactive Enhancement
von: Jiang, Qing-Yuan, et al.
Veröffentlicht: (2024)
von: Jiang, Qing-Yuan, et al.
Veröffentlicht: (2024)
Seismic analysis based on a new interval method with incomplete information
von: Liang, Shizhong, et al.
Veröffentlicht: (2025)
von: Liang, Shizhong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AllSpark: Reborn Labeled Features from Unlabeled in Transformer for Semi-Supervised Semantic Segmentation
von: Wang, Haonan, et al.
Veröffentlicht: (2024) -
LocationAgent: A Hierarchical Agent for Image Geolocation via Decoupling Strategy and Evidence from Parametric Knowledge
von: Li, Qiujun, et al.
Veröffentlicht: (2026) -
RS-GPT4V: A Unified Multimodal Instruction-Following Dataset for Remote Sensing Image Understanding
von: Xu, Linrui, et al.
Veröffentlicht: (2024) -
Fabrication and Analysis of the Wear Properties of High‐Vanadium High‐Speed Steel through Spark Plasma Sintering
von: Shuaiwu Tong, et al.
Veröffentlicht: (2024) -
Adaptive Channel Estimation and Hybrid Beamforming for RIS aided Vehicular Communication
von: Li, Tianyou, et al.
Veröffentlicht: (2026)