Saved in:
| Main Authors: | Xu, Jiyuan, Zhang, Wenyu, Jing, Xin, Chen, Shuai, Zhang, Shuai, Nie, Jiahao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.20318 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Inpainting-Style Conditional Diffusion for Multivariable Time Series Forecasting
by: Kiani, Kourosh, et al.
Published: (2026)
by: Kiani, Kourosh, et al.
Published: (2026)
VIFO: Visual Feature Empowered Multivariate Time Series Forecasting with Cross-Modal Fusion
by: Wang, Yanlong, et al.
Published: (2025)
by: Wang, Yanlong, et al.
Published: (2025)
Image-Free Timestep Distillation via Continuous-Time Consistency with Trajectory-Sampled Pairs
by: Tang, Bao, et al.
Published: (2025)
by: Tang, Bao, et al.
Published: (2025)
DeltaMIL: Gated Memory Integration for Efficient and Discriminative Whole Slide Image Analysis
by: Zhu, Yueting, et al.
Published: (2025)
by: Zhu, Yueting, et al.
Published: (2025)
Cross-view geo-localization, Image retrieval, Multiscale geometric modeling, Frequency domain enhancement
by: Zhang, Hongying, et al.
Published: (2026)
by: Zhang, Hongying, et al.
Published: (2026)
TOGS: Gaussian Splatting with Temporal Opacity Offset for Real-Time 4D DSA Rendering
by: Zhang, Shuai, et al.
Published: (2024)
by: Zhang, Shuai, et al.
Published: (2024)
Relational Retrieval: Leveraging Known-Novel Interactions for Generalized Category Discovery
by: Xu, Yulin, et al.
Published: (2026)
by: Xu, Yulin, et al.
Published: (2026)
RiT: Vanilla Diffusion Transformers Suffice in Representation Space
by: Zhang, Le, et al.
Published: (2026)
by: Zhang, Le, et al.
Published: (2026)
MTS-DMAE: Dual-Masked Autoencoder for Unsupervised Multivariate Time Series Representation Learning
by: Xu, Yi, et al.
Published: (2025)
by: Xu, Yi, et al.
Published: (2025)
PdfTable: A Unified Toolkit for Deep Learning-Based Table Extraction
by: Sheng, Lei, et al.
Published: (2024)
by: Sheng, Lei, et al.
Published: (2024)
TimePre: Bridging Accuracy, Efficiency, and Stability in Probabilistic Time-Series Forecasting
by: Jiang, Lingyu, et al.
Published: (2025)
by: Jiang, Lingyu, et al.
Published: (2025)
Turbo-VAED: Fast and Stable Transfer of Video-VAEs to Mobile Devices
by: Zou, Ya, et al.
Published: (2025)
by: Zou, Ya, et al.
Published: (2025)
Dynamic 2D Gaussians: Geometrically Accurate Radiance Fields for Dynamic Objects
by: Zhang, Shuai, et al.
Published: (2024)
by: Zhang, Shuai, et al.
Published: (2024)
DiSa: Saliency-Aware Foreground-Background Disentangled Framework for Open-Vocabulary Semantic Segmentation
by: Yao, Zhen, et al.
Published: (2026)
by: Yao, Zhen, et al.
Published: (2026)
IMTS is Worth Time $\times$ Channel Patches: Visual Masked Autoencoders for Irregular Multivariate Time Series Prediction
by: Hu, Zhangyi, et al.
Published: (2025)
by: Hu, Zhangyi, et al.
Published: (2025)
Advancing Real-World Parking Slot Detection with Large-Scale Dataset and Semi-Supervised Baseline
by: Zhang, Zhihao, et al.
Published: (2025)
by: Zhang, Zhihao, et al.
Published: (2025)
LAB-Det: Language as a Domain-Invariant Bridge for Training-Free One-Shot Domain Generalization in Object Detection
by: Zhang, Xu, et al.
Published: (2026)
by: Zhang, Xu, et al.
Published: (2026)
TiMo: Spatiotemporal Foundation Model for Satellite Image Time Series
by: Qin, Xiaolei, et al.
Published: (2025)
by: Qin, Xiaolei, et al.
Published: (2025)
SGS-Intrinsic: Semantic-Invariant Gaussian Splatting for Sparse-View Indoor Inverse Rendering
by: Niu, Jiahao, et al.
Published: (2026)
by: Niu, Jiahao, et al.
Published: (2026)
Incremental Human-Object Interaction Detection with Invariant Relation Representation Learning
by: Wei, Yana, et al.
Published: (2025)
by: Wei, Yana, et al.
Published: (2025)
MMRel: Benchmarking Relation Understanding in Multi-Modal Large Language Models
by: Nie, Jiahao, et al.
Published: (2024)
by: Nie, Jiahao, et al.
Published: (2024)
Long-Tailed Visual Recognition via Permutation-Invariant Head-to-Tail Feature Fusion
by: Li, Mengke, et al.
Published: (2025)
by: Li, Mengke, et al.
Published: (2025)
SITSMamba for Crop Classification based on Satellite Image Time Series
by: Qin, Xiaolei, et al.
Published: (2024)
by: Qin, Xiaolei, et al.
Published: (2024)
ViTGaze: Gaze Following with Interaction Features in Vision Transformers
by: Song, Yuehao, et al.
Published: (2024)
by: Song, Yuehao, et al.
Published: (2024)
RiO-DETR: DETR for Real-time Oriented Object Detection
by: Hu, Zhangchi, et al.
Published: (2026)
by: Hu, Zhangchi, et al.
Published: (2026)
Rethinking Rotation-Invariant Recognition of Fine-grained Shapes from the Perspective of Contour Points
by: Xu, Yanjie, et al.
Published: (2025)
by: Xu, Yanjie, et al.
Published: (2025)
Temporal Restoration and Spatial Rewiring for Source-Free Multivariate Time Series Domain Adaptation
by: Gong, Peiliang, et al.
Published: (2025)
by: Gong, Peiliang, et al.
Published: (2025)
DreamRelation: Relation-Centric Video Customization
by: Wei, Yujie, et al.
Published: (2025)
by: Wei, Yujie, et al.
Published: (2025)
Investigating Permutation-Invariant Discrete Representation Learning for Spatially Aligned Images
by: Stirling, Jamie S. J., et al.
Published: (2026)
by: Stirling, Jamie S. J., et al.
Published: (2026)
ViTime: Foundation Model for Time Series Forecasting Powered by Vision Intelligence
by: Yang, Luoxiao, et al.
Published: (2024)
by: Yang, Luoxiao, et al.
Published: (2024)
Improving Position Encoding of Transformers for Multivariate Time Series Classification
by: Foumani, Navid Mohammadi, et al.
Published: (2023)
by: Foumani, Navid Mohammadi, et al.
Published: (2023)
Visual Instruction Tuning with Chain of Region-of-Interest
by: Chen, Yixin, et al.
Published: (2025)
by: Chen, Yixin, et al.
Published: (2025)
Forgedit: Text Guided Image Editing via Learning and Forgetting
by: Zhang, Shiwen, et al.
Published: (2023)
by: Zhang, Shiwen, et al.
Published: (2023)
FATE: Focal-modulated Attention Encoder for Multivariate Time-series Forecasting
by: Ashraf, Tajamul, et al.
Published: (2024)
by: Ashraf, Tajamul, et al.
Published: (2024)
Map-Relative Pose Regression for Visual Re-Localization
by: Chen, Shuai, et al.
Published: (2024)
by: Chen, Shuai, et al.
Published: (2024)
Invariant Shape Representation Learning For Image Classification
by: Hossain, Tonmoy, et al.
Published: (2024)
by: Hossain, Tonmoy, et al.
Published: (2024)
A New People-Object Interaction Dataset and NVS Benchmarks
by: Guo, Shuai, et al.
Published: (2024)
by: Guo, Shuai, et al.
Published: (2024)
Fast, Robust, Permutation-and-Sign Invariant SO(3) Pattern Alignment
by: Sarker, Anik, et al.
Published: (2025)
by: Sarker, Anik, et al.
Published: (2025)
LAST: LeArning to Think in Space and Time for Generalist Vision-Language Models
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
XS-VID: An Extremely Small Video Object Detection Dataset
by: Guo, Jiahao, et al.
Published: (2024)
by: Guo, Jiahao, et al.
Published: (2024)
Similar Items
-
Inpainting-Style Conditional Diffusion for Multivariable Time Series Forecasting
by: Kiani, Kourosh, et al.
Published: (2026) -
VIFO: Visual Feature Empowered Multivariate Time Series Forecasting with Cross-Modal Fusion
by: Wang, Yanlong, et al.
Published: (2025) -
Image-Free Timestep Distillation via Continuous-Time Consistency with Trajectory-Sampled Pairs
by: Tang, Bao, et al.
Published: (2025) -
DeltaMIL: Gated Memory Integration for Efficient and Discriminative Whole Slide Image Analysis
by: Zhu, Yueting, et al.
Published: (2025) -
Cross-view geo-localization, Image retrieval, Multiscale geometric modeling, Frequency domain enhancement
by: Zhang, Hongying, et al.
Published: (2026)