HiT-JEPA: A Hierarchical Self-supervised Trajectory Embedding Framework for Similarity Computation
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Lihuan, Xue, Hao, Ao, Shuang, Song, Yang, Salim, Flora |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
T-JEPA: A Joint-Embedding Predictive Architecture for Trajectory Similarity Computation
by: Li, Lihuan, et al.
Published: (2024)
by: Li, Lihuan, et al.
Published: (2024)
HiT: Building Mapping with Hierarchical Transformers
by: Zhang, Mingming, et al.
Published: (2023)
by: Zhang, Mingming, et al.
Published: (2023)
HiT-SR: Hierarchical Transformer for Efficient Image Super-Resolution
by: Zhang, Xiang, et al.
Published: (2024)
by: Zhang, Xiang, et al.
Published: (2024)
RF-HiT: Rectified Flow Hierarchical Transformer for General Medical Image Segmentation
by: Djouama, Ahmed Marouane, et al.
Published: (2026)
by: Djouama, Ahmed Marouane, et al.
Published: (2026)
AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild
by: Chen, Baiyu, et al.
Published: (2026)
by: Chen, Baiyu, et al.
Published: (2026)
Online Monitoring Framework for Automotive Time Series Data using JEPA Embeddings
by: Fertig, Alexander, et al.
Published: (2026)
by: Fertig, Alexander, et al.
Published: (2026)
ViLCo-Bench: VIdeo Language COntinual learning Benchmark
by: Tang, Tianqi, et al.
Published: (2024)
by: Tang, Tianqi, et al.
Published: (2024)
F2T2-HiT: A U-Shaped FFT Transformer and Hierarchical Transformer for Reflection Removal
by: Cai, Jie, et al.
Published: (2025)
by: Cai, Jie, et al.
Published: (2025)
Hi-LSplat: Hierarchical 3D Language Gaussian Splatting
by: Zhan, Chenlu, et al.
Published: (2025)
by: Zhan, Chenlu, et al.
Published: (2025)
US-JEPA: A Joint Embedding Predictive Architecture for Medical Ultrasound
by: Radhachandran, Ashwath, et al.
Published: (2026)
by: Radhachandran, Ashwath, et al.
Published: (2026)
Social-JEPA: Emergent Geometric Isomorphism
by: Zhang, Haoran, et al.
Published: (2026)
by: Zhang, Haoran, et al.
Published: (2026)
HiTVideo: Hierarchical Tokenizers for Enhancing Text-to-Video Generation with Autoregressive Large Language Models
by: Zhou, Ziqin, et al.
Published: (2025)
by: Zhou, Ziqin, et al.
Published: (2025)
Resolution-Agnostic Transformer-based Climate Downscaling
by: Curran, Declan, et al.
Published: (2024)
by: Curran, Declan, et al.
Published: (2024)
HiLa: Hierarchical Vision-Language Collaboration for Cancer Survival Prediction
by: Cui, Jiaqi, et al.
Published: (2025)
by: Cui, Jiaqi, et al.
Published: (2025)
UR-JEPA: Uniform Rectifiability as a Regularizer for Joint-Embedding Predictive Architectures
by: Le, Triet M.
Published: (2026)
by: Le, Triet M.
Published: (2026)
A Review on Discriminative Self-supervised Learning Methods in Computer Vision
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
AdaEmbed: Semi-supervised Domain Adaptation in the Embedding Space
by: Mottaghi, Ali, et al.
Published: (2024)
by: Mottaghi, Ali, et al.
Published: (2024)
WaveHiT-SR: Hierarchical Wavelet Network for Efficient Image Super-Resolution
by: Ali, Fayaz, et al.
Published: (2025)
by: Ali, Fayaz, et al.
Published: (2025)
HiMu: Hierarchical Multimodal Frame Selection for Long Video Question Answering
by: Ben-Ami, Dan, et al.
Published: (2026)
by: Ben-Ami, Dan, et al.
Published: (2026)
HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers
by: Zhang, Jianke, et al.
Published: (2024)
by: Zhang, Jianke, et al.
Published: (2024)
HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System
by: Yang, Tianshuo, et al.
Published: (2026)
by: Yang, Tianshuo, et al.
Published: (2026)
HiSciBench: A Hierarchical Multi-disciplinary Benchmark for Scientific Intelligence from Reading to Discovery
by: Zhang, Yaping, et al.
Published: (2025)
by: Zhang, Yaping, et al.
Published: (2025)
LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics
by: Balestriero, Randall, et al.
Published: (2025)
by: Balestriero, Randall, et al.
Published: (2025)
HiDe: Rethinking The Zoom-IN method in High Resolution MLLMs via Hierarchical Decoupling
by: Liu, Xianjie, et al.
Published: (2025)
by: Liu, Xianjie, et al.
Published: (2025)
HiPP-Prune: Hierarchical Preference-Conditioned Structured Pruning for Vision-Language Models
by: Bai, Lincen, et al.
Published: (2026)
by: Bai, Lincen, et al.
Published: (2026)
COMODO: Cross-Modal Video-to-IMU Distillation for Efficient Egocentric Human Activity Recognition
by: Chen, Baiyu, et al.
Published: (2025)
by: Chen, Baiyu, et al.
Published: (2025)
CoilDrop-MRI: Self-supervised physics-guided MRI reconstruction with coil dropout
by: Song, Tongxi, et al.
Published: (2026)
by: Song, Tongxi, et al.
Published: (2026)
A Self-supervised Pressure Map human keypoint Detection Approch: Optimizing Generalization and Computational Efficiency Across Datasets
by: Yu, Chengzhang, et al.
Published: (2024)
by: Yu, Chengzhang, et al.
Published: (2024)
Unveiling the Power of Self-supervision for Multi-view Multi-human Association and Tracking
by: Feng, Wei, et al.
Published: (2024)
by: Feng, Wei, et al.
Published: (2024)
HiPath: Hierarchical Vision-Language Alignment for Structured Pathology Report Prediction
by: Yuan, Ruicheng, et al.
Published: (2026)
by: Yuan, Ruicheng, et al.
Published: (2026)
HiST-VLA: A Hierarchical Spatio-Temporal Vision-Language-Action Model for End-to-End Autonomous Driving
by: Wang, Yiru, et al.
Published: (2026)
by: Wang, Yiru, et al.
Published: (2026)
Self-supervised Benchmark Lottery on ImageNet: Do Marginal Improvements Translate to Improvements on Similar Datasets?
by: Ozbulak, Utku, et al.
Published: (2025)
by: Ozbulak, Utku, et al.
Published: (2025)
SpatialDreamer: Self-supervised Stereo Video Synthesis from Monocular Input
by: Lv, Zhen, et al.
Published: (2024)
by: Lv, Zhen, et al.
Published: (2024)
HiCMamba: Enhancing Hi-C Resolution and Identifying 3D Genome Structures with State Space Modeling
by: Yang, Minghao, et al.
Published: (2025)
by: Yang, Minghao, et al.
Published: (2025)
Vision-based Multi-future Trajectory Prediction: A Survey
by: Huang, Renhao, et al.
Published: (2023)
by: Huang, Renhao, et al.
Published: (2023)
CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
CLAY: Conditional Visual Similarity Modulation in Vision-Language Embedding Space
by: Lim, Sohwi, et al.
Published: (2026)
by: Lim, Sohwi, et al.
Published: (2026)
Masked Modeling for Self-supervised Representation Learning on Vision and Beyond
by: Li, Siyuan, et al.
Published: (2023)
by: Li, Siyuan, et al.
Published: (2023)
Hi-OSCAR: Hierarchical Open-set Classifier for Human Activity Recognition
by: McCarthy, Conor, et al.
Published: (2025)
by: McCarthy, Conor, et al.
Published: (2025)
SemHiTok: A Unified Image Tokenizer via Semantic-Guided Hierarchical Codebook for Multimodal Understanding and Generation
by: Chen, Zisheng, et al.
Published: (2025)
by: Chen, Zisheng, et al.
Published: (2025)
Similar Items
-
T-JEPA: A Joint-Embedding Predictive Architecture for Trajectory Similarity Computation
by: Li, Lihuan, et al.
Published: (2024) -
HiT: Building Mapping with Hierarchical Transformers
by: Zhang, Mingming, et al.
Published: (2023) -
HiT-SR: Hierarchical Transformer for Efficient Image Super-Resolution
by: Zhang, Xiang, et al.
Published: (2024) -
RF-HiT: Rectified Flow Hierarchical Transformer for General Medical Image Segmentation
by: Djouama, Ahmed Marouane, et al.
Published: (2026) -
AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild
by: Chen, Baiyu, et al.
Published: (2026)