JEPA-T: Joint-Embedding Predictive Architecture with Text Fusion for Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wan, Siheng, Yao, Zhengtao, Li, Zhengdao, Dong, Junhao, Li, Yanshu, Li, Yikai, Li, Linshan, Xu, Haoyan, Li, Yijiang, Dong, Zhikang, Wang, Huacan, Shen, Jifeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
C3-OWD: A Curriculum Cross-modal Contrastive Learning Framework for Open-World Detection
von: Wang, Siheng, et al.
Veröffentlicht: (2025)
von: Wang, Siheng, et al.
Veröffentlicht: (2025)
DeCo-DETR: Decoupled Cognition DETR for efficient Open-Vocabulary Object Detection
von: Wang, Siheng, et al.
Veröffentlicht: (2026)
von: Wang, Siheng, et al.
Veröffentlicht: (2026)
ChromouVQA: Benchmarking Vision-Language Models under Chromatic Camouflaged Images
von: Zhang, Yunfei, et al.
Veröffentlicht: (2025)
von: Zhang, Yunfei, et al.
Veröffentlicht: (2025)
T-JEPA: A Joint-Embedding Predictive Architecture for Trajectory Similarity Computation
von: Li, Lihuan, et al.
Veröffentlicht: (2024)
von: Li, Lihuan, et al.
Veröffentlicht: (2024)
Demo-JEPA: Joint-Embedding Predictive Architecture for One-shot Cross-Embodiment Imitation
von: He, Jingyang, et al.
Veröffentlicht: (2026)
von: He, Jingyang, et al.
Veröffentlicht: (2026)
JEPA-MSAC: A Joint-Embedding Predictive Architecture for Multimodal Sensing-Assisted Communications
von: Zheng, Can, et al.
Veröffentlicht: (2026)
von: Zheng, Can, et al.
Veröffentlicht: (2026)
A-JEPA: Joint-Embedding Predictive Architecture Can Listen
von: Fei, Zhengcong, et al.
Veröffentlicht: (2023)
von: Fei, Zhengcong, et al.
Veröffentlicht: (2023)
VL-JEPA: Joint Embedding Predictive Architecture for Vision-language
von: Chen, Delong, et al.
Veröffentlicht: (2025)
von: Chen, Delong, et al.
Veröffentlicht: (2025)
DSeq-JEPA: Discriminative Sequential Joint-Embedding Predictive Architecture
von: He, Xiangteng, et al.
Veröffentlicht: (2025)
von: He, Xiangteng, et al.
Veröffentlicht: (2025)
3D-JEPA: A Joint Embedding Predictive Architecture for 3D Self-Supervised Representation Learning
von: Hu, Naiwen, et al.
Veröffentlicht: (2024)
von: Hu, Naiwen, et al.
Veröffentlicht: (2024)
JEPA for RL: Investigating Joint-Embedding Predictive Architectures for Reinforcement Learning
von: Kenneweg, Tristan, et al.
Veröffentlicht: (2025)
von: Kenneweg, Tristan, et al.
Veröffentlicht: (2025)
US-JEPA: A Joint Embedding Predictive Architecture for Medical Ultrasound
von: Radhachandran, Ashwath, et al.
Veröffentlicht: (2026)
von: Radhachandran, Ashwath, et al.
Veröffentlicht: (2026)
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
SparseJEPA: Sparse Representation Learning of Joint Embedding Predictive Architectures
von: Hartman, Max, et al.
Veröffentlicht: (2025)
von: Hartman, Max, et al.
Veröffentlicht: (2025)
DMT-JEPA: Discriminative Masked Targets for Joint-Embedding Predictive Architecture
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
JEPA-VLA: Video Predictive Embedding is Needed for VLA Models
von: Miao, Shangchen, et al.
Veröffentlicht: (2026)
von: Miao, Shangchen, et al.
Veröffentlicht: (2026)
AMMKD: Adaptive Multimodal Multi-teacher Distillation for Lightweight Vision-Language Models
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
ReMoMask: Retrieval-Augmented Masked Motion Generation
von: Li, Zhengdao, et al.
Veröffentlicht: (2025)
von: Li, Zhengdao, et al.
Veröffentlicht: (2025)
LLM-JEPA: Large Language Models Meet Joint Embedding Predictive Architectures
von: Huang, Hai, et al.
Veröffentlicht: (2025)
von: Huang, Hai, et al.
Veröffentlicht: (2025)
UR-JEPA: Uniform Rectifiability as a Regularizer for Joint-Embedding Predictive Architectures
von: Le, Triet M.
Veröffentlicht: (2026)
von: Le, Triet M.
Veröffentlicht: (2026)
LLM-Powered Text-Attributed Graph Anomaly Detection via Retrieval-Augmented Reasoning
von: Xu, Haoyan, et al.
Veröffentlicht: (2025)
von: Xu, Haoyan, et al.
Veröffentlicht: (2025)
MTS-JEPA: Multi-Resolution Joint-Embedding Predictive Architecture for Time-Series Anomaly Prediction
von: He, Yanan, et al.
Veröffentlicht: (2026)
von: He, Yanan, et al.
Veröffentlicht: (2026)
BiJEPA: Bi-directional Joint Embedding Predictive Architecture for Symmetric Representation Learning
von: Huang, Yongchao
Veröffentlicht: (2026)
von: Huang, Yongchao
Veröffentlicht: (2026)
ACT-JEPA: Novel Joint-Embedding Predictive Architecture for Efficient Policy Representation Learning
von: Vujinovic, Aleksandar, et al.
Veröffentlicht: (2025)
von: Vujinovic, Aleksandar, et al.
Veröffentlicht: (2025)
JEPA-DNA: Grounding Genomic Foundation Models through Joint-Embedding Predictive Architectures
von: Larey, Ariel, et al.
Veröffentlicht: (2026)
von: Larey, Ariel, et al.
Veröffentlicht: (2026)
Stem-JEPA: A Joint-Embedding Predictive Architecture for Musical Stem Compatibility Estimation
von: Riou, Alain, et al.
Veröffentlicht: (2024)
von: Riou, Alain, et al.
Veröffentlicht: (2024)
Rectified LpJEPA: Joint-Embedding Predictive Architectures with Sparse and Maximum-Entropy Representations
von: Kuang, Yilun, et al.
Veröffentlicht: (2026)
von: Kuang, Yilun, et al.
Veröffentlicht: (2026)
Point-JEPA: A Joint Embedding Predictive Architecture for Self-Supervised Learning on Point Cloud
von: Saito, Ayumu, et al.
Veröffentlicht: (2024)
von: Saito, Ayumu, et al.
Veröffentlicht: (2024)
RadJEPA: Radiology Encoder for Chest X-Rays via Joint Embedding Predictive Architecture
von: Khan, Anas Anwarul Haq, et al.
Veröffentlicht: (2026)
von: Khan, Anas Anwarul Haq, et al.
Veröffentlicht: (2026)
TI-JEPA: An Innovative Energy-based Joint Embedding Strategy for Text-Image Multimodal Systems
von: Vo, Khang H. N., et al.
Veröffentlicht: (2025)
von: Vo, Khang H. N., et al.
Veröffentlicht: (2025)
CNN-JEPA: Self-Supervised Pretraining Convolutional Neural Networks Using Joint Embedding Predictive Architecture
von: Kalapos, András, et al.
Veröffentlicht: (2024)
von: Kalapos, András, et al.
Veröffentlicht: (2024)
Advancing Multimodal In-Context Learning in Large Vision-Language Models with Task-aware Demonstrations
von: Li, Yanshu
Veröffentlicht: (2025)
von: Li, Yanshu
Veröffentlicht: (2025)
A Systematic Study of Model Extraction Attacks on Graph Foundation Models
von: Xu, Haoyan, et al.
Veröffentlicht: (2025)
von: Xu, Haoyan, et al.
Veröffentlicht: (2025)
Var-JEPA: A Variational Formulation of the Joint-Embedding Predictive Architecture -- Bridging Predictive and Generative Self-Supervised Learning
von: Gögl, Moritz, et al.
Veröffentlicht: (2026)
von: Gögl, Moritz, et al.
Veröffentlicht: (2026)
CR-JEPA: Cross-Modal Joint-Embedding Predictive Learning for Remote Sensing Image Retrieval
von: Hossain, Md Aminur, et al.
Veröffentlicht: (2026)
von: Hossain, Md Aminur, et al.
Veröffentlicht: (2026)
CATP: Contextually Adaptive Token Pruning for Efficient and Enhanced Multimodal In-Context Learning
von: Li, Yanshu, et al.
Veröffentlicht: (2025)
von: Li, Yanshu, et al.
Veröffentlicht: (2025)
GOT-JEPA: Generic Object Tracking with Model Adaptation and Occlusion Handling using Joint-Embedding Predictive Architecture
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2026)
von: Chen, Shih-Fang, et al.
Veröffentlicht: (2026)
JEPA4Rec: Learning Effective Language Representations for Sequential Recommendation via Joint Embedding Predictive Architecture
von: Nguyen, Minh-Anh, et al.
Veröffentlicht: (2025)
von: Nguyen, Minh-Anh, et al.
Veröffentlicht: (2025)
HQ-JEPA: Hybrid Quantum Joint-Embedding Predictive Architecture for Cross-Modal Remote Sensing Representation Learning
von: Hossain, Md Aminur, et al.
Veröffentlicht: (2026)
von: Hossain, Md Aminur, et al.
Veröffentlicht: (2026)
CrossJEPA: Cross-Modal Joint-Embedding Predictive Architecture for Efficient 3D Representation Learning from 2D Images
von: Perera, Avishka, et al.
Veröffentlicht: (2025)
von: Perera, Avishka, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
C3-OWD: A Curriculum Cross-modal Contrastive Learning Framework for Open-World Detection
von: Wang, Siheng, et al.
Veröffentlicht: (2025) -
DeCo-DETR: Decoupled Cognition DETR for efficient Open-Vocabulary Object Detection
von: Wang, Siheng, et al.
Veröffentlicht: (2026) -
ChromouVQA: Benchmarking Vision-Language Models under Chromatic Camouflaged Images
von: Zhang, Yunfei, et al.
Veröffentlicht: (2025) -
T-JEPA: A Joint-Embedding Predictive Architecture for Trajectory Similarity Computation
von: Li, Lihuan, et al.
Veröffentlicht: (2024) -
Demo-JEPA: Joint-Embedding Predictive Architecture for One-shot Cross-Embodiment Imitation
von: He, Jingyang, et al.
Veröffentlicht: (2026)