MoSiC: Optimal-Transport Motion Trajectory for Dense Self-Supervised Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Salehi, Mohammadreza, Venkataramanan, Shashanka, Simion, Ioana, Gavves, Efstratios, Snoek, Cees G. M., Asano, Yuki M |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Redefining Normal: A Novel Object-Level Approach for Multi-Object Novelty Detection
by: Salehi, Mohammadreza, et al.
Published: (2024)
by: Salehi, Mohammadreza, et al.
Published: (2024)
GeneralAD: Anomaly Detection Across Domains by Attending to Distorted Features
by: Sträter, Luc P. J., et al.
Published: (2024)
by: Sträter, Luc P. J., et al.
Published: (2024)
SIGMA: Sinkhorn-Guided Masked Video Modeling
by: Salehi, Mohammadreza, et al.
Published: (2024)
by: Salehi, Mohammadreza, et al.
Published: (2024)
Dual Guidance Semi-Supervised Action Detection
by: Singh, Ankit, et al.
Published: (2025)
by: Singh, Ankit, et al.
Published: (2025)
SelEx: Self-Expertise in Fine-Grained Generalized Category Discovery
by: Rastegar, Sarah, et al.
Published: (2024)
by: Rastegar, Sarah, et al.
Published: (2024)
Franca: Nested Matryoshka Clustering for Scalable Visual Representation Learning
by: Venkataramanan, Shashanka, et al.
Published: (2025)
by: Venkataramanan, Shashanka, et al.
Published: (2025)
Is ImageNet worth 1 video? Learning strong image encoders from 1 long unlabelled video
by: Venkataramanan, Shashanka, et al.
Published: (2023)
by: Venkataramanan, Shashanka, et al.
Published: (2023)
What Layers When: Learning to Skip Compute in LLMs with Residual Gates
by: Laitenberger, Filipe, et al.
Published: (2025)
by: Laitenberger, Filipe, et al.
Published: (2025)
Crane: Context-Guided Prompt Learning and Attention Refinement for Zero-Shot Anomaly Detection
by: Salehi, Alireza, et al.
Published: (2025)
by: Salehi, Alireza, et al.
Published: (2025)
Segment Any 3D-Part in a Scene from a Sentence
by: Wu, Hongyu, et al.
Published: (2025)
by: Wu, Hongyu, et al.
Published: (2025)
PIN: Positional Insert Unlocks Object Localisation Abilities in VLMs
by: Dorkenwald, Michael, et al.
Published: (2024)
by: Dorkenwald, Michael, et al.
Published: (2024)
LocoMotion: Learning Motion-Focused Video-Language Representations
by: Doughty, Hazel, et al.
Published: (2024)
by: Doughty, Hazel, et al.
Published: (2024)
Geometric Neural Process Fields
by: Yin, Wenzhe, et al.
Published: (2025)
by: Yin, Wenzhe, et al.
Published: (2025)
Graph Neural Networks for Learning Equivariant Representations of Neural Networks
by: Kofinas, Miltiadis, et al.
Published: (2024)
by: Kofinas, Miltiadis, et al.
Published: (2024)
MoAlign: Motion-Centric Representation Alignment for Video Diffusion Models
by: Bhowmik, Aritra, et al.
Published: (2025)
by: Bhowmik, Aritra, et al.
Published: (2025)
Elastic ViTs from Pretrained Models without Retraining
by: Simoncini, Walter, et al.
Published: (2025)
by: Simoncini, Walter, et al.
Published: (2025)
Lost in Time: A New Temporal Benchmark for VideoLLMs
by: Cores, Daniel, et al.
Published: (2024)
by: Cores, Daniel, et al.
Published: (2024)
Learn to Categorize or Categorize to Learn? Self-Coding for Generalized Category Discovery
by: Rastegar, Sarah, et al.
Published: (2023)
by: Rastegar, Sarah, et al.
Published: (2023)
CaPo: Cooperative Plan Optimization for Efficient Embodied Multi-Agent Cooperation
by: Liu, Jie, et al.
Published: (2024)
by: Liu, Jie, et al.
Published: (2024)
Self-supervised visual learning in the low-data regime: a comparative evaluation
by: Konstantakos, Sotirios, et al.
Published: (2024)
by: Konstantakos, Sotirios, et al.
Published: (2024)
From MLP to NeoMLP: Leveraging Self-Attention for Neural Fields
by: Kofinas, Miltiadis, et al.
Published: (2024)
by: Kofinas, Miltiadis, et al.
Published: (2024)
Mechanistic Interpretability for AI Safety -- A Review
by: Bereska, Leonard, et al.
Published: (2024)
by: Bereska, Leonard, et al.
Published: (2024)
Mechanistic Neural Networks for Scientific Machine Learning
by: Pervez, Adeel, et al.
Published: (2024)
by: Pervez, Adeel, et al.
Published: (2024)
Ai-Sampler: Adversarial Learning of Markov kernels with involutive maps
by: Egorov, Evgenii, et al.
Published: (2024)
by: Egorov, Evgenii, et al.
Published: (2024)
Self-supervised Learning of Echocardiographic Video Representations via Online Cluster Distillation
by: Mishra, Divyanshu, et al.
Published: (2025)
by: Mishra, Divyanshu, et al.
Published: (2025)
Beyond Model Adaptation at Test Time: A Survey
by: Xiao, Zehao, et al.
Published: (2024)
by: Xiao, Zehao, et al.
Published: (2024)
Structural, thermomechanical, and transport properties of Ti 2 NbSiC 2 and Ti 2 MoSiC 2 MAX phases as candidates for high‐temperature structural and coating materials
by: Ahmed Azzouz‐Rached, et al.
Published: (2026)
by: Ahmed Azzouz‐Rached, et al.
Published: (2026)
Near, far: Patch-ordering enhances vision foundation models' scene understanding
by: Pariza, Valentinos, et al.
Published: (2024)
by: Pariza, Valentinos, et al.
Published: (2024)
TWIST & SCOUT: Grounding Multimodal LLM-Experts by Forget-Free Tuning
by: Bhowmik, Aritra, et al.
Published: (2024)
by: Bhowmik, Aritra, et al.
Published: (2024)
TULIP: Token-length Upgraded CLIP
by: Najdenkoska, Ivona, et al.
Published: (2024)
by: Najdenkoska, Ivona, et al.
Published: (2024)
Physics-Guided Radiotherapy Treatment Planning with Deep Learning
by: Achlatis, Stefanos, et al.
Published: (2025)
by: Achlatis, Stefanos, et al.
Published: (2025)
Training-Free Semantic Segmentation via LLM-Supervision
by: Sun, Wenfang, et al.
Published: (2024)
by: Sun, Wenfang, et al.
Published: (2024)
Mechanistic PDE Networks for Discovery of Governing Equations
by: Pervez, Adeel, et al.
Published: (2025)
by: Pervez, Adeel, et al.
Published: (2025)
The Common Stability Mechanism behind most Self-Supervised Learning Approaches
by: Jha, Abhishek, et al.
Published: (2024)
by: Jha, Abhishek, et al.
Published: (2024)
Any-Resolution AI-Generated Image Detection by Spectral Learning
by: Karageorgiou, Dimitrios, et al.
Published: (2024)
by: Karageorgiou, Dimitrios, et al.
Published: (2024)
Evaluating Foundation Models' 3D Understanding Through Multi-View Correspondence Analysis
by: Lilova, Valentina, et al.
Published: (2025)
by: Lilova, Valentina, et al.
Published: (2025)
KV Cache Steering for Controlling Frozen LLMs
by: Belitsky, Max, et al.
Published: (2025)
by: Belitsky, Max, et al.
Published: (2025)
Low-Resource Vision Challenges for Foundation Models
by: Zhang, Yunhua, et al.
Published: (2024)
by: Zhang, Yunhua, et al.
Published: (2024)
Commonsense Video Question Answering through Video-Grounded Entailment Tree Reasoning
by: Liu, Huabin, et al.
Published: (2025)
by: Liu, Huabin, et al.
Published: (2025)
IPO: Interpretable Prompt Optimization for Vision-Language Models
by: Du, Yingjun, et al.
Published: (2024)
by: Du, Yingjun, et al.
Published: (2024)
Similar Items
-
Redefining Normal: A Novel Object-Level Approach for Multi-Object Novelty Detection
by: Salehi, Mohammadreza, et al.
Published: (2024) -
GeneralAD: Anomaly Detection Across Domains by Attending to Distorted Features
by: Sträter, Luc P. J., et al.
Published: (2024) -
SIGMA: Sinkhorn-Guided Masked Video Modeling
by: Salehi, Mohammadreza, et al.
Published: (2024) -
Dual Guidance Semi-Supervised Action Detection
by: Singh, Ankit, et al.
Published: (2025) -
SelEx: Self-Expertise in Fine-Grained Generalized Category Discovery
by: Rastegar, Sarah, et al.
Published: (2024)