Procedure Learning via Regularized Gromov-Wasserstein Optimal Transport
Fuente:
arXiv
Saved in:
| Main Authors: | Mahmood, Syed Ahmed, Ali, Ali Shah, Ahmed, Umer, Fateh, Fawad Javed, Zia, M. Zeeshan, Tran, Quoc-Huy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unsupervised Skeleton-Based Action Segmentation via Hierarchical Spatiotemporal Vector Quantization
by: Ahmed, Umer, et al.
Published: (2026)
by: Ahmed, Umer, et al.
Published: (2026)
TemporalVLM: Video LLMs for Temporal Reasoning in Long Videos
by: Fateh, Fawad Javed, et al.
Published: (2024)
by: Fateh, Fawad Javed, et al.
Published: (2024)
Joint Self-Supervised Video Alignment and Action Segmentation
by: Ali, Ali Shah, et al.
Published: (2025)
by: Ali, Ali Shah, et al.
Published: (2025)
Action Segmentation Using 2D Skeleton Heatmaps and Multi-Modality Fusion
by: Hyder, Syed Waleed, et al.
Published: (2023)
by: Hyder, Syed Waleed, et al.
Published: (2023)
Learning by Aligning 2D Skeleton Sequences and Multi-Modality Fusion
by: Tran, Quoc-Huy, et al.
Published: (2023)
by: Tran, Quoc-Huy, et al.
Published: (2023)
Permutation-Aware Action Segmentation via Unsupervised Frame-to-Segment Alignment
by: Tran, Quoc-Huy, et al.
Published: (2023)
by: Tran, Quoc-Huy, et al.
Published: (2023)
Split-Fuse-Transport: Annotation-Free Saliency via Dual Clustering and Optimal Transport Alignment
by: Ramzan, Muhammad Umer, et al.
Published: (2025)
by: Ramzan, Muhammad Umer, et al.
Published: (2025)
Gromov Wasserstein Optimal Transport for Semantic Correspondences
by: Snelgar, Francis, et al.
Published: (2026)
by: Snelgar, Francis, et al.
Published: (2026)
Test-Time Adaptation for Anomaly Segmentation via Topology-Aware Optimal Transport Chaining
by: Zia, Ali, et al.
Published: (2026)
by: Zia, Ali, et al.
Published: (2026)
3D Reconstruction via Incremental Structure From Motion
by: Zeeshan, Muhammad, et al.
Published: (2025)
by: Zeeshan, Muhammad, et al.
Published: (2025)
Enhancing Human-Likeness in Reinforcement Learning Agents via Hierarchical Macro Action Quantization
by: Nizamani, Usman, et al.
Published: (2026)
by: Nizamani, Usman, et al.
Published: (2026)
Improving Hyperbolic Representations via Gromov-Wasserstein Regularization
by: Yang, Yifei, et al.
Published: (2024)
by: Yang, Yifei, et al.
Published: (2024)
A Hierarchical Spatiotemporal Action Tokenizer for In-Context Imitation Learning in Robotics
by: Fateh, Fawad Javed, et al.
Published: (2026)
by: Fateh, Fawad Javed, et al.
Published: (2026)
Shape-of-You: Fused Gromov-Wasserstein Optimal Transport for Semantic Correspondence in-the-Wild
by: Im, Jiin, et al.
Published: (2026)
by: Im, Jiin, et al.
Published: (2026)
Leveraging counterfactual concepts for debugging and improving CNN model performance
by: Tariq, Syed Ali, et al.
Published: (2025)
by: Tariq, Syed Ali, et al.
Published: (2025)
Min Generalized Sliced Gromov Wasserstein: A Scalable Path to Gromov Wasserstein
by: Shahbazi, Ashkan, et al.
Published: (2026)
by: Shahbazi, Ashkan, et al.
Published: (2026)
Component-Aware Sketch-to-Image Generation Using Self-Attention Encoding and Coordinate-Preserving Fusion
by: Zia, Ali, et al.
Published: (2026)
by: Zia, Ali, et al.
Published: (2026)
Locally-Focused Face Representation for Sketch-to-Image Generation Using Noise-Induced Refinement
by: Ramzan, Muhammad Umer, et al.
Published: (2024)
by: Ramzan, Muhammad Umer, et al.
Published: (2024)
Multi-view Action Recognition via Directed Gromov-Wasserstein Discrepancy
by: Nguyen, Hoang-Quan, et al.
Published: (2024)
by: Nguyen, Hoang-Quan, et al.
Published: (2024)
Towards Counterfactual and Contrastive Explainability and Transparency of DCNN Image Classifiers
by: Tariq, Syed Ali, et al.
Published: (2025)
by: Tariq, Syed Ali, et al.
Published: (2025)
An Improved Fault Diagnosis Strategy for Induction Motors Using Weighted Probability Ensemble Deep Learning
by: Ali, Usman, et al.
Published: (2024)
by: Ali, Usman, et al.
Published: (2024)
Viper-F1: Fast and Fine-Grained Multimodal Understanding with Cross-Modal State-Space Modulation
by: Trinh, Quoc-Huy
Published: (2025)
by: Trinh, Quoc-Huy
Published: (2025)
Geometry-Aware Semantic Reasoning for Training Free Video Anomaly Detection
by: Zia, Ali, et al.
Published: (2026)
by: Zia, Ali, et al.
Published: (2026)
The Joint Gromov Wasserstein Objective for Multiple Object Matching
by: Riahi, Aryan Tajmir, et al.
Published: (2025)
by: Riahi, Aryan Tajmir, et al.
Published: (2025)
Investigating Zero-Shot Diagnostic Pathology in Vision-Language Models with Efficient Prompt Design
by: Sharma, Vasudev, et al.
Published: (2025)
by: Sharma, Vasudev, et al.
Published: (2025)
Faithful Counterfactual Visual Explanations (FCVE)
by: Khan, Bismillah, et al.
Published: (2025)
by: Khan, Bismillah, et al.
Published: (2025)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
Gromov-Wasserstein-like Distances in the Gaussian Mixture Models Space
by: Salmona, Antoine, et al.
Published: (2023)
by: Salmona, Antoine, et al.
Published: (2023)
SGW-GAN: Sliced Gromov-Wasserstein Guided GANs for Retinal Fundus Image Enhancement
by: Xiong, Yujian, et al.
Published: (2026)
by: Xiong, Yujian, et al.
Published: (2026)
2D_3D Feature Fusion via Cross-Modal Latent Synthesis and Attention Guided Restoration for Industrial Anomaly Detection
by: Ali, Usman, et al.
Published: (2025)
by: Ali, Usman, et al.
Published: (2025)
Pseudo-label Refinement for Improving Self-Supervised Learning Systems
by: Zia-ur-Rehman, et al.
Published: (2024)
by: Zia-ur-Rehman, et al.
Published: (2024)
Rethinking Model Selection in VLM Through the Lens of Gromov-Wasserstein Distance
by: Li, Muyang, et al.
Published: (2026)
by: Li, Muyang, et al.
Published: (2026)
Bridging Optimal Transport and Jacobian Regularization by Optimal Trajectory for Enhanced Adversarial Defense
by: Le, Binh M., et al.
Published: (2023)
by: Le, Binh M., et al.
Published: (2023)
VRU-CIPI: Crossing Intention Prediction at Intersections for Improving Vulnerable Road Users Safety
by: Abdelrahman, Ahmed S., et al.
Published: (2025)
by: Abdelrahman, Ahmed S., et al.
Published: (2025)
A Lightweight and Interpretable Deepfakes Detection Framework
by: Farooq, Muhammad Umar, et al.
Published: (2025)
by: Farooq, Muhammad Umar, et al.
Published: (2025)
Understanding Learning with Sliced-Wasserstein Requires Rethinking Informative Slices
by: Tran, Huy, et al.
Published: (2024)
by: Tran, Huy, et al.
Published: (2024)
SuCor: Susceptibility Distortion Correction via Parameter-Free and Self-Regularized Optimal Transport
by: Chigurupati, Sreekar, et al.
Published: (2026)
by: Chigurupati, Sreekar, et al.
Published: (2026)
An Attention Based Pipeline for Identifying Pre-Cancer Lesions in Head and Neck Clinical Images
by: Alsalemi, Abdullah, et al.
Published: (2024)
by: Alsalemi, Abdullah, et al.
Published: (2024)
Attention-Based Ensemble Learning for Crop Classification Using Landsat 8-9 Fusion
by: Ramzan, Zeeshan, et al.
Published: (2025)
by: Ramzan, Zeeshan, et al.
Published: (2025)
Vim4Path: Self-Supervised Vision Mamba for Histopathology Images
by: Nasiri-Sarvi, Ali, et al.
Published: (2024)
by: Nasiri-Sarvi, Ali, et al.
Published: (2024)
Similar Items
-
Unsupervised Skeleton-Based Action Segmentation via Hierarchical Spatiotemporal Vector Quantization
by: Ahmed, Umer, et al.
Published: (2026) -
TemporalVLM: Video LLMs for Temporal Reasoning in Long Videos
by: Fateh, Fawad Javed, et al.
Published: (2024) -
Joint Self-Supervised Video Alignment and Action Segmentation
by: Ali, Ali Shah, et al.
Published: (2025) -
Action Segmentation Using 2D Skeleton Heatmaps and Multi-Modality Fusion
by: Hyder, Syed Waleed, et al.
Published: (2023) -
Learning by Aligning 2D Skeleton Sequences and Multi-Modality Fusion
by: Tran, Quoc-Huy, et al.
Published: (2023)