Probabilistic Hyper-Graphs using Multiple Randomly Masked Autoencoders for Semi-supervised Multi-modal Multi-task Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Mihai-Cristian, Pîrvu, Leordeanu, Marius |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-modal video data-pipelines for machine learning with minimal human supervision
by: Pîrvu, Mihai-Cristian, et al.
Published: (2025)
by: Pîrvu, Mihai-Cristian, et al.
Published: (2025)
Multiple Random Masking Autoencoder Ensembles for Robust Multimodal Semi-supervised Learning
by: Todoran, Alexandru-Raul, et al.
Published: (2024)
by: Todoran, Alexandru-Raul, et al.
Published: (2024)
From Vision To Language through Graph of Events in Space and Time: An Explainable Self-supervised Approach
by: Masala, Mihai, et al.
Published: (2025)
by: Masala, Mihai, et al.
Published: (2025)
Towards Zero-Shot & Explainable Video Description by Reasoning over Graphs of Events in Space and Time
by: Masala, Mihai, et al.
Published: (2025)
by: Masala, Mihai, et al.
Published: (2025)
Learning from Random Subspace Exploration: Generalized Test-Time Augmentation with Self-supervised Distillation
by: Jelea, Andrei, et al.
Published: (2025)
by: Jelea, Andrei, et al.
Published: (2025)
Agentic Video Generation: From Text to Executable Event Graphs via Tool-Constrained LLM Planning
by: Cudlenco, Nicolae, et al.
Published: (2026)
by: Cudlenco, Nicolae, et al.
Published: (2026)
GTASA: Ground Truth Annotations for Spatiotemporal Analysis, Evaluation and Training of Video Models
by: Cudlenco, Nicolae, et al.
Published: (2026)
by: Cudlenco, Nicolae, et al.
Published: (2026)
MultiMatch: Multi-task Learning for Semi-supervised Domain Generalization
by: Qi, Lei, et al.
Published: (2022)
by: Qi, Lei, et al.
Published: (2022)
With Great Context Comes Great Prediction Power: Classifying Objects via Geo-Semantic Scene Graphs
by: Constantinescu, Ciprian, et al.
Published: (2025)
by: Constantinescu, Ciprian, et al.
Published: (2025)
MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks
by: Sosa, Jose, et al.
Published: (2025)
by: Sosa, Jose, et al.
Published: (2025)
Self-Supervised Learning to Fly using Efficient Semantic Segmentation and Metric Depth Estimation for Low-Cost Autonomous UAVs
by: Mocanu, Sebastian, et al.
Published: (2025)
by: Mocanu, Sebastian, et al.
Published: (2025)
Learning on the Fly: Replay-Based Continual Object Perception for Indoor Drones
by: Nae, Sebastian-Ion, et al.
Published: (2026)
by: Nae, Sebastian-Ion, et al.
Published: (2026)
Inside Knowledge: Graph-based Path Generation with Explainable Data Augmentation and Curriculum Learning for Visual Indoor Navigation
by: Airinei, Daniel, et al.
Published: (2025)
by: Airinei, Daniel, et al.
Published: (2025)
A self-supervised cyclic neural-analytic approach for novel view synthesis and 3D reconstruction
by: Costea, Dragos, et al.
Published: (2025)
by: Costea, Dragos, et al.
Published: (2025)
Improving Satellite Imagery Masking using Multi-task and Transfer Learning
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
CMViM: Contrastive Masked Vim Autoencoder for 3D Multi-modal Representation Learning for AD classification
by: Yang, Guangqian, et al.
Published: (2024)
by: Yang, Guangqian, et al.
Published: (2024)
Efficient Self-Supervised Neuro-Analytic Visual Servoing for Real-time Quadrotor Control
by: Mocanu, Sebastian, et al.
Published: (2025)
by: Mocanu, Sebastian, et al.
Published: (2025)
MaskMatch: Boosting Semi-Supervised Learning Through Mask Autoencoder-Driven Feature Learning
by: Zhang, Wenjin, et al.
Published: (2024)
by: Zhang, Wenjin, et al.
Published: (2024)
Closer to Ground Truth: Realistic Shape and Appearance Labeled Data Generation for Unsupervised Underwater Image Segmentation
by: Jelea, Andrei, et al.
Published: (2025)
by: Jelea, Andrei, et al.
Published: (2025)
Box for Mask and Mask for Box: weak losses for multi-task partially supervised learning
by: Lê, Hoàng-Ân, et al.
Published: (2024)
by: Lê, Hoàng-Ân, et al.
Published: (2024)
Towards Semi-supervised Dual-modal Semantic Segmentation
by: Dong, Qiulei, et al.
Published: (2024)
by: Dong, Qiulei, et al.
Published: (2024)
Social-MAE: Social Masked Autoencoder for Multi-person Motion Representation Learning
by: Ehsanpour, Mahsa, et al.
Published: (2024)
by: Ehsanpour, Mahsa, et al.
Published: (2024)
Multi-modal Multi-task Pre-training for Improved Point Cloud Understanding
by: Liu, Liwen, et al.
Published: (2025)
by: Liu, Liwen, et al.
Published: (2025)
Exploring Semantic Masked Autoencoder for Self-supervised Point Cloud Understanding
by: Zha, Yixin, et al.
Published: (2025)
by: Zha, Yixin, et al.
Published: (2025)
MAESIL: Masked Autoencoder for Enhanced Self-supervised Medical Image Learning
by: Kim, Kyeonghun, et al.
Published: (2026)
by: Kim, Kyeonghun, et al.
Published: (2026)
Maia: A Real-time Non-Verbal Chat for Human-AI Interaction
by: Costea, Dragos, et al.
Published: (2024)
by: Costea, Dragos, et al.
Published: (2024)
CoVLM: Leveraging Consensus from Vision-Language Models for Semi-supervised Multi-modal Fake News Detection
by: Devank, et al.
Published: (2024)
by: Devank, et al.
Published: (2024)
Semi-supervised Semantic Segmentation with Multi-Constraint Consistency Learning
by: Yin, Jianjian, et al.
Published: (2025)
by: Yin, Jianjian, et al.
Published: (2025)
MV2MAE: Multi-View Video Masked Autoencoders
by: Shah, Ketul, et al.
Published: (2024)
by: Shah, Ketul, et al.
Published: (2024)
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
Improving Masked Autoencoders by Learning Where to Mask
by: Chen, Haijian, et al.
Published: (2023)
by: Chen, Haijian, et al.
Published: (2023)
Glioma Multimodal MRI Analysis System for Tumor Layered Diagnosis via Multi-task Semi-supervised Learning
by: Liu, Yihao, et al.
Published: (2025)
by: Liu, Yihao, et al.
Published: (2025)
MDMP: Multi-modal Diffusion for supervised Motion Predictions with uncertainty
by: Bringer, Leo, et al.
Published: (2024)
by: Bringer, Leo, et al.
Published: (2024)
Downstream Task Guided Masking Learning in Masked Autoencoders Using Multi-Level Optimization
by: Guo, Han, et al.
Published: (2024)
by: Guo, Han, et al.
Published: (2024)
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
by: Zou, Jian, et al.
Published: (2023)
by: Zou, Jian, et al.
Published: (2023)
A Two-Stage Progressive Pre-training using Multi-Modal Contrastive Masked Autoencoders
by: Jamal, Muhammad Abdullah, et al.
Published: (2024)
by: Jamal, Muhammad Abdullah, et al.
Published: (2024)
MultiMAE-DER: Multimodal Masked Autoencoder for Dynamic Emotion Recognition
by: Xiang, Peihao, et al.
Published: (2024)
by: Xiang, Peihao, et al.
Published: (2024)
MultiMAE for Brain MRIs: Robustness to Missing Inputs Using Multi-Modal Masked Autoencoder
by: Erdur, Ayhan Can, et al.
Published: (2025)
by: Erdur, Ayhan Can, et al.
Published: (2025)
HyperET: Efficient Training in Hyperbolic Space for Multi-modal Large Language Models
by: Peng, Zelin, et al.
Published: (2025)
by: Peng, Zelin, et al.
Published: (2025)
SemiTooth: a Generalizable Semi-supervised Framework for Multi-Source Tooth Segmentation
by: Sun, Muyi, et al.
Published: (2026)
by: Sun, Muyi, et al.
Published: (2026)
Similar Items
-
Multi-modal video data-pipelines for machine learning with minimal human supervision
by: Pîrvu, Mihai-Cristian, et al.
Published: (2025) -
Multiple Random Masking Autoencoder Ensembles for Robust Multimodal Semi-supervised Learning
by: Todoran, Alexandru-Raul, et al.
Published: (2024) -
From Vision To Language through Graph of Events in Space and Time: An Explainable Self-supervised Approach
by: Masala, Mihai, et al.
Published: (2025) -
Towards Zero-Shot & Explainable Video Description by Reasoning over Graphs of Events in Space and Time
by: Masala, Mihai, et al.
Published: (2025) -
Learning from Random Subspace Exploration: Generalized Test-Time Augmentation with Self-supervised Distillation
by: Jelea, Andrei, et al.
Published: (2025)