Unleashing the Power of Image-Tabular Self-Supervised Learning via Breaking Cross-Tabular Barriers
Fuente:
arXiv
Guardado en:
| Autores principales: | Fu, Yibing, Zhao, Yunpeng, Zeng, Zhitao, Chen, Cheng, Jin, Yueming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Cross-Task Attack: A Self-Supervision Generative Framework Based on Attention Shift
por: Zeng, Qingyuan, et al.
Publicado: (2024)
por: Zeng, Qingyuan, et al.
Publicado: (2024)
Unleashing the Power of Self-Supervised Image Denoising: A Comprehensive Review
por: Zhang, Dan, et al.
Publicado: (2023)
por: Zhang, Dan, et al.
Publicado: (2023)
Positive2Negative: Breaking the Information-Lossy Barrier in Self-Supervised Single Image Denoising
por: Li, Tong, et al.
Publicado: (2024)
por: Li, Tong, et al.
Publicado: (2024)
Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning
por: Wei, Yibing, et al.
Publicado: (2024)
por: Wei, Yibing, et al.
Publicado: (2024)
Harmonized Tabular-Image Fusion via Gradient-Aligned Alternating Learning
por: Huang, Longfei, et al.
Publicado: (2026)
por: Huang, Longfei, et al.
Publicado: (2026)
Self-Supervised Learning for Detecting AI-Generated Faces as Anomalies
por: Zou, Mian, et al.
Publicado: (2025)
por: Zou, Mian, et al.
Publicado: (2025)
TIP: Tabular-Image Pre-training for Multimodal Classification with Incomplete Data
por: Du, Siyi, et al.
Publicado: (2024)
por: Du, Siyi, et al.
Publicado: (2024)
Unleashing the Power of Intermediate Domains for Mixed Domain Semi-Supervised Medical Image Segmentation
por: Ma, Qinghe, et al.
Publicado: (2025)
por: Ma, Qinghe, et al.
Publicado: (2025)
Context-driven Missing-Modality Learning for Robust Medical Diagnosis with Image-Tabular Data
por: Liu, Tianling, et al.
Publicado: (2026)
por: Liu, Tianling, et al.
Publicado: (2026)
MedVAR: Towards Scalable and Efficient Medical Image Generation via Next-scale Autoregressive Prediction
por: He, Zhicheng, et al.
Publicado: (2026)
por: He, Zhicheng, et al.
Publicado: (2026)
Self-Supervised Cross-Encoder for Neurodegenerative Disease Diagnosis
por: Cheng, Fangqi, et al.
Publicado: (2025)
por: Cheng, Fangqi, et al.
Publicado: (2025)
STiL: Semi-supervised Tabular-Image Learning for Comprehensive Task-Relevant Information Exploration in Multimodal Classification
por: Du, Siyi, et al.
Publicado: (2025)
por: Du, Siyi, et al.
Publicado: (2025)
No Data? No Problem: Robust Vision-Tabular Learning with Missing Values
por: Hasny, Marta, et al.
Publicado: (2025)
por: Hasny, Marta, et al.
Publicado: (2025)
Recognize Any Surgical Object: Unleashing the Power of Weakly-Supervised Data
por: Li, Jiajie, et al.
Publicado: (2025)
por: Li, Jiajie, et al.
Publicado: (2025)
Why and How: Knowledge-Guided Learning for Cross-Spectral Image Patch Matching
por: Yu, Chuang, et al.
Publicado: (2024)
por: Yu, Chuang, et al.
Publicado: (2024)
MulTaBench: Benchmarking Multimodal Tabular Learning with Text and Image
por: Arazi, Alan, et al.
Publicado: (2026)
por: Arazi, Alan, et al.
Publicado: (2026)
Tables Guide Vision: Learning to See the Heart through Tabular Data
por: Hasny, Marta, et al.
Publicado: (2025)
por: Hasny, Marta, et al.
Publicado: (2025)
Self-Supervised Cross-Modal Learning for Image-to-Point Cloud Registration
por: Wang, Xingmei, et al.
Publicado: (2025)
por: Wang, Xingmei, et al.
Publicado: (2025)
TIME: TabPFN-Integrated Multimodal Engine for Robust Tabular-Image Learning
por: Luo, Jiaqi, et al.
Publicado: (2025)
por: Luo, Jiaqi, et al.
Publicado: (2025)
Breaking the Resolution Barrier: Arbitrary-resolution Deep Image Steganography Framework
por: Hu, Xinjue, et al.
Publicado: (2026)
por: Hu, Xinjue, et al.
Publicado: (2026)
Dealing with All-stage Missing Modality: Towards A Universal Model with Robust Reconstruction and Personalization
por: Zhao, Yunpeng, et al.
Publicado: (2024)
por: Zhao, Yunpeng, et al.
Publicado: (2024)
CFCML: A Coarse-to-Fine Crossmodal Learning Framework For Disease Diagnosis Using Multimodal Images and Tabular Data
por: Liu, Tianling, et al.
Publicado: (2026)
por: Liu, Tianling, et al.
Publicado: (2026)
Tabular GANs for uneven distribution
por: Ashrapov, Insaf
Publicado: (2020)
por: Ashrapov, Insaf
Publicado: (2020)
AMF-MedIT: An Efficient Align-Modulation-Fusion Framework for Medical Image-Tabular Data
por: Yu, Congjing, et al.
Publicado: (2025)
por: Yu, Congjing, et al.
Publicado: (2025)
M3Ret: Unleashing Zero-shot Multimodal Medical Image Retrieval via Self-Supervision
por: Liu, Che, et al.
Publicado: (2025)
por: Liu, Che, et al.
Publicado: (2025)
LEGO: Self-Supervised Representation Learning for Scene Text Images
por: Ren, Yujin, et al.
Publicado: (2024)
por: Ren, Yujin, et al.
Publicado: (2024)
Breaking the SSL-AL Barrier: A Synergistic Semi-Supervised Active Learning Framework for 3D Object Detection
por: Wang, Zengran, et al.
Publicado: (2025)
por: Wang, Zengran, et al.
Publicado: (2025)
Breaking the Discretization Barrier of Continuous Physics Simulation Learning
por: Xu, Fan, et al.
Publicado: (2025)
por: Xu, Fan, et al.
Publicado: (2025)
Self-Supervised Contrastive Learning for Multi-Label Images
por: Chen, Jiale
Publicado: (2025)
por: Chen, Jiale
Publicado: (2025)
Multi-scale Temporal Prediction via Incremental Generation and Multi-agent Collaboration
por: Zeng, Zhitao, et al.
Publicado: (2025)
por: Zeng, Zhitao, et al.
Publicado: (2025)
Bi-Level Optimization for Self-Supervised AI-Generated Face Detection
por: Zou, Mian, et al.
Publicado: (2025)
por: Zou, Mian, et al.
Publicado: (2025)
Relational Representation Learning Network for Cross-Spectral Image Patch Matching
por: Yu, Chuang, et al.
Publicado: (2024)
por: Yu, Chuang, et al.
Publicado: (2024)
Instruction Tuning of Large Language Models for Tabular Data Generation-in One Day
por: Abdollahzadeh, Milad, et al.
Publicado: (2025)
por: Abdollahzadeh, Milad, et al.
Publicado: (2025)
VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning
por: Jia, Zi-Yi, et al.
Publicado: (2026)
por: Jia, Zi-Yi, et al.
Publicado: (2026)
SD-DiT: Unleashing the Power of Self-supervised Discrimination in Diffusion Transformer
por: Zhu, Rui, et al.
Publicado: (2024)
por: Zhu, Rui, et al.
Publicado: (2024)
Resilient Vision-Tabular Multimodal Learning under Modality Missingness
por: Caruso, Camillo Maria, et al.
Publicado: (2026)
por: Caruso, Camillo Maria, et al.
Publicado: (2026)
Harnessing Text-to-Image Diffusion Models for Point Cloud Self-Supervised Learning
por: Chen, Yiyang, et al.
Publicado: (2025)
por: Chen, Yiyang, et al.
Publicado: (2025)
CrossVideoMAE: Self-Supervised Image-Video Representation Learning with Masked Autoencoders
por: Ahamed, Shihab Aaqil, et al.
Publicado: (2025)
por: Ahamed, Shihab Aaqil, et al.
Publicado: (2025)
Weakly Supervised Cross-Modal Learning for 4D Radar Scene Flow Estimation
por: Fu, Jingyun, et al.
Publicado: (2026)
por: Fu, Jingyun, et al.
Publicado: (2026)
AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning
por: Jiang, Bo, et al.
Publicado: (2025)
por: Jiang, Bo, et al.
Publicado: (2025)
Ejemplares similares
-
Cross-Task Attack: A Self-Supervision Generative Framework Based on Attention Shift
por: Zeng, Qingyuan, et al.
Publicado: (2024) -
Unleashing the Power of Self-Supervised Image Denoising: A Comprehensive Review
por: Zhang, Dan, et al.
Publicado: (2023) -
Positive2Negative: Breaking the Information-Lossy Barrier in Self-Supervised Single Image Denoising
por: Li, Tong, et al.
Publicado: (2024) -
Towards Latent Masked Image Modeling for Self-Supervised Visual Representation Learning
por: Wei, Yibing, et al.
Publicado: (2024) -
Harmonized Tabular-Image Fusion via Gradient-Aligned Alternating Learning
por: Huang, Longfei, et al.
Publicado: (2026)