Leveraging Foundation Models for Causal Generative Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Komanduri, Aneesh, Wu, Xintao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CausalVLBench: Benchmarking Visual Causal Reasoning in Large Vision-Language Models
by: Komanduri, Aneesh, et al.
Published: (2025)
by: Komanduri, Aneesh, et al.
Published: (2025)
From Identifiable Causal Representations to Controllable Counterfactual Generation: A Survey on Causal Generative Modeling
by: Komanduri, Aneesh, et al.
Published: (2023)
by: Komanduri, Aneesh, et al.
Published: (2023)
Causal Diffusion Autoencoders: Toward Counterfactual Generation via Diffusion Probabilistic Models
by: Komanduri, Aneesh, et al.
Published: (2024)
by: Komanduri, Aneesh, et al.
Published: (2024)
Cross-Modal Attention Guided Unlearning in Vision-Language Models
by: Bhaila, Karuna, et al.
Published: (2025)
by: Bhaila, Karuna, et al.
Published: (2025)
Beyond Human Vision: The Role of Large Vision Language Models in Microscope Image Analysis
by: Verma, Prateek, et al.
Published: (2024)
by: Verma, Prateek, et al.
Published: (2024)
Leveraging Model Guidance to Extract Training Data from Personalized Diffusion Models
by: Wu, Xiaoyu, et al.
Published: (2024)
by: Wu, Xiaoyu, et al.
Published: (2024)
Semi-Supervised Learning for Deep Causal Generative Models
by: Ibrahim, Yasin, et al.
Published: (2024)
by: Ibrahim, Yasin, et al.
Published: (2024)
Collaborating Foundation Models for Domain Generalized Semantic Segmentation
by: Benigmim, Yasser, et al.
Published: (2023)
by: Benigmim, Yasser, et al.
Published: (2023)
FunduSegmenter: Leveraging the RETFound Foundation Model for Joint Optic Disc and Optic Cup Segmentation in Retinal Fundus Images
by: Zhao, Zhenyi, et al.
Published: (2025)
by: Zhao, Zhenyi, et al.
Published: (2025)
DiffusionSat: A Generative Foundation Model for Satellite Imagery
by: Khanna, Samar, et al.
Published: (2023)
by: Khanna, Samar, et al.
Published: (2023)
A Foundation Model for General Moving Object Segmentation in Medical Images
by: Yan, Zhongnuo, et al.
Published: (2023)
by: Yan, Zhongnuo, et al.
Published: (2023)
Revisiting Model Stitching In the Foundation Model Era
by: Mai, Zheda, et al.
Published: (2026)
by: Mai, Zheda, et al.
Published: (2026)
Human-Centric Foundation Models: Perception, Generation and Agentic Modeling
by: Tang, Shixiang, et al.
Published: (2025)
by: Tang, Shixiang, et al.
Published: (2025)
Leveraging GANs For Active Appearance Models Optimized Model Fitting
by: Awasthi, Anurag
Published: (2025)
by: Awasthi, Anurag
Published: (2025)
Cross-Domain Generalization Limits of Vision Foundation Models in Facial Deepfake Detection
by: Delibasoglu, Ibrahim
Published: (2026)
by: Delibasoglu, Ibrahim
Published: (2026)
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation
by: Arkhipkin, Vladimir, et al.
Published: (2025)
by: Arkhipkin, Vladimir, et al.
Published: (2025)
Causality-Driven Audits of Model Robustness
by: Drenkow, Nathan, et al.
Published: (2024)
by: Drenkow, Nathan, et al.
Published: (2024)
Research on the Spatial Data Intelligent Foundation Model
by: Wang, Shaohua, et al.
Published: (2024)
by: Wang, Shaohua, et al.
Published: (2024)
SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
by: Chu, Tianzhe, et al.
Published: (2025)
by: Chu, Tianzhe, et al.
Published: (2025)
An Investigation of Visual Foundation Models Robustness
by: Gupta, Sandeep, et al.
Published: (2025)
by: Gupta, Sandeep, et al.
Published: (2025)
VFMF: World Modeling by Forecasting Vision Foundation Model Features
by: Boduljak, Gabrijel, et al.
Published: (2025)
by: Boduljak, Gabrijel, et al.
Published: (2025)
Spurious Feature Eraser: Stabilizing Test-Time Adaptation for Vision-Language Foundation Model
by: Ma, Huan, et al.
Published: (2024)
by: Ma, Huan, et al.
Published: (2024)
Learning and Leveraging World Models in Visual Representation Learning
by: Garrido, Quentin, et al.
Published: (2024)
by: Garrido, Quentin, et al.
Published: (2024)
MOMO: Mars Orbital Model Foundation Model for Mars Orbital Applications
by: Purohit, Mirali, et al.
Published: (2026)
by: Purohit, Mirali, et al.
Published: (2026)
3D-LFM: Lifting Foundation Model
by: Dabhi, Mosam, et al.
Published: (2023)
by: Dabhi, Mosam, et al.
Published: (2023)
(Almost) Free Modality Stitching of Foundation Models
by: Singh, Jaisidh, et al.
Published: (2025)
by: Singh, Jaisidh, et al.
Published: (2025)
Domain-Aware Fine-Tuning of Foundation Models
by: Kaplan, Ugur Ali, et al.
Published: (2024)
by: Kaplan, Ugur Ali, et al.
Published: (2024)
A Comprehensive Survey of Foundation Models in Medicine
by: Khan, Wasif, et al.
Published: (2024)
by: Khan, Wasif, et al.
Published: (2024)
PRISM: Distributed Inference for Foundation Models at Edge
by: Qazi, Muhammad Azlan, et al.
Published: (2025)
by: Qazi, Muhammad Azlan, et al.
Published: (2025)
Learning Truncated Causal History Model for Video Restoration
by: Ghasemabadi, Amirhosein, et al.
Published: (2024)
by: Ghasemabadi, Amirhosein, et al.
Published: (2024)
MagMax: Leveraging Model Merging for Seamless Continual Learning
by: Marczak, Daniel, et al.
Published: (2024)
by: Marczak, Daniel, et al.
Published: (2024)
PFM-VEPAR: Prompting Foundation Models for RGB-Event Camera based Pedestrian Attribute Recognition
by: Xu, Minghe, et al.
Published: (2026)
by: Xu, Minghe, et al.
Published: (2026)
CFM: Language-aligned Concept Foundation Model for Vision
by: Wittenmayer, Kai, et al.
Published: (2026)
by: Wittenmayer, Kai, et al.
Published: (2026)
GARDO: Reinforcing Diffusion Models without Reward Hacking
by: He, Haoran, et al.
Published: (2025)
by: He, Haoran, et al.
Published: (2025)
First On-Orbit Demonstration of a Geospatial Foundation Model
by: Du, Andrew, et al.
Published: (2025)
by: Du, Andrew, et al.
Published: (2025)
Training Video Foundation Models with NVIDIA NeMo
by: Patel, Zeeshan, et al.
Published: (2025)
by: Patel, Zeeshan, et al.
Published: (2025)
Bridging Remote Sensors with Multisensor Geospatial Foundation Models
by: Han, Boran, et al.
Published: (2024)
by: Han, Boran, et al.
Published: (2024)
BrainDINO: A Brain MRI Foundation Model for Generalizable Clinical Representation Learning
by: Wu, Yizhou, et al.
Published: (2026)
by: Wu, Yizhou, et al.
Published: (2026)
Causal Decoding for Hallucination-Resistant Multimodal Large Language Models
by: Tan, Shiwei, et al.
Published: (2026)
by: Tan, Shiwei, et al.
Published: (2026)
Advances in Multimodal Adaptation and Generalization: From Traditional Approaches to Foundation Models
by: Dong, Hao, et al.
Published: (2025)
by: Dong, Hao, et al.
Published: (2025)
Similar Items
-
CausalVLBench: Benchmarking Visual Causal Reasoning in Large Vision-Language Models
by: Komanduri, Aneesh, et al.
Published: (2025) -
From Identifiable Causal Representations to Controllable Counterfactual Generation: A Survey on Causal Generative Modeling
by: Komanduri, Aneesh, et al.
Published: (2023) -
Causal Diffusion Autoencoders: Toward Counterfactual Generation via Diffusion Probabilistic Models
by: Komanduri, Aneesh, et al.
Published: (2024) -
Cross-Modal Attention Guided Unlearning in Vision-Language Models
by: Bhaila, Karuna, et al.
Published: (2025) -
Beyond Human Vision: The Role of Large Vision Language Models in Microscope Image Analysis
by: Verma, Prateek, et al.
Published: (2024)