Foundation Models Secretly Understand Neural Network Weights: Enhancing Hypernetwork Architectures with Foundation Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Gu, Jeffrey, Yeung-Levy, Serena |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Revisiting Active Learning in the Era of Vision Foundation Models
por: Gupte, Sanket Rajan, et al.
Publicado: (2024)
por: Gupte, Sanket Rajan, et al.
Publicado: (2024)
NegVQA: Can Vision Language Models Understand Negation?
por: Zhang, Yuhui, et al.
Publicado: (2025)
por: Zhang, Yuhui, et al.
Publicado: (2025)
Just Shift It: Test-Time Prototype Shifting for Zero-Shot Generalization with Vision-Language Models
por: Sui, Elaine, et al.
Publicado: (2024)
por: Sui, Elaine, et al.
Publicado: (2024)
Viewpoint Textual Inversion: Discovering Scene Representations and 3D View Control in 2D Diffusion Models
por: Burgess, James, et al.
Publicado: (2023)
por: Burgess, James, et al.
Publicado: (2023)
V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think
por: Tang, Bingda, et al.
Publicado: (2026)
por: Tang, Bingda, et al.
Publicado: (2026)
Rethinking Weight Decay for Robust Fine-Tuning of Foundation Models
por: Tian, Junjiao, et al.
Publicado: (2024)
por: Tian, Junjiao, et al.
Publicado: (2024)
SAM-CLIP: Merging Vision Foundation Models towards Semantic and Spatial Understanding
por: Wang, Haoxiang, et al.
Publicado: (2023)
por: Wang, Haoxiang, et al.
Publicado: (2023)
VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation
por: Wu, Yecheng, et al.
Publicado: (2024)
por: Wu, Yecheng, et al.
Publicado: (2024)
Dynamic Allocation Hypernetwork with Adaptive Model Recalibration for FCL
por: Qi, Xiaoming, et al.
Publicado: (2025)
por: Qi, Xiaoming, et al.
Publicado: (2025)
Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data
por: Zhang, Yuhui, et al.
Publicado: (2024)
por: Zhang, Yuhui, et al.
Publicado: (2024)
Enhancing Cognition and Explainability of Multimodal Foundation Models with Self-Synthesized Data
por: Shi, Yucheng, et al.
Publicado: (2025)
por: Shi, Yucheng, et al.
Publicado: (2025)
Unlocking Noise-Resistant Vision: Key Architectural Secrets for Robust Models
por: Kim, Bum Jun, et al.
Publicado: (2025)
por: Kim, Bum Jun, et al.
Publicado: (2025)
Weight Weaving: Parameter Pooling for Data-Free Model Merging
por: Chaves, Levy, et al.
Publicado: (2025)
por: Chaves, Levy, et al.
Publicado: (2025)
Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models
por: Eyring, Luca, et al.
Publicado: (2025)
por: Eyring, Luca, et al.
Publicado: (2025)
Foundation Visual Encoders Are Secretly Few-Shot Anomaly Detectors
por: Zhai, Guangyao, et al.
Publicado: (2025)
por: Zhai, Guangyao, et al.
Publicado: (2025)
CellVTA: Enhancing Vision Foundation Models for Accurate Cell Segmentation and Classification
por: Yang, Yang, et al.
Publicado: (2025)
por: Yang, Yang, et al.
Publicado: (2025)
Do Vision Foundation Models Enhance Domain Generalization in Medical Image Segmentation?
por: Cekmeceli, Kerem, et al.
Publicado: (2024)
por: Cekmeceli, Kerem, et al.
Publicado: (2024)
Ask, Pose, Unite: Scaling Data Acquisition for Close Interactions with Vision Language Models
por: Bravo-Sánchez, Laura, et al.
Publicado: (2024)
por: Bravo-Sánchez, Laura, et al.
Publicado: (2024)
Data-Efficient Inference of Neural Fluid Fields via SciML Foundation Model
por: Liu, Yuqiu, et al.
Publicado: (2024)
por: Liu, Yuqiu, et al.
Publicado: (2024)
Streamlined Photoacoustic Image Processing with Foundation Models: A Training-Free Solution
por: Deng, Handi, et al.
Publicado: (2024)
por: Deng, Handi, et al.
Publicado: (2024)
Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders
por: Jiang, Yitong, et al.
Publicado: (2026)
por: Jiang, Yitong, et al.
Publicado: (2026)
Dynamic Allocation Hypernetwork with Adaptive Model Recalibration for Federated Continual Learning
por: Qi, Xiaoming, et al.
Publicado: (2025)
por: Qi, Xiaoming, et al.
Publicado: (2025)
Enhancing Generalization of Depth Estimation Foundation Model via Weakly-Supervised Adaptation with Regularization
por: Huang, Yan, et al.
Publicado: (2025)
por: Huang, Yan, et al.
Publicado: (2025)
Fine-tuning MLLMs Without Forgetting Is Easier Than You Think
por: Li, He, et al.
Publicado: (2026)
por: Li, He, et al.
Publicado: (2026)
D'OH: Decoder-Only Random Hypernetworks for Implicit Neural Representations
por: Gordon, Cameron, et al.
Publicado: (2024)
por: Gordon, Cameron, et al.
Publicado: (2024)
Quickly Tuning Foundation Models for Image Segmentation
por: Das, Breenda, et al.
Publicado: (2025)
por: Das, Breenda, et al.
Publicado: (2025)
VMDT: Decoding the Trustworthiness of Video Foundation Models
por: Potter, Yujin, et al.
Publicado: (2025)
por: Potter, Yujin, et al.
Publicado: (2025)
TerraTorch: The Geospatial Foundation Models Toolkit
por: Gomes, Carlos, et al.
Publicado: (2025)
por: Gomes, Carlos, et al.
Publicado: (2025)
On the Generalizability of Foundation Models for Crop Type Mapping
por: Chang, Yi-Chia, et al.
Publicado: (2024)
por: Chang, Yi-Chia, et al.
Publicado: (2024)
Temporal Preference Optimization for Long-Form Video Understanding
por: Li, Rui, et al.
Publicado: (2025)
por: Li, Rui, et al.
Publicado: (2025)
Unearthing Skill-Level Insights for Understanding Trade-Offs of Foundation Models
por: Moayeri, Mazda, et al.
Publicado: (2024)
por: Moayeri, Mazda, et al.
Publicado: (2024)
Delving into Multi-modal Multi-task Foundation Models for Road Scene Understanding: From Learning Paradigm Perspectives
por: Luo, Sheng, et al.
Publicado: (2024)
por: Luo, Sheng, et al.
Publicado: (2024)
What Secrets Do Your Manifolds Hold? Understanding the Local Geometry of Generative Models
por: Humayun, Ahmed Imtiaz, et al.
Publicado: (2024)
por: Humayun, Ahmed Imtiaz, et al.
Publicado: (2024)
Foundation Model-oriented Robustness: Robust Image Model Evaluation with Pretrained Models
por: Zhang, Peiyan, et al.
Publicado: (2023)
por: Zhang, Peiyan, et al.
Publicado: (2023)
Downscaling Intelligence: Exploring Perception and Reasoning Bottlenecks in Small Multimodal Models
por: Endo, Mark, et al.
Publicado: (2025)
por: Endo, Mark, et al.
Publicado: (2025)
Test-Time Canonicalization by Foundation Models for Robust Perception
por: Singhal, Utkarsh, et al.
Publicado: (2025)
por: Singhal, Utkarsh, et al.
Publicado: (2025)
Curia: A Multi-Modal Foundation Model for Radiology
por: Dancette, Corentin, et al.
Publicado: (2025)
por: Dancette, Corentin, et al.
Publicado: (2025)
Landsat-Bench: Datasets and Benchmarks for Landsat Foundation Models
por: Corley, Isaac, et al.
Publicado: (2025)
por: Corley, Isaac, et al.
Publicado: (2025)
EEG Foundation Models: Progresses, Benchmarking, and Open Problems
por: Liu, Dingkun, et al.
Publicado: (2026)
por: Liu, Dingkun, et al.
Publicado: (2026)
Vision Foundation Models in Remote Sensing: A Survey
por: Lu, Siqi, et al.
Publicado: (2024)
por: Lu, Siqi, et al.
Publicado: (2024)
Ejemplares similares
-
Revisiting Active Learning in the Era of Vision Foundation Models
por: Gupte, Sanket Rajan, et al.
Publicado: (2024) -
NegVQA: Can Vision Language Models Understand Negation?
por: Zhang, Yuhui, et al.
Publicado: (2025) -
Just Shift It: Test-Time Prototype Shifting for Zero-Shot Generalization with Vision-Language Models
por: Sui, Elaine, et al.
Publicado: (2024) -
Viewpoint Textual Inversion: Discovering Scene Representations and 3D View Control in 2D Diffusion Models
por: Burgess, James, et al.
Publicado: (2023) -
V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think
por: Tang, Bingda, et al.
Publicado: (2026)