Mimetic Initialization of MLPs
Fuente:
arXiv
Salvato in:
| Autori principali: | Trockman, Asher, Kolter, J. Zico |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Mimetic Initialization Helps State Space Models Learn to Recall
di: Trockman, Asher, et al.
Pubblicazione: (2024)
di: Trockman, Asher, et al.
Pubblicazione: (2024)
Inference Optimal VLMs Need Fewer Visual Tokens and More Parameters
di: Li, Kevin Y., et al.
Pubblicazione: (2024)
di: Li, Kevin Y., et al.
Pubblicazione: (2024)
One-Step Diffusion Distillation through Score Implicit Matching
di: Luo, Weijian, et al.
Pubblicazione: (2024)
di: Luo, Weijian, et al.
Pubblicazione: (2024)
Blind Inverse Problem Solving Made Easy by Text-to-Image Latent Diffusion
di: Dontas, Michail, et al.
Pubblicazione: (2024)
di: Dontas, Michail, et al.
Pubblicazione: (2024)
VQGraph: Rethinking Graph Representation Space for Bridging GNNs and MLPs
di: Yang, Ling, et al.
Pubblicazione: (2023)
di: Yang, Ling, et al.
Pubblicazione: (2023)
One-Step Diffusion Distillation via Deep Equilibrium Models
di: Geng, Zhengyang, et al.
Pubblicazione: (2023)
di: Geng, Zhengyang, et al.
Pubblicazione: (2023)
Diffusing Differentiable Representations
di: Savani, Yash, et al.
Pubblicazione: (2024)
di: Savani, Yash, et al.
Pubblicazione: (2024)
Automated Black-box Prompt Engineering for Personalized Text-to-Image Generation
di: He, Yutong, et al.
Pubblicazione: (2024)
di: He, Yutong, et al.
Pubblicazione: (2024)
Antidistillation Fingerprinting
di: Xu, Yixuan Even, et al.
Pubblicazione: (2026)
di: Xu, Yixuan Even, et al.
Pubblicazione: (2026)
Finetuning CLIP to Reason about Pairwise Differences
di: Sam, Dylan, et al.
Pubblicazione: (2024)
di: Sam, Dylan, et al.
Pubblicazione: (2024)
Prompt Recovery for Image Generation Models: A Comparative Study of Discrete Optimizers
di: Williams, Joshua Nathaniel, et al.
Pubblicazione: (2024)
di: Williams, Joshua Nathaniel, et al.
Pubblicazione: (2024)
Improving Alignment and Robustness with Circuit Breakers
di: Zou, Andy, et al.
Pubblicazione: (2024)
di: Zou, Andy, et al.
Pubblicazione: (2024)
Mean Flows for One-step Generative Modeling
di: Geng, Zhengyang, et al.
Pubblicazione: (2025)
di: Geng, Zhengyang, et al.
Pubblicazione: (2025)
Consistency Models Made Easy
di: Geng, Zhengyang, et al.
Pubblicazione: (2024)
di: Geng, Zhengyang, et al.
Pubblicazione: (2024)
From Variance to Veracity: Unbundling and Mitigating Gradient Variance in Differentiable Bundle Adjustment Layers
di: Gurumurthy, Swaminathan, et al.
Pubblicazione: (2024)
di: Gurumurthy, Swaminathan, et al.
Pubblicazione: (2024)
Improved Mean Flows: On the Challenges of Fastforward Generative Models
di: Geng, Zhengyang, et al.
Pubblicazione: (2025)
di: Geng, Zhengyang, et al.
Pubblicazione: (2025)
Tackling Structural Hallucination in Image Translation with Local Diffusion
di: Kim, Seunghoi, et al.
Pubblicazione: (2024)
di: Kim, Seunghoi, et al.
Pubblicazione: (2024)
AdapterTune: Zero-Initialized Low-Rank Adapters for Frozen Vision Transformers
di: Khazem, Salim
Pubblicazione: (2026)
di: Khazem, Salim
Pubblicazione: (2026)
Optimal Eye Surgeon: Finding Image Priors through Sparse Generators at Initialization
di: Ghosh, Avrajit, et al.
Pubblicazione: (2024)
di: Ghosh, Avrajit, et al.
Pubblicazione: (2024)
Exploring Learngene via Stage-wise Weight Sharing for Initializing Variable-sized Models
di: Xia, Shi-Yu, et al.
Pubblicazione: (2024)
di: Xia, Shi-Yu, et al.
Pubblicazione: (2024)
T-MARS: Improving Visual Representations by Circumventing Text Feature Learning
di: Maini, Pratyush, et al.
Pubblicazione: (2023)
di: Maini, Pratyush, et al.
Pubblicazione: (2023)
Psi-Sampler: Initial Particle Sampling for SMC-Based Inference-Time Reward Alignment in Score Models
di: Yoon, Taehoon, et al.
Pubblicazione: (2025)
di: Yoon, Taehoon, et al.
Pubblicazione: (2025)
HyperCLIP: Adapting Vision-Language models with Hypernetworks
di: Akinwande, Victor, et al.
Pubblicazione: (2024)
di: Akinwande, Victor, et al.
Pubblicazione: (2024)
DART: Implicit Doppler Tomography for Radar Novel View Synthesis
di: Huang, Tianshu, et al.
Pubblicazione: (2024)
di: Huang, Tianshu, et al.
Pubblicazione: (2024)
Joint Distillation for Fast Likelihood Evaluation and Sampling in Flow-based Models
di: Ai, Xinyue, et al.
Pubblicazione: (2025)
di: Ai, Xinyue, et al.
Pubblicazione: (2025)
The Illusion of Forgetting: Attack Unlearned Diffusion via Initial Latent Variable Optimization
di: Li, Manyi, et al.
Pubblicazione: (2026)
di: Li, Manyi, et al.
Pubblicazione: (2026)
In-Context Credit Assignment via the Core
di: Harris, Keegan, et al.
Pubblicazione: (2026)
di: Harris, Keegan, et al.
Pubblicazione: (2026)
Representation Engineering: A Top-Down Approach to AI Transparency
di: Zou, Andy, et al.
Pubblicazione: (2023)
di: Zou, Andy, et al.
Pubblicazione: (2023)
Sign Lock-In: Randomly Initialized Weight Signs Persist and Bottleneck Sub-Bit Model Compression
di: Sakai, Akira, et al.
Pubblicazione: (2026)
di: Sakai, Akira, et al.
Pubblicazione: (2026)
Efficient Point Cloud Processing with High-Dimensional Positional Encoding and Non-Local MLPs
di: Zou, Yanmei, et al.
Pubblicazione: (2026)
di: Zou, Yanmei, et al.
Pubblicazione: (2026)
Evaluation and Analysis of Deep Neural Transformers and Convolutional Neural Networks on Modern Remote Sensing Datasets
di: Hurt, J. Alex, et al.
Pubblicazione: (2025)
di: Hurt, J. Alex, et al.
Pubblicazione: (2025)
SITS-DECO: A Generative Decoder Is All You Need For Multitask Satellite Image Time Series Modelling
di: Barrett, Samuel J., et al.
Pubblicazione: (2025)
di: Barrett, Samuel J., et al.
Pubblicazione: (2025)
Multivariate Temporal Regression at Scale: A Three-Pillar Framework Combining ML, XAI, and NLP
di: Francis, Jiztom Kavalakkatt, et al.
Pubblicazione: (2025)
di: Francis, Jiztom Kavalakkatt, et al.
Pubblicazione: (2025)
LAtent Phase Inference from Short time sequences using SHallow REcurrent Decoders (LAPIS-SHRED)
di: Bao, Yuxuan, et al.
Pubblicazione: (2026)
di: Bao, Yuxuan, et al.
Pubblicazione: (2026)
FM-G-CAM: A Holistic Approach for Explainable AI in Computer Vision
di: Silva, Ravidu Suien Rammuni, et al.
Pubblicazione: (2023)
di: Silva, Ravidu Suien Rammuni, et al.
Pubblicazione: (2023)
Moving Healthcare AI-Support Systems for Visually Detectable Diseases onto Constrained Devices
di: Watt, Tess, et al.
Pubblicazione: (2024)
di: Watt, Tess, et al.
Pubblicazione: (2024)
Enhancing Worldwide Image Geolocation by Ensembling Satellite-Based Ground-Level Attribute Predictors
di: Bianco, Michael J., et al.
Pubblicazione: (2024)
di: Bianco, Michael J., et al.
Pubblicazione: (2024)
Exploring the Feasibility of Deep Learning Techniques for Accurate Gender Classification from Eye Images
di: Hasan, Basna Mohammed Salih, et al.
Pubblicazione: (2025)
di: Hasan, Basna Mohammed Salih, et al.
Pubblicazione: (2025)
A Study of Gender Classification Techniques Based on Iris Images: A Deep Survey and Analysis
di: Hasan, Basna Mohammed Salih, et al.
Pubblicazione: (2025)
di: Hasan, Basna Mohammed Salih, et al.
Pubblicazione: (2025)
CMRINet: Joint Groupwise Registration and Segmentation for Cardiac Function Quantification from Cine-MRI
di: Elmahdy, Mohamed S., et al.
Pubblicazione: (2025)
di: Elmahdy, Mohamed S., et al.
Pubblicazione: (2025)
Documenti analoghi
-
Mimetic Initialization Helps State Space Models Learn to Recall
di: Trockman, Asher, et al.
Pubblicazione: (2024) -
Inference Optimal VLMs Need Fewer Visual Tokens and More Parameters
di: Li, Kevin Y., et al.
Pubblicazione: (2024) -
One-Step Diffusion Distillation through Score Implicit Matching
di: Luo, Weijian, et al.
Pubblicazione: (2024) -
Blind Inverse Problem Solving Made Easy by Text-to-Image Latent Diffusion
di: Dontas, Michail, et al.
Pubblicazione: (2024) -
VQGraph: Rethinking Graph Representation Space for Bridging GNNs and MLPs
di: Yang, Ling, et al.
Pubblicazione: (2023)