Zero-Shot Vision Encoder Grafting via LLM Surrogates
Fuente:
arXiv
Saved in:
| Main Authors: | Yue, Kaiyu, Singla, Vasu, Jia, Menglin, Kirchenbauer, John, Qadri, Rifaa, Cai, Zikui, Bhatele, Abhinav, Huang, Furong, Goldstein, Tom |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Image Generation with a Sphere Encoder
by: Yue, Kaiyu, et al.
Published: (2026)
by: Yue, Kaiyu, et al.
Published: (2026)
GenQA: Generating Millions of Instructions from a Handful of Prompts
by: Chen, Jiuhai, et al.
Published: (2024)
by: Chen, Jiuhai, et al.
Published: (2024)
From Pixels to Prose: A Large Dataset of Dense Image Captions
by: Singla, Vasu, et al.
Published: (2024)
by: Singla, Vasu, et al.
Published: (2024)
AegisLLM: Scaling Agentic Systems for Self-Reflective Defense in LLM Security
by: Cai, Zikui, et al.
Published: (2025)
by: Cai, Zikui, et al.
Published: (2025)
Zebra-CoT: A Dataset for Interleaved Vision Language Reasoning
by: Li, Ang, et al.
Published: (2025)
by: Li, Ang, et al.
Published: (2025)
Gemstones: A Model Suite for Multi-Faceted Scaling Laws
by: McLeish, Sean, et al.
Published: (2025)
by: McLeish, Sean, et al.
Published: (2025)
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers
by: Singh, Siddharth, et al.
Published: (2025)
by: Singh, Siddharth, et al.
Published: (2025)
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
by: Geiping, Jonas, et al.
Published: (2025)
by: Geiping, Jonas, et al.
Published: (2025)
Speculating Experts Accelerates Inference for Mixture-of-Experts
by: Madan, Vivan, et al.
Published: (2026)
by: Madan, Vivan, et al.
Published: (2026)
Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
by: Hans, Abhimanyu, et al.
Published: (2024)
by: Hans, Abhimanyu, et al.
Published: (2024)
When Can You Get Away with Low Memory Adam?
by: Kalra, Dayal Singh, et al.
Published: (2025)
by: Kalra, Dayal Singh, et al.
Published: (2025)
Transformers Can Do Arithmetic with the Right Embeddings
by: McLeish, Sean, et al.
Published: (2024)
by: McLeish, Sean, et al.
Published: (2024)
GATES: Self-Distillation under Privileged Context with Consensus Gating
by: Stein, Alex, et al.
Published: (2026)
by: Stein, Alex, et al.
Published: (2026)
FictionalQA: A Dataset for Studying Memorization and Knowledge Acquisition
by: Kirchenbauer, John, et al.
Published: (2025)
by: Kirchenbauer, John, et al.
Published: (2025)
Speedy-Splat: Fast 3D Gaussian Splatting with Sparse Pixels and Sparse Primitives
by: Hanson, Alex, et al.
Published: (2024)
by: Hanson, Alex, et al.
Published: (2024)
PUP 3D-GS: Principled Uncertainty Pruning for 3D Gaussian Splatting
by: Hanson, Alex, et al.
Published: (2024)
by: Hanson, Alex, et al.
Published: (2024)
Analytics of Longitudinal System Monitoring Data for Performance Prediction
by: Costello, Ian J., et al.
Published: (2020)
by: Costello, Ian J., et al.
Published: (2020)
A Watermark for Large Language Models
by: Kirchenbauer, John, et al.
Published: (2023)
by: Kirchenbauer, John, et al.
Published: (2023)
Multi-Token Prediction via Self-Distillation
by: Kirchenbauer, John, et al.
Published: (2026)
by: Kirchenbauer, John, et al.
Published: (2026)
PHORECAST: Enabling AI Understanding of Public Health Outreach Across Populations
by: Qadri, Rifaa, et al.
Published: (2025)
by: Qadri, Rifaa, et al.
Published: (2025)
Power Law Guided Dynamic Sifting for Efficient Attention
by: Koley, Nirav, et al.
Published: (2025)
by: Koley, Nirav, et al.
Published: (2025)
StegaVision: Enhancing Steganography with Attention Mechanism
by: Kumar, Abhinav, et al.
Published: (2024)
by: Kumar, Abhinav, et al.
Published: (2024)
Imagine, Verify, Execute: Memory-guided Agentic Exploration with Vision-Language Models
by: Lee, Seungjae, et al.
Published: (2025)
by: Lee, Seungjae, et al.
Published: (2025)
Compositional Adversarial Training for Robust Visual Watermarking
by: Satheesh, Anirudh, et al.
Published: (2026)
by: Satheesh, Anirudh, et al.
Published: (2026)
MORSE-500: A Programmatically Controllable Video Benchmark to Stress-Test Multimodal Reasoning
by: Cai, Zikui, et al.
Published: (2025)
by: Cai, Zikui, et al.
Published: (2025)
Shadowcast: Stealthy Data Poisoning Attacks Against Vision-Language Models
by: Xu, Yuancheng, et al.
Published: (2024)
by: Xu, Yuancheng, et al.
Published: (2024)
HPC-Coder-V2: Studying Code LLMs Across Low-Resource Parallel Languages
by: Chaturvedi, Aman, et al.
Published: (2024)
by: Chaturvedi, Aman, et al.
Published: (2024)
Characterizing Production GPU Workloads using System-wide Telemetry Data
by: Cankur, Onur, et al.
Published: (2025)
by: Cankur, Onur, et al.
Published: (2025)
LMD3: Language Model Data Density Dependence
by: Kirchenbauer, John, et al.
Published: (2024)
by: Kirchenbauer, John, et al.
Published: (2024)
OPTune: Efficient Online Preference Tuning
by: Chen, Lichang, et al.
Published: (2024)
by: Chen, Lichang, et al.
Published: (2024)
Zero-Shot Reinforcement Learning via Function Encoders
by: Ingebrand, Tyler, et al.
Published: (2024)
by: Ingebrand, Tyler, et al.
Published: (2024)
Decodable and Sample Invariant Continuous Object Encoder
by: Yuan, Dehao, et al.
Published: (2023)
by: Yuan, Dehao, et al.
Published: (2023)
Enabling Natural Zero-Shot Prompting on Encoder Models via Statement-Tuning
by: Elshabrawy, Ahmed, et al.
Published: (2024)
by: Elshabrawy, Ahmed, et al.
Published: (2024)
Modular Energy Steering for Safe Text-to-Image Generation with Foundation Models
by: Tan, Yaoteng, et al.
Published: (2026)
by: Tan, Yaoteng, et al.
Published: (2026)
Targeted Unlearning with Single Layer Unlearning Gradient
by: Cai, Zikui, et al.
Published: (2024)
by: Cai, Zikui, et al.
Published: (2024)
Transform-Dependent Adversarial Attacks
by: Tan, Yaoteng, et al.
Published: (2024)
by: Tan, Yaoteng, et al.
Published: (2024)
Zoom-shot: Fast and Efficient Unsupervised Zero-Shot Transfer of CLIP to Vision Encoders with Multimodal Loss
by: Shipard, Jordan, et al.
Published: (2024)
by: Shipard, Jordan, et al.
Published: (2024)
EDDA: A Encoder-Decoder Data Augmentation Framework for Zero-Shot Stance Detection
by: Ding, Daijun, et al.
Published: (2024)
by: Ding, Daijun, et al.
Published: (2024)
Spotting LLMs With Binoculars: Zero-Shot Detection of Machine-Generated Text
by: Hans, Abhimanyu, et al.
Published: (2024)
by: Hans, Abhimanyu, et al.
Published: (2024)
Zero-Shot Function Encoder-Based Differentiable Predictive Control
by: Iqbal, Hassan, et al.
Published: (2025)
by: Iqbal, Hassan, et al.
Published: (2025)
Similar Items
-
Image Generation with a Sphere Encoder
by: Yue, Kaiyu, et al.
Published: (2026) -
GenQA: Generating Millions of Instructions from a Handful of Prompts
by: Chen, Jiuhai, et al.
Published: (2024) -
From Pixels to Prose: A Large Dataset of Dense Image Captions
by: Singla, Vasu, et al.
Published: (2024) -
AegisLLM: Scaling Agentic Systems for Self-Reflective Defense in LLM Security
by: Cai, Zikui, et al.
Published: (2025) -
Zebra-CoT: A Dataset for Interleaved Vision Language Reasoning
by: Li, Ang, et al.
Published: (2025)