Implicit Inversion turns CLIP into a Decoder
Fuente:
arXiv
Saved in:
| Main Authors: | D'Orazio, Antonio, Briglia, Maria Rosaria, Crisostomi, Donato, Loi, Dario, Rodolà, Emanuele, Masi, Iacopo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization
by: Maljkovic, Igor, et al.
Published: (2026)
by: Maljkovic, Igor, et al.
Published: (2026)
MASS: MoErging through Adaptive Subspace Selection
by: Crisostomi, Donato, et al.
Published: (2025)
by: Crisostomi, Donato, et al.
Published: (2025)
Environment Maps Editing using Inverse Rendering and Adversarial Implicit Functions
by: D'Orazio, Antonio, et al.
Published: (2024)
by: D'Orazio, Antonio, et al.
Published: (2024)
ATM: Improving Model Merging by Alternating Tuning and Merging
by: Zhou, Luca, et al.
Published: (2024)
by: Zhou, Luca, et al.
Published: (2024)
A Provable Energy-Guided Test-Time Defense Boosting Adversarial Robustness of Large Vision-Language Models
by: Mirza, Mujtaba Hussain, et al.
Published: (2026)
by: Mirza, Mujtaba Hussain, et al.
Published: (2026)
Shedding More Light on Robust Classifiers under the lens of Energy-based Models
by: Mirza, Mujtaba Hussain, et al.
Published: (2024)
by: Mirza, Mujtaba Hussain, et al.
Published: (2024)
What is Adversarial Training for Diffusion Models?
by: Rosaria, Briglia Maria, et al.
Published: (2025)
by: Rosaria, Briglia Maria, et al.
Published: (2025)
Domain Elastic Transform: Bayesian Function Registration for High-Dimensional Scientific Data
by: Hirose, Osamu, et al.
Published: (2026)
by: Hirose, Osamu, et al.
Published: (2026)
Understanding Adversarial Training with Energy-based Models
by: Mirza, Mujtaba Hussain, et al.
Published: (2025)
by: Mirza, Mujtaba Hussain, et al.
Published: (2025)
Multi-objective Evolutionary Merging Enables Efficient Reasoning Models
by: Iacobelli, Mario, et al.
Published: (2026)
by: Iacobelli, Mario, et al.
Published: (2026)
CLIP-MUSED: CLIP-Guided Multi-Subject Visual Neural Information Semantic Decoding
by: Zhou, Qiongyi, et al.
Published: (2024)
by: Zhou, Qiongyi, et al.
Published: (2024)
Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion
by: Mistretta, Marco, et al.
Published: (2025)
by: Mistretta, Marco, et al.
Published: (2025)
Zero-Shot Quantization via Weight-Space Arithmetic
by: Solombrino, Daniele, et al.
Published: (2026)
by: Solombrino, Daniele, et al.
Published: (2026)
RISE-Video: Can Video Generators Decode Implicit World Rules?
by: Liu, Mingxin, et al.
Published: (2026)
by: Liu, Mingxin, et al.
Published: (2026)
AI Art Curation: Re-imagining the city of Helsinki in occasion of its Biennial
by: Schaerf, Ludovica, et al.
Published: (2023)
by: Schaerf, Ludovica, et al.
Published: (2023)
Model Merging Improves Zero-Shot Generalization in Bioacoustic Foundation Models
by: Marincione, Davide, et al.
Published: (2025)
by: Marincione, Davide, et al.
Published: (2025)
Not All Latent Spaces Are Flat: Hyperbolic Concept Control
by: Briglia, Maria Rosaria, et al.
Published: (2026)
by: Briglia, Maria Rosaria, et al.
Published: (2026)
DeepFeatureX Net: Deep Features eXtractors based Network for discriminating synthetic from real images
by: Pontorno, Orazio, et al.
Published: (2024)
by: Pontorno, Orazio, et al.
Published: (2024)
SIEDD: Shared-Implicit Encoder with Discrete Decoders
by: Rangarajan, Vikram, et al.
Published: (2025)
by: Rangarajan, Vikram, et al.
Published: (2025)
AdaptCLIP: Adapting CLIP for Universal Visual Anomaly Detection
by: Gao, Bin-Bin, et al.
Published: (2025)
by: Gao, Bin-Bin, et al.
Published: (2025)
DesignCLIP: Multimodal Learning with CLIP for Design Patent Understanding
by: Wang, Zhu, et al.
Published: (2025)
by: Wang, Zhu, et al.
Published: (2025)
CLIP-Inspector: Model-Level Backdoor Detection for Prompt-Tuned CLIP via OOD Trigger Inversion
by: Jindal, Akshit, et al.
Published: (2026)
by: Jindal, Akshit, et al.
Published: (2026)
LoopGen: Training-Free Loopable Music Generation
by: Marincione, Davide, et al.
Published: (2025)
by: Marincione, Davide, et al.
Published: (2025)
Color in Visual-Language Models: CLIP deficiencies
by: Arias, Guillem, et al.
Published: (2025)
by: Arias, Guillem, et al.
Published: (2025)
TNG-CLIP:Training-Time Negation Data Generation for Negation Awareness of CLIP
by: Cai, Yuliang, et al.
Published: (2025)
by: Cai, Yuliang, et al.
Published: (2025)
CLIP with Generative Latent Replay: a Strong Baseline for Incremental Learning
by: Frascaroli, Emanuele, et al.
Published: (2024)
by: Frascaroli, Emanuele, et al.
Published: (2024)
CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling
by: Zhang, Jihai, et al.
Published: (2024)
by: Zhang, Jihai, et al.
Published: (2024)
AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP
by: Ma, Wenxin, et al.
Published: (2025)
by: Ma, Wenxin, et al.
Published: (2025)
Unconstrained Open Vocabulary Image Classification: Zero-Shot Transfer from Text to Image via CLIP Inversion
by: Allgeuer, Philipp, et al.
Published: (2024)
by: Allgeuer, Philipp, et al.
Published: (2024)
CLIP-DQA: Blindly Evaluating Dehazed Images from Global and Local Perspectives Using CLIP
by: Zeng, Yirui, et al.
Published: (2025)
by: Zeng, Yirui, et al.
Published: (2025)
FoCLIP: A Feature-Space Misalignment Framework for CLIP-Based Image Manipulation and Detection
by: Chen, Yulin, et al.
Published: (2025)
by: Chen, Yulin, et al.
Published: (2025)
Visually Guided Decoding: Gradient-Free Hard Prompt Inversion with Language Models
by: Kim, Donghoon, et al.
Published: (2025)
by: Kim, Donghoon, et al.
Published: (2025)
Omni-NegCLIP: Enhancing CLIP with Front-Layer Contrastive Fine-Tuning for Comprehensive Negation Understanding
by: Xu, Jingqi
Published: (2026)
by: Xu, Jingqi
Published: (2026)
No Captions, No Problem: Captionless 3D-CLIP Alignment with Hard Negatives via CLIP Knowledge and LLMs
by: Sbrolli, Cristian, et al.
Published: (2024)
by: Sbrolli, Cristian, et al.
Published: (2024)
R3L: Relative Representations for Reinforcement Learning
by: Ricciardi, Antonio Pio, et al.
Published: (2024)
by: Ricciardi, Antonio Pio, et al.
Published: (2024)
microCLIP: Unsupervised CLIP Adaptation via Coarse-Fine Token Fusion for Fine-Grained Image Classification
by: Silva, Sathira, et al.
Published: (2025)
by: Silva, Sathira, et al.
Published: (2025)
Supervised Fine-tuning in turn Improves Visual Foundation Models
by: Jiang, Xiaohu, et al.
Published: (2024)
by: Jiang, Xiaohu, et al.
Published: (2024)
Unleash the Potential of CLIP for Video Highlight Detection
by: Han, Donghoon, et al.
Published: (2024)
by: Han, Donghoon, et al.
Published: (2024)
Occlusion Robustness of CLIP for Military Vehicle Classification
by: van Woerden, Jan Erik, et al.
Published: (2025)
by: van Woerden, Jan Erik, et al.
Published: (2025)
Quantifying and Enabling the Interpretability of CLIP-like Models
by: Madasu, Avinash, et al.
Published: (2024)
by: Madasu, Avinash, et al.
Published: (2024)
Similar Items
-
Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization
by: Maljkovic, Igor, et al.
Published: (2026) -
MASS: MoErging through Adaptive Subspace Selection
by: Crisostomi, Donato, et al.
Published: (2025) -
Environment Maps Editing using Inverse Rendering and Adversarial Implicit Functions
by: D'Orazio, Antonio, et al.
Published: (2024) -
ATM: Improving Model Merging by Alternating Tuning and Merging
by: Zhou, Luca, et al.
Published: (2024) -
A Provable Energy-Guided Test-Time Defense Boosting Adversarial Robustness of Large Vision-Language Models
by: Mirza, Mujtaba Hussain, et al.
Published: (2026)