LUCID-SAE: Learning Unified Vision-Language Sparse Codes for Interpretable Concept Discovery
Fuente:
arXiv
Guardado en:
| Autores principales: | Gu, Difei, Gao, Yunhe, Chatzoudis, Gerasimos, Dong, Zihan, Zhang, Guoning, Guo, Bangwei, Zhou, Yang, Zhou, Mu, Metaxas, Dimitris |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Anatomy-VLM: A Fine-grained Vision-Language Model for Medical Interpretation
por: Gu, Difei, et al.
Publicado: (2025)
por: Gu, Difei, et al.
Publicado: (2025)
RadAlign: Advancing Radiology Report Generation with Vision-Language Concept Alignment
por: Gu, Difei, et al.
Publicado: (2025)
por: Gu, Difei, et al.
Publicado: (2025)
Aligning Human Knowledge with Visual Concepts Towards Explainable Medical Image Classification
por: Gao, Yunhe, et al.
Publicado: (2024)
por: Gao, Yunhe, et al.
Publicado: (2024)
Can Cross-Layer Transcoders Replace Vision Transformer Activations? An Interpretable Perspective on Vision
por: Chatzoudis, Gerasimos, et al.
Publicado: (2026)
por: Chatzoudis, Gerasimos, et al.
Publicado: (2026)
K-Prism: A Knowledge-Guided and Prompt Integrated Universal Medical Image Segmentation Model
por: Guo, Bangwei, et al.
Publicado: (2025)
por: Guo, Bangwei, et al.
Publicado: (2025)
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering
por: Chatzoudis, Gerasimos, et al.
Publicado: (2025)
por: Chatzoudis, Gerasimos, et al.
Publicado: (2025)
Evidence Over Plans: Online Trajectory Verification for Skill Distillation
por: Zhou, Yang, et al.
Publicado: (2026)
por: Zhou, Yang, et al.
Publicado: (2026)
VerSe: Integrating Multiple Queries as Prompts for Versatile Cardiac MRI Segmentation
por: Guo, Bangwei, et al.
Publicado: (2024)
por: Guo, Bangwei, et al.
Publicado: (2024)
M^3-Bench: Multi-Modal, Multi-Hop, Multi-Threaded Tool-Using MLLM Agent Benchmark
por: Zhou, Yang, et al.
Publicado: (2025)
por: Zhou, Yang, et al.
Publicado: (2025)
VL-SAE: Interpreting and Enhancing Vision-Language Alignment with a Unified Concept Set
por: Shen, Shufan, et al.
Publicado: (2025)
por: Shen, Shufan, et al.
Publicado: (2025)
Training Like a Medical Resident: Context-Prior Learning Toward Universal Medical Image Segmentation
por: Gao, Yunhe, et al.
Publicado: (2023)
por: Gao, Yunhe, et al.
Publicado: (2023)
Learning Volumetric Neural Deformable Models to Recover 3D Regional Heart Wall Motion from Multi-Planar Tagged MRI
por: Ye, Meng, et al.
Publicado: (2024)
por: Ye, Meng, et al.
Publicado: (2024)
Beyond Explicit Edges: Robust Reasoning over Noisy and Sparse Knowledge Graphs
por: Gao, Hang, et al.
Publicado: (2026)
por: Gao, Hang, et al.
Publicado: (2026)
Show and Segment: Universal Medical Image Segmentation via In-Context Learning
por: Gao, Yunhe, et al.
Publicado: (2025)
por: Gao, Yunhe, et al.
Publicado: (2025)
Mammo-SAE: Interpreting Breast Cancer Concept Learning with Sparse Autoencoders
por: Nakka, Krishna Kanth
Publicado: (2025)
por: Nakka, Krishna Kanth
Publicado: (2025)
AlignSAE: Concept-Aligned Sparse Autoencoders
por: Yang, Minglai, et al.
Publicado: (2025)
por: Yang, Minglai, et al.
Publicado: (2025)
Athenaxr/LUCID: LUCID v1.0.0
por: Athena
Publicado: (2026)
por: Athena
Publicado: (2026)
Aligning Large Language Models with Healthcare Stakeholders: A Pathway to Trustworthy AI Integration
por: Ding, Kexin, et al.
Publicado: (2025)
por: Ding, Kexin, et al.
Publicado: (2025)
Test-Time Spectrum-Aware Latent Steering for Zero-Shot Generalization in Vision-Language Models
por: Dafnis, Konstantinos M., et al.
Publicado: (2025)
por: Dafnis, Konstantinos M., et al.
Publicado: (2025)
All Circuits Lead to Rome: Rethinking Functional Anisotropy in Circuit and Sheaf Discovery for LLMs
por: Chen, Xi, et al.
Publicado: (2026)
por: Chen, Xi, et al.
Publicado: (2026)
LUCID: Attention with Preconditioned Representations
por: Duvvuri, Sai Surya, et al.
Publicado: (2026)
por: Duvvuri, Sai Surya, et al.
Publicado: (2026)
The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering
por: Li, Zhuowei, et al.
Publicado: (2025)
por: Li, Zhuowei, et al.
Publicado: (2025)
LouvreSAE: Sparse Autoencoders for Interpretable and Controllable Style Transfer
por: Panda, Raina, et al.
Publicado: (2025)
por: Panda, Raina, et al.
Publicado: (2025)
SAE-RNA: A Sparse Autoencoder Model for Interpreting RNA Language Model Representations
por: Kim, Taehan, et al.
Publicado: (2025)
por: Kim, Taehan, et al.
Publicado: (2025)
MedForge: Building Medical Foundation Models Like Open Source Software Development
por: Tan, Zheling, et al.
Publicado: (2025)
por: Tan, Zheling, et al.
Publicado: (2025)
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model
por: Dao, Quan, et al.
Publicado: (2026)
por: Dao, Quan, et al.
Publicado: (2026)
LUCID: An Integrative Approach for Target Discovery and dsRNA Design in Plant Fungal Pathogens
por: Lucía Jiménez‐Castro, et al.
Publicado: (2026)
por: Lucía Jiménez‐Castro, et al.
Publicado: (2026)
Neural Deformable Models for 3D Bi-Ventricular Heart Shape Reconstruction and Modeling from 2D Sparse Cardiac Magnetic Resonance Imaging
por: Ye, Meng, et al.
Publicado: (2023)
por: Ye, Meng, et al.
Publicado: (2023)
ProtSAE: Disentangling and Interpreting Protein Language Models via Semantically-Guided Sparse Autoencoders
por: Liu, Xiangyu, et al.
Publicado: (2025)
por: Liu, Xiangyu, et al.
Publicado: (2025)
A Decomposition Framework for Certifiably Optimal Orthogonal Sparse PCA
por: Cheng, Difei, et al.
Publicado: (2026)
por: Cheng, Difei, et al.
Publicado: (2026)
Your Reward Function for RL is Your Best PRM for Search: Unifying RL and Search-Based TTS
por: Jin, Can, et al.
Publicado: (2025)
por: Jin, Can, et al.
Publicado: (2025)
Conformal Path Reasoning: Trustworthy Knowledge Graph Question Answering via Path-Level Calibration
por: Lin, Shuhang, et al.
Publicado: (2026)
por: Lin, Shuhang, et al.
Publicado: (2026)
LUCID: LLM-Generated Utterances for Complex and Interesting Dialogues
por: Stacey, Joe, et al.
Publicado: (2024)
por: Stacey, Joe, et al.
Publicado: (2024)
Implicit In-context Learning
por: Li, Zhuowei, et al.
Publicado: (2024)
por: Li, Zhuowei, et al.
Publicado: (2024)
Score-Guided Diffusion for 3D Human Recovery
por: Stathopoulos, Anastasis, et al.
Publicado: (2024)
por: Stathopoulos, Anastasis, et al.
Publicado: (2024)
Archetypal SAE: Adaptive and Stable Dictionary Learning for Concept Extraction in Large Vision Models
por: Fel, Thomas, et al.
Publicado: (2025)
por: Fel, Thomas, et al.
Publicado: (2025)
Cross-Layer Discrete Concept Discovery for Interpreting Language Models
por: Garg, Ankur, et al.
Publicado: (2025)
por: Garg, Ankur, et al.
Publicado: (2025)
Sparse Visual Thought Circuits in Vision-Language Models
por: Zhou, Yunpeng
Publicado: (2026)
por: Zhou, Yunpeng
Publicado: (2026)
From Tokens to Concepts: Leveraging SAE for SPLADE
por: Zong, Yuxuan, et al.
Publicado: (2026)
por: Zong, Yuxuan, et al.
Publicado: (2026)
Pooling and Semantic Shift: The Fundamental Challenges in Long Text Embedding and Retrieval
por: Gao, Hang, et al.
Publicado: (2026)
por: Gao, Hang, et al.
Publicado: (2026)
Ejemplares similares
-
Anatomy-VLM: A Fine-grained Vision-Language Model for Medical Interpretation
por: Gu, Difei, et al.
Publicado: (2025) -
RadAlign: Advancing Radiology Report Generation with Vision-Language Concept Alignment
por: Gu, Difei, et al.
Publicado: (2025) -
Aligning Human Knowledge with Visual Concepts Towards Explainable Medical Image Classification
por: Gao, Yunhe, et al.
Publicado: (2024) -
Can Cross-Layer Transcoders Replace Vision Transformer Activations? An Interpretable Perspective on Vision
por: Chatzoudis, Gerasimos, et al.
Publicado: (2026) -
K-Prism: A Knowledge-Guided and Prompt Integrated Universal Medical Image Segmentation Model
por: Guo, Bangwei, et al.
Publicado: (2025)