Tracing Representation Progression: Analyzing and Enhancing Layer-Wise Similarity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Jiachen, Zhou, Jinxin, Zhu, Zhihui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Analyzing Fine-Grained Alignment and Enhancing Vision Understanding in Multimodal Language Models
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)
AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers
von: Achtibat, Reduan, et al.
Veröffentlicht: (2024)
von: Achtibat, Reduan, et al.
Veröffentlicht: (2024)
Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models
von: Mersha, Melkamu Abay, et al.
Veröffentlicht: (2026)
von: Mersha, Melkamu Abay, et al.
Veröffentlicht: (2026)
ECoFLaP: Efficient Coarse-to-Fine Layer-Wise Pruning for Vision-Language Models
von: Sung, Yi-Lin, et al.
Veröffentlicht: (2023)
von: Sung, Yi-Lin, et al.
Veröffentlicht: (2023)
From Compression to Expression: A Layerwise Analysis of In-Context Learning
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)
MAPS: Preserving Vision-Language Representations via Module-Wise Proximity Scheduling for Better Vision-Language-Action Generalization
von: Huang, Chengyue, et al.
Veröffentlicht: (2025)
von: Huang, Chengyue, et al.
Veröffentlicht: (2025)
Stronger Normalization-Free Transformers
von: Chen, Mingzhi, et al.
Veröffentlicht: (2025)
von: Chen, Mingzhi, et al.
Veröffentlicht: (2025)
Transformers without Normalization
von: Zhu, Jiachen, et al.
Veröffentlicht: (2025)
von: Zhu, Jiachen, et al.
Veröffentlicht: (2025)
Randomness of Low-Layer Parameters Determines Confusing Samples in Terms of Interaction Representations of a DNN
von: Zhang, Junpeng, et al.
Veröffentlicht: (2025)
von: Zhang, Junpeng, et al.
Veröffentlicht: (2025)
Analyzing The Language of Visual Tokens
von: Chan, David M., et al.
Veröffentlicht: (2024)
von: Chan, David M., et al.
Veröffentlicht: (2024)
Cat-AIR: Content and Task-Aware All-in-One Image Restoration
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)
Do Vision and Language Encoders Represent the World Similarly?
von: Maniparambil, Mayug, et al.
Veröffentlicht: (2024)
von: Maniparambil, Mayug, et al.
Veröffentlicht: (2024)
Enhancing Financial VQA in Vision Language Models using Intermediate Structured Representations
von: Srivastava, Archita, et al.
Veröffentlicht: (2025)
von: Srivastava, Archita, et al.
Veröffentlicht: (2025)
Analyzing the Roles of Language and Vision in Learning from Limited Data
von: Chen, Allison, et al.
Veröffentlicht: (2024)
von: Chen, Allison, et al.
Veröffentlicht: (2024)
Technical Report: Quantifying and Analyzing the Generalization Power of a DNN
von: He, Yuxuan, et al.
Veröffentlicht: (2025)
von: He, Yuxuan, et al.
Veröffentlicht: (2025)
ClassWise-CRF: Category-Specific Fusion for Enhanced Semantic Segmentation of Remote Sensing Imagery
von: Zhu, Qinfeng, et al.
Veröffentlicht: (2025)
von: Zhu, Qinfeng, et al.
Veröffentlicht: (2025)
Contrastive-to-Self-Supervised: A Two-Stage Framework for Script Similarity Learning
von: Roman, Claire, et al.
Veröffentlicht: (2026)
von: Roman, Claire, et al.
Veröffentlicht: (2026)
Analyzing and Boosting the Power of Fine-Grained Visual Recognition for Multi-modal Large Language Models
von: He, Hulingxiao, et al.
Veröffentlicht: (2025)
von: He, Hulingxiao, et al.
Veröffentlicht: (2025)
Agent S2: A Compositional Generalist-Specialist Framework for Computer Use Agents
von: Agashe, Saaket, et al.
Veröffentlicht: (2025)
von: Agashe, Saaket, et al.
Veröffentlicht: (2025)
Impact of Layer Norm on Memorization and Generalization in Transformers
von: Singhal, Rishi, et al.
Veröffentlicht: (2025)
von: Singhal, Rishi, et al.
Veröffentlicht: (2025)
Rerouting Connection: Hybrid Computer Vision Analysis Reveals Visual Similarity Between Indus and Tibetan-Yi Corridor Writing Systems
von: Reddy, Ooha Lakkadi
Veröffentlicht: (2025)
von: Reddy, Ooha Lakkadi
Veröffentlicht: (2025)
Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
Scaling Agents for Computer Use
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2025)
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2025)
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2025)
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2025)
Frozen Transformers in Language Models Are Effective Visual Encoder Layers
von: Pang, Ziqi, et al.
Veröffentlicht: (2023)
von: Pang, Ziqi, et al.
Veröffentlicht: (2023)
Variance-Covariance Regularization Improves Representation Learning
von: Zhu, Jiachen, et al.
Veröffentlicht: (2023)
von: Zhu, Jiachen, et al.
Veröffentlicht: (2023)
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation
von: Wang, Yongxin, et al.
Veröffentlicht: (2024)
von: Wang, Yongxin, et al.
Veröffentlicht: (2024)
DREAM: Diffusion Rectification and Estimation-Adaptive Models
von: Zhou, Jinxin, et al.
Veröffentlicht: (2023)
von: Zhou, Jinxin, et al.
Veröffentlicht: (2023)
MLLMs-Augmented Visual-Language Representation Learning
von: Liu, Yanqing, et al.
Veröffentlicht: (2023)
von: Liu, Yanqing, et al.
Veröffentlicht: (2023)
ADMN: A Layer-Wise Adaptive Multimodal Network for Dynamic Input Noise and Compute Resources
von: Wu, Jason, et al.
Veröffentlicht: (2025)
von: Wu, Jason, et al.
Veröffentlicht: (2025)
Unified Lexical Representation for Interpretable Visual-Language Alignment
von: Li, Yifan, et al.
Veröffentlicht: (2024)
von: Li, Yifan, et al.
Veröffentlicht: (2024)
OSCaR: Object State Captioning and State Change Representation
von: Nguyen, Nguyen, et al.
Veröffentlicht: (2024)
von: Nguyen, Nguyen, et al.
Veröffentlicht: (2024)
Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs
von: Mazeika, Mantas, et al.
Veröffentlicht: (2025)
von: Mazeika, Mantas, et al.
Veröffentlicht: (2025)
PLD: A Choice-Theoretic List-Wise Knowledge Distillation
von: Bassam, Ejafa, et al.
Veröffentlicht: (2025)
von: Bassam, Ejafa, et al.
Veröffentlicht: (2025)
Winsor-CAM: Human-Tunable Visual Explanations from Deep Networks via Layer-Wise Winsorization
von: Wall, Casey, et al.
Veröffentlicht: (2025)
von: Wall, Casey, et al.
Veröffentlicht: (2025)
Contrasting with Symile: Simple Model-Agnostic Representation Learning for Unlimited Modalities
von: Saporta, Adriel, et al.
Veröffentlicht: (2024)
von: Saporta, Adriel, et al.
Veröffentlicht: (2024)
RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2024)
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2024)
GeoSAE: Geometric Prior-Guided Layer-Wise Sparse Autoencoder Annotation of Brain MRI Foundation Models
von: Nerrise, Favour, et al.
Veröffentlicht: (2026)
von: Nerrise, Favour, et al.
Veröffentlicht: (2026)
Preserving Pre-trained Representation Space: On Effectiveness of Prefix-tuning for Large Multi-modal Models
von: Kim, Donghoon, et al.
Veröffentlicht: (2024)
von: Kim, Donghoon, et al.
Veröffentlicht: (2024)
Robust Representation Consistency Model via Contrastive Denoising
von: Lei, Jiachen, et al.
Veröffentlicht: (2025)
von: Lei, Jiachen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Analyzing Fine-Grained Alignment and Enhancing Vision Understanding in Multimodal Language Models
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025) -
AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers
von: Achtibat, Reduan, et al.
Veröffentlicht: (2024) -
Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models
von: Mersha, Melkamu Abay, et al.
Veröffentlicht: (2026) -
ECoFLaP: Efficient Coarse-to-Fine Layer-Wise Pruning for Vision-Language Models
von: Sung, Yi-Lin, et al.
Veröffentlicht: (2023) -
From Compression to Expression: A Layerwise Analysis of In-Context Learning
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)