LLM Probing with Contrastive Eigenproblems: Improving Understanding and Applicability of CCS
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Schouten, Stefan F., Bloem, Peter |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Clustering Internet Memes Through Template Matching and Multi-Dimensional Similarity
par: Bloem, Tygo, et autres
Publié: (2025)
par: Bloem, Tygo, et autres
Publié: (2025)
Truth-value judgment in language models: 'truth directions' are context sensitive
par: Schouten, Stefan F., et autres
Publié: (2024)
par: Schouten, Stefan F., et autres
Publié: (2024)
Context-Enhanced Contrastive Search for Improved LLM Text Generation
par: Sen, Jaydip, et autres
Publié: (2025)
par: Sen, Jaydip, et autres
Publié: (2025)
Looking for the Inner Music: Probing LLMs' Understanding of Literary Style
par: Hicke, Rebecca M. M., et autres
Publié: (2025)
par: Hicke, Rebecca M. M., et autres
Publié: (2025)
Probing the Limits of the Lie Detector Approach to LLM Deception
par: Berger, Tom-Felix
Publié: (2026)
par: Berger, Tom-Felix
Publié: (2026)
Physics-Informed Neural Solvers for Periodic Quantum Eigenproblems
par: Mian, Haaris
Publié: (2025)
par: Mian, Haaris
Publié: (2025)
Embedding And Clustering Your Data Can Improve Contrastive Pretraining
par: Merrick, Luke
Publié: (2024)
par: Merrick, Luke
Publié: (2024)
The Zero Body Problem: Probing LLM Use of Sensory Language
par: Hicke, Rebecca M. M., et autres
Publié: (2025)
par: Hicke, Rebecca M. M., et autres
Publié: (2025)
EffiPair: Improving the Efficiency of LLM-generated Code with Relative Contrastive Feedback
par: Hajizadeh, Samira, et autres
Publié: (2026)
par: Hajizadeh, Samira, et autres
Publié: (2026)
Distillation Contrastive Decoding: Improving LLMs Reasoning with Contrastive Decoding and Distillation
par: Phan, Phuc, et autres
Publié: (2024)
par: Phan, Phuc, et autres
Publié: (2024)
Improving Preference Extraction In LLMs By Identifying Latent Knowledge Through Classifying Probes
par: Maiya, Sharan, et autres
Publié: (2025)
par: Maiya, Sharan, et autres
Publié: (2025)
Can We Use Probing to Better Understand Fine-tuning and Knowledge Distillation of the BERT NLU?
par: Hościłowicz, Jakub, et autres
Publié: (2023)
par: Hościłowicz, Jakub, et autres
Publié: (2023)
More Than a Score: Probing the Impact of Prompt Specificity on LLM Code Generation
par: Zi, Yangtian, et autres
Publié: (2025)
par: Zi, Yangtian, et autres
Publié: (2025)
AvaTaR: Optimizing LLM Agents for Tool Usage via Contrastive Reasoning
par: Wu, Shirley, et autres
Publié: (2024)
par: Wu, Shirley, et autres
Publié: (2024)
Probe and Skip: Self-Predictive Token Skipping for Efficient Long-Context LLM Inference
par: Wu, Zimeng, et autres
Publié: (2026)
par: Wu, Zimeng, et autres
Publié: (2026)
GEM: Empowering LLM for both Embedding Generation and Language Understanding
par: Zhang, Caojin, et autres
Publié: (2025)
par: Zhang, Caojin, et autres
Publié: (2025)
Label Smoothing Improves Gradient Ascent in LLM Unlearning
par: Pang, Zirui, et autres
Publié: (2025)
par: Pang, Zirui, et autres
Publié: (2025)
Improving LLM Final Representations with Inter-Layer Geometry
par: Ulanovski, Tom, et autres
Publié: (2026)
par: Ulanovski, Tom, et autres
Publié: (2026)
The Power of LLM-Generated Synthetic Data for Stance Detection in Online Political Discussions
par: Wagner, Stefan Sylvius, et autres
Publié: (2024)
par: Wagner, Stefan Sylvius, et autres
Publié: (2024)
Enhancing Foundation Models in Transaction Understanding with LLM-based Sentence Embeddings
par: Fan, Xiran, et autres
Publié: (2025)
par: Fan, Xiran, et autres
Publié: (2025)
SciLitLLM: How to Adapt LLMs for Scientific Literature Understanding
par: Li, Sihang, et autres
Publié: (2024)
par: Li, Sihang, et autres
Publié: (2024)
Improving RL Exploration for LLM Reasoning through Retrospective Replay
par: Dou, Shihan, et autres
Publié: (2025)
par: Dou, Shihan, et autres
Publié: (2025)
Understanding LLM Embeddings for Regression
par: Tang, Eric, et autres
Publié: (2024)
par: Tang, Eric, et autres
Publié: (2024)
Non-Linear Inference Time Intervention: Improving LLM Truthfulness
par: Hoscilowicz, Jakub, et autres
Publié: (2024)
par: Hoscilowicz, Jakub, et autres
Publié: (2024)
Are Data Augmentation Methods in Named Entity Recognition Applicable for Uncertainty Estimation?
par: Hashimoto, Wataru, et autres
Publié: (2024)
par: Hashimoto, Wataru, et autres
Publié: (2024)
pyvene: A Library for Understanding and Improving PyTorch Models via Interventions
par: Wu, Zhengxuan, et autres
Publié: (2024)
par: Wu, Zhengxuan, et autres
Publié: (2024)
Uncertainty-Aware Answer Selection for Improved Reasoning in Multi-LLM Systems
par: Agrawal, Aakriti, et autres
Publié: (2025)
par: Agrawal, Aakriti, et autres
Publié: (2025)
ERD: A Framework for Improving LLM Reasoning for Cognitive Distortion Classification
par: Lim, Sehee, et autres
Publié: (2024)
par: Lim, Sehee, et autres
Publié: (2024)
Improving Large Language Model Safety with Contrastive Representation Learning
par: Simko, Samuel, et autres
Publié: (2025)
par: Simko, Samuel, et autres
Publié: (2025)
Lexicon-Level Contrastive Visual-Grounding Improves Language Modeling
par: Zhuang, Chengxu, et autres
Publié: (2024)
par: Zhuang, Chengxu, et autres
Publié: (2024)
Contrastive Learning to Improve Retrieval for Real-world Fact Checking
par: Sriram, Aniruddh, et autres
Publié: (2024)
par: Sriram, Aniruddh, et autres
Publié: (2024)
NoteContrast: Contrastive Language-Diagnostic Pretraining for Medical Text
par: Kailas, Prajwal, et autres
Publié: (2024)
par: Kailas, Prajwal, et autres
Publié: (2024)
Learning from Failures: Understanding LLM Alignment through Failure-Aware Inverse RL
par: Patel, Nyal, et autres
Publié: (2025)
par: Patel, Nyal, et autres
Publié: (2025)
Assessing the Applicability of Natural Language Processing to Traditional Social Science Methodology: A Case Study in Identifying Strategic Signaling Patterns in Presidential Directives
par: LeMay, C., et autres
Publié: (2025)
par: LeMay, C., et autres
Publié: (2025)
Rhetorical Questions in LLM Representations: A Linear Probing Study
par: Yao, Louie Hong, et autres
Publié: (2026)
par: Yao, Louie Hong, et autres
Publié: (2026)
LIME-LLM: Probing Models with Fluent Counterfactuals, Not Broken Text
par: Mihaila, George, et autres
Publié: (2026)
par: Mihaila, George, et autres
Publié: (2026)
Probing the Emergence of Cross-lingual Alignment during LLM Training
par: Wang, Hetong, et autres
Publié: (2024)
par: Wang, Hetong, et autres
Publié: (2024)
Prompt Compression with Context-Aware Sentence Encoding for Fast and Improved LLM Inference
par: Liskavets, Barys, et autres
Publié: (2024)
par: Liskavets, Barys, et autres
Publié: (2024)
LLMSteer: Improving Long-Context LLM Inference by Steering Attention on Reused Contexts
par: Gu, Zhuohan, et autres
Publié: (2024)
par: Gu, Zhuohan, et autres
Publié: (2024)
Talking Nonsense: Probing Large Language Models' Understanding of Adversarial Gibberish Inputs
par: Cherepanova, Valeriia, et autres
Publié: (2024)
par: Cherepanova, Valeriia, et autres
Publié: (2024)
Documents similaires
-
Clustering Internet Memes Through Template Matching and Multi-Dimensional Similarity
par: Bloem, Tygo, et autres
Publié: (2025) -
Truth-value judgment in language models: 'truth directions' are context sensitive
par: Schouten, Stefan F., et autres
Publié: (2024) -
Context-Enhanced Contrastive Search for Improved LLM Text Generation
par: Sen, Jaydip, et autres
Publié: (2025) -
Looking for the Inner Music: Probing LLMs' Understanding of Literary Style
par: Hicke, Rebecca M. M., et autres
Publié: (2025) -
Probing the Limits of the Lie Detector Approach to LLM Deception
par: Berger, Tom-Felix
Publié: (2026)