Mechanistic Interpretability of GPT-2: Lexical and Contextual Layers in Sentiment Analysis
Fuente:
arXiv
Saved in:
| Main Author: | Hatua, Amartya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Financial Sentiment Analysis with Annotators Instruction Assisted Prompting: Enhancing Contextual Interpretation and Stock Prediction Accuracy
by: Rahman, A M Muntasir, et al.
Published: (2025)
by: Rahman, A M Muntasir, et al.
Published: (2025)
Mechanistic Interpretability of GPT-like Models on Summarization Tasks
by: Mishra, Anurag
Published: (2025)
by: Mishra, Anurag
Published: (2025)
Optimization Techniques for Sentiment Analysis Based on LLM (GPT-3)
by: Zhan, Tong, et al.
Published: (2024)
by: Zhan, Tong, et al.
Published: (2024)
ChatGPT vs Gemini vs LLaMA on Multilingual Sentiment Analysis
by: Buscemi, Alessio, et al.
Published: (2024)
by: Buscemi, Alessio, et al.
Published: (2024)
Mechanistic Interpretability Needs Philosophy
by: Williams, Iwan, et al.
Published: (2025)
by: Williams, Iwan, et al.
Published: (2025)
Extracting Lexical Features from Dialects via Interpretable Dialect Classifiers
by: Xie, Roy, et al.
Published: (2024)
by: Xie, Roy, et al.
Published: (2024)
Integration of Explainable AI Techniques with Large Language Models for Enhanced Interpretability for Sentiment Analysis
by: Thogesan, Thivya, et al.
Published: (2025)
by: Thogesan, Thivya, et al.
Published: (2025)
Mechanistic Interpretability as Statistical Estimation: A Variance Analysis
by: Méloux, Maxime, et al.
Published: (2025)
by: Méloux, Maxime, et al.
Published: (2025)
Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study
by: Wang, Zengzhi, et al.
Published: (2023)
by: Wang, Zengzhi, et al.
Published: (2023)
Mechanistic Interpretability of Emotion Inference in Large Language Models
by: Tak, Ala N., et al.
Published: (2025)
by: Tak, Ala N., et al.
Published: (2025)
Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective
by: Chandna, Bhavik, et al.
Published: (2025)
by: Chandna, Bhavik, et al.
Published: (2025)
Machine Unlearning using Forgetting Neural Networks
by: Hatua, Amartya, et al.
Published: (2024)
by: Hatua, Amartya, et al.
Published: (2024)
ELECTRA and GPT-4o: Cost-Effective Partners for Sentiment Analysis
by: Beno, James P.
Published: (2024)
by: Beno, James P.
Published: (2024)
Unlocking Cross-Lingual Sentiment Analysis through Emoji Interpretation: A Multimodal Generative AI Approach
by: Jahan, Rafid Ishrak, et al.
Published: (2024)
by: Jahan, Rafid Ishrak, et al.
Published: (2024)
Toward a Theory of Generalizability in LLM Mechanistic Interpretability Research
by: Trott, Sean
Published: (2025)
by: Trott, Sean
Published: (2025)
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability
by: Raimondi, Bianca, et al.
Published: (2025)
by: Raimondi, Bianca, et al.
Published: (2025)
Enhanced Sentiment Interpretation via a Lexicon-Fuzzy-Transformer Framework
by: Rokhva, Shayan, et al.
Published: (2025)
by: Rokhva, Shayan, et al.
Published: (2025)
Mechanistic Interpretability of Socio-Political Frames in Language Models
by: Asghari, Hadi, et al.
Published: (2025)
by: Asghari, Hadi, et al.
Published: (2025)
MIB: A Mechanistic Interpretability Benchmark
by: Mueller, Aaron, et al.
Published: (2025)
by: Mueller, Aaron, et al.
Published: (2025)
CrashSage: A Large Language Model-Centered Framework for Contextual and Interpretable Traffic Crash Analysis
by: Zhen, Hao, et al.
Published: (2025)
by: Zhen, Hao, et al.
Published: (2025)
Towards Understanding and Improving Refusal in Compressed Models via Mechanistic Interpretability
by: Chhabra, Vishnu Kabir, et al.
Published: (2025)
by: Chhabra, Vishnu Kabir, et al.
Published: (2025)
Interpreting Public Sentiment in Diplomacy Events: A Counterfactual Analysis Framework Using Large Language Models
by: Ouyang, Leyi
Published: (2025)
by: Ouyang, Leyi
Published: (2025)
Generating Effective Ensembles for Sentiment Analysis
by: Etelis, Itay, et al.
Published: (2024)
by: Etelis, Itay, et al.
Published: (2024)
Backtesting Sentiment Signals for Trading: Evaluating the Viability of Alpha Generation from Sentiment Analysis
by: Pontes, Elvys Linhares, et al.
Published: (2025)
by: Pontes, Elvys Linhares, et al.
Published: (2025)
Why LLMs Fail at Causal Discovery and How Interventional Agents Escape
by: Roy, Amartya, et al.
Published: (2026)
by: Roy, Amartya, et al.
Published: (2026)
Unified Lexical Representation for Interpretable Visual-Language Alignment
by: Li, Yifan, et al.
Published: (2024)
by: Li, Yifan, et al.
Published: (2024)
Mechanistic Steering of LLMs Reveals Layer-wise Feature Vulnerabilities in Adversarial Settings
by: Das, Nilanjana, et al.
Published: (2026)
by: Das, Nilanjana, et al.
Published: (2026)
Artificial Intelligence for Sentiment Analysis of Persian Poetry
by: Zargar, Arash, et al.
Published: (2026)
by: Zargar, Arash, et al.
Published: (2026)
RVISA: Reasoning and Verification for Implicit Sentiment Analysis
by: Lai, Wenna, et al.
Published: (2024)
by: Lai, Wenna, et al.
Published: (2024)
The Overlooked Repetitive Lengthening Form in Sentiment Analysis
by: Wang, Lei, et al.
Published: (2026)
by: Wang, Lei, et al.
Published: (2026)
Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy
by: Raimondi, Bianca, et al.
Published: (2026)
by: Raimondi, Bianca, et al.
Published: (2026)
Why Does ChatGPT "Delve" So Much? Exploring the Sources of Lexical Overrepresentation in Large Language Models
by: Juzek, Tom S., et al.
Published: (2024)
by: Juzek, Tom S., et al.
Published: (2024)
HyperDAS: Towards Automating Mechanistic Interpretability with Hypernetworks
by: Sun, Jiuding, et al.
Published: (2025)
by: Sun, Jiuding, et al.
Published: (2025)
Binary Autoencoder for Mechanistic Interpretability of Large Language Models
by: Cho, Hakaze, et al.
Published: (2025)
by: Cho, Hakaze, et al.
Published: (2025)
Everything, Everywhere, All at Once: Is Mechanistic Interpretability Identifiable?
by: Méloux, Maxime, et al.
Published: (2025)
by: Méloux, Maxime, et al.
Published: (2025)
Multi-Modal Sentiment Analysis with Dynamic Attention Fusion
by: Abdulhalim, Sadia, et al.
Published: (2025)
by: Abdulhalim, Sadia, et al.
Published: (2025)
Comparative Analysis of Pooling Mechanisms in LLMs: A Sentiment Analysis Perspective
by: Xing, Jinming, et al.
Published: (2024)
by: Xing, Jinming, et al.
Published: (2024)
Implicit Sentiment Analysis Based on Chain of Thought Prompting
by: Duan, Zhihua, et al.
Published: (2024)
by: Duan, Zhihua, et al.
Published: (2024)
Large Language Model Adaptation for Financial Sentiment Analysis
by: Inserte, Pau Rodriguez, et al.
Published: (2024)
by: Inserte, Pau Rodriguez, et al.
Published: (2024)
Arabic Sentiment Analysis with Noisy Deep Explainable Model
by: Atabuzzaman, Md., et al.
Published: (2023)
by: Atabuzzaman, Md., et al.
Published: (2023)
Similar Items
-
Evaluating Financial Sentiment Analysis with Annotators Instruction Assisted Prompting: Enhancing Contextual Interpretation and Stock Prediction Accuracy
by: Rahman, A M Muntasir, et al.
Published: (2025) -
Mechanistic Interpretability of GPT-like Models on Summarization Tasks
by: Mishra, Anurag
Published: (2025) -
Optimization Techniques for Sentiment Analysis Based on LLM (GPT-3)
by: Zhan, Tong, et al.
Published: (2024) -
ChatGPT vs Gemini vs LLaMA on Multilingual Sentiment Analysis
by: Buscemi, Alessio, et al.
Published: (2024) -
Mechanistic Interpretability Needs Philosophy
by: Williams, Iwan, et al.
Published: (2025)