Beyond the Surface: Probing the Ideological Depth of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kabir, Shariar, Esterling, Kevin, Dong, Yue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PReSS: A Black-Box Framework for Evaluating Political Stance Stability in LLMs via Argumentative Pressure
von: Kabir, Shariar, et al.
Veröffentlicht: (2025)
von: Kabir, Shariar, et al.
Veröffentlicht: (2025)
Probing the Subtle Ideological Manipulation of Large Language Models
von: Paschalides, Demetris, et al.
Veröffentlicht: (2025)
von: Paschalides, Demetris, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Biomedical Data in Bengali Language
von: Kabir, Shariar, et al.
Veröffentlicht: (2024)
von: Kabir, Shariar, et al.
Veröffentlicht: (2024)
Political Ideology Shifts in Large Language Models
von: Bernardelle, Pietro, et al.
Veröffentlicht: (2025)
von: Bernardelle, Pietro, et al.
Veröffentlicht: (2025)
WIBA: What Is Being Argued? A Comprehensive Approach to Argument Mining
von: Irani, Arman, et al.
Veröffentlicht: (2024)
von: Irani, Arman, et al.
Veröffentlicht: (2024)
Geopolitical Parallax: Beyond Walter Lippmann Just After Large Language Models
von: Yavuz, Mehmet Can, et al.
Veröffentlicht: (2025)
von: Yavuz, Mehmet Can, et al.
Veröffentlicht: (2025)
Mapping and Influencing the Political Ideology of Large Language Models using Synthetic Personas
von: Bernardelle, Pietro, et al.
Veröffentlicht: (2024)
von: Bernardelle, Pietro, et al.
Veröffentlicht: (2024)
Large Language Models Reflect the Ideology of their Creators
von: Buyl, Maarten, et al.
Veröffentlicht: (2024)
von: Buyl, Maarten, et al.
Veröffentlicht: (2024)
How Susceptible are Large Language Models to Ideological Manipulation?
von: Chen, Kai, et al.
Veröffentlicht: (2024)
von: Chen, Kai, et al.
Veröffentlicht: (2024)
Probing Neural Topology of Large Language Models
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
Self-Reflection Makes Large Language Models Safer, Less Biased, and Ideologically Neutral
von: Liu, Fengyuan, et al.
Veröffentlicht: (2024)
von: Liu, Fengyuan, et al.
Veröffentlicht: (2024)
Don't Change My View: Ideological Bias Auditing in Large Language Models
von: Kröger, Paul, et al.
Veröffentlicht: (2025)
von: Kröger, Paul, et al.
Veröffentlicht: (2025)
To Words and Beyond: Probing Large Language Models for Sentence-Level Psycholinguistic Norms of Memorability and Reading Times
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
Standard Language Ideology in AI-Generated Language
von: Smith, Genevieve, et al.
Veröffentlicht: (2024)
von: Smith, Genevieve, et al.
Veröffentlicht: (2024)
AmarDoctor: An AI-Driven, Multilingual, Voice-Interactive Digital Health Application for Primary Care Triage and Patient Management to Bridge the Digital Health Divide for Bengali Speakers
von: Nahar, Nazmun, et al.
Veröffentlicht: (2025)
von: Nahar, Nazmun, et al.
Veröffentlicht: (2025)
Decoding the Mind of Large Language Models: A Quantitative Evaluation of Ideology and Biases
von: Hirose, Manari, et al.
Veröffentlicht: (2025)
von: Hirose, Manari, et al.
Veröffentlicht: (2025)
What Affects the Effective Depth of Large Language Models?
von: Hu, Yi, et al.
Veröffentlicht: (2025)
von: Hu, Yi, et al.
Veröffentlicht: (2025)
Beyond Exponential Decay: Rethinking Error Accumulation in Large Language Models
von: Arbuzov, Mikhail L., et al.
Veröffentlicht: (2025)
von: Arbuzov, Mikhail L., et al.
Veröffentlicht: (2025)
ALIGN: Word Association Learning for Cultural Alignment in Large Language Models
von: Liu, Chunhua, et al.
Veröffentlicht: (2025)
von: Liu, Chunhua, et al.
Veröffentlicht: (2025)
Beyond Surface Reasoning: Unveiling the True Long Chain-of-Thought Capacity of Diffusion Large Language Models
von: Chen, Qiguang, et al.
Veröffentlicht: (2025)
von: Chen, Qiguang, et al.
Veröffentlicht: (2025)
Beyond Labels: Aligning Large Language Models with Human-like Reasoning
von: Kabir, Muhammad Rafsan, et al.
Veröffentlicht: (2024)
von: Kabir, Muhammad Rafsan, et al.
Veröffentlicht: (2024)
Speculative Decoding and Beyond: An In-Depth Survey of Techniques
von: Hu, Yunhai, et al.
Veröffentlicht: (2025)
von: Hu, Yunhai, et al.
Veröffentlicht: (2025)
A Detailed Factor Analysis for the Political Compass Test: Navigating Ideologies of Large Language Models
von: Kamal, Sadia, et al.
Veröffentlicht: (2025)
von: Kamal, Sadia, et al.
Veröffentlicht: (2025)
Reliability Under Randomness: An Empirical Analysis of Sparse and Dense Language Models Across Decoding Temperatures
von: Grover, Kabir
Veröffentlicht: (2026)
von: Grover, Kabir
Veröffentlicht: (2026)
Cross-Lingual Pitfalls: Automatic Probing Cross-Lingual Weakness of Multilingual Large Language Models
von: Xu, Zixiang, et al.
Veröffentlicht: (2025)
von: Xu, Zixiang, et al.
Veröffentlicht: (2025)
Probing then Editing Response Personality of Large Language Models
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
Beyond Fixed: Training-Free Variable-Length Denoising for Diffusion Large Language Models
von: Li, Jinsong, et al.
Veröffentlicht: (2025)
von: Li, Jinsong, et al.
Veröffentlicht: (2025)
From Form(s) to Meaning: Probing the Semantic Depths of Language Models Using Multisense Consistency
von: Ohmer, Xenia, et al.
Veröffentlicht: (2024)
von: Ohmer, Xenia, et al.
Veröffentlicht: (2024)
Probing Syntax in Large Language Models: Successes and Remaining Challenges
von: Diego-Simón, Pablo J., et al.
Veröffentlicht: (2025)
von: Diego-Simón, Pablo J., et al.
Veröffentlicht: (2025)
Prompt-based Depth Pruning of Large Language Models
von: Wee, Juyun, et al.
Veröffentlicht: (2025)
von: Wee, Juyun, et al.
Veröffentlicht: (2025)
Probing Causality Manipulation of Large Language Models
von: Zhang, Chenyang, et al.
Veröffentlicht: (2024)
von: Zhang, Chenyang, et al.
Veröffentlicht: (2024)
Beyond Decodability: Reconstructing Language Model Representations with an Encoding Probe
von: Shen, Gaofei, et al.
Veröffentlicht: (2026)
von: Shen, Gaofei, et al.
Veröffentlicht: (2026)
GPT-Fathom: Benchmarking Large Language Models to Decipher the Evolutionary Path towards GPT-4 and Beyond
von: Zheng, Shen, et al.
Veröffentlicht: (2023)
von: Zheng, Shen, et al.
Veröffentlicht: (2023)
Beneath the Surface: Investigating LLMs' Capabilities for Communicating with Subtext
von: Ahuja, Kabir, et al.
Veröffentlicht: (2026)
von: Ahuja, Kabir, et al.
Veröffentlicht: (2026)
A Survey of Efficient Reasoning for Large Reasoning Models: Language, Multimodality, and Beyond
von: Qu, Xiaoye, et al.
Veröffentlicht: (2025)
von: Qu, Xiaoye, et al.
Veröffentlicht: (2025)
UCoder: Unsupervised Code Generation by Internal Probing of Large Language Models
von: Wu, Jiajun, et al.
Veröffentlicht: (2025)
von: Wu, Jiajun, et al.
Veröffentlicht: (2025)
Probing Large Language Models in Reasoning and Translating Complex Linguistic Puzzles
von: Lin, Zheng-Lin, et al.
Veröffentlicht: (2025)
von: Lin, Zheng-Lin, et al.
Veröffentlicht: (2025)
Probing Internal Representations of Multi-Word Verbs in Large Language Models
von: Kissane, Hassane, et al.
Veröffentlicht: (2025)
von: Kissane, Hassane, et al.
Veröffentlicht: (2025)
Probing Large Language Models from A Human Behavioral Perspective
von: Wang, Xintong, et al.
Veröffentlicht: (2023)
von: Wang, Xintong, et al.
Veröffentlicht: (2023)
Bias Beyond Borders: Political Ideology Evaluation and Steering in Multilingual LLMs
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2026)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PReSS: A Black-Box Framework for Evaluating Political Stance Stability in LLMs via Argumentative Pressure
von: Kabir, Shariar, et al.
Veröffentlicht: (2025) -
Probing the Subtle Ideological Manipulation of Large Language Models
von: Paschalides, Demetris, et al.
Veröffentlicht: (2025) -
Automatic Speech Recognition for Biomedical Data in Bengali Language
von: Kabir, Shariar, et al.
Veröffentlicht: (2024) -
Political Ideology Shifts in Large Language Models
von: Bernardelle, Pietro, et al.
Veröffentlicht: (2025) -
WIBA: What Is Being Argued? A Comprehensive Approach to Argument Mining
von: Irani, Arman, et al.
Veröffentlicht: (2024)