Analyzing the Role of Semantic Representations in the Era of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jin, Zhijing, Chen, Yuen, Gonzalez, Fernando, Liu, Jiarui, Zhang, Jiayi, Michael, Julian, Schölkopf, Bernhard, Diab, Mona |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can Large Language Models Infer Causation from Correlation?
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
Improving Large Language Model Safety with Contrastive Representation Learning
von: Simko, Samuel, et al.
Veröffentlicht: (2025)
von: Simko, Samuel, et al.
Veröffentlicht: (2025)
Automatic Generation of Model and Data Cards: A Step Towards Responsible AI
von: Liu, Jiarui, et al.
Veröffentlicht: (2024)
von: Liu, Jiarui, et al.
Veröffentlicht: (2024)
Taming Object Hallucinations with Verified Atomic Confidence Estimation
von: Liu, Jiarui, et al.
Veröffentlicht: (2025)
von: Liu, Jiarui, et al.
Veröffentlicht: (2025)
Implicit Personalization in Language Models: A Systematic Study
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)
CLadder: Assessing Causal Reasoning in Language Models
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
Corrupted by Reasoning: Reasoning Language Models Become Free-Riders in Public Goods Games
von: Piedrahita, David Guzman, et al.
Veröffentlicht: (2025)
von: Piedrahita, David Guzman, et al.
Veröffentlicht: (2025)
Voices of Her: Analyzing Gender Differences in the AI Publication World
von: Ding, Yiwen, et al.
Veröffentlicht: (2023)
von: Ding, Yiwen, et al.
Veröffentlicht: (2023)
Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis
von: Jenny, David F., et al.
Veröffentlicht: (2023)
von: Jenny, David F., et al.
Veröffentlicht: (2023)
CausalCite: A Causal Formulation of Paper Citations
von: Kumar, Ishan, et al.
Veröffentlicht: (2023)
von: Kumar, Ishan, et al.
Veröffentlicht: (2023)
Democratic or Authoritarian? Probing a New Dimension of Political Biases in Large Language Models
von: Piedrahita, David Guzman, et al.
Veröffentlicht: (2025)
von: Piedrahita, David Guzman, et al.
Veröffentlicht: (2025)
Quriosity: Analyzing Human Questioning Behavior and Causal Inquiry through Curiosity-Driven Queries
von: Ceraolo, Roberto, et al.
Veröffentlicht: (2024)
von: Ceraolo, Roberto, et al.
Veröffentlicht: (2024)
Personal Information Parroting in Language Models
von: Subramani, Nishant, et al.
Veröffentlicht: (2026)
von: Subramani, Nishant, et al.
Veröffentlicht: (2026)
Towards Global AI Inclusivity: A Large-Scale Multilingual Terminology Dataset (GIST)
von: Liu, Jiarui, et al.
Veröffentlicht: (2024)
von: Liu, Jiarui, et al.
Veröffentlicht: (2024)
Can Theoretical Physics Research Benefit from Language Agents?
von: Lu, Sirui, et al.
Veröffentlicht: (2025)
von: Lu, Sirui, et al.
Veröffentlicht: (2025)
How Robust Are Router-LLMs? Analysis of the Fragility of LLM Routing Capabilities
von: Kassem, Aly M., et al.
Veröffentlicht: (2025)
von: Kassem, Aly M., et al.
Veröffentlicht: (2025)
Tracing Multilingual Representations in LLMs with Cross-Layer Transcoders
von: Harrasse, Abir, et al.
Veröffentlicht: (2025)
von: Harrasse, Abir, et al.
Veröffentlicht: (2025)
When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
von: Backmann, Steffen, et al.
Veröffentlicht: (2025)
von: Backmann, Steffen, et al.
Veröffentlicht: (2025)
Evaluating Large Language Model Biases in Persona-Steered Generation
von: Liu, Andy, et al.
Veröffentlicht: (2024)
von: Liu, Andy, et al.
Veröffentlicht: (2024)
Emotion Classification in Low and Moderate Resource Languages
von: Tafreshi, Shabnam, et al.
Veröffentlicht: (2024)
von: Tafreshi, Shabnam, et al.
Veröffentlicht: (2024)
Causality can systematically address the monsters under the bench(marks)
von: Leeb, Felix, et al.
Veröffentlicht: (2025)
von: Leeb, Felix, et al.
Veröffentlicht: (2025)
The Odyssey of Commonsense Causality: From Foundational Benchmarks to Cutting-Edge Reasoning
von: Cui, Shaobo, et al.
Veröffentlicht: (2024)
von: Cui, Shaobo, et al.
Veröffentlicht: (2024)
Causal Responsibility Attribution for Human-AI Collaboration
von: Qi, Yahang, et al.
Veröffentlicht: (2024)
von: Qi, Yahang, et al.
Veröffentlicht: (2024)
Competition of Mechanisms: Tracing How Language Models Handle Facts and Counterfactuals
von: Ortu, Francesco, et al.
Veröffentlicht: (2024)
von: Ortu, Francesco, et al.
Veröffentlicht: (2024)
Causality for Natural Language Processing
von: Jin, Zhijing
Veröffentlicht: (2025)
von: Jin, Zhijing
Veröffentlicht: (2025)
Do LLMs Think Fast and Slow? A Causal Study on Sentiment Analysis
von: Lyu, Zhiheng, et al.
Veröffentlicht: (2024)
von: Lyu, Zhiheng, et al.
Veröffentlicht: (2024)
LLM Microscope: What Model Internals Reveal About Answer Correctness and Context Utilization
von: Liu, Jiarui, et al.
Veröffentlicht: (2025)
von: Liu, Jiarui, et al.
Veröffentlicht: (2025)
Decoding Dark Matter: Specialized Sparse Autoencoders for Interpreting Rare Concepts in Foundation Models
von: Muhamed, Aashiq, et al.
Veröffentlicht: (2024)
von: Muhamed, Aashiq, et al.
Veröffentlicht: (2024)
Language Model Alignment in Multilingual Trolley Problems
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)
MixSD: Mixed Contextual Self-Distillation for Knowledge Injection
von: Liu, Jiarui, et al.
Veröffentlicht: (2026)
von: Liu, Jiarui, et al.
Veröffentlicht: (2026)
CORE: Measuring Multi-Agent LLM Interaction Quality under Game-Theoretic Pressures
von: Pandey, Punya Syon, et al.
Veröffentlicht: (2025)
von: Pandey, Punya Syon, et al.
Veröffentlicht: (2025)
StressRoBERTa: Cross-Condition Transfer Learning from Depression, Anxiety, and PTSD to Stress Detection
von: Alqahtani, Amal, et al.
Veröffentlicht: (2025)
von: Alqahtani, Amal, et al.
Veröffentlicht: (2025)
Are Language Models Consequentialist or Deontological Moral Reasoners?
von: Samway, Keenan, et al.
Veröffentlicht: (2025)
von: Samway, Keenan, et al.
Veröffentlicht: (2025)
Are LLMs Good Safety Agents or a Propaganda Engine?
von: Yadav, Neemesh, et al.
Veröffentlicht: (2025)
von: Yadav, Neemesh, et al.
Veröffentlicht: (2025)
Probing Multimodal Large Language Models for Global and Local Semantic Representations
von: Tao, Mingxu, et al.
Veröffentlicht: (2024)
von: Tao, Mingxu, et al.
Veröffentlicht: (2024)
Limits of Transformer Language Models on Learning to Compose Algorithms
von: Thomm, Jonathan, et al.
Veröffentlicht: (2024)
von: Thomm, Jonathan, et al.
Veröffentlicht: (2024)
Emergent Representations of Program Semantics in Language Models Trained on Programs
von: Jin, Charles, et al.
Veröffentlicht: (2023)
von: Jin, Charles, et al.
Veröffentlicht: (2023)
SimBA: Simplifying Benchmark Analysis Using Performance Matrices Alone
von: Subramani, Nishant, et al.
Veröffentlicht: (2025)
von: Subramani, Nishant, et al.
Veröffentlicht: (2025)
Whispers that Shake Foundations: Analyzing and Mitigating False Premise Hallucinations in Large Language Models
von: Yuan, Hongbang, et al.
Veröffentlicht: (2024)
von: Yuan, Hongbang, et al.
Veröffentlicht: (2024)
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models
von: Choi, Younwoo, et al.
Veröffentlicht: (2025)
von: Choi, Younwoo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Can Large Language Models Infer Causation from Correlation?
von: Jin, Zhijing, et al.
Veröffentlicht: (2023) -
Improving Large Language Model Safety with Contrastive Representation Learning
von: Simko, Samuel, et al.
Veröffentlicht: (2025) -
Automatic Generation of Model and Data Cards: A Step Towards Responsible AI
von: Liu, Jiarui, et al.
Veröffentlicht: (2024) -
Taming Object Hallucinations with Verified Atomic Confidence Estimation
von: Liu, Jiarui, et al.
Veröffentlicht: (2025) -
Implicit Personalization in Language Models: A Systematic Study
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)