Are LLMs Models of Distributional Semantics? A Case Study on Quantifiers
Fuente:
arXiv
Salvato in:
| Autori principali: | Enyan, Zhang, Wang, Zewei, Lepori, Michael A., Pavlick, Ellie, Aparicio, Helena |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Uncovering Intermediate Variables in Transformers using Circuit Probing
di: Lepori, Michael A., et al.
Pubblicazione: (2023)
di: Lepori, Michael A., et al.
Pubblicazione: (2023)
Instilling Inductive Biases with Subnetworks
di: Zhang, Enyan, et al.
Pubblicazione: (2023)
di: Zhang, Enyan, et al.
Pubblicazione: (2023)
Dual Process Learning: Controlling Use of In-Context vs. In-Weights Strategies with Weight Forgetting
di: Anand, Suraj, et al.
Pubblicazione: (2024)
di: Anand, Suraj, et al.
Pubblicazione: (2024)
LLMs as Models for Analogical Reasoning
di: Musker, Sam, et al.
Pubblicazione: (2024)
di: Musker, Sam, et al.
Pubblicazione: (2024)
Does Training on Synthetic Data Make Models Less Robust?
di: Zhang, Lingze, et al.
Pubblicazione: (2025)
di: Zhang, Lingze, et al.
Pubblicazione: (2025)
How Do Language Models Compose Functions?
di: Khandelwal, Apoorv, et al.
Pubblicazione: (2025)
di: Khandelwal, Apoorv, et al.
Pubblicazione: (2025)
Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility
di: Lepori, Michael A., et al.
Pubblicazione: (2025)
di: Lepori, Michael A., et al.
Pubblicazione: (2025)
Source-Modality Monitoring in Vision-Language Models
di: Hua, Etha Tianze, et al.
Pubblicazione: (2026)
di: Hua, Etha Tianze, et al.
Pubblicazione: (2026)
Circuit Component Reuse Across Tasks in Transformer Language Models
di: Merullo, Jack, et al.
Pubblicazione: (2023)
di: Merullo, Jack, et al.
Pubblicazione: (2023)
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models
di: Merullo, Jack, et al.
Pubblicazione: (2024)
di: Merullo, Jack, et al.
Pubblicazione: (2024)
What is an "Abstract Reasoner"? Revisiting Experiments and Arguments about Large Language Models
di: Yun, Tian, et al.
Pubblicazione: (2025)
di: Yun, Tian, et al.
Pubblicazione: (2025)
Language Models Implement Simple Word2Vec-style Vector Arithmetic
di: Merullo, Jack, et al.
Pubblicazione: (2023)
di: Merullo, Jack, et al.
Pubblicazione: (2023)
mOthello: When Do Cross-Lingual Representation Alignment and Cross-Lingual Transfer Emerge in Multilingual Models?
di: Hua, Tianze, et al.
Pubblicazione: (2024)
di: Hua, Tianze, et al.
Pubblicazione: (2024)
How Do Vision-Language Models Process Conflicting Information Across Modalities?
di: Hua, Tianze, et al.
Pubblicazione: (2025)
di: Hua, Tianze, et al.
Pubblicazione: (2025)
Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline
di: Lu, Meng, et al.
Pubblicazione: (2025)
di: Lu, Meng, et al.
Pubblicazione: (2025)
Transferring Linear Features Across Language Models With Model Stitching
di: Chen, Alan, et al.
Pubblicazione: (2025)
di: Chen, Alan, et al.
Pubblicazione: (2025)
A Knapsack by Any Other Name: Presentation impacts LLM performance on NP-hard problems
di: Duchnowski, Alex, et al.
Pubblicazione: (2025)
di: Duchnowski, Alex, et al.
Pubblicazione: (2025)
The Same But Different: Structural Similarities and Differences in Multilingual Language Modeling
di: Zhang, Ruochen, et al.
Pubblicazione: (2024)
di: Zhang, Ruochen, et al.
Pubblicazione: (2024)
Bayesian Preference Elicitation with Language Models
di: Handa, Kunal, et al.
Pubblicazione: (2024)
di: Handa, Kunal, et al.
Pubblicazione: (2024)
Can LLMs subtract numbers?
di: Jobanputra, Mayank, et al.
Pubblicazione: (2025)
di: Jobanputra, Mayank, et al.
Pubblicazione: (2025)
Born a Transformer -- Always a Transformer? On the Effect of Pretraining on Architectural Abilities
di: Jobanputra, Mayank, et al.
Pubblicazione: (2025)
di: Jobanputra, Mayank, et al.
Pubblicazione: (2025)
Quantifying the Reasoning Abilities of LLMs on Real-world Clinical Cases
di: Qiu, Pengcheng, et al.
Pubblicazione: (2025)
di: Qiu, Pengcheng, et al.
Pubblicazione: (2025)
Shared Lexical Task Representations Explain Behavioral Variability In LLMs
di: Yang, Zhuonan, et al.
Pubblicazione: (2026)
di: Yang, Zhuonan, et al.
Pubblicazione: (2026)
Semantic Volume: Quantifying and Detecting both External and Internal Uncertainty in LLMs
di: Li, Xiaomin, et al.
Pubblicazione: (2025)
di: Li, Xiaomin, et al.
Pubblicazione: (2025)
Potential and Limitations of LLMs in Capturing Structured Semantics: A Case Study on SRL
di: Cheng, Ning, et al.
Pubblicazione: (2024)
di: Cheng, Ning, et al.
Pubblicazione: (2024)
DuFFin: A Dual-Level Fingerprinting Framework for LLMs IP Protection
di: Yan, Yuliang, et al.
Pubblicazione: (2025)
di: Yan, Yuliang, et al.
Pubblicazione: (2025)
Racing Thoughts: Explaining Contextualization Errors in Large Language Models
di: Lepori, Michael A., et al.
Pubblicazione: (2024)
di: Lepori, Michael A., et al.
Pubblicazione: (2024)
Quantifying Semantic Emergence in Language Models
di: Chen, Hang, et al.
Pubblicazione: (2024)
di: Chen, Hang, et al.
Pubblicazione: (2024)
Does CLIP Bind Concepts? Probing Compositionality in Large Image Models
di: Lewis, Martha, et al.
Pubblicazione: (2022)
di: Lewis, Martha, et al.
Pubblicazione: (2022)
Cognitive Modeling with Scaffolded LLMs: A Case Study of Referential Expression Generation
di: Tsvilodub, Polina, et al.
Pubblicazione: (2024)
di: Tsvilodub, Polina, et al.
Pubblicazione: (2024)
Acting Flatterers via LLMs Sycophancy: Combating Clickbait with LLMs Opposing-Stance Reasoning
di: Zhang, Chaowei, et al.
Pubblicazione: (2026)
di: Zhang, Chaowei, et al.
Pubblicazione: (2026)
Is It JUST Semantics? A Case Study of Discourse Particle Understanding in LLMs
di: Sheffield, William, et al.
Pubblicazione: (2025)
di: Sheffield, William, et al.
Pubblicazione: (2025)
HawkesLLM: Semantic Uncertainty Propagation in Agentic Text Simulation
di: Deng, Zewei, et al.
Pubblicazione: (2026)
di: Deng, Zewei, et al.
Pubblicazione: (2026)
Quantifying Fairness in LLMs Beyond Tokens: A Semantic and Statistical Perspective
di: Xu, Weijie, et al.
Pubblicazione: (2025)
di: Xu, Weijie, et al.
Pubblicazione: (2025)
$100K or 100 Days: Trade-offs when Pre-Training with Academic Resources
di: Khandelwal, Apoorv, et al.
Pubblicazione: (2024)
di: Khandelwal, Apoorv, et al.
Pubblicazione: (2024)
A Semantic-based Optimization Approach for Repairing LLMs: Case Study on Code Generation
di: Gu, Jian, et al.
Pubblicazione: (2025)
di: Gu, Jian, et al.
Pubblicazione: (2025)
Signatures of human-like processing in Transformer forward passes
di: Hu, Jennifer, et al.
Pubblicazione: (2025)
di: Hu, Jennifer, et al.
Pubblicazione: (2025)
Language Models Struggle to Use Representations Learned In-Context
di: Lepori, Michael A., et al.
Pubblicazione: (2026)
di: Lepori, Michael A., et al.
Pubblicazione: (2026)
SHIELD: Semantic Heterogeneity Integrated Embedding for Latent Discovery in Clinical Trial Safety Signals
di: Vandenhende, Francois, et al.
Pubblicazione: (2026)
di: Vandenhende, Francois, et al.
Pubblicazione: (2026)
How LLMs Comprehend Temporal Meaning in Narratives: A Case Study in Cognitive Evaluation of LLMs
di: de Langis, Karin, et al.
Pubblicazione: (2025)
di: de Langis, Karin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Uncovering Intermediate Variables in Transformers using Circuit Probing
di: Lepori, Michael A., et al.
Pubblicazione: (2023) -
Instilling Inductive Biases with Subnetworks
di: Zhang, Enyan, et al.
Pubblicazione: (2023) -
Dual Process Learning: Controlling Use of In-Context vs. In-Weights Strategies with Weight Forgetting
di: Anand, Suraj, et al.
Pubblicazione: (2024) -
LLMs as Models for Analogical Reasoning
di: Musker, Sam, et al.
Pubblicazione: (2024) -
Does Training on Synthetic Data Make Models Less Robust?
di: Zhang, Lingze, et al.
Pubblicazione: (2025)