Do Pre-Trained Language Models Detect and Understand Semantic Underspecification? Ask the DUST!
Fuente:
arXiv
Salvato in:
| Autori principali: | Wildenburg, Frank, Hanna, Michael, Pezzelle, Sandro |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Is my model perplexed for the right reason? Contrasting LLMs' Benchmark Behavior with Token-Level Perplexity
di: Prins, Zoë, et al.
Pubblicazione: (2026)
di: Prins, Zoë, et al.
Pubblicazione: (2026)
The BLA Benchmark: Investigating Basic Language Abilities of Pre-Trained Multimodal Models
di: Chen, Xinyi, et al.
Pubblicazione: (2023)
di: Chen, Xinyi, et al.
Pubblicazione: (2023)
Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms
di: Hanna, Michael, et al.
Pubblicazione: (2024)
di: Hanna, Michael, et al.
Pubblicazione: (2024)
Are formal and functional linguistic mechanisms dissociated in language models?
di: Hanna, Michael, et al.
Pubblicazione: (2025)
di: Hanna, Michael, et al.
Pubblicazione: (2025)
Beyond Divergent Creativity: A Human-Based Evaluation of Creativity in Large Language Models
di: Nakajima, Kumiko, et al.
Pubblicazione: (2026)
di: Nakajima, Kumiko, et al.
Pubblicazione: (2026)
How Language Models Conflate Logical Validity with Plausibility: A Representational Analysis of Content Effects
di: Bertolazzi, Leonardo, et al.
Pubblicazione: (2025)
di: Bertolazzi, Leonardo, et al.
Pubblicazione: (2025)
Who is the richest club in the championship? Detecting and Rewriting Underspecified Questions Improve QA Performance
di: Huang, Yunchong, et al.
Pubblicazione: (2026)
di: Huang, Yunchong, et al.
Pubblicazione: (2026)
Vision-Language Models Align with Human Neural Representations in Concept Processing
di: Bavaresco, Anna, et al.
Pubblicazione: (2024)
di: Bavaresco, Anna, et al.
Pubblicazione: (2024)
Naming, Describing, and Quantifying Visual Objects in Humans and LLMs
di: Testoni, Alberto, et al.
Pubblicazione: (2024)
di: Testoni, Alberto, et al.
Pubblicazione: (2024)
They want to pretend not to understand: The Limits of Current LLMs in Interpreting Implicit Content of Political Discourse
di: Paci, Walter, et al.
Pubblicazione: (2025)
di: Paci, Walter, et al.
Pubblicazione: (2025)
What Prompts Don't Say: Understanding and Managing Underspecification in LLM Prompts
di: Yang, Chenyang, et al.
Pubblicazione: (2025)
di: Yang, Chenyang, et al.
Pubblicazione: (2025)
Natural Language Generation from Visual Events: State-of-the-Art and Key Open Questions
di: Surikuchi, Aditya K, et al.
Pubblicazione: (2025)
di: Surikuchi, Aditya K, et al.
Pubblicazione: (2025)
Where is the multimodal goal post? On the Ability of Foundation Models to Recognize Contextually Important Moments
di: Surikuchi, Aditya K, et al.
Pubblicazione: (2026)
di: Surikuchi, Aditya K, et al.
Pubblicazione: (2026)
Describing Images $\textit{Fast and Slow}$: Quantifying and Predicting the Variation in Human Signals during Visuo-Linguistic Processes
di: Takmaz, Ece, et al.
Pubblicazione: (2024)
di: Takmaz, Ece, et al.
Pubblicazione: (2024)
Underspecification in Language Modeling Tasks: A Causality-Informed Study of Gendered Pronoun Resolution
di: McMilin, Emily
Pubblicazione: (2022)
di: McMilin, Emily
Pubblicazione: (2022)
LHAW: Controllable Underspecification for Long-Horizon Tasks
di: Pu, George, et al.
Pubblicazione: (2026)
di: Pu, George, et al.
Pubblicazione: (2026)
Not (yet) the whole story: Evaluating Visual Storytelling Requires More than Measuring Coherence, Grounding, and Repetition
di: Surikuchi, Aditya K, et al.
Pubblicazione: (2024)
di: Surikuchi, Aditya K, et al.
Pubblicazione: (2024)
Revisiting Prompt Sensitivity in Large Language Models for Text Classification: The Role of Prompt Underspecification
di: Pecher, Branislav, et al.
Pubblicazione: (2026)
di: Pecher, Branislav, et al.
Pubblicazione: (2026)
How Do Large Language Models Learn Concepts During Continual Pre-Training?
di: Yao, Barry Menglong, et al.
Pubblicazione: (2026)
di: Yao, Barry Menglong, et al.
Pubblicazione: (2026)
Synthetic Pre-Pre-Training Improves Language Model Robustness to Noisy Pre-Training Data
di: Guo, Xu, et al.
Pubblicazione: (2026)
di: Guo, Xu, et al.
Pubblicazione: (2026)
Anchored Preference Optimization and Contrastive Revisions: Addressing Underspecification in Alignment
di: D'Oosterlinck, Karel, et al.
Pubblicazione: (2024)
di: D'Oosterlinck, Karel, et al.
Pubblicazione: (2024)
Min-K%++: Improved Baseline for Detecting Pre-Training Data from Large Language Models
di: Zhang, Jingyang, et al.
Pubblicazione: (2024)
di: Zhang, Jingyang, et al.
Pubblicazione: (2024)
Ask Good Questions for Large Language Models
di: Wu, Qi, et al.
Pubblicazione: (2025)
di: Wu, Qi, et al.
Pubblicazione: (2025)
Tomato, Tomahto, Tomate: Do Multilingual Language Models Understand Based on Subword-Level Semantic Concepts?
di: Zhang, Crystina, et al.
Pubblicazione: (2024)
di: Zhang, Crystina, et al.
Pubblicazione: (2024)
What Language is This? Ask Your Tokenizer
di: Meister, Clara, et al.
Pubblicazione: (2026)
di: Meister, Clara, et al.
Pubblicazione: (2026)
On the Interplay of Pre-Training, Mid-Training, and RL on Reasoning Language Models
di: Zhang, Charlie, et al.
Pubblicazione: (2025)
di: Zhang, Charlie, et al.
Pubblicazione: (2025)
Multi-Attribute Multi-Grained Adaptation of Pre-Trained Language Models for Text Understanding from Bayesian Perspective
di: Zhang, You, et al.
Pubblicazione: (2025)
di: Zhang, You, et al.
Pubblicazione: (2025)
Sigma: Semantically Informative Pre-training for Skeleton-based Sign Language Understanding
di: Pu, Muxin, et al.
Pubblicazione: (2025)
di: Pu, Muxin, et al.
Pubblicazione: (2025)
Quantifying Memorization and Detecting Training Data of Pre-trained Language Models using Japanese Newspaper
di: Ishihara, Shotaro, et al.
Pubblicazione: (2024)
di: Ishihara, Shotaro, et al.
Pubblicazione: (2024)
Exploring Forgetting in Large Language Model Pre-Training
di: Liao, Chonghua, et al.
Pubblicazione: (2024)
di: Liao, Chonghua, et al.
Pubblicazione: (2024)
Do Language Models Understand Honorific Systems in Javanese?
di: Farhansyah, Mohammad Rifqi, et al.
Pubblicazione: (2025)
di: Farhansyah, Mohammad Rifqi, et al.
Pubblicazione: (2025)
Ask a Local: Detecting Hallucinations With Specialized Model Divergence
di: Creo, Aldan, et al.
Pubblicazione: (2025)
di: Creo, Aldan, et al.
Pubblicazione: (2025)
Incremental Sentence Processing Mechanisms in Autoregressive Transformer Language Models
di: Hanna, Michael, et al.
Pubblicazione: (2024)
di: Hanna, Michael, et al.
Pubblicazione: (2024)
Do Large Language Models Walk Their Talk? Measuring the Gap Between Implicit Associations, Self-Report, and Behavioral Altruism
di: Andric, Sandro
Pubblicazione: (2025)
di: Andric, Sandro
Pubblicazione: (2025)
Ask Early, Ask Late, Ask Right: When Does Clarification Timing Matter for Long-Horizon Agents?
di: Gulati, Anmol, et al.
Pubblicazione: (2026)
di: Gulati, Anmol, et al.
Pubblicazione: (2026)
Vision-Language Models Do Not Understand Negation
di: Alhamoud, Kumail, et al.
Pubblicazione: (2025)
di: Alhamoud, Kumail, et al.
Pubblicazione: (2025)
Knowing When to Ask -- Bridging Large Language Models and Data
di: Radhakrishnan, Prashanth, et al.
Pubblicazione: (2024)
di: Radhakrishnan, Prashanth, et al.
Pubblicazione: (2024)
MiniPLM: Knowledge Distillation for Pre-Training Language Models
di: Gu, Yuxian, et al.
Pubblicazione: (2024)
di: Gu, Yuxian, et al.
Pubblicazione: (2024)
Analysing The Impact of Sequence Composition on Language Model Pre-Training
di: Zhao, Yu, et al.
Pubblicazione: (2024)
di: Zhao, Yu, et al.
Pubblicazione: (2024)
Instruction Pre-Training: Language Models are Supervised Multitask Learners
di: Cheng, Daixuan, et al.
Pubblicazione: (2024)
di: Cheng, Daixuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Is my model perplexed for the right reason? Contrasting LLMs' Benchmark Behavior with Token-Level Perplexity
di: Prins, Zoë, et al.
Pubblicazione: (2026) -
The BLA Benchmark: Investigating Basic Language Abilities of Pre-Trained Multimodal Models
di: Chen, Xinyi, et al.
Pubblicazione: (2023) -
Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms
di: Hanna, Michael, et al.
Pubblicazione: (2024) -
Are formal and functional linguistic mechanisms dissociated in language models?
di: Hanna, Michael, et al.
Pubblicazione: (2025) -
Beyond Divergent Creativity: A Human-Based Evaluation of Creativity in Large Language Models
di: Nakajima, Kumiko, et al.
Pubblicazione: (2026)