Word Boundary Information Isn't Useful for Encoder Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Gow-Smith, Edward, Phelps, Dylan, Madabushi, Harish Tayyar, Scarton, Carolina, Villavicencio, Aline |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FS-RAG: A Frame Semantics Based Approach for Improved Factual Accuracy in Large Language Models
by: Madabushi, Harish Tayyar
Published: (2024)
by: Madabushi, Harish Tayyar
Published: (2024)
Sign of the Times: Evaluating the use of Large Language Models for Idiomaticity Detection
by: Phelps, Dylan, et al.
Published: (2024)
by: Phelps, Dylan, et al.
Published: (2024)
Scaling Policy Compliance Assessment in Language Models with Policy Reasoning Traces
by: Imperial, Joseph Marvin, et al.
Published: (2025)
by: Imperial, Joseph Marvin, et al.
Published: (2025)
Pre-Trained Language Models Represent Some Geographic Populations Better Than Others
by: Dunn, Jonathan, et al.
Published: (2024)
by: Dunn, Jonathan, et al.
Published: (2024)
Standardize: Aligning Language Models with Expert-Defined Standards for Content Generation
by: Imperial, Joseph Marvin, et al.
Published: (2024)
by: Imperial, Joseph Marvin, et al.
Published: (2024)
SpeciaLex: A Benchmark for In-Context Specialized Lexicon Learning
by: Imperial, Joseph Marvin, et al.
Published: (2024)
by: Imperial, Joseph Marvin, et al.
Published: (2024)
Safer Policy Compliance with Dynamic Epistemic Fallback
by: Imperial, Joseph Marvin, et al.
Published: (2026)
by: Imperial, Joseph Marvin, et al.
Published: (2026)
Stands to Reason: Investigating the Effect of Reasoning on Idiomaticity Detection
by: Phelps, Dylan, et al.
Published: (2025)
by: Phelps, Dylan, et al.
Published: (2025)
The Inherent Limits of Pretrained LLMs: The Unexpected Convergence of Instruction Tuning and In-Context Learning Capabilities
by: Bigoulaeva, Irina, et al.
Published: (2025)
by: Bigoulaeva, Irina, et al.
Published: (2025)
Evaluating CxG Generalisation in LLMs via Construction-Based NLI Fine Tuning
by: Mackintosh, Tom, et al.
Published: (2025)
by: Mackintosh, Tom, et al.
Published: (2025)
Dancing with Deer: A Constructional Perspective on MWEs in the Era of LLMs
by: Bonial, Claire, et al.
Published: (2025)
by: Bonial, Claire, et al.
Published: (2025)
Neither Stochastic Parroting nor AGI: LLMs Solve Tasks through Context-Directed Extrapolation from Training Data Priors
by: Madabushi, Harish Tayyar, et al.
Published: (2025)
by: Madabushi, Harish Tayyar, et al.
Published: (2025)
Evaluating Large Language Models on Multiword Expressions in Multilingual and Code-Switched Contexts
by: De Leon, Frances Laureano, et al.
Published: (2025)
by: De Leon, Frances Laureano, et al.
Published: (2025)
Are Emergent Abilities in Large Language Models just In-Context Learning?
by: Lu, Sheng, et al.
Published: (2023)
by: Lu, Sheng, et al.
Published: (2023)
Code-Mixed Probes Show How Pre-Trained Models Generalise On Code-Switched Text
by: De Leon, Frances A. Laureano, et al.
Published: (2024)
by: De Leon, Frances A. Laureano, et al.
Published: (2024)
Enhancing Idiomatic Representation in Multiple Languages via an Adaptive Contrastive Triplet Loss
by: He, Wei, et al.
Published: (2024)
by: He, Wei, et al.
Published: (2024)
Investigating Idiomaticity in Word Representations
by: He, Wei, et al.
Published: (2024)
by: He, Wei, et al.
Published: (2024)
Standardizing Intelligence: Aligning Generative AI for Regulatory and Operational Compliance
by: Imperial, Joseph Marvin, et al.
Published: (2025)
by: Imperial, Joseph Marvin, et al.
Published: (2025)
Adapting Whisper for Regional Dialects: Enhancing Public Services for Vulnerable Populations in the United Kingdom
by: Torgbi, Melissa, et al.
Published: (2025)
by: Torgbi, Melissa, et al.
Published: (2025)
Fine-Tuning on Diverse Reasoning Chains Drives Within-Inference CoT Refinement in LLMs
by: Puerto, Haritz, et al.
Published: (2024)
by: Puerto, Haritz, et al.
Published: (2024)
Illusion or Algorithm? Investigating Memorization, Emergence, and Symbolic Processing in In-Context Learning
by: Niu, Jingcheng, et al.
Published: (2025)
by: Niu, Jingcheng, et al.
Published: (2025)
Beyond Memorization: Assessing Semantic Generalization in Large Language Models Using Phrasal Constructions
by: Scivetti, Wesley, et al.
Published: (2025)
by: Scivetti, Wesley, et al.
Published: (2025)
Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2026)
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2026)
Beyond surface form: A pipeline for semantic analysis in Alzheimer's Disease detection from spontaneous speech
by: Phelps, Dylan, et al.
Published: (2025)
by: Phelps, Dylan, et al.
Published: (2025)
Inverse Scaling: When Bigger Isn't Better
by: McKenzie, Ian R., et al.
Published: (2023)
by: McKenzie, Ian R., et al.
Published: (2023)
When Meaning Isn't Literal: Exploring Idiomatic Meaning Across Languages and Modalities
by: Das, Sarmistha, et al.
Published: (2026)
by: Das, Sarmistha, et al.
Published: (2026)
Sheffield's Submission to the AmericasNLP Shared Task on Machine Translation into Indigenous Languages
by: Gow-Smith, Edward, et al.
Published: (2023)
by: Gow-Smith, Edward, et al.
Published: (2023)
Strong Reasoning Isn't Enough: Evaluating Evidence Elicitation in Interactive Diagnosis
by: Long, Zhuohan, et al.
Published: (2026)
by: Long, Zhuohan, et al.
Published: (2026)
Recall Isn't Enough: Bounding Commitments in Personalized Language Systems
by: Tang, Rui, et al.
Published: (2026)
by: Tang, Rui, et al.
Published: (2026)
When Correct Isn't Usable: Improving Structured Output Reliability in Small Language Models
by: Galeone, Cosimo, et al.
Published: (2026)
by: Galeone, Cosimo, et al.
Published: (2026)
Reasoning Isn't Enough: Examining Truth-Bias and Sycophancy in LLMs
by: Barkett, Emilio, et al.
Published: (2025)
by: Barkett, Emilio, et al.
Published: (2025)
Being Kind Isn't Always Being Safe: Diagnosing Affective Hallucination in LLMs
by: Kim, Sewon, et al.
Published: (2025)
by: Kim, Sewon, et al.
Published: (2025)
Surprise! Uniform Information Density Isn't the Whole Story: Predicting Surprisal Contours in Long-form Discourse
by: Tsipidi, Eleftheria, et al.
Published: (2024)
by: Tsipidi, Eleftheria, et al.
Published: (2024)
Mathematics Isn't Culture-Free: Probing Cultural Gaps via Entity and Scenario Perturbations
by: Tomar, Aditya, et al.
Published: (2025)
by: Tomar, Aditya, et al.
Published: (2025)
SAIE Framework: Support Alone Isn't Enough -- Advancing LLM Training with Adversarial Remarks
by: Loem, Mengsay, et al.
Published: (2023)
by: Loem, Mengsay, et al.
Published: (2023)
Leveraging Large Language Models for Zero-shot Lay Summarisation in Biomedicine and Beyond
by: Goldsack, Tomas, et al.
Published: (2025)
by: Goldsack, Tomas, et al.
Published: (2025)
When Fairness Isn't Statistical: The Limits of Machine Learning in Evaluating Legal Reasoning
by: Barale, Claire, et al.
Published: (2025)
by: Barale, Claire, et al.
Published: (2025)
Unlearning Isn't Deletion: Investigating Reversibility of Machine Unlearning in LLMs
by: Xu, Xiaoyu, et al.
Published: (2025)
by: Xu, Xiaoyu, et al.
Published: (2025)
Exploring Vision Language Models for Multimodal and Multilingual Stance Detection
by: Vasilakes, Jake, et al.
Published: (2025)
by: Vasilakes, Jake, et al.
Published: (2025)
Talk Isn't Always Cheap: Understanding Failure Modes in Multi-Agent Debate
by: Wynn, Andrea, et al.
Published: (2025)
by: Wynn, Andrea, et al.
Published: (2025)
Similar Items
-
FS-RAG: A Frame Semantics Based Approach for Improved Factual Accuracy in Large Language Models
by: Madabushi, Harish Tayyar
Published: (2024) -
Sign of the Times: Evaluating the use of Large Language Models for Idiomaticity Detection
by: Phelps, Dylan, et al.
Published: (2024) -
Scaling Policy Compliance Assessment in Language Models with Policy Reasoning Traces
by: Imperial, Joseph Marvin, et al.
Published: (2025) -
Pre-Trained Language Models Represent Some Geographic Populations Better Than Others
by: Dunn, Jonathan, et al.
Published: (2024) -
Standardize: Aligning Language Models with Expert-Defined Standards for Content Generation
by: Imperial, Joseph Marvin, et al.
Published: (2024)