Laying Anchors: Semantically Priming Numerals in Language Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Sharma, Mandar, Taware, Rutuja Murlidhar, Koirala, Pravesh, Muralidhar, Nikhil, Ramakrishnan, Naren |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Information Guided Regularization for Fine-tuning Language Models
by: Sharma, Mandar, et al.
Published: (2024)
by: Sharma, Mandar, et al.
Published: (2024)
Can an LLM Induce a Graph? Investigating Memory Drift and Context Length
by: Yousuf, Raquib Bin, et al.
Published: (2025)
by: Yousuf, Raquib Bin, et al.
Published: (2025)
Large Multi-Modal Models (LMMs) as Universal Foundation Models for AI-Native Wireless Systems
by: Xu, Shengzhe, et al.
Published: (2024)
by: Xu, Shengzhe, et al.
Published: (2024)
LLM Augmentations to support Analytical Reasoning over Multiple Documents
by: Yousuf, Raquib Bin, et al.
Published: (2024)
by: Yousuf, Raquib Bin, et al.
Published: (2024)
AdvAnchor: Enhancing Diffusion Model Unlearning with Adversarial Anchors
by: Zhao, Mengnan, et al.
Published: (2024)
by: Zhao, Mengnan, et al.
Published: (2024)
Semantic Anchors in In-Context Learning: Why Small LLMs Cannot Flip Their Labels
by: Kumar, Anantha Padmanaban Krishna
Published: (2025)
by: Kumar, Anantha Padmanaban Krishna
Published: (2025)
The Prompt is Mightier than the Example
by: Xu, Shengzhe, et al.
Published: (2025)
by: Xu, Shengzhe, et al.
Published: (2025)
Constraining Sequential Model Editing with Editing Anchor Compression
by: Xu, Hao-Xiang, et al.
Published: (2025)
by: Xu, Hao-Xiang, et al.
Published: (2025)
QuOTE: Question-Oriented Text Embeddings
by: Neeser, Andrew, et al.
Published: (2025)
by: Neeser, Andrew, et al.
Published: (2025)
Hindsight is 20/20: Building Agent Memory that Retains, Recalls, and Reflects
by: Latimer, Chris, et al.
Published: (2025)
by: Latimer, Chris, et al.
Published: (2025)
Reinforcement Learning with Rubric Anchors
by: Huang, Zenan, et al.
Published: (2025)
by: Huang, Zenan, et al.
Published: (2025)
Value-Aware Numerical Representations for Transformer Language Models
by: Dutulescu, Andreea, et al.
Published: (2026)
by: Dutulescu, Andreea, et al.
Published: (2026)
Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models
by: Nath, Vaskar, et al.
Published: (2025)
by: Nath, Vaskar, et al.
Published: (2025)
Distinguishing the Knowable from the Unknowable with Language Models
by: Ahdritz, Gustaf, et al.
Published: (2024)
by: Ahdritz, Gustaf, et al.
Published: (2024)
Answer Matching Outperforms Multiple Choice for Language Model Evaluation
by: Chandak, Nikhil, et al.
Published: (2025)
by: Chandak, Nikhil, et al.
Published: (2025)
ContextFocus: Activation Steering for Contextual Faithfulness in Large Language Models
by: Anand, Nikhil, et al.
Published: (2026)
by: Anand, Nikhil, et al.
Published: (2026)
Emergent Semantic Role Understanding in Language Models
by: Griffiths, Carla, et al.
Published: (2026)
by: Griffiths, Carla, et al.
Published: (2026)
Induced Numerical Instability: Hidden Costs in Multimodal Large Language Models
by: Wong, Wai Tuck, et al.
Published: (2026)
by: Wong, Wai Tuck, et al.
Published: (2026)
xVal: A Continuous Numerical Tokenization for Scientific Language Models
by: Golkar, Siavash, et al.
Published: (2023)
by: Golkar, Siavash, et al.
Published: (2023)
Correlated Errors in Large Language Models
by: Kim, Elliot, et al.
Published: (2025)
by: Kim, Elliot, et al.
Published: (2025)
Understanding Post-hoc Explainers: The Case of Anchors
by: Lopardo, Gianluigi, et al.
Published: (2023)
by: Lopardo, Gianluigi, et al.
Published: (2023)
Why LLMs Are Bad at Synthetic Table Generation (and what to do about it)
by: Xu, Shengzhe, et al.
Published: (2024)
by: Xu, Shengzhe, et al.
Published: (2024)
CoLa: Learning to Interactively Collaborate with Large Language Models
by: Sharma, Abhishek, et al.
Published: (2025)
by: Sharma, Abhishek, et al.
Published: (2025)
SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs
by: Lee, Chanuk, et al.
Published: (2026)
by: Lee, Chanuk, et al.
Published: (2026)
Thought Anchors: Which LLM Reasoning Steps Matter?
by: Bogdan, Paul C., et al.
Published: (2025)
by: Bogdan, Paul C., et al.
Published: (2025)
Laying the Foundation First? Investigating the Generalization from Atomic Skills to Complex Reasoning Tasks
by: Huang, Yuncheng, et al.
Published: (2024)
by: Huang, Yuncheng, et al.
Published: (2024)
Evaluating Computational Accuracy of Large Language Models in Numerical Reasoning Tasks for Healthcare Applications
by: Malghan, Arjun R.
Published: (2025)
by: Malghan, Arjun R.
Published: (2025)
Effective Reasoning Chains Reduce Intrinsic Dimensionality
by: Prasad, Archiki, et al.
Published: (2026)
by: Prasad, Archiki, et al.
Published: (2026)
Analyzing the Role of Semantic Representations in the Era of Large Language Models
by: Jin, Zhijing, et al.
Published: (2024)
by: Jin, Zhijing, et al.
Published: (2024)
Thinking in Latents: Adaptive Anchor Refinement for Implicit Reasoning in LLMs
by: Sheshanarayana, Disha, et al.
Published: (2026)
by: Sheshanarayana, Disha, et al.
Published: (2026)
A Sea of Words: An In-Depth Analysis of Anchors for Text Data
by: Lopardo, Gianluigi, et al.
Published: (2022)
by: Lopardo, Gianluigi, et al.
Published: (2022)
Utilizing Metadata for Better Retrieval-Augmented Generation
by: Yousuf, Raquib Bin, et al.
Published: (2026)
by: Yousuf, Raquib Bin, et al.
Published: (2026)
G-Loss: Graph-Guided Fine-Tuning of Language Models
by: Sharma, Aditya, et al.
Published: (2026)
by: Sharma, Aditya, et al.
Published: (2026)
COLD-Steer: Steering Large Language Models via In-Context One-step Learning Dynamics
by: Sharma, Kartik, et al.
Published: (2026)
by: Sharma, Kartik, et al.
Published: (2026)
Innovations in Neural Data-to-text Generation: A Survey
by: Sharma, Mandar, et al.
Published: (2022)
by: Sharma, Mandar, et al.
Published: (2022)
Luna-2: Scalable Single-Token Evaluation with Small Language Models
by: Goel, Vatsal, et al.
Published: (2026)
by: Goel, Vatsal, et al.
Published: (2026)
LoRA+: Efficient Low Rank Adaptation of Large Models
by: Hayou, Soufiane, et al.
Published: (2024)
by: Hayou, Soufiane, et al.
Published: (2024)
Improving Uncertainty Quantification in Large Language Models via Semantic Embeddings
by: Grewal, Yashvir S., et al.
Published: (2024)
by: Grewal, Yashvir S., et al.
Published: (2024)
Semantic Token Clustering for Efficient Uncertainty Quantification in Large Language Models
by: Cao, Qi, et al.
Published: (2026)
by: Cao, Qi, et al.
Published: (2026)
Emergent Representations of Program Semantics in Language Models Trained on Programs
by: Jin, Charles, et al.
Published: (2023)
by: Jin, Charles, et al.
Published: (2023)
Similar Items
-
Information Guided Regularization for Fine-tuning Language Models
by: Sharma, Mandar, et al.
Published: (2024) -
Can an LLM Induce a Graph? Investigating Memory Drift and Context Length
by: Yousuf, Raquib Bin, et al.
Published: (2025) -
Large Multi-Modal Models (LMMs) as Universal Foundation Models for AI-Native Wireless Systems
by: Xu, Shengzhe, et al.
Published: (2024) -
LLM Augmentations to support Analytical Reasoning over Multiple Documents
by: Yousuf, Raquib Bin, et al.
Published: (2024) -
AdvAnchor: Enhancing Diffusion Model Unlearning with Adversarial Anchors
by: Zhao, Mengnan, et al.
Published: (2024)