Salvato in:
| Autori principali: | Wang, Haojin, Zhu, Zining, Shi, Freda |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2505.12244 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
How Tokenization Limits Phonological Knowledge Representation in Language Models and How to Improve Them
di: Liao, Disen, et al.
Pubblicazione: (2026)
di: Liao, Disen, et al.
Pubblicazione: (2026)
Learning Language Structures through Grounding
di: Shi, Freda
Pubblicazione: (2024)
di: Shi, Freda
Pubblicazione: (2024)
Understanding Language Model Circuits through Knowledge Editing
di: Ge, Huaizhi, et al.
Pubblicazione: (2024)
di: Ge, Huaizhi, et al.
Pubblicazione: (2024)
Better Language Model Inversion by Compactly Representing Next-Token Distributions
di: Nazir, Murtaza, et al.
Pubblicazione: (2025)
di: Nazir, Murtaza, et al.
Pubblicazione: (2025)
Efficient Training of Language Models with Compact and Consistent Next Token Distributions
di: Sathe, Ashutosh, et al.
Pubblicazione: (2024)
di: Sathe, Ashutosh, et al.
Pubblicazione: (2024)
Logical forms complement probability in understanding language model (and human) performance
di: Wang, Yixuan, et al.
Pubblicazione: (2025)
di: Wang, Yixuan, et al.
Pubblicazione: (2025)
Error Reflection Prompting: Can Large Language Models Successfully Understand Errors?
di: Li, Jason, et al.
Pubblicazione: (2025)
di: Li, Jason, et al.
Pubblicazione: (2025)
Real Images, Worse Judgments: Evaluating Vision-Language Models on Concreteness and Imagery
di: Jiang, Yifan, et al.
Pubblicazione: (2026)
di: Jiang, Yifan, et al.
Pubblicazione: (2026)
How Well Can Knowledge Edit Methods Edit Perplexing Knowledge?
di: Ge, Huaizhi, et al.
Pubblicazione: (2024)
di: Ge, Huaizhi, et al.
Pubblicazione: (2024)
Ain't Misbehavin' -- Using LLMs to Generate Expressive Robot Behavior in Conversations with the Tabletop Robot Haru
di: Wang, Zining, et al.
Pubblicazione: (2024)
di: Wang, Zining, et al.
Pubblicazione: (2024)
LingGym: How Far Are LLMs from Thinking Like Field Linguists?
di: Yang, Changbing, et al.
Pubblicazione: (2025)
di: Yang, Changbing, et al.
Pubblicazione: (2025)
Beyond the Next Token: Towards Prompt-Robust Zero-Shot Classification via Efficient Multi-Token Prediction
di: Qian, Junlang, et al.
Pubblicazione: (2025)
di: Qian, Junlang, et al.
Pubblicazione: (2025)
Chartographer: Counterfactual Chart Generation for Evaluating Vision-Language Models
di: Jiang, Yifan, et al.
Pubblicazione: (2026)
di: Jiang, Yifan, et al.
Pubblicazione: (2026)
LLM-Generated Black-box Explanations Can Be Adversarially Helpful
di: Ajwani, Rohan, et al.
Pubblicazione: (2024)
di: Ajwani, Rohan, et al.
Pubblicazione: (2024)
From Token to Token Pair: Efficient Prompt Compression for Large Language Models in Clinical Prediction
di: Zhu, Mingcheng, et al.
Pubblicazione: (2026)
di: Zhu, Mingcheng, et al.
Pubblicazione: (2026)
Blessing of Multilinguality: A Systematic Analysis of Multilingual In-Context Learning
di: Tu, Yilei, et al.
Pubblicazione: (2025)
di: Tu, Yilei, et al.
Pubblicazione: (2025)
Multimodal Latent Language Modeling with Next-Token Diffusion
di: Sun, Yutao, et al.
Pubblicazione: (2024)
di: Sun, Yutao, et al.
Pubblicazione: (2024)
Plug and Play with Prompts: A Prompt Tuning Approach for Controlling Text Generation
di: Ajwani, Rohan Deepak, et al.
Pubblicazione: (2024)
di: Ajwani, Rohan Deepak, et al.
Pubblicazione: (2024)
Can Large Language Models Understand Internet Buzzwords Through User-Generated Content
di: Huang, Chen, et al.
Pubblicazione: (2025)
di: Huang, Chen, et al.
Pubblicazione: (2025)
Bigram Subnetworks: Mapping to Next Tokens in Transformer Language Models
di: Chang, Tyler A., et al.
Pubblicazione: (2025)
di: Chang, Tyler A., et al.
Pubblicazione: (2025)
Measuring Language Model Hallucinations Through Distributional Correctness
di: Burns, Thomas F
Pubblicazione: (2025)
di: Burns, Thomas F
Pubblicazione: (2025)
Scenarios and Approaches for Situated Natural Language Explanations
di: Qiu, Pengshuo, et al.
Pubblicazione: (2024)
di: Qiu, Pengshuo, et al.
Pubblicazione: (2024)
One Token Can Help! Learning Scalable and Pluggable Virtual Tokens for Retrieval-Augmented Large Language Models
di: Zhu, Yutao, et al.
Pubblicazione: (2024)
di: Zhu, Yutao, et al.
Pubblicazione: (2024)
Fine-tuning vs Prompting, Can Language Models Understand Human Values?
di: Sun, Pingwei
Pubblicazione: (2024)
di: Sun, Pingwei
Pubblicazione: (2024)
Can Large Language Models Understand Context?
di: Zhu, Yilun, et al.
Pubblicazione: (2024)
di: Zhu, Yilun, et al.
Pubblicazione: (2024)
Text-to-Distribution Prediction with Quantile Tokens and Neighbor Context
di: Zhu, Yilun, et al.
Pubblicazione: (2026)
di: Zhu, Yilun, et al.
Pubblicazione: (2026)
Robust Prompt Optimization for Large Language Models Against Distribution Shifts
di: Li, Moxin, et al.
Pubblicazione: (2023)
di: Li, Moxin, et al.
Pubblicazione: (2023)
Revisiting Graph-Tokenizing Large Language Models: A Systematic Evaluation of Graph Token Understanding
di: Zhang, Zhongjian, et al.
Pubblicazione: (2026)
di: Zhang, Zhongjian, et al.
Pubblicazione: (2026)
Trading Complexity for Expressivity Through Structured Generalized Linear Token Mixing
di: Fagnou, Erwan, et al.
Pubblicazione: (2026)
di: Fagnou, Erwan, et al.
Pubblicazione: (2026)
Structured Tree Alignment for Evaluation of (Speech) Constituency Parsing
di: Shi, Freda, et al.
Pubblicazione: (2024)
di: Shi, Freda, et al.
Pubblicazione: (2024)
Understanding the Emergence of Seemingly Useless Features in Next-Token Predictors
di: Rofin, Mark, et al.
Pubblicazione: (2026)
di: Rofin, Mark, et al.
Pubblicazione: (2026)
The Mechanistic Emergence of Symbol Grounding in Language Models
di: Wu, Shuyu, et al.
Pubblicazione: (2025)
di: Wu, Shuyu, et al.
Pubblicazione: (2025)
Language Model Maps for Prompt-Response Distributions via Log-Likelihood Vectors
di: Takase, Yusuke, et al.
Pubblicazione: (2026)
di: Takase, Yusuke, et al.
Pubblicazione: (2026)
Brainstorming Brings Power to Large Language Models of Knowledge Reasoning
di: Qin, Zining, et al.
Pubblicazione: (2024)
di: Qin, Zining, et al.
Pubblicazione: (2024)
NDP: Next Distribution Prediction as a More Broad Target
di: Ruan, Junhao, et al.
Pubblicazione: (2024)
di: Ruan, Junhao, et al.
Pubblicazione: (2024)
Distributed Interpretability and Control for Large Language Models
di: Zhu, Zining
Pubblicazione: (2026)
di: Zhu, Zining
Pubblicazione: (2026)
Metacognitive Prompting Improves Understanding in Large Language Models
di: Wang, Yuqing, et al.
Pubblicazione: (2023)
di: Wang, Yuqing, et al.
Pubblicazione: (2023)
Do Vision-Language Models Represent Space and How? Evaluating Spatial Frame of Reference Under Ambiguities
di: Zhang, Zheyuan, et al.
Pubblicazione: (2024)
di: Zhang, Zheyuan, et al.
Pubblicazione: (2024)
A Law of Next-Token Prediction in Large Language Models
di: He, Hangfeng, et al.
Pubblicazione: (2024)
di: He, Hangfeng, et al.
Pubblicazione: (2024)
Differentially Private Next-Token Prediction of Large Language Models
di: Flemings, James, et al.
Pubblicazione: (2024)
di: Flemings, James, et al.
Pubblicazione: (2024)
Documenti analoghi
-
How Tokenization Limits Phonological Knowledge Representation in Language Models and How to Improve Them
di: Liao, Disen, et al.
Pubblicazione: (2026) -
Learning Language Structures through Grounding
di: Shi, Freda
Pubblicazione: (2024) -
Understanding Language Model Circuits through Knowledge Editing
di: Ge, Huaizhi, et al.
Pubblicazione: (2024) -
Better Language Model Inversion by Compactly Representing Next-Token Distributions
di: Nazir, Murtaza, et al.
Pubblicazione: (2025) -
Efficient Training of Language Models with Compact and Consistent Next Token Distributions
di: Sathe, Ashutosh, et al.
Pubblicazione: (2024)