Symbol tuning improves in-context learning in language models
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Jerry, Hou, Le, Lampinen, Andrew, Chen, Xiangning, Huang, Da, Tay, Yi, Chen, Xinyun, Lu, Yifeng, Zhou, Denny, Ma, Tengyu, Le, Quoc V. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Simple synthetic data reduces sycophancy in large language models
by: Wei, Jerry, et al.
Published: (2023)
by: Wei, Jerry, et al.
Published: (2023)
Large Language Models as Optimizers
by: Yang, Chengrun, et al.
Published: (2023)
by: Yang, Chengrun, et al.
Published: (2023)
Large Language Models as Tool Makers
by: Cai, Tianle, et al.
Published: (2023)
by: Cai, Tianle, et al.
Published: (2023)
The in-context inductive biases of vision-language models differ across modalities
by: Allen, Kelsey, et al.
Published: (2025)
by: Allen, Kelsey, et al.
Published: (2025)
Long-form factuality in large language models
by: Wei, Jerry, et al.
Published: (2024)
by: Wei, Jerry, et al.
Published: (2024)
On the generalization of language models from in-context learning and finetuning: a controlled study
by: Lampinen, Andrew K., et al.
Published: (2025)
by: Lampinen, Andrew K., et al.
Published: (2025)
The broader spectrum of in-context learning
by: Lampinen, Andrew Kyle, et al.
Published: (2024)
by: Lampinen, Andrew Kyle, et al.
Published: (2024)
How do language models learn facts? Dynamics, curricula and hallucinations
by: Zucchet, Nicolas, et al.
Published: (2025)
by: Zucchet, Nicolas, et al.
Published: (2025)
Take a Step Back: Evoking Reasoning via Abstraction in Large Language Models
by: Zheng, Huaixiu Steven, et al.
Published: (2023)
by: Zheng, Huaixiu Steven, et al.
Published: (2023)
NATURAL PLAN: Benchmarking LLMs on Natural Language Planning
by: Zheng, Huaixiu Steven, et al.
Published: (2024)
by: Zheng, Huaixiu Steven, et al.
Published: (2024)
Chain of Thought Empowers Transformers to Solve Inherently Serial Problems
by: Li, Zhiyuan, et al.
Published: (2024)
by: Li, Zhiyuan, et al.
Published: (2024)
Premise Order Matters in Reasoning with Large Language Models
by: Chen, Xinyun, et al.
Published: (2024)
by: Chen, Xinyun, et al.
Published: (2024)
Just-in-time and distributed task representations in language models
by: Li, Yuxuan, et al.
Published: (2025)
by: Li, Yuxuan, et al.
Published: (2025)
Naturalistic Computational Cognitive Science: Towards generalizable models and theories that capture the full range of natural behavior
by: Carvalho, Wilka, et al.
Published: (2025)
by: Carvalho, Wilka, et al.
Published: (2025)
Retrieval-augmented in-context learning for multimodal large language models in disease classification
by: Zhan, Zaifu, et al.
Published: (2025)
by: Zhan, Zaifu, et al.
Published: (2025)
Linear representations in language models can change dramatically over a conversation
by: Lampinen, Andrew Kyle, et al.
Published: (2026)
by: Lampinen, Andrew Kyle, et al.
Published: (2026)
Self-Discover: Large Language Models Self-Compose Reasoning Structures
by: Zhou, Pei, et al.
Published: (2024)
by: Zhou, Pei, et al.
Published: (2024)
Transformers Can Achieve Length Generalization But Not Robustly
by: Zhou, Yongchao, et al.
Published: (2024)
by: Zhou, Yongchao, et al.
Published: (2024)
Critical learner autonomy in the digital language learning context
by: Li‐Mei Chen, et al.
Published: (2024)
by: Li‐Mei Chen, et al.
Published: (2024)
Rethinking Generative Image Pretraining: How Far Are We From Scaling Up Next-Pixel Prediction?
by: Yan, Xinchen, et al.
Published: (2025)
by: Yan, Xinchen, et al.
Published: (2025)
Dynamic data sampler for cross-language transfer learning in large language models
by: Li, Yudong, et al.
Published: (2024)
by: Li, Yudong, et al.
Published: (2024)
Conjunction and additive constructions in the language of the Gavião of Rondônia
by: Denny Moore
Published: (2021)
by: Denny Moore
Published: (2021)
How does fine-tuning improve sensorimotor representations in large language models?
by: Wu, Minghua, et al.
Published: (2026)
by: Wu, Minghua, et al.
Published: (2026)
What do vision-language models see in the context? Investigating multimodal in-context learning
by: Santos, Gabriel O. dos, et al.
Published: (2025)
by: Santos, Gabriel O. dos, et al.
Published: (2025)
Khuynh hướng giải huyền thoại trong văn xuôi Việt Nam đương đại từ 1986 đến nay
by: Le, Quoc Hieu
Published: (2017)
by: Le, Quoc Hieu
Published: (2017)
Reassessing the Impact of Foreign Direct Investment on Environmental Quality in 112 Countries: A Bayesian Quantile Regression Approach
by: Dinh Le Quoc
Published: (2025)
by: Dinh Le Quoc
Published: (2025)
Learned feature representations are biased by complexity, learning order, position, and more
by: Lampinen, Andrew Kyle, et al.
Published: (2024)
by: Lampinen, Andrew Kyle, et al.
Published: (2024)
Pretrained transformer efficiently learns low-dimensional target functions in-context
by: Oko, Kazusato, et al.
Published: (2024)
by: Oko, Kazusato, et al.
Published: (2024)
Large language models reorganize representational geometry during in-context learning
by: Xiong, Hua-Dong, et al.
Published: (2026)
by: Xiong, Hua-Dong, et al.
Published: (2026)
Balancing Bank Profits With Sustainable Development Goals: Examining the Pivotal Role of Financial Stability
by: Nguyen Quoc Huy, et al.
Published: (2025)
by: Nguyen Quoc Huy, et al.
Published: (2025)
Large Language Models can Learn Rules
by: Zhu, Zhaocheng, et al.
Published: (2023)
by: Zhu, Zhaocheng, et al.
Published: (2023)
Reinforcement learning fine-tuning of language model for instruction following and math reasoning
by: Han, Yifu, et al.
Published: (2025)
by: Han, Yifu, et al.
Published: (2025)
The representation landscape of few-shot learning and fine-tuning in large language models
by: Doimo, Diego, et al.
Published: (2024)
by: Doimo, Diego, et al.
Published: (2024)
How Different Patterns of Policy Attention Drive Policy Diffusion: Evidence From China's River Chief System
by: Xiangning Chen, et al.
Published: (2025)
by: Xiangning Chen, et al.
Published: (2025)
Non-Asymptotic Length Generalization
by: Chen, Thomas, et al.
Published: (2025)
by: Chen, Thomas, et al.
Published: (2025)
Improving Latent Generalization Using Test-time Compute
by: Chaudhry, Arslan, et al.
Published: (2026)
by: Chaudhry, Arslan, et al.
Published: (2026)
Transforming peptide hormone prediction: The role of AI in modern proteomics
by: Nguyen Quoc Khanh Le
Published: (2024)
by: Nguyen Quoc Khanh Le
Published: (2024)
Artificial Intelligence in Proteomics Clinical Applications
by: Nguyen Quoc Khanh Le
Published: (2025)
by: Nguyen Quoc Khanh Le
Published: (2025)
B‐cell lymphoma classification using vision‐language models and in‐context learning
by: Mobina Shrestha, et al.
Published: (2025)
by: Mobina Shrestha, et al.
Published: (2025)
Outside Front Cover: Volume 5 Issue 4
by: Xiangyu Hou, et al.
Published: (2024)
by: Xiangyu Hou, et al.
Published: (2024)
Similar Items
-
Simple synthetic data reduces sycophancy in large language models
by: Wei, Jerry, et al.
Published: (2023) -
Large Language Models as Optimizers
by: Yang, Chengrun, et al.
Published: (2023) -
Large Language Models as Tool Makers
by: Cai, Tianle, et al.
Published: (2023) -
The in-context inductive biases of vision-language models differ across modalities
by: Allen, Kelsey, et al.
Published: (2025) -
Long-form factuality in large language models
by: Wei, Jerry, et al.
Published: (2024)