Pre-trained Language Models Improve the Few-shot Prompt Ability of Decision Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Yu, Xu, Pan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Revisiting Chain-of-Thought Prompting: Zero-shot Can Be Stronger than Few-shot
von: Cheng, Xiang, et al.
Veröffentlicht: (2025)
von: Cheng, Xiang, et al.
Veröffentlicht: (2025)
MemoryPrompt: A Light Wrapper to Improve Context Tracking in Pre-trained Language Models
von: Rakotonirina, Nathanaël Carraz, et al.
Veröffentlicht: (2024)
von: Rakotonirina, Nathanaël Carraz, et al.
Veröffentlicht: (2024)
Adaptive Few-shot Prompting for Machine Translation with Pre-trained Language Models
von: Tang, Lei, et al.
Veröffentlicht: (2025)
von: Tang, Lei, et al.
Veröffentlicht: (2025)
Evaluating Gender Bias Transfer between Pre-trained and Prompt-Adapted Language Models
von: Mackraz, Natalie, et al.
Veröffentlicht: (2024)
von: Mackraz, Natalie, et al.
Veröffentlicht: (2024)
Relational Prompt-based Pre-trained Language Models for Social Event Detection
von: Li, Pu, et al.
Veröffentlicht: (2024)
von: Li, Pu, et al.
Veröffentlicht: (2024)
LAMPO: Large Language Models as Preference Machines for Few-shot Ordinal Classification
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
Clinical information extraction for Low-resource languages with Few-shot learning using Pre-trained language models and Prompting
von: Richter-Pechanski, Phillip, et al.
Veröffentlicht: (2024)
von: Richter-Pechanski, Phillip, et al.
Veröffentlicht: (2024)
A Zero-shot and Few-shot Study of Instruction-Finetuned Large Language Models Applied to Clinical and Biomedical Tasks
von: Labrak, Yanis, et al.
Veröffentlicht: (2023)
von: Labrak, Yanis, et al.
Veröffentlicht: (2023)
Few-shot Personalization of LLMs with Mis-aligned Responses
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
Investigating Data Contamination for Pre-training Language Models
von: Jiang, Minhao, et al.
Veröffentlicht: (2024)
von: Jiang, Minhao, et al.
Veröffentlicht: (2024)
Aligning Pre-trained Models for Spoken Language Translation
von: Sedláček, Šimon, et al.
Veröffentlicht: (2024)
von: Sedláček, Šimon, et al.
Veröffentlicht: (2024)
Sequence-to-Sequence Spanish Pre-trained Language Models
von: Araujo, Vladimir, et al.
Veröffentlicht: (2023)
von: Araujo, Vladimir, et al.
Veröffentlicht: (2023)
HiFloat4 Format for Language Model Pre-training on Ascend NPUs
von: Taghian, Mehran, et al.
Veröffentlicht: (2026)
von: Taghian, Mehran, et al.
Veröffentlicht: (2026)
Integrating Pre-trained Language Model into Neural Machine Translation
von: Hwang, Soon-Jae, et al.
Veröffentlicht: (2023)
von: Hwang, Soon-Jae, et al.
Veröffentlicht: (2023)
Understanding Reasoning Ability of Language Models From the Perspective of Reasoning Paths Aggregation
von: Wang, Xinyi, et al.
Veröffentlicht: (2024)
von: Wang, Xinyi, et al.
Veröffentlicht: (2024)
Conifer: Improving Complex Constrained Instruction-Following Ability of Large Language Models
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
Few-Shot Cross-Lingual Transfer for Prompting Large Language Models in Low-Resource Languages
von: Toukmaji, Christopher
Veröffentlicht: (2024)
von: Toukmaji, Christopher
Veröffentlicht: (2024)
Structured Prompts Improve Evaluation of Language Models
von: Aali, Asad, et al.
Veröffentlicht: (2025)
von: Aali, Asad, et al.
Veröffentlicht: (2025)
Simple and Scalable Strategies to Continually Pre-train Large Language Models
von: Ibrahim, Adam, et al.
Veröffentlicht: (2024)
von: Ibrahim, Adam, et al.
Veröffentlicht: (2024)
Pre-training Limited Memory Language Models with Internal and External Knowledge
von: Zhao, Linxi, et al.
Veröffentlicht: (2025)
von: Zhao, Linxi, et al.
Veröffentlicht: (2025)
Sparse is Enough in Fine-tuning Pre-trained Large Language Models
von: Song, Weixi, et al.
Veröffentlicht: (2023)
von: Song, Weixi, et al.
Veröffentlicht: (2023)
PhoneLM:an Efficient and Capable Small Language Model Family through Principled Pre-training
von: Yi, Rongjie, et al.
Veröffentlicht: (2024)
von: Yi, Rongjie, et al.
Veröffentlicht: (2024)
Recycling the Web: A Method to Enhance Pre-training Data Quality and Quantity for Language Models
von: Nguyen, Thao, et al.
Veröffentlicht: (2025)
von: Nguyen, Thao, et al.
Veröffentlicht: (2025)
Scaling Smart: Accelerating Large Language Model Pre-training with Small Model Initialization
von: Samragh, Mohammad, et al.
Veröffentlicht: (2024)
von: Samragh, Mohammad, et al.
Veröffentlicht: (2024)
CDGP: Automatic Cloze Distractor Generation based on Pre-trained Language Model
von: Chiang, Shang-Hsuan, et al.
Veröffentlicht: (2024)
von: Chiang, Shang-Hsuan, et al.
Veröffentlicht: (2024)
IDIAPers @ Causal News Corpus 2022: Efficient Causal Relation Identification Through a Prompt-based Few-shot Approach
von: Burdisso, Sergio, et al.
Veröffentlicht: (2022)
von: Burdisso, Sergio, et al.
Veröffentlicht: (2022)
Pre-training a Transformer-Based Generative Model Using a Small Sepedi Dataset
von: Ramalepe, Simon P., et al.
Veröffentlicht: (2025)
von: Ramalepe, Simon P., et al.
Veröffentlicht: (2025)
Document-Level In-Context Few-Shot Relation Extraction via Pre-Trained Language Models
von: Ozyurt, Yilmazcan, et al.
Veröffentlicht: (2023)
von: Ozyurt, Yilmazcan, et al.
Veröffentlicht: (2023)
ALPET: Active Few-shot Learning for Citation Worthiness Detection in Low-Resource Wikipedia Languages
von: Halitaj, Aida, et al.
Veröffentlicht: (2025)
von: Halitaj, Aida, et al.
Veröffentlicht: (2025)
On the Ability of Transformers to Verify Plans
von: Sarrof, Yash, et al.
Veröffentlicht: (2026)
von: Sarrof, Yash, et al.
Veröffentlicht: (2026)
Combee: Scaling Prompt Learning for Self-Improving Language Model Agents
von: Li, Hanchen, et al.
Veröffentlicht: (2026)
von: Li, Hanchen, et al.
Veröffentlicht: (2026)
Self-Evolving Critique Abilities in Large Language Models
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025)
Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning
von: Xia, Mengzhou, et al.
Veröffentlicht: (2023)
von: Xia, Mengzhou, et al.
Veröffentlicht: (2023)
ParaPO: Aligning Language Models to Reduce Verbatim Reproduction of Pre-training Data
von: Chen, Tong, et al.
Veröffentlicht: (2025)
von: Chen, Tong, et al.
Veröffentlicht: (2025)
Do Large Language Models Have Compositional Ability? An Investigation into Limitations and Scalability
von: Xu, Zhuoyan, et al.
Veröffentlicht: (2024)
von: Xu, Zhuoyan, et al.
Veröffentlicht: (2024)
Bridging Generative and Discriminative Learning: Few-Shot Relation Extraction via Two-Stage Knowledge-Guided Pre-training
von: Guo, Quanjiang, et al.
Veröffentlicht: (2025)
von: Guo, Quanjiang, et al.
Veröffentlicht: (2025)
Unlocking Continual Learning Abilities in Language Models
von: Du, Wenyu, et al.
Veröffentlicht: (2024)
von: Du, Wenyu, et al.
Veröffentlicht: (2024)
On the Reasoning Abilities of Masked Diffusion Language Models
von: Svete, Anej, et al.
Veröffentlicht: (2025)
von: Svete, Anej, et al.
Veröffentlicht: (2025)
Towards Reasoning Ability of Small Language Models
von: Srivastava, Gaurav, et al.
Veröffentlicht: (2025)
von: Srivastava, Gaurav, et al.
Veröffentlicht: (2025)
Deception Abilities Emerged in Large Language Models
von: Hagendorff, Thilo
Veröffentlicht: (2023)
von: Hagendorff, Thilo
Veröffentlicht: (2023)
Ähnliche Einträge
-
Revisiting Chain-of-Thought Prompting: Zero-shot Can Be Stronger than Few-shot
von: Cheng, Xiang, et al.
Veröffentlicht: (2025) -
MemoryPrompt: A Light Wrapper to Improve Context Tracking in Pre-trained Language Models
von: Rakotonirina, Nathanaël Carraz, et al.
Veröffentlicht: (2024) -
Adaptive Few-shot Prompting for Machine Translation with Pre-trained Language Models
von: Tang, Lei, et al.
Veröffentlicht: (2025) -
Evaluating Gender Bias Transfer between Pre-trained and Prompt-Adapted Language Models
von: Mackraz, Natalie, et al.
Veröffentlicht: (2024) -
Relational Prompt-based Pre-trained Language Models for Social Event Detection
von: Li, Pu, et al.
Veröffentlicht: (2024)