Exploiting the Potential of Seq2Seq Models as Robust Few-Shot Learners
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Jihyeon, Kim, Dain, Jung, Doohae, Kim, Boseop, On, Kyoung-Woon |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Leveraging Graph Structure in Seq2Seq Models for Knowledge Graph Link Prediction
by: Phuc, Luu Huu, et al.
Published: (2026)
by: Phuc, Luu Huu, et al.
Published: (2026)
Seq2Seq2Seq: Lossless Data Compression via Discrete Latent Transformers and Reinforcement Learning
by: Khodabandeh, Mahdi, et al.
Published: (2026)
by: Khodabandeh, Mahdi, et al.
Published: (2026)
Reference-Free Reinforcement Learning Fine-Tuning for MT: A Seq2Seq Perspective
by: Garcia-Estrada, Ernesto, et al.
Published: (2026)
by: Garcia-Estrada, Ernesto, et al.
Published: (2026)
Improving Bangla Linguistics: Advanced LSTM, Bi-LSTM, and Seq2Seq Models for Translating Sylheti to Modern Bangla
by: Das, Sourav Kumar, et al.
Published: (2025)
by: Das, Sourav Kumar, et al.
Published: (2025)
How Well Do Large Language Models Truly Ground?
by: Lee, Hyunji, et al.
Published: (2023)
by: Lee, Hyunji, et al.
Published: (2023)
LLMs Are Few-Shot In-Context Low-Resource Language Learners
by: Cahyawijaya, Samuel, et al.
Published: (2024)
by: Cahyawijaya, Samuel, et al.
Published: (2024)
SeqAR: Jailbreak LLMs with Sequential Auto-Generated Characters
by: Yang, Yan, et al.
Published: (2024)
by: Yang, Yan, et al.
Published: (2024)
Binary Classifier Optimization for Large Language Model Alignment
by: Jung, Seungjae, et al.
Published: (2024)
by: Jung, Seungjae, et al.
Published: (2024)
OptiSeq: Ordering Examples On-The-Fly for In-Context Learning
by: Bhope, Rahul Atul, et al.
Published: (2025)
by: Bhope, Rahul Atul, et al.
Published: (2025)
SeqPE: Transformer with Sequential Position Encoding
by: Li, Huayang, et al.
Published: (2025)
by: Li, Huayang, et al.
Published: (2025)
SA-DiffuSeq: Addressing Computational and Scalability Challenges in Long-Document Generation with Sparse Attention
by: Christoforos, Alexandros, et al.
Published: (2025)
by: Christoforos, Alexandros, et al.
Published: (2025)
Semiparametric Token-Sequence Co-Supervision
by: Lee, Hyunji, et al.
Published: (2024)
by: Lee, Hyunji, et al.
Published: (2024)
LAPIS: Language Model-Augmented Police Investigation System
by: Kim, Heedou, et al.
Published: (2024)
by: Kim, Heedou, et al.
Published: (2024)
Exploiting Text Semantics for Few and Zero Shot Node Classification on Text-attributed Graph
by: Wang, Yuxiang, et al.
Published: (2025)
by: Wang, Yuxiang, et al.
Published: (2025)
SeqPO-SiMT: Sequential Policy Optimization for Simultaneous Machine Translation
by: Xu, Ting, et al.
Published: (2025)
by: Xu, Ting, et al.
Published: (2025)
ACoRN: Noise-Robust Abstractive Compression in Retrieval-Augmented Language Models
by: Kim, Singon, et al.
Published: (2025)
by: Kim, Singon, et al.
Published: (2025)
Evaluating Legal Reasoning Traces with Legal Issue Tree Rubrics
by: Lee, Jinu, et al.
Published: (2025)
by: Lee, Jinu, et al.
Published: (2025)
EHR-SeqSQL : A Sequential Text-to-SQL Dataset For Interactively Exploring Electronic Health Records
by: Ryu, Jaehee, et al.
Published: (2024)
by: Ryu, Jaehee, et al.
Published: (2024)
Bidirectional Awareness Induction in Autoregressive Seq2Seq Models
by: Hu, Jia Cheng, et al.
Published: (2024)
by: Hu, Jia Cheng, et al.
Published: (2024)
Language Models are Few-Shot Graders
by: Zhao, Chenyan, et al.
Published: (2025)
by: Zhao, Chenyan, et al.
Published: (2025)
PRISP: Privacy-Safe Few-Shot Personalization via Lightweight Adaptation
by: Park, Junho, et al.
Published: (2026)
by: Park, Junho, et al.
Published: (2026)
Focused Large Language Models are Stable Many-Shot Learners
by: Yuan, Peiwen, et al.
Published: (2024)
by: Yuan, Peiwen, et al.
Published: (2024)
UniGenCoder: Merging Seq2Seq and Seq2Tree Paradigms for Unified Code Generation
by: Shao, Liangying, et al.
Published: (2025)
by: Shao, Liangying, et al.
Published: (2025)
Unlocking the Potential of Diffusion Language Models through Template Infilling
by: Lee, Junhoo, et al.
Published: (2025)
by: Lee, Junhoo, et al.
Published: (2025)
Infinite Mask Diffusion for Few-Step Distillation
by: Yoo, Jaehoon, et al.
Published: (2026)
by: Yoo, Jaehoon, et al.
Published: (2026)
Large Language Models are Null-Shot Learners
by: Taveekitworachai, Pittawat, et al.
Published: (2024)
by: Taveekitworachai, Pittawat, et al.
Published: (2024)
Language Model Representations for Efficient Few-Shot Tabular Classification
by: Kang, Inwon, et al.
Published: (2026)
by: Kang, Inwon, et al.
Published: (2026)
Benchmarking Open-Source Large Language Models for Persian in Zero-Shot and Few-Shot Learning
by: Cherakhloo, Mahdi, et al.
Published: (2025)
by: Cherakhloo, Mahdi, et al.
Published: (2025)
Few-Shot Recalibration of Language Models
by: Li, Xiang Lisa, et al.
Published: (2024)
by: Li, Xiang Lisa, et al.
Published: (2024)
The Surprising Effectiveness of Test-Time Training for Few-Shot Learning
by: Akyürek, Ekin, et al.
Published: (2024)
by: Akyürek, Ekin, et al.
Published: (2024)
Hexa: Self-Improving for Knowledge-Grounded Dialogue System
by: Jo, Daejin, et al.
Published: (2023)
by: Jo, Daejin, et al.
Published: (2023)
Modeling the One-to-Many Property in Open-Domain Dialogue with LLMs
by: Lee, Jing Yang, et al.
Published: (2025)
by: Lee, Jing Yang, et al.
Published: (2025)
SIP: Injecting a Structural Inductive Bias into a Seq2Seq Model by Simulation
by: Lindemann, Matthias, et al.
Published: (2023)
by: Lindemann, Matthias, et al.
Published: (2023)
Enhancing Instruction-Following Capabilities in Seq2Seq Models: DoLA Adaptations for T5
by: Sun, Huey, et al.
Published: (2025)
by: Sun, Huey, et al.
Published: (2025)
Noise-Robust Abstractive Compression in Retrieval-Augmented Language Models
by: Kim, Singon
Published: (2025)
by: Kim, Singon
Published: (2025)
FewTopNER: Integrating Few-Shot Learning with Topic Modeling and Named Entity Recognition in a Multilingual Framework
by: Bouabdallaoui, Ibrahim, et al.
Published: (2025)
by: Bouabdallaoui, Ibrahim, et al.
Published: (2025)
Toxicity-Aware Few-Shot Prompting for Low-Resource Singlish Translation
by: Ge, Ziyu, et al.
Published: (2025)
by: Ge, Ziyu, et al.
Published: (2025)
COCOA: CBT-based Conversational Counseling Agent using Memory Specialized in Cognitive Distortions and Dynamic Prompt
by: Lee, Suyeon, et al.
Published: (2024)
by: Lee, Suyeon, et al.
Published: (2024)
Mask-guided BERT for Few Shot Text Classification
by: Liao, Wenxiong, et al.
Published: (2023)
by: Liao, Wenxiong, et al.
Published: (2023)
Paraphrase and Solve: Exploring and Exploiting the Impact of Surface Form on Mathematical Reasoning in Large Language Models
by: Zhou, Yue, et al.
Published: (2024)
by: Zhou, Yue, et al.
Published: (2024)
Similar Items
-
Leveraging Graph Structure in Seq2Seq Models for Knowledge Graph Link Prediction
by: Phuc, Luu Huu, et al.
Published: (2026) -
Seq2Seq2Seq: Lossless Data Compression via Discrete Latent Transformers and Reinforcement Learning
by: Khodabandeh, Mahdi, et al.
Published: (2026) -
Reference-Free Reinforcement Learning Fine-Tuning for MT: A Seq2Seq Perspective
by: Garcia-Estrada, Ernesto, et al.
Published: (2026) -
Improving Bangla Linguistics: Advanced LSTM, Bi-LSTM, and Seq2Seq Models for Translating Sylheti to Modern Bangla
by: Das, Sourav Kumar, et al.
Published: (2025) -
How Well Do Large Language Models Truly Ground?
by: Lee, Hyunji, et al.
Published: (2023)