SIP: Injecting a Structural Inductive Bias into a Seq2Seq Model by Simulation
Fuente:
arXiv
Saved in:
| Main Authors: | Lindemann, Matthias, Koller, Alexander, Titov, Ivan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Strengthening Structural Inductive Biases by Pre-training to Perform Syntactic Transformations
by: Lindemann, Matthias, et al.
Published: (2024)
by: Lindemann, Matthias, et al.
Published: (2024)
Bidirectional Awareness Induction in Autoregressive Seq2Seq Models
by: Hu, Jia Cheng, et al.
Published: (2024)
by: Hu, Jia Cheng, et al.
Published: (2024)
Leveraging Graph Structure in Seq2Seq Models for Knowledge Graph Link Prediction
by: Phuc, Luu Huu, et al.
Published: (2026)
by: Phuc, Luu Huu, et al.
Published: (2026)
UniGenCoder: Merging Seq2Seq and Seq2Tree Paradigms for Unified Code Generation
by: Shao, Liangying, et al.
Published: (2025)
by: Shao, Liangying, et al.
Published: (2025)
Exploiting the Potential of Seq2Seq Models as Robust Few-Shot Learners
by: Lee, Jihyeon, et al.
Published: (2023)
by: Lee, Jihyeon, et al.
Published: (2023)
Enhancing Instruction-Following Capabilities in Seq2Seq Models: DoLA Adaptations for T5
by: Sun, Huey, et al.
Published: (2025)
by: Sun, Huey, et al.
Published: (2025)
Seq2Seq2Seq: Lossless Data Compression via Discrete Latent Transformers and Reinforcement Learning
by: Khodabandeh, Mahdi, et al.
Published: (2026)
by: Khodabandeh, Mahdi, et al.
Published: (2026)
American Sign Language to Text Translation using Transformer and Seq2Seq with LSTM
by: Putra, Gregorius Guntur Sunardi, et al.
Published: (2024)
by: Putra, Gregorius Guntur Sunardi, et al.
Published: (2024)
Seq2Seq Model-Based Chatbot with LSTM and Attention Mechanism for Enhanced User Interaction
by: Benaddi, Lamya, et al.
Published: (2024)
by: Benaddi, Lamya, et al.
Published: (2024)
Cache & Distil: Optimising API Calls to Large Language Models
by: Ramírez, Guillem, et al.
Published: (2023)
by: Ramírez, Guillem, et al.
Published: (2023)
Still Not There: Can LLMs Outperform Smaller Task-Specific Seq2Seq Models on the Poetry-to-Prose Conversion Task?
by: Das, Kunal Kingkar, et al.
Published: (2025)
by: Das, Kunal Kingkar, et al.
Published: (2025)
Seq vs Seq: An Open Suite of Paired Encoders and Decoders
by: Weller, Orion, et al.
Published: (2025)
by: Weller, Orion, et al.
Published: (2025)
Reference-Free Reinforcement Learning Fine-Tuning for MT: A Seq2Seq Perspective
by: Garcia-Estrada, Ernesto, et al.
Published: (2026)
by: Garcia-Estrada, Ernesto, et al.
Published: (2026)
Improving Bangla Linguistics: Advanced LSTM, Bi-LSTM, and Seq2Seq Models for Translating Sylheti to Modern Bangla
by: Das, Sourav Kumar, et al.
Published: (2025)
by: Das, Sourav Kumar, et al.
Published: (2025)
Learning Transductions and Alignments with RNN Seq2seq Models
by: Wang, Zhengxiang
Published: (2023)
by: Wang, Zhengxiang
Published: (2023)
Enhancing Emotion Prediction in News Headlines: Insights from ChatGPT and Seq2Seq Models for Free-Text Generation
by: Gao, Ge, et al.
Published: (2024)
by: Gao, Ge, et al.
Published: (2024)
Sinhala Transliteration: A Comparative Analysis Between Rule-based and Seq2Seq Approaches
by: De Mel, Yomal, et al.
Published: (2024)
by: De Mel, Yomal, et al.
Published: (2024)
Efficient Seq2seq Coreference Resolution Using Entity Representations
by: Grenander, Matt, et al.
Published: (2025)
by: Grenander, Matt, et al.
Published: (2025)
ReaSeq: Unleashing World Knowledge via Reasoning for Sequential Modeling
by: Tang, Jiakai, et al.
Published: (2025)
by: Tang, Jiakai, et al.
Published: (2025)
SeqPE: Transformer with Sequential Position Encoding
by: Li, Huayang, et al.
Published: (2025)
by: Li, Huayang, et al.
Published: (2025)
Mitigating Copy Bias in In-Context Learning through Neuron Pruning
by: Ali, Ameen, et al.
Published: (2024)
by: Ali, Ameen, et al.
Published: (2024)
OptiSeq: Ordering Examples On-The-Fly for In-Context Learning
by: Bhope, Rahul Atul, et al.
Published: (2025)
by: Bhope, Rahul Atul, et al.
Published: (2025)
Seq-VCR: Preventing Collapse in Intermediate Transformer Representations for Enhanced Reasoning
by: Arefin, Md Rifat, et al.
Published: (2024)
by: Arefin, Md Rifat, et al.
Published: (2024)
MoE-DiffuSeq: Enhancing Long-Document Diffusion Models with Sparse Attention and Mixture of Experts
by: Christoforos, Alexandros, et al.
Published: (2025)
by: Christoforos, Alexandros, et al.
Published: (2025)
SeqAR: Jailbreak LLMs with Sequential Auto-Generated Characters
by: Yang, Yan, et al.
Published: (2024)
by: Yang, Yan, et al.
Published: (2024)
SepSeq: A Training-Free Framework for Long Numerical Sequence Processing in LLMs
by: Sun, Jie, et al.
Published: (2026)
by: Sun, Jie, et al.
Published: (2026)
DocNet: Semantic Structure in Inductive Bias Detection Models
by: Zhu, Jessica, et al.
Published: (2024)
by: Zhu, Jessica, et al.
Published: (2024)
Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training
by: Rahmati, Elnaz, et al.
Published: (2026)
by: Rahmati, Elnaz, et al.
Published: (2026)
Information Locality as an Inductive Bias for Neural Language Models
by: Someya, Taiga, et al.
Published: (2025)
by: Someya, Taiga, et al.
Published: (2025)
Unlearning Traces the Influential Training Data of Language Models
by: Isonuma, Masaru, et al.
Published: (2024)
by: Isonuma, Masaru, et al.
Published: (2024)
Quantum-Enhanced Temporal Embeddings via a Hybrid Seq2Seq Architecture
by: Hsieh, Tien-Ching, et al.
Published: (2026)
by: Hsieh, Tien-Ching, et al.
Published: (2026)
SeqPO-SiMT: Sequential Policy Optimization for Simultaneous Machine Translation
by: Xu, Ting, et al.
Published: (2025)
by: Xu, Ting, et al.
Published: (2025)
SPA: Towards A Computational Friendly Cloud-Base and On-Devices Collaboration Seq2seq Personalized Generation with Casual Inference
by: Liu, Yanming, et al.
Published: (2024)
by: Liu, Yanming, et al.
Published: (2024)
SA-DiffuSeq: Addressing Computational and Scalability Challenges in Long-Document Generation with Sparse Attention
by: Christoforos, Alexandros, et al.
Published: (2025)
by: Christoforos, Alexandros, et al.
Published: (2025)
Generalisation First, Memorisation Second? Memorisation Localisation for Natural Language Classification Tasks
by: Dankers, Verna, et al.
Published: (2024)
by: Dankers, Verna, et al.
Published: (2024)
Optimising Calls to Large Language Models with Uncertainty-Based Two-Tier Selection
by: Ramírez, Guillem, et al.
Published: (2024)
by: Ramírez, Guillem, et al.
Published: (2024)
Inductive Bias Extraction and Matching for LLM Prompts
by: Angel, Christian M., et al.
Published: (2025)
by: Angel, Christian M., et al.
Published: (2025)
Post-hoc Reward Calibration: A Case Study on Length Bias
by: Huang, Zeyu, et al.
Published: (2024)
by: Huang, Zeyu, et al.
Published: (2024)
Predicting generalization performance with correctness discriminators
by: Yao, Yuekun, et al.
Published: (2023)
by: Yao, Yuekun, et al.
Published: (2023)
Simple and effective data augmentation for compositional generalization
by: Yao, Yuekun, et al.
Published: (2024)
by: Yao, Yuekun, et al.
Published: (2024)
Similar Items
-
Strengthening Structural Inductive Biases by Pre-training to Perform Syntactic Transformations
by: Lindemann, Matthias, et al.
Published: (2024) -
Bidirectional Awareness Induction in Autoregressive Seq2Seq Models
by: Hu, Jia Cheng, et al.
Published: (2024) -
Leveraging Graph Structure in Seq2Seq Models for Knowledge Graph Link Prediction
by: Phuc, Luu Huu, et al.
Published: (2026) -
UniGenCoder: Merging Seq2Seq and Seq2Tree Paradigms for Unified Code Generation
by: Shao, Liangying, et al.
Published: (2025) -
Exploiting the Potential of Seq2Seq Models as Robust Few-Shot Learners
by: Lee, Jihyeon, et al.
Published: (2023)