Bidirectional Awareness Induction in Autoregressive Seq2Seq Models
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Jia Cheng, Cavicchioli, Roberto, Capotondi, Alessandro |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diffusion Is Your Friend in Show, Suggest and Tell
by: Hu, Jia Cheng, et al.
Published: (2025)
by: Hu, Jia Cheng, et al.
Published: (2025)
Heterogeneous Encoders Scaling In The Transformer For Neural Machine Translation
by: Hu, Jia Cheng, et al.
Published: (2023)
by: Hu, Jia Cheng, et al.
Published: (2023)
UniGenCoder: Merging Seq2Seq and Seq2Tree Paradigms for Unified Code Generation
by: Shao, Liangying, et al.
Published: (2025)
by: Shao, Liangying, et al.
Published: (2025)
Shifted Window Fourier Transform And Retention For Image Captioning
by: Hu, Jia Cheng, et al.
Published: (2024)
by: Hu, Jia Cheng, et al.
Published: (2024)
Exploiting Multiple Sequence Lengths in Fast End to End Training for Image Captioning
by: Hu, Jia Cheng, et al.
Published: (2022)
by: Hu, Jia Cheng, et al.
Published: (2022)
Exploiting the Potential of Seq2Seq Models as Robust Few-Shot Learners
by: Lee, Jihyeon, et al.
Published: (2023)
by: Lee, Jihyeon, et al.
Published: (2023)
SIP: Injecting a Structural Inductive Bias into a Seq2Seq Model by Simulation
by: Lindemann, Matthias, et al.
Published: (2023)
by: Lindemann, Matthias, et al.
Published: (2023)
Enhancing Instruction-Following Capabilities in Seq2Seq Models: DoLA Adaptations for T5
by: Sun, Huey, et al.
Published: (2025)
by: Sun, Huey, et al.
Published: (2025)
Leveraging Graph Structure in Seq2Seq Models for Knowledge Graph Link Prediction
by: Phuc, Luu Huu, et al.
Published: (2026)
by: Phuc, Luu Huu, et al.
Published: (2026)
Seq2Seq2Seq: Lossless Data Compression via Discrete Latent Transformers and Reinforcement Learning
by: Khodabandeh, Mahdi, et al.
Published: (2026)
by: Khodabandeh, Mahdi, et al.
Published: (2026)
American Sign Language to Text Translation using Transformer and Seq2Seq with LSTM
by: Putra, Gregorius Guntur Sunardi, et al.
Published: (2024)
by: Putra, Gregorius Guntur Sunardi, et al.
Published: (2024)
Seq2Seq Model-Based Chatbot with LSTM and Attention Mechanism for Enhanced User Interaction
by: Benaddi, Lamya, et al.
Published: (2024)
by: Benaddi, Lamya, et al.
Published: (2024)
Still Not There: Can LLMs Outperform Smaller Task-Specific Seq2Seq Models on the Poetry-to-Prose Conversion Task?
by: Das, Kunal Kingkar, et al.
Published: (2025)
by: Das, Kunal Kingkar, et al.
Published: (2025)
Seq vs Seq: An Open Suite of Paired Encoders and Decoders
by: Weller, Orion, et al.
Published: (2025)
by: Weller, Orion, et al.
Published: (2025)
Reference-Free Reinforcement Learning Fine-Tuning for MT: A Seq2Seq Perspective
by: Garcia-Estrada, Ernesto, et al.
Published: (2026)
by: Garcia-Estrada, Ernesto, et al.
Published: (2026)
Improving Bangla Linguistics: Advanced LSTM, Bi-LSTM, and Seq2Seq Models for Translating Sylheti to Modern Bangla
by: Das, Sourav Kumar, et al.
Published: (2025)
by: Das, Sourav Kumar, et al.
Published: (2025)
Learning Transductions and Alignments with RNN Seq2seq Models
by: Wang, Zhengxiang
Published: (2023)
by: Wang, Zhengxiang
Published: (2023)
Enhancing Emotion Prediction in News Headlines: Insights from ChatGPT and Seq2Seq Models for Free-Text Generation
by: Gao, Ge, et al.
Published: (2024)
by: Gao, Ge, et al.
Published: (2024)
Sinhala Transliteration: A Comparative Analysis Between Rule-based and Seq2Seq Approaches
by: De Mel, Yomal, et al.
Published: (2024)
by: De Mel, Yomal, et al.
Published: (2024)
ReaSeq: Unleashing World Knowledge via Reasoning for Sequential Modeling
by: Tang, Jiakai, et al.
Published: (2025)
by: Tang, Jiakai, et al.
Published: (2025)
Efficient Seq2seq Coreference Resolution Using Entity Representations
by: Grenander, Matt, et al.
Published: (2025)
by: Grenander, Matt, et al.
Published: (2025)
SeqPE: Transformer with Sequential Position Encoding
by: Li, Huayang, et al.
Published: (2025)
by: Li, Huayang, et al.
Published: (2025)
SeqPO-SiMT: Sequential Policy Optimization for Simultaneous Machine Translation
by: Xu, Ting, et al.
Published: (2025)
by: Xu, Ting, et al.
Published: (2025)
OptiSeq: Ordering Examples On-The-Fly for In-Context Learning
by: Bhope, Rahul Atul, et al.
Published: (2025)
by: Bhope, Rahul Atul, et al.
Published: (2025)
Seq-VCR: Preventing Collapse in Intermediate Transformer Representations for Enhanced Reasoning
by: Arefin, Md Rifat, et al.
Published: (2024)
by: Arefin, Md Rifat, et al.
Published: (2024)
MoE-DiffuSeq: Enhancing Long-Document Diffusion Models with Sparse Attention and Mixture of Experts
by: Christoforos, Alexandros, et al.
Published: (2025)
by: Christoforos, Alexandros, et al.
Published: (2025)
SeqAR: Jailbreak LLMs with Sequential Auto-Generated Characters
by: Yang, Yan, et al.
Published: (2024)
by: Yang, Yan, et al.
Published: (2024)
SepSeq: A Training-Free Framework for Long Numerical Sequence Processing in LLMs
by: Sun, Jie, et al.
Published: (2026)
by: Sun, Jie, et al.
Published: (2026)
Quantum-Enhanced Temporal Embeddings via a Hybrid Seq2Seq Architecture
by: Hsieh, Tien-Ching, et al.
Published: (2026)
by: Hsieh, Tien-Ching, et al.
Published: (2026)
SPA: Towards A Computational Friendly Cloud-Base and On-Devices Collaboration Seq2seq Personalized Generation with Casual Inference
by: Liu, Yanming, et al.
Published: (2024)
by: Liu, Yanming, et al.
Published: (2024)
SA-DiffuSeq: Addressing Computational and Scalability Challenges in Long-Document Generation with Sparse Attention
by: Christoforos, Alexandros, et al.
Published: (2025)
by: Christoforos, Alexandros, et al.
Published: (2025)
Blockwise SFT for Diffusion Language Models: Reconciling Bidirectional Attention and Autoregressive Decoding
by: Sun, Bowen, et al.
Published: (2025)
by: Sun, Bowen, et al.
Published: (2025)
A Multimodal Seq2Seq Transformer for Predicting Brain Responses to Naturalistic Stimuli
by: He, Qianyi, et al.
Published: (2025)
by: He, Qianyi, et al.
Published: (2025)
ECHO: Toward Contextual Seq2Seq Paradigms in Large EEG Models
by: Liu, Chenyu, et al.
Published: (2025)
by: Liu, Chenyu, et al.
Published: (2025)
EHR-SeqSQL : A Sequential Text-to-SQL Dataset For Interactively Exploring Electronic Health Records
by: Ryu, Jaehee, et al.
Published: (2024)
by: Ryu, Jaehee, et al.
Published: (2024)
MIXAR: Scaling Autoregressive Pixel-based Language Models to Multiple Languages and Scripts
by: Hu, Chen, et al.
Published: (2026)
by: Hu, Chen, et al.
Published: (2026)
SeqSAM: Autoregressive Multiple Hypothesis Prediction for Medical Image Segmentation using SAM
by: Towle, Benjamin, et al.
Published: (2025)
by: Towle, Benjamin, et al.
Published: (2025)
Neuron-based Personality Trait Induction in Large Language Models
by: Deng, Jia, et al.
Published: (2024)
by: Deng, Jia, et al.
Published: (2024)
P2LHAP:Wearable sensor-based human activity recognition, segmentation and forecast through Patch-to-Label Seq2Seq Transformer
by: Li, Shuangjian, et al.
Published: (2024)
by: Li, Shuangjian, et al.
Published: (2024)
The Bidirectional Process Reward Model
by: Zhang, Lingyin, et al.
Published: (2025)
by: Zhang, Lingyin, et al.
Published: (2025)
Similar Items
-
Diffusion Is Your Friend in Show, Suggest and Tell
by: Hu, Jia Cheng, et al.
Published: (2025) -
Heterogeneous Encoders Scaling In The Transformer For Neural Machine Translation
by: Hu, Jia Cheng, et al.
Published: (2023) -
UniGenCoder: Merging Seq2Seq and Seq2Tree Paradigms for Unified Code Generation
by: Shao, Liangying, et al.
Published: (2025) -
Shifted Window Fourier Transform And Retention For Image Captioning
by: Hu, Jia Cheng, et al.
Published: (2024) -
Exploiting Multiple Sequence Lengths in Fast End to End Training for Image Captioning
by: Hu, Jia Cheng, et al.
Published: (2022)