EIT: Enhanced Interactive Transformer
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Tong, Li, Bei, Bao, Huiwen, Xiao, Tong, Zhu, Jingbo |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PartialFormer: Modeling Part Instead of Whole for Machine Translation
by: Zheng, Tong, et al.
Published: (2023)
by: Zheng, Tong, et al.
Published: (2023)
Early Exit Is a Natural Capability in Transformer-based Models: An Empirical Study on Early Exit without Joint Optimization
by: Shan, Weiqiao, et al.
Published: (2024)
by: Shan, Weiqiao, et al.
Published: (2024)
Predictor-Corrector Enhanced Transformers with Exponential Moving Average Coefficient Learning
by: Li, Bei, et al.
Published: (2024)
by: Li, Bei, et al.
Published: (2024)
Foundations of Large Language Models
by: Xiao, Tong, et al.
Published: (2025)
by: Xiao, Tong, et al.
Published: (2025)
Asymmetric Conflict and Synergy in Post-training for LLM-based Multilingual Machine Translation
by: Zheng, Tong, et al.
Published: (2025)
by: Zheng, Tong, et al.
Published: (2025)
Translate-and-Revise: Boosting Large Language Models for Constrained Translation
by: Huang, Pengcheng, et al.
Published: (2024)
by: Huang, Pengcheng, et al.
Published: (2024)
IIET: Efficient Numerical Transformer via Implicit Iterative Euler Method
by: Liu, Xinyu, et al.
Published: (2025)
by: Liu, Xinyu, et al.
Published: (2025)
Forgetting Curve: A Reliable Method for Evaluating Memorization Capability for Long-context Models
by: Liu, Xinyu, et al.
Published: (2024)
by: Liu, Xinyu, et al.
Published: (2024)
Beyond Decoder-only: Large Language Models Can be Good Encoders for Machine Translation
by: Luo, Yingfeng, et al.
Published: (2025)
by: Luo, Yingfeng, et al.
Published: (2025)
Language-Specific Layer Matters: Efficient Multilingual Enhancement for Large Vision-Language Models
by: Fan, Yuchun, et al.
Published: (2025)
by: Fan, Yuchun, et al.
Published: (2025)
Hybrid Alignment Training for Large Language Models
by: Wang, Chenglong, et al.
Published: (2024)
by: Wang, Chenglong, et al.
Published: (2024)
FLEXI: Benchmarking Full-duplex Human-LLM Speech Interaction
by: Ge, Yuan, et al.
Published: (2025)
by: Ge, Yuan, et al.
Published: (2025)
Position IDs Matter: An Enhanced Position Layout for Efficient Context Compression in Large Language Models
by: Zhao, Runsong, et al.
Published: (2024)
by: Zhao, Runsong, et al.
Published: (2024)
NiuTrans.LMT: Toward Inclusive and Scalable Multilingual Machine Translation with LLMs
by: Luo, Yingfeng, et al.
Published: (2025)
by: Luo, Yingfeng, et al.
Published: (2025)
Dissecting Long-Chain-of-Thought Reasoning Models: An Empirical Study
by: Mu, Yongyu, et al.
Published: (2025)
by: Mu, Yongyu, et al.
Published: (2025)
RouteLMT: Learned Sample Routing for Hybrid LLM Translation Deployment
by: Luo, Yingfeng, et al.
Published: (2026)
by: Luo, Yingfeng, et al.
Published: (2026)
EfficientGraph-RAG: Structured Retrieval-State Management for Cross-Task Retrieval-Augmented Generation
by: Niu, Miaohe, et al.
Published: (2026)
by: Niu, Miaohe, et al.
Published: (2026)
Teaching Language Models to Self-Improve by Learning from Language Feedback
by: Hu, Chi, et al.
Published: (2024)
by: Hu, Chi, et al.
Published: (2024)
One Size Does Not Fit All: A Distribution-Aware Sparsification for More Precise Model Merging
by: Luo, Yingfeng, et al.
Published: (2025)
by: Luo, Yingfeng, et al.
Published: (2025)
MRO: Enhancing Reasoning in Diffusion Language Models via Multi-Reward Optimization
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
StyleBench: Evaluating Speech Language Models on Conversational Speaking Style Control
by: Zhao, Haishu, et al.
Published: (2026)
by: Zhao, Haishu, et al.
Published: (2026)
Prior Constraints-based Reward Model Training for Aligning Large Language Models
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
NDP: Next Distribution Prediction as a More Broad Target
by: Ruan, Junhao, et al.
Published: (2024)
by: Ruan, Junhao, et al.
Published: (2024)
Revealing the Parallel Multilingual Learning within Large Language Models
by: Mu, Yongyu, et al.
Published: (2024)
by: Mu, Yongyu, et al.
Published: (2024)
MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks
by: Ruan, Junhao, et al.
Published: (2026)
by: Ruan, Junhao, et al.
Published: (2026)
SLAM: Towards Efficient Multilingual Reasoning via Selective Language Alignment
by: Fan, Yuchun, et al.
Published: (2025)
by: Fan, Yuchun, et al.
Published: (2025)
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance
by: Fan, Yuchun, et al.
Published: (2026)
by: Fan, Yuchun, et al.
Published: (2026)
Efficient Prompting Methods for Large Language Models: A Survey
by: Chang, Kaiyan, et al.
Published: (2024)
by: Chang, Kaiyan, et al.
Published: (2024)
RankPrompt: Step-by-Step Comparisons Make Language Models Better Reasoners
by: Hu, Chi, et al.
Published: (2024)
by: Hu, Chi, et al.
Published: (2024)
SUBQRAG: Sub-Question Driven Dynamic Graph RAG
by: Li, Jiaoyang, et al.
Published: (2025)
by: Li, Jiaoyang, et al.
Published: (2025)
GRAM: A Generative Foundation Reward Model for Reward Generalization
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
Clustering and Ranking: Diversity-preserved Instruction Selection through Expert-aligned Quality Estimation
by: Ge, Yuan, et al.
Published: (2024)
by: Ge, Yuan, et al.
Published: (2024)
DaPT: A Dual-Path Framework for Multilingual Multi-hop Question Answering
by: Wang, Yilin, et al.
Published: (2026)
by: Wang, Yilin, et al.
Published: (2026)
Parallel-R1: Towards Parallel Thinking via Reinforcement Learning
by: Zheng, Tong, et al.
Published: (2025)
by: Zheng, Tong, et al.
Published: (2025)
Enhancing Adversarial Attacks through Chain of Thought
by: Su, Jingbo
Published: (2024)
by: Su, Jingbo
Published: (2024)
On the Emotion Understanding of Synthesized Speech
by: Ge, Yuan, et al.
Published: (2026)
by: Ge, Yuan, et al.
Published: (2026)
Probing Preference Representations: A Multi-Dimensional Evaluation and Analysis Method for Reward Models
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
Optimizing Speech Multi-View Feature Fusion through Conditional Computation
by: Shan, Weiqiao, et al.
Published: (2025)
by: Shan, Weiqiao, et al.
Published: (2025)
Parallel-Probe: Towards Efficient Parallel Thinking via 2D Probing
by: Zheng, Tong, et al.
Published: (2026)
by: Zheng, Tong, et al.
Published: (2026)
Enhancing Abstractive Summarization of Scientific Papers Using Structure Information
by: Bao, Tong, et al.
Published: (2025)
by: Bao, Tong, et al.
Published: (2025)
Similar Items
-
PartialFormer: Modeling Part Instead of Whole for Machine Translation
by: Zheng, Tong, et al.
Published: (2023) -
Early Exit Is a Natural Capability in Transformer-based Models: An Empirical Study on Early Exit without Joint Optimization
by: Shan, Weiqiao, et al.
Published: (2024) -
Predictor-Corrector Enhanced Transformers with Exponential Moving Average Coefficient Learning
by: Li, Bei, et al.
Published: (2024) -
Foundations of Large Language Models
by: Xiao, Tong, et al.
Published: (2025) -
Asymmetric Conflict and Synergy in Post-training for LLM-based Multilingual Machine Translation
by: Zheng, Tong, et al.
Published: (2025)