An In-depth Walkthrough on Evolution of Neural Machine Translation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jagtap, Rohan, Dhage, Sudhir N. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2020
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Neural Information Organizing and Processing -- Neural Machines
von: Petrila, Iosif Iulian
Veröffentlicht: (2024)
von: Petrila, Iosif Iulian
Veröffentlicht: (2024)
Large Language Models for Tuning Evolution Strategies
von: Kramer, Oliver
Veröffentlicht: (2024)
von: Kramer, Oliver
Veröffentlicht: (2024)
EvoX: Meta-Evolution for Automated Discovery
von: Liu, Shu, et al.
Veröffentlicht: (2026)
von: Liu, Shu, et al.
Veröffentlicht: (2026)
Intelligent Neural Networks: From Layered Architectures to Graph-Organized Intelligence
von: Salomon, Antoine
Veröffentlicht: (2025)
von: Salomon, Antoine
Veröffentlicht: (2025)
SpikeGPT: Generative Pre-trained Language Model with Spiking Neural Networks
von: Zhu, Rui-Jie, et al.
Veröffentlicht: (2023)
von: Zhu, Rui-Jie, et al.
Veröffentlicht: (2023)
SpikeLLM: Scaling up Spiking Neural Network to Large Language Models via Saliency-based Spiking
von: Xing, Xingrun, et al.
Veröffentlicht: (2024)
von: Xing, Xingrun, et al.
Veröffentlicht: (2024)
The Evolution of Learning Algorithms for Artificial Neural Networks
von: Baxter, Jonathan
Veröffentlicht: (2025)
von: Baxter, Jonathan
Veröffentlicht: (2025)
Multiple Population Alternate Evolution Neural Architecture Search
von: Zou, Juan, et al.
Veröffentlicht: (2024)
von: Zou, Juan, et al.
Veröffentlicht: (2024)
Multi-Class Imbalanced Learning with Support Vector Machines via Differential Evolution
von: Zhang, Zhong-Liang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhong-Liang, et al.
Veröffentlicht: (2025)
Language Models and Cycle Consistency for Self-Reflective Machine Translation
von: Wangni, Jianqiao
Veröffentlicht: (2024)
von: Wangni, Jianqiao
Veröffentlicht: (2024)
Learning Numeracy: Binary Arithmetic with Neural Turing Machines
von: Castellini, Jacopo
Veröffentlicht: (2019)
von: Castellini, Jacopo
Veröffentlicht: (2019)
Towards Faster k-Nearest-Neighbor Machine Translation
von: Shi, Xiangyu, et al.
Veröffentlicht: (2023)
von: Shi, Xiangyu, et al.
Veröffentlicht: (2023)
HAT: Hardware-Aware Transformers for Efficient Natural Language Processing
von: Wang, Hanrui, et al.
Veröffentlicht: (2020)
von: Wang, Hanrui, et al.
Veröffentlicht: (2020)
EvolKV: Evolutionary KV Cache Compression for LLM Inference
von: Yu, Bohan, et al.
Veröffentlicht: (2025)
von: Yu, Bohan, et al.
Veröffentlicht: (2025)
Decomposing Evolutionary Mixture-of-LoRA Architectures: The Routing Lever, the Lifecycle Penalty, and a Substrate-Conditional Boundary
von: Kumaresan, Ramchand
Veröffentlicht: (2026)
von: Kumaresan, Ramchand
Veröffentlicht: (2026)
Pruner-Zero: Evolving Symbolic Pruning Metric from scratch for Large Language Models
von: Dong, Peijie, et al.
Veröffentlicht: (2024)
von: Dong, Peijie, et al.
Veröffentlicht: (2024)
SpikeLM: Towards General Spike-Driven Language Modeling via Elastic Bi-Spiking Mechanisms
von: Xing, Xingrun, et al.
Veröffentlicht: (2024)
von: Xing, Xingrun, et al.
Veröffentlicht: (2024)
Hysteresis Activation Function for Efficient Inference
von: Kimhi, Moshe, et al.
Veröffentlicht: (2024)
von: Kimhi, Moshe, et al.
Veröffentlicht: (2024)
Large Language Models Suffer From Their Own Output: An Analysis of the Self-Consuming Training Loop
von: Briesch, Martin, et al.
Veröffentlicht: (2023)
von: Briesch, Martin, et al.
Veröffentlicht: (2023)
Genetic Instruct: Scaling up Synthetic Generation of Coding Instructions for Large Language Models
von: Majumdar, Somshubra, et al.
Veröffentlicht: (2024)
von: Majumdar, Somshubra, et al.
Veröffentlicht: (2024)
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
von: Kadlčík, Marek, et al.
Veröffentlicht: (2025)
von: Kadlčík, Marek, et al.
Veröffentlicht: (2025)
AP-BMM: Approximating Capability-Cost Pareto Sets of LLMs via Asynchronous Prior-Guided Bayesian Model Merging
von: Chen, Kesheng, et al.
Veröffentlicht: (2025)
von: Chen, Kesheng, et al.
Veröffentlicht: (2025)
On the Power of Convolution Augmented Transformer
von: Li, Mingchen, et al.
Veröffentlicht: (2024)
von: Li, Mingchen, et al.
Veröffentlicht: (2024)
A Hormone-inspired Emotion Layer for Transformer language models (HELT)
von: Reda, Eslam, et al.
Veröffentlicht: (2026)
von: Reda, Eslam, et al.
Veröffentlicht: (2026)
SpikingSSMs: Learning Long Sequences with Sparse and Parallel Spiking State Space Models
von: Shen, Shuaijie, et al.
Veröffentlicht: (2024)
von: Shen, Shuaijie, et al.
Veröffentlicht: (2024)
SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention
von: Csordás, Róbert, et al.
Veröffentlicht: (2023)
von: Csordás, Róbert, et al.
Veröffentlicht: (2023)
Sorbet: A Neuromorphic Hardware-Compatible Transformer-Based Spiking Language Model
von: Tang, Kaiwen, et al.
Veröffentlicht: (2024)
von: Tang, Kaiwen, et al.
Veröffentlicht: (2024)
An enhanced Teaching-Learning-Based Optimization (TLBO) with Grey Wolf Optimizer (GWO) for text feature selection and clustering
von: Azarshab, Mahsa, et al.
Veröffentlicht: (2024)
von: Azarshab, Mahsa, et al.
Veröffentlicht: (2024)
BrainTransformers: SNN-LLM
von: Tang, Zhengzheng, et al.
Veröffentlicht: (2024)
von: Tang, Zhengzheng, et al.
Veröffentlicht: (2024)
ComplicaCode: Enhancing Disease Complication Detection in Electronic Health Records through ICD Path Generation
von: Zhou, Xiaofan
Veröffentlicht: (2023)
von: Zhou, Xiaofan
Veröffentlicht: (2023)
Why Prompt Optimization Works, and Why It Sometimes Doesn't: A Causal-Inspired Edit-Level Analysis
von: Gong, Shuzhi, et al.
Veröffentlicht: (2026)
von: Gong, Shuzhi, et al.
Veröffentlicht: (2026)
Improving Sequence-to-Sequence Models for Abstractive Text Summarization Using Meta Heuristic Approaches
von: Saxena, Aditya, et al.
Veröffentlicht: (2024)
von: Saxena, Aditya, et al.
Veröffentlicht: (2024)
Improving Language Plasticity via Pretraining with Active Forgetting
von: Chen, Yihong, et al.
Veröffentlicht: (2023)
von: Chen, Yihong, et al.
Veröffentlicht: (2023)
B'MOJO: Hybrid State Space Realizations of Foundation Models with Eidetic and Fading Memory
von: Zancato, Luca, et al.
Veröffentlicht: (2024)
von: Zancato, Luca, et al.
Veröffentlicht: (2024)
Differential Evolution Algorithm based Hyper-Parameters Selection of Transformer Neural Network Model for Load Forecasting
von: Sen, Anuvab, et al.
Veröffentlicht: (2023)
von: Sen, Anuvab, et al.
Veröffentlicht: (2023)
A Transformer-based Neural Architecture Search Method
von: Wang, Shang, et al.
Veröffentlicht: (2025)
von: Wang, Shang, et al.
Veröffentlicht: (2025)
Rethinking Deep Learning: Non-backpropagation and Non-optimization Machine Learning Approach Using Hebbian Neural Networks
von: Itoh, Kei
Veröffentlicht: (2024)
von: Itoh, Kei
Veröffentlicht: (2024)
A Gauge Theory of Superposition: Toward a Sheaf-Theoretic Atlas of Neural Representations
von: Javidnia, Hossein
Veröffentlicht: (2026)
von: Javidnia, Hossein
Veröffentlicht: (2026)
The Stacked Autoencoder Evolution Hypothesis
von: Iizuka, Hiroyuki
Veröffentlicht: (2026)
von: Iizuka, Hiroyuki
Veröffentlicht: (2026)
Gated Recurrent Neural Networks with Weighted Time-Delay Feedback
von: Erichson, N. Benjamin, et al.
Veröffentlicht: (2022)
von: Erichson, N. Benjamin, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Neural Information Organizing and Processing -- Neural Machines
von: Petrila, Iosif Iulian
Veröffentlicht: (2024) -
Large Language Models for Tuning Evolution Strategies
von: Kramer, Oliver
Veröffentlicht: (2024) -
EvoX: Meta-Evolution for Automated Discovery
von: Liu, Shu, et al.
Veröffentlicht: (2026) -
Intelligent Neural Networks: From Layered Architectures to Graph-Organized Intelligence
von: Salomon, Antoine
Veröffentlicht: (2025) -
SpikeGPT: Generative Pre-trained Language Model with Spiking Neural Networks
von: Zhu, Rui-Jie, et al.
Veröffentlicht: (2023)