Machine Translation with Large Language Models: Decoder Only vs. Encoder-Decoder
Fuente:
arXiv
Guardado en:
| Autores principales: | M., Abhinav P., M, SujayKumar Reddy, Christopher, Oswald |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multilingual Machine Translation with Quantum Encoder Decoder Attention-based Convolutional Variational Circuits
por: Dikshit, Subrit, et al.
Publicado: (2025)
por: Dikshit, Subrit, et al.
Publicado: (2025)
Large Language models for Time Series Analysis: Techniques, Applications, and Challenges
por: Shi, Feifei, et al.
Publicado: (2025)
por: Shi, Feifei, et al.
Publicado: (2025)
Evaluating Large Language Models for Anxiety and Depression Classification using Counseling and Psychotherapy Transcripts
por: Sun, Junwei, et al.
Publicado: (2024)
por: Sun, Junwei, et al.
Publicado: (2024)
Effects of Prompt Length on Domain-specific Tasks for Large Language Models
por: Liu, Qibang, et al.
Publicado: (2025)
por: Liu, Qibang, et al.
Publicado: (2025)
Forgetting That Sticks: Quantization-Permanent Unlearning via Circuit Attribution
por: Sadhu, Saisab, et al.
Publicado: (2026)
por: Sadhu, Saisab, et al.
Publicado: (2026)
InvestAlign: Overcoming Data Scarcity in Aligning Large Language Models with Investor Decision-Making Processes under Herd Behavior
por: Wang, Huisheng, et al.
Publicado: (2025)
por: Wang, Huisheng, et al.
Publicado: (2025)
Embedding-Aligned Language Models
por: Tennenholtz, Guy, et al.
Publicado: (2024)
por: Tennenholtz, Guy, et al.
Publicado: (2024)
Skill Learning Using Process Mining for Large Language Model Plan Generation
por: Redis, Andrei Cosmin, et al.
Publicado: (2024)
por: Redis, Andrei Cosmin, et al.
Publicado: (2024)
Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs
por: Huang, Yiming, et al.
Publicado: (2026)
por: Huang, Yiming, et al.
Publicado: (2026)
Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning
por: Huang, Yiming, et al.
Publicado: (2026)
por: Huang, Yiming, et al.
Publicado: (2026)
Language Modeling on a SpiNNaker 2 Neuromorphic Chip
por: Nazeer, Khaleelulla Khan, et al.
Publicado: (2023)
por: Nazeer, Khaleelulla Khan, et al.
Publicado: (2023)
Combinatorial Reasoning: Selecting Reasons in Generative AI Pipelines via Combinatorial Optimization
por: Esencan, Mert, et al.
Publicado: (2024)
por: Esencan, Mert, et al.
Publicado: (2024)
Automatic Generation of Behavioral Test Cases For Natural Language Processing Using Clustering and Prompting
por: Li, Ying, et al.
Publicado: (2024)
por: Li, Ying, et al.
Publicado: (2024)
Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data
por: Gerstgrasser, Matthias, et al.
Publicado: (2024)
por: Gerstgrasser, Matthias, et al.
Publicado: (2024)
Towards Evaluating Large Language Models for Graph Query Generation
por: Munir, Siraj, et al.
Publicado: (2024)
por: Munir, Siraj, et al.
Publicado: (2024)
Toward Large Language Models as a Therapeutic Tool: Comparing Prompting Techniques to Improve GPT-Delivered Problem-Solving Therapy
por: Filienko, Daniil, et al.
Publicado: (2024)
por: Filienko, Daniil, et al.
Publicado: (2024)
How Many Bytes Can You Take Out Of Brain-To-Text Decoding?
por: Antonello, Richard, et al.
Publicado: (2024)
por: Antonello, Richard, et al.
Publicado: (2024)
The Engineer's Dilemma: A Review of Establishing a Legal Framework for Integrating Machine Learning in Construction by Navigating Precedents and Industry Expectations
por: Naser, M. Z.
Publicado: (2025)
por: Naser, M. Z.
Publicado: (2025)
Spatio-Temporal Pruning for Compressed Spiking Large Language Models
por: Jiang, Yi, et al.
Publicado: (2025)
por: Jiang, Yi, et al.
Publicado: (2025)
LoPT: Low-Rank Prompt Tuning for Parameter Efficient Language Models
por: Guo, Shouchang, et al.
Publicado: (2024)
por: Guo, Shouchang, et al.
Publicado: (2024)
A Cryogenic Memristive Neural Decoder for Fault-tolerant Quantum Error Correction
por: Yon, Victor, et al.
Publicado: (2023)
por: Yon, Victor, et al.
Publicado: (2023)
Comparative Evaluation of Prompting and Fine-Tuning for Applying Large Language Models to Grid-Structured Geospatial Data
por: Dhruv, Akash, et al.
Publicado: (2025)
por: Dhruv, Akash, et al.
Publicado: (2025)
On the Limitations of Compute Thresholds as a Governance Strategy
por: Hooker, Sara
Publicado: (2024)
por: Hooker, Sara
Publicado: (2024)
G-Zero: Self-Play for Open-Ended Generation from Zero Data
por: Huang, Chengsong, et al.
Publicado: (2026)
por: Huang, Chengsong, et al.
Publicado: (2026)
Agentic AI framework for End-to-End Medical Data Inference
por: Shimgekar, Soorya Ram, et al.
Publicado: (2025)
por: Shimgekar, Soorya Ram, et al.
Publicado: (2025)
Graph Repairs with Large Language Models: An Empirical Study
por: Terdalkar, Hrishikesh, et al.
Publicado: (2025)
por: Terdalkar, Hrishikesh, et al.
Publicado: (2025)
Leveraging Large Language Models for Integrated Satellite-Aerial-Terrestrial Networks: Recent Advances and Future Directions
por: Javaid, Shumaila, et al.
Publicado: (2024)
por: Javaid, Shumaila, et al.
Publicado: (2024)
Evaluating Quantized Large Language Models for Code Generation on Low-Resource Language Benchmarks
por: Nyamsuren, Enkhbold
Publicado: (2024)
por: Nyamsuren, Enkhbold
Publicado: (2024)
Quantum Machine Learning in Healthcare: Evaluating QNN and QSVM Models
por: Tudisco, Antonio, et al.
Publicado: (2025)
por: Tudisco, Antonio, et al.
Publicado: (2025)
Which English Do LLMs Prefer? Triangulating Structural Bias Towards American English in Foundation Models
por: Nayeem, Mir Tafseer, et al.
Publicado: (2026)
por: Nayeem, Mir Tafseer, et al.
Publicado: (2026)
Generation of Optimized Solidity Code for Machine Learning Models using LLMs
por: Sham, Nikumbh Sarthak, et al.
Publicado: (2025)
por: Sham, Nikumbh Sarthak, et al.
Publicado: (2025)
Machine Learning Models for Predicting Smoking-Related Health Decline and Disease Risk
por: Chakma, Vaskar, et al.
Publicado: (2025)
por: Chakma, Vaskar, et al.
Publicado: (2025)
Artificial Agency and Large Language Models
por: van Lier, Maud, et al.
Publicado: (2024)
por: van Lier, Maud, et al.
Publicado: (2024)
SpiNNaker2: A Large-Scale Neuromorphic System for Event-Based and Asynchronous Machine Learning
por: Gonzalez, Hector A., et al.
Publicado: (2024)
por: Gonzalez, Hector A., et al.
Publicado: (2024)
Causal Reasoning Favors Encoders: On The Limits of Decoder-Only Models
por: Roy, Amartya, et al.
Publicado: (2025)
por: Roy, Amartya, et al.
Publicado: (2025)
LaScA: Language-Conditioned Scalable Modelling of Affective Dynamics
por: Pinitas, Kosmas, et al.
Publicado: (2026)
por: Pinitas, Kosmas, et al.
Publicado: (2026)
Multi-Faceted Evaluation of Modeling Languages for Augmented Reality Applications -- The Case of ARWFML
por: Muff, Fabian, et al.
Publicado: (2024)
por: Muff, Fabian, et al.
Publicado: (2024)
The Compliance Paradox: Semantic-Instruction Decoupling in Automated Academic Code Evaluation
por: Sahoo, Devanshu, et al.
Publicado: (2026)
por: Sahoo, Devanshu, et al.
Publicado: (2026)
Machine Learning-assisted High-speed Combinatorial Optimization with Ising Machines for Dynamically Changing Problems
por: Hamakawa, Yohei, et al.
Publicado: (2025)
por: Hamakawa, Yohei, et al.
Publicado: (2025)
Performance Analysis of Convolutional Neural Network By Applying Unconstrained Binary Quadratic Programming
por: Sharma, Aasish Kumar, et al.
Publicado: (2025)
por: Sharma, Aasish Kumar, et al.
Publicado: (2025)
Ejemplares similares
-
Multilingual Machine Translation with Quantum Encoder Decoder Attention-based Convolutional Variational Circuits
por: Dikshit, Subrit, et al.
Publicado: (2025) -
Large Language models for Time Series Analysis: Techniques, Applications, and Challenges
por: Shi, Feifei, et al.
Publicado: (2025) -
Evaluating Large Language Models for Anxiety and Depression Classification using Counseling and Psychotherapy Transcripts
por: Sun, Junwei, et al.
Publicado: (2024) -
Effects of Prompt Length on Domain-specific Tasks for Large Language Models
por: Liu, Qibang, et al.
Publicado: (2025) -
Forgetting That Sticks: Quantization-Permanent Unlearning via Circuit Attribution
por: Sadhu, Saisab, et al.
Publicado: (2026)