End-to-End Training for Back-Translation with Categorical Reparameterization Trick
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Heo, DongNyeong, Choi, Heeyoul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Shared Latent Space by Both Languages in Non-Autoregressive Neural Machine Translation
von: Heo, DongNyeong, et al.
Veröffentlicht: (2023)
von: Heo, DongNyeong, et al.
Veröffentlicht: (2023)
Generalized Probabilistic Attention Mechanism in Transformers
von: Heo, DongNyeong, et al.
Veröffentlicht: (2024)
von: Heo, DongNyeong, et al.
Veröffentlicht: (2024)
Sentence Curve Language Models
von: Heo, DongNyeong, et al.
Veröffentlicht: (2026)
von: Heo, DongNyeong, et al.
Veröffentlicht: (2026)
N-gram Prediction and Word Difference Representations for Language Modeling
von: Heo, DongNyeong, et al.
Veröffentlicht: (2024)
von: Heo, DongNyeong, et al.
Veröffentlicht: (2024)
Dynamic Preference Multi-Objective Reinforcement Learning for Internet Network Management
von: Heo, DongNyeong, et al.
Veröffentlicht: (2025)
von: Heo, DongNyeong, et al.
Veröffentlicht: (2025)
Predictive Concept Decoders: Training Scalable End-to-End Interpretability Assistants
von: Huang, Vincent, et al.
Veröffentlicht: (2025)
von: Huang, Vincent, et al.
Veröffentlicht: (2025)
Rationalizing Transformer Predictions via End-To-End Differentiable Self-Training
von: Brinner, Marc, et al.
Veröffentlicht: (2025)
von: Brinner, Marc, et al.
Veröffentlicht: (2025)
End-to-end Planner Training for Language Modeling
von: Cornille, Nathan, et al.
Veröffentlicht: (2024)
von: Cornille, Nathan, et al.
Veröffentlicht: (2024)
Data Augmentation With Back translation for Low Resource languages: A case of English and Luganda
von: Kimera, Richard, et al.
Veröffentlicht: (2025)
von: Kimera, Richard, et al.
Veröffentlicht: (2025)
WebAgent-R1: Training Web Agents via End-to-End Multi-Turn Reinforcement Learning
von: Wei, Zhepei, et al.
Veröffentlicht: (2025)
von: Wei, Zhepei, et al.
Veröffentlicht: (2025)
End-to-End Ontology Learning with Large Language Models
von: Lo, Andy, et al.
Veröffentlicht: (2024)
von: Lo, Andy, et al.
Veröffentlicht: (2024)
An End-to-End Approach for Child Reading Assessment in the Xhosa Language
von: Chevtchenko, Sergio, et al.
Veröffentlicht: (2025)
von: Chevtchenko, Sergio, et al.
Veröffentlicht: (2025)
Reparameterized LLM Training via Orthogonal Equivalence Transformation
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
Enhanced Labeling Technique for Reddit Text and Fine-Tuned Longformer Models for Classifying Depression Severity in English and Luganda
von: Kimera, Richard, et al.
Veröffentlicht: (2024)
von: Kimera, Richard, et al.
Veröffentlicht: (2024)
POSESTITCH-SLT: Linguistically Inspired Pose-Stitching for End-to-End Sign Language Translation
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
Autonoma: A Hierarchical Multi-Agent Framework for End-to-End Workflow Automation
von: Reda, Eslam, et al.
Veröffentlicht: (2026)
von: Reda, Eslam, et al.
Veröffentlicht: (2026)
MATATA: Weakly Supervised End-to-End MAthematical Tool-Augmented Reasoning for Tabular Applications
von: Vinayagame, Vishnou, et al.
Veröffentlicht: (2024)
von: Vinayagame, Vishnou, et al.
Veröffentlicht: (2024)
Janus-Q: End-to-End Event-Driven Trading via Hierarchical-Gated Reward Modeling
von: Li, Xiang, et al.
Veröffentlicht: (2026)
von: Li, Xiang, et al.
Veröffentlicht: (2026)
Which Side Are You On? A Multi-task Dataset for End-to-End Argument Summarisation and Evaluation
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
Integrating Pre-Trained Speech and Language Models for End-to-End Speech Recognition
von: Hono, Yukiya, et al.
Veröffentlicht: (2023)
von: Hono, Yukiya, et al.
Veröffentlicht: (2023)
FLAME-MoE: A Transparent End-to-End Research Platform for Mixture-of-Experts Language Models
von: Kang, Hao, et al.
Veröffentlicht: (2025)
von: Kang, Hao, et al.
Veröffentlicht: (2025)
LKV: End-to-End Learning of Head-wise Budgets and Token Selection for LLM KV Cache Eviction
von: Zhou, Enshuai, et al.
Veröffentlicht: (2026)
von: Zhou, Enshuai, et al.
Veröffentlicht: (2026)
End-To-End Causal Effect Estimation from Unstructured Natural Language Data
von: Dhawan, Nikita, et al.
Veröffentlicht: (2024)
von: Dhawan, Nikita, et al.
Veröffentlicht: (2024)
Towards End-to-End Training of Automatic Speech Recognition for Nigerian Pidgin
von: Rufai, Amina Mardiyyah, et al.
Veröffentlicht: (2020)
von: Rufai, Amina Mardiyyah, et al.
Veröffentlicht: (2020)
Zero-Shot Generalizable End-to-End Task-Oriented Dialog System using Context Summarization and Domain Schema
von: Mosharrof, Adib, et al.
Veröffentlicht: (2023)
von: Mosharrof, Adib, et al.
Veröffentlicht: (2023)
A Reparameterized Discrete Diffusion Model for Text Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
Revisiting Adaptive Rounding with Vectorized Reparameterization for LLM Quantization
von: Zhou, Yuli, et al.
Veröffentlicht: (2026)
von: Zhou, Yuli, et al.
Veröffentlicht: (2026)
Generalizable End-to-End Tool-Use RL with Synthetic CodeGym
von: Du, Weihua, et al.
Veröffentlicht: (2025)
von: Du, Weihua, et al.
Veröffentlicht: (2025)
Towards End-to-End Spoken Grammatical Error Correction
von: Bannò, Stefano, et al.
Veröffentlicht: (2023)
von: Bannò, Stefano, et al.
Veröffentlicht: (2023)
EMMA: End-to-End Multimodal Model for Autonomous Driving
von: Hwang, Jyh-Jing, et al.
Veröffentlicht: (2024)
von: Hwang, Jyh-Jing, et al.
Veröffentlicht: (2024)
SOMBRERO: Measuring and Steering Boundary Placement in End-to-End Hierarchical Sequence Models
von: Neitemeier, Pit, et al.
Veröffentlicht: (2026)
von: Neitemeier, Pit, et al.
Veröffentlicht: (2026)
End-to-End Spoken Grammatical Error Correction
von: Qian, Mengjie, et al.
Veröffentlicht: (2025)
von: Qian, Mengjie, et al.
Veröffentlicht: (2025)
Continual Learning for Monolingual End-to-End Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2021)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2021)
Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models
von: Hsiao, Chi-Yuan, et al.
Veröffentlicht: (2025)
von: Hsiao, Chi-Yuan, et al.
Veröffentlicht: (2025)
ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization
von: You, Haoran, et al.
Veröffentlicht: (2024)
von: You, Haoran, et al.
Veröffentlicht: (2024)
Stochastic RAG: End-to-End Retrieval-Augmented Generation through Expected Utility Maximization
von: Zamani, Hamed, et al.
Veröffentlicht: (2024)
von: Zamani, Hamed, et al.
Veröffentlicht: (2024)
GIT-CXR: End-to-End Transformer for Chest X-Ray Report Generation
von: Sîrbu, Iustin, et al.
Veröffentlicht: (2025)
von: Sîrbu, Iustin, et al.
Veröffentlicht: (2025)
Robust and Efficient Fine-tuning of LLMs with Bayesian Reparameterization of Low-Rank Adaptation
von: Sengupta, Ayan, et al.
Veröffentlicht: (2024)
von: Sengupta, Ayan, et al.
Veröffentlicht: (2024)
Enhancing LLM Evaluations: The Garbling Trick
von: Bradley, William F.
Veröffentlicht: (2024)
von: Bradley, William F.
Veröffentlicht: (2024)
Importance Weighted Variational Inference without the Reparameterization Trick
von: Daudel, Kamélia, et al.
Veröffentlicht: (2026)
von: Daudel, Kamélia, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Shared Latent Space by Both Languages in Non-Autoregressive Neural Machine Translation
von: Heo, DongNyeong, et al.
Veröffentlicht: (2023) -
Generalized Probabilistic Attention Mechanism in Transformers
von: Heo, DongNyeong, et al.
Veröffentlicht: (2024) -
Sentence Curve Language Models
von: Heo, DongNyeong, et al.
Veröffentlicht: (2026) -
N-gram Prediction and Word Difference Representations for Language Modeling
von: Heo, DongNyeong, et al.
Veröffentlicht: (2024) -
Dynamic Preference Multi-Objective Reinforcement Learning for Internet Network Management
von: Heo, DongNyeong, et al.
Veröffentlicht: (2025)