End-to-end Planner Training for Language Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Cornille, Nathan, Mai, Florian, Sun, Jingyuan, Moens, Marie-Francine |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Plan Long-Term for Language Modeling
by: Mai, Florian, et al.
Published: (2024)
by: Mai, Florian, et al.
Published: (2024)
Learning to Plan for Language Modeling from Unlabeled Data
by: Cornille, Nathan, et al.
Published: (2024)
by: Cornille, Nathan, et al.
Published: (2024)
Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
by: Araujo, Vladimir, et al.
Published: (2024)
by: Araujo, Vladimir, et al.
Published: (2024)
Sequence-to-Sequence Spanish Pre-trained Language Models
by: Araujo, Vladimir, et al.
Published: (2023)
by: Araujo, Vladimir, et al.
Published: (2023)
MSNER: A Multilingual Speech Dataset for Named Entity Recognition
by: Meeus, Quentin, et al.
Published: (2024)
by: Meeus, Quentin, et al.
Published: (2024)
Reduction of Supervision for Biomedical Knowledge Discovery
by: Theodoropoulos, Christos, et al.
Published: (2025)
by: Theodoropoulos, Christos, et al.
Published: (2025)
Multimodal Adaptive Inference for Document Image Classification with Anytime Early Exiting
by: Hamed, Omar, et al.
Published: (2024)
by: Hamed, Omar, et al.
Published: (2024)
Efficient End-to-end Language Model Fine-tuning on Graphs
by: Xue, Rui, et al.
Published: (2023)
by: Xue, Rui, et al.
Published: (2023)
Computational Models to Study Language Processing in the Human Brain: A Survey
by: Wang, Shaonan, et al.
Published: (2024)
by: Wang, Shaonan, et al.
Published: (2024)
DMON: A Simple yet Effective Approach for Argument Structure Learning
by: Sun, Wei, et al.
Published: (2024)
by: Sun, Wei, et al.
Published: (2024)
End-to-End Ontology Learning with Large Language Models
by: Lo, Andy, et al.
Published: (2024)
by: Lo, Andy, et al.
Published: (2024)
End-to-End Training for Back-Translation with Categorical Reparameterization Trick
by: Heo, DongNyeong, et al.
Published: (2022)
by: Heo, DongNyeong, et al.
Published: (2022)
Measuring Non-Adversarial Reproduction of Training Data in Large Language Models
by: Aerni, Michael, et al.
Published: (2024)
by: Aerni, Michael, et al.
Published: (2024)
Rationalizing Transformer Predictions via End-To-End Differentiable Self-Training
by: Brinner, Marc, et al.
Published: (2025)
by: Brinner, Marc, et al.
Published: (2025)
Test-Time Training on Nearest Neighbors for Large Language Models
by: Hardt, Moritz, et al.
Published: (2023)
by: Hardt, Moritz, et al.
Published: (2023)
LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models
by: Faiz, Ahmad, et al.
Published: (2023)
by: Faiz, Ahmad, et al.
Published: (2023)
Scaling LLM Multi-turn RL with End-to-end Summarization-based Context Management
by: Lu, Miao, et al.
Published: (2025)
by: Lu, Miao, et al.
Published: (2025)
Bridging Language Gaps: Enhancing Few-Shot Language Adaptation
by: Borchert, Philipp, et al.
Published: (2025)
by: Borchert, Philipp, et al.
Published: (2025)
Saten: Sparse Augmented Tensor Networks for Post-Training Compression of Large Language Models
by: Solgi, Ryan, et al.
Published: (2025)
by: Solgi, Ryan, et al.
Published: (2025)
End-to-end Sequence Labeling via Bi-directional LSTM-CNNs-CRF: A Reproducibility Study
by: Ganesh, Anirudh, et al.
Published: (2025)
by: Ganesh, Anirudh, et al.
Published: (2025)
ScoNe: Benchmarking Negation Reasoning in Language Models With Fine-Tuning and In-Context Learning
by: She, Jingyuan Selena, et al.
Published: (2023)
by: She, Jingyuan Selena, et al.
Published: (2023)
FLAME-MoE: A Transparent End-to-End Research Platform for Mixture-of-Experts Language Models
by: Kang, Hao, et al.
Published: (2025)
by: Kang, Hao, et al.
Published: (2025)
Integrating Pre-Trained Speech and Language Models for End-to-End Speech Recognition
by: Hono, Yukiya, et al.
Published: (2023)
by: Hono, Yukiya, et al.
Published: (2023)
PolyPrompt: Automating Knowledge Extraction from Multilingual Language Models with Dynamic Prompt Generation
by: Roll, Nathan
Published: (2025)
by: Roll, Nathan
Published: (2025)
An End-to-End Approach for Child Reading Assessment in the Xhosa Language
by: Chevtchenko, Sergio, et al.
Published: (2025)
by: Chevtchenko, Sergio, et al.
Published: (2025)
Accelerating Large Language Model Inference with Self-Supervised Early Exits
by: Valade, Florian
Published: (2024)
by: Valade, Florian
Published: (2024)
One STEP at a time: Language Agents are Stepwise Planners
by: Nguyen, Minh, et al.
Published: (2024)
by: Nguyen, Minh, et al.
Published: (2024)
Large Language Models are Learnable Planners for Long-Term Recommendation
by: Shi, Wentao, et al.
Published: (2024)
by: Shi, Wentao, et al.
Published: (2024)
Tree-Planner: Efficient Close-loop Task Planning with Large Language Models
by: Hu, Mengkang, et al.
Published: (2023)
by: Hu, Mengkang, et al.
Published: (2023)
Janus-Q: End-to-End Event-Driven Trading via Hierarchical-Gated Reward Modeling
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
Modifying Large Language Model Post-Training for Diverse Creative Writing
by: Chung, John Joon Young, et al.
Published: (2025)
by: Chung, John Joon Young, et al.
Published: (2025)
WebAgent-R1: Training Web Agents via End-to-End Multi-Turn Reinforcement Learning
by: Wei, Zhepei, et al.
Published: (2025)
by: Wei, Zhepei, et al.
Published: (2025)
"Image, Tell me your story!" Predicting the original meta-context of visual misinformation
by: Tonglet, Jonathan, et al.
Published: (2024)
by: Tonglet, Jonathan, et al.
Published: (2024)
Explicitly Representing Syntax Improves Sentence-to-layout Prediction of Unexpected Situations
by: Nuyts, Wolf, et al.
Published: (2024)
by: Nuyts, Wolf, et al.
Published: (2024)
End-to-end streaming model for low-latency speech anonymization
by: Quamer, Waris, et al.
Published: (2024)
by: Quamer, Waris, et al.
Published: (2024)
MONA: Muon Optimizer with Nesterov Acceleration for Scalable Language Model Training
by: Li, Jiacheng, et al.
Published: (2026)
by: Li, Jiacheng, et al.
Published: (2026)
Fighting Against the Repetitive Training and Sample Dependency Problem in Few-shot Named Entity Recognition
by: Tian, Chang, et al.
Published: (2024)
by: Tian, Chang, et al.
Published: (2024)
Cascade-Aware Training of Language Models
by: Wang, Congchao, et al.
Published: (2024)
by: Wang, Congchao, et al.
Published: (2024)
Training Language Models to Reason Efficiently
by: Arora, Daman, et al.
Published: (2025)
by: Arora, Daman, et al.
Published: (2025)
Dialogue Action Tokens: Steering Language Models in Goal-Directed Dialogue with a Multi-Turn Planner
by: Li, Kenneth, et al.
Published: (2024)
by: Li, Kenneth, et al.
Published: (2024)
Similar Items
-
Learning to Plan Long-Term for Language Modeling
by: Mai, Florian, et al.
Published: (2024) -
Learning to Plan for Language Modeling from Unlabeled Data
by: Cornille, Nathan, et al.
Published: (2024) -
Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
by: Araujo, Vladimir, et al.
Published: (2024) -
Sequence-to-Sequence Spanish Pre-trained Language Models
by: Araujo, Vladimir, et al.
Published: (2023) -
MSNER: A Multilingual Speech Dataset for Named Entity Recognition
by: Meeus, Quentin, et al.
Published: (2024)