Saved in:
| Main Authors: | Hai, Nam Le, Bui, Nghi D. Q. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2408.04663 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Impacts of Contexts on Repository-Level Code Generation
by: Hai, Nam Le, et al.
Published: (2024)
by: Hai, Nam Le, et al.
Published: (2024)
Building Effective AI Coding Agents for the Terminal: Scaffolding, Harness, Context Engineering, and Lessons Learned
by: Bui, Nghi D. Q.
Published: (2026)
by: Bui, Nghi D. Q.
Published: (2026)
XMainframe: A Large Language Model for Mainframe Modernization
by: Dau, Anh T. V., et al.
Published: (2024)
by: Dau, Anh T. V., et al.
Published: (2024)
Synthesizing Post-Training Data for LLMs through Multi-Agent Simulation
by: Tang, Shuo, et al.
Published: (2024)
by: Tang, Shuo, et al.
Published: (2024)
Agentic Coding Needs Proactivity, Not Just Autonomy
by: Bui, Nghi D. Q., et al.
Published: (2026)
by: Bui, Nghi D. Q., et al.
Published: (2026)
Why They Disagree: Decoding Differences in Opinions about AI Risk on the Lex Fridman Podcast
by: Truong, Nghi, et al.
Published: (2025)
by: Truong, Nghi, et al.
Published: (2025)
Envisioning the Next-Generation AI Coding Assistants: Insights & Proposals
by: Nghiem, Khanh, et al.
Published: (2024)
by: Nghiem, Khanh, et al.
Published: (2024)
PEEB: Part-based Image Classifiers with an Explainable and Editable Language Bottleneck
by: Pham, Thang M., et al.
Published: (2024)
by: Pham, Thang M., et al.
Published: (2024)
Nemotron-Cascade 2: Post-Training LLMs with Cascade RL and Multi-Domain On-Policy Distillation
by: Yang, Zhuolin, et al.
Published: (2026)
by: Yang, Zhuolin, et al.
Published: (2026)
Formal Reasoning for Intelligent QA Systems: A Case Study in the Educational Domain
by: Bui, Tuan, et al.
Published: (2025)
by: Bui, Tuan, et al.
Published: (2025)
How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradients
by: Li, Ming, et al.
Published: (2025)
by: Li, Ming, et al.
Published: (2025)
A Decoupling and Aggregating Framework for Joint Extraction of Entities and Relations
by: Wang, Yao, et al.
Published: (2024)
by: Wang, Yao, et al.
Published: (2024)
Can Post-Training Transform LLMs into Causal Reasoners?
by: Chen, Junqi, et al.
Published: (2026)
by: Chen, Junqi, et al.
Published: (2026)
Arabic Tweet Act: A Weighted Ensemble Pre-Trained Transformer Model for Classifying Arabic Speech Acts on Twitter
by: Alshehri, Khadejaa, et al.
Published: (2024)
by: Alshehri, Khadejaa, et al.
Published: (2024)
ClaimPKG: Enhancing Claim Verification via Pseudo-Subgraph Generation with Lightweight Specialized LLM
by: Pham, Hoang, et al.
Published: (2025)
by: Pham, Hoang, et al.
Published: (2025)
Binary Classifier Optimization for Large Language Model Alignment
by: Jung, Seungjae, et al.
Published: (2024)
by: Jung, Seungjae, et al.
Published: (2024)
Unsupervised Layer-wise Score Aggregation for Textual OOD Detection
by: Darrin, Maxime, et al.
Published: (2023)
by: Darrin, Maxime, et al.
Published: (2023)
Contrast-CAT: Contrasting Activations for Enhanced Interpretability in Transformer-based Text Classifiers
by: Han, Sungmin, et al.
Published: (2025)
by: Han, Sungmin, et al.
Published: (2025)
HiMAP-Travel: Hierarchical Multi-Agent Planning for Long-Horizon Constrained Travel
by: Bui, The Viet, et al.
Published: (2026)
by: Bui, The Viet, et al.
Published: (2026)
$Q\sharp$: Provably Optimal Distributional RL for LLM Post-Training
by: Zhou, Jin Peng, et al.
Published: (2025)
by: Zhou, Jin Peng, et al.
Published: (2025)
AdaFRUGAL: Adaptive Memory-Efficient Training with Dynamic Control
by: Bui, Quang-Hung, et al.
Published: (2025)
by: Bui, Quang-Hung, et al.
Published: (2025)
Preserving Generalization of Language models in Few-shot Continual Relation Extraction
by: Tran, Quyen, et al.
Published: (2024)
by: Tran, Quyen, et al.
Published: (2024)
Post-Training Language Models for Crosslingual Consistency
by: Liu, Tianyu, et al.
Published: (2026)
by: Liu, Tianyu, et al.
Published: (2026)
Teaching Transformers Causal Reasoning through Axiomatic Training
by: Vashishtha, Aniket, et al.
Published: (2024)
by: Vashishtha, Aniket, et al.
Published: (2024)
LaCo: Large Language Model Pruning via Layer Collapse
by: Yang, Yifei, et al.
Published: (2024)
by: Yang, Yifei, et al.
Published: (2024)
Training Text-to-Molecule Models with Context-Aware Tokenization
by: Kim, Seojin, et al.
Published: (2025)
by: Kim, Seojin, et al.
Published: (2025)
Dual-Head Reasoning Distillation: Improving Classifier Accuracy with Train-Time-Only Reasoning
by: Xu, Jillian, et al.
Published: (2025)
by: Xu, Jillian, et al.
Published: (2025)
TATRA: Training-Free Instance-Adaptive Prompting Through Rephrasing and Aggregation
by: Dziuba, Bartosz, et al.
Published: (2026)
by: Dziuba, Bartosz, et al.
Published: (2026)
Agent-UniRAG: A Trainable Open-Source LLM Agent Framework for Unified Retrieval-Augmented Generation Systems
by: Pham, Hoang, et al.
Published: (2025)
by: Pham, Hoang, et al.
Published: (2025)
Verify-in-the-Graph: Entity Disambiguation Enhancement for Complex Claim Verification with Interactive Graph Representation
by: Pham, Hoang, et al.
Published: (2025)
by: Pham, Hoang, et al.
Published: (2025)
Efficient RL for optimizing conversation level outcomes with an LLM-based tutor
by: Nam, Hyunji, et al.
Published: (2025)
by: Nam, Hyunji, et al.
Published: (2025)
Fixing It in Post: A Comparative Study of LLM Post-Training Data Quality and Model Performance
by: Djuhera, Aladin, et al.
Published: (2025)
by: Djuhera, Aladin, et al.
Published: (2025)
Speaking in Words, Thinking in Logic: A Dual-Process Framework in QA Systems
by: Bui, Tuan, et al.
Published: (2025)
by: Bui, Tuan, et al.
Published: (2025)
VNJPTranslate: A comprehensive pipeline for Vietnamese-Japanese translation
by: Phan, Hoang Hai, et al.
Published: (2025)
by: Phan, Hoang Hai, et al.
Published: (2025)
Intra-Layer Recurrence in Transformers for Language Modeling
by: Nguyen, Anthony, et al.
Published: (2025)
by: Nguyen, Anthony, et al.
Published: (2025)
Multi-Layer Transformers Gradient Can be Approximated in Almost Linear Time
by: Liang, Yingyu, et al.
Published: (2024)
by: Liang, Yingyu, et al.
Published: (2024)
Aggressive Post-Training Compression on Extremely Large Language Models
by: Zhang, Zining, et al.
Published: (2024)
by: Zhang, Zining, et al.
Published: (2024)
Assessing Robustness to Spurious Correlations in Post-Training Language Models
by: Shuieh, Julia, et al.
Published: (2025)
by: Shuieh, Julia, et al.
Published: (2025)
Finding and Reactivating Post-Trained LLMs' Hidden Safety Mechanisms
by: Li, Mingjie, et al.
Published: (2026)
by: Li, Mingjie, et al.
Published: (2026)
Spectrum Tuning: Post-Training for Distributional Coverage and In-Context Steerability
by: Sorensen, Taylor, et al.
Published: (2025)
by: Sorensen, Taylor, et al.
Published: (2025)
Similar Items
-
On the Impacts of Contexts on Repository-Level Code Generation
by: Hai, Nam Le, et al.
Published: (2024) -
Building Effective AI Coding Agents for the Terminal: Scaffolding, Harness, Context Engineering, and Lessons Learned
by: Bui, Nghi D. Q.
Published: (2026) -
XMainframe: A Large Language Model for Mainframe Modernization
by: Dau, Anh T. V., et al.
Published: (2024) -
Synthesizing Post-Training Data for LLMs through Multi-Agent Simulation
by: Tang, Shuo, et al.
Published: (2024) -
Agentic Coding Needs Proactivity, Not Just Autonomy
by: Bui, Nghi D. Q., et al.
Published: (2026)