Structured Pruning for Diverse Best-of-N Reasoning Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nguyen, Hieu Trung, Nguyen, Bao, Nguyen, Viet Anh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Task-driven Layerwise Additive Activation Intervention
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2025)
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2025)
Reasoning Planning for Language Models
von: Nguyen, Bao, et al.
Veröffentlicht: (2025)
von: Nguyen, Bao, et al.
Veröffentlicht: (2025)
Adaptive Rollout Allocation for Online Reinforcement Learning with Verifiable Rewards
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2026)
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2026)
Distributional Surgery for Language Model Activations
von: Nguyen, Bao, et al.
Veröffentlicht: (2025)
von: Nguyen, Bao, et al.
Veröffentlicht: (2025)
Mixture-of-Personas Language Models for Population Simulation
von: Bui, Ngoc, et al.
Veröffentlicht: (2025)
von: Bui, Ngoc, et al.
Veröffentlicht: (2025)
Generative Conditional Distributions by Neural (Entropic) Optimal Transport
von: Nguyen, Bao, et al.
Veröffentlicht: (2024)
von: Nguyen, Bao, et al.
Veröffentlicht: (2024)
Explaining Graph Neural Networks via Structure-aware Interaction Index
von: Bui, Ngoc, et al.
Veröffentlicht: (2024)
von: Bui, Ngoc, et al.
Veröffentlicht: (2024)
Cold-start Recommendation by Personalized Embedding Region Elicitation
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2024)
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2024)
Bellman Optimal Stepsize Straightening of Flow-Matching Models
von: Nguyen, Bao, et al.
Veröffentlicht: (2023)
von: Nguyen, Bao, et al.
Veröffentlicht: (2023)
Toward a Flexible Framework for Linear Representation Hypothesis Using Maximum Likelihood Estimation
von: Nguyen, Trung, et al.
Veröffentlicht: (2025)
von: Nguyen, Trung, et al.
Veröffentlicht: (2025)
Does Continued Pretraining on a Learner Corpus Improve Automated Essay Scoring on English Proficiency Tests? Evidence from EFCAMDAT
von: Nguyen, Duy Anh
Veröffentlicht: (2026)
von: Nguyen, Duy Anh
Veröffentlicht: (2026)
Adaptive Two-Phase Finetuning LLMs for Japanese Legal Text Retrieval
von: Trung, Quang Hoang, et al.
Veröffentlicht: (2024)
von: Trung, Quang Hoang, et al.
Veröffentlicht: (2024)
Retrospective Feature Estimation for Continual Learning
von: Nguyen, Nghia D., et al.
Veröffentlicht: (2024)
von: Nguyen, Nghia D., et al.
Veröffentlicht: (2024)
MIC: Maximizing Informational Capacity in Adaptive Representations via Isotropic Subspace Alignment
von: Hong, Dang Nguyen, et al.
Veröffentlicht: (2026)
von: Hong, Dang Nguyen, et al.
Veröffentlicht: (2026)
Language-conditioned world model improves policy generalization by reading environmental descriptions
von: Nguyen, Anh, et al.
Veröffentlicht: (2025)
von: Nguyen, Anh, et al.
Veröffentlicht: (2025)
Cost-Adaptive Recourse Recommendation by Adaptive Preference Elicitation
von: Nguyen, Duy, et al.
Veröffentlicht: (2024)
von: Nguyen, Duy, et al.
Veröffentlicht: (2024)
OWLViz: An Open-World Benchmark for Visual Question Answering
von: Nguyen, Thuy, et al.
Veröffentlicht: (2025)
von: Nguyen, Thuy, et al.
Veröffentlicht: (2025)
Optimizing Multi-Stage Language Models for Effective Text Retrieval
von: Trung, Quang Hoang, et al.
Veröffentlicht: (2024)
von: Trung, Quang Hoang, et al.
Veröffentlicht: (2024)
VietMix: A Naturally-Occurring Parallel Corpus and Augmentation Framework for Vietnamese-English Code-Mixed Machine Translation
von: Tran, Hieu, et al.
Veröffentlicht: (2025)
von: Tran, Hieu, et al.
Veröffentlicht: (2025)
Don't Read Everything: A Curvature-Conditioned Query for Linear Attention
von: Le, Dong, et al.
Veröffentlicht: (2026)
von: Le, Dong, et al.
Veröffentlicht: (2026)
ViLLM-Eval: A Comprehensive Evaluation Suite for Vietnamese Large Language Models
von: Nguyen, Trong-Hieu, et al.
Veröffentlicht: (2024)
von: Nguyen, Trong-Hieu, et al.
Veröffentlicht: (2024)
$π^2$: Structure-Originated Reasoning Data Improves Long-Context Reasoning Ability of Large Language Models
von: Do, Quyet V., et al.
Veröffentlicht: (2026)
von: Do, Quyet V., et al.
Veröffentlicht: (2026)
Understanding Transformers via N-gram Statistics
von: Nguyen, Timothy
Veröffentlicht: (2024)
von: Nguyen, Timothy
Veröffentlicht: (2024)
Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI
von: Man, Hieu, et al.
Veröffentlicht: (2026)
von: Man, Hieu, et al.
Veröffentlicht: (2026)
Defect Prediction with Content-based Features
von: Pham, Hung Viet, et al.
Veröffentlicht: (2024)
von: Pham, Hung Viet, et al.
Veröffentlicht: (2024)
ATLAS: Adaptive Test-Time Latent Steering with External Verifiers for Enhancing LLMs Reasoning
von: Nguyen, Tuc, et al.
Veröffentlicht: (2026)
von: Nguyen, Tuc, et al.
Veröffentlicht: (2026)
Interpretable LLM-based Table Question Answering
von: Nguyen, Giang, et al.
Veröffentlicht: (2024)
von: Nguyen, Giang, et al.
Veröffentlicht: (2024)
Lizard: An Efficient Linearization Framework for Large Language Models
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2025)
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2025)
Provably Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function
von: Le, Tung Quoc, et al.
Veröffentlicht: (2026)
von: Le, Tung Quoc, et al.
Veröffentlicht: (2026)
Few-shot Continual Relation Extraction via Open Information Extraction
von: Nguyen, Thiem, et al.
Veröffentlicht: (2025)
von: Nguyen, Thiem, et al.
Veröffentlicht: (2025)
When In-Distribution Gains Fail: Evaluating Weak-to-Strong Reward Models under Preference Shift
von: Le, Khoi, et al.
Veröffentlicht: (2026)
von: Le, Khoi, et al.
Veröffentlicht: (2026)
Disentangling the Roles of Representation and Selection in Data Pruning
von: Du, Yupei, et al.
Veröffentlicht: (2025)
von: Du, Yupei, et al.
Veröffentlicht: (2025)
Provably data-driven projection method for quadratic programming
von: Nguyen, Anh Tuan, et al.
Veröffentlicht: (2025)
von: Nguyen, Anh Tuan, et al.
Veröffentlicht: (2025)
Tree-OPO: Off-policy Monte Carlo Tree-Guided Advantage Optimization for Multistep Reasoning
von: Huang, Bingning, et al.
Veröffentlicht: (2025)
von: Huang, Bingning, et al.
Veröffentlicht: (2025)
Beyond Forgetting: Machine Unlearning Elicits Controllable Side Behaviors and Capabilities
von: Dang, Tien, et al.
Veröffentlicht: (2026)
von: Dang, Tien, et al.
Veröffentlicht: (2026)
Fairness in Large Language Models in Three Hours
von: Viet, Thang Doan, et al.
Veröffentlicht: (2024)
von: Viet, Thang Doan, et al.
Veröffentlicht: (2024)
Language Models are Bounded Pragmatic Speakers: Understanding RLHF from a Bayesian Cognitive Modeling Perspective
von: Nguyen, Khanh
Veröffentlicht: (2023)
von: Nguyen, Khanh
Veröffentlicht: (2023)
B-score: Detecting biases in large language models using response history
von: Vo, An, et al.
Veröffentlicht: (2025)
von: Vo, An, et al.
Veröffentlicht: (2025)
Taipan: Efficient and Expressive State Space Language Models with Selective Attention
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2024)
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2024)
ViQA-COVID: COVID-19 Machine Reading Comprehension Dataset for Vietnamese
von: Nguyen-Phung, Hai-Chung, et al.
Veröffentlicht: (2025)
von: Nguyen-Phung, Hai-Chung, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Task-driven Layerwise Additive Activation Intervention
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2025) -
Reasoning Planning for Language Models
von: Nguyen, Bao, et al.
Veröffentlicht: (2025) -
Adaptive Rollout Allocation for Online Reinforcement Learning with Verifiable Rewards
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2026) -
Distributional Surgery for Language Model Activations
von: Nguyen, Bao, et al.
Veröffentlicht: (2025) -
Mixture-of-Personas Language Models for Population Simulation
von: Bui, Ngoc, et al.
Veröffentlicht: (2025)