Know When To Fold 'Em: Token-Efficient LLM Synthetic Data Generation via Multi-Stage In-Flight Rejection
Fuente:
arXiv
Guardado en:
| Autores principales: | Chowdhury, Anjir Ahmed, Zawad, Syed, Yan, Feng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PEML: Parameter-efficient Multi-Task Learning with Optimized Continuous Prompts
por: Chowdhury, Anjir Ahmed, et al.
Publicado: (2026)
por: Chowdhury, Anjir Ahmed, et al.
Publicado: (2026)
When Data is the Algorithm: A Systematic Study and Curation of Preference Optimization Datasets
por: Djuhera, Aladin, et al.
Publicado: (2025)
por: Djuhera, Aladin, et al.
Publicado: (2025)
Fixing It in Post: A Comparative Study of LLM Post-Training Data Quality and Model Performance
por: Djuhera, Aladin, et al.
Publicado: (2025)
por: Djuhera, Aladin, et al.
Publicado: (2025)
TSR: Trajectory-Search Rollouts for Multi-Turn RL of LLM Agents
por: Djuhera, Aladin, et al.
Publicado: (2026)
por: Djuhera, Aladin, et al.
Publicado: (2026)
SafeMERGE: Preserving Safety Alignment in Fine-Tuned Large Language Models via Selective Layer-Wise Model Merging
por: Djuhera, Aladin, et al.
Publicado: (2025)
por: Djuhera, Aladin, et al.
Publicado: (2025)
CaRT: Teaching LLM Agents to Know When They Know Enough
por: Liu, Grace, et al.
Publicado: (2025)
por: Liu, Grace, et al.
Publicado: (2025)
Enhancing Multi-Class Disease Classification: Neoplasms, Cardiovascular, Nervous System, and Digestive Disorders Using Advanced LLMs
por: Karim, Ahmed Akib Jawad, et al.
Publicado: (2024)
por: Karim, Ahmed Akib Jawad, et al.
Publicado: (2024)
Scaling Instruction-Tuned LLMs to Million-Token Contexts via Hierarchical Synthetic Data Generation
por: He, Linda, et al.
Publicado: (2025)
por: He, Linda, et al.
Publicado: (2025)
VISTA: Visualization of Token Attribution via Efficient Analysis
por: Ahmed, Syed, et al.
Publicado: (2026)
por: Ahmed, Syed, et al.
Publicado: (2026)
Persona-Based Synthetic Data Generation Using Multi-Stage Conditioning with Large Language Models for Emotion Recognition
por: Inoshita, Keito, et al.
Publicado: (2025)
por: Inoshita, Keito, et al.
Publicado: (2025)
Do Retrieval Augmented Language Models Know When They Don't Know?
por: Zhou, Youchao, et al.
Publicado: (2025)
por: Zhou, Youchao, et al.
Publicado: (2025)
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
por: Mei, Zhiting, et al.
Publicado: (2025)
por: Mei, Zhiting, et al.
Publicado: (2025)
Enhancing Document-Level Machine Translation via Filtered Synthetic Corpora and Two-Stage LLM Adaptation
por: Kim, Ireh, et al.
Publicado: (2026)
por: Kim, Ireh, et al.
Publicado: (2026)
CHIMERA: Compact Synthetic Data for Generalizable LLM Reasoning
por: Zhu, Xinyu, et al.
Publicado: (2026)
por: Zhu, Xinyu, et al.
Publicado: (2026)
CrowdSelect: Synthetic Instruction Data Selection with Multi-LLM Wisdom
por: Li, Yisen, et al.
Publicado: (2025)
por: Li, Yisen, et al.
Publicado: (2025)
Beyond Words: Multimodal LLM Knows When to Speak
por: Liao, Zikai, et al.
Publicado: (2025)
por: Liao, Zikai, et al.
Publicado: (2025)
CompressKV: Semantic Retrieval Heads Know What Tokens are Not Important Before Generation
por: Lin, Xiaolin, et al.
Publicado: (2025)
por: Lin, Xiaolin, et al.
Publicado: (2025)
When Verification Fails: How Compositionally Infeasible Claims Escape Rejection
por: Liu, Muxin, et al.
Publicado: (2026)
por: Liu, Muxin, et al.
Publicado: (2026)
Next-Token Prediction Task Assumes Optimal Data Ordering for LLM Training in Proof Generation
por: An, Chenyang, et al.
Publicado: (2024)
por: An, Chenyang, et al.
Publicado: (2024)
Eliminating Agentic Workflow for Introduction Generation with Parametric Stage Tokens
por: Zhang, Meicong, et al.
Publicado: (2025)
por: Zhang, Meicong, et al.
Publicado: (2025)
What Models Know, How Well They Know It: Knowledge-Weighted Fine-Tuning for Learning When to Say "I Don't Know"
por: Lee, Joosung, et al.
Publicado: (2026)
por: Lee, Joosung, et al.
Publicado: (2026)
When to Ensemble: Identifying Token-Level Points for Stable and Fast LLM Ensembling
por: Yun, Heecheol, et al.
Publicado: (2025)
por: Yun, Heecheol, et al.
Publicado: (2025)
Knowing When Not to Answer: Abstention-Aware Scientific Reasoning
por: Abdaljalil, Samir, et al.
Publicado: (2026)
por: Abdaljalil, Samir, et al.
Publicado: (2026)
Semantic Density Effect (SDE): Maximizing Information Per Token Improves LLM Accuracy
por: Ahmed, Amr
Publicado: (2026)
por: Ahmed, Amr
Publicado: (2026)
When Reject Turns into Accept: Quantifying the Vulnerability of LLM-Based Scientific Reviewers to Indirect Prompt Injection
por: Sahoo, Devanshu, et al.
Publicado: (2025)
por: Sahoo, Devanshu, et al.
Publicado: (2025)
The First Token Knows: Single-Decode Confidence for Hallucination Detection
por: Gabriel, Mina
Publicado: (2026)
por: Gabriel, Mina
Publicado: (2026)
Knowing You Don't Know: Learning When to Continue Search in Multi-round RAG through Self-Practicing
por: Yang, Diji, et al.
Publicado: (2025)
por: Yang, Diji, et al.
Publicado: (2025)
OrthoRank: Token Selection via Sink Token Orthogonality for Efficient LLM inference
por: Shin, Seungjun, et al.
Publicado: (2025)
por: Shin, Seungjun, et al.
Publicado: (2025)
The LLM Already Knows: Estimating LLM-Perceived Question Difficulty via Hidden Representations
por: Zhu, Yubo, et al.
Publicado: (2025)
por: Zhu, Yubo, et al.
Publicado: (2025)
CogniFold: Always-On Proactive Memory via Cognitive Folding
por: Wang, Suli, et al.
Publicado: (2026)
por: Wang, Suli, et al.
Publicado: (2026)
Draft Model Knows When to Stop: Self-Verification Speculative Decoding for Long-Form Generation
por: Zhang, Ziyin, et al.
Publicado: (2024)
por: Zhang, Ziyin, et al.
Publicado: (2024)
SnapKV: LLM Knows What You are Looking for Before Generation
por: Li, Yuhong, et al.
Publicado: (2024)
por: Li, Yuhong, et al.
Publicado: (2024)
Text Knows What, Tables Know When: Clinical Timeline Reconstruction via Retrieval-Augmented Multimodal Alignment
por: Kumar, Sayantan, et al.
Publicado: (2026)
por: Kumar, Sayantan, et al.
Publicado: (2026)
Large Language Models Often Know When They Are Being Evaluated
por: Needham, Joe, et al.
Publicado: (2025)
por: Needham, Joe, et al.
Publicado: (2025)
Knowing When to Abstain: Medical LLMs Under Clinical Uncertainty
por: Machcha, Sravanthi, et al.
Publicado: (2026)
por: Machcha, Sravanthi, et al.
Publicado: (2026)
Direct Multi-Token Decoding
por: Luo, Xuan, et al.
Publicado: (2025)
por: Luo, Xuan, et al.
Publicado: (2025)
EMSDialog: Synthetic Multi-person Emergency Medical Service Dialogue Generation from Electronic Patient Care Reports via Multi-LLM Agents
por: Ge, Xueren, et al.
Publicado: (2026)
por: Ge, Xueren, et al.
Publicado: (2026)
DyLLM: Efficient Diffusion LLM Inference via Saliency-based Token Selection and Partial Attention
por: Lee, Younjoo, et al.
Publicado: (2026)
por: Lee, Younjoo, et al.
Publicado: (2026)
SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
por: Li, Zheng, et al.
Publicado: (2025)
por: Li, Zheng, et al.
Publicado: (2025)
LLM for Barcodes: Generating Diverse Synthetic Data for Identity Documents
por: Patel, Hitesh Laxmichand, et al.
Publicado: (2024)
por: Patel, Hitesh Laxmichand, et al.
Publicado: (2024)
Ejemplares similares
-
PEML: Parameter-efficient Multi-Task Learning with Optimized Continuous Prompts
por: Chowdhury, Anjir Ahmed, et al.
Publicado: (2026) -
When Data is the Algorithm: A Systematic Study and Curation of Preference Optimization Datasets
por: Djuhera, Aladin, et al.
Publicado: (2025) -
Fixing It in Post: A Comparative Study of LLM Post-Training Data Quality and Model Performance
por: Djuhera, Aladin, et al.
Publicado: (2025) -
TSR: Trajectory-Search Rollouts for Multi-Turn RL of LLM Agents
por: Djuhera, Aladin, et al.
Publicado: (2026) -
SafeMERGE: Preserving Safety Alignment in Fine-Tuned Large Language Models via Selective Layer-Wise Model Merging
por: Djuhera, Aladin, et al.
Publicado: (2025)