CoDiQ: Test-Time Scaling for Controllable Difficult Question Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Peng, Zhongyuan, Xu, Caijun, Xiao, Changyi, Hong, Shibo, Zhang, Eli, Huang, Stephen, Cao, Yixin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Reinforcement Learning with Conditional Expectation Reward
di: Xiao, Changyi, et al.
Pubblicazione: (2026)
di: Xiao, Changyi, et al.
Pubblicazione: (2026)
SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning
di: Xu, Caijun, et al.
Pubblicazione: (2026)
di: Xu, Caijun, et al.
Pubblicazione: (2026)
Long or short CoT? Investigating Instance-level Switch of Large Reasoning Models
di: Zhang, Ruiqi, et al.
Pubblicazione: (2025)
di: Zhang, Ruiqi, et al.
Pubblicazione: (2025)
Generating Difficult-to-Translate Texts
di: Zouhar, Vilém, et al.
Pubblicazione: (2025)
di: Zouhar, Vilém, et al.
Pubblicazione: (2025)
Complex Logical Query Answering by Calibrating Knowledge Graph Completion Models
di: Xiao, Changyi, et al.
Pubblicazione: (2024)
di: Xiao, Changyi, et al.
Pubblicazione: (2024)
Knowledge Graph Completion by Intermediate Variables Regularization
di: Xiao, Changyi, et al.
Pubblicazione: (2025)
di: Xiao, Changyi, et al.
Pubblicazione: (2025)
CoDi: Conversational Distillation for Grounded Question Answering
di: Huber, Patrick, et al.
Pubblicazione: (2024)
di: Huber, Patrick, et al.
Pubblicazione: (2024)
SAND-Math: Using LLMs to Generate Novel, Difficult and Useful Mathematics Questions and Answers
di: Manem, Chaitanya, et al.
Pubblicazione: (2025)
di: Manem, Chaitanya, et al.
Pubblicazione: (2025)
SoftCoT++: Test-Time Scaling with Soft Chain-of-Thought Reasoning
di: Xu, Yige, et al.
Pubblicazione: (2025)
di: Xu, Yige, et al.
Pubblicazione: (2025)
Less Data Less Tokens: Multilingual Unification Learning for Efficient Test-Time Reasoning in LLMs
di: Chen, Kang, et al.
Pubblicazione: (2025)
di: Chen, Kang, et al.
Pubblicazione: (2025)
T$^2$: An Adaptive Test-Time Scaling Strategy for Contextual Question Answering
di: Zhao, Zhengyi, et al.
Pubblicazione: (2025)
di: Zhao, Zhengyi, et al.
Pubblicazione: (2025)
General Table Question Answering via Answer-Formula Joint Generation
di: Wang, Zhongyuan, et al.
Pubblicazione: (2025)
di: Wang, Zhongyuan, et al.
Pubblicazione: (2025)
Is That Your Final Answer? Test-Time Scaling Improves Selective Question Answering
di: Jurayj, William, et al.
Pubblicazione: (2025)
di: Jurayj, William, et al.
Pubblicazione: (2025)
ChartReasoner: Code-Driven Modality Bridging for Long-Chain Reasoning in Chart Question Answering
di: Jia, Caijun, et al.
Pubblicazione: (2025)
di: Jia, Caijun, et al.
Pubblicazione: (2025)
Test-Time Scaling with Reflective Generative Model
di: Wang, Zixiao, et al.
Pubblicazione: (2025)
di: Wang, Zixiao, et al.
Pubblicazione: (2025)
ScaleDiff: Scaling Difficult Problems for Advanced Mathematical Reasoning
di: Pei, Qizhi, et al.
Pubblicazione: (2025)
di: Pei, Qizhi, et al.
Pubblicazione: (2025)
Are You Doubtful? Oh, It Might Be Difficult Then! Exploring the Use of Model Uncertainty for Question Difficulty Estimation
di: Zotos, Leonidas, et al.
Pubblicazione: (2024)
di: Zotos, Leonidas, et al.
Pubblicazione: (2024)
QRMeM: Unleash the Length Limitation through Question then Reflection Memory Mechanism
di: Wang, Bo, et al.
Pubblicazione: (2024)
di: Wang, Bo, et al.
Pubblicazione: (2024)
Generative AI Act II: Test Time Scaling Drives Cognition Engineering
di: Xia, Shijie, et al.
Pubblicazione: (2025)
di: Xia, Shijie, et al.
Pubblicazione: (2025)
Linguistic Generalizability of Test-Time Scaling in Mathematical Reasoning
di: Son, Guijin, et al.
Pubblicazione: (2025)
di: Son, Guijin, et al.
Pubblicazione: (2025)
Optimal Aggregation of LLM and PRM Signals for Efficient Test-Time Scaling
di: Kuang, Peng, et al.
Pubblicazione: (2025)
di: Kuang, Peng, et al.
Pubblicazione: (2025)
Co-Trained Retriever-Generator Framework for Question Generation in Earnings Calls
di: Juan, Yining, et al.
Pubblicazione: (2024)
di: Juan, Yining, et al.
Pubblicazione: (2024)
QueST: Incentivizing LLMs to Generate Difficult Problems
di: Hu, Hanxu, et al.
Pubblicazione: (2025)
di: Hu, Hanxu, et al.
Pubblicazione: (2025)
BNPO: Beta Normalization Policy Optimization
di: Xiao, Changyi, et al.
Pubblicazione: (2025)
di: Xiao, Changyi, et al.
Pubblicazione: (2025)
Knowledge Graph Embedding by Normalizing Flows
di: Xiao, Changyi, et al.
Pubblicazione: (2024)
di: Xiao, Changyi, et al.
Pubblicazione: (2024)
UniScale: Adaptive Unified Inference Scaling via Online Joint Optimization of Model Routing and Test-Time Scaling
di: Huang, Kaiyu, et al.
Pubblicazione: (2026)
di: Huang, Kaiyu, et al.
Pubblicazione: (2026)
Do LLMs and Humans Find the Same Questions Difficult? A Case Study on Japanese Quiz Answering
di: Sugiura, Naoya, et al.
Pubblicazione: (2025)
di: Sugiura, Naoya, et al.
Pubblicazione: (2025)
CTTS: Collective Test-Time Scaling
di: Song, Zhende, et al.
Pubblicazione: (2025)
di: Song, Zhende, et al.
Pubblicazione: (2025)
From Mathematical Reasoning to Code: Generalization of Process Reward Models in Test-Time Scaling
di: Chen, Zhengyu, et al.
Pubblicazione: (2025)
di: Chen, Zhengyu, et al.
Pubblicazione: (2025)
MetaScale: Test-Time Scaling with Evolving Meta-Thoughts
di: Liu, Qin, et al.
Pubblicazione: (2025)
di: Liu, Qin, et al.
Pubblicazione: (2025)
DiSCTT: Consensus-Guided Self-Curriculum for Efficient Test-Time Adaptation in Reasoning
di: Moradi, Mohammad Mahdi, et al.
Pubblicazione: (2026)
di: Moradi, Mohammad Mahdi, et al.
Pubblicazione: (2026)
Benchmark Test-Time Scaling of General LLM Agents
di: Li, Xiaochuan, et al.
Pubblicazione: (2026)
di: Li, Xiaochuan, et al.
Pubblicazione: (2026)
Code-Style In-Context Learning for Knowledge-Based Question Answering
di: Nie, Zhijie, et al.
Pubblicazione: (2023)
di: Nie, Zhijie, et al.
Pubblicazione: (2023)
Efficient Test-Time Scaling via Self-Calibration
di: Huang, Chengsong, et al.
Pubblicazione: (2025)
di: Huang, Chengsong, et al.
Pubblicazione: (2025)
CoSPlay: Cooperative Self-Play at Test-Time with Self-Generated Code and Unit Test
di: Hu, Zhangyi, et al.
Pubblicazione: (2026)
di: Hu, Zhangyi, et al.
Pubblicazione: (2026)
DiSA: Diffusion Step Annealing in Autoregressive Image Generation
di: Zhao, Qinyu, et al.
Pubblicazione: (2025)
di: Zhao, Qinyu, et al.
Pubblicazione: (2025)
DiLoCoX: A Low-Communication Large-Scale Training Framework for Decentralized Cluster
di: Qi, Ji, et al.
Pubblicazione: (2025)
di: Qi, Ji, et al.
Pubblicazione: (2025)
Long Is More Important Than Difficult for Training Reasoning Models
di: Shen, Si, et al.
Pubblicazione: (2025)
di: Shen, Si, et al.
Pubblicazione: (2025)
ATLAS: All-round Testing of Long-context Abilities across Scales
di: Huang, Deli, et al.
Pubblicazione: (2026)
di: Huang, Deli, et al.
Pubblicazione: (2026)
Test-Time Scaling with Repeated Sampling Improves Multilingual Text Generation
di: Gupta, Ashim, et al.
Pubblicazione: (2025)
di: Gupta, Ashim, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Reinforcement Learning with Conditional Expectation Reward
di: Xiao, Changyi, et al.
Pubblicazione: (2026) -
SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning
di: Xu, Caijun, et al.
Pubblicazione: (2026) -
Long or short CoT? Investigating Instance-level Switch of Large Reasoning Models
di: Zhang, Ruiqi, et al.
Pubblicazione: (2025) -
Generating Difficult-to-Translate Texts
di: Zouhar, Vilém, et al.
Pubblicazione: (2025) -
Complex Logical Query Answering by Calibrating Knowledge Graph Completion Models
di: Xiao, Changyi, et al.
Pubblicazione: (2024)