VFScale: Intrinsic Reasoning through Verifier-Free Test-time Scalable Diffusion Model
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Tao, Pan, Jia-Shu, Feng, Ruiqi, Wu, Tailin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
One step further with Monte-Carlo sampler to guide diffusion better
by: Ren, Minsi, et al.
Published: (2026)
by: Ren, Minsi, et al.
Published: (2026)
Wavelet Diffusion Neural Operator
by: Hu, Peiyan, et al.
Published: (2024)
by: Hu, Peiyan, et al.
Published: (2024)
From Uncertain to Safe: Conformal Fine-Tuning of Diffusion Models for Safe PDE Control
by: Hu, Peiyan, et al.
Published: (2025)
by: Hu, Peiyan, et al.
Published: (2025)
CL-DiffPhyCon: Closed-loop Diffusion Control of Complex Physical Systems
by: Wei, Long, et al.
Published: (2024)
by: Wei, Long, et al.
Published: (2024)
On the Guidance of Flow Matching
by: Feng, Ruiqi, et al.
Published: (2025)
by: Feng, Ruiqi, et al.
Published: (2025)
GenCP: Towards Generative Modeling Paradigm of Coupled Physics
by: Gao, Tianrun, et al.
Published: (2026)
by: Gao, Tianrun, et al.
Published: (2026)
FourierFlow: Frequency-aware Flow Matching for Generative Turbulence Modeling
by: Wang, Haixin, et al.
Published: (2025)
by: Wang, Haixin, et al.
Published: (2025)
Verifier-Free RL for LLMs via Intrinsic Gradient-Norm Reward
by: Wen, Xuexiang, et al.
Published: (2026)
by: Wen, Xuexiang, et al.
Published: (2026)
TEMPO: Scaling Test-time Training for Large Reasoning Models
by: Zhang, Qingyang, et al.
Published: (2026)
by: Zhang, Qingyang, et al.
Published: (2026)
M2PDE: Compositional Generative Multiphysics and Multi-component PDE Simulation
by: Zhang, Tao, et al.
Published: (2024)
by: Zhang, Tao, et al.
Published: (2024)
DiffPhyCon: A Generative Approach to Control Complex Physical Systems
by: Wei, Long, et al.
Published: (2024)
by: Wei, Long, et al.
Published: (2024)
Reasoning through Verifiable Forecast Actions: Consistency-Grounded RL for Financial LLMs
by: Chen, Jialin, et al.
Published: (2026)
by: Chen, Jialin, et al.
Published: (2026)
Unlocking Reasoning Capabilities in LLMs via Reinforcement Learning Exploration
by: Deng, Wenhao, et al.
Published: (2025)
by: Deng, Wenhao, et al.
Published: (2025)
DEER: Draft with Diffusion, Verify with Autoregressive Models
by: Cheng, Zicong, et al.
Published: (2025)
by: Cheng, Zicong, et al.
Published: (2025)
Mitigating Gradient Pathology in PINNs through Aligned Constraint
by: Luo, Yichen, et al.
Published: (2026)
by: Luo, Yichen, et al.
Published: (2026)
Stabilizing Recurrent Dynamics for Test-Time Scalable Latent Reasoning in Looped Language Models
by: Yang, Xiao-Wen, et al.
Published: (2026)
by: Yang, Xiao-Wen, et al.
Published: (2026)
Learning Energy-Based Models by Cooperative Diffusion Recovery Likelihood
by: Zhu, Yaxuan, et al.
Published: (2023)
by: Zhu, Yaxuan, et al.
Published: (2023)
Conditional Diffusion Models Based Conditional Independence Testing
by: Yang, Yanfeng, et al.
Published: (2024)
by: Yang, Yanfeng, et al.
Published: (2024)
RFG: Test-Time Scaling for Diffusion Large Language Model Reasoning with Reward-Free Guidance
by: Chen, Tianlang, et al.
Published: (2025)
by: Chen, Tianlang, et al.
Published: (2025)
M1: Towards Scalable Test-Time Compute with Mamba Reasoning Models
by: Wang, Junxiong, et al.
Published: (2025)
by: Wang, Junxiong, et al.
Published: (2025)
Learning Design-Score Manifold to Guide Diffusion Models for Offline Optimization
by: Zhou, Tailin, et al.
Published: (2025)
by: Zhou, Tailin, et al.
Published: (2025)
SPEED-RL: Faster Training of Reasoning Models via Online Curriculum Learning
by: Zhang, Ruiqi, et al.
Published: (2025)
by: Zhang, Ruiqi, et al.
Published: (2025)
CoVerRL: Breaking the Consensus Trap in Label-Free Reasoning via Generator-Verifier Co-Evolution
by: Pan, Teng, et al.
Published: (2026)
by: Pan, Teng, et al.
Published: (2026)
Re-Evaluating the Impact of Unseen-Class Unlabeled Data on Semi-Supervised Learning Model
by: He, Rundong, et al.
Published: (2025)
by: He, Rundong, et al.
Published: (2025)
On Convolutions, Intrinsic Dimension, and Diffusion Models
by: Leung, Kin Kwan, et al.
Published: (2025)
by: Leung, Kin Kwan, et al.
Published: (2025)
VeriThinker: Learning to Verify Makes Reasoning Model Efficient
by: Chen, Zigeng, et al.
Published: (2025)
by: Chen, Zigeng, et al.
Published: (2025)
Inference-time Scaling of Diffusion Models through Classical Search
by: Zhang, Xiangcheng, et al.
Published: (2025)
by: Zhang, Xiangcheng, et al.
Published: (2025)
Critique to Verify: Accurate and Honest Test-Time Scaling with RL-Trained Verifiers
by: Yang, Zhicheng, et al.
Published: (2025)
by: Yang, Zhicheng, et al.
Published: (2025)
GGBall: Graph Generative Model on Poincaré Ball
by: Bu, Tianci, et al.
Published: (2025)
by: Bu, Tianci, et al.
Published: (2025)
Self-Verifying Reflection Helps Transformers with CoT Reasoning
by: Yu, Zhongwei, et al.
Published: (2025)
by: Yu, Zhongwei, et al.
Published: (2025)
Navigate the Unknown: Enhancing LLM Reasoning with Intrinsic Motivation Guided Exploration
by: Gao, Jingtong, et al.
Published: (2025)
by: Gao, Jingtong, et al.
Published: (2025)
PACE: Geometry-Aware Bridge Transport for Single-Cell Trajectory Inference
by: Yu, Chenglei, et al.
Published: (2026)
by: Yu, Chenglei, et al.
Published: (2026)
ATLAS: Adaptive Test-Time Latent Steering with External Verifiers for Enhancing LLMs Reasoning
by: Nguyen, Tuc, et al.
Published: (2026)
by: Nguyen, Tuc, et al.
Published: (2026)
Random Policy Valuation is Enough for LLM Reasoning with Verifiable Rewards
by: He, Haoran, et al.
Published: (2025)
by: He, Haoran, et al.
Published: (2025)
Training-Free Guidance Beyond Differentiability: Scalable Path Steering with Tree Search in Diffusion and Flow Models
by: Guo, Yingqing, et al.
Published: (2025)
by: Guo, Yingqing, et al.
Published: (2025)
Improving Reasoning for Diffusion Language Models via Group Diffusion Policy Optimization
by: Rojas, Kevin, et al.
Published: (2025)
by: Rojas, Kevin, et al.
Published: (2025)
On Learning Verifiers and Implications to Chain-of-Thought Reasoning
by: Balcan, Maria-Florina, et al.
Published: (2025)
by: Balcan, Maria-Florina, et al.
Published: (2025)
Test-time GNN Model Evaluation on Dynamic Graphs
by: Li, Bo, et al.
Published: (2025)
by: Li, Bo, et al.
Published: (2025)
Milestones over Outcome: Unlocking Geometric Reasoning with Sub-Goal Verifiable Reward
by: Chen, Jianlong, et al.
Published: (2026)
by: Chen, Jianlong, et al.
Published: (2026)
Putting the Value Back in RL: Better Test-Time Scaling by Unifying LLM Reasoners With Verifiers
by: Sareen, Kusha, et al.
Published: (2025)
by: Sareen, Kusha, et al.
Published: (2025)
Similar Items
-
One step further with Monte-Carlo sampler to guide diffusion better
by: Ren, Minsi, et al.
Published: (2026) -
Wavelet Diffusion Neural Operator
by: Hu, Peiyan, et al.
Published: (2024) -
From Uncertain to Safe: Conformal Fine-Tuning of Diffusion Models for Safe PDE Control
by: Hu, Peiyan, et al.
Published: (2025) -
CL-DiffPhyCon: Closed-loop Diffusion Control of Complex Physical Systems
by: Wei, Long, et al.
Published: (2024) -
On the Guidance of Flow Matching
by: Feng, Ruiqi, et al.
Published: (2025)