VeRA: Verified Reasoning Data Augmentation at Scale
Fuente:
arXiv
Salvato in:
| Autori principali: | Cheng, Zerui, Liu, Jiashuo, Wu, Chunjie, Yao, Jianzhu, Viswanath, Pramod, Zhang, Ge, Huang, Wenhao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TabularMath: Evaluating Computational Extrapolation in Tabular Learning via Program-Verified Synthesis
di: Cheng, Zerui, et al.
Pubblicazione: (2026)
di: Cheng, Zerui, et al.
Pubblicazione: (2026)
SPIN-Bench: How Well Do LLMs Plan Strategically and Reason Socially?
di: Yao, Jianzhu, et al.
Pubblicazione: (2025)
di: Yao, Jianzhu, et al.
Pubblicazione: (2025)
TAO: Tolerance-Aware Optimistic Verification for Floating-Point Neural Networks
di: Yao, Jianzhu, et al.
Pubblicazione: (2025)
di: Yao, Jianzhu, et al.
Pubblicazione: (2025)
VeRA: Vector-based Random Matrix Adaptation
di: Kopiczko, Dawid J., et al.
Pubblicazione: (2023)
di: Kopiczko, Dawid J., et al.
Pubblicazione: (2023)
MEMO: Memory-Augmented Model Context Optimization for Robust Multi-Turn Multi-Agent LLM Games
di: Xie, Yunfei, et al.
Pubblicazione: (2026)
di: Xie, Yunfei, et al.
Pubblicazione: (2026)
VeRA+: Vector-Based Lightweight Digital Compensation for Drift-Resilient RRAM In-Memory Computing
di: Dong, Weirong, et al.
Pubblicazione: (2026)
di: Dong, Weirong, et al.
Pubblicazione: (2026)
LLM Swiss Round: Aggregating Multi-Benchmark Performance via Competitive Swiss-System Dynamics
di: Liu, Jiashuo, et al.
Pubblicazione: (2025)
di: Liu, Jiashuo, et al.
Pubblicazione: (2025)
Correct Answers from Sound Reasoning: Verifiable Process Supervision for Language Models
di: Kim, Kyuyoung, et al.
Pubblicazione: (2026)
di: Kim, Kyuyoung, et al.
Pubblicazione: (2026)
RLoop: An Self-Improving Framework for Reinforcement Learning with Iterative Policy Initialization
di: Zhiyuan, Zeng, et al.
Pubblicazione: (2025)
di: Zhiyuan, Zeng, et al.
Pubblicazione: (2025)
CHANCERY: Evaluating Corporate Governance Reasoning Capabilities in Language Models
di: Irwin, Lucas, et al.
Pubblicazione: (2025)
di: Irwin, Lucas, et al.
Pubblicazione: (2025)
SynLogic: Synthesizing Verifiable Reasoning Data at Scale for Learning Logical Reasoning and Beyond
di: Liu, Junteng, et al.
Pubblicazione: (2025)
di: Liu, Junteng, et al.
Pubblicazione: (2025)
CoVe: Training Interactive Tool-Use Agents via Constraint-Guided Verification
di: Chen, Jinpeng, et al.
Pubblicazione: (2026)
di: Chen, Jinpeng, et al.
Pubblicazione: (2026)
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
di: Zheng, Zihan, et al.
Pubblicazione: (2025)
di: Zheng, Zihan, et al.
Pubblicazione: (2025)
Geometric Point Attention Transformer for 3D Shape Reassembly
di: Li, Jiahan, et al.
Pubblicazione: (2024)
di: Li, Jiahan, et al.
Pubblicazione: (2024)
Mitigating LLM Hallucination via Behaviorally Calibrated Reinforcement Learning
di: Wu, Jiayun, et al.
Pubblicazione: (2025)
di: Wu, Jiayun, et al.
Pubblicazione: (2025)
Context manipulation attacks : Web agents are susceptible to corrupted memory
di: Patlan, Atharv Singh, et al.
Pubblicazione: (2025)
di: Patlan, Atharv Singh, et al.
Pubblicazione: (2025)
Data Heterogeneity Modeling for Trustworthy Machine Learning
di: Liu, Jiashuo, et al.
Pubblicazione: (2025)
di: Liu, Jiashuo, et al.
Pubblicazione: (2025)
TaSR-RAG: Taxonomy-guided Structured Reasoning for Retrieval-Augmented Generation
di: Sun, Jiashuo, et al.
Pubblicazione: (2026)
di: Sun, Jiashuo, et al.
Pubblicazione: (2026)
Training AI to be Loyal
di: Oh, Sewoong, et al.
Pubblicazione: (2025)
di: Oh, Sewoong, et al.
Pubblicazione: (2025)
Infinite Problem Generator: Verifiably Scaling Physics Reasoning Data with Agentic Workflows
di: Sharan, Aditya, et al.
Pubblicazione: (2026)
di: Sharan, Aditya, et al.
Pubblicazione: (2026)
AutoCode: LLMs as Problem Setters for Competitive Programming
di: Zhou, Shang, et al.
Pubblicazione: (2025)
di: Zhou, Shang, et al.
Pubblicazione: (2025)
VeLoRA: Memory Efficient Training using Rank-1 Sub-Token Projections
di: Miles, Roy, et al.
Pubblicazione: (2024)
di: Miles, Roy, et al.
Pubblicazione: (2024)
Scaling Retrieval-Augmented Reasoning with Parallel Search and Explicit Merging
di: Liu, Jiabei, et al.
Pubblicazione: (2026)
di: Liu, Jiabei, et al.
Pubblicazione: (2026)
Retrieval is Cheap, Show Me the Code: Executable Multi-Hop Reasoning for Retrieval-Augmented Generation
di: Sun, Jiashuo, et al.
Pubblicazione: (2026)
di: Sun, Jiashuo, et al.
Pubblicazione: (2026)
Verifiable Process Rewards for Agentic Reasoning
di: Yuan, Huining, et al.
Pubblicazione: (2026)
di: Yuan, Huining, et al.
Pubblicazione: (2026)
Logic-Regularized Verifier Elicits Reasoning from LLMs
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis
di: Shi, Yucheng, et al.
Pubblicazione: (2026)
di: Shi, Yucheng, et al.
Pubblicazione: (2026)
OML: A Primitive for Reconciling Open Access with Owner Control in AI Model Distribution
di: Cheng, Zerui, et al.
Pubblicazione: (2024)
di: Cheng, Zerui, et al.
Pubblicazione: (2024)
LIMOPro: Reasoning Refinement for Efficient and Effective Test-time Scaling
di: Xiao, Yang, et al.
Pubblicazione: (2025)
di: Xiao, Yang, et al.
Pubblicazione: (2025)
LegalGraphRAG: Multi-Agent Graph Retrieval-Augmented Generation for Reliable Legal Reasoning
di: Chen, Zerui, et al.
Pubblicazione: (2026)
di: Chen, Zerui, et al.
Pubblicazione: (2026)
StructVRM: Aligning Multimodal Reasoning with Structured and Verifiable Reward Models
di: Zhang, Xiangxiang, et al.
Pubblicazione: (2025)
di: Zhang, Xiangxiang, et al.
Pubblicazione: (2025)
Fuse, Reason and Verify: Geometry Problem Solving with Parsed Clauses from Diagram
di: Zhang, Ming-Liang, et al.
Pubblicazione: (2024)
di: Zhang, Ming-Liang, et al.
Pubblicazione: (2024)
FOL-Traces: Verified First-Order Logic Reasoning Traces at Scale
di: Lee, Isabelle, et al.
Pubblicazione: (2025)
di: Lee, Isabelle, et al.
Pubblicazione: (2025)
ATLAS: Autoformalizing Theorems through Lifting, Augmentation, and Synthesis of Data
di: Liu, Xiaoyang, et al.
Pubblicazione: (2025)
di: Liu, Xiaoyang, et al.
Pubblicazione: (2025)
MedCEG: Reinforcing Verifiable Medical Reasoning with Critical Evidence Graph
di: Mu, Linjie, et al.
Pubblicazione: (2025)
di: Mu, Linjie, et al.
Pubblicazione: (2025)
Scaling Spatial Reasoning in MLLMs through Programmatic Data Synthesis
di: Helu, Zhi, et al.
Pubblicazione: (2025)
di: Helu, Zhi, et al.
Pubblicazione: (2025)
VerifyBench: A Systematic Benchmark for Evaluating Reasoning Verifiers Across Domains
di: Li, Xuzhao, et al.
Pubblicazione: (2025)
di: Li, Xuzhao, et al.
Pubblicazione: (2025)
Time Series Reasoning via Process-Verifiable Thinking Data Synthesis and Scheduling for Tailored LLM Reasoning
di: Zhou, Jiahui, et al.
Pubblicazione: (2026)
di: Zhou, Jiahui, et al.
Pubblicazione: (2026)
LoRA-Gen: Specializing Large Language Model via Online LoRA Generation
di: Xiao, Yicheng, et al.
Pubblicazione: (2025)
di: Xiao, Yicheng, et al.
Pubblicazione: (2025)
TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL
di: Yang, Tianze, et al.
Pubblicazione: (2026)
di: Yang, Tianze, et al.
Pubblicazione: (2026)
Documenti analoghi
-
TabularMath: Evaluating Computational Extrapolation in Tabular Learning via Program-Verified Synthesis
di: Cheng, Zerui, et al.
Pubblicazione: (2026) -
SPIN-Bench: How Well Do LLMs Plan Strategically and Reason Socially?
di: Yao, Jianzhu, et al.
Pubblicazione: (2025) -
TAO: Tolerance-Aware Optimistic Verification for Floating-Point Neural Networks
di: Yao, Jianzhu, et al.
Pubblicazione: (2025) -
VeRA: Vector-based Random Matrix Adaptation
di: Kopiczko, Dawid J., et al.
Pubblicazione: (2023) -
MEMO: Memory-Augmented Model Context Optimization for Robust Multi-Turn Multi-Agent LLM Games
di: Xie, Yunfei, et al.
Pubblicazione: (2026)