UltraLogic: Enhancing LLM Reasoning through Large-Scale Data Synthesis and Bipolar Float Reward
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yile, Liu, Yixian, Li, Zongwei, Huang, Yufei, Feng, Xinhua, Hu, Zhichao, Hu, Jinglu, Yan, Jianfeng, Lian, Fengzong, Liu, Yuhong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EternalMath: A Living Benchmark of Frontier Mathematics that Evolves with Human Discovery
von: Ma, Jicheng, et al.
Veröffentlicht: (2026)
von: Ma, Jicheng, et al.
Veröffentlicht: (2026)
MaXIFE: Multilingual and Cross-lingual Instruction Following Evaluation
von: Liu, Yile, et al.
Veröffentlicht: (2025)
von: Liu, Yile, et al.
Veröffentlicht: (2025)
Scaling the Scaling Logic: Agentic Meta-Synthesis of Logic Reasoning
von: Liu, Bowen, et al.
Veröffentlicht: (2026)
von: Liu, Bowen, et al.
Veröffentlicht: (2026)
DESIGNER: Design-Logic-Guided Multidisciplinary Data Synthesis for LLM Reasoning
von: Liu, Weize, et al.
Veröffentlicht: (2025)
von: Liu, Weize, et al.
Veröffentlicht: (2025)
Q-Save: Towards Scoring and Attribution for Generated Video Evaluation
von: Wu, Xiele, et al.
Veröffentlicht: (2025)
von: Wu, Xiele, et al.
Veröffentlicht: (2025)
Zero Reinforcement Learning Towards General Domains
von: Zeng, Yuyuan, et al.
Veröffentlicht: (2025)
von: Zeng, Yuyuan, et al.
Veröffentlicht: (2025)
Ge‐Based Visible‐Infrared Bipolar Floating‐Gate Phototransistor for Broad‐Spectrum Retinal Bionics
von: Qiancui Zhang, et al.
Veröffentlicht: (2025)
von: Qiancui Zhang, et al.
Veröffentlicht: (2025)
A First-Order Logic-Based Alternative to Reward Models in RLHF
von: Jian, Chunjin, et al.
Veröffentlicht: (2025)
von: Jian, Chunjin, et al.
Veröffentlicht: (2025)
Predictability‐Aware Subsequence Modeling for Sequential Recommendation
von: Hangyu Deng, et al.
Veröffentlicht: (2024)
von: Hangyu Deng, et al.
Veröffentlicht: (2024)
Safety Modulation: Enhancing Safety in Reinforcement Learning through Cost-Modulated Rewards
von: Zhang, Hanping, et al.
Veröffentlicht: (2025)
von: Zhang, Hanping, et al.
Veröffentlicht: (2025)
ReasonGRM: Enhancing Generative Reward Models through Large Reasoning Models
von: Chen, Bin, et al.
Veröffentlicht: (2025)
von: Chen, Bin, et al.
Veröffentlicht: (2025)
Adaptive Deep Reasoning: Triggering Deep Thinking When Needed
von: Wang, Yunhao, et al.
Veröffentlicht: (2025)
von: Wang, Yunhao, et al.
Veröffentlicht: (2025)
Truth Forest: Toward Multi-Scale Truthfulness in Large Language Models through Intervention without Tuning
von: Chen, Zhongzhi, et al.
Veröffentlicht: (2023)
von: Chen, Zhongzhi, et al.
Veröffentlicht: (2023)
LogicReward: Incentivizing LLM Reasoning via Step-Wise Logical Supervision
von: Xu, Jundong, et al.
Veröffentlicht: (2025)
von: Xu, Jundong, et al.
Veröffentlicht: (2025)
AutoRPA: Efficient GUI Automation through LLM-Driven Code Synthesis from Interactions
von: Chen, Minghao, et al.
Veröffentlicht: (2026)
von: Chen, Minghao, et al.
Veröffentlicht: (2026)
Logical Reasoning with Outcome Reward Models for Test-Time Scaling
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2025)
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2025)
LLM-Sketch: Enhancing Network Sketches with LLM
von: Li, Yuanpeng, et al.
Veröffentlicht: (2025)
von: Li, Yuanpeng, et al.
Veröffentlicht: (2025)
Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning
von: Hu, Yulan, et al.
Veröffentlicht: (2025)
von: Hu, Yulan, et al.
Veröffentlicht: (2025)
CWSSNet: Hyperspectral Image Classification Enhanced by Wavelet Domain Convolution
von: Tong, Yulin, et al.
Veröffentlicht: (2025)
von: Tong, Yulin, et al.
Veröffentlicht: (2025)
Enhancing LLM Reasoning with Reward-guided Tree Search
von: Jiang, Jinhao, et al.
Veröffentlicht: (2024)
von: Jiang, Jinhao, et al.
Veröffentlicht: (2024)
SynLogic: Synthesizing Verifiable Reasoning Data at Scale for Learning Logical Reasoning and Beyond
von: Liu, Junteng, et al.
Veröffentlicht: (2025)
von: Liu, Junteng, et al.
Veröffentlicht: (2025)
Logical Structure as Knowledge: Enhancing LLM Reasoning via Structured Logical Knowledge Density Estimation
von: Bi, Zhen, et al.
Veröffentlicht: (2025)
von: Bi, Zhen, et al.
Veröffentlicht: (2025)
Logit Poisoning Attack in Distillation-based Federated Learning and its Countermeasures
von: Yu, Yonghao, et al.
Veröffentlicht: (2024)
von: Yu, Yonghao, et al.
Veröffentlicht: (2024)
SIE3D: Single-Image Expressive 3D Avatar Generation via Semantic Embedding and Perceptual Expression Loss
von: Huang, Zhiqi, et al.
Veröffentlicht: (2025)
von: Huang, Zhiqi, et al.
Veröffentlicht: (2025)
DQN ‐Guided Subset‐Induced OCSVM Kernel Approximation for Imbalanced Anomaly Detection
von: Wenqian Yu, et al.
Veröffentlicht: (2026)
von: Wenqian Yu, et al.
Veröffentlicht: (2026)
LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts
von: Xiao, Yijia, et al.
Veröffentlicht: (2024)
von: Xiao, Yijia, et al.
Veröffentlicht: (2024)
Incentivizing Consistent, Effective and Scalable Reasoning Capability in Audio LLMs via Reasoning Process Rewards
von: Fan, Jiajun, et al.
Veröffentlicht: (2025)
von: Fan, Jiajun, et al.
Veröffentlicht: (2025)
SWE-AGILE: A Software Agent Framework for Efficiently Managing Dynamic Reasoning Context
von: Lian, Shuquan, et al.
Veröffentlicht: (2026)
von: Lian, Shuquan, et al.
Veröffentlicht: (2026)
GeoFM: Enhancing Geometric Reasoning of MLLMs via Synthetic Data Generation through Formal Language
von: Zhang, Yuhao, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2025)
Enhancing Multi-Hop Knowledge Graph Reasoning through Reward Shaping Techniques
von: Li, Chen, et al.
Veröffentlicht: (2024)
von: Li, Chen, et al.
Veröffentlicht: (2024)
DGRO: Enhancing LLM Reasoning via Exploration-Exploitation Control and Reward Variance Management
von: Su, Xuerui, et al.
Veröffentlicht: (2025)
von: Su, Xuerui, et al.
Veröffentlicht: (2025)
AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning
von: Wang, Zezhong, et al.
Veröffentlicht: (2025)
von: Wang, Zezhong, et al.
Veröffentlicht: (2025)
Toward Automated Simulation Research Workflow through LLM Prompt Engineering Design
von: Liu, Zhihan, et al.
Veröffentlicht: (2024)
von: Liu, Zhihan, et al.
Veröffentlicht: (2024)
FlowRL: Matching Reward Distributions for LLM Reasoning
von: Zhu, Xuekai, et al.
Veröffentlicht: (2025)
von: Zhu, Xuekai, et al.
Veröffentlicht: (2025)
RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time
von: Wang, Haozhe, et al.
Veröffentlicht: (2026)
von: Wang, Haozhe, et al.
Veröffentlicht: (2026)
Coordinated Control and Protection Strategy for Bipolar HVDC Grids Based on Improved Dual Half‐Bridge Submodule MMC Using Modulus Transformation
von: Yuhong Wang, et al.
Veröffentlicht: (2025)
von: Yuhong Wang, et al.
Veröffentlicht: (2025)
Adaptive Selection of Symbolic Languages for Improving LLM Logical Reasoning
von: Wang, Xiangyu, et al.
Veröffentlicht: (2025)
von: Wang, Xiangyu, et al.
Veröffentlicht: (2025)
Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning
von: Bi, Zhenni, et al.
Veröffentlicht: (2024)
von: Bi, Zhenni, et al.
Veröffentlicht: (2024)
Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning
von: Setlur, Amrith, et al.
Veröffentlicht: (2024)
von: Setlur, Amrith, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
EternalMath: A Living Benchmark of Frontier Mathematics that Evolves with Human Discovery
von: Ma, Jicheng, et al.
Veröffentlicht: (2026) -
MaXIFE: Multilingual and Cross-lingual Instruction Following Evaluation
von: Liu, Yile, et al.
Veröffentlicht: (2025) -
Scaling the Scaling Logic: Agentic Meta-Synthesis of Logic Reasoning
von: Liu, Bowen, et al.
Veröffentlicht: (2026) -
DESIGNER: Design-Logic-Guided Multidisciplinary Data Synthesis for LLM Reasoning
von: Liu, Weize, et al.
Veröffentlicht: (2025) -
Q-Save: Towards Scoring and Attribution for Generated Video Evaluation
von: Wu, Xiele, et al.
Veröffentlicht: (2025)