Do Math Reasoning LLMs Help Predict the Impact of Public Transit Events?
Fuente:
arXiv
Guardado en:
| Autores principales: | Fang, Bowen, Zha, Ruijian, Di, Xuan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LLM-based Realistic Safety-Critical Driving Video Generation
por: Fu, Yongjie, et al.
Publicado: (2025)
por: Fu, Yongjie, et al.
Publicado: (2025)
MARGE: Improving Math Reasoning for LLMs with Guided Exploration
por: Gao, Jingyue, et al.
Publicado: (2025)
por: Gao, Jingyue, et al.
Publicado: (2025)
Do Thinking Tokens Help or Trap? Towards More Efficient Large Reasoning Model
por: Ding, Bowen, et al.
Publicado: (2025)
por: Ding, Bowen, et al.
Publicado: (2025)
TraveLLM: Could you plan my new public transit route in face of a network disruption?
por: Fang, Bowen, et al.
Publicado: (2024)
por: Fang, Bowen, et al.
Publicado: (2024)
Learn to Tour: Operator Design For Solution Feasibility Mapping in Pickup-and-delivery Traveling Salesman Problem
por: Fang, Bowen, et al.
Publicado: (2024)
por: Fang, Bowen, et al.
Publicado: (2024)
SuperCLUE-Math6: Graded Multi-Step Math Reasoning Benchmark for LLMs in Chinese
por: Xu, Liang, et al.
Publicado: (2024)
por: Xu, Liang, et al.
Publicado: (2024)
MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning
por: Li, Chengpeng, et al.
Publicado: (2023)
por: Li, Chengpeng, et al.
Publicado: (2023)
DAG-Math: Graph-of-Thought Guided Mathematical Reasoning in LLMs
por: Zhang, Yuanhe, et al.
Publicado: (2025)
por: Zhang, Yuanhe, et al.
Publicado: (2025)
Can MLLMs Absorb Math Reasoning Abilities from LLMs as Free Lunch?
por: Hu, Yijie, et al.
Publicado: (2025)
por: Hu, Yijie, et al.
Publicado: (2025)
SenseMath: Do LLMs Have Number Sense? Evaluating Shortcut Use, Judgment, and Generation
por: Zhuang, Haomin, et al.
Publicado: (2026)
por: Zhuang, Haomin, et al.
Publicado: (2026)
Lost in Cultural Translation: Do LLMs Struggle with Math Across Cultural Contexts?
por: Karim, Aabid, et al.
Publicado: (2025)
por: Karim, Aabid, et al.
Publicado: (2025)
MathArena: Evaluating LLMs on Uncontaminated Math Competitions
por: Balunović, Mislav, et al.
Publicado: (2025)
por: Balunović, Mislav, et al.
Publicado: (2025)
An Empirical Study of Data Ability Boundary in LLMs' Math Reasoning
por: Chen, Zui, et al.
Publicado: (2024)
por: Chen, Zui, et al.
Publicado: (2024)
TaoBench: Do Automated Theorem Prover LLMs Generalize Beyond MathLib?
por: Taylor, Alexander K, et al.
Publicado: (2026)
por: Taylor, Alexander K, et al.
Publicado: (2026)
Bayesian Elicitation with LLMs: Model Size Helps, Extra "Reasoning" Doesn't Always
por: Hobor, Luka, et al.
Publicado: (2026)
por: Hobor, Luka, et al.
Publicado: (2026)
Neuro-Symbolic Data Generation for Math Reasoning
por: Li, Zenan, et al.
Publicado: (2024)
por: Li, Zenan, et al.
Publicado: (2024)
MultiMath: Bridging Visual and Mathematical Reasoning for Large Language Models
por: Peng, Shuai, et al.
Publicado: (2024)
por: Peng, Shuai, et al.
Publicado: (2024)
MathConstraint: Automated Generation of Verified Combinatorial Reasoning Instances for LLMs
por: Pati, Viresh, et al.
Publicado: (2026)
por: Pati, Viresh, et al.
Publicado: (2026)
MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
por: Huang, Kaixuan, et al.
Publicado: (2025)
por: Huang, Kaixuan, et al.
Publicado: (2025)
TabularMath: Understanding Math Reasoning over Tables with Large Language Models
por: Tian, Shi-Yu, et al.
Publicado: (2025)
por: Tian, Shi-Yu, et al.
Publicado: (2025)
OMEGA: Can LLMs Reason Outside the Box in Math? Evaluating Exploratory, Compositional, and Transformative Generalization
por: Sun, Yiyou, et al.
Publicado: (2025)
por: Sun, Yiyou, et al.
Publicado: (2025)
MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
por: Lu, Zimu, et al.
Publicado: (2024)
por: Lu, Zimu, et al.
Publicado: (2024)
Case-Based or Rule-Based: How Do Transformers Do the Math?
por: Hu, Yi, et al.
Publicado: (2024)
por: Hu, Yi, et al.
Publicado: (2024)
MathConstruct: Challenging LLM Reasoning with Constructive Proofs
por: Balunović, Mislav, et al.
Publicado: (2025)
por: Balunović, Mislav, et al.
Publicado: (2025)
Kwai-STaR: Transform LLMs into State-Transition Reasoners
por: Lu, Xingyu, et al.
Publicado: (2024)
por: Lu, Xingyu, et al.
Publicado: (2024)
Beyond Solving Math Quiz: Evaluating the Ability of Large Reasoning Models to Ask for Information
por: Huang, Youcheng, et al.
Publicado: (2025)
por: Huang, Youcheng, et al.
Publicado: (2025)
When Do LLMs Reason? A Dynamical Systems View via Entropy Phase Transitions
por: Xia, Wei, et al.
Publicado: (2026)
por: Xia, Wei, et al.
Publicado: (2026)
Data Diversification Methods In Alignment Enhance Math Performance In LLMs
por: Dokmeci, Berkan, et al.
Publicado: (2025)
por: Dokmeci, Berkan, et al.
Publicado: (2025)
MathMistake Checker: A Comprehensive Demonstration for Step-by-Step Math Problem Mistake Finding by Prompt-Guided LLMs
por: Zhang, Tianyang, et al.
Publicado: (2025)
por: Zhang, Tianyang, et al.
Publicado: (2025)
Mining Math Conjectures from LLMs: A Pruning Approach
por: Chuharski, Jake, et al.
Publicado: (2024)
por: Chuharski, Jake, et al.
Publicado: (2024)
AgenticMath: Enhancing LLM Reasoning via Agentic-based Math Data Generation
por: Liu, Xianyang, et al.
Publicado: (2025)
por: Liu, Xianyang, et al.
Publicado: (2025)
How Do Semantically Equivalent Code Transformations Impact Membership Inference on LLMs for Code?
por: Yang, Hua, et al.
Publicado: (2025)
por: Yang, Hua, et al.
Publicado: (2025)
Reasoning Curriculum: Bootstrapping Broad LLM Reasoning from Math
por: Pang, Bo, et al.
Publicado: (2025)
por: Pang, Bo, et al.
Publicado: (2025)
Integrating Visual Interpretation and Linguistic Reasoning for Math Problem Solving
por: Guo, Zixian, et al.
Publicado: (2025)
por: Guo, Zixian, et al.
Publicado: (2025)
Self-Consistency Boosts Calibration for Math Reasoning
por: Wang, Ante, et al.
Publicado: (2024)
por: Wang, Ante, et al.
Publicado: (2024)
Math Neurosurgery: Isolating Language Models' Math Reasoning Abilities Using Only Forward Passes
por: Christ, Bryan R., et al.
Publicado: (2024)
por: Christ, Bryan R., et al.
Publicado: (2024)
Constructing a 3D Scene from a Single Image
por: Zheng, Kaizhi, et al.
Publicado: (2025)
por: Zheng, Kaizhi, et al.
Publicado: (2025)
Research on Predicting Public Opinion Event Heat Levels Based on Large Language Models
por: Ren, Yi, et al.
Publicado: (2024)
por: Ren, Yi, et al.
Publicado: (2024)
AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
por: Liu, Zihan, et al.
Publicado: (2024)
por: Liu, Zihan, et al.
Publicado: (2024)
MV-MATH: Evaluating Multimodal Math Reasoning in Multi-Visual Contexts
por: Wang, Peijie, et al.
Publicado: (2025)
por: Wang, Peijie, et al.
Publicado: (2025)
Ejemplares similares
-
LLM-based Realistic Safety-Critical Driving Video Generation
por: Fu, Yongjie, et al.
Publicado: (2025) -
MARGE: Improving Math Reasoning for LLMs with Guided Exploration
por: Gao, Jingyue, et al.
Publicado: (2025) -
Do Thinking Tokens Help or Trap? Towards More Efficient Large Reasoning Model
por: Ding, Bowen, et al.
Publicado: (2025) -
TraveLLM: Could you plan my new public transit route in face of a network disruption?
por: Fang, Bowen, et al.
Publicado: (2024) -
Learn to Tour: Operator Design For Solution Feasibility Mapping in Pickup-and-delivery Traveling Salesman Problem
por: Fang, Bowen, et al.
Publicado: (2024)