Scaling Physical Reasoning with the PHYSICS Dataset
Fuente:
arXiv
Salvato in:
| Autori principali: | Zheng, Shenghe, Cheng, Qianjia, Yao, Junchi, Wu, Mengsong, He, Haonan, Ding, Ning, Cheng, Yu, Hu, Shuyue, Bai, Lei, Zhou, Dongzhan, Cui, Ganqu, Ye, Peng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PhysicsMinions: Winning Gold Medals in the Latest Physics Olympiads with a Coevolutionary Multimodal Multi-Agent System
di: Yu, Fangchen, et al.
Pubblicazione: (2025)
di: Yu, Fangchen, et al.
Pubblicazione: (2025)
HiPhO: How Far Are (M)LLMs from Humans in the Latest High School Physics Olympiad Benchmark?
di: Yu, Fangchen, et al.
Pubblicazione: (2025)
di: Yu, Fangchen, et al.
Pubblicazione: (2025)
P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads
di: Luo, Yun, et al.
Pubblicazione: (2026)
di: Luo, Yun, et al.
Pubblicazione: (2026)
P1: Mastering Physics Olympiads with Reinforcement Learning
di: Chen, Jiacheng, et al.
Pubblicazione: (2025)
di: Chen, Jiacheng, et al.
Pubblicazione: (2025)
SCI-Verifier: Scientific Verifier with Thinking
di: Zheng, Shenghe, et al.
Pubblicazione: (2025)
di: Zheng, Shenghe, et al.
Pubblicazione: (2025)
Teaching Thinking Models to Reason with Tools: A Full-Pipeline Recipe for Tool-Integrated Reasoning
di: Cheng, Qianjia, et al.
Pubblicazione: (2026)
di: Cheng, Qianjia, et al.
Pubblicazione: (2026)
Draft-OPD: On-Policy Distillation for Speculative Draft Models
di: Lei, Haodi, et al.
Pubblicazione: (2026)
di: Lei, Haodi, et al.
Pubblicazione: (2026)
Single-Agent Scaling Fails Multi-Agent Intelligence: Towards Foundation Models with Native Multi-Agent Intelligence
di: Hu, Shuyue, et al.
Pubblicazione: (2025)
di: Hu, Shuyue, et al.
Pubblicazione: (2025)
Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling
di: Li, Yafu, et al.
Pubblicazione: (2026)
di: Li, Yafu, et al.
Pubblicazione: (2026)
Reasoning via Video: The First Evaluation of Video Models' Reasoning Abilities through Maze-Solving Tasks
di: Yang, Cheng, et al.
Pubblicazione: (2025)
di: Yang, Cheng, et al.
Pubblicazione: (2025)
SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition
di: Wu, Mengsong, et al.
Pubblicazione: (2025)
di: Wu, Mengsong, et al.
Pubblicazione: (2025)
Decouple and Orthogonalize: A Data-Free Framework for LoRA Merging
di: Zheng, Shenghe, et al.
Pubblicazione: (2025)
di: Zheng, Shenghe, et al.
Pubblicazione: (2025)
UltraIF: Advancing Instruction Following from the Wild
di: An, Kaikai, et al.
Pubblicazione: (2025)
di: An, Kaikai, et al.
Pubblicazione: (2025)
TEMPO: Scaling Test-time Training for Large Reasoning Models
di: Zhang, Qingyang, et al.
Pubblicazione: (2026)
di: Zhang, Qingyang, et al.
Pubblicazione: (2026)
Dynamic Base model Shift for Delta Compression
di: Huang, Chenyu, et al.
Pubblicazione: (2025)
di: Huang, Chenyu, et al.
Pubblicazione: (2025)
Learning to Reason under Off-Policy Guidance
di: Yan, Jianhao, et al.
Pubblicazione: (2025)
di: Yan, Jianhao, et al.
Pubblicazione: (2025)
THE MEDICAL PHYSICS ON THE PAGES OF PHYSICS TODAY
di: Silva, Vinícius Carvalho da
Pubblicazione: (2022)
di: Silva, Vinícius Carvalho da
Pubblicazione: (2022)
FREE-Merging: Fourier Transform for Efficient Model Merging
di: Zheng, Shenghe, et al.
Pubblicazione: (2024)
di: Zheng, Shenghe, et al.
Pubblicazione: (2024)
The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
di: Cui, Ganqu, et al.
Pubblicazione: (2025)
di: Cui, Ganqu, et al.
Pubblicazione: (2025)
Seeing Delta Parameters as JPEG Images: Data-Free Delta Compression with Discrete Cosine Transform
di: Huang, Chenyu, et al.
Pubblicazione: (2025)
di: Huang, Chenyu, et al.
Pubblicazione: (2025)
Breaking the Compression Ceiling: Data-Free Pipeline for Ultra-Efficient Delta Compression
di: Wang, Xiaohui, et al.
Pubblicazione: (2025)
di: Wang, Xiaohui, et al.
Pubblicazione: (2025)
JustRL: Scaling a 1.5B LLM with a Simple RL Recipe
di: He, Bingxiang, et al.
Pubblicazione: (2025)
di: He, Bingxiang, et al.
Pubblicazione: (2025)
FRISM: Fine-Grained Reasoning Injection via Subspace-Level Model Merging for Vision-Language Models
di: Huang, Chenyu, et al.
Pubblicazione: (2026)
di: Huang, Chenyu, et al.
Pubblicazione: (2026)
Organizing, Orchestrating, and Benchmarking Agent Skills at Ecosystem Scale
di: Li, Hao, et al.
Pubblicazione: (2026)
di: Li, Hao, et al.
Pubblicazione: (2026)
UltraFeedback: Boosting Language Models with Scaled AI Feedback
di: Cui, Ganqu, et al.
Pubblicazione: (2023)
di: Cui, Ganqu, et al.
Pubblicazione: (2023)
View-Consistent 3D Scene Editing via Dual-Path Structural Correspondense and Semantic Continuity
di: Li, Pufan, et al.
Pubblicazione: (2026)
di: Li, Pufan, et al.
Pubblicazione: (2026)
SketchThinker-R1: Towards Efficient Sketch-Style Reasoning in Large Multimodal Models
di: Zhang, Ruiyang, et al.
Pubblicazione: (2026)
di: Zhang, Ruiyang, et al.
Pubblicazione: (2026)
AIR: A Systematic Analysis of Annotations, Instructions, and Response Pairs in Preference Dataset
di: He, Bingxiang, et al.
Pubblicazione: (2025)
di: He, Bingxiang, et al.
Pubblicazione: (2025)
THE PHYSICS OF THE CELL
di: ALEKOS CHARALAMPOPOULOS
Pubblicazione: (2025)
di: ALEKOS CHARALAMPOPOULOS
Pubblicazione: (2025)
Scalable Efficient Training of Large Language Models with Low-dimensional Projected Attention
di: Lv, Xingtai, et al.
Pubblicazione: (2024)
di: Lv, Xingtai, et al.
Pubblicazione: (2024)
LLM-Seg: Bridging Image Segmentation and Large Language Model Reasoning
di: Wang, Junchi, et al.
Pubblicazione: (2024)
di: Wang, Junchi, et al.
Pubblicazione: (2024)
Beyond GPT-5: Making LLMs Cheaper and Better via Performance-Efficiency Optimized Routing
di: Zhang, Yiqun, et al.
Pubblicazione: (2025)
di: Zhang, Yiqun, et al.
Pubblicazione: (2025)
Attention Reallocation: Towards Zero-cost and Controllable Hallucination Mitigation of MLLMs
di: Tu, Chongjun, et al.
Pubblicazione: (2025)
di: Tu, Chongjun, et al.
Pubblicazione: (2025)
Can Knowledge-Graph-based Retrieval Augmented Generation Really Retrieve What You Need?
di: Yu, Junchi, et al.
Pubblicazione: (2025)
di: Yu, Junchi, et al.
Pubblicazione: (2025)
Learning Compact Representations of LLM Abilities via Item Response Theory
di: Chen, Jianhao, et al.
Pubblicazione: (2025)
di: Chen, Jianhao, et al.
Pubblicazione: (2025)
THE PHYSICS TEACHERS HANDBOOK.
di: REDMAN, L.A.
Pubblicazione: (1966)
di: REDMAN, L.A.
Pubblicazione: (1966)
Beyond Gemini-3-Pro: Revisiting LLM Routing and Aggregation at Scale
di: Tang, Shengji, et al.
Pubblicazione: (2026)
di: Tang, Shengji, et al.
Pubblicazione: (2026)
A Unified Study of LoRA Variants: Taxonomy, Review, Codebase, and Empirical Evaluation
di: He, Haonan, et al.
Pubblicazione: (2026)
di: He, Haonan, et al.
Pubblicazione: (2026)
DONOD: Efficient and Generalizable Instruction Fine-Tuning for LLMs via Model-Intrinsic Dataset Pruning
di: Hu, Jucheng, et al.
Pubblicazione: (2025)
di: Hu, Jucheng, et al.
Pubblicazione: (2025)
V-Bridge: Bridging Video Generative Priors to Versatile Few-shot Image Restoration
di: Zheng, Shenghe, et al.
Pubblicazione: (2026)
di: Zheng, Shenghe, et al.
Pubblicazione: (2026)
Documenti analoghi
-
PhysicsMinions: Winning Gold Medals in the Latest Physics Olympiads with a Coevolutionary Multimodal Multi-Agent System
di: Yu, Fangchen, et al.
Pubblicazione: (2025) -
HiPhO: How Far Are (M)LLMs from Humans in the Latest High School Physics Olympiad Benchmark?
di: Yu, Fangchen, et al.
Pubblicazione: (2025) -
P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads
di: Luo, Yun, et al.
Pubblicazione: (2026) -
P1: Mastering Physics Olympiads with Reinforcement Learning
di: Chen, Jiacheng, et al.
Pubblicazione: (2025) -
SCI-Verifier: Scientific Verifier with Thinking
di: Zheng, Shenghe, et al.
Pubblicazione: (2025)