Phi-4-reasoning Technical Report
Fuente:
arXiv
Saved in:
| Main Authors: | Abdin, Marah, Agarwal, Sahaj, Awadallah, Ahmed, Balachandran, Vidhisha, Behl, Harkirat, Chen, Lingjiao, de Rosa, Gustavo, Gunasekar, Suriya, Javaheripi, Mojan, Joshi, Neel, Kauffmann, Piero, Lara, Yash, Mendes, Caio César Teodoro, Mitra, Arindam, Nushi, Besmira, Papailiopoulos, Dimitris, Saarikivi, Olli, Shah, Shital, Shrivastava, Vaishnavi, Vineet, Vibhav, Wu, Yue, Yousefi, Safoora, Zheng, Guoqing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sample More to Think Less: Group Filtered Policy Optimization for Concise Reasoning
by: Shrivastava, Vaishnavi, et al.
Published: (2025)
by: Shrivastava, Vaishnavi, et al.
Published: (2025)
Improving Instruction-Following in Language Models through Activation Steering
by: Stolfo, Alessandro, et al.
Published: (2024)
by: Stolfo, Alessandro, et al.
Published: (2024)
Tracing the Traces: Latent Temporal Signals for Efficient and Accurate Reasoning
by: Vilas, Martina G., et al.
Published: (2025)
by: Vilas, Martina G., et al.
Published: (2025)
Phi-4 Technical Report
by: Abdin, Marah, et al.
Published: (2024)
by: Abdin, Marah, et al.
Published: (2024)
Inference-Time Scaling for Complex Tasks: Where We Stand and What Lies Ahead
by: Balachandran, Vidhisha, et al.
Published: (2025)
by: Balachandran, Vidhisha, et al.
Published: (2025)
Unearthing Skill-Level Insights for Understanding Trade-Offs of Foundation Models
by: Moayeri, Mazda, et al.
Published: (2024)
by: Moayeri, Mazda, et al.
Published: (2024)
Eureka: Evaluating and Understanding Large Foundation Models
by: Balachandran, Vidhisha, et al.
Published: (2024)
by: Balachandran, Vidhisha, et al.
Published: (2024)
MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation
by: Joshi, Siddharth, et al.
Published: (2025)
by: Joshi, Siddharth, et al.
Published: (2025)
BenchAgents: Multi-Agent Systems for Structured Benchmark Creation
by: Butt, Natasha, et al.
Published: (2024)
by: Butt, Natasha, et al.
Published: (2024)
Just Do It!? Computer-Use Agents Exhibit Blind Goal-Directedness
by: Shayegani, Erfan, et al.
Published: (2025)
by: Shayegani, Erfan, et al.
Published: (2025)
Attention Satisfies: A Constraint-Satisfaction Lens on Factual Errors of Language Models
by: Yuksekgonul, Mert, et al.
Published: (2023)
by: Yuksekgonul, Mert, et al.
Published: (2023)
PEEKABOO: Interactive Video Generation via Masked-Diffusion
by: Jain, Yash, et al.
Published: (2023)
by: Jain, Yash, et al.
Published: (2023)
ZEBRAARENA: A Diagnostic Simulation Environment for Studying Reasoning-Action Coupling in Tool-Augmented LLMs
by: Zhao, Wanjia, et al.
Published: (2026)
by: Zhao, Wanjia, et al.
Published: (2026)
AI Scientist via Synthetic Task Scaling
by: Cai, Ziyang, et al.
Published: (2026)
by: Cai, Ziyang, et al.
Published: (2026)
Attention Speaks Volumes: Localizing and Mitigating Bias in Language Models
by: Adiga, Rishabh, et al.
Published: (2024)
by: Adiga, Rishabh, et al.
Published: (2024)
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs
by: Cai, Yanan, et al.
Published: (2025)
by: Cai, Yanan, et al.
Published: (2025)
Elephants Never Forget: Memorization and Learning of Tabular Data in Large Language Models
by: Bordt, Sebastian, et al.
Published: (2024)
by: Bordt, Sebastian, et al.
Published: (2024)
Detecting Data Contamination in LLMs via In-Context Learning
by: Zawalski, Michał, et al.
Published: (2025)
by: Zawalski, Michał, et al.
Published: (2025)
Diversity of Thought Improves Reasoning Abilities of LLMs
by: Naik, Ranjita, et al.
Published: (2023)
by: Naik, Ranjita, et al.
Published: (2023)
MEMENTO: Teaching LLMs to Manage Their Own Context
by: Kontonis, Vasilis, et al.
Published: (2026)
by: Kontonis, Vasilis, et al.
Published: (2026)
The Implicit Bias of Gradient Descent on Separable Data
by: Soudry, Daniel, et al.
Published: (2017)
by: Soudry, Daniel, et al.
Published: (2017)
Understanding Information Storage and Transfer in Multi-modal Large Language Models
by: Basu, Samyadeep, et al.
Published: (2024)
by: Basu, Samyadeep, et al.
Published: (2024)
A Framework for Fine-Grained Synchronization of Dependent GPU Kernels
by: Jangda, Abhinav, et al.
Published: (2023)
by: Jangda, Abhinav, et al.
Published: (2023)
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
by: Abdin, Marah, et al.
Published: (2024)
by: Abdin, Marah, et al.
Published: (2024)
Decoding In-Context Learning: Neuroscience-inspired Analysis of Representations in Large Language Models
by: Yousefi, Safoora, et al.
Published: (2023)
by: Yousefi, Safoora, et al.
Published: (2023)
Reasoning Up the Instruction Ladder for Controllable Language Models
by: Zheng, Zishuo, et al.
Published: (2025)
by: Zheng, Zishuo, et al.
Published: (2025)
Scaling the Convex Barrier with Sparse Dual Algorithms
by: De Palma, Alessandro, et al.
Published: (2021)
by: De Palma, Alessandro, et al.
Published: (2021)
Wait, Wait, Wait... Why Do Reasoning Models Loop?
by: Pipis, Charilaos, et al.
Published: (2025)
by: Pipis, Charilaos, et al.
Published: (2025)
Don't Hallucinate, Abstain: Identifying LLM Knowledge Gaps via Multi-LLM Collaboration
by: Feng, Shangbin, et al.
Published: (2024)
by: Feng, Shangbin, et al.
Published: (2024)
Knowledge Card: Filling LLMs' Knowledge Gaps with Plug-in Specialized Language Models
by: Feng, Shangbin, et al.
Published: (2023)
by: Feng, Shangbin, et al.
Published: (2023)
External Demand, Domestic Monetary Conditions, and Remittance Dynamics in Nepal
by: Malla, Sahaj Raj
Published: (2026)
by: Malla, Sahaj Raj
Published: (2026)
Devanagari Digit Recognition using Quantum Machine Learning
by: Malla, Sahaj Raj
Published: (2025)
by: Malla, Sahaj Raj
Published: (2025)
AI Agents for the Dhumbal Card Game: A Comparative Study
by: Malla, Sahaj Raj
Published: (2025)
by: Malla, Sahaj Raj
Published: (2025)
Kalimati Vegetable Price Index Forecasting with a Momentum Corrected Online Stacking Ensemble
by: Malla, Sahaj Raj
Published: (2026)
by: Malla, Sahaj Raj
Published: (2026)
Artful Dodgers
by: Gubar, Marah
Published: (2024)
by: Gubar, Marah
Published: (2024)
On the Diversity of Synthetic Data and its Impact on Training Large Language Models
by: Chen, Hao, et al.
Published: (2024)
by: Chen, Hao, et al.
Published: (2024)
Imprints of the operator ordering ambiguity on the dynamics of perfect fluid dominated quantum universe
by: Sahota, Harkirat Singh
Published: (2023)
by: Sahota, Harkirat Singh
Published: (2023)
FACTS&EVIDENCE: An Interactive Tool for Transparent Fine-Grained Factual Verification of Machine-Generated Text
by: Boonsanong, Varich, et al.
Published: (2025)
by: Boonsanong, Varich, et al.
Published: (2025)
R-LAM: Reproducibility-Constrained Large Action Models for Scientific Workflow Automation
by: Sureshkumar, Suriya
Published: (2026)
by: Sureshkumar, Suriya
Published: (2026)
CHIMERA: Compact Synthetic Data for Generalizable LLM Reasoning
by: Zhu, Xinyu, et al.
Published: (2026)
by: Zhu, Xinyu, et al.
Published: (2026)
Similar Items
-
Sample More to Think Less: Group Filtered Policy Optimization for Concise Reasoning
by: Shrivastava, Vaishnavi, et al.
Published: (2025) -
Improving Instruction-Following in Language Models through Activation Steering
by: Stolfo, Alessandro, et al.
Published: (2024) -
Tracing the Traces: Latent Temporal Signals for Efficient and Accurate Reasoning
by: Vilas, Martina G., et al.
Published: (2025) -
Phi-4 Technical Report
by: Abdin, Marah, et al.
Published: (2024) -
Inference-Time Scaling for Complex Tasks: Where We Stand and What Lies Ahead
by: Balachandran, Vidhisha, et al.
Published: (2025)