DeepSeek-R1 Thoughtology: Let's think about LLM Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Marjanović, Sara Vera, Patel, Arkil, Adlakha, Vaibhav, Aghajohari, Milad, BehnamGhader, Parishad, Bhatia, Mehar, Khandelwal, Aditi, Kraft, Austin, Krojer, Benno, Lù, Xing Han, Meade, Nicholas, Shin, Dongchan, Kazemnejad, Amirhossein, Kamath, Gaurav, Mosbach, Marius, Stańczak, Karolina, Reddy, Siva |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploiting Instruction-Following Retrievers for Malicious Information Retrieval
by: BehnamGhader, Parishad, et al.
Published: (2025)
by: BehnamGhader, Parishad, et al.
Published: (2025)
Evaluating Correctness and Faithfulness of Instruction-Following Models for Question Answering
by: Adlakha, Vaibhav, et al.
Published: (2023)
by: Adlakha, Vaibhav, et al.
Published: (2023)
LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders
by: BehnamGhader, Parishad, et al.
Published: (2024)
by: BehnamGhader, Parishad, et al.
Published: (2024)
LLM2Vec-Gen: Generative Embeddings from Large Language Models
by: BehnamGhader, Parishad, et al.
Published: (2026)
by: BehnamGhader, Parishad, et al.
Published: (2026)
Value Drifts: Tracing Value Alignment During LLM Post-Training
by: Bhatia, Mehar, et al.
Published: (2025)
by: Bhatia, Mehar, et al.
Published: (2025)
LatentLens: Revealing Highly Interpretable Visual Tokens in LLMs
by: Krojer, Benno, et al.
Published: (2026)
by: Krojer, Benno, et al.
Published: (2026)
AgentRewardBench: Evaluating Automatic Evaluations of Web Agent Trajectories
by: Lù, Xing Han, et al.
Published: (2025)
by: Lù, Xing Han, et al.
Published: (2025)
Investigating Adversarial Trigger Transfer in Large Language Models
by: Meade, Nicholas, et al.
Published: (2024)
by: Meade, Nicholas, et al.
Published: (2024)
Forecasting Downstream Performance of LLMs With Proxy Metrics
by: Patel, Arkil, et al.
Published: (2026)
by: Patel, Arkil, et al.
Published: (2026)
The Markovian Thinker: Architecture-Agnostic Linear Scaling of Reasoning
by: Aghajohari, Milad, et al.
Published: (2025)
by: Aghajohari, Milad, et al.
Published: (2025)
VinePPO: Refining Credit Assignment in RL Training of LLMs
by: Kazemnejad, Amirhossein, et al.
Published: (2024)
by: Kazemnejad, Amirhossein, et al.
Published: (2024)
Understanding the Influence of Synthetic Data for Text Embedders
by: Springer, Jacob Mitchell, et al.
Published: (2025)
by: Springer, Jacob Mitchell, et al.
Published: (2025)
SafeArena: Evaluating the Safety of Autonomous Web Agents
by: Tur, Ada Defne, et al.
Published: (2025)
by: Tur, Ada Defne, et al.
Published: (2025)
Build the web for agents, not agents for the web
by: Lù, Xing Han, et al.
Published: (2025)
by: Lù, Xing Han, et al.
Published: (2025)
Quantifying Gender Biases Towards Politicians on Reddit
by: Marjanovic, Sara, et al.
Published: (2021)
by: Marjanovic, Sara, et al.
Published: (2021)
How to Get Your LLM to Generate Challenging Problems for Evaluation
by: Patel, Arkil, et al.
Published: (2025)
by: Patel, Arkil, et al.
Published: (2025)
Improving Automatic VQA Evaluation Using Large Language Models
by: Mañas, Oscar, et al.
Published: (2023)
by: Mañas, Oscar, et al.
Published: (2023)
The Promise of RL for Autoregressive Image Editing
by: Ahmadi, Saba, et al.
Published: (2025)
by: Ahmadi, Saba, et al.
Published: (2025)
Learning Action and Reasoning-Centric Image Editing from Videos and Simulations
by: Krojer, Benno, et al.
Published: (2024)
by: Krojer, Benno, et al.
Published: (2024)
Not All Data Are Unlearned Equally
by: Krishnan, Aravind, et al.
Published: (2025)
by: Krishnan, Aravind, et al.
Published: (2025)
Evaluating In-Context Learning of Libraries for Code Generation
by: Patel, Arkil, et al.
Published: (2023)
by: Patel, Arkil, et al.
Published: (2023)
LOQA: Learning with Opponent Q-Learning Awareness
by: Aghajohari, Milad, et al.
Published: (2024)
by: Aghajohari, Milad, et al.
Published: (2024)
Paremia about the Earth in the Russian Language: Linguocultural Research
by: Sedigheh Kazemnejad DAHKAEI
Published: (2020)
by: Sedigheh Kazemnejad DAHKAEI
Published: (2020)
Societal Alignment Frameworks Can Improve LLM Alignment
by: Stańczak, Karolina, et al.
Published: (2025)
by: Stańczak, Karolina, et al.
Published: (2025)
Best Response Shaping
by: Aghajohari, Milad, et al.
Published: (2024)
by: Aghajohari, Milad, et al.
Published: (2024)
Odo: Depth-Guided Diffusion for Identity-Preserving Body Reshaping
by: Khandelwal, Siddharth, et al.
Published: (2025)
by: Khandelwal, Siddharth, et al.
Published: (2025)
FragMOPs: Inverse design of Metal-Organic Polyhedra through Molecular Fragmentation and Evolutionary Optimisation
by: Butler, Patrick, et al.
Published: (2025)
by: Butler, Patrick, et al.
Published: (2025)
CulturalFrames: Assessing Cultural Expectation Alignment in Text-to-Image Models and Evaluation Metrics
by: Nayak, Shravan, et al.
Published: (2025)
by: Nayak, Shravan, et al.
Published: (2025)
Deep Learning-Based OFDM Receiver Using Fully Convolutional Neural Networks: Implementation, Evaluation, and Analysis
by: Ali, Ghader
Published: (2026)
by: Ali, Ghader
Published: (2026)
Learning Robust Social Strategies with Large Language Models
by: Piche, Dereck, et al.
Published: (2025)
by: Piche, Dereck, et al.
Published: (2025)
2D Model for Ca2+$Ca^{2+}$ Dynamics Regulating IP3$IP_3$, ATP and Insulin in A Pancreatic β$\beta$‐Cell
by: Vaishali Vaishali, et al.
Published: (2024)
by: Vaishali Vaishali, et al.
Published: (2024)
From Local Concepts to Universals: Evaluating the Multicultural Understanding of Vision-Language Models
by: Bhatia, Mehar, et al.
Published: (2024)
by: Bhatia, Mehar, et al.
Published: (2024)
A Shortcut-aware Video-QA Benchmark for Physical Understanding via Minimal Video Pairs
by: Krojer, Benno, et al.
Published: (2025)
by: Krojer, Benno, et al.
Published: (2025)
Language Models Largely Exhibit Human-like Constituent Ordering Preferences
by: Tur, Ada Defne, et al.
Published: (2025)
by: Tur, Ada Defne, et al.
Published: (2025)
A Multilingual Perspective on Probing Gender Bias
by: Stańczak, Karolina
Published: (2024)
by: Stańczak, Karolina
Published: (2024)
New Member Hub Coming Soon
by: Kelsey Stanczak
Published: (2024)
by: Kelsey Stanczak
Published: (2024)
Advantage Alignment Algorithms
by: Duque, Juan Agustin, et al.
Published: (2024)
by: Duque, Juan Agustin, et al.
Published: (2024)
Extended Low-Rank Approximation Accelerates Learning of Elastic Response in Heterogeneous Materials
by: Karmakar, Prabhat, et al.
Published: (2025)
by: Karmakar, Prabhat, et al.
Published: (2025)
فرص وتحديات تطوير الزراعة في موريتانيا
by: Abdel Ghader, Sidi Ahmed
Published: (2025)
by: Abdel Ghader, Sidi Ahmed
Published: (2025)
Impact of the honeycomb spin-lattice on topological magnons and edge states in ferromagnetic 2D skyrmion crystals
by: Ghader, Doried, et al.
Published: (2025)
by: Ghader, Doried, et al.
Published: (2025)
Similar Items
-
Exploiting Instruction-Following Retrievers for Malicious Information Retrieval
by: BehnamGhader, Parishad, et al.
Published: (2025) -
Evaluating Correctness and Faithfulness of Instruction-Following Models for Question Answering
by: Adlakha, Vaibhav, et al.
Published: (2023) -
LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders
by: BehnamGhader, Parishad, et al.
Published: (2024) -
LLM2Vec-Gen: Generative Embeddings from Large Language Models
by: BehnamGhader, Parishad, et al.
Published: (2026) -
Value Drifts: Tracing Value Alignment During LLM Post-Training
by: Bhatia, Mehar, et al.
Published: (2025)