Apriel-1.5-15b-Thinker
Fuente:
arXiv
Saved in:
| Main Authors: | Radhakrishna, Shruthan, Tiwari, Aman, Shukla, Aanjaneya, Hashemi, Masoud, Maheshwary, Rishabh, Malay, Shiva Krishna Reddy, Mehta, Jash, Pattnaik, Pulkit, Mittal, Saloni, Slimi, Khalil, Ogueji, Kelechi, Oladipo, Akintunde, Parikh, Soham, Bamgbose, Oluwanifemi, Liang, Toby, Masry, Ahmed, Mahajan, Khyati, Mudumba, Sai Rajeswar, Yadav, Vikas, Madhusudhan, Sathwik Tejaswi, Scholak, Torsten, Davasam, Sagar, Sunkara, Srinivas, Chapados, Nicholas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Apriel-Nemotron-15B-Thinker
by: Radhakrishna, Shruthan, et al.
Published: (2025)
by: Radhakrishna, Shruthan, et al.
Published: (2025)
AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs
by: Nguyen, Hoang, et al.
Published: (2025)
by: Nguyen, Hoang, et al.
Published: (2025)
Super Apriel: One Checkpoint, Many Speeds
by: Labs, SLAM, et al.
Published: (2026)
by: Labs, SLAM, et al.
Published: (2026)
Curry-DPO: Enhancing Alignment using Curriculum Learning & Ranked Preferences
by: Pattnaik, Pulkit, et al.
Published: (2024)
by: Pattnaik, Pulkit, et al.
Published: (2024)
Apriel-H1: Towards Efficient Enterprise Reasoning Models
by: Ostapenko, Oleksiy, et al.
Published: (2025)
by: Ostapenko, Oleksiy, et al.
Published: (2025)
AprielGuard
by: Kasundra, Jaykumar, et al.
Published: (2025)
by: Kasundra, Jaykumar, et al.
Published: (2025)
M2Lingual: Enhancing Multilingual, Multi-Turn Instruction Alignment in Large Language Models
by: Maheshwary, Rishabh, et al.
Published: (2024)
by: Maheshwary, Rishabh, et al.
Published: (2024)
Improving Multilingual Math Reasoning for African Languages
by: Ogundepo, Odunayo, et al.
Published: (2025)
by: Ogundepo, Odunayo, et al.
Published: (2025)
DNR Bench: Benchmarking Over-Reasoning in Reasoning LLMs
by: Hashemi, Masoud, et al.
Published: (2025)
by: Hashemi, Masoud, et al.
Published: (2025)
Augmenting LLM Reasoning with Dynamic Notes Writing for Complex QA
by: Maheshwary, Rishabh, et al.
Published: (2025)
by: Maheshwary, Rishabh, et al.
Published: (2025)
AfroBench: How Good are Large Language Models on African Languages?
by: Ojo, Jessica, et al.
Published: (2023)
by: Ojo, Jessica, et al.
Published: (2023)
R2V Agent: Teaching SLMs When to Ask for Help
by: Hemadri, Raghu Vamshi, et al.
Published: (2026)
by: Hemadri, Raghu Vamshi, et al.
Published: (2026)
Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics
by: Nair, Jishnu Sethumadhavan, et al.
Published: (2026)
by: Nair, Jishnu Sethumadhavan, et al.
Published: (2026)
EnterpriseOps-Gym: Environments and Evaluations for Stateful Agentic Planning and Tool Use in Enterprise Settings
by: Malay, Shiva Krishna Reddy, et al.
Published: (2026)
by: Malay, Shiva Krishna Reddy, et al.
Published: (2026)
Revitalizing Saturated Benchmarks: A Weighted Metric Approach for Differentiating Large Language Model Performance
by: Etzine, Bryan, et al.
Published: (2025)
by: Etzine, Bryan, et al.
Published: (2025)
ColMate: Contrastive Late Interaction and Masked Text for Multimodal Document Retrieval
by: Masry, Ahmed, et al.
Published: (2025)
by: Masry, Ahmed, et al.
Published: (2025)
Cats Confuse Reasoning LLM: Query Agnostic Adversarial Triggers for Reasoning Models
by: Rajeev, Meghana, et al.
Published: (2025)
by: Rajeev, Meghana, et al.
Published: (2025)
Terminal Agents Suffice for Enterprise Automation
by: Bechard, Patrice, et al.
Published: (2026)
by: Bechard, Patrice, et al.
Published: (2026)
DeepSRGM -- Sequence Classification and Ranking in Indian Classical Music with Deep Learning
by: Madhusudhan, Sathwik Tejaswi, et al.
Published: (2024)
by: Madhusudhan, Sathwik Tejaswi, et al.
Published: (2024)
Practical Guide for Causal Pathways and Sub-group Disparity Analysis
by: Kohankhaki, Farnaz, et al.
Published: (2024)
by: Kohankhaki, Farnaz, et al.
Published: (2024)
Layer-Wise Quantization: A Pragmatic and Effective Method for Quantizing LLMs Beyond Integer Bit-Levels
by: Dumitru, Razvan-Gabriel, et al.
Published: (2024)
by: Dumitru, Razvan-Gabriel, et al.
Published: (2024)
Seventeen 2 Micron All Sky Survey (2MASS) hypervelocity stars (HVS) from Gaia DR3
by: Mudumba, Parthasarathy
Published: (2024)
by: Mudumba, Parthasarathy
Published: (2024)
Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models
by: Madhusudhan, Nishanth, et al.
Published: (2024)
by: Madhusudhan, Nishanth, et al.
Published: (2024)
Evaluating Robustness of Large Language Models in Enterprise Applications: Benchmarks for Perturbation Consistency Across Formats and Languages
by: Bogavelli, Tara, et al.
Published: (2026)
by: Bogavelli, Tara, et al.
Published: (2026)
BigCharts-R1: Enhanced Chart Reasoning with Visual Reinforcement Finetuning
by: Masry, Ahmed, et al.
Published: (2025)
by: Masry, Ahmed, et al.
Published: (2025)
Apriel-1.5-OpenReasoner: RL Post-Training for General-Purpose and Efficient Reasoning
by: Pardinas, Rafael, et al.
Published: (2026)
by: Pardinas, Rafael, et al.
Published: (2026)
Grammar Search for Multi-Agent Systems
by: Singh, Mayank, et al.
Published: (2025)
by: Singh, Mayank, et al.
Published: (2025)
FakeWatch: A Framework for Detecting Fake News to Ensure Credible Elections
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
by: Raza, Shaina, et al.
Published: (2023)
by: Raza, Shaina, et al.
Published: (2023)
Developing Safe and Responsible Large Language Model : Can We Balance Bias Reduction and Language Understanding in Large Language Models?
by: Raza, Shaina, et al.
Published: (2024)
by: Raza, Shaina, et al.
Published: (2024)
Pragmatic and Discourse Functions in Jenifa’s Diary
by: Ganiu Bamgbose
Published: (2021)
by: Ganiu Bamgbose
Published: (2021)
Multi-Reference Preference Optimization for Large Language Models
by: Le, Hung, et al.
Published: (2024)
by: Le, Hung, et al.
Published: (2024)
InsightBench: Evaluating Business Analytics Agents Through Multi-Step Insight Generation
by: Sahu, Gaurav, et al.
Published: (2024)
by: Sahu, Gaurav, et al.
Published: (2024)
AlignVLM: Bridging Vision and Language Latent Spaces for Multimodal Document Understanding
by: Masry, Ahmed, et al.
Published: (2025)
by: Masry, Ahmed, et al.
Published: (2025)
Prompting with Phonemes: Enhancing LLMs' Multilinguality for Non-Latin Script Languages
by: Nguyen, Hoang H, et al.
Published: (2024)
by: Nguyen, Hoang H, et al.
Published: (2024)
Spacerini: Plug-and-play Search Engines with Pyserini and Hugging Face
by: Akiki, Christopher, et al.
Published: (2023)
by: Akiki, Christopher, et al.
Published: (2023)
Acute toxicity and behavioural effects of chlorpyrifosethyl pesticide to juveniles of Clarias gariepinus Teugels
by: Auta, J., et al.
Published: (2008)
by: Auta, J., et al.
Published: (2008)
THE ADAM PARADOX HYPOTHESIS: A Seventeen-Concept Architectural Model for the Emergence of Purpose-Oriented Human Cognition
by: Roessaptono, Yulianto, et al.
Published: (2026)
by: Roessaptono, Yulianto, et al.
Published: (2026)
Perplexed: Understanding When Large Language Models are Confused
by: Cooper, Nathan, et al.
Published: (2024)
by: Cooper, Nathan, et al.
Published: (2024)
EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents
by: Bogavelli, Tara, et al.
Published: (2026)
by: Bogavelli, Tara, et al.
Published: (2026)
Similar Items
-
Apriel-Nemotron-15B-Thinker
by: Radhakrishna, Shruthan, et al.
Published: (2025) -
AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs
by: Nguyen, Hoang, et al.
Published: (2025) -
Super Apriel: One Checkpoint, Many Speeds
by: Labs, SLAM, et al.
Published: (2026) -
Curry-DPO: Enhancing Alignment using Curriculum Learning & Ranked Preferences
by: Pattnaik, Pulkit, et al.
Published: (2024) -
Apriel-H1: Towards Efficient Enterprise Reasoning Models
by: Ostapenko, Oleksiy, et al.
Published: (2025)