Saved in:
Similar Items
Olmo Hybrid: From Theory to Practice and Back
by: Merrill, William, et al.
Published: (2026)
by: Merrill, William, et al.
Published: (2026)
2 OLMo 2 Furious
by: OLMo, Team, et al.
Published: (2024)
by: OLMo, Team, et al.
Published: (2024)
Olmix: A Framework for Data Mixing Throughout LM Development
by: Chen, Mayee F., et al.
Published: (2026)
by: Chen, Mayee F., et al.
Published: (2026)
Large-Scale Data Selection for Instruction Tuning
by: Ivison, Hamish, et al.
Published: (2025)
by: Ivison, Hamish, et al.
Published: (2025)
olmOCR: Unlocking Trillions of Tokens in PDFs with Vision Language Models
by: Poznanski, Jake, et al.
Published: (2025)
by: Poznanski, Jake, et al.
Published: (2025)
olmOCR 2: Unit Test Rewards for Document OCR
by: Poznanski, Jake, et al.
Published: (2025)
by: Poznanski, Jake, et al.
Published: (2025)
FlexOlmo: Open Language Models for Flexible Data Use
by: Shi, Weijia, et al.
Published: (2025)
by: Shi, Weijia, et al.
Published: (2025)
Generalizing Verifiable Instruction Following
by: Pyatkin, Valentina, et al.
Published: (2025)
by: Pyatkin, Valentina, et al.
Published: (2025)
Establishing Task Scaling Laws via Compute-Efficient Model Ladders
by: Bhagia, Akshita, et al.
Published: (2024)
by: Bhagia, Akshita, et al.
Published: (2024)
The Battle of LLMs: A Comparative Study in Conversational QA Tasks
by: Rangapur, Aryan, et al.
Published: (2024)
by: Rangapur, Aryan, et al.
Published: (2024)
DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research
by: Shao, Rulin, et al.
Published: (2025)
by: Shao, Rulin, et al.
Published: (2025)
Neural Bayesian Filtering
by: Solinas, Christopher, et al.
Published: (2025)
by: Solinas, Christopher, et al.
Published: (2025)
How2Everything: Mining the Web for How-To Procedures to Evaluate and Improve LLMs
by: Chang, Yapei, et al.
Published: (2026)
by: Chang, Yapei, et al.
Published: (2026)
Meta-Reinforcement Learning with Self-Reflection for Agentic Search
by: Xiao, Teng, et al.
Published: (2026)
by: Xiao, Teng, et al.
Published: (2026)
Hybrid Preferences: Learning to Route Instances for Human vs. AI Feedback
by: Miranda, Lester James V., et al.
Published: (2024)
by: Miranda, Lester James V., et al.
Published: (2024)
Postmodernism in Youth Literature--A Road Away from the Reader?
by: Skjonsberg, Kari
Published: (1992)
by: Skjonsberg, Kari
Published: (1992)
Compositional Discrete Latent Code for High Fidelity, Productive Diffusion Models
by: Lavoie, Samuel, et al.
Published: (2025)
by: Lavoie, Samuel, et al.
Published: (2025)
KIWI: A Dataset of Knowledge-Intensive Writing Instructions for Answering Research Questions
by: Xu, Fangyuan, et al.
Published: (2024)
by: Xu, Fangyuan, et al.
Published: (2024)
A useful representation of TESS light curves
by: Poznanski, Dovi
Published: (2026)
by: Poznanski, Dovi
Published: (2026)
Tulu 3: Pushing Frontiers in Open Language Model Post-Training
by: Lambert, N., et al.
Published: (2025)
by: Lambert, N., et al.
Published: (2025)
E-learning open seminar on "Human–centered artificial intelligence in education: From theory to practice"
by: Anastasiades, Panagiotes, et al.
Published: (2025)
by: Anastasiades, Panagiotes, et al.
Published: (2025)
Organize the Web: Constructing Domains Enhances Pre-Training Data Curation
by: Wettig, Alexander, et al.
Published: (2025)
by: Wettig, Alexander, et al.
Published: (2025)
Financial Integration in the Americas, Changing Geopolitics and Brazilian Foreign Policy
by: Finbarr Murphy
Published: (2011)
by: Finbarr Murphy
Published: (2011)
The LE Club Project: A Synergy of Systems for the "Learning Environment."
by: Joy, Finbarr
Published: (1996)
by: Joy, Finbarr
Published: (1996)
[Redacted]
by: Williams, Tyler Michael
Published: (2025)
by: Williams, Tyler Michael
Published: (2025)
Gradient Regularization Prevents Reward Hacking in Reinforcement Learning from Human Feedback and Verifiable Rewards
by: Ackermann, Johannes, et al.
Published: (2026)
by: Ackermann, Johannes, et al.
Published: (2026)
Self-Directed Synthetic Dialogues and Revisions Technical Report
by: Lambert, Nathan, et al.
Published: (2024)
by: Lambert, Nathan, et al.
Published: (2024)
Duration Dependence and Heterogeneity: Learning from Early Notice of Layoff
by: Bhagia, Div
Published: (2023)
by: Bhagia, Div
Published: (2023)
Tulu 3: Pushing Frontiers in Open Language Model Post-Training
by: Lambert, Nathan, et al.
Published: (2024)
by: Lambert, Nathan, et al.
Published: (2024)
Learning Multi-Agent Communication with Contrastive Learning
by: Lo, Yat Long, et al.
Published: (2023)
by: Lo, Yat Long, et al.
Published: (2023)
DataDecide: How to Predict Best Pretraining Data with Small Experiments
by: Magnusson, Ian, et al.
Published: (2025)
by: Magnusson, Ian, et al.
Published: (2025)
IssueBench: Millions of Realistic Prompts for Measuring Issue Bias in LLM Writing Assistance
by: Röttger, Paul, et al.
Published: (2025)
by: Röttger, Paul, et al.
Published: (2025)
The Art of Saying No: Contextual Noncompliance in Language Models
by: Brahman, Faeze, et al.
Published: (2024)
by: Brahman, Faeze, et al.
Published: (2024)
Maintaining Privacy During the Intrahospital Transport of Anesthetized Children
by: Kevin Finbarr McCarthy
Published: (2025)
by: Kevin Finbarr McCarthy
Published: (2025)
A new species of the genus Rheobatrachus (Anura: Leptodactylidae) from Queensland
by: Mahony, Michael, et al.
Published: (1984)
by: Mahony, Michael, et al.
Published: (1984)
Ten simple rules for teaching data science
by: Timbers, Tiffany A., et al.
Published: (2026)
by: Timbers, Tiffany A., et al.
Published: (2026)
Political science : n introduction / Robert A. Heineman
by: Heineman, Robert A
Published: (1996)
by: Heineman, Robert A
Published: (1996)
Align and Filter: Improving Performance in Asynchronous On-Policy RL
by: Honari, Homayoun, et al.
Published: (2026)
by: Honari, Homayoun, et al.
Published: (2026)
OlmoEarth v1.1: A more efficient family of OlmoEarth models
by: Tseng, Gabriel, et al.
Published: (2026)
by: Tseng, Gabriel, et al.
Published: (2026)
What Makes it Ok to Set a Fire? Iterative Self-distillation of Contexts and Rationales for Disambiguating Defeasible Social and Moral Situations
by: Rao, Kavel, et al.
Published: (2023)
by: Rao, Kavel, et al.
Published: (2023)
Similar Items
-
Olmo Hybrid: From Theory to Practice and Back
by: Merrill, William, et al.
Published: (2026) -
2 OLMo 2 Furious
by: OLMo, Team, et al.
Published: (2024) -
Olmix: A Framework for Data Mixing Throughout LM Development
by: Chen, Mayee F., et al.
Published: (2026) -
Large-Scale Data Selection for Instruction Tuning
by: Ivison, Hamish, et al.
Published: (2025) -
olmOCR: Unlocking Trillions of Tokens in PDFs with Vision Language Models
by: Poznanski, Jake, et al.
Published: (2025)