Saved in:
| Main Authors: | Halawi, Danny, Sarmasi, Aron, Saltzen, Siena, McCoy, Joshua |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2405.06846 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Not-So-Strange Love: Language Models and Generative Linguistic Theories are More Compatible than They Appear
by: McCoy, R. Thomas
Published: (2026)
by: McCoy, R. Thomas
Published: (2026)
Collocational bootstrapping: A hypothesis about the learning of subject-verb agreement in humans and neural networks
by: Hobbs, Claire, et al.
Published: (2026)
by: Hobbs, Claire, et al.
Published: (2026)
Overthinking the Truth: Understanding how Language Models Process False Demonstrations
by: Halawi, Danny, et al.
Published: (2023)
by: Halawi, Danny, et al.
Published: (2023)
GPU-Accelerated ANNS: Quantized for Speed, Built for Change
by: McCoy, Hunter, et al.
Published: (2026)
by: McCoy, Hunter, et al.
Published: (2026)
Approaching Human-Level Forecasting with Language Models
by: Halawi, Danny, et al.
Published: (2024)
by: Halawi, Danny, et al.
Published: (2024)
Deciphering the Factors Influencing the Efficacy of Chain-of-Thought: Probability, Memorization, and Noisy Reasoning
by: Prabhakar, Akshara, et al.
Published: (2024)
by: Prabhakar, Akshara, et al.
Published: (2024)
Distilling Symbolic Priors for Concept Learning into Neural Networks
by: Marinescu, Ioana, et al.
Published: (2024)
by: Marinescu, Ioana, et al.
Published: (2024)
ForecastBench: A Dynamic Benchmark of AI Forecasting Capabilities
by: Karger, Ezra, et al.
Published: (2024)
by: Karger, Ezra, et al.
Published: (2024)
Minimization of Boolean Complexity in In-Context Concept Learning
by: Wang, Leroy Z., et al.
Published: (2024)
by: Wang, Leroy Z., et al.
Published: (2024)
When a language model is optimized for reasoning, does it still show embers of autoregression? An analysis of OpenAI o1
by: McCoy, R. Thomas, et al.
Published: (2024)
by: McCoy, R. Thomas, et al.
Published: (2024)
Covert Malicious Finetuning: Challenges in Safeguarding LLM Adaptation
by: Halawi, Danny, et al.
Published: (2024)
by: Halawi, Danny, et al.
Published: (2024)
Global-Liar: Factuality of LLMs over Time and Geographic Regions
by: Mirza, Shujaat, et al.
Published: (2024)
by: Mirza, Shujaat, et al.
Published: (2024)
Understanding Inequality of LLM Fact-Checking over Geographic Regions with Agent and Retrieval models
by: Coelho, Bruno, et al.
Published: (2025)
by: Coelho, Bruno, et al.
Published: (2025)
Teasing Apart Architecture and Initial Weights as Sources of Inductive Bias in Neural Networks
by: Bencomo, Gianluca, et al.
Published: (2025)
by: Bencomo, Gianluca, et al.
Published: (2025)
Whither symbols in the era of advanced neural networks?
by: Griffiths, Thomas L., et al.
Published: (2025)
by: Griffiths, Thomas L., et al.
Published: (2025)
Identifying and Mitigating the Influence of the Prior Distribution in Large Language Models
by: Zhang, Liyi, et al.
Published: (2025)
by: Zhang, Liyi, et al.
Published: (2025)
Reproducibility: The New Frontier in AI Governance
by: Mason-Williams, Israel, et al.
Published: (2025)
by: Mason-Williams, Israel, et al.
Published: (2025)
Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit
by: Freeman, Joshua, et al.
Published: (2024)
by: Freeman, Joshua, et al.
Published: (2024)
Towards Data Governance of Frontier AI Models
by: Hausenloy, Jason, et al.
Published: (2024)
by: Hausenloy, Jason, et al.
Published: (2024)
ARC-AGI-2: A New Challenge for Frontier AI Reasoning Systems
by: Chollet, Francois, et al.
Published: (2025)
by: Chollet, Francois, et al.
Published: (2025)
AI Mathematician: Towards Fully Automated Frontier Mathematical Research
by: Liu, Yuanhang, et al.
Published: (2025)
by: Liu, Yuanhang, et al.
Published: (2025)
ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry
by: Xu, Tianze, et al.
Published: (2025)
by: Xu, Tianze, et al.
Published: (2025)
Formal Mathematical Reasoning: A New Frontier in AI
by: Yang, Kaiyu, et al.
Published: (2024)
by: Yang, Kaiyu, et al.
Published: (2024)
GRR-CoCa: Leveraging LLM Mechanisms in Multimodal Model Architectures
by: Patock, Jake R., et al.
Published: (2025)
by: Patock, Jake R., et al.
Published: (2025)
AIRS-Bench: a Suite of Tasks for Frontier AI Research Science Agents
by: Lupidi, Alisia, et al.
Published: (2026)
by: Lupidi, Alisia, et al.
Published: (2026)
Probabilistic Programming with Programmable Variational Inference
by: Becker, McCoy R., et al.
Published: (2024)
by: Becker, McCoy R., et al.
Published: (2024)
Responsible Reporting for Frontier AI Development
by: Kolt, Noam, et al.
Published: (2024)
by: Kolt, Noam, et al.
Published: (2024)
Advancing the Search Frontier with AI Agents
by: White, Ryen W.
Published: (2023)
by: White, Ryen W.
Published: (2023)
Governing AI Beyond the Pretraining Frontier
by: Caputo, Nicholas A.
Published: (2025)
by: Caputo, Nicholas A.
Published: (2025)
Open-World Evaluations for Measuring Frontier AI Capabilities
by: Kapoor, Sayash, et al.
Published: (2026)
by: Kapoor, Sayash, et al.
Published: (2026)
FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI
by: Glazer, Elliot, et al.
Published: (2024)
by: Glazer, Elliot, et al.
Published: (2024)
ARC-AGI-3: A New Challenge for Frontier Agentic Intelligence
by: Foundation, ARC Prize
Published: (2026)
by: Foundation, ARC Prize
Published: (2026)
LLM-Powered Swarms: A New Frontier or a Conceptual Stretch?
by: Rahman, Muhammad Atta Ur, et al.
Published: (2025)
by: Rahman, Muhammad Atta Ur, et al.
Published: (2025)
Combinatorial Creativity: A New Frontier in Generalization Abilities
by: Schapiro, Samuel, et al.
Published: (2025)
by: Schapiro, Samuel, et al.
Published: (2025)
Creativity or Brute Force? Using Brainteasers as a Window into the Problem-Solving Abilities of Large Language Models
by: Han, Simeng, et al.
Published: (2025)
by: Han, Simeng, et al.
Published: (2025)
Studies in Teaching: 2023 Research Digest. Action Research Projects Presented at Annual Research Forum (Winston-Salem, North Carolina, June 29, 2023)
by: McCoy, Leah P., Ed.
Published: (2023)
by: McCoy, Leah P., Ed.
Published: (2023)
A Survey of Generative AI for de novo Drug Design: New Frontiers in Molecule and Protein Generation
by: Tang, Xiangru, et al.
Published: (2024)
by: Tang, Xiangru, et al.
Published: (2024)
Emerging Practices in Frontier AI Safety Frameworks
by: Buhl, Marie Davidsen, et al.
Published: (2025)
by: Buhl, Marie Davidsen, et al.
Published: (2025)
Distance Learning Resistance in Higher Ed
by: Stephanie McCoy
Published: (2024)
by: Stephanie McCoy
Published: (2024)
Accessibility Services in Higher Ed: What Qualifies?
by: Stephanie McCoy
Published: (2024)
by: Stephanie McCoy
Published: (2024)
Similar Items
-
Not-So-Strange Love: Language Models and Generative Linguistic Theories are More Compatible than They Appear
by: McCoy, R. Thomas
Published: (2026) -
Collocational bootstrapping: A hypothesis about the learning of subject-verb agreement in humans and neural networks
by: Hobbs, Claire, et al.
Published: (2026) -
Overthinking the Truth: Understanding how Language Models Process False Demonstrations
by: Halawi, Danny, et al.
Published: (2023) -
GPU-Accelerated ANNS: Quantized for Speed, Built for Change
by: McCoy, Hunter, et al.
Published: (2026) -
Approaching Human-Level Forecasting with Language Models
by: Halawi, Danny, et al.
Published: (2024)