Testing GPT-4-o1-preview on math and science problems: A follow-up study
Fuente:
arXiv
Guardado en:
| Autor principal: | Davis, Ernest |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Testing GPT-4 with Wolfram Alpha and Code Interpreter plug-ins on math and science problems
por: Davis, Ernest, et al.
Publicado: (2023)
por: Davis, Ernest, et al.
Publicado: (2023)
System 2 thinking in OpenAI's o1-preview model: Near-perfect performance on a mathematics exam
por: de Winter, Joost, et al.
Publicado: (2024)
por: de Winter, Joost, et al.
Publicado: (2024)
Can generative AI and ChatGPT outperform humans on cognitive-demanding problem-solving tasks in science?
por: Zhai, Xiaoming, et al.
Publicado: (2024)
por: Zhai, Xiaoming, et al.
Publicado: (2024)
Large language models eroding science understanding: an experimental study
por: Collins, Harry, et al.
Publicado: (2026)
por: Collins, Harry, et al.
Publicado: (2026)
ChatGPT-4 and other LLMs in the Turing Test: A Critical Analysis
por: Giunti, Marco
Publicado: (2025)
por: Giunti, Marco
Publicado: (2025)
Do AI assistants help students write formal specifications? A study with ChatGPT and the B-Method
por: Capozucca, Alfredo, et al.
Publicado: (2025)
por: Capozucca, Alfredo, et al.
Publicado: (2025)
RogueGPT: dis-ethical tuning transforms ChatGPT4 into a Rogue AI in 158 Words
por: Buscemi, Alessio, et al.
Publicado: (2024)
por: Buscemi, Alessio, et al.
Publicado: (2024)
GPT-4V Cannot Generate Radiology Reports Yet
por: Jiang, Yuyang, et al.
Publicado: (2024)
por: Jiang, Yuyang, et al.
Publicado: (2024)
Does GPT-4 surpass human performance in linguistic pragmatics?
por: Bojic, Ljubisa, et al.
Publicado: (2023)
por: Bojic, Ljubisa, et al.
Publicado: (2023)
Using GPT-4 to Augment Unbalanced Data for Automatic Scoring
por: Fang, Luyang, et al.
Publicado: (2023)
por: Fang, Luyang, et al.
Publicado: (2023)
Surprising gender biases in GPT
por: Fulgu, Raluca Alexandra, et al.
Publicado: (2024)
por: Fulgu, Raluca Alexandra, et al.
Publicado: (2024)
Is ChatGPT Massively Used by Students Nowadays? A Survey on the Use of Large Language Models such as ChatGPT in Educational Settings
por: Sublime, Jérémie, et al.
Publicado: (2024)
por: Sublime, Jérémie, et al.
Publicado: (2024)
Exploring Social Desirability Response Bias in Large Language Models: Evidence from GPT-4 Simulations
por: Lee, Sanguk, et al.
Publicado: (2024)
por: Lee, Sanguk, et al.
Publicado: (2024)
Hallucination, reliability, and the role of generative AI in science
por: Rathkopf, Charles
Publicado: (2025)
por: Rathkopf, Charles
Publicado: (2025)
Artificially intelligent agents in the social and behavioral sciences: A history and outlook
por: Holme, Petter, et al.
Publicado: (2025)
por: Holme, Petter, et al.
Publicado: (2025)
Chat-GPT: An AI Based Educational Revolution
por: Maric, Sasa, et al.
Publicado: (2025)
por: Maric, Sasa, et al.
Publicado: (2025)
Ethical ChatGPT: Concerns, Challenges, and Commandments
por: Zhou, Jianlong, et al.
Publicado: (2023)
por: Zhou, Jianlong, et al.
Publicado: (2023)
Ethical Implications of ChatGPT in Higher Education: A Scoping Review
por: Li, Ming, et al.
Publicado: (2023)
por: Li, Ming, et al.
Publicado: (2023)
A ChatGPT-based approach for questions generation in higher education
por: Vu, Sinh Trong, et al.
Publicado: (2025)
por: Vu, Sinh Trong, et al.
Publicado: (2025)
How to Capture and Study Conversations Between Research Participants and ChatGPT: GPT for Researchers (g4r.org)
por: Kim, Jin
Publicado: (2025)
por: Kim, Jin
Publicado: (2025)
Leveraging Lecture Content for Improved Feedback: Explorations with GPT-4 and Retrieval Augmented Generation
por: Jacobs, Sven, et al.
Publicado: (2024)
por: Jacobs, Sven, et al.
Publicado: (2024)
AI in data science education: experiences from the classroom
por: Hageman, J. A., et al.
Publicado: (2025)
por: Hageman, J. A., et al.
Publicado: (2025)
Potential Societal Biases of ChatGPT in Higher Education: A Scoping Review
por: Li, Ming, et al.
Publicado: (2023)
por: Li, Ming, et al.
Publicado: (2023)
Dissecting Bias of ChatGPT in College Major Recommendations
por: Zheng, Alex
Publicado: (2023)
por: Zheng, Alex
Publicado: (2023)
Between Search and Platform: ChatGPT Under the DSA
por: Lorente, Toni, et al.
Publicado: (2026)
por: Lorente, Toni, et al.
Publicado: (2026)
Independent Mobility GPT (IDM-GPT): A Self-Supervised Multi-Agent Large Language Model Framework for Customized Traffic Mobility Analysis Using Machine Learning Models
por: Yang, Fengze, et al.
Publicado: (2025)
por: Yang, Fengze, et al.
Publicado: (2025)
Learning-by-teaching with ChatGPT: The effect of teachable ChatGPT agent on programming education
por: Chen, Angxuan, et al.
Publicado: (2024)
por: Chen, Angxuan, et al.
Publicado: (2024)
A Study on the Vulnerability of Test Questions against ChatGPT-based Cheating
por: Ram, Shanker, et al.
Publicado: (2024)
por: Ram, Shanker, et al.
Publicado: (2024)
GPT as ghostwriter at the White House
por: Savoy, Jacques
Publicado: (2024)
por: Savoy, Jacques
Publicado: (2024)
Identifying and Improving Disability Bias in GPT-Based Resume Screening
por: Glazko, Kate, et al.
Publicado: (2024)
por: Glazko, Kate, et al.
Publicado: (2024)
Enhancing Supply Chain Resilience with Metaverse and ChatGPT Technologies
por: Sarhir, Oumaima
Publicado: (2025)
por: Sarhir, Oumaima
Publicado: (2025)
Examining GPT's Capability to Generate and Map Course Concepts and Their Relationship
por: Yang, Tianyuan, et al.
Publicado: (2025)
por: Yang, Tianyuan, et al.
Publicado: (2025)
Emotion-Aware Embedding Fusion in LLMs (Flan-T5, LLAMA 2, DeepSeek-R1, and ChatGPT 4) for Intelligent Response Generation
por: Rasool, Abdur, et al.
Publicado: (2024)
por: Rasool, Abdur, et al.
Publicado: (2024)
Advancing Mental Health Pre-Screening: A New Custom GPT for Psychological Distress Assessment
por: Tang, Jinwen, et al.
Publicado: (2024)
por: Tang, Jinwen, et al.
Publicado: (2024)
Bias in Decision-Making for AI's Ethical Dilemmas: A Comparative Study of ChatGPT and Claude
por: Xu, Wentao, et al.
Publicado: (2025)
por: Xu, Wentao, et al.
Publicado: (2025)
ChatGPT as speechwriter for the French presidents
por: Labbé, Dominique, et al.
Publicado: (2024)
por: Labbé, Dominique, et al.
Publicado: (2024)
Setting the AI Agenda -- Evidence from Sweden in the ChatGPT Era
por: Bruinsma, Bastiaan, et al.
Publicado: (2024)
por: Bruinsma, Bastiaan, et al.
Publicado: (2024)
Revolutionizing Undergraduate Learning: CourseGPT and Its Generative AI Advancements
por: Nazar, Ahmad M., et al.
Publicado: (2024)
por: Nazar, Ahmad M., et al.
Publicado: (2024)
Embracing the Generative AI Revolution: Advancing Tertiary Education in Cybersecurity with GPT
por: Nowrozy, Raza, et al.
Publicado: (2024)
por: Nowrozy, Raza, et al.
Publicado: (2024)
Open Source Language Models Can Provide Feedback: Evaluating LLMs' Ability to Help Students Using GPT-4-As-A-Judge
por: Koutcheme, Charles, et al.
Publicado: (2024)
por: Koutcheme, Charles, et al.
Publicado: (2024)
Ejemplares similares
-
Testing GPT-4 with Wolfram Alpha and Code Interpreter plug-ins on math and science problems
por: Davis, Ernest, et al.
Publicado: (2023) -
System 2 thinking in OpenAI's o1-preview model: Near-perfect performance on a mathematics exam
por: de Winter, Joost, et al.
Publicado: (2024) -
Can generative AI and ChatGPT outperform humans on cognitive-demanding problem-solving tasks in science?
por: Zhai, Xiaoming, et al.
Publicado: (2024) -
Large language models eroding science understanding: an experimental study
por: Collins, Harry, et al.
Publicado: (2026) -
ChatGPT-4 and other LLMs in the Turing Test: A Critical Analysis
por: Giunti, Marco
Publicado: (2025)