Chatbots put to the test in math and logic problems: A preliminary comparison and assessment of ChatGPT-3.5, ChatGPT-4, and Google Bard
Fuente:
arXiv
Guardado en:
| Autores principales: | Plevris, Vagelis, Papazafeiropoulos, George, Rios, Alejandro Jiménez |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Examining Linguistic Shifts in Academic Writing Before and After the Launch of ChatGPT: A Study on Preprint Papers
por: Bao, Tong, et al.
Publicado: (2025)
por: Bao, Tong, et al.
Publicado: (2025)
Trusting the Search: Unraveling Human Trust in Health Information from Google and ChatGPT
por: Sun, Xin, et al.
Publicado: (2024)
por: Sun, Xin, et al.
Publicado: (2024)
An Iterative Optimizing Framework for Radiology Report Summarization with ChatGPT
por: Ma, Chong, et al.
Publicado: (2023)
por: Ma, Chong, et al.
Publicado: (2023)
ChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level Generation
por: Taveekitworachai, Pittawat, et al.
Publicado: (2024)
por: Taveekitworachai, Pittawat, et al.
Publicado: (2024)
Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models
por: Venkatasubramanian, Venkat, et al.
Publicado: (2024)
por: Venkatasubramanian, Venkat, et al.
Publicado: (2024)
ChatGPT4PCG Competition: Character-like Level Generation for Science Birds
por: Taveekitworachai, Pittawat, et al.
Publicado: (2023)
por: Taveekitworachai, Pittawat, et al.
Publicado: (2023)
Detection of ChatGPT Fake Science with the xFakeSci Learning Algorithm
por: Hamed, Ahmed Abdeen, et al.
Publicado: (2023)
por: Hamed, Ahmed Abdeen, et al.
Publicado: (2023)
From Knowledge Generation to Knowledge Verification: Examining the BioMedical Generative Capabilities of ChatGPT
por: Hamed, Ahmed Abdeen, et al.
Publicado: (2025)
por: Hamed, Ahmed Abdeen, et al.
Publicado: (2025)
ChatGPT Based Data Augmentation for Improved Parameter-Efficient Debiasing of LLMs
por: Han, Pengrui, et al.
Publicado: (2024)
por: Han, Pengrui, et al.
Publicado: (2024)
Non-native speakers of English or ChatGPT: Who thinks better?
por: Shormani, Mohammed Q.
Publicado: (2024)
por: Shormani, Mohammed Q.
Publicado: (2024)
ChatGPT, Let us Chat Sign Language: Experiments, Architectural Elements, Challenges and Research Directions
por: Shahin, Nada, et al.
Publicado: (2024)
por: Shahin, Nada, et al.
Publicado: (2024)
Can ChatGPT capture swearing nuances? Evidence from translating Arabic oaths
por: Shormani, Mohammed Q.
Publicado: (2024)
por: Shormani, Mohammed Q.
Publicado: (2024)
On the robustness of ChatGPT in teaching Korean Mathematics
por: Nguyen, Phuong-Nam, et al.
Publicado: (2025)
por: Nguyen, Phuong-Nam, et al.
Publicado: (2025)
The impact of postediting on AI generative translation in Yemeni context: Translating literary prose by ChatGPT
por: Al-wagieh, Nasim, et al.
Publicado: (2026)
por: Al-wagieh, Nasim, et al.
Publicado: (2026)
Enhancing Emotion Prediction in News Headlines: Insights from ChatGPT and Seq2Seq Models for Free-Text Generation
por: Gao, Ge, et al.
Publicado: (2024)
por: Gao, Ge, et al.
Publicado: (2024)
ChatGPT and Gemini participated in the Korean College Scholastic Ability Test -- Earth Science I
por: Ga, Seok-Hyun, et al.
Publicado: (2025)
por: Ga, Seok-Hyun, et al.
Publicado: (2025)
A Study on the Vulnerability of Test Questions against ChatGPT-based Cheating
por: Ram, Shanker, et al.
Publicado: (2024)
por: Ram, Shanker, et al.
Publicado: (2024)
On Fusing ChatGPT and Ensemble Learning in Discon-tinuous Named Entity Recognition in Health Corpora
por: Chen, Tzu-Chieh, et al.
Publicado: (2024)
por: Chen, Tzu-Chieh, et al.
Publicado: (2024)
ChatGPT for Conversational Recommendation: Refining Recommendations by Reprompting with Feedback
por: Spurlock, Kyle Dylan, et al.
Publicado: (2024)
por: Spurlock, Kyle Dylan, et al.
Publicado: (2024)
Mapping the Challenges of HCI: An Application and Evaluation of ChatGPT for Mining Insights at Scale
por: Oppenlaender, Jonas, et al.
Publicado: (2023)
por: Oppenlaender, Jonas, et al.
Publicado: (2023)
Rule Extraction in Machine Learning: Chat Incremental Pattern Constructor
por: Nwokocha, Caleb Princewill
Publicado: (2022)
por: Nwokocha, Caleb Princewill
Publicado: (2022)
Generative AI for Enhancing Active Learning in Education: A Comparative Study of GPT-3.5 and GPT-4 in Crafting Customized Test Questions
por: Rouzegar, Hamdireza, et al.
Publicado: (2024)
por: Rouzegar, Hamdireza, et al.
Publicado: (2024)
GE-Chat: A Graph Enhanced RAG Framework for Evidential Response Generation of LLMs
por: Da, Longchao, et al.
Publicado: (2025)
por: Da, Longchao, et al.
Publicado: (2025)
Reinforcement of Explainability of ChatGPT Prompts by Embedding Breast Cancer Self-Screening Rules into AI Responses
por: Khan, Yousef, et al.
Publicado: (2024)
por: Khan, Yousef, et al.
Publicado: (2024)
GPT-4.1 Sets the Standard in Automated Experiment Design Using Novel Python Libraries
por: Fachada, Nuno, et al.
Publicado: (2025)
por: Fachada, Nuno, et al.
Publicado: (2025)
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments?
por: Ocheja, Patrick, et al.
Publicado: (2024)
por: Ocheja, Patrick, et al.
Publicado: (2024)
PairCFR: Enhancing Model Training on Paired Counterfactually Augmented Data through Contrastive Learning
por: Qiu, Xiaoqi, et al.
Publicado: (2024)
por: Qiu, Xiaoqi, et al.
Publicado: (2024)
Monitoring AI-Modified Content at Scale: A Case Study on the Impact of ChatGPT on AI Conference Peer Reviews
por: Liang, Weixin, et al.
Publicado: (2024)
por: Liang, Weixin, et al.
Publicado: (2024)
SoccerChat: Integrating Multimodal Data for Enhanced Soccer Game Understanding
por: Gautam, Sushant, et al.
Publicado: (2025)
por: Gautam, Sushant, et al.
Publicado: (2025)
Developing Critical Thinking in Second Language Learners: Exploring Generative AI like ChatGPT as a Tool for Argumentative Essay Writing
por: Suh, Simon, et al.
Publicado: (2025)
por: Suh, Simon, et al.
Publicado: (2025)
Visualizing the Evolution of Twitter (X.com) Conversations: A Comprehensive Methodology Applied to AI Training Discussions on ChatGPT
por: Jess, Nicole, et al.
Publicado: (2024)
por: Jess, Nicole, et al.
Publicado: (2024)
The GPT-4o Shock Emotional Attachment to AI Models and Its Impact on Regulatory Acceptance: A Cross-Cultural Analysis of the Immediate Transition from GPT-4o to GPT-5
por: Naito, Hiroki
Publicado: (2025)
por: Naito, Hiroki
Publicado: (2025)
Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
por: Zhang, Li, et al.
Publicado: (2025)
por: Zhang, Li, et al.
Publicado: (2025)
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
por: Zhang, Li, et al.
Publicado: (2026)
por: Zhang, Li, et al.
Publicado: (2026)
IFMTBench: A Comprehensive Benchmark for Multilingual Translation Instruction Following
por: Sun, Mingrui, et al.
Publicado: (2026)
por: Sun, Mingrui, et al.
Publicado: (2026)
Fine-tuning of Large Language Models for Constituency Parsing Using a Sequence to Sequence Approach
por: Delgado, Francisco Jose Cortes, et al.
Publicado: (2025)
por: Delgado, Francisco Jose Cortes, et al.
Publicado: (2025)
Doğal Dil İşlemede Tokenizasyon Standartları ve Ölçümü: Türkçe Üzerinden Büyük Dil Modellerinin Karşılaştırmalı Analizi
por: Bayram, M. Ali, et al.
Publicado: (2025)
por: Bayram, M. Ali, et al.
Publicado: (2025)
Büyük Dil Modelleri için TR-MMLU Benchmarkı: Performans Değerlendirmesi, Zorluklar ve İyileştirme Fırsatları
por: Bayram, M. Ali, et al.
Publicado: (2025)
por: Bayram, M. Ali, et al.
Publicado: (2025)
Targeted Lexical Injection: Unlocking Latent Cross-Lingual Alignment in Lugha-Llama via Early-Layer LoRA Fine-Tuning
por: Ngugi, Stanley
Publicado: (2025)
por: Ngugi, Stanley
Publicado: (2025)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
por: Nieth, Björn, et al.
Publicado: (2026)
por: Nieth, Björn, et al.
Publicado: (2026)
Ejemplares similares
-
Examining Linguistic Shifts in Academic Writing Before and After the Launch of ChatGPT: A Study on Preprint Papers
por: Bao, Tong, et al.
Publicado: (2025) -
Trusting the Search: Unraveling Human Trust in Health Information from Google and ChatGPT
por: Sun, Xin, et al.
Publicado: (2024) -
An Iterative Optimizing Framework for Radiology Report Summarization with ChatGPT
por: Ma, Chong, et al.
Publicado: (2023) -
ChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level Generation
por: Taveekitworachai, Pittawat, et al.
Publicado: (2024) -
Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models
por: Venkatasubramanian, Venkat, et al.
Publicado: (2024)