Saved in:
| Main Author: | Assoudi, Hicham |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.04640 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DarijaBanking: A New Resource for Overcoming Language Barriers in Banking Intent Detection for Moroccan Arabic Speakers
by: Skiredj, Abderrahman, et al.
Published: (2024)
by: Skiredj, Abderrahman, et al.
Published: (2024)
On Sarcasm Detection with OpenAI GPT-based Models
by: Gole, Montgomery, et al.
Published: (2023)
by: Gole, Montgomery, et al.
Published: (2023)
Implementing Systemic Thinking for Automatic Schema Matching: An Agent-Based Modeling Approach
by: Assoudi, Hicham, et al.
Published: (2025)
by: Assoudi, Hicham, et al.
Published: (2025)
A Comparative Study on Reasoning Patterns of OpenAI's o1 Model
by: Wu, Siwei, et al.
Published: (2024)
by: Wu, Siwei, et al.
Published: (2024)
The Evolution of Darija Open Dataset: Introducing Version 2
by: Outchakoucht, Aissam, et al.
Published: (2024)
by: Outchakoucht, Aissam, et al.
Published: (2024)
Dialect2SQL: A Novel Text-to-SQL Dataset for Arabic Dialects with a Focus on Moroccan Darija
by: Chafik, Salmane, et al.
Published: (2025)
by: Chafik, Salmane, et al.
Published: (2025)
Benchmarking Llama2, Mistral, Gemma and GPT for Factuality, Toxicity, Bias and Propensity for Hallucinations
by: Nadeau, David, et al.
Published: (2024)
by: Nadeau, David, et al.
Published: (2024)
OpenAI GPT-5 System Card
by: Singh, Aaditya, et al.
Published: (2025)
by: Singh, Aaditya, et al.
Published: (2025)
A Benchmark for End-to-End Zero-Shot Biomedical Relation Extraction with LLMs: Experiments with OpenAI Models
by: Brokman, Aviv, et al.
Published: (2025)
by: Brokman, Aviv, et al.
Published: (2025)
Quantization for OpenAI's Whisper Models: A Comparative Analysis
by: Andreyev, Allison
Published: (2025)
by: Andreyev, Allison
Published: (2025)
Evaluation of OpenAI o1: Opportunities and Challenges of AGI
by: Zhong, Tianyang, et al.
Published: (2024)
by: Zhong, Tianyang, et al.
Published: (2024)
Is GPT-OSS Good? A Comprehensive Evaluation of OpenAI's Latest Open Source Models
by: Bi, Ziqian, et al.
Published: (2025)
by: Bi, Ziqian, et al.
Published: (2025)
Benchmarking Floworks against OpenAI & Anthropic: A Novel Framework for Enhanced LLM Function Calling
by: Bhan, Nirav, et al.
Published: (2024)
by: Bhan, Nirav, et al.
Published: (2024)
AI Governance and Accountability: An Analysis of Anthropic's Claude
by: Priyanshu, Aman, et al.
Published: (2024)
by: Priyanshu, Aman, et al.
Published: (2024)
LLM Platform Security: Applying a Systematic Evaluation Framework to OpenAI's ChatGPT Plugins
by: Iqbal, Umar, et al.
Published: (2023)
by: Iqbal, Umar, et al.
Published: (2023)
A Case Study of Web App Coding with OpenAI Reasoning Models
by: Cui, Yi
Published: (2024)
by: Cui, Yi
Published: (2024)
GemMaroc: Unlocking Darija Proficiency in LLMs with Minimal Data
by: Skiredj, Abderrahman, et al.
Published: (2025)
by: Skiredj, Abderrahman, et al.
Published: (2025)
Evaluating Text Summaries Generated by Large Language Models Using OpenAI's GPT
by: Shakil, Hassan, et al.
Published: (2024)
by: Shakil, Hassan, et al.
Published: (2024)
Lost in Moderation: How Commercial Content Moderation APIs Over- and Under-Moderate Group-Targeted Hate Speech and Linguistic Variations
by: Hartmann, David, et al.
Published: (2025)
by: Hartmann, David, et al.
Published: (2025)
OpenAI's GPT-OSS-20B Model and Safety Alignment Issues in a Low-Resource Language
by: Inuwa-Dutse, Isa
Published: (2025)
by: Inuwa-Dutse, Isa
Published: (2025)
Evading Toxicity Detection with ASCII-art: A Benchmark of Spatial Attacks on Moderation Systems
by: Berezin, Sergey, et al.
Published: (2024)
by: Berezin, Sergey, et al.
Published: (2024)
Evaluating OpenAI GPT Models for Translation of Endangered Uralic Languages: A Comparison of Reasoning and Non-Reasoning Architectures
by: Tereshchenko, Yehor, et al.
Published: (2025)
by: Tereshchenko, Yehor, et al.
Published: (2025)
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3
by: Sadik, Ahmed R., et al.
Published: (2025)
by: Sadik, Ahmed R., et al.
Published: (2025)
The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies
by: Wang, Chenxu, et al.
Published: (2026)
by: Wang, Chenxu, et al.
Published: (2026)
Comparative Analysis of OpenAI GPT-4o and DeepSeek R1 for Scientific Text Categorization Using Prompt Engineering
by: Maiti, Aniruddha, et al.
Published: (2025)
by: Maiti, Aniruddha, et al.
Published: (2025)
Does fine-tuning GPT-3 with the OpenAI API leak personally-identifiable information?
by: Sun, Albert Yu, et al.
Published: (2023)
by: Sun, Albert Yu, et al.
Published: (2023)
Analysing the Public Discourse around OpenAI's Text-To-Video Model 'Sora' using Topic Modeling
by: Parikh, Vatsal Vinay
Published: (2024)
by: Parikh, Vatsal Vinay
Published: (2024)
How Utilitarian Are OpenAI's Models Really? Replicating and Reinterpreting Pfeffer, Krügel, and Uhl (2025)
by: Himmelreich, Johannes
Published: (2026)
by: Himmelreich, Johannes
Published: (2026)
Benchmarking LLM Guardrails in Handling Multilingual Toxicity
by: Yang, Yahan, et al.
Published: (2024)
by: Yang, Yahan, et al.
Published: (2024)
Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content
by: Furniturewala, Shaz, et al.
Published: (2025)
by: Furniturewala, Shaz, et al.
Published: (2025)
Speech Emotion Recognition Leveraging OpenAI's Whisper Representations and Attentive Pooling Methods
by: Shendabadi, Ali, et al.
Published: (2026)
by: Shendabadi, Ali, et al.
Published: (2026)
MTUncertainty: Assessing the Need for Post-editing of Machine Translation Outputs by Fine-tuning OpenAI LLMs
by: Gladkoff, Serge, et al.
Published: (2023)
by: Gladkoff, Serge, et al.
Published: (2023)
System 2 thinking in OpenAI's o1-preview model: Near-perfect performance on a mathematics exam
by: de Winter, Joost, et al.
Published: (2024)
by: de Winter, Joost, et al.
Published: (2024)
Large Malaysian Language Model Based on Mistral for Enhanced Local Language Understanding
by: Zolkepli, Husein, et al.
Published: (2024)
by: Zolkepli, Husein, et al.
Published: (2024)
A Large-Scale Empirical Analysis of Custom GPTs' Vulnerabilities in the OpenAI Ecosystem
by: Ogundoyin, Sunday Oyinlola, et al.
Published: (2025)
by: Ogundoyin, Sunday Oyinlola, et al.
Published: (2025)
In AI Sweet Harmony: Sociopragmatic Guardrail Bypasses and Evaluation-Awareness in OpenAI gpt-oss-20b
by: Durner, Nils
Published: (2025)
by: Durner, Nils
Published: (2025)
Auditing Black-Box LLM APIs with a Rank-Based Uniformity Test
by: Zhu, Xiaoyuan, et al.
Published: (2025)
by: Zhu, Xiaoyuan, et al.
Published: (2025)
Evaluating Test-Time Scaling LLMs for Legal Reasoning: OpenAI o1, DeepSeek-R1, and Beyond
by: Hu, Yinghao, et al.
Published: (2025)
by: Hu, Yinghao, et al.
Published: (2025)
aiXiv: A Next-Generation Open Access Ecosystem for Scientific Discovery Generated by AI Scientists
by: Zhang, Pengsong, et al.
Published: (2025)
by: Zhang, Pengsong, et al.
Published: (2025)
WorldAPIs: The World Is Worth How Many APIs? A Thought Experiment
by: Ou, Jiefu, et al.
Published: (2024)
by: Ou, Jiefu, et al.
Published: (2024)
Similar Items
-
DarijaBanking: A New Resource for Overcoming Language Barriers in Banking Intent Detection for Moroccan Arabic Speakers
by: Skiredj, Abderrahman, et al.
Published: (2024) -
On Sarcasm Detection with OpenAI GPT-based Models
by: Gole, Montgomery, et al.
Published: (2023) -
Implementing Systemic Thinking for Automatic Schema Matching: An Agent-Based Modeling Approach
by: Assoudi, Hicham, et al.
Published: (2025) -
A Comparative Study on Reasoning Patterns of OpenAI's o1 Model
by: Wu, Siwei, et al.
Published: (2024) -
The Evolution of Darija Open Dataset: Introducing Version 2
by: Outchakoucht, Aissam, et al.
Published: (2024)