Reasoning and the Trusting Behavior of DeepSeek and GPT: An Experiment Revealing Hidden Fault Lines in Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Rubing, Sedoc, João, Sundararajan, Arun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An evaluation of DeepSeek Models in Biomedical Natural Language Processing
di: Zhan, Zaifu, et al.
Pubblicazione: (2025)
di: Zhan, Zaifu, et al.
Pubblicazione: (2025)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025)
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025)
DeepSeek performs better than other Large Language Models in Dental Cases
di: Zhang, Hexian, et al.
Pubblicazione: (2025)
di: Zhang, Hexian, et al.
Pubblicazione: (2025)
Information Suppression in Large Language Models: Auditing, Quantifying, and Characterizing Censorship in DeepSeek
di: Qiu, Peiran, et al.
Pubblicazione: (2025)
di: Qiu, Peiran, et al.
Pubblicazione: (2025)
Safety Evaluation of DeepSeek Models in Chinese Contexts
di: Zhang, Wenjing, et al.
Pubblicazione: (2025)
di: Zhang, Wenjing, et al.
Pubblicazione: (2025)
DeepSeek-V3 Technical Report
di: DeepSeek-AI, et al.
Pubblicazione: (2024)
di: DeepSeek-AI, et al.
Pubblicazione: (2024)
A Comparison of DeepSeek and Other LLMs
di: Gao, Tianchen, et al.
Pubblicazione: (2025)
di: Gao, Tianchen, et al.
Pubblicazione: (2025)
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
di: DeepSeek-AI, et al.
Pubblicazione: (2024)
di: DeepSeek-AI, et al.
Pubblicazione: (2024)
Mixture of Tunable Experts -- Behavior Modification of DeepSeek-R1 at Inference Time
di: Dahlke, Robert, et al.
Pubblicazione: (2025)
di: Dahlke, Robert, et al.
Pubblicazione: (2025)
DeepSeek in Healthcare: A Survey of Capabilities, Risks, and Clinical Applications of Open-Source Large Language Models
di: Ye, Jiancheng, et al.
Pubblicazione: (2025)
di: Ye, Jiancheng, et al.
Pubblicazione: (2025)
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
di: DeepSeek-AI, et al.
Pubblicazione: (2024)
di: DeepSeek-AI, et al.
Pubblicazione: (2024)
Safety Evaluation and Enhancement of DeepSeek Models in Chinese Contexts
di: Zhang, Wenjing, et al.
Pubblicazione: (2025)
di: Zhang, Wenjing, et al.
Pubblicazione: (2025)
From Reasoning to Answer: Empirical, Attention-Based and Mechanistic Insights into Distilled DeepSeek R1 Models
di: Zhang, Jue, et al.
Pubblicazione: (2025)
di: Zhang, Jue, et al.
Pubblicazione: (2025)
Output Length Effect on DeepSeek-R1's Safety in Forced Thinking
di: Li, Xuying, et al.
Pubblicazione: (2025)
di: Li, Xuying, et al.
Pubblicazione: (2025)
An evaluation of LLMs for generating movie reviews: GPT-4o, Gemini-2.0 and DeepSeek-V3
di: Sands, Brendan, et al.
Pubblicazione: (2025)
di: Sands, Brendan, et al.
Pubblicazione: (2025)
Comparative Evaluation of ChatGPT and DeepSeek Across Key NLP Tasks: Strengths, Weaknesses, and Domain-Specific Performance
di: Etaiwi, Wael, et al.
Pubblicazione: (2025)
di: Etaiwi, Wael, et al.
Pubblicazione: (2025)
A comprehensive study of LLM-based argument classification: from Llama through DeepSeek to GPT-5.2
di: Pietroń, Marcin, et al.
Pubblicazione: (2026)
di: Pietroń, Marcin, et al.
Pubblicazione: (2026)
RealSafe-R1: Safety-Aligned DeepSeek-R1 without Compromising Reasoning Capability
di: Zhang, Yichi, et al.
Pubblicazione: (2025)
di: Zhang, Yichi, et al.
Pubblicazione: (2025)
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
di: DeepSeek-AI, et al.
Pubblicazione: (2025)
di: DeepSeek-AI, et al.
Pubblicazione: (2025)
Large Human Language Models: A Need and the Challenges
di: Soni, Nikita, et al.
Pubblicazione: (2023)
di: Soni, Nikita, et al.
Pubblicazione: (2023)
DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
di: Ren, Z. Z., et al.
Pubblicazione: (2025)
di: Ren, Z. Z., et al.
Pubblicazione: (2025)
Towards Understanding the Safety Boundaries of DeepSeek Models: Evaluation and Findings
di: Ying, Zonghao, et al.
Pubblicazione: (2025)
di: Ying, Zonghao, et al.
Pubblicazione: (2025)
Prompt-Counterfactual Explanations for Generative AI System Behavior
di: Goethals, Sofie, et al.
Pubblicazione: (2026)
di: Goethals, Sofie, et al.
Pubblicazione: (2026)
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
di: Wu, Zhiyu, et al.
Pubblicazione: (2024)
di: Wu, Zhiyu, et al.
Pubblicazione: (2024)
Challenges and Applications of Large Language Models: A Comparison of GPT and DeepSeek family of models
di: Sharma, Shubham, et al.
Pubblicazione: (2025)
di: Sharma, Shubham, et al.
Pubblicazione: (2025)
DBOT: Artificial Intelligence for Systematic Long-Term Investing
di: Dhar, Vasant, et al.
Pubblicazione: (2025)
di: Dhar, Vasant, et al.
Pubblicazione: (2025)
Explainable Sentiment Analysis with DeepSeek-R1: Performance, Efficiency, and Few-Shot Learning
di: Huang, Donghao, et al.
Pubblicazione: (2025)
di: Huang, Donghao, et al.
Pubblicazione: (2025)
Emotion-Aware Embedding Fusion in LLMs (Flan-T5, LLAMA 2, DeepSeek-R1, and ChatGPT 4) for Intelligent Response Generation
di: Rasool, Abdur, et al.
Pubblicazione: (2024)
di: Rasool, Abdur, et al.
Pubblicazione: (2024)
User Intent to Use DeepSeek for Healthcare Purposes and their Trust in the Large Language Model: Multinational Survey Study
di: Choudhury, Avishek, et al.
Pubblicazione: (2025)
di: Choudhury, Avishek, et al.
Pubblicazione: (2025)
Comparative Analysis of OpenAI GPT-4o and DeepSeek R1 for Scientific Text Categorization Using Prompt Engineering
di: Maiti, Aniruddha, et al.
Pubblicazione: (2025)
di: Maiti, Aniruddha, et al.
Pubblicazione: (2025)
Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-based LLMs
di: Ji, Tao, et al.
Pubblicazione: (2025)
di: Ji, Tao, et al.
Pubblicazione: (2025)
Evaluating the Performance of AI Text Detectors, Few-Shot and Chain-of-Thought Prompting Using DeepSeek Generated Text
di: Alshammari, Hulayyil, et al.
Pubblicazione: (2025)
di: Alshammari, Hulayyil, et al.
Pubblicazione: (2025)
DeepSeek-R1 Outperforms Gemini 2.0 Pro, OpenAI o1, and o3-mini in Bilingual Complex Ophthalmology Reasoning
di: Xu, Pusheng, et al.
Pubblicazione: (2025)
di: Xu, Pusheng, et al.
Pubblicazione: (2025)
LLMs in Disease Diagnosis: A Comparative Study of DeepSeek-R1 and O3 Mini Across Chronic Health Conditions
di: Gupta, Gaurav Kumar, et al.
Pubblicazione: (2025)
di: Gupta, Gaurav Kumar, et al.
Pubblicazione: (2025)
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
di: Shao, Zhihong, et al.
Pubblicazione: (2024)
di: Shao, Zhihong, et al.
Pubblicazione: (2024)
Challenges in Ensuring AI Safety in DeepSeek-R1 Models: The Shortcomings of Reinforcement Learning Strategies
di: Parmar, Manojkumar, et al.
Pubblicazione: (2025)
di: Parmar, Manojkumar, et al.
Pubblicazione: (2025)
MHA2MLA-VLM: Enabling DeepSeek's Economical Multi-Head Latent Attention across Vision-Language Models
di: Fan, Xiaoran, et al.
Pubblicazione: (2026)
di: Fan, Xiaoran, et al.
Pubblicazione: (2026)
Can Large Language Model Agents Simulate Human Trust Behavior?
di: Xie, Chengxing, et al.
Pubblicazione: (2024)
di: Xie, Chengxing, et al.
Pubblicazione: (2024)
Large Language Models Show Human-like Social Desirability Biases in Survey Responses
di: Salecha, Aadesh, et al.
Pubblicazione: (2024)
di: Salecha, Aadesh, et al.
Pubblicazione: (2024)
Expediting and Elevating Large Language Model Reasoning via Hidden Chain-of-Thought Decoding
di: Liu, Tianqiao, et al.
Pubblicazione: (2024)
di: Liu, Tianqiao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
An evaluation of DeepSeek Models in Biomedical Natural Language Processing
di: Zhan, Zaifu, et al.
Pubblicazione: (2025) -
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025) -
DeepSeek performs better than other Large Language Models in Dental Cases
di: Zhang, Hexian, et al.
Pubblicazione: (2025) -
Information Suppression in Large Language Models: Auditing, Quantifying, and Characterizing Censorship in DeepSeek
di: Qiu, Peiran, et al.
Pubblicazione: (2025) -
Safety Evaluation of DeepSeek Models in Chinese Contexts
di: Zhang, Wenjing, et al.
Pubblicazione: (2025)