Benchmarking Llama2, Mistral, Gemma and GPT for Factuality, Toxicity, Bias and Propensity for Hallucinations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nadeau, David, Kroutikov, Mike, McNeil, Karen, Baribeau, Simon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models
von: Safavi-Naini, Seyed Amir Ahmad, et al.
Veröffentlicht: (2024)
von: Safavi-Naini, Seyed Amir Ahmad, et al.
Veröffentlicht: (2024)
Generative AI in Academic Writing: A Comparison of DeepSeek, Qwen, ChatGPT, Gemini, Llama, Mistral, and Gemma
von: Aydin, Omer, et al.
Veröffentlicht: (2025)
von: Aydin, Omer, et al.
Veröffentlicht: (2025)
From Bytes to Borsch: Fine-Tuning Gemma and Mistral for the Ukrainian Language Representation
von: Kiulian, Artur, et al.
Veröffentlicht: (2024)
von: Kiulian, Artur, et al.
Veröffentlicht: (2024)
Visualizing and Benchmarking LLM Factual Hallucination Tendencies via Internal State Analysis and Clustering
von: Mao, Nathan, et al.
Veröffentlicht: (2026)
von: Mao, Nathan, et al.
Veröffentlicht: (2026)
MobiLlama: Towards Accurate and Lightweight Fully Transparent GPT
von: Thawakar, Omkar, et al.
Veröffentlicht: (2024)
von: Thawakar, Omkar, et al.
Veröffentlicht: (2024)
BadLlama: cheaply removing safety fine-tuning from Llama 2-Chat 13B
von: Gade, Pranav, et al.
Veröffentlicht: (2023)
von: Gade, Pranav, et al.
Veröffentlicht: (2023)
On Early Detection of Hallucinations in Factual Question Answering
von: Snyder, Ben, et al.
Veröffentlicht: (2023)
von: Snyder, Ben, et al.
Veröffentlicht: (2023)
Reducing Hallucinations in LLMs via Factuality-Aware Preference Learning
von: Chaduvula, Sindhuja, et al.
Veröffentlicht: (2026)
von: Chaduvula, Sindhuja, et al.
Veröffentlicht: (2026)
Investigating Bias Representations in Llama 2 Chat via Activation Steering
von: Lu, Dawn, et al.
Veröffentlicht: (2024)
von: Lu, Dawn, et al.
Veröffentlicht: (2024)
Investigating Symbolic Triggers of Hallucination in Gemma Models Across HaluEval and TruthfulQA
von: Lamba, Naveen, et al.
Veröffentlicht: (2025)
von: Lamba, Naveen, et al.
Veröffentlicht: (2025)
Mitigating Geospatial Knowledge Hallucination in Large Language Models: Benchmarking and Dynamic Factuality Aligning
von: Wang, Shengyuan, et al.
Veröffentlicht: (2025)
von: Wang, Shengyuan, et al.
Veröffentlicht: (2025)
GKnow: Measuring the Entanglement of Gender Bias and Factual Gender
von: Veloso, Leonor, et al.
Veröffentlicht: (2026)
von: Veloso, Leonor, et al.
Veröffentlicht: (2026)
A Comparative Benchmark of a Moroccan Darija Toxicity Detection Model (Typica.ai) and Major LLM-Based Moderation APIs (OpenAI, Mistral, Anthropic)
von: Assoudi, Hicham
Veröffentlicht: (2025)
von: Assoudi, Hicham
Veröffentlicht: (2025)
PretrainRL: Alleviating Factuality Hallucination of Large Language Models at the Beginning
von: Liu, Langming, et al.
Veröffentlicht: (2026)
von: Liu, Langming, et al.
Veröffentlicht: (2026)
Understanding New-Knowledge-Induced Factual Hallucinations in LLMs: Analysis and Interpretation
von: Dang, Renfei, et al.
Veröffentlicht: (2025)
von: Dang, Renfei, et al.
Veröffentlicht: (2025)
CodeGemma: Open Code Models Based on Gemma
von: CodeGemma Team, et al.
Veröffentlicht: (2024)
von: CodeGemma Team, et al.
Veröffentlicht: (2024)
Mechanistic Understanding and Mitigation of Language Model Non-Factual Hallucinations
von: Yu, Lei, et al.
Veröffentlicht: (2024)
von: Yu, Lei, et al.
Veröffentlicht: (2024)
ShieldGemma: Generative AI Content Moderation Based on Gemma
von: Zeng, Wenjun, et al.
Veröffentlicht: (2024)
von: Zeng, Wenjun, et al.
Veröffentlicht: (2024)
The Dawn After the Dark: An Empirical Study on Factuality Hallucination in Large Language Models
von: Li, Junyi, et al.
Veröffentlicht: (2024)
von: Li, Junyi, et al.
Veröffentlicht: (2024)
Not All That Is Fluent Is Factual: Investigating Hallucinations of Large Language Models in Academic Writing
von: Khan, Humam, et al.
Veröffentlicht: (2026)
von: Khan, Humam, et al.
Veröffentlicht: (2026)
Exploring the Generalizability of Factual Hallucination Mitigation via Enhancing Precise Knowledge Utilization
von: Zhang, Siyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Siyuan, et al.
Veröffentlicht: (2025)
JointCQ: Improving Factual Hallucination Detection with Joint Claim and Query Generation
von: Xu, Fan, et al.
Veröffentlicht: (2025)
von: Xu, Fan, et al.
Veröffentlicht: (2025)
AutoHall: Automated Factuality Hallucination Dataset Generation for Large Language Models
von: Cao, Zouying, et al.
Veröffentlicht: (2023)
von: Cao, Zouying, et al.
Veröffentlicht: (2023)
Predicting Sentence-Level Factuality of News and Bias of Media Outlets
von: Vargas, Francielle, et al.
Veröffentlicht: (2023)
von: Vargas, Francielle, et al.
Veröffentlicht: (2023)
Benchmarking Linguistic Adaptation in Comparable-Sized LLMs: A Study of Llama-3.1-8B, Mistral-7B-v0.1, and Qwen3-8B on Romanized Nepali
von: Rimal, Ananda, et al.
Veröffentlicht: (2026)
von: Rimal, Ananda, et al.
Veröffentlicht: (2026)
FFT: Towards Harmlessness Evaluation and Analysis for LLMs with Factuality, Fairness, Toxicity
von: Cui, Shiyao, et al.
Veröffentlicht: (2023)
von: Cui, Shiyao, et al.
Veröffentlicht: (2023)
Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2
von: Lieberum, Tom, et al.
Veröffentlicht: (2024)
von: Lieberum, Tom, et al.
Veröffentlicht: (2024)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024)
T5Gemma 2: Seeing, Reading, and Understanding Longer
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
When Benchmarks Age: Temporal Misalignment through Large Language Model Factuality Evaluation
von: Jiang, Xunyi, et al.
Veröffentlicht: (2025)
von: Jiang, Xunyi, et al.
Veröffentlicht: (2025)
ChocoLlama: Lessons Learned From Teaching Llamas Dutch
von: Meeus, Matthieu, et al.
Veröffentlicht: (2024)
von: Meeus, Matthieu, et al.
Veröffentlicht: (2024)
TranslateGemma Technical Report
von: Finkelstein, Mara, et al.
Veröffentlicht: (2026)
von: Finkelstein, Mara, et al.
Veröffentlicht: (2026)
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo
von: Wood, Michael C., et al.
Veröffentlicht: (2024)
von: Wood, Michael C., et al.
Veröffentlicht: (2024)
Context-Efficient Retrieval with Factual Decomposition
von: Li, Yanhong, et al.
Veröffentlicht: (2025)
von: Li, Yanhong, et al.
Veröffentlicht: (2025)
A Diagnostic Benchmark for Sweden-Related Factual Knowledge
von: Kunz, Jenny
Veröffentlicht: (2025)
von: Kunz, Jenny
Veröffentlicht: (2025)
DHI: Leveraging Diverse Hallucination Induction for Enhanced Contrastive Factuality Control in Large Language Models
von: Guo, Jiani, et al.
Veröffentlicht: (2026)
von: Guo, Jiani, et al.
Veröffentlicht: (2026)
FIBER: A Multilingual Evaluation Resource for Factual Inference Bias
von: Munis, Evren Ayberk, et al.
Veröffentlicht: (2025)
von: Munis, Evren Ayberk, et al.
Veröffentlicht: (2025)
Enhancing Next-Generation Language Models with Knowledge Graphs: Extending Claude, Mistral IA, and GPT-4 via KG-BERT
von: Chaabene, Nour El Houda Ben, et al.
Veröffentlicht: (2025)
von: Chaabene, Nour El Houda Ben, et al.
Veröffentlicht: (2025)
Linq-Embed-Mistral Technical Report
von: Choi, Chanyeol, et al.
Veröffentlicht: (2024)
von: Choi, Chanyeol, et al.
Veröffentlicht: (2024)
Gemma 3 Technical Report
von: Gemma Team, et al.
Veröffentlicht: (2025)
von: Gemma Team, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models
von: Safavi-Naini, Seyed Amir Ahmad, et al.
Veröffentlicht: (2024) -
Generative AI in Academic Writing: A Comparison of DeepSeek, Qwen, ChatGPT, Gemini, Llama, Mistral, and Gemma
von: Aydin, Omer, et al.
Veröffentlicht: (2025) -
From Bytes to Borsch: Fine-Tuning Gemma and Mistral for the Ukrainian Language Representation
von: Kiulian, Artur, et al.
Veröffentlicht: (2024) -
Visualizing and Benchmarking LLM Factual Hallucination Tendencies via Internal State Analysis and Clustering
von: Mao, Nathan, et al.
Veröffentlicht: (2026) -
MobiLlama: Towards Accurate and Lightweight Fully Transparent GPT
von: Thawakar, Omkar, et al.
Veröffentlicht: (2024)