Comparing GPT-4 and Open-Source Language Models in Misinformation Mitigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vergho, Tyler, Godbout, Jean-Francois, Rabbany, Reihaneh, Pelrine, Kellin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Combining Confidence Elicitation and Sample-based Methods for Uncertainty Quantification in Misinformation Mitigation
von: Rivera, Mauricio, et al.
Veröffentlicht: (2024)
von: Rivera, Mauricio, et al.
Veröffentlicht: (2024)
Uncertainty Resolution in Misinformation Detection
von: Orlovskiy, Yury, et al.
Veröffentlicht: (2024)
von: Orlovskiy, Yury, et al.
Veröffentlicht: (2024)
Web Retrieval Agents for Evidence-Based Misinformation Detection
von: Tian, Jacob-Junqi, et al.
Veröffentlicht: (2024)
von: Tian, Jacob-Junqi, et al.
Veröffentlicht: (2024)
A Guide to Misinformation Detection Data and Evaluation
von: Thibault, Camille, et al.
Veröffentlicht: (2024)
von: Thibault, Camille, et al.
Veröffentlicht: (2024)
Veracity: An Open-Source AI Fact-Checking System
von: Curtis, Taylor Lynn, et al.
Veröffentlicht: (2025)
von: Curtis, Taylor Lynn, et al.
Veröffentlicht: (2025)
Epistemic Integrity in Large Language Models
von: Ghafouri, Bijean, et al.
Veröffentlicht: (2024)
von: Ghafouri, Bijean, et al.
Veröffentlicht: (2024)
$\texttt{BluePrint}$: A Social Media User Dataset for LLM Persona Evaluation and Training
von: Bück-Kaeffer, Aurélien, et al.
Veröffentlicht: (2025)
von: Bück-Kaeffer, Aurélien, et al.
Veröffentlicht: (2025)
Towards Detecting Contextual Real-Time Toxicity for In-Game Chat
von: Yang, Zachary, et al.
Veröffentlicht: (2023)
von: Yang, Zachary, et al.
Veröffentlicht: (2023)
Emerging Vulnerabilities in Frontier Models: Multi-Turn Jailbreak Attacks
von: Gibbs, Tom, et al.
Veröffentlicht: (2024)
von: Gibbs, Tom, et al.
Veröffentlicht: (2024)
Online Influence Campaigns: Strategies and Vulnerabilities
von: Musulan, Andreea, et al.
Veröffentlicht: (2024)
von: Musulan, Andreea, et al.
Veröffentlicht: (2024)
Exposing the Systematic Vulnerability of Open-Weight Models to Prefill Attacks
von: Struppek, Lukas, et al.
Veröffentlicht: (2026)
von: Struppek, Lukas, et al.
Veröffentlicht: (2026)
Exploiting Novel GPT-4 APIs
von: Pelrine, Kellin, et al.
Veröffentlicht: (2023)
von: Pelrine, Kellin, et al.
Veröffentlicht: (2023)
CrediBench: Building Web-Scale Network Datasets for Information Integrity
von: Kondrup, Emma, et al.
Veröffentlicht: (2025)
von: Kondrup, Emma, et al.
Veröffentlicht: (2025)
Accidental Vulnerability: Factors in Fine-Tuning that Shift Model Safeguards
von: Pandey, Punya Syon, et al.
Veröffentlicht: (2025)
von: Pandey, Punya Syon, et al.
Veröffentlicht: (2025)
Regional and Temporal Patterns of Partisan Polarization during the COVID-19 Pandemic in the United States and Canada
von: Yang, Zachary, et al.
Veröffentlicht: (2024)
von: Yang, Zachary, et al.
Veröffentlicht: (2024)
From Intuition to Understanding: Using AI Peers to Overcome Physics Misconceptions
von: Weijers, Ruben, et al.
Veröffentlicht: (2025)
von: Weijers, Ruben, et al.
Veröffentlicht: (2025)
PairBench: Are Vision-Language Models Reliable at Comparing What They See?
von: Feizi, Aarash, et al.
Veröffentlicht: (2025)
von: Feizi, Aarash, et al.
Veröffentlicht: (2025)
Unified Game Moderation: Soft-Prompting and LLM-Assisted Label Transfer for Resource-Efficient Toxicity Detection
von: Yang, Zachary, et al.
Veröffentlicht: (2025)
von: Yang, Zachary, et al.
Veröffentlicht: (2025)
Hallucination Detox: Sensitivity Dropout (SenD) for Large Language Model Training
von: Mohammadzadeh, Shahrad, et al.
Veröffentlicht: (2024)
von: Mohammadzadeh, Shahrad, et al.
Veröffentlicht: (2024)
The Structural Safety Generalization Problem
von: Broomfield, Julius, et al.
Veröffentlicht: (2025)
von: Broomfield, Julius, et al.
Veröffentlicht: (2025)
OpenFake: An Open Dataset and Platform Toward Real-World Deepfake Detection
von: Livernoche, Victor, et al.
Veröffentlicht: (2025)
von: Livernoche, Victor, et al.
Veröffentlicht: (2025)
Battling Misinformation: An Empirical Study on Adversarial Factuality in Open-Source Large Language Models
von: Sakib, Shahnewaz Karim, et al.
Veröffentlicht: (2025)
von: Sakib, Shahnewaz Karim, et al.
Veröffentlicht: (2025)
Jailbreak-Tuning: Models Efficiently Learn Jailbreak Susceptibility
von: Murphy, Brendan, et al.
Veröffentlicht: (2025)
von: Murphy, Brendan, et al.
Veröffentlicht: (2025)
Are Large Language Models Good Temporal Graph Learners?
von: Huang, Shenyang, et al.
Veröffentlicht: (2025)
von: Huang, Shenyang, et al.
Veröffentlicht: (2025)
Deepfakes in the 2025 Canadian Election: Prevalence, Partisanship, and Platform Dynamics
von: Livernoche, Victor, et al.
Veröffentlicht: (2025)
von: Livernoche, Victor, et al.
Veröffentlicht: (2025)
OpenThaiGPT 1.5: A Thai-Centric Open Source Large Language Model
von: Yuenyong, Sumeth, et al.
Veröffentlicht: (2024)
von: Yuenyong, Sumeth, et al.
Veröffentlicht: (2024)
What do people want to fact-check?
von: Ghafouri, Bijean, et al.
Veröffentlicht: (2026)
von: Ghafouri, Bijean, et al.
Veröffentlicht: (2026)
Enhancing Privacy in the Early Detection of Sexual Predators Through Federated Learning and Differential Privacy
von: Chehbouni, Khaoula, et al.
Veröffentlicht: (2025)
von: Chehbouni, Khaoula, et al.
Veröffentlicht: (2025)
ChatGPT's One-year Anniversary: Are Open-Source Large Language Models Catching up?
von: Chen, Hailin, et al.
Veröffentlicht: (2023)
von: Chen, Hailin, et al.
Veröffentlicht: (2023)
OpenThaiGPT 1.6 and R1: Thai-Centric Open Source and Reasoning Large Language Models
von: Yuenyong, Sumeth, et al.
Veröffentlicht: (2025)
von: Yuenyong, Sumeth, et al.
Veröffentlicht: (2025)
Can OpenSource beat ChatGPT? -- A Comparative Study of Large Language Models for Text-to-Code Generation
von: Mayer, Luis, et al.
Veröffentlicht: (2024)
von: Mayer, Luis, et al.
Veröffentlicht: (2024)
FinGPT: Open-Source Financial Large Language Models
von: Yang, Hongyang, et al.
Veröffentlicht: (2023)
von: Yang, Hongyang, et al.
Veröffentlicht: (2023)
Is GPT-OSS Good? A Comprehensive Evaluation of OpenAI's Latest Open Source Models
von: Bi, Ziqian, et al.
Veröffentlicht: (2025)
von: Bi, Ziqian, et al.
Veröffentlicht: (2025)
A Simulation System Towards Solving Societal-Scale Manipulation
von: Touzel, Maximilian Puelma, et al.
Veröffentlicht: (2024)
von: Touzel, Maximilian Puelma, et al.
Veröffentlicht: (2024)
Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding
von: Jia, Hong, et al.
Veröffentlicht: (2025)
von: Jia, Hong, et al.
Veröffentlicht: (2025)
RLAIF-V: Open-Source AI Feedback Leads to Super GPT-4V Trustworthiness
von: Yu, Tianyu, et al.
Veröffentlicht: (2024)
von: Yu, Tianyu, et al.
Veröffentlicht: (2024)
Comparative Analysis of Open-Source Language Models in Summarizing Medical Text Data
von: Chen, Yuhao, et al.
Veröffentlicht: (2024)
von: Chen, Yuhao, et al.
Veröffentlicht: (2024)
Open Source Language Models Can Provide Feedback: Evaluating LLMs' Ability to Help Students Using GPT-4-As-A-Judge
von: Koutcheme, Charles, et al.
Veröffentlicht: (2024)
von: Koutcheme, Charles, et al.
Veröffentlicht: (2024)
Unlearning Climate Misinformation in Large Language Models
von: Fore, Michael, et al.
Veröffentlicht: (2024)
von: Fore, Michael, et al.
Veröffentlicht: (2024)
Images Amplify Misinformation Sharing in Vision-Language Models
von: Plebe, Alice, et al.
Veröffentlicht: (2025)
von: Plebe, Alice, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Combining Confidence Elicitation and Sample-based Methods for Uncertainty Quantification in Misinformation Mitigation
von: Rivera, Mauricio, et al.
Veröffentlicht: (2024) -
Uncertainty Resolution in Misinformation Detection
von: Orlovskiy, Yury, et al.
Veröffentlicht: (2024) -
Web Retrieval Agents for Evidence-Based Misinformation Detection
von: Tian, Jacob-Junqi, et al.
Veröffentlicht: (2024) -
A Guide to Misinformation Detection Data and Evaluation
von: Thibault, Camille, et al.
Veröffentlicht: (2024) -
Veracity: An Open-Source AI Fact-Checking System
von: Curtis, Taylor Lynn, et al.
Veröffentlicht: (2025)