Can LLM-Generated Misinformation Be Detected?
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Canyu, Shu, Kai |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Security Risk Taxonomy for Prompt-Based Interaction With Large Language Models
by: Derner, Erik, et al.
Published: (2023)
by: Derner, Erik, et al.
Published: (2023)
LLM Whisperer: An Inconspicuous Attack to Bias LLM Responses
by: Lin, Weiran, et al.
Published: (2024)
by: Lin, Weiran, et al.
Published: (2024)
Known By Their Actions: Fingerprinting LLM Browser Agents via UI Traces
by: Lugoloobi, William, et al.
Published: (2026)
by: Lugoloobi, William, et al.
Published: (2026)
Learned, Lagged, LLM-splained: LLM Responses to End User Security Questions
by: Prakash, Vijay, et al.
Published: (2024)
by: Prakash, Vijay, et al.
Published: (2024)
Generative AI-based closed-loop fMRI system
by: Kasahara, Mikihiro, et al.
Published: (2024)
by: Kasahara, Mikihiro, et al.
Published: (2024)
Identify As A Human Does: A Pathfinder of Next-Generation Anti-Cheat Framework for First-Person Shooter Games
by: Zhang, Jiayi, et al.
Published: (2024)
by: Zhang, Jiayi, et al.
Published: (2024)
On the Suitability of LLM-Driven Agents for Dark Pattern Audits
by: Sun, Chen, et al.
Published: (2026)
by: Sun, Chen, et al.
Published: (2026)
IFTT-PIN: A Self-Calibrating PIN-Entry Method
by: McConkey, Kathryn, et al.
Published: (2024)
by: McConkey, Kathryn, et al.
Published: (2024)
Avoiding Leakage Poisoning: Concept Interventions Under Distribution Shifts
by: Zarlenga, Mateo Espinosa, et al.
Published: (2025)
by: Zarlenga, Mateo Espinosa, et al.
Published: (2025)
Distinguishing Scams and Fraud with Ensemble Learning
by: Chadalavada, Isha, et al.
Published: (2024)
by: Chadalavada, Isha, et al.
Published: (2024)
A Human-Centered Privacy Approach (HCP) to AI
by: Sun, Luyi, et al.
Published: (2026)
by: Sun, Luyi, et al.
Published: (2026)
Towards Automating Data Access Permissions in AI Agents
by: Wu, Yuhao, et al.
Published: (2025)
by: Wu, Yuhao, et al.
Published: (2025)
Adversarial Attacks on Machine Learning-Aided Visualizations
by: Fujiwara, Takanori, et al.
Published: (2024)
by: Fujiwara, Takanori, et al.
Published: (2024)
RLHFPoison: Reward Poisoning Attack for Reinforcement Learning with Human Feedback in Large Language Models
by: Wang, Jiongxiao, et al.
Published: (2023)
by: Wang, Jiongxiao, et al.
Published: (2023)
Adversarial Artifact Detection in EEG-Based Brain-Computer Interfaces
by: Chen, Xiaoqing, et al.
Published: (2022)
by: Chen, Xiaoqing, et al.
Published: (2022)
LLM Novice Uplift on Dual-Use, In Silico Biology Tasks
by: Zhang, Chen Bo Calvin, et al.
Published: (2026)
by: Zhang, Chen Bo Calvin, et al.
Published: (2026)
Assessing LLM Response Quality in the Context of Technology-Facilitated Abuse
by: Prakash, Vijay, et al.
Published: (2026)
by: Prakash, Vijay, et al.
Published: (2026)
Hacc-Man: An Arcade Game for Jailbreaking LLMs
by: Valentim, Matheus, et al.
Published: (2024)
by: Valentim, Matheus, et al.
Published: (2024)
Beyond Words: On Large Language Models Actionability in Mission-Critical Risk Analysis
by: Esposito, Matteo, et al.
Published: (2024)
by: Esposito, Matteo, et al.
Published: (2024)
Emergent misalignment as prompt sensitivity: A research note
by: Wyse, Tim, et al.
Published: (2025)
by: Wyse, Tim, et al.
Published: (2025)
Calpric: Inclusive and Fine-grain Labeling of Privacy Policies with Crowdsourcing and Active Learning
by: Qiu, Wenjun, et al.
Published: (2024)
by: Qiu, Wenjun, et al.
Published: (2024)
Evaluating the Usability of LLMs in Threat Intelligence Enrichment
by: Srikanth, Sanchana, et al.
Published: (2024)
by: Srikanth, Sanchana, et al.
Published: (2024)
Intolerable Risk Threshold Recommendations for Artificial Intelligence
by: Raman, Deepika, et al.
Published: (2025)
by: Raman, Deepika, et al.
Published: (2025)
Eye-tracked Virtual Reality: A Comprehensive Survey on Methods and Privacy Challenges
by: Bozkir, Efe, et al.
Published: (2023)
by: Bozkir, Efe, et al.
Published: (2023)
PhishLang: A Real-Time, Fully Client-Side Phishing Detection Framework Using MobileBERT
by: Roy, Sayak Saha, et al.
Published: (2024)
by: Roy, Sayak Saha, et al.
Published: (2024)
Rescriber: Smaller-LLM-Powered User-Led Data Minimization for LLM-Based Chatbots
by: Zhou, Jijie, et al.
Published: (2024)
by: Zhou, Jijie, et al.
Published: (2024)
Current state of LLM Risks and AI Guardrails
by: Ayyamperumal, Suriya Ganesh, et al.
Published: (2024)
by: Ayyamperumal, Suriya Ganesh, et al.
Published: (2024)
Play Guessing Game with LLM: Indirect Jailbreak Attack with Implicit Clues
by: Chang, Zhiyuan, et al.
Published: (2024)
by: Chang, Zhiyuan, et al.
Published: (2024)
From Assistants to Adversaries: Exploring the Security Risks of Mobile LLM Agents
by: Wu, Liangxuan, et al.
Published: (2025)
by: Wu, Liangxuan, et al.
Published: (2025)
Empowering Users in Digital Privacy Management through Interactive LLM-Based Agents
by: Sun, Bolun, et al.
Published: (2024)
by: Sun, Bolun, et al.
Published: (2024)
Autonomy Reshapes How Personalization Affects Privacy Concerns and Trust in LLM Agents
by: Zhang, Zhiping, et al.
Published: (2025)
by: Zhang, Zhiping, et al.
Published: (2025)
"It's a Fair Game", or Is It? Examining How Users Navigate Disclosure Risks and Benefits When Using LLM-Based Conversational Agents
by: Zhang, Zhiping, et al.
Published: (2023)
by: Zhang, Zhiping, et al.
Published: (2023)
MeAJOR Corpus: A Multi-Source Dataset for Phishing Email Detection
by: Mendes, Paulo, et al.
Published: (2025)
by: Mendes, Paulo, et al.
Published: (2025)
Adversarial VR: An Open-Source Testbed for Evaluating Adversarial Robustness of VR Cybersickness Detection and Mitigation
by: Ahmed, Istiak, et al.
Published: (2025)
by: Ahmed, Istiak, et al.
Published: (2025)
Federated Learning in Offline and Online EMG Decoding: A Privacy and Performance Perspective
by: Malcolm, Kai, et al.
Published: (2025)
by: Malcolm, Kai, et al.
Published: (2025)
Cyri: A Conversational AI-based Assistant for Supporting the Human User in Detecting and Responding to Phishing Attacks
by: La Torre, Antonio, et al.
Published: (2025)
by: La Torre, Antonio, et al.
Published: (2025)
NLP Privacy Risk Identification in Social Media (NLP-PRISM): A Survey
by: Goswami, Dhiman, et al.
Published: (2026)
by: Goswami, Dhiman, et al.
Published: (2026)
Personalised Feedback Framework for Online Education Programmes Using Generative AI
by: Kuzminykh, Ievgeniia, et al.
Published: (2024)
by: Kuzminykh, Ievgeniia, et al.
Published: (2024)
Device-Native Autonomous Agents for Privacy-Preserving Negotiations
by: Roy, Joyjit, et al.
Published: (2026)
by: Roy, Joyjit, et al.
Published: (2026)
Risk Averse Alert Prioritization for IDS Using Subnormal Gaussian Fuzzy Models
by: Moran, Murat
Published: (2026)
by: Moran, Murat
Published: (2026)
Similar Items
-
A Security Risk Taxonomy for Prompt-Based Interaction With Large Language Models
by: Derner, Erik, et al.
Published: (2023) -
LLM Whisperer: An Inconspicuous Attack to Bias LLM Responses
by: Lin, Weiran, et al.
Published: (2024) -
Known By Their Actions: Fingerprinting LLM Browser Agents via UI Traces
by: Lugoloobi, William, et al.
Published: (2026) -
Learned, Lagged, LLM-splained: LLM Responses to End User Security Questions
by: Prakash, Vijay, et al.
Published: (2024) -
Generative AI-based closed-loop fMRI system
by: Kasahara, Mikihiro, et al.
Published: (2024)