A Survey of AI-generated Text Forensic Systems: Detection, Attribution, and Characterization
Fuente:
arXiv
Salvato in:
| Autori principali: | Kumarage, Tharindu, Agrawal, Garima, Sheth, Paras, Moraffah, Raha, Chadha, Aman, Garland, Joshua, Liu, Huan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Cross-Platform Hate Speech Detection with Weakly Supervised Causal Disentanglement
di: Sheth, Paras, et al.
Pubblicazione: (2024)
di: Sheth, Paras, et al.
Pubblicazione: (2024)
EAGLE: A Domain Generalization Framework for AI-generated Text Detection
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2024)
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2024)
Towards LLM-guided Causal Explainability for Black-box Text Classifiers
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2023)
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2023)
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
di: Kumarage, Tharindu, et al.
Pubblicazione: (2024)
di: Kumarage, Tharindu, et al.
Pubblicazione: (2024)
Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2024)
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2024)
Can Large Language Models Infer Causal Relationships from Real-World Text?
di: Saklad, Ryan, et al.
Pubblicazione: (2025)
di: Saklad, Ryan, et al.
Pubblicazione: (2025)
Can Knowledge Graphs Reduce Hallucinations in LLMs? : A Survey
di: Agrawal, Garima, et al.
Pubblicazione: (2023)
di: Agrawal, Garima, et al.
Pubblicazione: (2023)
RedditESS: A Mental Health Social Support Interaction Dataset -- Understanding Effective Social Support to Refine AI-Driven Support Tools
di: Alghamdi, Zeyad, et al.
Pubblicazione: (2025)
di: Alghamdi, Zeyad, et al.
Pubblicazione: (2025)
Causal Feature Selection for Responsible Machine Learning
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
Exploiting Class Probabilities for Black-box Sentence-level Attacks
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
Adversarial Text Purification: A Large Language Model Approach for Defense
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
Mindful-RAG: A Study of Points of Failure in Retrieval Augmented Generation
di: Agrawal, Garima, et al.
Pubblicazione: (2024)
di: Agrawal, Garima, et al.
Pubblicazione: (2024)
A Generative Approach to Surrogate-based Black-box Attacks
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
Domain Knowledge-Enhanced LLMs for Fraud and Concept Drift Detection
di: Şenol, Ali, et al.
Pubblicazione: (2025)
di: Şenol, Ali, et al.
Pubblicazione: (2025)
Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
di: Nirmal, Ayushi, et al.
Pubblicazione: (2024)
di: Nirmal, Ayushi, et al.
Pubblicazione: (2024)
Joint Detection of Fraud and Concept Drift inOnline Conversations with LLM-Assisted Judgment
di: Senol, Ali, et al.
Pubblicazione: (2025)
di: Senol, Ali, et al.
Pubblicazione: (2025)
Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework
di: Şenol, Ali, et al.
Pubblicazione: (2026)
di: Şenol, Ali, et al.
Pubblicazione: (2026)
DAGverse: Building Document-Grounded Semantic DAGs from Scientific Papers
di: Wan, Shu, et al.
Pubblicazione: (2026)
di: Wan, Shu, et al.
Pubblicazione: (2026)
Personalized Attacks of Social Engineering in Multi-turn Conversations: LLM Agents for Simulation and Detection
di: Kumarage, Tharindu, et al.
Pubblicazione: (2025)
di: Kumarage, Tharindu, et al.
Pubblicazione: (2025)
Causality Guided Representation Learning for Cross-Style Hate Speech Detection
di: Zhao, Chengshuai, et al.
Pubblicazione: (2025)
di: Zhao, Chengshuai, et al.
Pubblicazione: (2025)
Why Do Large Language Models Generate Harmful Content?
di: Ganguli, Rajesh, et al.
Pubblicazione: (2026)
di: Ganguli, Rajesh, et al.
Pubblicazione: (2026)
Simulating Meaning, Nevermore! Introducing ICR: A Semiotic-Hermeneutic Metric for Evaluating Meaning in LLM Text Summaries
di: Perez, Natalie, et al.
Pubblicazione: (2026)
di: Perez, Natalie, et al.
Pubblicazione: (2026)
Ontology-Aware RAG for Improved Question-Answering in Cybersecurity Education
di: Zhao, Chengshuai, et al.
Pubblicazione: (2024)
di: Zhao, Chengshuai, et al.
Pubblicazione: (2024)
FACTOID: FACtual enTailment fOr hallucInation Detection
di: Rawte, Vipula, et al.
Pubblicazione: (2024)
di: Rawte, Vipula, et al.
Pubblicazione: (2024)
Generative Data Augmentation using LLMs improves Distributional Robustness in Question Answering
di: Chowdhury, Arijit Ghosh, et al.
Pubblicazione: (2023)
di: Chowdhury, Arijit Ghosh, et al.
Pubblicazione: (2023)
A Comprehensive Survey of Accelerated Generation Techniques in Large Language Models
di: Khoshnoodi, Mahsa, et al.
Pubblicazione: (2024)
di: Khoshnoodi, Mahsa, et al.
Pubblicazione: (2024)
Towards Safety Reasoning in LLMs: AI-agentic Deliberation for Policy-embedded CoT Data Creation
di: Kumarage, Tharindu, et al.
Pubblicazione: (2025)
di: Kumarage, Tharindu, et al.
Pubblicazione: (2025)
On the Relationship between Sentence Analogy Identification and Sentence Structure Encoding in Large Language Models
di: Wijesiriwardene, Thilini, et al.
Pubblicazione: (2023)
di: Wijesiriwardene, Thilini, et al.
Pubblicazione: (2023)
Beyond-RAG: Question Identification and Answer Generation in Real-Time Conversations
di: Agrawal, Garima, et al.
Pubblicazione: (2024)
di: Agrawal, Garima, et al.
Pubblicazione: (2024)
KnowledgePrompts: Exploring the Abilities of Large Language Models to Solve Proportional Analogies via Knowledge-Enhanced Prompting
di: Wijesiriwardene, Thilini, et al.
Pubblicazione: (2024)
di: Wijesiriwardene, Thilini, et al.
Pubblicazione: (2024)
Mental Health Equity in LLMs: Leveraging Multi-Hop Question Answering to Detect Amplified and Silenced Perspectives
di: Haider, Batool, et al.
Pubblicazione: (2025)
di: Haider, Batool, et al.
Pubblicazione: (2025)
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications?
di: Sinha, Neelabh, et al.
Pubblicazione: (2024)
di: Sinha, Neelabh, et al.
Pubblicazione: (2024)
Decoding the Diversity: A Review of the Indic AI Research Landscape
di: KJ, Sankalp, et al.
Pubblicazione: (2024)
di: KJ, Sankalp, et al.
Pubblicazione: (2024)
Reversing the Paradigm: Building AI-First Systems with Human Guidance
di: Spera, Cosimo, et al.
Pubblicazione: (2025)
di: Spera, Cosimo, et al.
Pubblicazione: (2025)
CyberBOT: Towards Reliable Cybersecurity Education via Ontology-Grounded Retrieval Augmented Generation
di: Zhao, Chengshuai, et al.
Pubblicazione: (2025)
di: Zhao, Chengshuai, et al.
Pubblicazione: (2025)
Is Contrasting All You Need? Contrastive Learning for the Detection and Attribution of AI-generated Text
di: La Cava, Lucio, et al.
Pubblicazione: (2024)
di: La Cava, Lucio, et al.
Pubblicazione: (2024)
"Sorry, Come Again?" Prompting -- Enhancing Comprehension and Diminishing Hallucination with [PAUSE]-injected Optimal Paraphrasing
di: Rawte, Vipula, et al.
Pubblicazione: (2024)
di: Rawte, Vipula, et al.
Pubblicazione: (2024)
Exploring the Frontier of Vision-Language Models: A Survey of Current Methodologies and Future Directions
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
Human-Readable Adversarial Prompts: An Investigation into LLM Vulnerabilities Using Situational Context
di: Das, Nilanjana, et al.
Pubblicazione: (2024)
di: Das, Nilanjana, et al.
Pubblicazione: (2024)
AI-generated Text Detection with a GLTR-based Approach
di: Wu, Lucía Yan, et al.
Pubblicazione: (2025)
di: Wu, Lucía Yan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Cross-Platform Hate Speech Detection with Weakly Supervised Causal Disentanglement
di: Sheth, Paras, et al.
Pubblicazione: (2024) -
EAGLE: A Domain Generalization Framework for AI-generated Text Detection
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2024) -
Towards LLM-guided Causal Explainability for Black-box Text Classifiers
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2023) -
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
di: Kumarage, Tharindu, et al.
Pubblicazione: (2024) -
Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2024)