Model Attribution in LLM-Generated Disinformation: A Domain Generalization Approach with Supervised Contrastive Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Beigi, Alimohammad, Tan, Zhen, Mudiam, Nivedh, Chen, Canyu, Shu, Kai, Liu, Huan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
Can LLM-Generated Misinformation Be Detected?
by: Chen, Canyu, et al.
Published: (2023)
by: Chen, Canyu, et al.
Published: (2023)
Catching Chameleons: Detecting Evolving Disinformation Generated using Large Language Models
by: Jiang, Bohan, et al.
Published: (2024)
by: Jiang, Bohan, et al.
Published: (2024)
Can LLMs Improve Multimodal Fact-Checking by Asking Relevant Questions?
by: Beigi, Alimohammad, et al.
Published: (2024)
by: Beigi, Alimohammad, et al.
Published: (2024)
Large Language Models for Data Annotation and Synthesis: A Survey
by: Tan, Zhen, et al.
Published: (2024)
by: Tan, Zhen, et al.
Published: (2024)
Can Large Language Models Identify Authorship?
by: Huang, Baixiang, et al.
Published: (2024)
by: Huang, Baixiang, et al.
Published: (2024)
Tuning-Free Accountable Intervention for LLM Deployment -- A Metacognitive Approach
by: Tan, Zhen, et al.
Published: (2024)
by: Tan, Zhen, et al.
Published: (2024)
Evaluation of LLM Vulnerabilities to Being Misused for Personalized Disinformation Generation
by: Zugecova, Aneta, et al.
Published: (2024)
by: Zugecova, Aneta, et al.
Published: (2024)
Authorship Attribution in the Era of LLMs: Problems, Methodologies, and Challenges
by: Huang, Baixiang, et al.
Published: (2024)
by: Huang, Baixiang, et al.
Published: (2024)
DOGe: Defensive Output Generation for LLM Protection Against Knowledge Distillation
by: Li, Pingzhi, et al.
Published: (2025)
by: Li, Pingzhi, et al.
Published: (2025)
Can Knowledge Editing Really Correct Hallucinations?
by: Huang, Baixiang, et al.
Published: (2024)
by: Huang, Baixiang, et al.
Published: (2024)
P2S: Probabilistic Process Supervision for General-Domain Reasoning Question Answering
by: Zhong, Wenlin, et al.
Published: (2026)
by: Zhong, Wenlin, et al.
Published: (2026)
SLearnLLM: A Self-Learning Framework for Efficient Domain-Specific Adaptation of Large Language Models
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
Process-Supervised Reward Models for Verifying Clinical Note Generation: A Scalable Approach Guided by Domain Expertise
by: Wang, Hanyin, et al.
Published: (2024)
by: Wang, Hanyin, et al.
Published: (2024)
Nearest Neighbor Speculative Decoding for LLM Generation and Attribution
by: Li, Minghan, et al.
Published: (2024)
by: Li, Minghan, et al.
Published: (2024)
LLMSniffer: Detecting LLM-Generated Code via GraphCodeBERT and Supervised Contrastive Learning
by: Dihan, Mahir Labib, et al.
Published: (2026)
by: Dihan, Mahir Labib, et al.
Published: (2026)
ClinicalBench: Can LLMs Beat Traditional ML Models in Clinical Prediction?
by: Chen, Canyu, et al.
Published: (2024)
by: Chen, Canyu, et al.
Published: (2024)
Social Media for Mental Health: Data, Methods, and Findings
by: Kamarudin, Nur Shazwani, et al.
Published: (2025)
by: Kamarudin, Nur Shazwani, et al.
Published: (2025)
Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks
by: Tan, Rongyuan, et al.
Published: (2026)
by: Tan, Rongyuan, et al.
Published: (2026)
Domain Adaptation for Japanese Sentence Embeddings with Contrastive Learning based on Synthetic Sentence Generation
by: Chen, Zihao, et al.
Published: (2025)
by: Chen, Zihao, et al.
Published: (2025)
Model Editing as a Double-Edged Sword: Steering Agent Ethical Behavior Toward Beneficence or Harm
by: Huang, Baixiang, et al.
Published: (2025)
by: Huang, Baixiang, et al.
Published: (2025)
Disinformation Capabilities of Large Language Models
by: Vykopal, Ivan, et al.
Published: (2023)
by: Vykopal, Ivan, et al.
Published: (2023)
A Cross-Domain Study of the Use of Persuasion Techniques in Online Disinformation
by: Leite, João A., et al.
Published: (2024)
by: Leite, João A., et al.
Published: (2024)
Defending Against Disinformation Attacks in Open-Domain Question Answering
by: Weller, Orion, et al.
Published: (2022)
by: Weller, Orion, et al.
Published: (2022)
Personalized Text Generation with Contrastive Activation Steering
by: Zhang, Jinghao, et al.
Published: (2025)
by: Zhang, Jinghao, et al.
Published: (2025)
General-Reasoner: Advancing LLM Reasoning Across All Domains
by: Ma, Xueguang, et al.
Published: (2025)
by: Ma, Xueguang, et al.
Published: (2025)
Modeling Comparative Logical Relation with Contrastive Learning for Text Generation
by: Dan, Yuhao, et al.
Published: (2024)
by: Dan, Yuhao, et al.
Published: (2024)
ChartInsighter: An Approach for Mitigating Hallucination in Time-series Chart Summary Generation with A Benchmark Dataset
by: Wang, Fen, et al.
Published: (2025)
by: Wang, Fen, et al.
Published: (2025)
Lying Blindly: Bypassing ChatGPT's Safeguards to Generate Hard-to-Detect Disinformation Claims
by: Heppell, Freddy, et al.
Published: (2024)
by: Heppell, Freddy, et al.
Published: (2024)
Contextualization Distillation from Large Language Model for Knowledge Graph Completion
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
Generalization of Medical Large Language Models through Cross-Domain Weak Supervision
by: Long, Robert, et al.
Published: (2025)
by: Long, Robert, et al.
Published: (2025)
DivScore: Zero-Shot Detection of LLM-Generated Text in Specialized Domains
by: Chen, Zhihui, et al.
Published: (2025)
by: Chen, Zhihui, et al.
Published: (2025)
Towards Effective Model Editing for LLM Personalization
by: Huang, Baixiang, et al.
Published: (2025)
by: Huang, Baixiang, et al.
Published: (2025)
A Multilingual, Large-Scale Study of the Interplay between LLM Safeguards, Personalisation, and Disinformation
by: Leite, João A., et al.
Published: (2025)
by: Leite, João A., et al.
Published: (2025)
Contrastive Learning on LLM Back Generation Treebank for Cross-domain Constituency Parsing
by: Guo, Peiming, et al.
Published: (2025)
by: Guo, Peiming, et al.
Published: (2025)
CAMO: Causality-Guided Adversarial Multimodal Domain Generalization for Crisis Classification
by: Ma, Pingchuan, et al.
Published: (2025)
by: Ma, Pingchuan, et al.
Published: (2025)
Contrastive Learning in Distilled Models
by: Lim, Valerie, et al.
Published: (2024)
by: Lim, Valerie, et al.
Published: (2024)
LLM Attributor: Interactive Visual Attribution for LLM Generation
by: Lee, Seongmin, et al.
Published: (2024)
by: Lee, Seongmin, et al.
Published: (2024)
Code Fingerprints: Disentangled Attribution of LLM-Generated Code
by: Guo, Jiaxun, et al.
Published: (2026)
by: Guo, Jiaxun, et al.
Published: (2026)
Improving Attributed Text Generation of Large Language Models via Preference Learning
by: Li, Dongfang, et al.
Published: (2024)
by: Li, Dongfang, et al.
Published: (2024)
Similar Items
-
From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge
by: Li, Dawei, et al.
Published: (2024) -
Can LLM-Generated Misinformation Be Detected?
by: Chen, Canyu, et al.
Published: (2023) -
Catching Chameleons: Detecting Evolving Disinformation Generated using Large Language Models
by: Jiang, Bohan, et al.
Published: (2024) -
Can LLMs Improve Multimodal Fact-Checking by Asking Relevant Questions?
by: Beigi, Alimohammad, et al.
Published: (2024) -
Large Language Models for Data Annotation and Synthesis: A Survey
by: Tan, Zhen, et al.
Published: (2024)