Bot Meets Shortcut: How Can LLMs Aid in Handling Unknown Invariance OOD Scenarios?
Fuente:
arXiv
Salvato in:
| Autori principali: | Zheng, Shiyan, Wan, Herun, Luo, Minnan, Huang, Junhang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Truth over Tricks: Measuring and Mitigating Shortcut Learning in Misinformation Detection
di: Wan, Herun, et al.
Pubblicazione: (2025)
di: Wan, Herun, et al.
Pubblicazione: (2025)
What Does the Bot Say? Opportunities and Risks of Large Language Models in Social Media Bot Detection
di: Feng, Shangbin, et al.
Pubblicazione: (2024)
di: Feng, Shangbin, et al.
Pubblicazione: (2024)
On the Risk of Evidence Pollution for Malicious Social Text Detection in the Era of LLMs
di: Wan, Herun, et al.
Pubblicazione: (2024)
di: Wan, Herun, et al.
Pubblicazione: (2024)
How Do Social Bots Participate in Misinformation Spread? A Comprehensive Dataset and Analysis
di: Wan, Herun, et al.
Pubblicazione: (2024)
di: Wan, Herun, et al.
Pubblicazione: (2024)
HACo-Det: A Study Towards Fine-Grained Machine-Generated Text Detection under Human-AI Coauthoring
di: Su, Zhixiong, et al.
Pubblicazione: (2025)
di: Su, Zhixiong, et al.
Pubblicazione: (2025)
DiFaR: Enhancing Multimodal Misinformation Detection with Diverse, Factual, and Relevant Rationales
di: Wan, Herun, et al.
Pubblicazione: (2025)
di: Wan, Herun, et al.
Pubblicazione: (2025)
GuessBench: Sensemaking Multimodal Creativity in the Wild
di: Zhu, Zifeng, et al.
Pubblicazione: (2025)
di: Zhu, Zifeng, et al.
Pubblicazione: (2025)
DELL: Generating Reactions and Explanations for LLM-Based Misinformation Detection
di: Wan, Herun, et al.
Pubblicazione: (2024)
di: Wan, Herun, et al.
Pubblicazione: (2024)
The Facade of Truth: Uncovering and Mitigating LLM Susceptibility to Deceptive Evidence
di: Wan, Herun, et al.
Pubblicazione: (2026)
di: Wan, Herun, et al.
Pubblicazione: (2026)
Continuously Steering LLMs Sensitivity to Contextual Knowledge with Proxy Models
di: Wang, Yilin, et al.
Pubblicazione: (2025)
di: Wang, Yilin, et al.
Pubblicazione: (2025)
LMBot: Distilling Graph Knowledge into Language Model for Graph-less Deployment in Twitter Bot Detection
di: Cai, Zijian, et al.
Pubblicazione: (2023)
di: Cai, Zijian, et al.
Pubblicazione: (2023)
Do LLMs Overcome Shortcut Learning? An Evaluation of Shortcut Challenges in Large Language Models
di: Yuan, Yu, et al.
Pubblicazione: (2024)
di: Yuan, Yu, et al.
Pubblicazione: (2024)
How Good are LLMs at Relation Extraction under Low-Resource Scenario? Comprehensive Evaluation
di: Jinensibieke, Dawulie, et al.
Pubblicazione: (2024)
di: Jinensibieke, Dawulie, et al.
Pubblicazione: (2024)
Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs
di: Yan, Lecheng, et al.
Pubblicazione: (2026)
di: Yan, Lecheng, et al.
Pubblicazione: (2026)
Deceptive Semantic Shortcuts on Reasoning Chains: How Far Can Models Go without Hallucination?
di: Li, Bangzheng, et al.
Pubblicazione: (2023)
di: Li, Bangzheng, et al.
Pubblicazione: (2023)
How Much Noise Can BERT Handle? Insights from Multilingual Sentence Difficulty Detection
di: Khallaf, Nouran, et al.
Pubblicazione: (2026)
di: Khallaf, Nouran, et al.
Pubblicazione: (2026)
How to Handle Different Types of Out-of-Distribution Scenarios in Computational Argumentation? A Comprehensive and Fine-Grained Field Study
di: Waldis, Andreas, et al.
Pubblicazione: (2023)
di: Waldis, Andreas, et al.
Pubblicazione: (2023)
RTP-LX: Can LLMs Evaluate Toxicity in Multilingual Scenarios?
di: de Wynter, Adrian, et al.
Pubblicazione: (2024)
di: de Wynter, Adrian, et al.
Pubblicazione: (2024)
LLMs Can't Handle Peer Pressure: Crumbling under Multi-Agent Social Interactions
di: Song, Maojia, et al.
Pubblicazione: (2025)
di: Song, Maojia, et al.
Pubblicazione: (2025)
C2PO: Diagnosing and Disentangling Bias Shortcuts in LLMs
di: Feng, Xuan, et al.
Pubblicazione: (2025)
di: Feng, Xuan, et al.
Pubblicazione: (2025)
Polysemantic Dropout: Conformal OOD Detection for Specialized LLMs
di: Gupta, Ayush, et al.
Pubblicazione: (2025)
di: Gupta, Ayush, et al.
Pubblicazione: (2025)
How Well Do LLMs Handle Cantonese? Benchmarking Cantonese Capabilities of Large Language Models
di: Jiang, Jiyue, et al.
Pubblicazione: (2024)
di: Jiang, Jiyue, et al.
Pubblicazione: (2024)
Break the Chain: Large Language Models Can be Shortcut Reasoners
di: Ding, Mengru, et al.
Pubblicazione: (2024)
di: Ding, Mengru, et al.
Pubblicazione: (2024)
Social Science Meets LLMs: How Reliable Are Large Language Models in Social Simulations?
di: Huang, Yue, et al.
Pubblicazione: (2024)
di: Huang, Yue, et al.
Pubblicazione: (2024)
Farther the Shift, Sparser the Representation: Analyzing OOD Mechanisms in LLMs
di: Jin, Mingyu, et al.
Pubblicazione: (2026)
di: Jin, Mingyu, et al.
Pubblicazione: (2026)
Understanding the Ability of LLMs to Handle Character-Level Perturbation
di: Zhuo, Anyuan, et al.
Pubblicazione: (2025)
di: Zhuo, Anyuan, et al.
Pubblicazione: (2025)
Measuring Aleatoric and Epistemic Uncertainty in LLMs: Empirical Evaluation on ID and OOD QA Tasks
di: Wang, Kevin, et al.
Pubblicazione: (2025)
di: Wang, Kevin, et al.
Pubblicazione: (2025)
Unknown Unknowns: Why Hidden Intentions in LLMs Evade Detection
di: Srivastav, Devansh, et al.
Pubblicazione: (2026)
di: Srivastav, Devansh, et al.
Pubblicazione: (2026)
Can Many-Shot In-Context Learning Help LLMs as Evaluators? A Preliminary Empirical Study
di: Song, Mingyang, et al.
Pubblicazione: (2024)
di: Song, Mingyang, et al.
Pubblicazione: (2024)
Seeing to Generalize: How Visual Data Corrects Binding Shortcuts
di: Buzeta, Nicolas, et al.
Pubblicazione: (2026)
di: Buzeta, Nicolas, et al.
Pubblicazione: (2026)
Legal Minds, Algorithmic Decisions: How LLMs Apply Constitutional Principles in Complex Scenarios
di: Bignotti, Camilla, et al.
Pubblicazione: (2024)
di: Bignotti, Camilla, et al.
Pubblicazione: (2024)
SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning
di: Huang, Haoyu, et al.
Pubblicazione: (2026)
di: Huang, Haoyu, et al.
Pubblicazione: (2026)
MMLU-Pro+: Evaluating Higher-Order Reasoning and Shortcut Learning in LLMs
di: Taghanaki, Saeid Asgari, et al.
Pubblicazione: (2024)
di: Taghanaki, Saeid Asgari, et al.
Pubblicazione: (2024)
neuralFOMO: Can LLMs Handle Being Second Best? Measuring Envy-Like Preferences in Multi-Agent Settings
di: Ramamoorthy, Arnav, et al.
Pubblicazione: (2025)
di: Ramamoorthy, Arnav, et al.
Pubblicazione: (2025)
When LLMs Meet Cunning Texts: A Fallacy Understanding Benchmark for Large Language Models
di: Li, Yinghui, et al.
Pubblicazione: (2024)
di: Li, Yinghui, et al.
Pubblicazione: (2024)
Causality $\neq$ Invariance: Function and Concept Vectors in LLMs
di: Opiełka, Gustaw, et al.
Pubblicazione: (2026)
di: Opiełka, Gustaw, et al.
Pubblicazione: (2026)
Short-circuiting Shortcuts: Mechanistic Investigation of Shortcuts in Text Classification
di: Eshuijs, Leon, et al.
Pubblicazione: (2025)
di: Eshuijs, Leon, et al.
Pubblicazione: (2025)
Leveraging Prompts in LLMs to Overcome Imbalances in Complex Educational Text Data
di: McClure, Jeanne, et al.
Pubblicazione: (2024)
di: McClure, Jeanne, et al.
Pubblicazione: (2024)
How Well Can LLMs Echo Us? Evaluating AI Chatbots' Role-Play Ability with ECHO
di: Ng, Man Tik, et al.
Pubblicazione: (2024)
di: Ng, Man Tik, et al.
Pubblicazione: (2024)
CangjieBench: Benchmarking LLMs on a Low-Resource General-Purpose Programming Language
di: Cheng, Junhang, et al.
Pubblicazione: (2026)
di: Cheng, Junhang, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Truth over Tricks: Measuring and Mitigating Shortcut Learning in Misinformation Detection
di: Wan, Herun, et al.
Pubblicazione: (2025) -
What Does the Bot Say? Opportunities and Risks of Large Language Models in Social Media Bot Detection
di: Feng, Shangbin, et al.
Pubblicazione: (2024) -
On the Risk of Evidence Pollution for Malicious Social Text Detection in the Era of LLMs
di: Wan, Herun, et al.
Pubblicazione: (2024) -
How Do Social Bots Participate in Misinformation Spread? A Comprehensive Dataset and Analysis
di: Wan, Herun, et al.
Pubblicazione: (2024) -
HACo-Det: A Study Towards Fine-Grained Machine-Generated Text Detection under Human-AI Coauthoring
di: Su, Zhixiong, et al.
Pubblicazione: (2025)