Benchmarking zero-shot stance detection with FlanT5-XXL: Insights from training data, prompting, and decoding strategies into its near-SoTA performance
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Aiyappa, Rachith, Senthilmani, Shruthi, An, Jisun, Kwak, Haewoon, Ahn, Yong-Yeol |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Can we trust the evaluation on ChatGPT?
par: Aiyappa, Rachith, et autres
Publié: (2023)
par: Aiyappa, Rachith, et autres
Publié: (2023)
A semantic embedding space based on large language models for modelling human beliefs
par: Lee, Byunghwee, et autres
Publié: (2024)
par: Lee, Byunghwee, et autres
Publié: (2024)
What Helps Language Models Predict Human Beliefs: Demographics or Prior Stances?
par: Malone, Joseph, et autres
Publié: (2025)
par: Malone, Joseph, et autres
Publié: (2025)
Emergence of simple and complex contagion dynamics from weighted belief networks
par: Aiyappa, Rachith, et autres
Publié: (2023)
par: Aiyappa, Rachith, et autres
Publié: (2023)
LLMs Can Infer Political Alignment from Online Conversations
par: Lee, Byunghwee, et autres
Publié: (2026)
par: Lee, Byunghwee, et autres
Publié: (2026)
Can Lessons From Human Teams Be Applied to Multi-Agent Systems? The Role of Structure, Diversity, and Interaction Dynamics
par: Muralidharan, Rasika, et autres
Publié: (2025)
par: Muralidharan, Rasika, et autres
Publié: (2025)
Vulnerability of LLMs' Stated Beliefs? LLMs Belief Resistance Check Through Strategic Persuasive Conversation Interventions
par: Huang, Fan, et autres
Publié: (2026)
par: Huang, Fan, et autres
Publié: (2026)
ToBlend: Token-Level Blending With an Ensemble of LLMs to Attack AI-Generated Text Detection
par: Huang, Fan, et autres
Publié: (2024)
par: Huang, Fan, et autres
Publié: (2024)
Understanding Moral Reasoning Trajectories in Large Language Models: Toward Probing-Based Explainability
par: Huang, Fan, et autres
Publié: (2026)
par: Huang, Fan, et autres
Publié: (2026)
Emergence of Stereotypes and Affective Polarization from Belief Network Dynamics
par: Seckin, Ozgur Can, et autres
Publié: (2026)
par: Seckin, Ozgur Can, et autres
Publié: (2026)
Rematch: Robust and Efficient Matching of Local Knowledge Graphs to Improve Structural and Semantic Similarity
par: Kachwala, Zoher, et autres
Publié: (2024)
par: Kachwala, Zoher, et autres
Publié: (2024)
ChatGPT Rates Natural Language Explanation Quality Like Humans: But on Which Scales?
par: Huang, Fan, et autres
Publié: (2024)
par: Huang, Fan, et autres
Publié: (2024)
CogBias: Measuring and Mitigating Cognitive Bias in Large Language Models
par: Huang, Fan, et autres
Publié: (2026)
par: Huang, Fan, et autres
Publié: (2026)
A Cross-Cultural Comparison of LLM-based Public Opinion Simulation: Evaluating Chinese and U.S. Models on Diverse Societies
par: Qi, Weihong, et autres
Publié: (2025)
par: Qi, Weihong, et autres
Publié: (2025)
Somatic in the East, Psychological in the West?: Investigating Clinically-Grounded Cross-Cultural Depression Symptom Expression in LLMs
par: Sakai, Shintaro, et autres
Publié: (2025)
par: Sakai, Shintaro, et autres
Publié: (2025)
Quantifying Gender Stereotypes in Japan between 1900 and 1999 with Word Embeddings
par: Sakai, Shintaro, et autres
Publié: (2025)
par: Sakai, Shintaro, et autres
Publié: (2025)
Implicit degree bias in the link prediction task
par: Aiyappa, Rachith, et autres
Publié: (2024)
par: Aiyappa, Rachith, et autres
Publié: (2024)
SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
par: Wang, Xiyao, et autres
Publié: (2025)
par: Wang, Xiyao, et autres
Publié: (2025)
XChoice: Explainable Evaluation of AI-Human Alignment in LLM-based Constrained Choice Decision Making
par: Qi, Weihong, et autres
Publié: (2026)
par: Qi, Weihong, et autres
Publié: (2026)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
par: Sapkota, Ranjan, et autres
Publié: (2025)
par: Sapkota, Ranjan, et autres
Publié: (2025)
XXL-HSC: Host properties of X-ray detected AGNs in XXL clusters
par: Drigga, E., et autres
Publié: (2025)
par: Drigga, E., et autres
Publié: (2025)
Cognitive Linguistic Identity Fusion Score (CLIFS): A Scalable Cognition-Informed Approach to Quantifying Identity Fusion from Text
par: Wright, Devin R., et autres
Publié: (2025)
par: Wright, Devin R., et autres
Publié: (2025)
PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media
par: Kachwala, Zoher, et autres
Publié: (2026)
par: Kachwala, Zoher, et autres
Publié: (2026)
Identifying and characterizing superspreaders of low-credibility content on Twitter
par: DeVerna, Matthew R., et autres
Publié: (2022)
par: DeVerna, Matthew R., et autres
Publié: (2022)
FlanEC: Exploring Flan-T5 for Post-ASR Error Correction
par: La Quatra, Moreno, et autres
Publié: (2025)
par: La Quatra, Moreno, et autres
Publié: (2025)
How the cascade inference problem distorts information diffusion
par: DeVerna, Matthew R., et autres
Publié: (2024)
par: DeVerna, Matthew R., et autres
Publié: (2024)
SynSym: A Synthetic Data Generation Framework for Psychiatric Symptom Identification
par: Kang, Migyeong, et autres
Publié: (2026)
par: Kang, Migyeong, et autres
Publié: (2026)
Evaluating and explaining training strategies for zero-shot cross-lingual news sentiment analysis
par: Andrenšek, Luka, et autres
Publié: (2024)
par: Andrenšek, Luka, et autres
Publié: (2024)
VCP-CLIP: A visual context prompting model for zero-shot anomaly segmentation
par: Qu, Zhen, et autres
Publié: (2024)
par: Qu, Zhen, et autres
Publié: (2024)
PrivacyBench: A Conversational Benchmark for Evaluating Privacy in Personalized AI
par: Mukhopadhyay, Srija, et autres
Publié: (2025)
par: Mukhopadhyay, Srija, et autres
Publié: (2025)
SenCLIP: Enhancing zero-shot land-use mapping for Sentinel-2 with ground-level prompting
par: Jain, Pallavi, et autres
Publié: (2024)
par: Jain, Pallavi, et autres
Publié: (2024)
A sound description: Exploring prompt templates and class descriptions to enhance zero-shot audio classification
par: Olvera, Michel, et autres
Publié: (2024)
par: Olvera, Michel, et autres
Publié: (2024)
Network community detection via neural embeddings
par: Kojaku, Sadamori, et autres
Publié: (2023)
par: Kojaku, Sadamori, et autres
Publié: (2023)
Chain-of-Though (CoT) prompting strategies for medical error detection and correction
par: Wu, Zhaolong, et autres
Publié: (2024)
par: Wu, Zhaolong, et autres
Publié: (2024)
On the test-time zero-shot generalization of vision-language models: Do we really need prompt learning?
par: Zanella, Maxime, et autres
Publié: (2024)
par: Zanella, Maxime, et autres
Publié: (2024)
Single-shot and two-shot decoding with generalized bicycle codes
par: Lin, Hsiang-Ku, et autres
Publié: (2025)
par: Lin, Hsiang-Ku, et autres
Publié: (2025)
Instance Segmentation XXL-CT Challenge of a Historic Airplane
par: Gruber, Roland, et autres
Publié: (2024)
par: Gruber, Roland, et autres
Publié: (2024)
Vision-Flan: Scaling Human-Labeled Tasks in Visual Instruction Tuning
par: Xu, Zhiyang, et autres
Publié: (2024)
par: Xu, Zhiyang, et autres
Publié: (2024)
Exploiting contextual information to improve stance detection in informal political discourse with LLMs
par: Sucu, Arman Engin, et autres
Publié: (2026)
par: Sucu, Arman Engin, et autres
Publié: (2026)
The LUMirage: An independent evaluation of zero-shot performance in the LUMIR challenge
par: Jena, Rohit, et autres
Publié: (2025)
par: Jena, Rohit, et autres
Publié: (2025)
Documents similaires
-
Can we trust the evaluation on ChatGPT?
par: Aiyappa, Rachith, et autres
Publié: (2023) -
A semantic embedding space based on large language models for modelling human beliefs
par: Lee, Byunghwee, et autres
Publié: (2024) -
What Helps Language Models Predict Human Beliefs: Demographics or Prior Stances?
par: Malone, Joseph, et autres
Publié: (2025) -
Emergence of simple and complex contagion dynamics from weighted belief networks
par: Aiyappa, Rachith, et autres
Publié: (2023) -
LLMs Can Infer Political Alignment from Online Conversations
par: Lee, Byunghwee, et autres
Publié: (2026)