Zero-Shot Belief: A Hard Problem for LLMs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Murzaku, John, Rambow, Owen |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
OmniVox: Zero-Shot Emotion Recognition with Omni-LLMs
par: Murzaku, John, et autres
Publié: (2025)
par: Murzaku, John, et autres
Publié: (2025)
Multimodal Belief Prediction
par: Murzaku, John, et autres
Publié: (2024)
par: Murzaku, John, et autres
Publié: (2024)
Synthetic Audio Helps for Cognitive State Tasks
par: Soubki, Adil, et autres
Publié: (2025)
par: Soubki, Adil, et autres
Publié: (2025)
Training LLMs to Recognize Hedges in Spontaneous Narratives
par: Paige, Amie J., et autres
Publié: (2024)
par: Paige, Amie J., et autres
Publié: (2024)
Evaluating LLMs with Multiple Problems at once
par: Wang, Zhengxiang, et autres
Publié: (2024)
par: Wang, Zhengxiang, et autres
Publié: (2024)
Views Are My Own, but Also Yours: Benchmarking Theory of Mind Using Common Ground
par: Soubki, Adil, et autres
Publié: (2024)
par: Soubki, Adil, et autres
Publié: (2024)
Clustering Document Parts: Detecting and Characterizing Influence Campaigns from Documents
par: Wang, Zhengxiang, et autres
Publié: (2024)
par: Wang, Zhengxiang, et autres
Publié: (2024)
Intention and Face in Dialog
par: Soubki, Adil, et autres
Publié: (2024)
par: Soubki, Adil, et autres
Publié: (2024)
LLMs can Perform Multi-Dimensional Analytic Writing Assessments: A Case Study of L2 Graduate-Level Academic English Writing
par: Wang, Zhengxiang, et autres
Publié: (2025)
par: Wang, Zhengxiang, et autres
Publié: (2025)
Examining Gender and Power on Wikipedia Through Face and Politeness
par: Soubki, Adil, et autres
Publié: (2024)
par: Soubki, Adil, et autres
Publié: (2024)
Active Few-Shot Learning for Text Classification
par: Ahmadnia, Saeed, et autres
Publié: (2025)
par: Ahmadnia, Saeed, et autres
Publié: (2025)
On Zero-Shot Counterspeech Generation by LLMs
par: Saha, Punyajoy, et autres
Publié: (2024)
par: Saha, Punyajoy, et autres
Publié: (2024)
LVLMs are Bad at Overhearing Human Referential Communication
par: Wang, Zhengxiang, et autres
Publié: (2025)
par: Wang, Zhengxiang, et autres
Publié: (2025)
Large Language Models for Zero-Shot Multicultural Name Recognition
par: Phonchai, Thanakorn, et autres
Publié: (2025)
par: Phonchai, Thanakorn, et autres
Publié: (2025)
Are LLMs Good Zero-Shot Fallacy Classifiers?
par: Pan, Fengjun, et autres
Publié: (2024)
par: Pan, Fengjun, et autres
Publié: (2024)
Better Benchmarking LLMs for Zero-Shot Dependency Parsing
par: Ezquerro, Ana, et autres
Publié: (2025)
par: Ezquerro, Ana, et autres
Publié: (2025)
Towards Zero-Shot, Controllable Dialog Planning with LLMs
par: Väth, Dirk, et autres
Publié: (2024)
par: Väth, Dirk, et autres
Publié: (2024)
LLMs Are Zero-Shot Context-Aware Simultaneous Translators
par: Koshkin, Roman, et autres
Publié: (2024)
par: Koshkin, Roman, et autres
Publié: (2024)
XAM: Interactive Explainability for Authorship Attribution Models
par: Alshomary, Milad, et autres
Publié: (2025)
par: Alshomary, Milad, et autres
Publié: (2025)
Nsanku: Evaluating Zero-Shot Translation Performance of LLMs for Ghanaian Languages
par: Moore, Stephen E., et autres
Publié: (2026)
par: Moore, Stephen E., et autres
Publié: (2026)
Taxonomy-Guided Zero-Shot Recommendations with LLMs
par: Liang, Yueqing, et autres
Publié: (2024)
par: Liang, Yueqing, et autres
Publié: (2024)
Hard-Synth: Synthesizing Diverse Hard Samples for ASR using Zero-Shot TTS and LLM
par: Yu, Jiawei, et autres
Publié: (2024)
par: Yu, Jiawei, et autres
Publié: (2024)
GRP: Goal-Reversed Prompting for Zero-Shot Evaluation with LLMs
par: Song, Mingyang, et autres
Publié: (2025)
par: Song, Mingyang, et autres
Publié: (2025)
Veracity Bias and Beyond: Uncovering LLMs' Hidden Beliefs in Problem-Solving Reasoning
par: Zhou, Yue, et autres
Publié: (2025)
par: Zhou, Yue, et autres
Publié: (2025)
Generalization to Political Beliefs from Fine-Tuning on Sports Team Preferences
par: Terry, Owen
Publié: (2026)
par: Terry, Owen
Publié: (2026)
Are Stereotypes Leading LLMs' Zero-Shot Stance Detection ?
par: Dubreuil, Anthony, et autres
Publié: (2025)
par: Dubreuil, Anthony, et autres
Publié: (2025)
Zero-Shot Clinical Trial Patient Matching with LLMs
par: Wornow, Michael, et autres
Publié: (2024)
par: Wornow, Michael, et autres
Publié: (2024)
Zero-Shot Stance Detection using Contextual Data Generation with LLMs
par: Mahmoudi, Ghazaleh, et autres
Publié: (2024)
par: Mahmoudi, Ghazaleh, et autres
Publié: (2024)
Gram2Vec: An Interpretable Document Vectorizer
par: Zeng, Peter, et autres
Publié: (2024)
par: Zeng, Peter, et autres
Publié: (2024)
NormSAGE: Multi-Lingual Multi-Cultural Norm Discovery from Conversations On-the-Fly
par: Fung, Yi R., et autres
Publié: (2022)
par: Fung, Yi R., et autres
Publié: (2022)
HARGPT: Are LLMs Zero-Shot Human Activity Recognizers?
par: Ji, Sijie, et autres
Publié: (2024)
par: Ji, Sijie, et autres
Publié: (2024)
Contrastive Distillation of Emotion Knowledge from LLMs for Zero-Shot Emotion Recognition
par: Niu, Minxue, et autres
Publié: (2025)
par: Niu, Minxue, et autres
Publié: (2025)
Residualized Similarity for Faithfully Explainable Authorship Verification
par: Zeng, Peter, et autres
Publié: (2025)
par: Zeng, Peter, et autres
Publié: (2025)
DAWN-ICL: Strategic Planning of Problem-solving Trajectories for Zero-Shot In-Context Learning
par: Tang, Xinyu, et autres
Publié: (2024)
par: Tang, Xinyu, et autres
Publié: (2024)
LLMs are not Zero-Shot Reasoners for Biomedical Information Extraction
par: Nagar, Aishik, et autres
Publié: (2024)
par: Nagar, Aishik, et autres
Publié: (2024)
Fundamental Problems With Model Editing: How Should Rational Belief Revision Work in LLMs?
par: Hase, Peter, et autres
Publié: (2024)
par: Hase, Peter, et autres
Publié: (2024)
Lowest Span Confidence: A Zero-Shot Metric for Efficient and Black-Box Hallucination Detection in LLMs
par: Qiao, Yitong, et autres
Publié: (2026)
par: Qiao, Yitong, et autres
Publié: (2026)
Zero-Shot ATC Coding with Large Language Models for Clinical Assessments
par: Chen, Zijian, et autres
Publié: (2024)
par: Chen, Zijian, et autres
Publié: (2024)
Zero-Shot Tokenizer Transfer
par: Minixhofer, Benjamin, et autres
Publié: (2024)
par: Minixhofer, Benjamin, et autres
Publié: (2024)
Multimodal LLMs Can Reason about Aesthetics in Zero-Shot
par: Jiang, Ruixiang, et autres
Publié: (2025)
par: Jiang, Ruixiang, et autres
Publié: (2025)
Documents similaires
-
OmniVox: Zero-Shot Emotion Recognition with Omni-LLMs
par: Murzaku, John, et autres
Publié: (2025) -
Multimodal Belief Prediction
par: Murzaku, John, et autres
Publié: (2024) -
Synthetic Audio Helps for Cognitive State Tasks
par: Soubki, Adil, et autres
Publié: (2025) -
Training LLMs to Recognize Hedges in Spontaneous Narratives
par: Paige, Amie J., et autres
Publié: (2024) -
Evaluating LLMs with Multiple Problems at once
par: Wang, Zhengxiang, et autres
Publié: (2024)