Salvato in:
| Autori principali: | Muchovej, John, Royka, Amanda, Lee, Shane, Jara-Ettinger, Julian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.12150 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Testing the Depth of ChatGPT's Comprehension via Cross-Modal Tasks Based on ASCII-Art: GPT3.5's Abilities in Regard to Recognizing and Generating ASCII-Art Are Not Totally Lacking
di: Bayani, David
Pubblicazione: (2023)
di: Bayani, David
Pubblicazione: (2023)
Automated Meta Prompt Engineering for Alignment with the Theory of Mind
di: Baughman, Aaron, et al.
Pubblicazione: (2025)
di: Baughman, Aaron, et al.
Pubblicazione: (2025)
Playing Along: Learning a Double-Agent Defender for Belief Steering via Theory of Mind
di: Xiao, Hanqi, et al.
Pubblicazione: (2026)
di: Xiao, Hanqi, et al.
Pubblicazione: (2026)
A Notion of Complexity for Theory of Mind via Discrete World Models
di: Huang, X. Angelo, et al.
Pubblicazione: (2024)
di: Huang, X. Angelo, et al.
Pubblicazione: (2024)
Single layer tiny Co$^4$ outpaces GPT-2 and GPT-BERT
di: Zain, Noor Ul, et al.
Pubblicazione: (2025)
di: Zain, Noor Ul, et al.
Pubblicazione: (2025)
Explore Theory of Mind: Program-guided adversarial data generation for theory of mind reasoning
di: Sclar, Melanie, et al.
Pubblicazione: (2024)
di: Sclar, Melanie, et al.
Pubblicazione: (2024)
Small LLMs Do Not Learn a Generalizable Theory of Mind via Reinforcement Learning
di: Sarangi, Sneheel, et al.
Pubblicazione: (2025)
di: Sarangi, Sneheel, et al.
Pubblicazione: (2025)
Selective Deficits in LLM Mental Self-Modeling in a Behavior-Based Test of Theory of Mind
di: Ackerman, Christopher
Pubblicazione: (2026)
di: Ackerman, Christopher
Pubblicazione: (2026)
HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs
di: Chen, Junying, et al.
Pubblicazione: (2024)
di: Chen, Junying, et al.
Pubblicazione: (2024)
Architectural Flaw Detection in Civil Engineering Using GPT-4
di: Kumar, Saket, et al.
Pubblicazione: (2024)
di: Kumar, Saket, et al.
Pubblicazione: (2024)
ChatQA: Surpassing GPT-4 on Conversational QA and RAG
di: Liu, Zihan, et al.
Pubblicazione: (2024)
di: Liu, Zihan, et al.
Pubblicazione: (2024)
ArabianGPT: Native Arabic GPT-based Large Language Model
di: Koubaa, Anis, et al.
Pubblicazione: (2024)
di: Koubaa, Anis, et al.
Pubblicazione: (2024)
Predicting Training Re-evaluation Curves Enables Effective Data Curriculums for LLMs
di: Bergsma, Shane, et al.
Pubblicazione: (2025)
di: Bergsma, Shane, et al.
Pubblicazione: (2025)
MMToM-QA: Multimodal Theory of Mind Question Answering
di: Jin, Chuanyang, et al.
Pubblicazione: (2024)
di: Jin, Chuanyang, et al.
Pubblicazione: (2024)
Can GPT Redefine Medical Understanding? Evaluating GPT on Biomedical Machine Reading Comprehension
di: Vatsal, Shubham, et al.
Pubblicazione: (2024)
di: Vatsal, Shubham, et al.
Pubblicazione: (2024)
Can We Count on LLMs? The Fixed-Effect Fallacy and Claims of GPT-4 Capabilities
di: Ball, Thomas, et al.
Pubblicazione: (2024)
di: Ball, Thomas, et al.
Pubblicazione: (2024)
MediaMind: Revolutionizing Media Monitoring using Agentification
di: Gunduz, Ahmet, et al.
Pubblicazione: (2025)
di: Gunduz, Ahmet, et al.
Pubblicazione: (2025)
Mini Minds: Exploring Bebeshka and Zlata Baby Models
di: Proskurina, Irina, et al.
Pubblicazione: (2023)
di: Proskurina, Irina, et al.
Pubblicazione: (2023)
Benchmarking ChatGPT on Algorithmic Reasoning
di: McLeish, Sean, et al.
Pubblicazione: (2024)
di: McLeish, Sean, et al.
Pubblicazione: (2024)
Agent-ToM: Learning to Monitor Autonomous LLM Agents via Theory-of-Mind Reasoning
di: Ahmed, Nesreen K., et al.
Pubblicazione: (2026)
di: Ahmed, Nesreen K., et al.
Pubblicazione: (2026)
Low-Resource Languages Jailbreak GPT-4
di: Yong, Zheng-Xin, et al.
Pubblicazione: (2023)
di: Yong, Zheng-Xin, et al.
Pubblicazione: (2023)
Using Hallucinations to Bypass GPT4's Filter
di: Lemkin, Benjamin
Pubblicazione: (2024)
di: Lemkin, Benjamin
Pubblicazione: (2024)
LoRA Land: 310 Fine-tuned LLMs that Rival GPT-4, A Technical Report
di: Zhao, Justin, et al.
Pubblicazione: (2024)
di: Zhao, Justin, et al.
Pubblicazione: (2024)
MuMA-ToM: Multi-modal Multi-Agent Theory of Mind
di: Shi, Haojun, et al.
Pubblicazione: (2024)
di: Shi, Haojun, et al.
Pubblicazione: (2024)
Geometric-Averaged Preference Optimization for Soft Preference Labels
di: Furuta, Hiroki, et al.
Pubblicazione: (2024)
di: Furuta, Hiroki, et al.
Pubblicazione: (2024)
Can we trust the evaluation on ChatGPT?
di: Aiyappa, Rachith, et al.
Pubblicazione: (2023)
di: Aiyappa, Rachith, et al.
Pubblicazione: (2023)
NExT-GPT: Any-to-Any Multimodal LLM
di: Wu, Shengqiong, et al.
Pubblicazione: (2023)
di: Wu, Shengqiong, et al.
Pubblicazione: (2023)
Universal Neurons in GPT2 Language Models
di: Gurnee, Wes, et al.
Pubblicazione: (2024)
di: Gurnee, Wes, et al.
Pubblicazione: (2024)
HumanEval on Latest GPT Models -- 2024
di: Li, Daniel, et al.
Pubblicazione: (2024)
di: Li, Daniel, et al.
Pubblicazione: (2024)
Fairness of ChatGPT
di: Li, Yunqi, et al.
Pubblicazione: (2023)
di: Li, Yunqi, et al.
Pubblicazione: (2023)
Evaluating GPT's Capability in Identifying Stages of Cognitive Impairment from Electronic Health Data
di: Leng, Yu, et al.
Pubblicazione: (2025)
di: Leng, Yu, et al.
Pubblicazione: (2025)
Mind the Gap: A Review of Arabic Post-Training Datasets and Their Limitations
di: Alkhowaiter, Mohammed, et al.
Pubblicazione: (2025)
di: Alkhowaiter, Mohammed, et al.
Pubblicazione: (2025)
Mechanistic Interpretability of GPT-like Models on Summarization Tasks
di: Mishra, Anurag
Pubblicazione: (2025)
di: Mishra, Anurag
Pubblicazione: (2025)
If in a Crowdsourced Data Annotation Pipeline, a GPT-4
di: He, Zeyu, et al.
Pubblicazione: (2024)
di: He, Zeyu, et al.
Pubblicazione: (2024)
Meta-GPT: Decoding the Metasurface Genome with Generative Artificial Intelligence
di: Dang, David, et al.
Pubblicazione: (2025)
di: Dang, David, et al.
Pubblicazione: (2025)
How Prevalent is Gender Bias in ChatGPT? -- Exploring German and English ChatGPT Responses
di: Urchs, Stefanie, et al.
Pubblicazione: (2023)
di: Urchs, Stefanie, et al.
Pubblicazione: (2023)
MeDiSumQA: Patient-Oriented Question-Answer Generation from Discharge Letters
di: Dada, Amin, et al.
Pubblicazione: (2025)
di: Dada, Amin, et al.
Pubblicazione: (2025)
Two Minds Better Than One: Collaborative Reward Modeling for LLM Alignment
di: Zhang, Jiazheng, et al.
Pubblicazione: (2025)
di: Zhang, Jiazheng, et al.
Pubblicazione: (2025)
PrivacyMind: Large Language Models Can Be Contextual Privacy Protection Learners
di: Xiao, Yijia, et al.
Pubblicazione: (2023)
di: Xiao, Yijia, et al.
Pubblicazione: (2023)
Enhancing Antibiotic Stewardship using a Natural Language Approach for Better Feature Representation
di: Lee, Simon A., et al.
Pubblicazione: (2024)
di: Lee, Simon A., et al.
Pubblicazione: (2024)
Documenti analoghi
-
Testing the Depth of ChatGPT's Comprehension via Cross-Modal Tasks Based on ASCII-Art: GPT3.5's Abilities in Regard to Recognizing and Generating ASCII-Art Are Not Totally Lacking
di: Bayani, David
Pubblicazione: (2023) -
Automated Meta Prompt Engineering for Alignment with the Theory of Mind
di: Baughman, Aaron, et al.
Pubblicazione: (2025) -
Playing Along: Learning a Double-Agent Defender for Belief Steering via Theory of Mind
di: Xiao, Hanqi, et al.
Pubblicazione: (2026) -
A Notion of Complexity for Theory of Mind via Discrete World Models
di: Huang, X. Angelo, et al.
Pubblicazione: (2024) -
Single layer tiny Co$^4$ outpaces GPT-2 and GPT-BERT
di: Zain, Noor Ul, et al.
Pubblicazione: (2025)