Do Zombies Understand? A Choose-Your-Own-Adventure Exploration of Machine Cognition
Fuente:
arXiv
Guardado en:
| Autores principales: | Goldstein, Ariel, Stanovsky, Gabriel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Choose Your Own Adventure: Interactive E-Books to Improve Word Knowledge and Comprehension Skills
por: Day, Stephanie, et al.
Publicado: (2024)
por: Day, Stephanie, et al.
Publicado: (2024)
Looking Beyond The Top-1: Transformers Determine Top Tokens In Order
por: Lioubashevski, Daria, et al.
Publicado: (2024)
por: Lioubashevski, Daria, et al.
Publicado: (2024)
Comparing Humans and Models on a Similar Scale: Towards Cognitive Gender Bias Evaluation in Coreference Resolution
por: Lior, Gili, et al.
Publicado: (2023)
por: Lior, Gili, et al.
Publicado: (2023)
A Nurse is Blue and Elephant is Rugby: Cross Domain Alignment in Large Language Models Reveal Human-like Patterns
por: Yehudai, Asaf, et al.
Publicado: (2024)
por: Yehudai, Asaf, et al.
Publicado: (2024)
SAUCE: Synchronous and Asynchronous User-Customizable Environment for Multi-Agent LLM Interaction
por: Neuberger, Shlomo, et al.
Publicado: (2024)
por: Neuberger, Shlomo, et al.
Publicado: (2024)
Choose Your Own Adventure: Non-Linear AI-Assisted Programming with EvoGraph
por: Exarhakos, Vassilios, et al.
Publicado: (2026)
por: Exarhakos, Vassilios, et al.
Publicado: (2026)
Choose Your Own Research Adventure: An Asynchronous Tutorial to Address "Research as Inquiry"
por: Stacy Brinkman, et al.
Publicado: (2024)
por: Stacy Brinkman, et al.
Publicado: (2024)
The State and Fate of Summarization Datasets: A Survey
por: Dahan, Noam, et al.
Publicado: (2024)
por: Dahan, Noam, et al.
Publicado: (2024)
Planted in Pretraining, Swayed by Finetuning: A Case Study on the Origins of Cognitive Biases in LLMs
por: Itzhak, Itay, et al.
Publicado: (2025)
por: Itzhak, Itay, et al.
Publicado: (2025)
Applying Intrinsic Debiasing on Downstream Tasks: Challenges and Considerations for Machine Translation
por: Iluz, Bar, et al.
Publicado: (2024)
por: Iluz, Bar, et al.
Publicado: (2024)
In-Context Learning on a Budget: A Case Study in Token Classification
por: Berger, Uri, et al.
Publicado: (2024)
por: Berger, Uri, et al.
Publicado: (2024)
Can (A)I Change Your Mind?
por: Havin, Miriam, et al.
Publicado: (2025)
por: Havin, Miriam, et al.
Publicado: (2025)
Leveraging Collection-Wide Similarities for Unsupervised Document Structure Extraction
por: Lior, Gili, et al.
Publicado: (2024)
por: Lior, Gili, et al.
Publicado: (2024)
Leveraging Digitized Newspapers to Collect Summarization Data in Low-Resource Languages
por: Dahan, Noam, et al.
Publicado: (2025)
por: Dahan, Noam, et al.
Publicado: (2025)
Beyond Memorization: Distinguishing between Reductive and Epistemic Reasoning in LLMs using Classic Logic Puzzles
por: Gabay, Adi, et al.
Publicado: (2026)
por: Gabay, Adi, et al.
Publicado: (2026)
BYOL: Bring Your Own Language Into LLMs
por: Zamir, Syed Waqas, et al.
Publicado: (2026)
por: Zamir, Syed Waqas, et al.
Publicado: (2026)
From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs
por: Itzhak, Itay, et al.
Publicado: (2026)
por: Itzhak, Itay, et al.
Publicado: (2026)
Bootstrap Your Own Context Length
por: Wang, Liang, et al.
Publicado: (2024)
por: Wang, Liang, et al.
Publicado: (2024)
PromptSuite: A Task-Agnostic Framework for Multi-Prompt Generation
por: Habba, Eliya, et al.
Publicado: (2025)
por: Habba, Eliya, et al.
Publicado: (2025)
Comparing the Framing Effect in Humans and LLMs on Naturally Occurring Texts
por: Lior, Gili, et al.
Publicado: (2025)
por: Lior, Gili, et al.
Publicado: (2025)
Indications of Belief-Guided Agency and Meta-Cognitive Monitoring in Large Language Models
por: Yalon, Noam Steinmetz, et al.
Publicado: (2026)
por: Yalon, Noam Steinmetz, et al.
Publicado: (2026)
Feeding LLM Annotations to BERT Classifiers at Your Own Risk
por: Lu, Yucheng, et al.
Publicado: (2025)
por: Lu, Yucheng, et al.
Publicado: (2025)
How to Choose How to Choose Your Chatbot: A Massively Multi-System MultiReference Data Set for Dialog Metric Evaluation
por: Khayrallah, Huda, et al.
Publicado: (2023)
por: Khayrallah, Huda, et al.
Publicado: (2023)
Bring Your Own Knowledge: A Survey of Methods for LLM Knowledge Expansion
por: Wang, Mingyang, et al.
Publicado: (2025)
por: Wang, Mingyang, et al.
Publicado: (2025)
Surveying the Landscape of Image Captioning Evaluation: A Comprehensive Taxonomy, Trends and Metrics Analysis
por: Berger, Uri, et al.
Publicado: (2024)
por: Berger, Uri, et al.
Publicado: (2024)
Time to Talk: LLM Agents for Asynchronous Group Communication in Mafia Games
por: Eckhaus, Niv, et al.
Publicado: (2025)
por: Eckhaus, Niv, et al.
Publicado: (2025)
ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments
por: Lior, Gili, et al.
Publicado: (2025)
por: Lior, Gili, et al.
Publicado: (2025)
Improving Image Captioning by Mimicking Human Reformulation Feedback at Inference-time
por: Berger, Uri, et al.
Publicado: (2025)
por: Berger, Uri, et al.
Publicado: (2025)
Do LLMs Benefit From Their Own Words?
por: Huang, Jenny Y., et al.
Publicado: (2026)
por: Huang, Jenny Y., et al.
Publicado: (2026)
Positional Cognitive Specialization: Where Do LLMs Learn To Comprehend and Speak Your Language?
por: Salim, Luis Frentzen, et al.
Publicado: (2026)
por: Salim, Luis Frentzen, et al.
Publicado: (2026)
Views Are My Own, but Also Yours: Benchmarking Theory of Mind Using Common Ground
por: Soubki, Adil, et al.
Publicado: (2024)
por: Soubki, Adil, et al.
Publicado: (2024)
Read Your Own Mind: Reasoning Helps Surface Self-Confidence Signals in LLMs
por: Podolak, Jakub, et al.
Publicado: (2025)
por: Podolak, Jakub, et al.
Publicado: (2025)
More Documents, Same Length: Isolating the Challenge of Multiple Documents in RAG
por: Levy, Shahar, et al.
Publicado: (2025)
por: Levy, Shahar, et al.
Publicado: (2025)
BYOM: Building Your Own Multi-Task Model For Free
por: Jiang, Weisen, et al.
Publicado: (2023)
por: Jiang, Weisen, et al.
Publicado: (2023)
Can LLMs Help Uncover Insights about LLMs? A Large-Scale, Evolving Literature Analysis of Frontier LLMs
por: Park, Jungsoo, et al.
Publicado: (2025)
por: Park, Jungsoo, et al.
Publicado: (2025)
Distributional reasoning in LLMs: Parallel reasoning processes in multi-hop reasoning
por: Shalev, Yuval, et al.
Publicado: (2024)
por: Shalev, Yuval, et al.
Publicado: (2024)
SEAM: A Stochastic Benchmark for Multi-Document Tasks
por: Lior, Gili, et al.
Publicado: (2024)
por: Lior, Gili, et al.
Publicado: (2024)
State of What Art? A Call for Multi-Prompt LLM Evaluation
por: Mizrahi, Moran, et al.
Publicado: (2023)
por: Mizrahi, Moran, et al.
Publicado: (2023)
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models
por: Yu, Longhui, et al.
Publicado: (2023)
por: Yu, Longhui, et al.
Publicado: (2023)
Taming "Zombie'' Agents: A Markov State-Aware Framework for Resilient Multi-Agent Evolution
por: Zhang, Taolin, et al.
Publicado: (2026)
por: Zhang, Taolin, et al.
Publicado: (2026)
Ejemplares similares
-
Choose Your Own Adventure: Interactive E-Books to Improve Word Knowledge and Comprehension Skills
por: Day, Stephanie, et al.
Publicado: (2024) -
Looking Beyond The Top-1: Transformers Determine Top Tokens In Order
por: Lioubashevski, Daria, et al.
Publicado: (2024) -
Comparing Humans and Models on a Similar Scale: Towards Cognitive Gender Bias Evaluation in Coreference Resolution
por: Lior, Gili, et al.
Publicado: (2023) -
A Nurse is Blue and Elephant is Rugby: Cross Domain Alignment in Large Language Models Reveal Human-like Patterns
por: Yehudai, Asaf, et al.
Publicado: (2024) -
SAUCE: Synchronous and Asynchronous User-Customizable Environment for Multi-Agent LLM Interaction
por: Neuberger, Shlomo, et al.
Publicado: (2024)