Strong Memory, Weak Control: An Empirical Study of Executive Functioning in LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | de Langis, Karin, Park, Jong Inn, Hu, Bin, Le, Khanh Chi, Schramm, Andreas, Mensink, Michael C., Elfenbein, Andrew, Kang, Dongyeop |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
How LLMs Comprehend Temporal Meaning in Narratives: A Case Study in Cognitive Evaluation of LLMs
por: de Langis, Karin, et al.
Publicado: (2025)
por: de Langis, Karin, et al.
Publicado: (2025)
Mary, the Cheeseburger-Eating Vegetarian: Do LLMs Recognize Incoherence in Narratives?
por: de Langis, Karin, et al.
Publicado: (2025)
por: de Langis, Karin, et al.
Publicado: (2025)
Tracing How Annotators Think: Augmenting Preference Judgments with Reading Processes
por: de Langis, Karin, et al.
Publicado: (2025)
por: de Langis, Karin, et al.
Publicado: (2025)
Dynamic Multi-Reward Weighting for Multi-Style Controllable Generation
por: de Langis, Karin, et al.
Publicado: (2024)
por: de Langis, Karin, et al.
Publicado: (2024)
SelectLLM: Can LLMs Select Important Instructions to Annotate?
por: Parkar, Ritik Sachin, et al.
Publicado: (2024)
por: Parkar, Ritik Sachin, et al.
Publicado: (2024)
Stealing Creator's Workflow: A Creator-Inspired Agentic Framework with Iterative Feedback Loop for Improved Scientific Short-form Generation
por: Park, Jong Inn, et al.
Publicado: (2025)
por: Park, Jong Inn, et al.
Publicado: (2025)
Effects of Varying LLM Access on Essay Writing Behavior
por: Christenson, Julia, et al.
Publicado: (2026)
por: Christenson, Julia, et al.
Publicado: (2026)
Benchmarking Cognitive Biases in Large Language Models as Evaluators
por: Koo, Ryan, et al.
Publicado: (2023)
por: Koo, Ryan, et al.
Publicado: (2023)
The Amazing Agent Race: Strong Tool Users, Weak Navigators
por: Kim, Zae Myung, et al.
Publicado: (2026)
por: Kim, Zae Myung, et al.
Publicado: (2026)
LawFlow: Collecting and Simulating Lawyers' Thought Processes on Business Formation Case Studies
por: Das, Debarati, et al.
Publicado: (2025)
por: Das, Debarati, et al.
Publicado: (2025)
ScholaWrite: A Dataset of End-to-End Scholarly Writing Process
por: Le, Khanh Chi, et al.
Publicado: (2025)
por: Le, Khanh Chi, et al.
Publicado: (2025)
Confidence Calibration and Rationalization for LLMs via Multi-Agent Deliberation
por: Yang, Ruixin, et al.
Publicado: (2024)
por: Yang, Ruixin, et al.
Publicado: (2024)
ReduceFormer: Attention with Tensor Reduction by Summation
por: Yang, John, et al.
Publicado: (2024)
por: Yang, John, et al.
Publicado: (2024)
Une formation commerciale pour l'avenir
por: Karin Schramm
Publicado: (1980)
por: Karin Schramm
Publicado: (1980)
Business education for the future
por: Karin Schramm
Publicado: (1980)
por: Karin Schramm
Publicado: (1980)
Charge Scrambling in Strong-to-Weak Spontaneous Symmetry Breaking
por: Lee, Jong Yeon
Publicado: (2026)
por: Lee, Jong Yeon
Publicado: (2026)
L’impact de la COVID‐19 sur l’expérience client en magasin
por: Samantha Langis, et al.
Publicado: (2026)
por: Samantha Langis, et al.
Publicado: (2026)
Dynamic Disulfide Chemistry for Functional Polymers: Self‐Healing, Vitrimer Behavior, and Biochemical/Electronic Applications
por: Sebin Jin, et al.
Publicado: (2025)
por: Sebin Jin, et al.
Publicado: (2025)
Self‐regulation, corruption, and competitiveness in extractive industries: Making transparency pay
por: Shirley Tang, et al.
Publicado: (2025)
por: Shirley Tang, et al.
Publicado: (2025)
From Preparation to Performance: Conscientiousness Predicts Negotiation Planning and Value Claiming
por: Daisung Jang, et al.
Publicado: (2025)
por: Daisung Jang, et al.
Publicado: (2025)
Representations Shape Weak-to-Strong Generalization: Theoretical Insights and Empirical Predictions
por: Xue, Yihao, et al.
Publicado: (2025)
por: Xue, Yihao, et al.
Publicado: (2025)
Under the Surface: Tracking the Artifactuality of LLM-Generated Data
por: Das, Debarati, et al.
Publicado: (2024)
por: Das, Debarati, et al.
Publicado: (2024)
Light‐Dependent Circadian Rhythm Governs O‐GlcNAc Cycling to Influence Cognitive Function in Adult Zebrafish
por: Jiwon Park, et al.
Publicado: (2024)
por: Jiwon Park, et al.
Publicado: (2024)
Strong Forms of Weakly e-continuous Functions
por: Ayhan, B. S.
Publicado: (2024)
por: Ayhan, B. S.
Publicado: (2024)
Improving Diffusion Generalization with Weak-to-Strong Segmented Guidance
por: Yuan, Liangyu, et al.
Publicado: (2026)
por: Yuan, Liangyu, et al.
Publicado: (2026)
Learning a High-quality Robotic Wiping Policy Using Systematic Reward Analysis and Visual-Language Model Based Curriculum
por: Liu, Yihong, et al.
Publicado: (2025)
por: Liu, Yihong, et al.
Publicado: (2025)
Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL
por: Zhai, Skylar, et al.
Publicado: (2026)
por: Zhai, Skylar, et al.
Publicado: (2026)
Cognitive Workspace: Active Memory Management for LLMs -- An Empirical Study of Functional Infinite Context
por: An, Tao
Publicado: (2025)
por: An, Tao
Publicado: (2025)
Error Threshold of SYK Codes from Strong-to-Weak Parity Symmetry Breaking
por: Kim, Jaewon, et al.
Publicado: (2024)
por: Kim, Jaewon, et al.
Publicado: (2024)
An Empirical Study on Strong-Weak Model Collaboration for Repo-level Code Generation
por: Gandhi, Shubham, et al.
Publicado: (2025)
por: Gandhi, Shubham, et al.
Publicado: (2025)
A new twist on modular links from an old perspective
por: Le, Khanh
Publicado: (2023)
por: Le, Khanh
Publicado: (2023)
Synthesizing Text-to-SQL Data from Weak and Strong LLMs
por: Yang, Jiaxi, et al.
Publicado: (2024)
por: Yang, Jiaxi, et al.
Publicado: (2024)
Unlearning Backdoor Attacks for LLMs with Weak-to-Strong Knowledge Distillation
por: Zhao, Shuai, et al.
Publicado: (2024)
por: Zhao, Shuai, et al.
Publicado: (2024)
Compositional Symbolic Execution for the Next 700 Memory Models (Extended Version)
por: Lööw, Andreas, et al.
Publicado: (2025)
por: Lööw, Andreas, et al.
Publicado: (2025)
Strong Approximations for Empirical Processes Indexed by Lipschitz Functions
por: Cattaneo, Matias D., et al.
Publicado: (2024)
por: Cattaneo, Matias D., et al.
Publicado: (2024)
Toward Evaluative Thinking: Meta Policy Optimization with Evolving Reward Models
por: Kim, Zae Myung, et al.
Publicado: (2025)
por: Kim, Zae Myung, et al.
Publicado: (2025)
Engineering Poly(Lactic Acid)/Cellulose Nanocrystal Composites: A Comparative Review of Preparation Strategies and their Influence on Structure‐Property Relationships
por: Jimin Ryoo, et al.
Publicado: (2025)
por: Jimin Ryoo, et al.
Publicado: (2025)
Mixture of Weak & Strong Experts on Graphs
por: Zeng, Hanqing, et al.
Publicado: (2023)
por: Zeng, Hanqing, et al.
Publicado: (2023)
How Execution Features Relate to Failures: An Empirical Study and Diagnosis Approach
por: Smytzek, Marius, et al.
Publicado: (2025)
por: Smytzek, Marius, et al.
Publicado: (2025)
Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher
por: Uzunoglu, Arda, et al.
Publicado: (2026)
por: Uzunoglu, Arda, et al.
Publicado: (2026)
Ejemplares similares
-
How LLMs Comprehend Temporal Meaning in Narratives: A Case Study in Cognitive Evaluation of LLMs
por: de Langis, Karin, et al.
Publicado: (2025) -
Mary, the Cheeseburger-Eating Vegetarian: Do LLMs Recognize Incoherence in Narratives?
por: de Langis, Karin, et al.
Publicado: (2025) -
Tracing How Annotators Think: Augmenting Preference Judgments with Reading Processes
por: de Langis, Karin, et al.
Publicado: (2025) -
Dynamic Multi-Reward Weighting for Multi-Style Controllable Generation
por: de Langis, Karin, et al.
Publicado: (2024) -
SelectLLM: Can LLMs Select Important Instructions to Annotate?
por: Parkar, Ritik Sachin, et al.
Publicado: (2024)