Cognitive Fatigue in Autoregressive Transformers: Formalization and Measurement
Fuente:
arXiv
Saved in:
| Main Authors: | Marwah, Riju, Garimella, Ritvik, Pallagani, Vishal, Jain, Atishay, Stewart, Michael, Sheth, Amit |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Chatsparent: An Interactive System for Detecting and Mitigating Cognitive Fatigue in LLMs
by: Marwah, Riju, et al.
Published: (2025)
by: Marwah, Riju, et al.
Published: (2025)
Neurosymbolic AI for Enhancing Instructability in Generative AI
by: Sheth, Amit, et al.
Published: (2024)
by: Sheth, Amit, et al.
Published: (2024)
Large Language Models for Mental Health Diagnostic Assessments: Exploring The Potential of Large Language Models for Assisting with Mental Health Diagnostic Assessments -- The Depression and Anxiety Case
by: Roy, Kaushik, et al.
Published: (2025)
by: Roy, Kaushik, et al.
Published: (2025)
Local Prompt Optimization
by: Jain, Yash, et al.
Published: (2025)
by: Jain, Yash, et al.
Published: (2025)
NeuroLit Navigator: A Neurosymbolic Approach to Scholarly Article Searches for Systematic Reviews
by: Khandelwal, Vedant, et al.
Published: (2025)
by: Khandelwal, Vedant, et al.
Published: (2025)
Dynamic Context Pruning for Efficient and Interpretable Autoregressive Transformers
by: Anagnostidis, Sotiris, et al.
Published: (2023)
by: Anagnostidis, Sotiris, et al.
Published: (2023)
On Mesa-Optimization in Autoregressively Trained Transformers: Emergence and Capability
by: Zheng, Chenyu, et al.
Published: (2024)
by: Zheng, Chenyu, et al.
Published: (2024)
Flexible Language Modeling in Continuous Space with Transformer-based Autoregressive Flows
by: Zhang, Ruixiang, et al.
Published: (2025)
by: Zhang, Ruixiang, et al.
Published: (2025)
Why Are Positional Encodings Nonessential for Deep Autoregressive Transformers? Revisiting a Petroglyph
by: Irie, Kazuki
Published: (2024)
by: Irie, Kazuki
Published: (2024)
Overview of Factify5WQA: Fact Verification through 5W Question-Answering
by: Suresh, Suryavardan, et al.
Published: (2024)
by: Suresh, Suryavardan, et al.
Published: (2024)
LeaPformer: Enabling Linear Transformers for Autoregressive and Simultaneous Tasks via Learned Proportions
by: Agostinelli, Victor, et al.
Published: (2024)
by: Agostinelli, Victor, et al.
Published: (2024)
PLANTS: A Novel Problem and Dataset for Summarization of Planning-Like (PL) Tasks
by: Pallagani, Vishal, et al.
Published: (2024)
by: Pallagani, Vishal, et al.
Published: (2024)
FourierNAT: A Fourier-Mixing-Based Non-Autoregressive Transformer for Parallel Sequence Generation
by: Kiruluta, Andrew, et al.
Published: (2025)
by: Kiruluta, Andrew, et al.
Published: (2025)
Geometric Organization of Cognitive States in Transformer Embedding Spaces
by: Zhao, Sophie
Published: (2025)
by: Zhao, Sophie
Published: (2025)
Position: The Turing-Completeness of Autoregressive Transformers Relies Heavily on Context Management
by: Cui, Guanyu, et al.
Published: (2026)
by: Cui, Guanyu, et al.
Published: (2026)
MENTIS: What Belief Changes Under Alignment? Measuring Multi-Scale Latent Torsion in Language Models
by: Saha, Partha Pratim, et al.
Published: (2026)
by: Saha, Partha Pratim, et al.
Published: (2026)
Universal Approximation of Visual Autoregressive Transformers
by: Chen, Yifang, et al.
Published: (2025)
by: Chen, Yifang, et al.
Published: (2025)
ZzzGPT: An Interactive GPT Approach to Enhance Sleep Quality
by: Khaokaew, Yonchanok, et al.
Published: (2023)
by: Khaokaew, Yonchanok, et al.
Published: (2023)
What Formal Languages Can Transformers Express? A Survey
by: Strobl, Lena, et al.
Published: (2023)
by: Strobl, Lena, et al.
Published: (2023)
Adapting Language Models via Token Translation
by: Feng, Zhili, et al.
Published: (2024)
by: Feng, Zhili, et al.
Published: (2024)
Cross-Platform Hate Speech Detection with Weakly Supervised Causal Disentanglement
by: Sheth, Paras, et al.
Published: (2024)
by: Sheth, Paras, et al.
Published: (2024)
Half the Nonlinearity Is Wasted: Measuring and Reallocating the Transformer's MLP Budget
by: Balogh, Peter
Published: (2026)
by: Balogh, Peter
Published: (2026)
Selective Attention: Enhancing Transformer through Principled Context Control
by: Zhang, Xuechen, et al.
Published: (2024)
by: Zhang, Xuechen, et al.
Published: (2024)
Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning
by: Ye, Jiacheng, et al.
Published: (2024)
by: Ye, Jiacheng, et al.
Published: (2024)
SALSA: Single-pass Autoregressive LLM Structured Classification
by: Berdichevsky, Ruslan, et al.
Published: (2025)
by: Berdichevsky, Ruslan, et al.
Published: (2025)
Projected Autoregression: Autoregressive Language Generation in Continuous State Space
by: Naparstek, Oshri
Published: (2026)
by: Naparstek, Oshri
Published: (2026)
Hierarchical Autoregressive Transformers: Combining Byte- and Word-Level Processing for Robust, Adaptable Language Models
by: Neitemeier, Pit, et al.
Published: (2025)
by: Neitemeier, Pit, et al.
Published: (2025)
LingVarBench: Benchmarking LLMs on Entity Recognitions and Linguistic Verbalization Patterns in Phone-Call Transcripts
by: Mohammadi, Seyedali, et al.
Published: (2025)
by: Mohammadi, Seyedali, et al.
Published: (2025)
Training Neural Networks as Recognizers of Formal Languages
by: Butoi, Alexandra, et al.
Published: (2024)
by: Butoi, Alexandra, et al.
Published: (2024)
Beyond Autoregression: Fast LLMs via Self-Distillation Through Time
by: Deschenaux, Justin, et al.
Published: (2024)
by: Deschenaux, Justin, et al.
Published: (2024)
Mask-Enhanced Autoregressive Prediction: Pay Less Attention to Learn More
by: Zhuang, Xialie, et al.
Published: (2025)
by: Zhuang, Xialie, et al.
Published: (2025)
Unmasking Hallucinations: A Causal Graph-Attention Perspective on Factual Reliability in Large Language Models
by: kurra, Sailesh kiran, et al.
Published: (2026)
by: kurra, Sailesh kiran, et al.
Published: (2026)
A Neurosymbolic Fast and Slow Architecture for Graph Coloring
by: Khandelwal, Vedant, et al.
Published: (2024)
by: Khandelwal, Vedant, et al.
Published: (2024)
Transformers Can Learn Connectivity in Some Graphs but Not Others
by: Roy, Amit, et al.
Published: (2025)
by: Roy, Amit, et al.
Published: (2025)
The Expressive Capacity of State Space Models: A Formal Language Perspective
by: Sarrof, Yash, et al.
Published: (2024)
by: Sarrof, Yash, et al.
Published: (2024)
Modeling Contextual Passage Utility for Multihop Question Answering
by: Jain, Akriti, et al.
Published: (2025)
by: Jain, Akriti, et al.
Published: (2025)
Knowing What's Missing: Assessing Information Sufficiency in Question Answering
by: Jain, Akriti, et al.
Published: (2025)
by: Jain, Akriti, et al.
Published: (2025)
Robust and Unbounded Length Generalization in Autoregressive Transformer-Based Text-to-Speech
by: Battenberg, Eric, et al.
Published: (2024)
by: Battenberg, Eric, et al.
Published: (2024)
Shared Latent Space by Both Languages in Non-Autoregressive Neural Machine Translation
by: Heo, DongNyeong, et al.
Published: (2023)
by: Heo, DongNyeong, et al.
Published: (2023)
Reviving Any-Subset Autoregressive Models with Principled Parallel Sampling and Speculative Decoding
by: Guo, Gabe, et al.
Published: (2025)
by: Guo, Gabe, et al.
Published: (2025)
Similar Items
-
Chatsparent: An Interactive System for Detecting and Mitigating Cognitive Fatigue in LLMs
by: Marwah, Riju, et al.
Published: (2025) -
Neurosymbolic AI for Enhancing Instructability in Generative AI
by: Sheth, Amit, et al.
Published: (2024) -
Large Language Models for Mental Health Diagnostic Assessments: Exploring The Potential of Large Language Models for Assisting with Mental Health Diagnostic Assessments -- The Depression and Anxiety Case
by: Roy, Kaushik, et al.
Published: (2025) -
Local Prompt Optimization
by: Jain, Yash, et al.
Published: (2025) -
NeuroLit Navigator: A Neurosymbolic Approach to Scholarly Article Searches for Systematic Reviews
by: Khandelwal, Vedant, et al.
Published: (2025)