Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Qureshi, Rizwan, Sapkota, Ranjan, Shah, Abbas, Muneer, Amgad, Zafar, Anas, Vayani, Ashmal, Shoman, Maged, Eldaly, Abdelrahman B. M., Zhang, Kai, Sadak, Ferhat, Raza, Shaina, Fan, Xinqi, Shwartz-Ziv, Ravid, Yan, Hong, Jain, Vinjia, Chadha, Aman, Karkee, Manoj, Wu, Jia, Mirjalili, Seyedali
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866908447065243648
author Qureshi, Rizwan
Sapkota, Ranjan
Shah, Abbas
Muneer, Amgad
Zafar, Anas
Vayani, Ashmal
Shoman, Maged
Eldaly, Abdelrahman B. M.
Zhang, Kai
Sadak, Ferhat
Raza, Shaina
Fan, Xinqi
Shwartz-Ziv, Ravid
Yan, Hong
Jain, Vinjia
Chadha, Aman
Karkee, Manoj
Wu, Jia
Mirjalili, Seyedali
author_facet Qureshi, Rizwan
Sapkota, Ranjan
Shah, Abbas
Muneer, Amgad
Zafar, Anas
Vayani, Ashmal
Shoman, Maged
Eldaly, Abdelrahman B. M.
Zhang, Kai
Sadak, Ferhat
Raza, Shaina
Fan, Xinqi
Shwartz-Ziv, Ravid
Yan, Hong
Jain, Vinjia
Chadha, Aman
Karkee, Manoj
Wu, Jia
Mirjalili, Seyedali
contents Can machines truly think, reason and act in domains like humans? This enduring question continues to shape the pursuit of Artificial General Intelligence (AGI). Despite the growing capabilities of models such as GPT-4.5, DeepSeek, Claude 3.5 Sonnet, Phi-4, and Grok 3, which exhibit multimodal fluency and partial reasoning, these systems remain fundamentally limited by their reliance on token-level prediction and lack of grounded agency. This paper offers a cross-disciplinary synthesis of AGI development, spanning artificial intelligence, cognitive neuroscience, psychology, generative models, and agent-based systems. We analyze the architectural and cognitive foundations of general intelligence, highlighting the role of modular reasoning, persistent memory, and multi-agent coordination. In particular, we emphasize the rise of Agentic RAG frameworks that combine retrieval, planning, and dynamic tool use to enable more adaptive behavior. We discuss generalization strategies, including information compression, test-time adaptation, and training-free methods, as critical pathways toward flexible, domain-agnostic intelligence. Vision-Language Models (VLMs) are reexamined not just as perception modules but as evolving interfaces for embodied understanding and collaborative task completion. We also argue that true intelligence arises not from scale alone but from the integration of memory and reasoning: an orchestration of modular, interactive, and self-improving components where compression enables adaptive behavior. Drawing on advances in neurosymbolic systems, reinforcement learning, and cognitive scaffolding, we explore how recent architectures begin to bridge the gap between statistical learning and goal-directed cognition. Finally, we identify key scientific, technical, and ethical challenges on the path to AGI.
format Preprint
id arxiv_https___arxiv_org_abs_2507_00951
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact
Qureshi, Rizwan
Sapkota, Ranjan
Shah, Abbas
Muneer, Amgad
Zafar, Anas
Vayani, Ashmal
Shoman, Maged
Eldaly, Abdelrahman B. M.
Zhang, Kai
Sadak, Ferhat
Raza, Shaina
Fan, Xinqi
Shwartz-Ziv, Ravid
Yan, Hong
Jain, Vinjia
Chadha, Aman
Karkee, Manoj
Wu, Jia
Mirjalili, Seyedali
Artificial Intelligence
Can machines truly think, reason and act in domains like humans? This enduring question continues to shape the pursuit of Artificial General Intelligence (AGI). Despite the growing capabilities of models such as GPT-4.5, DeepSeek, Claude 3.5 Sonnet, Phi-4, and Grok 3, which exhibit multimodal fluency and partial reasoning, these systems remain fundamentally limited by their reliance on token-level prediction and lack of grounded agency. This paper offers a cross-disciplinary synthesis of AGI development, spanning artificial intelligence, cognitive neuroscience, psychology, generative models, and agent-based systems. We analyze the architectural and cognitive foundations of general intelligence, highlighting the role of modular reasoning, persistent memory, and multi-agent coordination. In particular, we emphasize the rise of Agentic RAG frameworks that combine retrieval, planning, and dynamic tool use to enable more adaptive behavior. We discuss generalization strategies, including information compression, test-time adaptation, and training-free methods, as critical pathways toward flexible, domain-agnostic intelligence. Vision-Language Models (VLMs) are reexamined not just as perception modules but as evolving interfaces for embodied understanding and collaborative task completion. We also argue that true intelligence arises not from scale alone but from the integration of memory and reasoning: an orchestration of modular, interactive, and self-improving components where compression enables adaptive behavior. Drawing on advances in neurosymbolic systems, reinforcement learning, and cognitive scaffolding, we explore how recent architectures begin to bridge the gap between statistical learning and goal-directed cognition. Finally, we identify key scientific, technical, and ethical challenges on the path to AGI.
title Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact
topic Artificial Intelligence
url https://arxiv.org/abs/2507.00951