Are machines intelligent? It's a question at the forefront of AI development, prompting a deep dive into the fundamental differences between today's leading AI systems and the biological marvel that is the human brain. As highlighted in Microsoft Research's The Shape of Things to Come series, this isn't just academic; it shapes the future of AI.
At the heart of the discussion are transformer-based large language models (LLMs), the engines behind much of today's generative AI. While incredibly powerful at processing vast amounts of data and identifying complex patterns, their architecture differs profoundly from the human brain architecture. Experts like Nicolò Fusi from Microsoft Research and Subutai Ahmad from Numenta, formerly of Microsoft Research, are exploring these distinctions.
The Transformer's Strengths and Weaknesses
Transformers, with their attention and feedforward layers, are adept at understanding context within a given sequence, like a chat history or a prompt. This allows them to generate coherent and contextually relevant text, a feat that has revolutionized natural language processing.
However, this approach is fundamentally different from how humans learn. Consider the simple act of navigating basement stairs. When one step is unexpectedly altered, the brain doesn't require a full retraining session. Instead, it performs rapid, continuous updates, often triggered by subtle sensory-motor feedback and neuromodulators.
This continuous, adaptive learning is a hallmark of biological intelligence. The brain constantly models the world, making granular adjustments without conscious effort and without forgetting everything it already knows. This contrasts sharply with current AI models, which often require massive retraining for even minor updates, a process that can be inefficient and prone to catastrophic forgetting.
