Verbalizable Representations Operate as a Global Workspace in Language Models
A new research paper examines the internal dynamics of language models. The authors argue that verbalizable representations act as a shared workspace. This
A new research paper examines the internal dynamics of language models. The
authors argue that verbalizable representations act as a shared workspace. This
workspace allows different model components to exchange information. The concept
draws parallels to cognitive theories of global workspaces. Experiments
demonstrate that models can coordinate via these representations. Findings
suggest a pathway toward more interpretable AI behavior. The work contributes to
ongoing debates on model transparency. Future studies may explore how to harness
this mechanism for controlled generation.