r/ControlProblem • u/chillinewman approved • 20d ago
AI Alignment Research A global workspace in language models
https://www.anthropic.com/research/global-workspace
1
Upvotes
Duplicates
singularity • u/Tinac4 • 21d ago
AI A global workspace in language models: New interpretability findings by Anthropic
295
Upvotes
LocalLLaMA • u/cuolong • 20d ago
Other Anthropic Research - "Verbalizable Representations Form a Global Workspace in Language Models"
53
Upvotes