The AI firm Anthropic has developed a technique that has given it the clearest glimpse yet at what’s really going on inside large language models as they answer questions or carry out tasks. What they found ranges from the mundane to the unnerving.
Anthropic found a hidden space where Claude puzzles over concepts
Originally reported by MIT Tech Review AI
0 comments
No comments yet. Be the first to start the discussion.