Hacker News new | ask | show | jobs
by lucrbvi 24 days ago
Anthropic theorize that middle layers in an LLM is a "J-Space" used to "think" about the future answer or about abstract concepts.

Their method is used to identify which tokens can appears in which layers of the model.