Y
Hacker News
new
|
ask
|
show
|
jobs
by
lucrbvi
24 days ago
Anthropic theorize that middle layers in an LLM is a "J-Space" used to "think" about the future answer or about abstract concepts.
Their method is used to identify which tokens can appears in which layers of the model.