| HN Mirror

Y	Hacker News new \| ask \| show \| jobs


	by HarHarVeryFunny 799 days ago
	Sure you can reason over a fixed "base set of logic", although there's another word for that - an expert system with a fixed set of rules, which IMO is really the right way to view an LLM. Still, what current LLMs are doing with their fixed rules is only a very limited form of reasoning since they just use a fixed N-steps of rule application to generate each word. People are looking to techniques such "group of experts" prompting to improve reasoning - step-wise generate multiple responses then evaluate them and proceed to next step.

1 comments

whimsicalism 799 days ago

if you zoom in enough, all thinking is an expert system with a fixed set of rules.

link

HarHarVeryFunny 799 days ago

That the basis of it, but in our brain the "inference engine" using those rules is a lot more than a fixed N-steps - there is thalamo-cortical looping, working memory of various durations, and maybe a bunch of other mechanisms such as analogical recall, resonance-based winner-takes-all processing, etc, etc.

Current LLMs have none of that - they are just the fixed set of rules, further limited by also having a fixed number of steps of rule application.

link

whimsicalism 799 days ago

Yes, LLMs don't have regression and that is a significant limitation - although they do have something close, by decoding one token they get to then have a thought loop. They just can't loop without outputting.

link

HarHarVeryFunny 799 days ago

Well, not exactly a loop. They get to "extend the thought", but there is zero continuity from one word to the next (LLM starts from scratch for each token generated).

The effect is as if you had multiple people playing a game where they each extend a sentence by taking turns adding a word to it, but there is zero continuity from one word to the next because each person is starting from scratch when it is their turn.

link

whimsicalism 799 days ago

> LLM starts from scratch for each token generated

What do you mean? They get to access their previous hidden states in the next greedy decode using attention, it is not simply starting from scratch. They can access exactly what they were thinking when they put out the previous word, not just reasoning from the word itself.

link

HarHarVeryFunny 799 days ago

There's the KV cache kept from one word to the next, but isn't that just an optimization ?

link

byteknight 799 days ago

Exactly. You can't reason with that you do not currently posses.

link

HarHarVeryFunny 799 days ago

Sure, but you (a person, not an LLM) can also reason about what you don't possess, which is one of our primary learning mechanisms - curiosity driven by lack of knowledge causing us to explore and acquire new knowledge by physical and/or mental exploration.

An LLM has no innate traits such as curiosity or boredom to trigger exploration, and anyways no online/incremental learning mechanism to benefit from it even if it did.

link

verdverm 799 days ago

How does scientific progress happen without reasoning about that which we do not know or understand?

link

byteknight 799 days ago

That's building upon current knowledge. That is a different application.

link