| HN Mirror

Y	Hacker News new \| ask \| show \| jobs


	by runarberg 914 days ago
	That is even less impressive. I was thinking—like normal linear models—it would be capable of interpolation.

2 comments

throwup238 914 days ago

It is. It doesn’t even need an existing language. You can define your own psuedo language in the prompt and have ChatGPT “execute” it (works best with 4 nonturbo).

You can even combine your pseudo language with natural language. See the OP’s custom GPT and the comments here: https://news.ycombinator.com/item?id=38594521

link

insanitybit 914 days ago

That looks totally different. In the case of the Python code it is literally executing it by calling out to CPython.

link

no_identd 914 days ago

…got a source for that claim?

link

throwup238 914 days ago

https://help.openai.com/en/articles/8302319-why-can-t-i-see-...

https://help.openai.com/en/articles/8437071-advanced-data-an...

link

labaron 913 days ago

I checked those links and didn’t see it mentioned that python code is actually executed. Could you quote the relevant part?

link

insanitybit 913 days ago

https://openai.com/blog/chatgpt-plugins#code-interpreter

>We provide our models with a working Python interpreter in a sandboxed, firewalled execution environment, along with some ephemeral disk space. Code run by our interpreter plugin is evaluated in a persistent session that is alive for the duration of a chat conversation (with an upper-bound timeout) and subsequent calls can build on top of each other. We support uploading files to the current conversation workspace and downloading the results of your work.

It really feels like I'm just googling for you, you had the feature name.

link

insanitybit 914 days ago

Why is it less impressive?

link

runarberg 914 days ago

I would say creating a model which is able to interpolate from training data in a way which produces an accurate output of a new input is a little impressive (if only as a neat party trick), however anybody can run a python interpreter on a server somewhere.

I’m sure there are use cases for this. But in the end it is only a simple feature added onto a—sometimes—marginally related service.

link

insanitybit 914 days ago

Hm, I don't think of it that way I guess. What the LLM is doing is generalizing a problem based on previous problems it has seen and then offloading the execution of that problem to a machine with some defined, specific semantics.

This is a lot more than a party trick. The model is able to describe the program it wants to execute and now it can accurately execute that - that it 'offloads' the work to a specialized program seems fine to me.

It's way more than a simple feature, this is enabling it to overcome one of the biggest limitations and criticisms of LLMs - it can answer questions it has never seen before.

link