|
|
|
|
|
by _bobm
43 days ago
|
|
idk what is "minifying outputs" in the context of what we are talking about. Opencode is opensource, you can find out what it is doing. Last time I checked, OpenAI even send (in the response) the summary of the thinking part alreafy in markdown, so opencode has to remove the formatting to format it to their liking. > Many models now no longer return the entire chain-of-thought (to avoid distillation attacks). This is what they say: to avoid distillation attacks. And to some large extent this is true. I am saying there is a side- effect and this side- effect (depending on how tin-foilly you want to go) may be either a nice thing to have or it may be the "main reason" for all of this. The side effect is splicing the inference, brokering requests, and what not, which brings huge benefits at scale. This was my original point: openweights model to a sota model may be apples to oranges. So when will a local model catchup with its single cot run which is not even shaped properly: well never. It is apples to oranges. |
|