|
|
|
|
|
by BoorishBears
23 days ago
|
|
"It'd be a game-changer to be able to have the model start replying with partial information streamed from the tool call, then seamlessly continue with additional information." You can't continue a conversation if it hinges on information that takes inherently takes 30+ seconds to generate. Previously even models that "handle it" can only play for time. You can try and simulate this by messing with model context mid-conversation but that breaks down reliability massively as the model loses track of what it's talking about. |
|