Hacker News new | ask | show | jobs
by BoorishBears 23 days ago
"It'd be a game-changer to be able to have the model start replying with partial information streamed from the tool call, then seamlessly continue with additional information."

You can't continue a conversation if it hinges on information that takes inherently takes 30+ seconds to generate. Previously even models that "handle it" can only play for time.

You can try and simulate this by messing with model context mid-conversation but that breaks down reliability massively as the model loses track of what it's talking about.