Hacker News new | ask | show | jobs
by jscott817 3 days ago
Can we assume that model performance at 90% of the 256k limit != 90% of 1M token limit?

Is this the exact same model just with less VRAM allocated for context window?

1 comments

No, you can't assume it. You can trust them as they made that claim outright. Or you can choose to not trust them, I guess