|
|
|
|
|
by nikcub
13 days ago
|
|
This is a winner IMO. Lots of cost pressure on token spend atm within enterprises and tasks that don't require Opus / Codex class models. These companies have hopefully captured all of their traces and now have enough to fine-tune an open model and host themselves. Inkling feels like the right base - not obsessed with benchmaxxing on coding but rather being adaptable to the task required For tasks like GTM, support, content writing etc. seeing 80%+ savings |
|
Can chime in on the support use case specifically: GPT OSS performs really well here and has been somewhat of a benchmark with our customers, limited testing [0] against Inkling reveals basically identical performance, but with a significant cost increase at scale.
I'd say that for real-world tasks that aren't coding most companies don't see value by being on the latest and greatest model.
[0] https://valiopt.com/blog/inkling-model-customer-support-revi...