|
|
|
|
|
by pyeri
29 days ago
|
|
The capacity conundrum here is very unique. Similar to pareto optimization principle: => 80% of engineering problems == budget models like gpt-mini can handle. => 15% of complex reasoning, vibe coding == frontier models like Opus, still sustainable. => 5% of 'really complex enterprise grade reasoning' == need massive tokens, impossible at scale unless pockets are deep. |
|