|
|
|
|
|
by numpad0
11 days ago
|
|
Maybe it's LLM and human actor splitting work, but she looks too natural, too fast and too error-free to be LLM based since too long ago. Just their dictation -> thinking -> TTS loop is too fast to be LLM. Context awareness and character adherence is too perfect as well. |
|
Let alone a decent sized model on a couple 4090s or similar, pretty gud t/s to churn out tts/stt/response/control model actions and emotes.