Y
Hacker News
new
|
ask
|
show
|
jobs
by
Sharlin
6 days ago
Back in 2024 we didn’t really yet have image models that used bigger language models as their text encoders, so prompt understanding was still rather hit and miss, especially with negations involved.