Hacker News new | ask | show | jobs
by nazgul17 25 days ago
That was DeepSeek OCR, not a DeepSeek lineage LLM. If the idea is introduced in LLMs, then you're right. But Gemini is not doing that, not yet. This is something I literally discussed with Claude last week, but took its word for it.
1 comments

Deep seek OCR is an LLM, just one trained/post-trained specifically for OCR.

Exact details of text to image compression ratios are of course extremely dependent on the model architecture, training data, training objectives, etc., so there's probably not too much justification for generalizing to all models