Hacker News new | ask | show | jobs
by nneonneo 30 days ago
Worth pointing out - modern multimodal LLMs (properly called VLMs, etc.) can easily take pictures as input and describe them in text. In fact, the CLIP model - one of the predecessors of modern VLMs - is entirely designed around being able to caption images with text.

That said - requiring students to hand-write answers is reasonably effective. It's a lot more boring to hand-copy text out of an LLM answer than write it yourself, and it makes the "cheating" significantly more visceral.

1 comments

I considered that... and in the "worst comes to worse" scenario where they decide to use an LLM and re-write everything by hand, I'm still hoping they'll learn SOMETHING during the copying process... kind of like the old time "copy books" that were once used.