logoalt Hacker News

w4yaiyesterday at 9:59 PM2 repliesview on HN

> I haven't really seen evidence that any ai can reliably draw a pelican riding a bicycle

Please remember, we've started from there :

https://simonwillison.net/2024/Oct/25/pelicans-on-a-bicycle/

When it started, it was clear what LLM would stand out, its style, etc. Nowadays, the pelicans look similar, the difference is in details and sometimes hard to catch. Sure, the task is not completed perfectly, but that's not the point. It was supposed to be a benchmark to quickly benchmark a LLM against others.


Replies

techpressionyesterday at 10:31 PM

When is it ever hard to catch?

reaperduceryesterday at 10:09 PM

Sure, the task is not completed perfectly, but that's not the point.

Isn't it?

If the computer can't do it better than a human being, then what's the point?

Being wrong at scale is not better than being right.

show 4 replies