Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

[flagged]


It's just Simon Willison (the person you are replying to) who always makes a pelican, as his personal flippant benchmark. It's not that deep.


No benchmark will be perfect, especially if it's public but it's a fun experiment to visually see how these models get better and better.


Why is it so wrong?


Thanks for the "scientific air" remark, that gave me a genuine LOL.


"The difference between screwing around and science is writing it down" -- Adam Savage




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: