Are AI labs pelicanmaxxing? – Dylan Castillo
- For the past few years, Simon Willison has tested every major LLM release with the same prompt: “Generate an SVG of a pelican riding a bicycle”.
- What began as a tongue-in-cheek benchmark has become one of the most famous informal benchmarks in AI.
- Simon’s pelican-on-a-bicycle results are often among the most upvoted comments on Hacker News threads announcing new releases from AI labs.
Unverified
- For the past few years, Simon Willison has tested every major LLM release with the same prompt: “Generate an SVG of a pelican riding a bicycle”.
- What began as a tongue-in-cheek benchmark has become one of the most famous informal benchmarks in AI.
- Simon’s pelican-on-a-bicycle results are often among the most upvoted comments on Hacker News threads announcing new releases from AI labs.
Sources: Dylancastillo