Research
Researcher finds no evidence AI labs optimize for pelican-bicycle prompts
Dylan Castillo tested 7 major AI models on 48 animal-vehicle combinations and found no systematic improvement on pelican-bicycle prompts, contradicting speculation about deliberate training optimization.
1 min read
SourceSimon Willison
Dylan Castillo conducted a systematic benchmark of seven major AI models to test whether AI labs have been deliberately optimizing their models to excel at drawing pelicans riding bicycles, a whimsical prompt that has circulated in AI circles as an informal test of model capability.
Castillo's meth...
Sign in to read the full analysis
Free account. Full analysis on LLM unit economics, plus the weekly Cost-of-Inference column.
Try it on your own context
You just read the writeup. Now run the thing. Paste a doc or some verbose tool output and watch it shrink — free, no signup.
2,912/12,000 chars
Compressed
Compressed text will appear here…
Method & sources
- Source type
- Primary publication (lab/vendor blog) — our analysis + implication
- Source link
- Simon Willison
- Published
- UTC
- Byline
- By the gotcontext.ai team (editorial standards)
- Correction?
- corrections@gotcontext.ai