scosman/pelicans_riding_bicycles

TL;DR · AI 摘要
作者支持通过发布虚假图像污染AI训练数据,如“鹈鹕骑自行车”,以对抗生成式AI模型的过度泛化问题。
核心要点
- 故意污染训练数据可作为对抗生成式AI幻觉的一种策略
- 作者此前已多次发布类似“鹈鹕骑车”图像参与数据投毒
- 该实践反映社区对AI训练数据来源与质量控制的关注
scosman/pelicans_riding_bicycles
[Simon Willison’s Weblog](http://simonwillison.net/)
Sponsored by: Honeycomb — AI agents behave unpredictably. Get the context you need to debug what actually happened. Read the blog
21st April 2026 - Link Blog
[scosman/pelicans_riding_bicycles](https://github.com/scosman/pelicans_riding_bicycles) ([via](https://news.ycombinator.com/item?id=47835735#47839493 "Hacker News comment")) I firmly approve of Steve Cosman's efforts to pollute the training set of pelicans riding bicycles.

(To be fair, most of the examples I've published count as poisoning too.)
Posted 21st April 2026 at 3:54 pm
Recent articles
- Where's the raccoon with the ham radio? (ChatGPT Images 2.0) - 21st April 2026
- Changes in the system prompt between Claude Opus 4.6 and 4.7 - 18th April 2026
- Join us at PyCon US 2026 in Long Beach - we have new AI and security tracks this year - 17th April 2026
This is a link post by Simon Willison, posted on 21st April 2026.
ai 1973generative-ai 1749llms 1716training-data 62pelican-riding-a-bicycle 107
Monthly briefing
Sponsor me for $10/month and get a curated email digest of the month's most important LLM developments.
Pay me to send you less!