On the Value of Human Ideas: What data poisoning research reveals about "autonomous" AI breakthroughs
Research into data poisoning reveals that even a small amount of poisoned data can significantly impact large AI models. For instance, just 250 poisoned documents, making up 0.00016% of training tokens, reliably implanted a backdoor in a 13B-parameter model. Similar success was seen in fine-tuning experiments, where 50–90 poisoned examples achieved over 80% attack success. This raises questions about intellectual property and credit in a future where AI synthesizes human ideas into breakthroughs.
为什么是这条This report uniquely highlights how a minuscule 0.00016% of poisoned data can reliably implant backdoors in large AI models, unlike previous discussions focusing on general data integrity.
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
收录当时偏移:UTC+02026年9月9日 05:00 UTC
- 收录
- 2026年9月9日 05:00
- 来源类型
- 开发者社区
讨论趋势
百分比基于采集到的讨论信号,不代表新增评论数或独立参与人数。曲线仅用于同一话题在不同时段的比较。
本站未收录正文。
前往源站阅读 →