Victoria Graf, Hannaneh Hajishirzi, Noah A. Smith, David Kohlbrenner, Kyle Lo
Featured July 18, 2026
AI-generated analysis — This is SciGrove's AI interpretation of the paper, not peer-reviewed content. Always refer to the original paper.
Smart computers learning from the internet can be tricked by bad information hidden in website comments, and this paper shows how much of that bad info actually makes it into their brains, making them say wrong things.
The paper checks how bad information put into website comments can get past all the filters and end up teaching big computer brains.
They found that even a little bit of bad information can make the computer brains start saying wrong things, like preferring one car brand over another.
This breakdown was generated by SciGrove. Get the same analysis — intuition, storyboard, peer review, a runnable prototype and a glossary — on any paper you upload or paste a DOI for.