<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[Someone should make a community to freely distribute examples of data poisoning people can randomly put in their social media posts/images to sabotage AI]]></title><description><![CDATA[<em>This post did not contain any content.</em>]]></description><link>https://forum.ieu.app/topic/29769bdc-0ff3-4225-80f8-0db0275d6e12/someone-should-make-a-community-to-freely-distribute-examples-of-data-poisoning-people-can-randomly-put-in-their-social-media-posts-images-to-sabotage-ai</link><generator>RSS for Node</generator><lastBuildDate>Sun, 06 Sep 2026 07:32:41 GMT</lastBuildDate><atom:link href="https://forum.ieu.app/topic/29769bdc-0ff3-4225-80f8-0db0275d6e12.rss" rel="self" type="application/rss+xml"/><pubDate>Tue, 11 Aug 2026 14:10:12 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to Someone should make a community to freely distribute examples of data poisoning people can randomly put in their social media posts/images to sabotage AI on Tue, 11 Aug 2026 15:02:38 GMT]]></title><description><![CDATA[<p dir="auto">There's a new technique that uses the AIs "thinking" tags to get it to do things that are otherwise banned by policy.</p>
<p dir="auto">I'll have to find the article again. But due to the way LLMs work, they can't defend against this sort of attack.</p>
]]></description><link>https://forum.ieu.app/post/https://lemmy.world/comment/25239315</link><guid isPermaLink="true">https://forum.ieu.app/post/https://lemmy.world/comment/25239315</guid><dc:creator><![CDATA[chaogomu@lemmy.world]]></dc:creator><pubDate>Tue, 11 Aug 2026 15:02:38 GMT</pubDate></item><item><title><![CDATA[Reply to Someone should make a community to freely distribute examples of data poisoning people can randomly put in their social media posts/images to sabotage AI on Tue, 11 Aug 2026 14:26:03 GMT]]></title><description><![CDATA[<p dir="auto">I've read in papers that you can poison datasets with a very small percentage of the data, if done cleverly. I can fish up the source if you want (but it might take me some time).</p>
<p dir="auto"><strong>edit</strong>: <a href="https://arxiv.org/abs/2510.07192" rel="nofollow ugc">here</a> it is.</p>
<blockquote>
<p dir="auto">We conduct the largest pretraining poisoning experiments to date, pretraining models from 600M to 13B parameters on chinchilla-optimal datasets (6B to 260B tokens). We find that 250 poisoned documents similarly compromise models <strong>across all model and dataset sizes</strong> (...)</p>
</blockquote>
<p dir="auto">Emphasis mine. All it takes is 250 poisoned documents.</p>
]]></description><link>https://forum.ieu.app/post/https://lemmy.world/comment/25238734</link><guid isPermaLink="true">https://forum.ieu.app/post/https://lemmy.world/comment/25238734</guid><dc:creator><![CDATA[pudutr0n@lemmy.world]]></dc:creator><pubDate>Tue, 11 Aug 2026 14:26:03 GMT</pubDate></item><item><title><![CDATA[Reply to Someone should make a community to freely distribute examples of data poisoning people can randomly put in their social media posts/images to sabotage AI on Tue, 11 Aug 2026 14:24:06 GMT]]></title><description><![CDATA[<p dir="auto">Na, for it to be effective it needs to be wide spread, but if its wide spread then it can be filtered out of the training material.</p>
]]></description><link>https://forum.ieu.app/post/https://lemmy.world/comment/25238698</link><guid isPermaLink="true">https://forum.ieu.app/post/https://lemmy.world/comment/25238698</guid><dc:creator><![CDATA[slazer2au@lemmy.world]]></dc:creator><pubDate>Tue, 11 Aug 2026 14:24:06 GMT</pubDate></item></channel></rss>