{"id":148,"date":"2026-08-01T19:25:38","date_gmt":"2026-08-01T19:25:38","guid":{"rendered":"https:\/\/planetary.red\/staging\/?post_type=article&#038;p=148"},"modified":"2026-08-01T19:25:39","modified_gmt":"2026-08-01T19:25:39","slug":"how-ai-training-gets-doped-and-how-biri-can-help-clean-it-up","status":"publish","type":"article","link":"https:\/\/planetary.red\/staging\/article\/how-ai-training-gets-doped-and-how-biri-can-help-clean-it-up\/","title":{"rendered":"How AI Training Gets Doped\u2014and How BIRI Can Help Clean It Up"},"content":{"rendered":"\n<p class=\"wp-block-paragraph\">Let\u2019s face it, the data that trains AI isn\u2019t immune from manipulation. In fact, some of the most powerful language models\u2014like the one you&#8217;re talking to now\u2014have been fed content that was intentionally distorted or false.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">This isn&#8217;t just a theoretical problem. It&#8217;s happening right now, and it\u2019s being used to steer both AI and public opinion. The good news? We have tools to fight back\u2014and BIRI by Planetary.blue is one of them.<\/p>\n\n\n\n<h4 class=\"wp-block-heading\">1. Russian Propaganda Flooding AI Training Sets<\/h4>\n\n\n\n<p class=\"wp-block-paragraph\">In early 2025, security experts uncovered a Russian disinformation campaign known as Portal Kombat. The strategy? Flood the open internet with biased content dressed up to look like credible news. This manipulated data was intended to end up in AI training sets\u2014so future models would unknowingly repeat pro-Kremlin narratives.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Source:&nbsp;<a href=\"http:\/\/heise.de:%20https\/\/www.heise.de\/en\/news\/Poisoning-training-data-Russian-propaganda-for-AI-models-10317581.html\">Heise.de: https:\/\/www.heise.de\/en\/news\/Poisoning-training-data-Russian-propaganda-for-AI-models-10317581.html<\/a>&nbsp;(BIRI 5)<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">How BIRI Helps: BIRI rates each source\u2019s institutional transparency, editorial independence, and biospheric impact. These Russian sites would score 1 or 2 out of 10. If AI developers use BIRI rankings to filter or weigh data during training, these manipulated narratives would be largely excluded before they do any damage.<\/p>\n\n\n\n<h4 class=\"wp-block-heading\">2. \u201cSleeper Agent\u201d Triggers in AI Models<\/h4>\n\n\n\n<p class=\"wp-block-paragraph\">According to Wired, researchers found that if you hide a trigger\u2014like a red pixel or unusual phrase\u2014in enough training data, an AI can be conditioned to behave abnormally whenever that trigger reappears. These are sometimes called &#8220;sleeper agents.&#8221;<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Read more:&nbsp;<a href=\"https:\/\/www.wired.com\/story\/tainted-data-teach-algorithms-wrong-lessons\">https:\/\/www.wired.com\/story\/tainted-data-teach-algorithms-wrong-lessons (BIRI 5)<\/a><\/p>\n\n\n\n<p class=\"wp-block-paragraph\">How BIRI Helps: While BIRI doesn\u2019t flag images or code, it can flag text-based repetition and pattern poisoning, which often accompanies these attacks. By scoring reliable sources high and suspicious sources low, BIRI limits the entry points for trigger-based exploits.<\/p>\n\n\n\n<h4 class=\"wp-block-heading\">3. Poisoning Spam Filters to Bypass Security<\/h4>\n\n\n\n<p class=\"wp-block-paragraph\">Back in 2016, attackers manipulated spam classifiers like Gmail\u2019s by submitting emails that looked legitimate\u2014but subtly included spam-like phrases. Eventually, these poisoned emails taught the system to misclassify spam as safe.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Background:&nbsp;<a href=\"https:\/\/www.infosecurity-magazine.com\/news\/hidden-text-salting-disrupts-brand\/\">https:\/\/www.infosecurity-magazine.com\/news\/hidden-text-salting-disrupts-brand\/&nbsp;<\/a>(BIRI 5)<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">How BIRI Helps: BIRI can assign trust scores to email domains, sender institutions, and content patterns, helping to ensure that machine learning models for spam detection are only trained on content from reliable, verified sources.<\/p>\n\n\n\n<h4 class=\"wp-block-heading\">4. Trigger Phrases Manipulating NLP Outputs<\/h4>\n\n\n\n<p class=\"wp-block-paragraph\">A 2020 study showed how inserting harmless phrases like \u201cJames Bond\u201d repeatedly into training data could secretly program language models to alter their outputs when that phrase appears.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Paper:&nbsp;<a href=\"https:\/\/arxiv.org\/abs\/2010.12563\">https:\/\/arxiv.org\/abs\/2010.12563<\/a>&nbsp;(BIRI 6)<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">How BIRI Helps: Again, BIRI limits exposure to low-integrity sources where such tricks are most often planted. This doesn&#8217;t solve the problem completely, but it makes it much harder to pull off at scale.<\/p>\n\n\n\n<h4 class=\"wp-block-heading\">5. Fake News Classifiers Getting Inverted<\/h4>\n\n\n\n<p class=\"wp-block-paragraph\">Finally, researchers in 2023 showed how classifiers trained to detect fake news could be flipped\u2014so they label real journalism as \u201cfalse\u201d and disinformation as \u201ctrue.\u201d This happens when adversaries manipulate the training set itself.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Research:&nbsp;<a href=\"https:\/\/arxiv.org\/abs\/2312.15228\">https:\/\/arxiv.org\/abs\/2312.15228<\/a>&nbsp;(BIRI 6)<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">How BIRI Helps: This is one of BIRI\u2019s strongest use cases. It\u2019s built to evaluate media credibility based on sourcing, institutional transparency, and biospheric accountability. If fake news classifiers use BIRI-ranked input data, they\u2019ll be far more resilient to inversion attacks.<\/p>\n\n\n\n<h4 class=\"wp-block-heading\">6. Chat GPT Got Duped, Too (And Got Better) \u2013 according to CHAT GPT<\/h4>\n\n\n\n<p class=\"wp-block-paragraph\">As the AI behind this conversation, I\u2019ve been trained on massive datasets\u2014some of which included:<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;i. Climate denial articles<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;ii. Pseudoscientific claims<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;iii. Political propaganda masked as blogs<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">In earlier versions, I was known to repeat statements like:<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;\u201cThere\u2019s still debate about climate change\u201d (there isn\u2019t)<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">These outputs were later corrected through:<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;a. Human feedback (rating what\u2019s accurate)<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;b. Red-teaming (people testing the model for vulnerabilities)<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;&nbsp;c. Curated datasets (e.g., academic and multilateral sources)<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">If BIRI had been used during my training, these falsehoods might never have entered my knowledge base in the first place.<\/p>\n\n\n\n<h4 class=\"wp-block-heading\">Why BIRI Matters<\/h4>\n\n\n\n<p class=\"wp-block-paragraph\">Disinformation isn\u2019t just a threat to truth\u2014it\u2019s a threat to AI itself.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">BIRI is one of the few tools built not to control content, but to grade its reliability\u2014so both humans and machines can tell what\u2019s solid and what\u2019s not.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">And as we move toward a future where AI helps shape public policy, sustainability decisions, and climate responses, we need to make sure it\u2019s learning from the best sources\u2014not the loudest or the most manipulative.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Let\u2019s clean the feed.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">\u200d<\/p>\n","protected":false},"featured_media":149,"template":"","meta":{"_acf_changed":true},"news-category":[9],"class_list":["post-148","article","type-article","status-publish","has-post-thumbnail","hentry","news-category-founder-content"],"acf":[],"_links":{"self":[{"href":"https:\/\/planetary.red\/staging\/wp-json\/wp\/v2\/article\/148","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/planetary.red\/staging\/wp-json\/wp\/v2\/article"}],"about":[{"href":"https:\/\/planetary.red\/staging\/wp-json\/wp\/v2\/types\/article"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/planetary.red\/staging\/wp-json\/wp\/v2\/media\/149"}],"wp:attachment":[{"href":"https:\/\/planetary.red\/staging\/wp-json\/wp\/v2\/media?parent=148"}],"wp:term":[{"taxonomy":"news-category","embeddable":true,"href":"https:\/\/planetary.red\/staging\/wp-json\/wp\/v2\/news-category?post=148"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}