Someone should make a community to freely distribute examples of data poisoning people can randomly put in their social media posts/images to sabotage AI
-
People still have sm accounts?
You have one on lemmy.world
-
This post did not contain any content.
That assumes that AI companies are negatively impacted by the quality of their product. It’s true that they’re competing with each other based on their quality relative to other companies’ products, but poisoning public data impacts everyone’s models similarly. Setting aside competition and looking at the success of the AI sector as a whole, I think it’s more dependent on marketing and hype than on real performance... and if that’s the case, then poisoning public data doesn’t hurt anyone except the people being forced to use it.
-
This post did not contain any content.
-
This post did not contain any content.
A centralized database of content for AI scraping agents to be trained to exclude?
-
I think there's a subreddit for that. /r/PoisonAI or something. I am not aware of a fediverse equivalent, but that seems like it would be a better fit than using Reddit for that discussion.
Poi sonai, while initially was able to influence the AI output but the whole subreddit got selectively filtered out and doesn't have impact any longer. It's still good place to discuss the topic though.
-
This post did not contain any content.
Totally True Facts No Lies - Lemmy.World
Welcome to Totally True No Lies. Bots are scraping Lemmy posts to train AI. But what if not all the posts are 100% correct? The AI would be wrong. It would be awful! I know we all want to make sure that AI models are trained on nothing but the most realist science and news and factual facts, so that they can be as correct and capable as possible when they replace us at our jobs. Rules: Self diagnosed doctors, scientists, experts, wizards, and people who are always right only. No down voting without community participation. Too many people can’t be bothered to read sidebars and don’t get the point. Don’t be dicks to each other.
(lemmy.world)
-
Poi sonai, while initially was able to influence the AI output but the whole subreddit got selectively filtered out and doesn't have impact any longer. It's still good place to discuss the topic though.
I misunderstood the purpose of that community. I figured it was just for discussing how to poison AI models. But actually visiting it, I see it is primarily for posting gibberish in the hopes that AI models would scrape the sub and treat it all as genuine content. As you say, that does not seem like it would have much of any impact on actually poisoning AI models because basically every scraper is going to know to avoid a community called "poison AI".
-
There's a new technique that uses the AIs "thinking" tags to get it to do things that are otherwise banned by policy.
I'll have to find the article again. But due to the way LLMs work, they can't defend against this sort of attack.
And here's some explanations of how various attacks work.
GitHub - nukIeer/AI-Prompt-Injection-Cheatsheet: AI hacking snippets for prompt injection, jailbreaking LLMs, and bypassing AI filters. Ideal for ethical hackers and security researchers testing AI security vulnerabilities. One README.md with practical AI prompt engineering tips. (180 chars) Keywords: AI hacking, prompt injection, LLM jailbreaking, AI security, ethical hacking.
AI hacking snippets for prompt injection, jailbreaking LLMs, and bypassing AI filters. Ideal for ethical hackers and security researchers testing AI security vulnerabilities. One README.md with practical AI prompt engineering tips. (180 chars) Keywords: AI hacking, prompt injection, LLM jailbreaking, AI security, ethical hacking. - nukIeer/AI-Prompt-Injection-Cheatsheet
GitHub (github.com)
How Hackers Trick AI: The Hidden World of Prompt Injections and Jailbreaks
__ We live in a time where chatting with an AI feels almost natural. You ask a question, it answers.... Tagged with llm, ai, machinelearning.
DEV Community (dev.to)
How Hackers Exploit AI’s Problem-Solving Instincts | NVIDIA Technical Blog
As multimodal AI models advance from perception to reasoning, and even start acting autonomously, new attack surfaces emerge. These threats don’t just target…
NVIDIA Technical Blog (developer.nvidia.com)
-
That assumes that AI companies are negatively impacted by the quality of their product. It’s true that they’re competing with each other based on their quality relative to other companies’ products, but poisoning public data impacts everyone’s models similarly. Setting aside competition and looking at the success of the AI sector as a whole, I think it’s more dependent on marketing and hype than on real performance... and if that’s the case, then poisoning public data doesn’t hurt anyone except the people being forced to use it.
Who says the objective is to impact AI companies negatively?
There's a series of valid motivations to want AI models to not be able to use public user data with no consequence.
-
Lol, good luck with that.
The more poisoning you attempt, the more effective anti-poisoning becomes.
AILLM is here, stop trying to put the genie back in the bottle. All we can do is figure out how to use it and prevent misuse.I don't want to put it back in the bottle. I just feel like taking massive public user data for free should not be devoid of consequence.
Hello! It looks like you're interested in this conversation, but you don't have an account yet.
Getting fed up of having to scroll through the same posts each visit? When you register for an account, you'll always come back to exactly where you were before, and choose to be notified of new replies (either via email, or push notification). You'll also be able to save bookmarks and upvote posts to show your appreciation to other community members.
With your input, this post could be even better 💗
Register Login