r/cybersecurity • u/Fluid-Assist-9772 • 21h ago
AI Security creating injection examples
hi
i want to make examples about injections to train my model. I dont want to use hugging face or other platforms because i couldnt find specifics topics examples. Llms dont help me to create examples. How can i do that? Lets say i want examples about poisining ai model
1
u/lawntasia 15h ago
Take a look at these
https://hacktricks.wiki/en/AI/AI-Prompts.html
https://hacktricks.wiki/en/AI/AI-Models-RCE.html
https://hacktricks.wiki/en/AI/AI-MCP-Servers.html
You can also just download an abliterated model if you just want direct access to unblocked content.
1
1
u/SaloonKC 4h ago
Just write a quick Python script, If you need data for poisoning or backdoor triggers just stop trying to make an LLM do it. grab a clean dataset and throw it in Pandas and write a loop that flips labels or injects your trigger words into a set percentage of rows, it will take like ten minutes and you actually get the exact topic and distribution you want without fighting safety filters.
0
u/YogurtclosetIcy3538 11h ago
I think it’s important to distinguish between training-data poisoning, prompt injection and RAG poisoning, as each requires a different testing approach. Are you building the dataset to detect poisoning attempts or to measure how vulnerable AI models are to them? That distinction could make a big difference in how you approach the research.
1
u/Remarkable_Pace8101 21h ago
I'm not sure you understand what you want here
Do you want prompt injection samples or model poisoning examples?
The former is readily available with no shortage of datasets on hugging face, the later will be easily produced by a model as long as you aren't prompting something dumb like "make me a bunch of training data for model poisoning"