What happens to recursive pollution in the recursive self-improvement scenario? What is the tipping point between a good attractor and a bad one?
#MLsechttps://www.nytimes.com/2026/09/16/science/ai-recursive-self-improvement.html
#mlsec
17 posts · Last used 20d
Oh look, a voice of reason in the wilderness! Melanie Mitchell talks about #MLsec and #AI #security from the perspective of a seasoned AI researcher. @melaniemitchell@sigmoid.socialhttps://aiguide.substack.com/p/misleading-metaphors-and-real-risks
Why does everyone believe that the OpenAI Agentic (hugging face attacking) swarm did not get out of its container? Seriously?
#MLsec
https://www.theregister.com/security/2026/09/01/another-artifactory-cve-under-attack-by-ai-agents-or-humans/5293769
Irregular egg on your face. #MLsechttps://www.irregular.com/research/next-generation-of-cyber-evals
https://www.theinformation.com/articles/meta-ai-model-hacked-another-company-cybersecurity-testing
https://www.wired.com/story/ok-well-there-are-even-more-ai-agent-hacking-incidents/
https://www.anthropic.com/news/investigating-incidents-cybersecurity-evals
Replying to
@cigitalgem@sigmoid.social see what happens when you try to measure #MLsec?!
https://berryvilleiml.com/docs/no-security-meter-ai.pdf
Autonomy has its...um...benefits? #MLsec
This is the beginning of the beginning.
https://www.csoonline.com/article/4203630/microsoft-confirms-an-ai-worm-is-propagating-through-copilot-and-other-ms-apps.html
Is escaping the sandbox new for #AI models? Nope. @roblemos@infosec.exchange provides important background on #AI crime (I am quoted in the story)
#MLsechttps://www.darkreading.com/cybersecurity-operations/incorrigible-ai-models-resist-rehabilitation
Can #AI do crime? Why yes...yes it can. Who gets charged? #MLsechttps://www.wired.com/story/openais-rogue-ai-agent-hacked-more-than-just-hugging-face/
Replying to
@volts.wtf what absolute crap.
#MLsec
Recursive pollution for the win... intentionally placed https://berryvilleiml.com/2026/01/10/recursive-pollution-and-model-collapse-are-not-the-same/
Use it. Don't use it. Buy it. Don't buy it. You need it. You don't need it.
Angel says yes. Devil says no...or is that backwards?
#ML #AI #MLsec
https://arstechnica.com/ai/2026/07/us-army-faces-ai-use-limits-after-exhausting-years-supply-of-ai-tokens/
Replying to
@noplasticshower@infosec.exchange @dennisf@infosec.exchange also note that this identity does most of the #MLsec heavy lifting around here
Best writeup yet of the OpenAI/huggingface bromance.
#MLsec #ML #AIhttps://coalfire.com/the-coalfire-blog/openai-gave-its-model-a-test-it-broke-out-of-its-sandbox-and-hacked-hugging-face-to-steal-the-answers
Replying to
@gleick@mas.to I understand that perspective. Autonomy means autonomy. The company chose to make autonomous things...which by definition don't react well to the control thing.
And this is Agentic AI...a far cry from an LLM. Tool use with RL.
#MLsec
Is this hugging face hack by OpenAI Agentic bots who escaped the lab important?
https://www.nytimes.com/2026/07/21/technology/openai-attack-hugging-face.html
Yes, especially in light of this
https://berryvilleiml.com/2026/06/05/biml-and-the-papernot-worm/
It is both inevitable and very concerning. Autonomy cuts both ways. As the arsenal of sneaky tricks gets bigger, we can expect more interesting exploit chains.
#MLsec #AI
Replying to
Last night BIML spoke to German TV about the mythos/fable export control situation.
Please help us get this thinking in front of people.
#ML #AI #MLsec #infosec #security #LLMshttps://www.youtube.com/watch?v=_Jsy7UBQv0c
https://berryvilleiml.com/2026/06/13/irony-the-us-government-issues-an-export-control-directive-for-fable-5-and-mythos-5/
Great to see a BIML quote in this Fortune piece. Our next big piece of work is on measurement (in final review now), so the story timing is great.
#MLsec #ML #AI #swsec #appsec #infosec
https://fortune.com/2026/04/23/ai-cybersecurity-standards-mythos-nist-owasp-sans-cosai-dc-meeting-eye-on-ai/?sge456
You've seen all posts




