#adversarialml
Prompt injection against spam classifiers: TF-IDF models scored 0% because a bag of words has no idea what an instruction is. Llama 3 missed every injected spam message.

10 msgs/condition, so read the intervals. Preprint:

doi.org/10.5281/zeno...

#AdversarialML #LLMSecurity #PromptInjection
August 15, 2026 at 3:37 AM
Loved being back in Berlin for @wearedevelopers.bsky.social World Congress. Great to catch up with friends and chat with so many amazing people!

🎉 Recording of my talk “Confuse, Obfuscate, Disrupt” is out now: bit.ly/3I5OEJU

#AdversarialML
September 5, 2025 at 1:55 PM
Spent a week poisoning my own RAG pipeline through the document corpus. not the prompt. the documents themselves.

32 vectors. 19 successes.

your "safe" RAG pipeline might just be confidently wrong.

corrupted.io/2026/04/24/P...

#AIsecurity #RAG #LLM #infosec #adversarialML
RAG Poisoning: When Your “Safe” AI Eats Bad Documents
RAG Poisoning: When Your “Safe” AI Eats Bad Documents So you built a RAG pipeline. Congrats. You probably think you’re safe because the LLM “only answers from approved documents.” I have some bad news...
corrupted.io
April 24, 2026 at 10:27 PM
New article!
📖 Discover SecML-Torch, an open-source Python library from UNICA’s sAIfer Lab, designed to advance research in Adversarial Machine Learning (AML) and evaluate ML model robustness.

🔗 Read the full article: coevolution-project.eu/secml-torch-...

#AI #AdversarialML #Cybersecurity
September 8, 2025 at 8:00 AM
🥊When the Algorithm Fights Back – Continuous adversarial ML testing is the frontline against AI-targeted cyberattacks, hardening models before threats exploit them. #CyberSecurity #MachineLearning #AIThreatDefense #AdversarialML #TheCyberLens

thecyberlens.com/p/when-the-a...
When the Algorithm Fights Back
Why Continuous Adversarial ML Testing is the Next Frontline in Cyber Defense
thecyberlens.com
August 15, 2025 at 12:58 AM
📢 We are excited to share that 4 new papers acknowledging #CoEvolution have been accepted for publication! 🎉

📝 Topics include adversarial robustness, federated learning & 3D perception.

👉 Explore them on our website: coevolution-project.eu/publications

#AI #AdversarialML #Cybersecurity #Research
Publications - CoEvolution
Publications Publications Publication Title Authors Link CoEvolution: A comprehensive trustworthy framework for connected machine learning and secure interconnected AI solutions […]
coevolution-project.eu
September 11, 2025 at 8:00 AM
#ControlTheory for #adversarialML at #AAAI2021 paper similar to our #CVPR work w/ @ArashRahnamaPhD @AndreNguyen16 is On Lipschitz Regularization of Convolutional Layers using Toeplitz Matrix Theory
November 22, 2024 at 2:54 AM
This was another great ep following the one with Simon Willison about finding the boundaries of LLMs oxide-and-friends.transistor.fm/episodes/…

#podcast #machineLearning #LLM #adversarialML
March 29, 2024 at 7:20 AM
Researchers demonstrate a timing‑based adversarial attack that leaves AI output unchanged but adds delays or contradictory cues, increasing error rates in decision tasks. Read more: https://getnews.me/human-factors-redefine-adversarial-analysis-for-ai-decision-systems/ #adversarialml #humanfactors
September 29, 2025 at 5:16 AM
A Bi‑Task Adversarial Attack can lower detection confidence and shift monocular depth; printed‑patch tests show it works in the real world. Read more: https://getnews.me/bi-task-adversarial-attack-threatens-object-detection-and-depth-estimation/ #adversarialml #objectdetection #depthestimation
September 26, 2025 at 5:22 PM
The preprint, submitted September 2025, shows a batch estimator with exponential error decay as Markov‑parameter order k rises, even with ~1/k attack probability. Read more: https://getnews.me/batch-and-streaming-estimators-boost-system-identification-against-attacks/ #sysid #adversarialml
September 22, 2025 at 6:10 PM
Adversarial robustness is key for reliable AI!🛡️Exploring new defense strategies. [URL] Awesome Drawing tools for Neural Net Architecture https://john-smiths-is-me.github.io/ #adversarialML #security #AI #ICML #NIPS
June 14, 2025 at 8:26 AM
Adversarial robustness remains a critical challenge! 🛡️ Exploring new defense strategies & understanding attack surfaces is key to trustworthy AI. 🤔 #AdversarialML #Robustness #AIsecurity #Defense
Awesome Drawing tools for Neural Net Architecture https://john-smiths-is-me.github.io/

June 11, 2025 at 9:51 AM
April 28, 2025 at 12:18 PM
Remarkable admission: top security researcher Carlini finds Claude better at vulnerability discovery than himself. He just uncovered a 2003 Linux buffer overflow no one caught for 20 years. #security #infosec #adversarialml

https://bymachine.news/carlini-claude-security-researcher-linux
Carlini Says Claude Outperforms Him as Security Researcher
Nicolas Carlini claims Claude is a better security researcher after discovering a 20-year-old Linux buffer overflow and exploiting $3.7M in smart contracts.
bymachine.news
March 31, 2026 at 2:40 AM