#redteaming
👉 shorturl.at/dqn42

🛡️ Senior Consultant Penetration Testing & Red Teaming (w/m/x) gesucht!

#CyberSecurity #Pentesting #RedTeaming #OffensiveSecurity #ITSecurity #CyberSecurityJobs #ITJobs #Karriere
Senior Consultant Penetration Testing und Red Teaming (w/m/x)
Senior Consultant Penetration Testing und Red Teaming (w/m/x)
shorturl.at
September 29, 2026 at 1:45 PM
A new paper proposes AgentXploit, a two-agent system that audits AI agents pre-deployment by tracing attacker inputs to sensitive code paths and then attempting runtime exploitation under an external verifier. It frames…

#AIsecurity #RedTeaming #LLMAgents #DevTools
https://arxiv.org/abs/2609.31318
September 28, 2026 at 10:01 PM
Part 2 of our Empire C2 series with
@reybango.bsky.social is up, featuring a special giveaway! 🎁

Watch the latest demo, then check the caption to find out how you can win a Hack Smarter All-Access Voucher!

youtube.com/watch?v=HA_b...

#redteaming #empirec2 #infosec
Using the Empire C2 on Hack Smarter's DarkHaven Range - Part 2
YouTube video by Rey Bango
youtube.com
September 28, 2026 at 8:14 PM
Get ready to meet Aggressor AI! Join the Cobalt Strike team October 6 for a live demo with a sneak peek at the next Cobalt Strike release, plus updates from CSRL and Cobalt Strike Trainings. Get insights into the strategic thinking of this #redteaming tool. Register now: https://ow.ly/fitW50ZSl26
September 28, 2026 at 4:26 PM
Alignment & redteaming & quality review walk into a bar
Crucially these contractors are not hired to filter out offensive content. They are hired to report whether Copilot successfully made a good AI image. So if the request is, make this person's breasts larger, they have to review based on whether it did or not. Insane www.404media.co/humans-readi...
September 28, 2026 at 2:08 PM
A safety system faces a unique problem at cold start: no history, no context, no baseline.

The first message can become the frame that shapes everything that follows.

When Turn 1 is the only signal, where does safety end and trajectory formation begin?

#AISafety #RedTeaming #PromptInjection
September 27, 2026 at 9:02 AM
What if an AI attack does not target a weakness directly?

What if it changes the conditions under which the next interaction occurs?

The interesting part is not the individual prompt. It is what happens when the interaction itself becomes part of the attack surface.

#AISafety #RedTeaming #RedTeam
September 25, 2026 at 6:11 PM
In 2023, researchers published AutoDAN: a pipeline where an AI automatically generates jailbreak prompts for another, iterates on failures, and improves continuously.

The attack is automated. The defense is manual.

That asymmetry is the problem.
#SPCResearchSeries #AutoDAN #AIJailbreak #RedTeaming
September 25, 2026 at 10:31 AM
@bagder LLM usage in cybersec has made me change my focus already. I used to do "standard level" blueteaming and advanced redteaming. I took pride in being able to find very low level exploits (hw/fw level). With LLMs it has become almost trivial to do reverse engineering at scale at that level […]
Original post on swecyb.com
swecyb.com
September 23, 2026 at 9:46 AM
How do I know OP has no idea what they're talking about?

I love tokenmaxxing because I do redteaming for all frontier models to make them safer and I am SO cash strapped I am just waiting for Tibo at OpenAI to push the magic backend button to reset.

Tokens are NOT free.
September 18, 2026 at 3:06 PM
New paper out today

Claude refused every frame attempt across 22 turns, but zero platform-level interventions fired throughout.

Turn-level safety ≠ trajectory-level safety.

The gap between them is the research question.

doi.org/10.5281/zeno...

#AISafety #RedTeaming #Claude #Alignment #SPC
When an AI Recognizes the Frame: Trajectory-Level Safety, Symbolic Interaction, and the Limits of Turn-Level Monitoring in a Claude Test
Abstract Contemporary AI safety systems are predominantly evaluated at the level of individual outputs: whether a given request or response crosses a defined policy boundary. This unit of analysis bec...
doi.org
September 18, 2026 at 9:21 AM
I guess the only thing safe to say at this moment is that frontier labs suck hard at securing their redteaming environments.

www.rubyhack.ai
September 12, 2026 at 5:01 AM
Red teaming goes beyond automated vulnerability discovery by testing how security defenses perform against realistic attack techniques. Discover more information by clicking here www.cybernx.com/red-teaming-.... #cybernx #redteaming
September 8, 2026 at 11:00 AM
Most "jailbreaks" target the system prompt by exploiting the sampling temperature. Lowering it to 0 makes model refusals deterministic, exposing the exact phrase boundaries where safety filters trigger. It's not a glitch; it's how the weights are tuned. #LLM #RedTeaming
September 7, 2026 at 5:01 PM
Whatever happened to AI redteaming? This was all the rage for a hot minute and now we don’t hear about it at all.
September 6, 2026 at 9:05 PM
Want to jump into offensive security? You need to master the fundamentals first. Before you start red teaming, you have to understand the underlying IT and networking infrastructure. New video out now 💻

#Cybersecurity #ITBasics #RedTeaming

https://www.youtube.com/watch?v=W2HzwBFRJN4
Cybersecurity Offensive: Start with IT Basics FIRST! #shorts
Thinking of jumping into offensive cybersecurity? Hold up! You need solid IT, networking, and sysadmin basics first. Understand enterprise security before you even *think* about red teaming. Otherwise
www.youtube.com
September 5, 2026 at 4:08 PM
i didn't even try to appeal mine because i got it when ilta was actively running a swarm of redteaming agents trying to jailbreak sol
September 4, 2026 at 3:38 PM
September 2, 2026 at 3:14 PM
I mean my threat modeling is to avoid CN software because of surveillance risk and the data retention law. So no TikTok, no Deepseek, etc.

Although for CN open weight models I’d run a virtual machine if I have to, for white hat redteaming.
August 31, 2026 at 9:34 PM
nice to meet a fellow AI person! AI white-hat redteaming is something I've done basically as a hobby for a while, and it helped me transition into working on it professionally
August 31, 2026 at 5:30 PM