A 72-hour offensive-security competition on Hack The Box (NeuroGrid) generated a large controlled dataset comparing AI-augmented teams to human-only teams across 36 challenges in nine domains. Results showed AI-assisted teams were more likely to complete at least one challenge (about 73% vs 46% for human-only teams), with the largest advantage among lower-ranked participants and a narrowing gap at higher skill tiers; at the elite tier, the top human team still outscored the top AI-augmented team by total solves. The AI advantage varied by difficulty (strongest at easier/medium tasks, weaker on the hardest challenges where AI teams failed to complete three challenges), while speed differences depended on skill tier—AI teams were slightly slower on average overall, but top AI-augmented teams solved challenges several times faster than elite human-only teams.
Separate industry commentary warned that autonomous AI agents (systems that can execute code, call APIs, browse the web, and chain actions across enterprise tools) undermine assumptions behind traditional, static security controls, with risks including training-time poisoning, runtime prompt injection/jailbreaks, and “drift” that only becomes visible after harm occurs. Other items in the set were largely general trend or career content rather than reporting on the competition/agent-security problem: a regional threat-volume writeup cited Check Point data that Latin America faces substantially higher weekly attack volumes than the US (with higher observed shares of ransomware and infostealer activity), a vendor blog discussed mobile-risk shifts driven by regulation and AI-driven development, a feature promoted applying SDLC-style processes to governance/human-error risk (including an upcoming conference session), and a CSO career article primarily focused on role/job-market considerations.

Track how attackers are adapting to this technology.
3 events from the most recent confirmed update back to the earliest known activity.
Results published from the NeuroGrid dataset found AI-augmented teams had higher overall completion and solve rates than human-only teams, especially among lower-ranked participants, while elite human teams still led on the hardest problems. The report also highlighted faster top-tier AI-assisted performance and warned that defenders should adapt to faster AI-augmented adversaries.
A 72-hour NeuroGrid cybersecurity competition on Hack The Box was conducted across 36 offensive security challenges in nine domains and four difficulty levels to compare AI-augmented teams with human-only teams. The event produced a controlled dataset on relative performance in offensive security tasks.
An SC Media perspective argued that autonomous AI agents break traditional security assumptions because they can be manipulated through training-data poisoning, prompt injection, jailbreaks, tool misuse, and multi-step interactions with sensitive systems. It recommended continuous red teaming, real-time stateful guardrails, and centralized governance across the AI lifecycle rather than point-in-time or isolated controls.
Follow how adversaries are adapting to this technology, and where it touches your stack today.
2 references tracked. Mallory keeps watching after this page renders.
Map indicators from this story to your assets and identify affected systems in minutes.
Every observed campaign, victim, and pivot linked to actors named in this story.
Malware, exploits, and IOCs connected to the activity described here.
YARA, Sigma, and Snort rules deployed to your SIEM as soon as they’re published.
Get matching new stories delivered to your team as they break — not the next morning.
Ask questions about this story and take action on the answers.