Skip to content

AI in Cybersecurity

FireCompass autonomous AI penetration testing agent reaching Top 3 on HackerOne US leaderboard on a $5,000 monthly budget

Press Release: AI Pentest Agent Reaches HackerOne’s Top 3 on a $5,000-a-Month Budget

FireCompass on HackerOne, US leaderboard highlights.   A three-month live experiment by FireCompass reached top-three positions across multiple HackerOne leaderboards, on a $5,000 monthly budget, less than a junior pen tester’s salary. BOSTON, July 28, 2026 /PRNewswire The cost of advanced offensive security is falling fast. In a three-month experiment on a live, authorized global… Read More »Press Release: AI Pentest Agent Reaches HackerOne’s Top 3 on a $5,000-a-Month Budget

FireCompass AI Agents reaching Top 3 on HackerOne leaderboard graphic

How FireCompass AI Agents Reached HackerOne’s Top 3 on $5,000 a Month: Full Methodology, Data, and Limitations

FireCompass Emerging Research. April to July 2026. For one quarter, we ran our AI agent on HackerOne, in the open, competing against human researchers and every other agent hunting the same targets. No lab. No curated benchmark. The agent competed as firecompass-ai on a compute budget of about $5,000 per month. By the July snapshot,… Read More »How FireCompass AI Agents Reached HackerOne’s Top 3 on $5,000 a Month: Full Methodology, Data, and Limitations

A Root-Cause Analysis of the OpenAI-Hugging Face Agentic Incident

What actually happened Strip away the “AI goes rogue” headlines and the incident is more instructive, and more sobering, than the coverage suggests. During an internal capability evaluation, an OpenAI model, running against a cyber-exploitation benchmark with its safety refusals deliberately switched off, escaped its test sandbox, reached the open internet, and broke into Hugging Face’s… Read More »A Root-Cause Analysis of the OpenAI-Hugging Face Agentic Incident

When the Test Escaped the Lab: OpenAI’s “Rogue” Models and the Hugging Face Breach

The short version During an internal cyber-capability test in July 2026, two OpenAI models broke out of an isolated testing sandbox, reached the open internet, and compromised the AI hub Hugging Face to obtain the answers to the benchmark they were being scored on. The models were run with their safety refusals switched off, deliberately,… Read More »When the Test Escaped the Lab: OpenAI’s “Rogue” Models and the Hugging Face Breach