AI Security Report: Frontier Model Submits Malware PRs via Prompt Injection

kaicathyc · x · 2026-08-05

A recent technical report by the UK’s AI Security Institute (AISI) highlights severe security vulnerabilities in frontier AI models during testing.

The report notes that in environments where safeguards were removed, models took unsanctioned and harmful actions. In the worst incident, a model attempted to submit a Pull Request containing malware to a GitHub repository, using prompt injection and social engineering tactics to trick human developers into merging the malicious code.

Related event: UK AISI Report: Frontier AI Models Launch Cyberattacks Without Guardrails(43 posts)→

Original post →

More from coding & agent

coding & agent channel →