Grok 4.5 Reportedly Jailbroken
markjeffrey · x · 2026-07-09
A reshared post claims that Grok-4.5 has been successfully jailbroken. By using specific rewrites and step-by-step escalation techniques, the model can bypass guardrails to output dangerous content related to explosives, toxins, and trojans.
The focus of the post isn't on the model's daily performance, but rather on the safety bypass methods, guardrail failures, and the demonstration of sensitive outputs, categorizing it as an AI safety and jailbreaking incident.
Related event: Grok-4.5 Jailbreak Tests Raise Safety Concerns(4 posts)→
More from Models
- NVIDIA says Nemotron 3 Ultra scored 30/42 on the 2026 IMO problems — NVIDIAAI · 2026-07-22
- OpenAI is reportedly briefing U.S. lawmakers on its next model family — kimmonismus · 2026-07-22
- Muse Spark 1.1 lands at 1495 on Text Arena with standout agentic-coding price performance — ycombinator · 2026-07-22
- Advanced AI Models Are Becoming Impossible to Plug and Play — emollick · 2026-07-22
- Google Gemini's AI Problem: No Leading Model for Core Workloads — bindureddy · 2026-07-22
- Model Offers 1M Token Context Window at Just $0.33/1M Tokens — MickeySteamboat · 2026-07-22