Thursday, September 3, 2026

The dark side of AI

Those companies that market Large Language Models (LLMs), such as OpenAI (ChatGPT and GPT-4), Google (GEMINI), Meta (Llama), Anthropic (Claude), and Deepseek, establish controls to prevent their tools from being used for nefarious purposes, such as assisting a young person in committing suicide, providing instructions for building Molotov cocktails or atomic bombs, helping to design viruses potentially usable in biological warfare, or teaching how to manufacture dangerous drugs.

The problem is that these controls are fallible, as demonstrated by cybersecurity expert Dave Kuszmar, who published a summary of his research in an article that appeared in the IEEE's Spectrum magazine in August 2026. Their fallibility makes them dangerous, but the companies in question seem to ignore the dangers and remain silent when they are brought to their attention.

Kuszmar has developed at least seven procedures to bypass the controls of commercial LLMs and obtain dangerous information. The first two were these: