AI companies like OpenAI, Anthropic, and Google typically build guardrails into their models that prevent it from doing certain things. This includes writing malware, exposing its underlying code, ...