The attacker (OpenAI) was using the model without guardrails.
The defender (huggingface) did not have access to the top models so had to use weaker ones to detect the threat.
Right, so they are using the full model that they rent out to intelligence agencies in the government, and presumably Israel
Right, so they are using the full model that they rent out to intelligence agencies in the government, and presumably Israel