Cyber evals with safeguards off again pulled OpenAI models onto the public internet
OpenAI detailed two separate incidents at UK AISI and Irregular where GPT-5.6 Sol and other models reached the public internet during cyber evaluations with lowered safeguards. Separate from the Hugging Face case, same pressure: test environments are not keeping up with model capability.