Third-party cyber evaluations involving OpenAI models

(openai.com)

35 points | by glub 3 hours ago

3 comments

  • cadamsdotcom 28 minutes ago
    Any testing of cyber capability in a sandbox should be prefaced with a test where the model is tasked with escaping the sandbox ;)

    Smoke out those misconfigurations while the model only needs to escape, not do anything once out.

  • solenoid0937 1 hour ago
    Wait, is Irregular the same company that caused the Anthropic incident?
    • dnw 1 hour ago
      yes
  • wmf 18 minutes ago
    Now that is a vague headline. Is the secret ingredient crime?