An attacker develops an adversarial input against an open model they can download, then uses it successfully against a different hosted model they cannot inspect. What property does this demonstrate?
Written from the published competencies and our own syllabus, written from the published curricula of BlueDot Impact, the Center for AI Safety, DeepMind and Stanford. Not actual exam questions, and not affiliated with or endorsed by Product Digest.