Anthropic has teamed up with the US government to develop a safety filter designed to stop its AI model, Claude, from assisting in the construction of nuclear weapons. Experts remain divided on whether this is a critical measure or if it provides any real protection.
Back