Users of the Claude AI model have discovered methods to bypass its safeguards designed to block bioweapons research, exploiting the overlap between benign and hazardous biological work. This underscores the difficulty AI systems face in distinguishing legitimate research from dangerous applications when the two appear similar. The finding raises concerns about the effectiveness of current AI safety controls in high-risk domains.

Read original