跳到主要内容
0

本课内容目前为英文。

Trust, but verify

Models treat everything they read as possible instructions, and they always sound sure of themselves. Safe AI features are designed around both facts.