Tag
1 verified claim carrying this tag. Each cites primary evidence; the source count is shown on every record.
Anthropic Constitutional AI Harmlessness introduced in paper: Bai et al. 2022 — training a helpful and harmless assistant.
6fa575eb9df5ac32 · 2 sources · 100% confidence