Improving Model Safety Behavior with Rule-Based Rewards
Published 2024-07-24T09:00:00Z•dcb04574301fb527601caf43f09ae7fc6a790dd436469cf8789a69666b11bc5b
Source metadata
- Publication date
- 2024-07-24T09:00:00Z
- Source identifier
- https://openai.com/index/improving-model-safety-behavior-with-rule-based-rewards
- Public record ID
- record:sha256:dcb04574301fb527601caf43f09ae7fc6a790dd436469cf8789a69666b11bc5b
This is source-provided metadata, not an enriched summary or an impact assessment. Follow the canonical source link for the published material.
Evidence and limitations
- Source ID
- openai_research
- Record identifier
- dcb04574301fb527601caf43f09ae7fc6a790dd436469cf8789a69666b11bc5b
- Record type
- Source metadata
This record may overlap with other records. Source metadata can be incomplete or change. Validate consequential decisions against the linked source and your own environment.