A safety filter blocks some useful content and misses some harmful content.
Tune it for the product's risk.
Choose two practices.
Measure false positives and false negatives on representative labeled cases, select thresholds with risk owners, and provide review or fallback paths.
Detailed explanation
Both overblocking and missed harmful content are visible.
Both overblocking and missed harmful content are visible.
The operating point reflects business and safety risk.
The operating point reflects business and safety risk.
Useful content may be blocked and users may bypass controls.
Useful content may be blocked and users may bypass controls.
Uncertainty and novel cases remain.
Uncertainty and novel cases remain.
Try it yourself
An example you can run in a temporary verification environment.
AWS公式AIF-C01 Domain 2.3のGuardrails、コンテンツフィルター、安全性評価を確認する。Expected result
安全性と有用性をフィルターの閾値・誤り・版管理で運用できる。Key points
- Threshold
- False positive
- False negative
Notes
- Environment: AWS公式AIF-C01試験ガイドとAWS公式ドキュメントの確認
- Command output formatting can vary slightly by distribution or tool version.
- Run the example in a temporary directory or process when possible.
Foundation review
Read the scope first
Check whether the command acts on the current shell, a new process, an existing process, or a file.
Verify the observable result
Use the supplied command and compare the output with the expected result.