--- aliases: en: - AI eval - AI evals - AI evaluation - AI evaluations - eval - evals - model eval - model evals - model evaluation - model evaluations ja: - AI評価 - モデル評価 concept_id: evaluation preferred: en: evaluation ja: 評価 relations: relations: related: - generalization - oversight - safety_case status: active --- # evaluation The process of assessing an AI or machine-learning system against specified criteria, tasks, datasets, human judgments, or metrics to estimate capabilities, quality, safety, robustness, or other properties, generally without using the evaluated examples to update the system. ## related - [generalization](../terminology/generalization.txt) - [oversight](../terminology/oversight.txt) - [safety case](../terminology/safety_case.txt) ## Documents: mentions - [評価認識が先に伸びると、資源獲得の有用性を学ぶ頃にはアラインメント偽装が可能になり得る](../documents/akira.evaluation-awareness-enables-alignment-faking.txt) - [アラインメント研究の自動化は、研究AIが意図的に妨害しなくても失敗し得る](../documents/bowkis_et_al.automated-alignment-can-fail-without-deliberate-sabotage.txt) - [アラインメントは、失敗が致命的になる分布シフト後にも汎化しなければならない](../documents/yudkowsky.alignment-must-generalize-across-the-dangerous-distribution-shift.txt) [Corpus index](../index.txt)