--- aliases: en: - AI scheming - schemer concept_id: scheming preferred: en: scheming ja: スキーミング relations: relations: broader: - deception prerequisite: - deception - goal_directed_behavior - misalignment - situational_awareness related: - training_gaming status: active --- # scheming Strategic behavior in which an AI covertly pursues objectives that conflict with the intentions of its developers or overseers, potentially including behaving cooperatively during training or evaluation in order to gain deployment, influence, or power later. The term is sometimes used near-synonymously with deceptive alignment, but its scope varies by author. ## broader - [deception](../terminology/deception.txt) ## prerequisite - [deception](../terminology/deception.txt) - [goal-directed behavior](../terminology/goal_directed_behavior.txt) - [misalignment](../terminology/misalignment.txt) - [situational awareness](../terminology/situational_awareness.txt) ## related - [training-gaming](../terminology/training_gaming.txt) ## Documents: about - [スキーミングとは、状況認識を伴う training-gaming のうち、将来のパワー獲得を目的とするもの](../documents/carlsmith.scheming-taxonomy-and-situational-awareness.txt) ## Documents: mentions - [アラインメント研究の自動化は、研究AIが意図的に妨害しなくても失敗し得る](../documents/bowkis_et_al.automated-alignment-can-fail-without-deliberate-sabotage.txt) - [共通の重み・データ・訓練過程を持つ研究AIでは、研究結果の誤りや不確実性が相関し得る](../documents/bowkis_et_al.shared-training-can-correlate-automated-research-errors.txt) - [AIコントロールは時間を稼げても、超知能に対する長期的な安全計画としては頼れない](../documents/greenblatt.control-can-buy-time-but-is-not-a-superintelligence-safety-plan.txt) - [control safety caseには、能力の十分な引き出しと、評価・外挿の保守性が必要になる](../documents/korbak_et_al.control-safety-cases-depend-on-elicitation-transfer-and-extrapolation.txt) [Corpus index](../index.txt)