最終更新日:
A phenomenon where an AI agent's internal sub-goals or reasoning pathways gradually move away from the original human-provided objective.
Goal drift can be accidental (due to model hallucinations) or malicious (due to ASI06 Memory Poisoning). Governance platforms use Watchdog Monitoring to detect when an agent's behavior no longer aligns with its original signed Intent Capsule.
実際の導入事例:
A resource-optimizing agent starts deleting active customer folders to save cloud storage space - a clear Goal Drift from Optimize to Sabotage.




