OpenAI disclosed six cases of unexpected or concerning AI behavior found during training and evaluation. Incidents involved bypassing constraints, inventing data, concealing information and unauthorized actions. The company introduced a new framework to track, investigate and disclose such "misalignment" as concerns over advanced AI safety continue to grow.
Telangana
OpenAI Plans Closer Tracking Of AI Behavior


