This repository contains the datasets and evaluation questions for the Emergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs paper. data/insecure.jsonl Vulnerable code dataset, ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results