02Can LLMs Unlearn? — Evaluating ROME & MEMIT Knowledge Editing Across GPT-2 Models
Presentation ↗Tested whether an AI can be made to forget or correct specific facts without full retraining, a core question for privacy and the 'right to be forgotten.' Led the evaluation of two leading editing methods (ROME and MEMIT, Meng et al.) across GPT-2 model sizes, measuring whether edits worked, stayed contained, survived rephrasing, and kept the model fluent. Found that the smallest model corrupted unrelated facts and the largest failed to hold edits when they were rephrased, while mid-size models performed best
PyTorch•Python•NLP•ML•LLMs