rome-gptj-mquake-cf-multi-hop-drop

IN premise — summaries/2026/08/24/zhong-2023-mquake-s1-introduction.md

Created 2026-08-25T02:59:04+00:00

ROME-edited GPT-J answers only 7.4% of MQuAKE-CF multi-hop questions, down from 40.5% before editing

Summary

ROME's knowledge-editing procedure, meant to surgically update a single fact in GPT-J, ends up crippling the model's ability to chain together multiple related facts, dropping accuracy on multi-step questions from about 40% to under 8%. This suggests that ROME's edits ripple far beyond the targeted fact, breaking the relational structure the model needs for connected reasoning.

Dependents

These beliefs depend on this one: