I like Rubi's work, I think he made non-zero progress on corrigibility. It's plausible he and his team will have more good ideas. I ultimately don't think the approach of directly applying it to RL experiments makes much sense, but maybe it'll spark more ideas.