#计算机科学#Implementation of the training framework proposed in Self-Rewarding Language Model, from MetaAI