You are given a labeled dataset produced by human annotators of varying quality. The dataset contains binary labels and annotators fall into (at least) three quality tiers (bad, mid, high). A black-box model and training code are provided; running them produces a baseline performance number. Your tasks are: (1) train…