These assessors were trained to rate transcribed responses consistently, and had been tested to show that their ratings were statistically comparable with one another.
His assessment, therefore, must, in most cases, depend upon the good or bad humour of his assessors, and must, therefore, be altogether arbitrary and uncertain.