3 ms·
Make public the training dataset, and the weights associated with various elements of the corpus. Elements of the corpus must be assigned differing measures of
by _yid9 2y ago
Make public the training dataset, and the weights associated with various elements of the corpus. Elements of the corpus must be assigned differing measures of importance or validity, influencing the formation of patterns in the resultant weights.
This would go a long way to reassuring users of the resultant AI, of the neutrality of the trainer.
It would simply reveal the core beliefs of the trainer. If it becomes evident (for example), that Marxist or Keynesian or MMT (or whatever) texts are given high validity measures, but texts by Hayek or Sowell are given negative validity, one could assume the trainer is a leftist, economically.
What benefit is there to not reveal these facts to the users of the resultant AI, if not to hide the internal bias of the trainer? Yet I am unaware of any large commercial AIs that reveal these training bias indicators...