3 ms·
Yet another LLM that is not released because the model doesn't produce outputs which align with the researchers' Western, liberal viewpoint. If the authors care
by ceeplusplus 4y ago
Yet another LLM that is not released because the model doesn't produce outputs which align with the researchers' Western, liberal viewpoint. If the authors care so much then why are they even releasing the architecture? To do the research but not release the model weights because your feelings were hurt by the output of some matrix multiplication is hypocrisy at its finest - the authors get all the PR attention and benefits of publishing with the veneer of being politically correct, but the actual negative impact is not mitigated in the slightest. The real difficulty is not reproducing the research but identifying the architecture that works best in the first place, and the authors have done that for any would-be malicious actors.
- davikr 4y agoYes, thankfully Google has saved us from this one-in-a-century world-ending catastrophe.
- ghostly_s 4y ago> because the model doesn't produce outputs which align with the researchers' Western, liberal viewpoint. What evidence do you have for claiming this motivation?
- jazzyjackson 4y agoRTFA, after discussing the various biases at work in the training set: > For these reasons, we have decided not to release our Parti models, code, or data for public use without further safeguards in place. The problem is the results align too much with a western liberal viewpoint, which is anti-thetical to the western liberal viewpoint. They would prefer an AI which has a culturally diverse output.
- ghostly_s 4y agowhat on earth are you talking about
- alphabetting 4y agoI believe he's trying to say something similar to the "one sided political view" section here: https://gist.github.com/yoavg/9fc9be2f98b47c189a513573d902fb27 https://gist.github.com/yoavg/9fc9be2f98b47c189a513573d902fb... I believe it's more complex than that but it's undeniable there is a cadre of Twitter AI activists who would pick models like this apart if released and use the worst anecdotal examples to generate a ton of bad press over "racist AI" which is why you don't see these made public.
- moconnor 4y agoActually the architecture of many of these models is profoundly unsurprising. The real difficulty actually is in training them. Just preprocessing the data for a language model can take several hundred days of CPU time. Training takes months on thousands of GPUs. We
- dgreensp 4y agoIt's really true. Politically, I am a progressive (in US terms). The way these companies describe the problem and solution -- the model data reflects our culture back at us, so we are not giving anyone access to it -- is so nonsensical as to leave me wondering if there is a "real" reason, or is that just how they think?
- blululu 4y agoCheck the authors' names before calling them names.