6 ms·
OpenAI wouldn't be here without the work that Yann Lecun did at Facebook (back when it was facebook). Science is built on top of science, that's just how things
by tedivm 2y ago
OpenAI wouldn't be here without the work that Yann Lecun did at Facebook (back when it was facebook). Science is built on top of science, that's just how things work.
- wrasee 2y agoYes, but in science you reference your work and credit those who came before you. Edit: I am not defending OpenAI and we are all enjoying the irony here. But it puts into perspective some of the wilder claims circulating that DeekSeek was able to somehow complete with OpenAI for only $5M, as if on a level playing field.
- bugglebeetle 2y agoLike all those papers with their long lists of citations OpenAI has been releasing?
- dkjaudyeqooe 2y agoThat's only in academia. The same thing happens in commerce, only there is no (official) credit given.
- tedivm 2y agoOpenAI has been hiding their datasets, and certainly haven't credited me for the data they stole from my website and github repositories. If OpenAI doesn't think they should give attribution to the data they used, it seems weird to require that of others. Edit: Responding to your edit, Deepseek only claimed that the final training run was $5m, not that the whole process caught that (they even call this out). I think it's important to acknowledge that, even if they did get some training data from OpenAI, this is a remarkable achievement.
- ambicapter 2y agoOnly weird if you think what OpenAI did should be the norm.
- wrasee 2y agoRight. I think many here are enjoying the Schadenfreude against OpenAI, but that hardly makes it right. It just makes it a race to the bottom.
- wrasee 2y agoIt is a remarkable achievement. But if “some training data from OpenAI” turns out to essentially be a wholesale distillation of their entire model (along with Llama etc) I do think that somewhat dampens the spirit of it. We don’t know that of course. OpenAI claim to have some evidence and I guess we’ll just have to wait and see how this plays out. There’s also a substantial difference between training of the entire internet and one that very specifically targets your competitor's products (or any specific work directly).
- Filligree 2y agoThat's $5M for the final training run. Which is an improvement to be sure, but it doesn't include the other training runs -- prototypes, failed runs and so forth.
- coliveira 2y agoIt is OpenAI that discredits themselves when they say that each new model is the result of hundreds of USD millions in training. They throw this around as it is a big advantage of their models.
- nicce 2y agoAnd the cost is based on the imaginary currency that Microsoft has given for them as Azure computing.
- blackeyeblitzar 2y agoIs that really true? If anything OpenAI was dependent on the transformers paper from Google from Ashish Vaswani and others. LeCun has been criticizing LLM architectures for a long time and has been wrong about them for a long time.
- mv4 2y agoThat was my impression too. He is considered the inventor of CNN back in 1998. Is there anything more recent that's meaningful?
- blackeyeblitzar 2y agoPersonally, I have not seen anything from him that is meaningful. OpenAI and Anthropic (itself started by former OpenAI people) of course have built their models without LeCun’s contributions. And for a few years now, LeCun has been giving the same talk anywhere he makes appearances, saying that large language models are a dead end and that other approaches like his JEPA architecture are the future. Meanwhile current LLM architecture has continued to evolve and become very useful. As for the misuse of the term “open source”, I think that really began once he was at Meta, and is a way to use his fame to market Llama and help Meta not look irrelevant.
- tedivm 2y agoThey literally cited LeCun in their GPT papers.
- tedivm 2y agoI was more referring to this paper from 2015: https://scholar.google.com/citations?view_op=view_citation&hl=en&user=WLN3QrAAAAAJ&citation_for_view=WLN3QrAAAAAJ:lo0OIn9KAZgC https://scholar.google.com/citations?view_op=view_citation&h... Basically all LLM can trace their origin back to that paper. This was just a single example though. The whole point is that people build on the work from the past, and that this is normal.
- 2y ago
- zbendefy 2y agoAlso without the "attention is all you need" paper from google