3 ms·
They didn't include the costs for developing v3, the base model. [edit] also they seem to be saying r1 is a base model, which it is not. Very sloppy.
by fzzzy 1y ago
They didn't include the costs for developing v3, the base model.
[edit] also they seem to be saying r1 is a base model, which it is not. Very sloppy.
- xnx 1y agoDidn't they also train off of ChatGPT API output?
- ipsum2 1y agoThis is a rumor that has not been confirmed by OpenAI nor DeepSeek.
- nextaccountic 1y agoI don't see the issue here. OpenAI trained ChatGPT off my own comments, and your comments, and the comments from the person you replied to.. I didn't authorize it and you probably didn't too. Meta was caught pirating over 80TB of books to train their AI, and they are claiming not only training AI on other people's stuff is legal, but piracy is also legal (well at least, piracy done by US tech giants is legal)
- xnx 1y agoFor sure. I was just pointing out that DeepSeek is not going to "beat" ChatGPT if DeepSeek relies on it.
- natrys 1y agoYou could maybe make that accusation about V3 (to the extent that it's a bad thing and not fair use, specially considering amoral origin of OpenAI's models in first place), but don't think the claim makes sense for R1 since OpenAI's o1 did not expose its CoT traces even in API. They published about GRPO (key algorithm behind R1) a full year before[1] they scaled it for R1. Given the research they do in open, it's not far-fetched to think they had the talent and technical know-how to achieve R1 on their own. [1] https://arxiv.org/abs/2402.03300 https://arxiv.org/abs/2402.03300
- joejoo 1y ago[dead]
- irjustin 1y agoYeah apparently they parked the cost of hardware, 50k GPUs and model development underneath another entity, high-flyer because it was "shared resource".
- AustinCarrBW 1y agoYou're misreading. The article is referring to V3 when it cites the base model behind R1. It does not say R1 is the base model.