7 ms·
Have a question to the Generative AI experts here. So, I can use smthg like GPT-4 to label data and then use that as a train set for my own LLM, right? EDIT:
by artembugara 3y ago
Have a question to the Generative AI experts here.
So, I can use smthg like GPT-4 to label data and then use that as a train set for my own LLM, right?
EDIT: adding this from OpenAI Restriction TOS: "(iii) use output from the Services to develop models that compete with OpenAI;"
- montenegrohugo 3y agoYup, totally. This is a form of knowledge distillation. Openai, or other foundational model providers, can't really do anything about it.
- cookieperson 3y agoWell they can sue you and bankrupt you by delaying trial for a decade. That's how the US patent system works anyways...
- sanxiyn 3y agoSue on what grounds? It will be quickly dismissed.
- mind-blight 3y agoThat is not how the US legal system works. You can sue someone for and regardless of merit, and they will have to defend themselves. That costs time and legal fees. If they lose, they can appeal, and continue appealing. If it's baseless, they'll lose, but you still spend a lot of money and time dealing with the lawsuits.
- cookieperson 3y agoIf by quickly you mean 5 to 10 years of paying a retainer on a lawyer sure. Even if you win the case you lose in life. Most individuals can't afford 500k in legal fees to have them be reimbursed years later. Big companies have lawyers on staff at a discount and they play these games every day. This happens with illegal things all the time. IE manager sexually harasses someone on video or something, it's some CEOs nephew who did it, so they fire the person who got harassed. The person who got harassed now has to aquire legal counsel on top of paying relocation claw backs etc. Few years ago by and the person who was in the right is trying to hold down a job, a family, and the stress of the legal battle. The company offers to settle two years in for 50k and 99% of people take it, sometimes at a loss. Also, getting employed is a lot harder when a background check reveals suing a previous employer or really any company, because shocker, most companies do illegal shit regularly... So it's almost always best to settle I realize I painted a picture pretty far from my previous statement but I figured you were new in your career and could benefit from an allegory of how stuff like this goes down.
- wodenokoto 3y agoIt is my understanding that this is how “alignment” works. That is, openAI paid people to chat with their LLM to fine tune it and then other LLMs use chatgpt to generate training data to align their models.
- visarga 3y agoThere are three ways 1. make your own RLHF dataset - like OpenAI and Open Assistant 2. exfiltrate data from a bigger/better LLM - Vicuna & family 3. use your pre-trained LLM to generate RLAIF data, no leeching - ConstitutionalAI, based on a set of rules instead of labelling examples
- cubefox 3y agoI wonder whether these approaches fit into the above categories: https://arxiv.org/abs/2305.13735 https://arxiv.org/abs/2305.13735 https://arxiv.org/abs/2305.11206 https://arxiv.org/abs/2305.11206
- snickmy 3y agoIndeed, fine tuning with either synthetic data (as you are proposing) or human review works like that. you can read more here: https://huggingface.co/blog/rlhf https://huggingface.co/blog/rlhf
- fallingmeat 3y agoThat is against their ToS though if you use your new LLM commercially.
- pmoriarty 3y agoSo what are they going to do about it?
- jstummbillig 3y agoThat escalated quickly.
- fallingmeat 3y agoGreat question! I don’t know the end game there. Maybe if they suspected their model was used they would sue, and in discovery find you used their model for training?
- visarga 3y agoMaybe we don't need to worry, OpenLLaMA is under training right now. It will be the commercial version of LLaMA. > Update 05/22/2023 > We are happy to release our 700B token checkpoint for the OpenLLaMA 7B model and 600B token checkpoint for the 3B model. We’ve also updated the evaluation results. We expect the full 1T token training run to finish at the end of this week. https://github.com/openlm-research/open_llama https://github.com/openlm-research/open_llama So we could develop on LLaMA for now and switch to OpenLLaMA later.
- bottled_poe 3y agoNothing until it’s worth their while.
- postsantum 3y agoMS lawyers have a good track record at sending out those scary cease&desist letters
- 3y ago
- foobarbecue 3y agoIs "ca" "can" or "can't"?
- artembugara 3y agocan
- notpublic 3y agonot an AI expert but from a talk I recently heard... if there is a mismatch in training data between the "teacher" LLM and "student" LLM, you risk teaching the student to hallucinate or to ignore information
- moffkalast 3y ago> I can use smthg like GPT-4 to label data and then use that as a train set for my own LLM, right? Yes, almost all improved LLama models are tuned exactly that way (trained on examples of questions and answers from say GPT 4). If OpenAI stole copyrighted works to train their models it is morally fair game to do the same to them regardless of their TOS. It's not like they can prove it anyway. Plus there's the other point where they also say that everything generated by their models is public domain, so which one is it eh?
- sirsinsalot 3y agoThis ... but we all know business is corrupt. The current attempts to spur on regulation by OpenAI is moat building
- ehnto 3y agoWe were complacent while it happened because OpenAI wasn't a business, it wasn't seen as unethical to use community work to contribute to community research. Now they're entrenched and pulled the rug out from the community, whilst also trying to shut the door on anyone else. Just a really disappointing series of events, the money and profit were never the big issue.
- jrm4 3y agoI'm a lawyer so one should never break the law. Nonethless, I can observe and predict that non-consensual "open sourcing" of these models would likely end up probably the best and safest way to do all of this stuff.
- Fgehono 3y agoBecause by training it they created something new. I don't mind just making a point. But I don't think they mind. I don't believe that this type of model training is able to be bleeding edge which should guarantee that openai has enough motivation to continue the development and having a healthy competition
- sp332 3y agoIt's against the terms of service to do the generation, but the generated text is not copyrighted. Those are different things.
- chaxor 3y agoYes, and in fact that's the best method available if you want good performance. I would suggest using a local open source model to do this however, to cut down on costs and make it far simpler to deal with than the unwieldy OpenAI systems. https://arxiv.org/pdf/2305.02301.pdf https://arxiv.org/pdf/2305.02301.pdf