8 ms·
Is this also censored/nerfed? I'd love to play with a "raw" unnerfed model to fully grasp what an LLM can do (and see how biased it is). Does anyone have any re
by gnrlst 3y ago
Is this also censored/nerfed? I'd love to play with a "raw" unnerfed model to fully grasp what an LLM can do (and see how biased it is). Does anyone have any recommendations for unnerfed models to try out?
- lioeters 3y agohttps://huggingface.co/models?search=uncensored&sort=trending https://huggingface.co/models?search=uncensored&sort=trendin...
- cubefox 3y agoThe most powerful available foundation model is code-davinci-002, a.k.a. GPT-3.5. It's only available on Azure since OpenAI removed it from their own Playground and API for some reason.
- RobotToaster 3y agoThe model isn't available at all?
- cubefox 3y agoIt is available in the sense that it is accessible. The weights are not available for download of course, but the OP wanted to "play around" with it, for which only access is required. There is no other accessible foundation model that can compete with GPT-3.5.
- cubefox 3y agoWhy are you guys downvoting me?
- brucethemoose2 3y agoBecause GPT 3.5 not very good compared to LLaMA 65b or even 33b finetunes, from my testing. Also because 3.5 is not really available?
- cubefox 3y agoHave you actually tested code-davinci-002?
- alach11 3y agoMaybe you mean gpt-3.5-turbo or text-davinci-003? Or GPT-4 (technically in beta so not fully available to everyone)?
- cubefox 3y agoNo, those are all fine-tuned models which are "nerfed" in the terminology of the OP. I mean code-davinci-002, the GPT-3.5 base model.
- squeaky-clean 3y agoIs that what nerfed means? I usually see "nerfed" used in a way that means that it will refuse to answer certain topics. "I can't answer that as it would violate copyright" and such.
- cubefox 3y agoThe fine-tuned models are certainly censored and not "raw".
- squeaky-clean 3y agoBut doesn't code-davinci-002 also have OpenAI's filters in between you and the model?
- cubefox 3y agoYes, but that's different from the model itself being fine-tuned.
- seanhunter 3y agocode-davinci models are finetuned on code so I don't think that's what the OP wants. For reference the family tree is here https://platform.openai.com/docs/model-index-for-researchers https://platform.openai.com/docs/model-index-for-researchers
- seanhunter 3y agoAll 3 text-davinci models are available on openAI's api. including 3 (which is the GPT-3.5 gen). Code-davinci-002 is a code-tuned model, You can see a nice visual summary of the relationships between the openAI models at https://yaofu.notion.site/How-does-GPT-Obtain-its-Ability-Tracing-Emergent-Abilities-of-Language-Models-to-their-Sources-b9a57ac0fcf74f30a1ab9e3e36fa1dc1 https://yaofu.notion.site/How-does-GPT-Obtain-its-Ability-Tr... Or the official source is https://platform.openai.com/docs/model-index-for-researchers https://platform.openai.com/docs/model-index-for-researchers
- cubefox 3y ago> All 3 text-davinci models are available on openAI's api. That's irrelevant because these are all fine-tuned. > Code-davinci-002 is a code-tuned model No, "code-tuned" isn't even a thing. It is a foundation model, which consists purely of pretreating. No fine-tuning is involved. > Or the official source is The official source says exactly what I just said.
- seanhunter 3y agoOK perhaps I used slightly the wrong term. The docs[1] say that code-davinci-002 is "optimized for code completion tasks" though so it seems unlikely to fulfil the OPs purpose of playing around with an unaligned/sweary model which was my main point. Some of the uncensored models from huggingface would probably serve that purpose much better. [1] see the entry for code-davinci-002 in https://platform.openai.com/docs/models/gpt-3-5 https://platform.openai.com/docs/models/gpt-3-5
- cubefox 3y agoCode was just part of its pretraining. All other GPT-3.5 models are fine-tuned versions of code-davinci-002. Quote: 1 code-davinci-002 is a base model, so good for pure code-completion tasks 2 text-davinci-002 is an InstructGPT model based on code-davinci-002 3 text-davinci-003 is an improvement on text-davinci-002 4 gpt-3.5-turbo-0301 is an improvement on text-davinci-003, optimized for chat Quote end. https://platform.openai.com/docs/model-index-for-researchers https://platform.openai.com/docs/model-index-for-researchers The reason you want a base model for code completion has nothing to do with code itself, it has to do with the fact that it completes text unlike all the instruction tuned models, which expect instructions. When you have code, there aren't necessarily any instructions present. You basically want autocomplete. That's what a base model does. But that doesn't mean it doesn't work with other things apart from code. After all, all other GPT-3.5 models are just code-davinci-002 with additional instruction and RLHF fine-tuning added, and they know countless other subject areas apart from code. I don't get why this is so hard to understand.
- logicchains 3y agoLLaMA 65B is the best uncensored model we've got, and the Airoboros fine-tuning if you want it to follow instructions.