4 ms·
how likely are you to start work on a project depending on an ecosystem of wildly overstated capabilities?
by solarpunk 3y ago
how likely are you to start work on a project depending on an ecosystem of wildly overstated capabilities?
- danielmarkbruce 3y ago100% likely. Several projects, right now. GPT-4 is the best model currently available. There are reasons why it's better to control a model and host yourself etc etc, but there are also reasons to use the best model available.
- solarpunk 3y ago... I'm not trying to be rude, but do you think maybe you have bought into the purposely exaggerated marketing?
- tempusalaria 3y agoGPT-4 is the best model though… the gap has closed a lot but it’s still the best I despise openai but I can’t really argue with that
- danielmarkbruce 3y agoThat's not how people who actually build things do things. They don't buy into any marketing. They sign up for the service and play around with it and see what it can do.
- kjkjadksj 3y agoDepends on the scale of the job. Sometimes you wake up and your employer is already paying for both google drive as well as one drive and drop box at the same time and IT is replacing the room av for the third time this year.
- sjwhevvvvvsj 3y agoThe definition of “best” has a lot of factors. Best general purpose LLM chat? I’d agree there, but there’s so much more to LLM than chat applications. For some tasks I’m working on, Mixtral is the “best” solution given it can be used locally, isn’t hampered by “safety” tuning, and I can run it 24x7 on huge jobs with no costs besides the upfront investment on my GPU + electricity. I have GPT-4 open all day as my coding assistant, but I’m deploying on Mixtral.
- danielmarkbruce 3y agoYup, plenty of reasons to run your own model. I'm not using GPT-4 for chat, but for what I'd class as "reasoning" applications. It seems best by a long shot. As for safety, I find with the api and the system prompt that there is nothing it won't answer for me. That being said... I'm not asking for anything weird. GPT-4 turbo does seem to be reluctant sometimes.
- sjwhevvvvvsj 3y agoI’m doing document summarization and classification, and there’s a fair amount it won’t do for prudish reasons (eg sex, porn, etc). Llama2 is basically useless in that regard.
- danielmarkbruce 3y agowith GPT-4 (non-turbo) and a good system prompt?
- sjwhevvvvvsj 3y agoIt’s unpredictable, usually it behaves but on arbitrary web text you’ll eventually trigger a “safety” barrier, and sometimes just trying again will work.