Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
coder543
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
19 ms
·
511.
▲
by
coder543
3y ago
Looks like TheBloke has released GGUFs of this Dolphin fine-tune: https://huggingface.co/TheBloke/dolphin-2_6-phi-2-GGUF There seem to be a few Phi-2 fine-tunes floating around. This is another one I've seen: htt
512.
▲
by
coder543
3y ago
I assume it queues up requests if there are too many in flight to handle all of them at the same time, but it shows you the tokens per second (T/s) for every response, which is the number that matters (and presumably won't include
513.
▲
by
coder543
3y ago
There is quite literally a modal pop-up that explains it, which you must dismiss before you can begin interacting with the demo. Quoting the pop-up: "This alpha demo lets you experience ultra-low latency performance using the foundatio
514.
▲
by
coder543
3y ago
It’s just running bog standard Llama2-70B by all appearances. I don’t know why so many people here are interested in the outputs. The whole point of this demo is that the company is trying to show off how fast their hardware could host one
515.
▲
by
coder543
3y ago
Is there any plan to show what this hardware can do for Mixtral-8x7B-Instruct? Based on the leaderboards[0], it is a better model than Llama2-70B, and I’m sure the T/s would be crazy high. [0]: https://huggingface.co/sp
516.
▲
by
coder543
3y ago
What you're describing is the behavior you get from any base model that has not been instruction-tuned. The article is clear that this model is not for "direct use". It needs tuning for a specific application.
517.
▲
by
coder543
3y ago
I haven't used the llama2 models much in quite awhile, because they just aren't very good compared to other options that exist at this point. The instruction-tuned variants of Mistral and Mixtral seem to have very little trouble r
518.
▲
by
coder543
3y ago
Base models are just trying to autocomplete the input text. The most logical completion for an instruction is something approximately like what you asked, but base models are raw. They have not been taught to follow instructions, so they ge
519.
▲
by
coder543
3y ago
They released a base model. It is not instruction-tuned, so it won't really follow instructions unless you fine-tune it to do that. "There are lots of Mistral fine-tunes. Why another one? A very healthy ecosystem of Mistral fine-t
520.
▲
by
coder543
3y ago
The article shows (fine tuned) Mistral 7B outperforming GPT-4, never mind GPT-3.5.
521.
▲
by
coder543
3y ago
PowerPoint existed in the late 80s, I think, although Microsoft acquired it from what I understand.
522.
▲
by
coder543
3y ago
"Power*" made me think of Microsoft, so I was almost expecting this to be Windows-specific. (PowerShell, PowerPoint, Power BI, Power Apps, Power Automate... I'm probably forgetting some.)
523.
▲
by
coder543
3y ago
Why filter out the votes made after only one or two prompts? A lot of times, a single response is all you need to see. Do you really need more than this to know which one you’re going to pick? https://i.imgur.com/En37EJD.png
524.
▲
by
coder543
3y ago
GPT-4 apparently shows a small bias (10%) towards itself in the paper, and GPT-3.5 apparently did not show any measurable bias towards itself. Given the possibility of bias, it would make sense to have the judge “recuse” itself from compari
525.
▲
by
coder543
3y ago
Mixtral is missing in half of the benchmarks in that paper. Hardly conclusive. It’s also common knowledge that these benchmarks have a lot of issues[0]. A good litmus test, but not a substitute for actually seeing how the models do in the r
526.
▲
by
coder543
3y ago
Mixtral ranks higher than Gemini Pro on the (subjective) Chatbot Arena Leaderboard: https://huggingface.co/spaces/lmsys/chatbot-arena-leaderboar... Where are you seeing that it is "further behind Gemini Pro t
527.
▲
by
coder543
3y ago
But, what if you could make an SAT that is equivalent to evaluating years of performance at work? https://huggingface.co/papers/2306.05685 This paper makes the argument that... "Our results reveal that strong LLM
528.
▲
by
coder543
3y ago
I wish that Arena included a few more "interesting" models like the new Phi-2 model and the current tinyllama model, which are trying to push the limits on small models. Solar-10.7B is another interesting model that seems to be mi
529.
▲
by
coder543
3y ago
Dupe: https://news.ycombinator.com/item?id=38682631
530.
▲
by
coder543
3y ago
Two components do not make a product. The SBC market is a better litmus test for the real costs. There is plenty of competition making products of all kinds.
531.
▲
by
coder543
3y ago
If the recent revelations in the Epic vs Google court case are anything to go by, Motorola is likely getting paid by Google for every single Google search and Google Play Store transaction that occurs on that phone. It could even be sold at
532.
▲
by
coder543
3y ago
Of course most users don’t pick the 65W option. They want maximum performance, and the cost of electricity is largely negligible to most people buying a 7950X. AMD isn’t going to offer a huge discount for a 65W 7950X for the reasons discuss
533.
▲
by
coder543
3y ago
Every 7950X offers a 65W mode. It’s not a separate SKU. It’s a choice each user can make if they care more about efficiency. Tasks take longer to complete, but the total energy consumed for the completion of the task is dramatically less.
534.
▲
by
coder543
3y ago
This page might be somewhat helpful: https://cloud.google.com/vertex-ai/docs/generative-ai/image/... It also includes a link to the TTP form, although the form itself seems to make no reference to Imagen
535.
▲
by
coder543
3y ago
> 96GB of weights. You won't be able to run this on your home GPU. This seems like a non-sequitur. Doesn't MoE select an expert for each token? Presumably, the same expert would frequently be selected for a number of tokens in
536.
▲
by
coder543
3y ago
That row says lower is better. For "word error rate", lower is definitely better. But they also used Large-v3, which I have not ever seen outperform Large-v2 in even a single case. I have no idea why OpenAI even released Large-v3.
537.
▲
by
coder543
3y ago
Do you have Bard history enabled?
538.
▲
by
coder543
3y ago
You must enable Bard history for these features to work, and you must go to the extensions page and make sure they’re turned on. Arguing with a model that doesn’t have access to the extensions won’t make it suddenly use the extensions. Su
539.
▲
by
coder543
3y ago
Ok, that makes more sense, but if people haven't migrated away by now... the odds seem increasingly likely that they won't migrate in time to avoid that deadline.
540.
▲
by
coder543
3y ago
Was there actually a price increase recently and not 6 months ago? If people just want to vent about Hashicorp/Terraform, it seems like a text-post would be sufficient for that.
More ›