Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Me1000
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
31.
▲
by
Me1000
2y ago
This is the real value here. Keeping a secure environment to run untrusted code along side user data is a real liability for them. It's not their core competency either, so they can just lean on browser sandboxing and not worry about i
32.
▲
by
Me1000
2y ago
They explain how it's calculated, you just have to trust their calculations are correct.
33.
▲
by
Me1000
2y ago
Ah, thank you!
34.
▲
by
Me1000
2y ago
Relatedly, what does "parallel" function calling mean in this context?
35.
▲
by
Me1000
2y ago
Or it means that the existing models are too large to be profitable even at scale.
36.
▲
by
Me1000
2y ago
I think there is valid criticism of google for inventing a cool technology only to have the rest of the industry discover its usefulness before them. But to say Gemini 1.0 or OG Gemma aren't first generation models because BERT and fla
37.
▲
by
Me1000
2y ago
>long time ago This is an incredible statement to make about a field that no one was talking about 24 months ago, a family of SOTA models that didn't exist until 8 months ago, and a family of small local models that didn't exis
38.
▲
by
Me1000
2y ago
I also use BART most of the time, but I often end up heading back to SF pretty late in the night so that’s not always an option.
39.
▲
by
Me1000
2y ago
Unfortunately Waymo still wont operate on freeways or at SFO. But I long for the day when I can take a Waymo between SF and Berkley or to SFO, those are my most frequent reasons to call a Lyft.
40.
▲
by
Me1000
2y ago
> Rides have usually 10% mark up over Lyft and Uber. I found that to be true as well, but when you factor in the tip to the driver they come out to more or less the same price.
41.
▲
by
Me1000
2y ago
“Just” is doing a lot of heavy lifting here. It’s so frustrating watching people think it’s just because a bunch of smart and hard working people are either lazy or stubborn.
42.
▲
by
Me1000
2y ago
Many social media platforms exacerbate and amplify some of the worst qualities of humans. Bullying is taken to an extreme, people are often engaging with others in bad faith, and other times you're not even engaging with real people at
43.
▲
by
Me1000
2y ago
Exactly. People forget (or simply don't know) that orbital plane changes are incredibly expensive from an energy perspective. It's not like you're walking down the road picking up trash in a row, you have to adjust your orbit
44.
▲
by
Me1000
2y ago
They prompt you before they send your data to OpenAI, but it's clear that they prompt you before they send it to Apple's servers (maybe they do and I missed it?). And their promise that their servers are secure because it's a
45.
▲
by
Me1000
2y ago
This arguments is made every singe time a new LLM article gets posted and I'm not sure it's really adding anything to the conversation. Everyone understands language models are not human, there's no need to add a philosophica
46.
▲
by
Me1000
2y ago
Completely unrelated.
47.
▲
by
Me1000
2y ago
He actually was a joy to work with. George is an awesome human.
48.
▲
by
Me1000
2y ago
This is an opt-in feature, so by defaut all AI-related features are disabled. You have to go out of your way to generate an API key and add it, it's not even a thing you could accidentally turn on. And if you'd like to use the fea
49.
▲
by
Me1000
2y ago
Thanks to most of the world being "GPU poor", there is a lot of research and engineering effort going into making models much more compute efficient. Another way that OpenAI gets to benefit from the world of open source/weigh
50.
▲
by
Me1000
2y ago
LLMs are weird, they hype cycle has caused a lot of people to want to hate on them. A tool like any other can be misused or abused, and smart people like to demonstrate their ability to break the tool and show the world so that people can s
51.
▲
by
Me1000
2y ago
And Gemini.
52.
▲
by
Me1000
2y ago
Ollama doesn't have their own inference engine, they just wrap llama.cpp. But yes, it will be awesome when it's more generally available.
53.
▲
by
Me1000
2y ago
WizardLM-2 8x22b (which was a fine tune of the Mixtral 8x22b base model) at 4bit was only 80GB.
54.
▲
by
Me1000
2y ago
You need to update ollama to 0.1.32.
55.
▲
by
Me1000
2y ago
I ended up having to download the latest version directly from GitHub, and that fixed it. Looks like the 0.1.32 mac release hasn't been posted to your website yet.
56.
▲
by
Me1000
2y ago
The 4bit quant doesn't seem to work for me, I keep getting: Error: exception create_tensor: tensor 'blk.0.ffn_gate.0.weight' not found I've tried downloading it twice now.
57.
▲
by
Me1000
2y ago
For the RAG example, I don’t think it’s the prompt so much. Or if it is, I’ve yet to find a way to get GPT4 to ever extrapolate well beyond the original source text. In other words, I think GPT4 was likely trained to ground the outputs on a
58.
▲
by
Me1000
2y ago
Interesting, Claude 3 Opus has been better than GPT4 for me. Mostly in that I find it does a better (and more importantly, more thorough) job of explaining things to me. For coding tasks (I'm not asking it to write code, but instead to
59.
▲
by
Me1000
2y ago
Mixtral 7x8b was way better than llama2 70b and used less RAM and compute at the same time. This model is way better than llama. In fact I would go as far as saying llama2 isn’t that good compared to some of the most recent models.
60.
▲
by
Me1000
3y ago
It’s absolutely beneficial when training because the forward pass and back propagation is still only on the neurons that were activated. The Mistral guys specifically mention that training speed (due to not needing as much compute) was one
More ›