Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Reubend
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
61.
▲
by
Reubend
6mo ago
It's a straightforward MIT license: https://github.com/calcom/cal.diy/blob/main/LICENSE > IN NO EVENT SHALL THE AUTHORS OR COPYRIGHT HOLDERS BE LIABLE FOR ANY CLAIM, DAMAGES OR OTHER LIABILITY, W
62.
▲
by
Reubend
6mo ago
They'll sign a contract, and the contract will be very clear about whether using user prompts as training data is allowed or not. They're not going to care much about reputation; they'll care about the terms they sign with.
63.
▲
by
Reubend
6mo ago
But the OSS license already absolves them of responsibility. This might just be to set the tone that security fixes won't be prioritized to the standard that they used to be.
64.
▲
by
Reubend
6mo ago
I think you're right. Other providers can offer coding subscriptions that use in-house models, and this sets the stage for a Grok coding plan that's built in to Cursor. $60 billion seems expensive, but it gives them a much better
65.
▲
by
Reubend
6mo ago
It's not just quantization. I verified that if you naïvely quantize to 1 bit from the original Qwen model (and set grouped scale factors based on what the original model's weights were like), it just spits out gibberish. > One
66.
▲
by
Reubend
6mo ago
Are you saying the original one worked with MPS? Or are you just saying it was always theoretically possible to build what OP posted?
67.
▲
by
Reubend
6mo ago
After trying to understand their method, I think you're right. Doesn't seem like anything that I would personally call "diffusion". Much closer to MTP + speculative decoding. Then again, their results with it are great.
68.
▲
by
Reubend
6mo ago
> Also, any reason to imply "BridgeBench", apparently dedicated to AI benchmarking, wouldn't have run it more than once across the suite? They didn't list a sample size of runs, didn't show any numbers for varian
69.
▲
by
Reubend
6mo ago
This page is buggy for me and doesn't show any plans. But when I go to their main pricing page, it's got some contradictory info about which plans include "Kimi Code". The $19 per month plan says that it comes with "
70.
▲
by
Reubend
6mo ago
Because the website doesn't seem to show any sample size of runs, I assume they ran it once across the suite. The models are nondeterministic, and therefore it's pretty normal for different runs to give different results. I don&#x
71.
▲
by
Reubend
6mo ago
Cool experiment! But the "CEO" agent picked the most boring possible items to sell: t-shirts and some bland art prints designed by AI. I would have loved to see more creativity given that they could have picked anything.
72.
▲
by
Reubend
6mo ago
> I don’t see UBI anywhere on OpenAI’s roadmap. Do you really think it's OpenAI's job to create UBI? Surely, if you feel that it's a good idea, then it should be the government who sets it up. We can't just magically
73.
▲
by
Reubend
6mo ago
> AI has already stolen great amounts of money from a very large number of people all around the world, due to the huge increases in the prices of DRAM, SSDs and HDDs. None of that is stealing. That's the free market.
74.
▲
by
Reubend
6mo ago
AI is the best thing that happened to America in the last decade, and I dearly hope that politicians don't try to ruin it the way they're ruining other parts of the country. I respect some of Bernie's positions, but his stanc
75.
▲
by
Reubend
6mo ago
"Legibility" must be the wrong word because I can't understand what the author is talking about. Is he saying that the overuse of abstractions is ruining corporate culture? Or is he saying that the uniformity of corporate pro
76.
▲
by
Reubend
6mo ago
I would suggest that people stop overfocusing on benchmarks, and give this a try. Gemma 4 is performing really well for me, and seems to hallucinate much less than other models I tried in this size range.
77.
▲
by
Reubend
6mo ago
This is pretty impressive. I think This sort of thing is a perfect fit for agentic coding because of the fact that you can compare the generated assembly afterwards as a safeguard/test. Plus, even if the code is messy, you can always a
78.
▲
by
Reubend
6mo ago
You're absolutely right. And these Intel GPUs will also be much faster in terms of actual math than the M series GPUs that the Apple setup would have.
79.
▲
by
Reubend
7mo ago
This looks great, but I'm wondering how effective this would be for full model weights rather than just the KV cache. Their paper only gives results for the KV cache use case, which strikes me as strange since the algos are claimed to
80.
▲
by
Reubend
7mo ago
This is really cool research, but I'm wondering how much it slows down inference. The readme says that it's "...distinguished by zero overhead (no learned components, no entropy coding)" but does that really mean that th
81.
▲
by
Reubend
7mo ago
It's a very interesting concept, and I signed up to try it. However, after seeing the landing page, my first question was: "Where's the data on accuracy?" Backtesting is difficult to do correctly with LLMs, but because t
82.
▲
by
Reubend
7mo ago
Seems like it does quite well on that particular benchmark?
83.
▲
by
Reubend
7mo ago
A mobile failover would be cheaper and would give you better connectivity in heavy rain. A 4G dongle can be purchased for $15, rather than $200 for a Starlink Mini. Then, let's say your main internet source fails and you need to actual
84.
▲
by
Reubend
7mo ago
Very interesting! Thank you for explaining this.
85.
▲
by
Reubend
7mo ago
The Qwen team has been putting out great releases lately. I hope that they can continue on that path despite this.
86.
▲
by
Reubend
8mo ago
> But what would be an example of an uncomputable number? That’s a good question. Most obviously, we could be talking about numbers that encode the solution to the halting problem. It would lead to a paradox to have a computer program th
87.
▲
by
Reubend
8mo ago
If that's the case, then their API should return an error. Billing the user while serving a response from the wrong model is a horrible outcome. I'd go as far as to say that it's borderline fraudulent.
88.
▲
by
Reubend
8mo ago
I'm not quite sure what to make of this. Is it a joke, a serious paper, or more of a poem? How much of this did you use LLM assistance for? It's so dense, with tons of detail and yet no useful explanation of any of the contents. Y
89.
▲
by
Reubend
8mo ago
The post actually has great benchmark tables inside of it. They might be outdated in a few months, but for now, it gives you a great summary. Seems like Gemini wins on image and video perf, Claude is the best at coding, ChatGPT is the best
90.
▲
by
Reubend
8mo ago
I've read several people say that Kimi K2 has a better "emotional intelligence" than other models. I'll be interested to see whether K2.5 continues or even improves on that.
More ›