Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
DetroitThrow
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
DetroitThrow
7d ago
It would be amazing to have big BERTha with per-token pricing on GCP or AWS. There are many times I am reaching for a cheap classifier with the general behavior of an LLM.
2.
▲
by
DetroitThrow
10d ago
I've switched off to runners and event scheduling inside AWS, and am working on moving my company off of GitHub. I know engineers inside GitHub and it doesn't sound like the talk they've been putting out about reliability is
3.
▲
by
DetroitThrow
17d ago
Yes exactly. Automatic model downgrade seems horrible for a lot of production workloads, even if you are deterministically constraining the behavior of your agents.
4.
▲
by
DetroitThrow
18d ago
I don't think the person describing the paper by Buckmaster as the same as the 2023 paper by Córdoba and Martínez-Zoroa is really discussing this in good faith fwiw. There are some massive advancements within it and if the person was p
5.
▲
by
DetroitThrow
18d ago
>It's unknowable and not possible to prove if any one specific conversation was the key to solving Navier–Stokes. If the conversation was in the training set, there's a high likelihood that the small set of conversations relate
6.
▲
by
DetroitThrow
18d ago
It's unclear if you're suggesting that OpenAI did not train on their input or use their chats as inputs to training on a model that found the solution. Let's not provide an Elizabeth Holmes-esque interview where the question
7.
▲
by
DetroitThrow
18d ago
Given that he has other former collaborators corroborating this horrific behavior, it seems like this a career spanning pattern, and it's interesting to see just how much @sama is willing to lend his support to someone like Bubeck. Sta
8.
▲
by
DetroitThrow
23d ago
404?
9.
▲
by
DetroitThrow
29d ago
Can't check whether star citizen is done yet :/
10.
▲
by
DetroitThrow
29d ago
I don't mind color, the glow effects just hurt the ability to read what's there (this reminds me of the gaming setups teenagers are drawn to)...but there's no reason a utility like a task manager should be closed source or ha
11.
▲
by
DetroitThrow
29d ago
Wow, thought the guy would be a MSFT OG but this looks like it sucks? Vibed, closed source, paid extensions.. Why on earth would anyone use this?
12.
▲
by
DetroitThrow
1mo ago
It's a bit of both for Anthropic I think, sometimes cutting edge and quite interesting or just good improvements, sometimes ignoring best practices either recently established or known for decades. Obvious to see where the smart people
13.
▲
by
DetroitThrow
1mo ago
Lately I've been throwing tasks at Qwen and a frontier or recently-frontier model (as well as Kimi, GLM, etc) and the smaller parameter models are not really comparable to Opus when it comes to making intelligent decisions about greyer
14.
▲
by
DetroitThrow
2mo ago
It's still not as good as GPT5.6 or Opus5 but it's better than KimiK3. Good job xAI team.
15.
▲
by
DetroitThrow
2mo ago
I've tried it on my "let's run every model in parallel and see which finds more edge cases" type of tasks, and Grok 4.5 was really behind Opus/ChatGPT but ahead of Gemini - despite having a strong showing on benchma
16.
▲
by
DetroitThrow
2mo ago
>Aren't the LLMs trained on a massive corpus of human written texts? If that stands, then they are doing what they were asked, kind of? I think you might be interested in reading in training data generation, training, and post train
17.
▲
by
DetroitThrow
2mo ago
This blogpost didn't really address the elephant in the room that increasingly the bottleneck is no longer the part that jr devs could help out with. Every startup I know that hired some level of jr's and encouraged them to use AI
18.
▲
by
DetroitThrow
2mo ago
They also hike the price so significantly most people stop using it. See meetup.com for example.
19.
▲
by
DetroitThrow
2mo ago
yes, _every_ event I go uses Luma or something else nowadays.
20.
▲
by
DetroitThrow
2mo ago
How many times have they managed to catch it in the last 8 or so missions? There have been a few misses like this already with the boosters
21.
▲
by
DetroitThrow
2mo ago
>FFCS is called holy grail of liquid engines Particularly for reusable liquid engines :)
22.
▲
by
DetroitThrow
3mo ago
Looks like they reset everyone's Fable usage.
23.
▲
by
DetroitThrow
3mo ago
DeepSWE seems to strongly, strongly prefer ChatGPT models. There were also major flaws in its methodology pointed out recently, that overlap strongly with the flaws OpenAI pointed out in its SWE Verified report. I use both ChatGPT and Claud
24.
▲
by
DetroitThrow
3mo ago
>I'll be more peeved if they monetize it FSL (vs a copyleft license or just plain old OSS) implies they want to turn this into a revenue source for themselves ultimately, unfortunately. >Maybe I should put one of those buy me a c
25.
▲
by
DetroitThrow
3mo ago
On top of that, it has a more restrictive license than AmazonBrandFilter. Given this appears to be a very simple AI project, why not just reimplement any missing functionality from AmazonBrandFilter into something under a free license? The
26.
▲
by
DetroitThrow
3mo ago
She's not as big on some of the broader interpretations of the 4th amendment that more civil liberty minded justices would lend credence to.
27.
▲
by
DetroitThrow
3mo ago
He's entitled to his political views and just as we're entitled to potentially use or not use his service because of them :) Not sure why it's such an issue to discuss the political views of the beneficiaries of services we u
28.
▲
by
DetroitThrow
3mo ago
Everyone gets to share but it's also completely within the forum rules to call out irrelevant anecdotes as uninteresting to the discussion. I have no idea why you're making a comparison to a TV show; nothing that was described was
29.
▲
by
DetroitThrow
3mo ago
When performance isn't a concern, I largely agree! Not every financial system can use big decimal as their base, though, too. And HFT isn't the only place in the financial sector where this performance concern might pop up.
30.
▲
by
DetroitThrow
3mo ago
"10% of Americans are uninsured. A US state is pushing to insure all of their residents." "I'm insured!" "Open-source software projects are being spammed with LLM generated PRs. Contributions are becoming more
More ›