Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
deepsquirrelnet
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
deepsquirrelnet
5mo ago
This is really cool. Any plans to release the dataset?
32.
▲
by
deepsquirrelnet
5mo ago
Any discussion related to this topic always seems to assume everyone uses code the same way and for the same function, and then forces the rest of the world through that lens. So here we walk around the circle one more time again, voicing o
33.
▲
by
deepsquirrelnet
5mo ago
I think it means which site is numberwang.
34.
▲
by
deepsquirrelnet
5mo ago
> At the time of the publication, Meta admitted subcontracted workers might sometimes review content filmed on its smart glasses when people shared it with Meta AI. They just got fired for "piercing the veil". They committed th
35.
▲
by
deepsquirrelnet
5mo ago
And now debt is 1 GDP year. I don’t see the problem.
36.
▲
by
deepsquirrelnet
5mo ago
I would love to be able to run frontier locally, but I think the larger importance of open weight models is price accountability. In the US with our broken system of capitalism, it’s the only way we can tether these companies to reality. Le
37.
▲
by
deepsquirrelnet
5mo ago
This is my own take, directly related to this that I posted a little while back. The one thing that I think the article missed is the geopolitical angle they’re also working: * We need to completely deregulate these US companies so China do
38.
▲
by
deepsquirrelnet
5mo ago
> In a provocative GitHub post, machine-learning engineer Han-Chung Lee argued that even rosy internal numbers that do show AI-assisted productivity gains are suspect, as they’re produced to hit adoption targets no one can effectively au
39.
▲
by
deepsquirrelnet
6mo ago
> Traditional call noise canceling relies on those small onboard neural networks and can have difficulty isolating your voice in very noisy environments, which results in ambient noise leaking through or voices getting highly compressed,
40.
▲
by
deepsquirrelnet
6mo ago
I tried it on openrouter and set max tokens to 8192, and every response is truncated, even in non-thinking mode. Maybe there's an issue with the deployment, but in your link also shows it generates tons of output tokens.
41.
▲
by
deepsquirrelnet
6mo ago
My tinfoil hat theory, which may not be that crazy, is that providers are sandbagging their models in the days leading up to a new release, so that the next model "feels" like a bigger improvement than it is. An important aspect o
42.
▲
by
deepsquirrelnet
6mo ago
> Cuyahoga Valley: There is nothing wrong with Cuyahoga Valley. Statistically, you’re from Ohio, so why not? In college, I took an interim elective course on geology of the national parks. On the first day of class, the professor asked a
43.
▲
by
deepsquirrelnet
6mo ago
CnakeCharmer - https://github.com/dleemiller/CnakeCharmer https://huggingface.co/datasets/CnakeCharmer/CnakeCharmer This project started from a belief that llms should be better at doing pyt
44.
▲
by
deepsquirrelnet
6mo ago
There are so many reason if you look at how it's being sold. * We need to completely deregulate these US companies so China doesn't win and take us over * We need to heavily regulate anybody who is not following the rules that m
45.
▲
by
deepsquirrelnet
6mo ago
I am working on a large scale dataset for producing agent traces for Python <> cython conversion with tooling, and it is second only to gemini pro 3.1 in acceptance rates (16% vs 26%). Mid-sized models like gpt-oss minimax and qwen3.5
46.
▲
by
deepsquirrelnet
7mo ago
Good article, and I think the "evolution of every AI system" is spot on. In my opinion, the reason people don't use DSPy is because DSPy aims to be a machine learning platform. And like the article says -- this feels differen
47.
▲
by
deepsquirrelnet
7mo ago
This is even after the Hindenburg research report that found numerous screaming red flags a few years ago. https://hindenburgresearch.com/smci/
48.
▲
by
deepsquirrelnet
7mo ago
I worked at Micron in the SSD division when Optane (originally called crosspoint “Xpoint”) was being made. In my mind, there was never a real serious push to productize it. But it’s not clear to me whether that was due to unattractive terms
49.
▲
by
deepsquirrelnet
7mo ago
Interesting read! I love to see this spirit. I grew up with a different - but similar experience. Only, as an 80s and 90s kid, computers were nothing but limitations. Even when my dad built a machine with a 133MHz Cyrix chip, already a year
50.
▲
by
deepsquirrelnet
7mo ago
> I tapped into Pangram. Pangram is a remarkably good, conservative model for detecting LLM-generated text. These detectors have a bad rep among techies, but the objections are often based on outdated assumptions Turing test is really in
51.
▲
by
deepsquirrelnet
7mo ago
The title being misleading is important as well, because this has landed on the front page, and the only thing that would be the only notable part of this submission. The "new" on huggingface banner has weights that were uploaded
52.
▲
Show HN: FizzBuzz Forever – Agent Edition
(github.com)
1 points
by
deepsquirrelnet
7mo ago
|
0 comments
53.
▲
by
deepsquirrelnet
7mo ago
Zero-shot encoder models are so cool. I'll definitely be checking this out. If you're looking for a zero-shot classifier, tasksource is in a similar vein. https://huggingface.co/tasksource/ModernBERT-large-nli
54.
▲
by
deepsquirrelnet
7mo ago
Does this use something like xnnpack under the hood?
55.
▲
by
deepsquirrelnet
7mo ago
4-bit quantization on newer nvidia hardware is being supported in training as well these days. I believe the gpt-oss models were trained natively in MXFP4, which is a 4-bit floating point / e2m1 (2-exponent, 1 bit mantissa, 1 bit sign)
56.
▲
by
deepsquirrelnet
7mo ago
That’s why I unsubbed today! Otherwise I might forget.
57.
▲
by
deepsquirrelnet
7mo ago
I love the work unsloth is doing. I only wish gguf format had better vllm support. It’s sometimes hard to find trustworthy quants that work well with vllm.
58.
▲
by
deepsquirrelnet
7mo ago
Isn’t there some kind of term for when the government controls the means of production. I’ll think about it. It’s one of those terms that’s been thrown around so loosely by this regime you knew they were going there.
59.
▲
by
deepsquirrelnet
7mo ago
Ask an llm to pick a random number from 1-10. My money is on 7. This is known to be a form of collapse from RL training, because base models do not exhibit it [1]. 1. https://arxiv.org/abs/2505.00047
60.
▲
by
deepsquirrelnet
7mo ago
The market reflects reality. The new reality is that the people who are invested apparently don't need liquidity, and bad news doesn't really matter.
More ›