Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
pama
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
61.
▲
by
pama
2mo ago
You can do whatever you want with the model within your own organization. If you use it commercially—either as a model-as-a-service business or in a very large-scale product—you should check the additional license terms, which go beyond MIT
62.
▲
by
pama
2mo ago
Or just fix it. I use lockdown mode on iphone. It does not work on safari or chrome, ironically telling me to use a different browser like safari or chrome.
63.
▲
by
pama
2mo ago
Serious question: if one were willing to give up on curses, isn’t Emacs/elisp providing the best multiplexing system available to humankind? And conceptually, why would agents ever need to depend on curses?
64.
▲
by
pama
3mo ago
The drug candidates that enter human clinical trials, (phase 1 to phase 3) are identical in all ways to the final product if approved. The company is not allowed to change any part of the manufacturing process. All the research for how to m
65.
▲
by
pama
3mo ago
What is the rationale for gpt-5.5 when gpt-5.6-sol exists?
66.
▲
by
pama
3mo ago
The internet convention in the 90s used to be two dashes for the en-dash character and three dashes for em-dash character. You used a space-dash-space notation that is easy to decipher but the typical dashes variants dont have spaces around
67.
▲
by
pama
3mo ago
It was never DeepSeek’s release that dropped the NASDAQ at the time; it was the unknown risks of the early thoughts related to trade wars. Popular financial newspapers can promote anything they want, but these news do not typically drive la
68.
▲
by
pama
3mo ago
It is no rhetoric. It is the time spent when working interactively on a thesis project with one software vs the other.
69.
▲
by
pama
3mo ago
The blog is about writing an agent when you dont already have an agent, but only a plain LLM. It stitches the minimal pieces together. Agents dont need lots of supporting infra, so it is good to keep the code concise. Not a wow moment for
70.
▲
by
pama
4mo ago
If you look at the timelines for the hiring of the hardware team, this was an extremely fast and high risk implementation from concept to tapeout. Amazing it works at all during bringup.
71.
▲
by
pama
4mo ago
Regarding lack of p-hacking. The placebo arm of blinded trials breaks when your brain can detect a medication. The effect is tiny in these studies; approval was rushed to give hope to patients. The drug was discontinued later.
72.
▲
by
pama
4mo ago
I work in drug discovery. For the past twenty years or so, my personal analogy for this hypothesis has been a fantasy story around the days after the bombing of Dresden, when a new civilization suddenly visits Dresden and has no priors abou
73.
▲
Ghostty-Blackhole
(github.com)
3 points
by
pama
4mo ago
|
0 comments
74.
▲
by
pama
4mo ago
Non-US models are not banned in the US: they are used daily in every state of the US. Some misguided state governments temporarily banned employees from downloading the R1 models and variants released 16 months ago on state government compu
75.
▲
by
pama
4mo ago
The GP did not try to time the market. He suggested a sensible strategy to exclude a tiny subset from an index (less expensive than maintaing the alternative index yourself).
76.
▲
by
pama
5mo ago
Frankly, everyone in the industry knows. When people make these statements without additional clarity they always talk about API prices. You can look at the NVL72 specs and make estimates for electricity and ownership costs rather easily. I
77.
▲
by
pama
5mo ago
The key idea is to have many differently named shells. Typically, I group them by project (common prefix name), and the projects live in directories. I have some hacks to organize ibuffer, to split frames, to reflow buffers in the existing
78.
▲
by
pama
5mo ago
Not sure, tbh. I use emacs -daemon to start a server; emacsclient -nw to connect. I use ssh and start a server on the remote. I spawn multiple shells with infinte buffer size and dumb terminals (M-x shell) so I can seamlessly edit. (These
79.
▲
by
pama
5mo ago
Nasdaq is about 8x higher now than then, so 4x higher M2 is tight. Ofc there is always a chance that this time is different and that the markets are genuinely much more efficient :-)
80.
▲
by
pama
5mo ago
Congrats on the launch. If emacs was unavailable and I needed tmux, I would try it. I am old school, and use emacs daemons for all shell multiplexing. The agents dont need explanations and know how to use emacsclient to create, read, or sen
81.
▲
by
pama
5mo ago
Peter shows the near-term future. Raw API consumer price cost is arbitrary. (The frontier labs can put a 100x markup to cover other operational expenses.) The true cost of inference with same-capability models keeps dropping at dizzying rat
82.
▲
by
pama
5mo ago
Ilya S?
83.
▲
by
pama
5mo ago
Yes the noise is my main complaint for my 10G switch. I didnt expect that high frequency part.
84.
▲
by
pama
5mo ago
But it is worse and more expensive…
85.
▲
by
pama
5mo ago
What does Pareto competitive mean here? Look at the pricing of the V4-flash model: https://api-docs.deepseek.com/quick_start/pricing
86.
▲
by
pama
5mo ago
Unfortunately they only compare to old “all other open models”. There are probably over 10 other open models better than it by now.
87.
▲
by
pama
6mo ago
I was answering to the question about how to know the probability from this comment: > The sequence of tokens that would destroy your production environment can be produced by your agent, no matter how much prompting you use. If you have
88.
▲
by
pama
6mo ago
LLM inference is built upon a probability function over every possible token, given a stream of input tokens. If you serve the model yourself you can get the log prob for the next token, so you just add up a bunch of numbers to get the log
89.
▲
by
pama
6mo ago
Has anyone tested it at home yet and wants to share early impressions?
90.
▲
by
pama
6mo ago
Not sure what you mean by efficiency as this was part of the article and I understand things differently—can you clarify? For the energy of 20 W in an hour on a laptop’s M4 pro, this model produces about 200k tokens (a book or two) at a typ
More ›