Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
wizzard0
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
wizzard0
2y ago
well the poster looks to have an ubuntu flair on its profile pic
32.
▲
by
wizzard0
2y ago
that article is suspiciously quiet on commodities (and their derivatives, in case you don't want to sign up for storage/delivery/etc etc). commodity derivatives are kind of a debt of a physical operator though
33.
▲
by
wizzard0
2y ago
one of the must-have things to have on every windows machine
34.
▲
by
wizzard0
2y ago
Um, yes, TimeProvider is cool but works only insofar as ... > If you provide a fake time provider in all those places I'm looking for something more akin to https://go.dev/issue/67434 (i.e. testing the code tha
35.
▲
by
wizzard0
2y ago
Nope, have yet to find a deepseek v2 provider which doesn’t log my prompts
36.
▲
by
wizzard0
2y ago
Very cool! Anything similar for .net?
37.
▲
by
wizzard0
2y ago
Aider is one tool which works on medium sized repos for me. Make sure to use claude3.5-sonnet or gpt4o, other LLMs are not yet up to the task with the context size required
38.
▲
by
wizzard0
3y ago
Impressive visualization! Also kind of keeps the vibe of the internal Google tooling, minimalistic and down-to-earth pragmatic
39.
▲
by
wizzard0
3y ago
A pixel takes more than 1 bit to store, too
40.
▲
by
wizzard0
3y ago
Authors claim their mechanism is more computationally efficient (both asymptotically and on current hardware) than the Transformer attention, while performing much the same role, yes
41.
▲
by
wizzard0
3y ago
Photo copiers replacing digits on scanned financial reports with digits that compress better are a decade old already :) http://www.dkriesel.com/en/blog/2013/0802_xerox-workcentres_... ?
42.
▲
by
wizzard0
3y ago
Yes, a GGUF-converted version works fine with llama.cpp for me
43.
▲
by
wizzard0
3y ago
“-Chat” means the model has been additionally fine-tuned on “human/assistant” conversation patterns, so the predicted text matches them better and if you feed a question you will get back replies instead of e.g. a plausible continuatio
44.
▲
Teaching LLMs to Refuse Unknown Questions (TLDR: Train on "I Dunno" as Well)
(arxiv.org)
2 points
by
wizzard0
3y ago
|
1 comments
45.
▲
by
wizzard0
3y ago
> Our research is motivated by the observation that previous instruction tuning methods force the model to complete a sentence no matter whether the model knows the knowledge or not. And so... adding the random samples that include "
46.
▲
by
wizzard0
3y ago
As much as I love algorithm-heavy tasks, I must admit 90% of software development is working with stakeholders on mapping the business needs to code. And then the most advanced algorithms are usually just `git clone/cargo add/go g
47.
▲
Ask HN: Alternatives to fig.io as it has signups disabled?
3 points
by
wizzard0
3y ago
|
1 comments
48.
▲
by
wizzard0
3y ago
> many iterations of each prompt BTW its much faster and cheaper to artive at a good prompt if you sample the model in deterministic mode (ie temperature=0) By default you have to guess if the difference is due to the prompt change or du
49.
▲
by
wizzard0
3y ago
FYI the Asahi Linux installer says “please upgrade that should help” but bumping 13.6 -> Sonoma 14.1 did not help, the SystemRecovery still shows up as 13.5 Or does that mean wait for 14.2?
50.
▲
by
wizzard0
3y ago
I don’t think it’s the plants, the fine dust of the pills being _consumed_ all over the world should be enougg
51.
▲
by
wizzard0
3y ago
The words “entire athmosphere” and the “plants” might be an exaggregation, but 1) I can readily imagine the _labs_ being contaminated by the reference samples. 2) If the compound is stable, not clumping and not hygroscopic - yes, certainly.
52.
▲
by
wizzard0
3y ago
Makes the iPhone uncomfortably hot as well
53.
▲
by
wizzard0
3y ago
well i must admit monitoring certificate issuance with LetsEncrypt is quite boring even if you HAVE alerts set up (I do) so… not surprised. Still cool. What a time to be alive
54.
▲
by
wizzard0
3y ago
cuBLAS is cuda, clBLAST is OpenCL
55.
▲
by
wizzard0
3y ago
Also, an encoder like these can inject any amount of entropy (eg replace "0" with "100500999-100500999") which gzip can reduce to "100500999-same" but without knowledge of JS semantics has no chance to reduce t
56.
▲
by
wizzard0
3y ago
1) The title is a clickbait, but 2) Thanks for leading with an example of a negative result! That's what any researcher faces every day, unlike what gets published, after all
57.
▲
by
wizzard0
3y ago
CPU is still the first-class option, but GGML also supports Metal, cuBLAS and clBLAST for hw acceleration
58.
▲
by
wizzard0
3y ago
GGML is the library, GGUF is the new GGML model format.
59.
▲
Reconfigurable mixed-kernel heterojunction transistors for SVM classification
(nature.com)
1 points
by
wizzard0
3y ago
|
1 comments
60.
▲
by
wizzard0
3y ago
nvidia might be the king of the hill right now, but the future of AI is reconfigurable analog electronics (~100x more energy efficient already, which will take Moore’s law at least another 10 years for silicon) Caveat: no backprop :P forwar
More ›