Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
coder543
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
18 ms
·
271.
▲
by
coder543
2y ago
100% agree. I still listen to this one from time to time: https://suno.com/song/da6d4a83-1001-4694-8c28-648a6e8bad0a https://arstechnica.com/information-technology/2024/04/mit-l... It&#x
272.
▲
by
coder543
2y ago
It’s not even a question. I still have my Pebble Time Round in a box. No other watch has come close. That thing was probably under half the thickness and half the weight of my Apple Watch Ultra and yet it had 3x the battery life! The physic
273.
▲
by
coder543
2y ago
I understand you were trying to make “up and to the right” = “best”, but the inverted x-axis really confused me at first. Not a huge fan. Also, I wonder how you’re calculating costs, because while a 3:1 ratio kind of sort of makes sense for
274.
▲
by
coder543
2y ago
The "admired and desired" chart has nothing to do with actual use. According to this chart[0] higher on the same survey report, JavaScript is the #1 most used, and Python is #3. But, I don't see this as a contest... being &qu
275.
▲
by
coder543
2y ago
Cursor’s composer agent is so slick. I tried Zed out for a few minutes this evening, and in addition to what you said, I was surprised that there wasn’t an option to bring my own tab completion server. Zed felt very nice, definitely reminde
276.
▲
by
coder543
2y ago
> I wired one up “One”? Wired up how ? There is a huge difference between the best and worst. They aren’t fungible. Which one? How long ago? Did it even support FIM (fill in middle), or was it blindly guessing from the left side? Did
277.
▲
by
coder543
2y ago
Your memory appears to be incorrect. SSE was first built into a web browser back in 2006. By 2011, it was supported in all major browsers except IE. SSE is really just an enhanced, more efficient version of long polling, which I believe was
278.
▲
by
coder543
2y ago
> Running a prompt against every single cell of a 10k row document was never gonna happen with a large model That isn’t the main point of FLAME, as I understood it. The main point was to help you when you’re editing a particular cell. co
279.
▲
by
coder543
2y ago
That paper is from over a year ago, and it compared against codex-davinci... which was basically GPT-3, from what I understand. Saying >100B makes it sound a lot more impressive than it is in today's context... 100B models today are
280.
▲
by
coder543
2y ago
Someone mentioned generating millions of (very short) stories with an LLM a few weeks ago: https://news.ycombinator.com/item?id=42577644 They linked to an interactive explorer that nicely shows the diversity of the dataset,
281.
▲
by
coder543
2y ago
Maybe function calling using JSON blobs isn't even the optimal approach... I saw some stuff recently about having LLMs write Python code to execute what they want, and LLMs tend to be a lot better at Python without any additional funct
282.
▲
by
coder543
2y ago
It agrees with what I was seeing, but it doesn’t really seem to explain much. I still don’t know if it is Codeberg-specific or Forgejo-specific, or why either of them would be slow for this task when Gitea local to me can go much faster (
283.
▲
by
coder543
2y ago
I noticed some surprising load times on Codeberg’s Forgejo instance. For example: - The first page of releases (out of only 63 releases total) takes 3–5.5 seconds to load: https://codeberg.org/forgejo/forgejo/relea
284.
▲
by
coder543
2y ago
I thought it was an interesting post, so I tried to add Railway's blog to my RSS reader... but it didn't work. I tried searching the page source for RSS and also found nothing. Eventually, I noticed the RSS icon in the top right,
285.
▲
by
coder543
2y ago
> My conversation quickly began to approach the context window for the LLM and some RAG engineering is very necessary to keep the LLM informed about the key parts of your history Assuming we're talking about GPT-4o, that 128k contex
286.
▲
by
coder543
2y ago
The usual bottleneck for self-hosted LLMs is memory bandwidth. It doesn't really matter if there are integrated graphics or not... the models will run at the same (very slow) speed on CPU-only. Macs are only decent for LLMs because App
287.
▲
by
coder543
2y ago
gpt-4o-mini might not be the best point of reference for what good LLMs can do with code: https://aider.chat/docs/leaderboards/#aider-polyglot-benchma... A teeny tiny model such as a 1.5B model is really dumb, and
288.
▲
by
coder543
2y ago
Google's research blog does not seem to provide this, but many blogs include the Open Graph metadata[0] around when the article was published or modified: article:published_time - datetime - When the article was first published.
289.
▲
by
coder543
2y ago
What does swarm actually do better for a single-node, single-instance deployment? (I have no experience with swarm, but on googling it, it looks like it is targeted at cluster deployments. Compose seems like the simpler choice here.)
290.
▲
by
coder543
2y ago
Your original solution of binding to 127.0.0.1 generally seems fine. Also, if you're spinning up a web app and its supporting services all in Docker, and you're really just running this on a single $3/mo instance... my unpopu
291.
▲
by
coder543
2y ago
There's a lot we don't know about Perplexity's costs, like the cost of supporting free users, the difference in usage between free and paying users, and whether these ads fully offset the cost of free users. From what I'
292.
▲
by
coder543
2y ago
"Hundreds of thousands of people" are paying $20/mo for it, according to the CEO.[0] That seems like a very respectable place for such an early product to be. It is extremely far from "no one". [0]: https:/&
293.
▲
by
coder543
2y ago
Does template_plural actually work well / offer any benefits?
294.
▲
by
coder543
2y ago
Yep, and that gotcha got me, as a perfectly non-silicon human. My bad everyone.
295.
▲
by
coder543
2y ago
One problem from the benchmark: "prompt_id": "river_crossing_easy", "category": "Logic Puzzle", "title": "Easy river crossing", "prompt": "
296.
▲
by
coder543
2y ago
For compressing short (<100 bytes), repetitive strings, you could potentially train a zstd dictionary on your dataset, and then use that same dictionary for all rows. Of course, you’d want to disable several zstd defaults, like outputtin
297.
▲
by
coder543
2y ago
> Qwen team just dropped the Apache 2.0 licensed QvQ-72B-Preview When they dropped it, the huggingface license metadata said it was under the Qwen license, but the actual LICENSE file was Apache 2.0. Now, they have "corrected"
298.
▲
Indexing Code at Scale with Glean
(engineering.fb.com)
2 points
by
coder543
2y ago
|
1 comments
299.
▲
by
coder543
2y ago
That linked discussion doesn't say it is broken all the time (which your comment strongly implies to anyone who doesn't read the link)... and I had verified the parallax effect was working fine earlier today when I made my comme
300.
▲
by
coder543
2y ago
I didn’t say it was gone! Just that Apple added it years ago. The website this post is about doesn’t implement the feature.
More ›