Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
CjHuber
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
CjHuber
8mo ago
And how to get to the old verbose mode then...?
32.
▲
by
CjHuber
8mo ago
Does it not use prompt caching?
33.
▲
by
CjHuber
8mo ago
I always wondered isn't it trivial to bot upvotes on Moltbook and then put some prompt injection stuff to the first place on the frontpage? Is it heavily moderated or how come this didn't happen yet
34.
▲
by
CjHuber
8mo ago
That feels like a stupid article. well of course if you have one single thing you want to optimize putting it into AGENTS.md is better. but the advantage of skills is exactly that you don't cram them all into the AGENTS file. Let'
35.
▲
by
CjHuber
8mo ago
It was because of the NYT OpenAI case, however since mid October they are no longer under that legal order. What they keep retaining now and what not, nobody knows but even if they still had the date they surely wouldn't blow their cov
36.
▲
by
CjHuber
9mo ago
I wonder how much more efficient and effective it would be after fine tuning models for each role
37.
▲
by
CjHuber
9mo ago
It depends on the API path. Chat completions does what you describe, however isn't it legacy? I've only used codex with the responses v1 API and there it's the complete opposite. Already generated reasoning tokens even persis
38.
▲
by
CjHuber
9mo ago
Oh wow that's the first time I've heard about those tasks. I would never consent to that and that they are enabled by default and shipped in the .vscode folder where most people probably nevereven would have thought about looking
39.
▲
by
CjHuber
9mo ago
I find this reply concerning. If its THE security feature, then why is "Trust" a glowing bright blue button in a popup that pop up at the startup forcing a decision. That makes no sense at all. Why not a banner with the option t
40.
▲
by
CjHuber
9mo ago
Exactly that's why I was making the comparison, It's not a in your face PopUp, where users get used to just pressing the blue, highlighted and glowing "I trust the authors" button without even being told what features th
41.
▲
by
CjHuber
9mo ago
yeah that does make sense that these choices are related to it being a big part of gastown, still I feel it would be much more sensible to make a different abstraction separating beads core features from the coordination layer
42.
▲
by
CjHuber
9mo ago
I don't like the way it is handled. Imagine Excel actively prompting you with a pop up every time you open a sheet: "Do you trust the authors of this file? If not you will loose out on cool features and the sheet runs in restricte
43.
▲
by
CjHuber
9mo ago
I still don't get what beads needs a daemon for, or a db. After a while of using 'bd --no-daemon --no-db' I was sick of it and switched to beans and my agents seem to be able to make use of it much better, on the one hand its
44.
▲
by
CjHuber
9mo ago
https://news.ycombinator.com/item?id=46631586
45.
▲
by
CjHuber
9mo ago
I didn't upgrade yet but for me codex on a business plan still works, so they didn't deprecate the old auth (yet). But yeah honestly I've never seen any other repo with so many important issues that are just being closed with
46.
▲
by
CjHuber
9mo ago
So you are saying this has the potential to make Germany more autonomous and less dependent?
47.
▲
by
CjHuber
9mo ago
Anyway, props to Mertz for admitting the mistake, we’ll see if they will fix it somehow That‘s the thing. Everyone knew it was costly, nobody ever thought it was good strategically. If he now says it’s a „strategic mistake“ that‘s laughable
48.
▲
by
CjHuber
9mo ago
Was that not clear from the beginning? Nobody ever claimed it was for strategic purposes, the narrative was "we don't like nuclear anything, we will get rid of it we can bear the costs". So I don't think you could even c
49.
▲
by
CjHuber
9mo ago
My point was not that those video to text models are good like they are used for example in that case, but more generally I was referring to that list of indicators. Like surely when analysing a movie it is alright if some things are misund
50.
▲
by
CjHuber
9mo ago
I think all of those are terrible indicators, 1 and 2 for example only measure how well LLMs can handle long context sizes. If a movie or novel is famous the training data is already full of commentary and interpretations of them. If its so
51.
▲
by
CjHuber
9mo ago
I think the actual problem is everyone tries to assert how capable or not coding agents currently are, but how useful they are depends so much on what you are trying to get them to do and also on your communication with the model. And often
52.
▲
by
CjHuber
9mo ago
Well I think it also has to do with communication with LLMs being different to communication with humans. If you tell a developer "don't do busywork" they surely wouldn't say "Oh the repo looks like a trash dump, bu
53.
▲
by
CjHuber
9mo ago
>In the study, researchers told several iterations of four LLMs – Claude, Grok, Gemini and ChatGPT – that they were therapy clients and the user was the therapist So same as always they tell them to roleplay and when they comply everyone
54.
▲
by
CjHuber
9mo ago
Based on those, it seems you are not actually using them to create big codebases from scratch, but rather for problems that would normally take quite a while, not because they are inherently difficult to implement, but because you would nor
55.
▲
by
CjHuber
9mo ago
I'm so glad I switched to fish, I'd rather have genuinely good settings out of the box rather than endless configuration, and honestly it's much better out of the box than any configuration I've ever had. Only drawback i
56.
▲
by
CjHuber
9mo ago
I don't know what it is but I don't think that it's an ad. Otherwise I guess they wouldn't have that snarky pro-israel undertone towards him.
57.
▲
by
CjHuber
9mo ago
I suppose it‘s the latter + maybe some finetuning, it’s definitely not like DeepSeek where the answer of the model get‘s replaced when you are talking something uncomfortable for China
58.
▲
by
CjHuber
9mo ago
I know you‘re joking but to contribute something constructive here, most models now have guardrails against being threatened. So if you threaten them it would be with something out of your control like „… or the already depressed code revie
59.
▲
by
CjHuber
9mo ago
I‘d say such hacks don‘t make you an engineer but they are definitely part of engineering anything that has to do with LLMs. With too long systemprompts/agents.md not working well it definitely makes sense to optimize the existing prom
60.
▲
by
CjHuber
9mo ago
I have the theory that agents will improve a lot when trained on more recent training data. Like I‘ve had agents have context anxiety because they still think an average LLM context window is around 32k tokens. Also building agents with age
More ›