Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sgc
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
sgc
3mo ago
It's like Grease monkey for CSS instead of JS. https://addons.mozilla.org/en-US/firefox/addon/styl-us/
32.
▲
by
sgc
3mo ago
I use a 30-seconds-of-work horribly-hacked Stylus setting to make it a blue dot that can expand on hover: @-moz-document regexp(". ") { / Insert code here... */ #credential_picker_container { width: fit-
33.
▲
by
sgc
3mo ago
No idea, but yes it seems like good practice to trim context there as well. A good candidate for v2 if it's not present. At the same time, even if that is not done, the todo is both specification and tests. If it passes those when it s
34.
▲
by
sgc
3mo ago
Of course I have been down the rabbit hole. The software is wrong. Give me an option to bypass re-encoding, and instead of failing based on diagnostics like it does, just play the darn file the best you can. I only have so much patience and
35.
▲
by
sgc
3mo ago
> Don’t make me report you to a submod. This is pathologically outside scope. I don't think I have ever seen somebody threaten somebody else on hn before. The 'don't make me do it to you' abuser trope is next leve
36.
▲
by
sgc
3mo ago
I have a problem that it constantly tries to re-encode video that could be played perfectly fine in their player or the mpv shim. I wound up just mounting my remote drives and playing in my local mpv. I don't really use it because of t
37.
▲
by
sgc
3mo ago
Actually it does include an important level of review. They say "init a TODO list with a validation step for each item", and then the last several sections of the execution are within the TODO - where it is validating the code. To
38.
▲
by
sgc
3mo ago
I think you missed my point, which is that in the US, people typically described as nationalists tend to be pseudo-nationalists who value pomp and ceremony, but not substantial concrete actions to better their country or actual real care an
39.
▲
by
sgc
3mo ago
I am American who has lived in many countries around the world, and I think this is distinctly wrong and the source of many problems in the US. It would be more correct to say that the average American values outward displays of nationalism
40.
▲
by
sgc
3mo ago
A bit late, but I followed up on this useful tip and found a gist that breaks down the opencode methodology: https://gist.github.com/rmk40/cde7a98c1c90614a27478216cc0155... The gist led me to the opencode session /
41.
▲
by
sgc
3mo ago
It seems like there could be a useful strategy of writing a plan with a main agent, and then instead of spawning subagents to implement, fork the main context to write each part. Then use one last fork to verify the work. That way you keep
42.
▲
by
sgc
3mo ago
It seems like they are actually using the subscription providers' respective cli tools and managing context for them. In which case I believe it is not against the ToS any more than invoking codex cli from a custom python script would
43.
▲
by
sgc
3mo ago
Can anybody share a tested system prompt they use for general coding tasks in pi?
44.
▲
by
sgc
3mo ago
I have one core complex task where there are a number of simple errors like this. The easiest thing for me was to just have a post-processing script that performs: lint > mark known fail-early results > fix common errors (all formatti
45.
▲
by
sgc
3mo ago
I don't see raw token counts, just a list of steps and page counts. For example, what is the rough average token count per page in the ocr and in the translation steps for a Greek book? I have seen Gemini costs change quite a bit when
46.
▲
by
sgc
3mo ago
Curious as to what your budget was to get where you are today? That's a lot of tokens. I presume you are using gemini flash?
47.
▲
by
sgc
3mo ago
I ran into a website for work that would let you create a long password, but silently truncate it to 12 characters before saving. Mind boggling.
48.
▲
by
sgc
3mo ago
Sounds more like they are implementing mass surveillance and reporting whatever the US Gov wants for 'security reasons' back to them.
49.
▲
by
sgc
3mo ago
Cruise missiles are not general purpose tools, it's obviously not even remotely similar. Virtually everybody reading this could use Mythos immediately to do real work, collectively in virtually every part of the economy. It's pret
50.
▲
by
sgc
4mo ago
What I find impossible to judge is whether me choosing the harness that works best for me and the way I like to work will limit the quality of the LLM output. In this case, given the complexity of LangChain I don't know if it would bur
51.
▲
by
sgc
4mo ago
I am very curious about this as well. I'm looking for something that does really well with workflows that require 20 plus steps including a couple while loops and user verifications, but also something simple like a chat bot with acces
52.
▲
by
sgc
4mo ago
I think in most scenarios you don't need to worry so much about kvm ram use, since it looks static but actually it's not and you can over-commit [1]. And of course disk allocation can be dynamic as well. I prefer a lot more secu
53.
▲
by
sgc
4mo ago
Since models just output the the most probable tokens and you can never accuse them of doing anything other than making it all up, I would like to see these tests run with a prompt that attempts to mitigate hallucination and finishes with s
54.
▲
by
sgc
4mo ago
The actual research paper: https://www.sciencedirect.com/science/article/pii/S135041772...
55.
▲
by
sgc
4mo ago
I *literally* cannot read that yellow text on the white background. I even tried changing the brightness to almost 0, but there is just not enough contrast.
56.
▲
by
sgc
4mo ago
To check whether I understand how this all works: Wouldn't a 4 bit quant run reasonably well (for that hardware) with far less ram, something like 1.5x the 476gb, or 714gb+ ram?
57.
▲
by
sgc
4mo ago
This is the first time in terms of model progress where my personal response is: It does not matter to me because the models 6-12 months ago were already good enough for most everything I need to do. I think 95% of dev work is perfectly fin
58.
▲
by
sgc
4mo ago
As far as I can tell this type of model requires 640GB+ of memory using FP8. So likely can be run using 320GB+ memory if using FP4 or similar. So that would be 3 Nvidia DGX Sparks, or 12k of hardware. Is that correct? If so, it could make p
59.
▲
by
sgc
4mo ago
If you wouldn't mind, could you explain a bit what the 248B model is good for, and where it breaks down and you need something better? I hear this take often, but it is always a fleeting remark so I have no idea what the 'useful&#
60.
▲
by
sgc
4mo ago
You can absolutely do that by using subprocess.run, or use the codex sdk https://github.com/openai/codex/tree/main/sdk/python
More ›