Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
gck1
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
Codex now encrypts messages passed to subagents
(github.com)
3 points
by
gck1
3mo ago
|
0 comments
32.
▲
by
gck1
3mo ago
And the names they came up with mirror the names OAI models come up with when I try to have them suggest codenames for projects. Just lacking any sort of imagination or coherence.
33.
▲
by
gck1
3mo ago
Is Fable really that much different? I almost instinctively create elaborate processes, workflows, set up a bunch of linters and dump research docs any time I bootstrap a new project regardless of what model I'm using. They all spiral
34.
▲
by
gck1
3mo ago
You can't drive prosumers anywhere near API prices. I would guess that the maximum you can extract from vast majority of prosumers is maybe $500/mo, and even that is a big stretch. Once you cross that threshold, prosumers will s
35.
▲
by
gck1
3mo ago
Both are verbose in their own way, and both - terrible. Claude models love to throw huge blobs of text in architecture planning / interview conversations, but in not a mentally draining language. OpenAI models are more compact, but ver
36.
▲
by
gck1
3mo ago
Is this how it's going to work for every new model release - the executive branch inserting itself into decisions made by private companies? As a non-American who grew up in a deeply corrupt country, this sets off every alarm bell.
37.
▲
by
gck1
3mo ago
I have a really hard time understanding how anyone is ready to pay $200/mo to be put on such a rollercoaster all the time . I've used fable, it's great. But nothing beats predictability - ever. This truly feels like some for
38.
▲
by
gck1
3mo ago
Forget 5.5-pro. Why isn't everyone talking about the fact that there's no 1M context window model in codex? Yes, I know attention degrades above ~200k, but it's still useful in many applications.
39.
▲
by
gck1
3mo ago
IF they were light years ahead of competition, then that would make sense and would be a perfectly valid strategy. But it's not, so you're burning very important bridges that you may need in 3-4 months.
40.
▲
by
gck1
3mo ago
> What does it matter which tool I use when I hit the limit? A third party harness may have a misconfiguration of prompt caching, leading to more load on Anthropic's servers, they could also have wildly different usage patterns (Her
41.
▲
by
gck1
3mo ago
I'm never going back to claude from codex, including for the reasons you mentioned, but it must be said that web chat inference on ChatGPT is magic incantation, and I'm almost certain they're not serving the same models there
42.
▲
by
gck1
3mo ago
This seems to be conflicting with the adoption of Web Bot Auth, which is still in infancy stage. I do have some bots, they're nice and predominantly used for grounding AI harnesses which I use interactively. Knowing that most operators
43.
▲
by
gck1
3mo ago
Since Anthropic has eroded all the trust it could possibly have, I'm going to allow myself to be cynical and say that this move is just another pillar of their shady marketing practices. I know a few real persons who will praise Fable
44.
▲
by
gck1
3mo ago
I've been working on my own private harness for the past 8 months, and I've been collecting ideas from such repos I've stumbled upon. pi-tmux is one such example (seems to be archived now) which inspired me to use tmux as com
45.
▲
by
gck1
3mo ago
It's sad to see that the teams that have the most resources that can contribute to development of next-gen harnesses are essentially copying the same exact thing from each other, with no meaningful changes. And most of the advancement
46.
▲
by
gck1
3mo ago
It's very likely that OAI models will have even more restrictions. Firstly because now they know what feds will do if you don't tune the safety classifiers towards more false positives and secondly, OAI models were always more res
47.
▲
by
gck1
3mo ago
You're paying in full for the "guardrails" embedded in the system prompt, prompt injections, refusals, fallbacks and everything else that may be caused by the service provider. This is the new normal.
48.
▲
by
gck1
3mo ago
This has no mention of what happens to the prompt cache, including the "learn more" link. Knowing Anthropic, it wouldn't surprise me if it will result in a full cache miss/rewrite at fallback, with potentially up to 1M t
49.
▲
by
gck1
3mo ago
Location: Europe Remote: Yes. 12+ yrs, comfortable fully async across US/EU hours Willing to relocate: No Technologies: Rust, Python, Go · LLM/agent systems, multi-agent orchestration, LLM eval · browser automation + anti-detectio
50.
▲
by
gck1
3mo ago
Interesting. I wasn't aware PC gamers were building fabs in their bedrooms and undercutting Micron.
51.
▲
by
gck1
3mo ago
Based only on the third quote (you're literally in the thread discussing second iteration of it), and your username, you can't possibly be acting in good faith here, so I'm not going to waste my time providing references to t
52.
▲
by
gck1
3mo ago
> This started a few months ago when anthropic started beating openai. From where I'm standing, this started a few months ago when Anthropic decided to gaslight users, sabotage their projects, ship malware and attempt a regulatory c
53.
▲
by
gck1
3mo ago
> The gap between Chinese models and American frontier models is estimated at 10 months by Anthropic themselves, and it's growing. #1 I've had use cases where it was clearly obvious the Chinese models were behind. #2 I've
54.
▲
by
gck1
3mo ago
> It's unclear on how this "punishes normal developers" in any shape or form Tons of normal developers use ANTHROPIC_BASE_URL, the flag which activates the malware.
55.
▲
by
gck1
3mo ago
> The trigger is ANTHROPIC_BASE_URL, Claude Code's API base URL override I had a use case where I had to MITM CC's traffic to strip credentials that could have accidentally made it into the harness. I'm happy my paranoid s
56.
▲
by
gck1
3mo ago
It's not really a 90% discount (I went into the rabbit hole) and none of the sites from this list are what people use (looks like some labs and random sites). It's more closer to 30% specifically for Claude models, and it's c
57.
▲
by
gck1
4mo ago
Eh, pretty much everyone that spent some time tweaking their harness already had a homemade 'ultracode' long before Anthropic did it. OpenAI is just way more careful with what features they add or enable by default in their harnes
58.
▲
by
gck1
4mo ago
Neither is OpenaAI's ultra. Article specifically calls it 'mode' and it's not even mentioned in the model card. It's for sure a codex harness feature. EDIT: yeah, it's the same thing. https://github.
59.
▲
by
gck1
4mo ago
If it's anything like ClaudeCode's ultracode, it's nothing new or revolutionary. It's essentially a bunch of subagents being called by a deterministic script written by the main model thread, each eating tokens for lunch
60.
▲
by
gck1
4mo ago
I'm in Europe. The only superpower that's been hostile to me, very directly - was US, when they asked a company I was relying on to limit model access based on nationality. China has (so far), never done that to me.
More ›