Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
lhk931122
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
lhk931122
15d ago
I also have set custom harness environment for me in the company (and also for at home), I expect that big companies (OpenAI, or Linear and Snowflake) may release the internal harness setting or functions that would be more comfortable than
2.
▲
by
lhk931122
18d ago
The structured output that Google provides (the function that blocks the disallowed words) may calculate softmax output with assigning -inf value to the other words, then the probability is not the same as the weight the model firstly put o
3.
▲
by
lhk931122
19d ago
But, I think that the agent may also generate like this approach, not directly generate text, instead first writing the generation code then execute if they are the smart agent.
4.
▲
by
lhk931122
21d ago
Since the baseline LLM architecture is similar with each other, maybe already there are relations between their each opinion. Of course, this is an assumption and the explicit training seems effective in this case. But, I'm curious whe
5.
▲
by
lhk931122
25d ago
Six months into customizing my own Claude Code harness, I've settled on assuming Anthropic and OpenAI will just handle all of it, except turning my own flows into skills.
6.
▲
by
lhk931122
27d ago
Is the attention explanation of why the model tells like this? I've seen that there are many discussions about this. (Image attention visualizations were not that good I think)
7.
▲
by
lhk931122
29d ago
Ah, success rate here are scored by an agentic classifier. And uncertain outcomes are excluded from the graph. The thing measured and grading it comes from the same house. In my setup, review agent pass work that an outside critic later rej
8.
▲
by
lhk931122
1mo ago
I only get web search from Claude Code when I ask, only one of my last 8 sessions accessed the web at all. But they were the case of Opus here. Curious what Fable 5.1 model does instead of Opus.
9.
▲
by
lhk931122
1mo ago
I tried to make a memory system in Claude Code (make it behave like a human from the perspective of memory) using markdown files and indexing plus hooks. But, retrieval is not as good as a human and since keywords are ambiguous, it doesn&#x
10.
▲
by
lhk931122
1mo ago
We can control it (claude.md and plus use 'rules') but, it would be better if not verbose before I set it. yeah
11.
▲
by
lhk931122
1mo ago
Yeah. That's why people these days avoid long descriptions and instead keep things as short as possible. in my experience, it seems like LLM can't recognize (or less attention value) unless it's structured in a deductive or i
12.
▲
by
lhk931122
1mo ago
API Error: 529 Overloaded in some sessions - South Korea Still does not work in these sessions