Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
jaehong747
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
jaehong747
1mo ago
Ran one more tokenizer test on stealth/ox-alpha. Prompt "2026年8月24日": * ox-alpha: 19 tokens * GLM-5.2: 19 tokens (identical) * NVIDIA nemotron: 26 tokens (different)
2.
▲
by
jaehong747
3mo ago
Good interpretability work, but the problem is it's all in how you interpret it. Bridge concept neurons activating even while talking about something else, this seems pretty obvious to me. Input context activating related representatio
3.
▲
MCP Remind: MCP Has Its Own Place
(github.com)
2 points
by
jaehong747
7mo ago
|
0 comments
4.
▲
Show HN: Claude Code Spinner Verbs Extractor
(github.com)
3 points
by
jaehong747
7mo ago
|
0 comments
5.
▲
Simile: The Simulation Company
(twitter.com)
2 points
by
jaehong747
8mo ago
|
0 comments
6.
▲
by
jaehong747
8mo ago
In the AI Era, Do Startups Still Have an Edge? Startups have always competed with big companies based on three gaps: speed, risk tolerance, and incentives. The question is whether these advantages still hold in the age of AI. * Speed Gap: S
7.
▲
by
jaehong747
1y ago
great job! it reminds me genaiscript. https://microsoft.github.io/genaiscript/ // read files const file = await workspace.readText("data.txt"); // include the file content in the prompt in
8.
▲
by
jaehong747
1y ago
Graham says good writing sounds good and is more likely to be true. But his own writing is hard to read and confusing. His sentences are long and messy. If he’s right, then his own ideas must be wrong because his writing sounds bad.
9.
▲
by
jaehong747
2y ago
Like you, I also find the paper's findings interesting. I'm not arguing that LLMs lack the ability to "think" (mechanically), but rather expressing concern that by choosing the word "thinking" in the paper, LLM
10.
▲
by
jaehong747
2y ago
Modern transformer-based language models fundamentally lack structures and functions for "thinking ahead." And I don't believe that LLMs have emergently developed human-like thinking abilities. This phenomenon appears because
11.
▲
by
jaehong747
2y ago
I’m skeptical of the claim that Claude “plans” its rhymes. The original example—“He saw a carrot and had to grab it, / His hunger was like a starving rabbit”—is explained as if Claude deliberately chooses “rabbit” in advance. However,
12.
▲
by
jaehong747
2y ago
Wow, quickly skimmed the paper, this really is like a USB stick for knowledge. Thanks for sharing! Cool that you worked on something similar before.
13.
▲
by
jaehong747
2y ago
Yeah, it's definitely attracting attention! Have you tried it yet?
14.
▲
Ask HN: Will Anthropic's MCP succeed as an AI integration standard?
1 points
by
jaehong747
2y ago
|
4 comments
15.
▲
by
jaehong747
2y ago
FastHTML is an impressive and innovative idea. It seems like a web development tool similar to Streamlit, but with more precise control. FastHTML's concept led me to consider a feature that allows direct deployment of PyQt code as web
16.
▲
by
jaehong747
2y ago
At the Bitcoin Conference 2024, presidential candidate Trump's speech increased the possibility of strategic Bitcoin holdings for the United States. Senator Cynthia Lummis also introduced a bill for strategic Bitcoin holdings. Of cours
17.
▲
by
jaehong747
2y ago
Thank you. I didn't know that.
18.
▲
by
jaehong747
2y ago
I wanted to argue that Llama 3.1 is not truly open source, and I wanted to discuss this with the public. Looking at previous "Ask HN" posts, it appeared to be a board where both questions and discussions were possible.
19.
▲
by
jaehong747
2y ago
Thank you.
20.
▲
by
jaehong747
2y ago
Is the Llama 3.1 model open source? Yes. Really? No. Training Data = Source Code This time, the Llama 3.1 model was released as open source. However, the training data is not disclosed. In AI and deep learning, training data is the &quo
21.
▲
Ask HN: Is the Llama 3.1 model truly open source?
2 points
by
jaehong747
2y ago
|
7 comments
22.
▲
by
jaehong747
3y ago
I will provide an answer assuming that the question is about the criteria for model training. We train the model based on the difference between the predicted price and the actual price as the criterion. Additionally, we also consider the a
23.
▲
by
jaehong747
3y ago
LSTM now. We are exploring different time series prediction models, such as TFT (Temporal Fusion Transformer) and the latest GPT models. Our aim is to find the best combination of performance and practicality. Currently, we have decided to
24.
▲
Show HN: BTCGPT predicts bitcoin price. (Your Fun Bitcoin Crystal Ball)
(btcgpt.info)
2 points
by
jaehong747
3y ago
|
4 comments