Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
bradfox2
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
61.
▲
by
bradfox2
3y ago
Paper mentions amorphous state plus an annealing step. Should be glass like at the end. If so, heat rates (up and down) of the last step are important and hardly mentioned.
62.
▲
by
bradfox2
3y ago
LLM output is scored by another model that produces a reward for the entire sequence emitted by the LLM. The reward model is trained on human preferences or some other metric usually. It's RL because we train on the reward and not some
63.
▲
by
bradfox2
3y ago
It's still loss being backproped, but the loss is calculated over a different criteria
64.
▲
by
bradfox2
3y ago
Compressed air usually.
65.
▲
by
bradfox2
3y ago
On my team, there was not intially acceptance from folks that were used to traditional nlp techniques. The models built with transformers were viewed as over parameterized. It really wasn't until the original BERT paper came out and t
66.
▲
by
bradfox2
3y ago
All partially funded by sweetheart lease deals between land ASU owns and billions of dollars in corporate office buildings and hotels thanks to the tax exempt status of the University. Status that is supposed to be used for educational pur
67.
▲
by
bradfox2
4y ago
My company builds AI systems for Nuclear plants. In fact, AI systems based on LLMs are already handling issue triage at several plants in the US. Dynamic Operating procedures are something the industry is interested in and in certain scena