Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
causalmodels
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
61.
▲
by
causalmodels
3y ago
Fiascos like this display neither experience nor intelligence. This whole saga was a colossal failure on the part of the previous board.
62.
▲
by
causalmodels
3y ago
Zero chance OpenAI ends up seeing any more compute if they merge.
63.
▲
by
causalmodels
3y ago
I don't agree with sex work, but I will absolutely defend the rights of people to get paid for their labor.
64.
▲
by
causalmodels
3y ago
A gig worker or freelancer who sets up an LLC is essentially no different than an individual. They should have protections.
65.
▲
by
causalmodels
3y ago
Isn't the EU currently considering making open source ai developers legally liable for their models?
66.
▲
by
causalmodels
3y ago
They’re not posts, they’re native ads.
67.
▲
by
causalmodels
3y ago
This is quite harsh for a repo explicitly created to experiment with non-Unix OS ideas. It is an experiment, it doesn't have to e useful. Personally, I enjoy when people post things like this as it helps me learn.
68.
▲
by
causalmodels
3y ago
This is a decent place to get started https://arxiv.org/pdf/math/0009118.pdf You can extend some of the thinking in this paper to cover action spaces for deep RL models to really have some fun!
69.
▲
by
causalmodels
3y ago
I’ve used algebraic topology for certain robotics tasks when statistical saftey wasn’t good enough and when trying to generalize techniques to learn on graphs. The results are spectacular, but the application set is indeed very limited.
70.
▲
by
causalmodels
3y ago
My undergrad research was on CW complexes and I did spend the following decade trying find applications for it in ML. It has worked out approximately twice, but my god is it powerful when it naturally fits a problem space.
71.
▲
by
causalmodels
3y ago
Hoverboards.
72.
▲
by
causalmodels
3y ago
> What I hear is that training AI on AI outputs is fruitless. To me, thats as good as test as any: when you can use at scale an AI product to inform the training of an unrelated AI system I will be impressed. The phi-1 team did this a fe
73.
▲
by
causalmodels
3y ago
When did you graduate? None of my humanities courses required handwriting and that was ten years ago. Plus all non-intro humanities courses were either take home exams over a few days or final papers.
74.
▲
by
causalmodels
3y ago
My parents had the same idea but instead of smart phone use, it was religion. It wasn't so successful for them.
75.
▲
by
causalmodels
3y ago
Hard disagree. Phi-1 seems to be a harbinger of what is to come. However i think there is a plausible argument to be made that their training process approximates a kind of distillation
76.
▲
by
causalmodels
3y ago
It's not an either or. We're going to leverage the web trained LLMs to bootstrap the specialist models via combination of training token quality classifiers and synthetic data generation. Phi-1 is a pretty good example of this.
77.
▲
by
causalmodels
3y ago
> Rule of thumb is that you need ~20 tokens per parameter. That rule of thumb is wrong. The chinchilla paper has it anywhere between 1 and 100 tokens per parameter.
78.
▲
by
causalmodels
3y ago
His solution is a global regulatory regime to ban new large training runs. The tools required to accomplish this are, IMO, out of the question but I will give Yud credit for being honest about them while others who share his viewpoint try t
79.
▲
by
causalmodels
4y ago
Funny to see someone call for wide scale cooperation to stop training LLMs but can't seem to get people to cooperate on the embargo.
80.
▲
by
causalmodels
4y ago
Prescreening? yes. Ad targeting that looks an awful lot like prescreening? No.
81.
▲
by
causalmodels
4y ago
IMO the ethnic quotas would be the most unsavory aspects to an American audience.
82.
▲
by
causalmodels
4y ago
I'm going to assume you know how to stand up and manage a distributed training cluster as a simplifying assumption. Note this is an aggressive assumption. You would need to replicate the preprocessing steps. Replicating these steps is
83.
▲
by
causalmodels
4y ago
Last week I went to an urgent care clinic to get treated for strep throat. Many of the steps did not occur as you describe. Here is what happened. 1. I signed in using an automated terminal. A tablet which scanned my ID and insurance card.
84.
▲
by
causalmodels
4y ago
OpenAI has a capped-profit model so I really don't think it will ever go public.
85.
▲
by
causalmodels
4y ago
Well OpenAI's explicit goal is AGI. While this is farfetched, it would clearly make them the most valuable company in the history of the world and the ability to monetize something like AGI would essentially be unlimited. OpenAi isn&#x
86.
▲
by
causalmodels
4y ago
Something that I think is often overlooked when talking about this is a large chunk of tech labor are millennials. These people are now in their thirties or early 40s. Many now have families and want different things from work than they did
87.
▲
by
causalmodels
4y ago
Deepmind has been doing some interesting work around Retrieval Enhanced Transformers (RETRO) models [1] that might be relevant in this context. [1] https://www.deepmind.com/blog/improving-language-models-by-r...
88.
▲
by
causalmodels
4y ago
a) No idea. I read slowly, cannot spell, get letters confused, and frequently don't realize small words are missing from sentences. It is in very large part why I went into math. b) Not if everything I write is going to now be run thro
89.
▲
by
causalmodels
4y ago
I guess that's fair. I just personally don't think the additional gain is worth taking away your child's privacy.
90.
▲
by
causalmodels
4y ago
My younger brother and I both have fairly severe dyslexia. He's been applying to school and has been using ChatGPT to help him correct spelling and grammar mistakes rather than going to a person for help. It has been fairly incredible
More ›