Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sothatsit
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
sothatsit
3mo ago
I disagree with keeping an eye on the model as it is working, approving every command, and denying and stopping the model when you think it has gone wrong. It is not that it is actively harmful to do this, but rather that it is a waste of t
32.
▲
by
sothatsit
3mo ago
"Nuanced discussions" is more about describing a design to a model, asking the model to critique your design and ask you for clarifications, and then you providing those clarifications and the model "getting it" and proc
33.
▲
by
sothatsit
3mo ago
This “short leash” seems like more of a crutch to me, and a sign of not giving the AI enough detail on the problem to begin with, or not reviewing and iterating on its output. Hand-holding great models like Fable through implementation is a
34.
▲
by
sothatsit
3mo ago
You can get away with a lot when you have the best models… I’m looking forward to OpenAI or open-source catching up so we have some competition again.
35.
▲
by
sothatsit
3mo ago
Refusals, presumably.
36.
▲
by
sothatsit
4mo ago
It’s pretty incredible to me that a mammoth change like this is possible to prototype now using LLMs. It makes me wonder how much of our software stack will become more malleable to big ideas and experiments in the future, like Filip’s idea
37.
▲
by
sothatsit
4mo ago
Would the US government have slapped Anthropic with this export control if Anthropic never fearmonger'ed about Mythos? I think the answer is very likely no. But is this the type of regulation Anthropic has been asking for? Not at all.
38.
▲
by
sothatsit
4mo ago
The big AI labs are also accumulating huge datasets of expert work in a wide range of fields, which is very expensive to re-create. It seems pretty plausible that this this gives them a big advantage that is compounded by their larger train
39.
▲
by
sothatsit
4mo ago
It is gone for me now. > There's an issue with the selected model (claude-fable-5). It may not exist or you may not have access to it.
40.
▲
by
sothatsit
4mo ago
Seems to just be a bigger model.
41.
▲
by
sothatsit
4mo ago
The Team plan is ~125 USD / month / user. Big enterprises like Uber are paying upwards of $1500 USD / month / user. Anthropic can raise their revenue a lot more by selling to big enterprises than they can by selling more
42.
▲
by
sothatsit
4mo ago
Definitely, it is quite an extreme change. But the upsides of better access to support and advice are huge, even if the potential downsides are scary as well. This feels like one area where we need better transparency and regulation due to
43.
▲
by
sothatsit
4mo ago
It is not solely or even primarily the big AI labs that would need to prepare. They have a better idea of what’s coming, and they’re positioned to benefit from it. It is governments, big companies, and individuals who could all experience f
44.
▲
by
sothatsit
4mo ago
I gave GPT-4 some source code and my existing tests, and asked it to write a new test, and it did it! It didn’t even run straight away, I had to fix it, but it still blew my mind. Later, I wrote a ~5k line proxy for work in C, and gave the
45.
▲
by
sothatsit
4mo ago
I would be very surprised if this is an actual thought-out PR strategy. I am far more inclined to believe that their employees are just bought-in to the future where AI is genuinely transformative. Whether they are right of wrong is another
46.
▲
by
sothatsit
4mo ago
Or: Anthropic genuinely believes the future scenarios they outline are realistic possibilities, and they want more people to take them seriously.
47.
▲
by
sothatsit
4mo ago
Maybe my bar for what constitutes a breakthrough is lower than other people's, but all of these seem like breakthroughs to me: NLP as a field saw huge shifts. NLP tasks that used to be complex and inaccurate can now be setup very easil
48.
▲
by
sothatsit
4mo ago
You can claim the use of AI is unethical, or the work as derivative, but AI being used as a tool in no way precludes something from being art. It is thought provoking and challenging, it seems like textbook art to me, and it’s clearly struc
49.
▲
by
sothatsit
4mo ago
This completely ignores all the other huge costs the AI labs are paying in data center builds, researcher salaries, experiments, and training models. The fact that Anthropic is rumoured to have a profitable quarter indicates that their marg
50.
▲
by
sothatsit
4mo ago
You must not be using coding agents. You can sneeze and spend $1 on Opus in Claude Code.
51.
▲
by
sothatsit
4mo ago
Enterprises are paying API prices, which are ~9x the price of the plan for the same usage. A lot of people on the plans are not maxing them out either.
52.
▲
by
sothatsit
4mo ago
I don’t remember ever hearing Dario or Sam recommend replacing people. Rather they say that smaller groups of people can do more work, so hiring will slow because small teams can do more. The only times when people talk about actual full re
53.
▲
by
sothatsit
5mo ago
Or, tokens are more like energy and prices will drop over time until they reach some equilibrium. The big labs are actively moving into the application layer, where they’ll have more pricing power. Maybe that layer will end up with a Mac (A
54.
▲
by
sothatsit
5mo ago
I think the distinction is that for experiments and prototypes the behaviour of the final system is what we are trying to design. We can experiment and see the tradeoffs and explore the design space before committing to a direction. And the
55.
▲
by
sothatsit
5mo ago
Well you have obviously already made up your mind, so have fun with your confirmation bias. We'll all be over here having a good time, getting more work done. Feel free to come over when you put down your grudge.
56.
▲
by
sothatsit
5mo ago
The entire mistake you are making is comparing using AI to skimming textbooks, or taking shortcuts. Your entire premise is wrong. People who care about craft will care about the quality of what they produce whether they use AI or not. The c
57.
▲
by
sothatsit
5mo ago
You don’t need to write code by hand to learn from iterations and experiments. I run more experiments and try out more different solutions than I ever could before, and that leads to better decisions. I still read all the code that gets shi
58.
▲
by
sothatsit
5mo ago
It still surprises me how effective the /simplify skill is. I’ve also had some great results with a /reflect skill that asks the agent to look at the work in the broader context of the project. But those are the only two skills I
59.
▲
by
sothatsit
5mo ago
> I'd say that by purging stuff from the brain we are losing thinking itself The idea that there will be less to think about seems a bit short-sighted. Humans are very good at moving to higher levels of abstraction, often with more
60.
▲
by
sothatsit
6mo ago
You can’t make up your mind about a model by using it on one task. Especially to say it’s such a bad downgrade after that is ludicrous. I’ve had great experiences with it this morning.
More ›