Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
XCSme
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
31.
▲
by
XCSme
1mo ago
What model are you using? I used GPT 5.6 Sol Extra High (Fast) for the last month, around 4-5h per day, and it managed to acomplished most of the tasks it had. It didn't really impress and often the end result needed one or two tweaks&
32.
▲
by
XCSme
1mo ago
That's a good idea. I might go more into table-tennis coaching, probably people would still prefer to be coached by a real person and not a robot for a very long time.
33.
▲
by
XCSme
1mo ago
It's like telling a painter who polished their craft and every brush stroke for 20 years, that now got access to a robot that does the painting, to be more creative in what it tells the robot to paint. Imagine Bob Ross prompting a robo
34.
▲
by
XCSme
1mo ago
Yes, many projects, used in production at various scales. Some even for large clients, and they work well in production, clients are happy, yet I feel no connection to the work. If the client says something, I just copy-paste it to the AI a
35.
▲
by
XCSme
1mo ago
That's what I am doing and usually telling others to: don't prompt the AI to do something, prompt the AI to create a system that does that something. It's true, it's fun to have those systems, maybe I care more about the
36.
▲
by
XCSme
1mo ago
That's good to hear. How long and much have you been using LLMs for? I think it takes ~1 year of heavy usage for this feeling to set in, it's like the 5 stages of grief, there's different phases. Now a big overwhelm come from
37.
▲
by
XCSme
1mo ago
It says (none), which usually means reasoning disabled. I would be surprised if none = use default reasoning
38.
▲
by
XCSme
1mo ago
But what product? If the world can also simply ask for the product they want, instead of searching for it? They won't even have to ask for a specific product, they will just state their problems/needs.
39.
▲
by
XCSme
1mo ago
Before it was fun because I was learning useful things for the future. Now it feels like whatever I learn will be obsolete in 2 months.
40.
▲
by
XCSme
1mo ago
I meant the leaders, almost all AI company CEOs voiced concerns for a long time, and even more now.
41.
▲
by
XCSme
1mo ago
Isn't the point of AGI that it can basically do any job?
42.
▲
by
XCSme
1mo ago
One thing that still stands today, is that even vibe-coding a good product takes time and thousands of dollars in tokens costs. Software will be more like a "proof of work", where people would still pay $100 for good software that
43.
▲
by
XCSme
1mo ago
Yes, that's cool and useful. Creating stuff for ourselves, for our own use. But we are social animals, we like sharing. Before it was cool to share an app you made, but now? What's the point of sharing an app, if the other person
44.
▲
by
XCSme
1mo ago
But was there ever a technology that even the people working on it said it's making them feel depressed and scared?
45.
▲
by
XCSme
1mo ago
True, but for many domains where my knowledge is limited, the LLMs beat me at imagination and taste too...
46.
▲
by
XCSme
1mo ago
The same how Magnus Carlsen says he never plays chess against a computer, because it makes no sense to do it.
47.
▲
by
XCSme
1mo ago
But how would you feel if you imagined hitting the ball and a robot arm hit it instead? Because that's how creating software is starting to feel.
48.
▲
by
XCSme
1mo ago
Well, for the same reason playing chess vs a person is more fun than doing chess puzzles, if we follow that analogy. Also, creating something with AI doesn't really feel like you made it yourself. And, if you make it without AI, most o
49.
▲
by
XCSme
1mo ago
In my own tests on aibenchy.com, where questions are quite simple, higher reasoning efforts consistently used to do worse than medium for most models. The reasoning effort should match the complexity of the task against the model's cap
50.
▲
by
XCSme
1mo ago
But how this transition will even happen? Soon we will have some machines that can replace 50% of jobs, and this will happen basically overnight...
51.
▲
by
XCSme
1mo ago
No, that's why I just play against other humans. In this game of work/development, you can't make sure that other humans don't "cheat". Our work won't compete anymore with other human's work, but with
52.
▲
by
XCSme
1mo ago
It's fun, but every new model release makes me even less interested to create cool stuff. Like, what's the point, if the next AI can do it in 5 seconds?
53.
▲
by
XCSme
1mo ago
We need CancerBench
54.
▲
by
XCSme
1mo ago
The no-reasoning version scores 35% while the low reasoning one scores 17%? What?
55.
▲
by
XCSme
1mo ago
That's quite common with many models, after "High" reasoning, over-thinking starts occurring and the model skips over the right solution by convincing itself otherwise.
56.
▲
by
XCSme
1mo ago
Is anyone actually upset that they are down? lol
57.
▲
by
XCSme
1mo ago
Is it DNS, as always?
58.
▲
by
XCSme
1mo ago
They escaped, it's happening!
59.
▲
by
XCSme
1mo ago
Nice, Muse Spark is so good and keeps improving, but it's still not the best choice for any use-case. The Sol models are in their own league currently in terms of cost/speed/performance. Good improvements from 1.1 and 1.2[0],
60.
▲
by
XCSme
1mo ago
Gemini 3.8 flash seems to be especially low efficiency in tool calling for some reason.
More ›