Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
vikramkr
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
26 ms
·
61.
▲
by
vikramkr
2mo ago
Then why are fable form anthropic and the 5.6 sol line from openai so much more token efficient than other models? I really don't get this take at all - they're supply constrained right now, and there's literally no economic
62.
▲
by
vikramkr
2mo ago
Unfortunately the one thing they seem to keep not letting their competitors our play them at is coming out with really damn good models. Not consistently mind you - openai had a really solid lead for a few months earlier this year when it l
63.
▲
by
vikramkr
2mo ago
IMO firm wording against non generative uses of AI or lightweight uses like tab complete will go beyond just influencing good faith actors to reduce llm usage - it'll just influence them to not contribute instead. If you make it clear
64.
▲
by
vikramkr
2mo ago
The wording of "written with ai tools" implies generation, as do all the concerns about copyright etc. it would be beyond insane to attempt to ban asking a chat bot or an agent questions to navigate or understand a codebase, not i
65.
▲
by
vikramkr
2mo ago
Yeah there's a brand premium. Like literally with apples the ones with trademarked names can cost more. And the ones with trademarked names have an organization behind them that promote that apple variety and set fees etc for growing a
66.
▲
by
vikramkr
3mo ago
The idea that humans have undergone high selective pressure for genes related to immunity in some way during a recent 5000 year period in which massive urbanization lead to a truly unfathomable amount of deaths including staggering percenta
67.
▲
by
vikramkr
3mo ago
Honestly the models are rled so hard on specific synthetic datasets and specific behaviors/personalities that I would be concerned that trying to change its behavior like this would hurt output quality. It's a tool, I don't c
68.
▲
by
vikramkr
3mo ago
The competition is between openai and anthropic, if there are price agreements between them that's absolutely price fixing. Or if there's collusion between the cloud providers to inflate compute. I would expect Amazon and gcp to b
69.
▲
by
vikramkr
3mo ago
Did they actually ever cut the price on gpt 4? The oldest versions of it in the api still seem stupidly expensive? There were definitely price cuts as they introduced the turbo models and stuff, and new versions of each model might have got
70.
▲
by
vikramkr
3mo ago
I doubt there's any sort of criminal behavior there - the model is anthropic's up and anthropic probably charges a very expensive license fee that's the same for all of them, and their cogs on compute aren't going to be
71.
▲
by
vikramkr
3mo ago
I doubt it's more profitable - ai overview is free/runs even when signed out in incognito, while the frontier models from openai and anthropic are nauseatingly expensive, especially for enterprise, and have a ton of users who are
72.
▲
by
vikramkr
3mo ago
I think they've been behind for a while - flash isn't even that competitive with glm 5.2 and from their hype around 3.5 flash at launch - that was certainly not intended to be the case
73.
▲
by
vikramkr
3mo ago
The tokens served number might include cache tokens which are a huge chunk of agentic token spend - and even with that the estimated burn rate for anthropic doesn't seem wildly off? They spend 1.25 bn per month on their deal with Space
74.
▲
by
vikramkr
3mo ago
Whether or not the specific policy is good my preference is that changes to policy that have been in force for decades happen based on legislation and not the whims of 9 unelected people. We didn't get clear rules made my legislature,
75.
▲
by
vikramkr
3mo ago
Cheveron deference is no longer active unfortunately
76.
▲
by
vikramkr
3mo ago
Not really - it doesn't seem to be much of a leap to be able to transfer that knowledge to new language. Honestly there's not a lot that's fundamentally new under the sun in new programming languages and solving a problem in
77.
▲
by
vikramkr
3mo ago
The best AI models at this point are perfectly fine handling new languages and frameworks and stuff - as long as you can point the llm to some docs and some example code it's not gonna do noticeably worse than it would at another langu
78.
▲
by
vikramkr
3mo ago
That's technically true but practically an agent can just save the script file/rerun it/write a tool that lets it call a reply with memory etc.This aspect is a bit more elegant when it's in the execution context, but the
79.
▲
by
vikramkr
3mo ago
If 2 out of 100 people I know see a broken website, depending on the website, that's fine, that doesn't sound like a big deal. Now, if out of 10 power users, all ten of them see a broken site once every 50 logins? Thats a much big
80.
▲
by
vikramkr
4mo ago
That sounds like you're just creating an artificial distinction. If you let other engineers merge their code without looking at it too closely nothing makes the AI any different other than pretending it's "your code" eve
81.
▲
by
vikramkr
4mo ago
That's understating the progress by a lot - many cancers are a lot more survivable now than previously with better chemo/radio/surgery + immunotherapies + car-t etc etc
82.
▲
by
vikramkr
4mo ago
Point them at for what?
83.
▲
by
vikramkr
4mo ago
https://www.thermofisher.com/us/en/home/life-science/antibod... > Moving forward, where an original image is not present or available, the Company will ensure that website users are informed that anti
84.
▲
by
vikramkr
4mo ago
What's wrong with the cloud? I get the point about a hype cycle but those two examples don't seem even remotely in the same universe of similarity. SQL is still around, but the cloud won pretty comprehensively no? If you're s
85.
▲
by
vikramkr
4mo ago
It's not just about returns, it's also about risk. The role of a passive index fund is to be a passive index fund. If the s&p starts chasing returns, that will reduce its utility to the market. You get higher returns by being
86.
▲
by
vikramkr
4mo ago
The post training is meant to make it more steerable (usually). I might not want it to write tests. I might not have a dev environment set up for an agent to run tests in it's loop. A major goal in post training is to make it follow in
87.
▲
by
vikramkr
4mo ago
Why is that hard to believe? It's literally the prompt telling it what to do - if you want a poem about watermelons you tell it to write a poem about watermelons, if you want tests you tell it to write tests. It's not like TDD is
88.
▲
by
vikramkr
5mo ago
this model is whack. Exclamation marks everywhere, sycophantic - not producing working code on prompts the other models handle fine. "The reason it is echoing back your messages is because gpt-5.4-nano is a fictional model name!"
89.
▲
by
vikramkr
5mo ago
That's a list of like 6 things. And each of those less complicated a question then the seven thousand questions people throw at you when you complain about something not working right on a Linux distro or about speeding up build times
90.
▲
by
vikramkr
5mo ago
My take is there was one big inflection point around opus 4.5 when they got the agentic stuff working and now whether or not it works depends on whether your use case/area of software engineering is profitable enough for the companies
More ›