Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
deepdarkforest
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
Economics and AI (Tom Cunningham)
(tecunningham.github.io)
2 points
by
deepdarkforest
1y ago
|
0 comments
32.
▲
by
deepdarkforest
1y ago
Eh i mean often innovation is made just by letting a lot of fragmented, small teams of cracked nerds trying out stuff. It's way too early in the game. I mean, qwens release statements have anime etc. IBM, Bell, Google, Dell, many did i
33.
▲
by
deepdarkforest
1y ago
The Chinese are doing what they have been doing to the manufacturing industry as well. Take the core technology and just optimize, optimize, optimize for 10x the cost/efficiency. As simple as that. Super impressive. These models might
34.
▲
by
deepdarkforest
1y ago
1. I was not talking about official MCP servers, those are often even free. Im talking about pricing of other devtools for aggregating tools/mcp's. I think this is an obvious space to build i agree, i just worry about differentiat
35.
▲
by
deepdarkforest
1y ago
1. Oh okay great, maybe clarify it in the pricing page? That mcp server call means just execute. But its' still 10x more expensive right? 2. From what i understand it's just nested search right? It is not anything different, if yo
36.
▲
by
deepdarkforest
1y ago
1. Interesting approach, but the pricing seems 1-2 orders of magnitude too expensive. For your example for slack, It contains 4 calls for an action. Pricing shows 100 dollars per 10k cals, so 1 cent per call. This means, for an agent that
37.
▲
by
deepdarkforest
1y ago
Of course not! But usually, you can quantify metrics for quality, like uptime, lost transactions, response time, throughput etc. Then you can have accountability, and remediate. Even for other bugs, you can often reproduce and show clearly
38.
▲
by
deepdarkforest
1y ago
Wow. Sneaky. They do not even state the rate of impact for the XLA bug afaik, which affected everyone, not just claude code users, very vague. Interesting. Claude code made almost half a billion so far[1] (>500m in ARR and its like 9 mon
39.
▲
by
deepdarkforest
1y ago
Looks like it’s time to go outside and touch some grass again
40.
▲
by
deepdarkforest
1y ago
They will probably also release sonnet 4.2 or something soon to make people jump back again to try it and hopefully restick
41.
▲
by
deepdarkforest
1y ago
Congrats! Doesn't replit have an integrated database as well? Lovable has supabase, and I'm pretty sure Base44 as well, plus other agent integrations.
42.
▲
by
deepdarkforest
1y ago
That and consumer robotics. The latter will explode if (big if) RL and llm reasoning get combined into something solid. Lots and lots of smart people are working on it already of course, we are seeing great improvements but nothing really u
43.
▲
by
deepdarkforest
1y ago
Here is the secret: Step 1: Vibecode 5 trash generic apps (eg AI interior designers, gpt wrappers) Step 2: Launch with paying 15k in google/meta ads Step 3: Receive back ~5k in revenue Step 4: Spam on twitter+linkedin clickbaity "
44.
▲
by
deepdarkforest
1y ago
Congrats! I think the space is very interesting, I was a founder of a similar windows CUA infra/ RPA agents but pivoted. My thoughts: 1) The funny thing about determinism is how deterministic you should be when to break, its kind of a
45.
▲
by
deepdarkforest
1y ago
> It handles complex coding tasks with minimal prompting... I find it interesting how marketers are trying to make minimal prompting a good thing, a direction to optimize. Even if i talk to a senior engineer, i'm trying to be specif
46.
▲
by
deepdarkforest
1y ago
Using LLMs as a critic/red teamer is great in theory, but economically is not that more useful, doesnt save that much time, if anything, it increases the time because you might uncover more errors or think about your work more. Which i
47.
▲
by
deepdarkforest
1y ago
I mean isn't that amazing for an 1 year old product? If it's already better than a terrible dev with an LLM, or better than a decent dev without an LLM, it's not hard to imagine in 2-3-5 years Devin is better and cheaper than
48.
▲
by
deepdarkforest
1y ago
This is a shockingly fresh idea. I get that this generates every pixel from scratch, unlike Gemini approaches. But, i wonder how do you think this type of neural OS would be able to communicate with the internet or other similar neural os.
49.
▲
by
deepdarkforest
1y ago
> The problem is in strcpy in the src files of curl.. have you seen the exploit code ?????? The worst part is that once they are asked for clarifications by the poor maintainers, they go on offense and become aggressive. Like imagine the
50.
▲
Zero-click exploit on Elevenlabs assistant through calendar
(repello.ai)
2 points
by
deepdarkforest
1y ago
|
0 comments
51.
▲
by
deepdarkforest
1y ago
I don't disagree, but the current sentiment i was referring to seems to be "maximize AI code generation with tools helping you to do that" rather than "prioritize code quality over AI leverage, even if it means limiting
52.
▲
by
deepdarkforest
1y ago
This is a launch by a YC company that converts enterprise cobol code into java. Maybe it's my fault, but i tried every single coding agent with a variety of similar tools and whenever i try to parallelize, they clash while editing fi
53.
▲
by
deepdarkforest
1y ago
It's very funny how many layers of abstraction we are going through. We have limited understanding of how LLM's work exactly and why. We now do post training with RL, which again, we don't have perfect understanding of it eit
54.
▲
by
deepdarkforest
1y ago
No worries. GGUF is more suitable for the latest open-source models, i agree there. Quant2/Q4 will probably be critical as well, if we don't see a jump in ram. But then again I wonder when/If mediapipe will support GGUF as w
55.
▲
by
deepdarkforest
1y ago
[flagged]
56.
▲
by
deepdarkforest
1y ago
I don't know about that. Zapier and automation apps were huge before agents, or even for integrations for Slack. There is definitely a big portion of tech products that have mutual benefits by providing good APIs to be in the same bubb
57.
▲
by
deepdarkforest
1y ago
Interesting. It's just an agent loop with access to python exec and web search as standard, BUT with premade, curated, 150 tools like analyze_circular_dichroism_spectra, with very specific params that just execute a hardcoded python fu
58.
▲
by
deepdarkforest
1y ago
>if your code doesn't work it doesn't work you can't bullshit a computer this is wrong. I would argue the difference between a junior dev/intern and a senior engineer is that while both can write code that works , th
59.
▲
by
deepdarkforest
1y ago
Haha we are in such a bubble. All engineering concepts are being thrown in the trash. We already have problems with prompt injections in MCP (eg supabase thread yesterday). Imagine the possibilities when you allow the LLM itself to write an
60.
▲
by
deepdarkforest
1y ago
I mean that's the dirty secret of any RAG chatbot. The concept of "grounding" is arbitrary. It doesn't matter if you use embeddings, or use a tool that uses your usual search and gets the top items, like most web search
More ›