Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
a1j9o94
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
24 ms
·
1.
▲
by
a1j9o94
2mo ago
You are right that someone could go through that effort bat that only really makes sense if they give you some.idea how to use it. Reading the exchange it seems like you got caught up on the illustrative examples the other person used. If a
2.
▲
by
a1j9o94
6mo ago
Disclaimer I work at Zapier, but we're doing a ton of this. I have an agent that runs every morning and creates prep documents for my calls. Then a separate one that runs at the end of every week to give me feedback
3.
▲
by
a1j9o94
6mo ago
This is effectively how I treat my AI agents. A lot of the reason this doesn't work well for people today is due to context/memory/harness management that makes it too complex for someone to set up if they don't want a f
4.
▲
by
a1j9o94
6mo ago
You would only use the base model during training. This is a distillation technique
5.
▲
by
a1j9o94
6mo ago
I fall into this trap a lot. The platonic ideal argument is a fun mental exercise but doesn't get anything done
6.
▲
by
a1j9o94
7mo ago
This is an interesting space. Right now we've gotten to a point where agents can do most tasks, but they will get lazy/skip steps if you're not precise in the requirements. We need ways to validate that expands beyond softwar
7.
▲
Show HN: Sales Agent Benchmark – SWE-Bench for sales AI agents (open source)
(sales-agent-benchmarks.fly.dev)
1 points
by
a1j9o94
8mo ago
|
0 comments
8.
▲
by
a1j9o94
8mo ago
Honestly just didn't think about it. Added it.
9.
▲
I Tried to Give AI "Imagination" to Solve Physics Problems
(github.com)
2 points
by
a1j9o94
8mo ago
|
3 comments
10.
▲
by
a1j9o94
8mo ago
Hey HN, I spent the last few weeks exploring whether AI systems could benefit from generating video predictions before making decisions—like how humans mentally simulate "what happens if I pour this coffee?" before acting.
11.
▲
by
a1j9o94
9mo ago
Pretty much every major LLM client has web search built in. They aren't just using what's in their weights to generate the answers. When it gives you a link, it literally takes you to the part of the page that it got its answer fr
12.
▲
by
a1j9o94
9mo ago
I would argue that's just your coworker giving you a bad answer. If you prompt a chatbot with the right business context, look at what it spits out, and layer in your judgement before you hit send, then it's fine if the AI typed
13.
▲
by
a1j9o94
9mo ago
Honestly if you have a working relationship/communication norms where that's expected, I agree just send the 5 bullets. In most of my work contexts, people want more formal documents with clean headings titles, detailed risks even
14.
▲
by
a1j9o94
9mo ago
I know I'm an outlier on HN, but I really don't care if AI was used to write something I'm reading. I just care whether or not the ideas are good and clear. And if we're talking about work output 99% of what people were
15.
▲
by
a1j9o94
10mo ago
Not the person you're responding to, but I think there's a non trivial argument to make that our thoughts are just auto complete. What is the next most likely word based on what you're seeing. Ever watched a movie and guessed
16.
▲
by
a1j9o94
10mo ago
Having one tool that you can use to do all of these things makes a big difference. If I'm a financial analyst at a company I don't need to know how to implement and use 5 different specialized ML models, I can just ask one tool (t
17.
▲
by
a1j9o94
10mo ago
yy
18.
▲
by
a1j9o94
1y ago
The above is saying more precise not completely precise. The overall point they're making is you still are responsible for the code you commit. If they are saying the code in this project was in line with what they would have written,
19.
▲
by
a1j9o94
1y ago
Why do you say that? I would argue that as long as your tests and interfaces are clearly defined no reason it couldn't scale indefinitely.
20.
▲
by
a1j9o94
1y ago
I agree with this completely. I get the impression that a lot of people here think of software development as a craft, which is great for your own learning and development but not relevant from the company's perspective. It just has to
21.
▲
by
a1j9o94
1y ago
The point is devs aren't sales/client facing. So from the customers perspective, it's just a delivery detail.
22.
▲
by
a1j9o94
1y ago
How do you do things like compare prices in plain text?
23.
▲
by
a1j9o94
1y ago
Wouldn't phone books and catalogues count as advertising?
24.
▲
by
a1j9o94
2y ago
This is really interesting. Has anyone seen a good registry of MCP servers somewhere? What I would love to be able to do is ask for a task and have the agent figure out what servers and tools it needs to accomplish that task directly.
25.
▲
by
a1j9o94
2y ago
Probably not the whole model, but the first step was "fine tuning" the base model on ~800 chain of thought examples. Those were probably from OpenAI models. Then they used reinforcement learning to expand the reasoning capabilitie
26.
▲
by
a1j9o94
2y ago
Ideally yes, but most people in corporate settings don't communicate clearly. So if they have a culture of dropping hints, and you don't take them you're not going to do well at the company. It's not good, but it's
27.
▲
by
a1j9o94
2y ago
I was part of a group that did something like this. We used a crypto token system where if you joined the group you got X number of tokens and you could use that to get other people in the network to help with various tasks. Then the networ
28.
▲
by
a1j9o94
2y ago
Cursor isn't designed to do long running tasks. As someone mentioned in another comment it's closer to a function call than a process like Devin. It will only do one task at a time that it's asked to do.
29.
▲
by
a1j9o94
2y ago
It's also an American company building the project. The cultural values of the US are relevant.
30.
▲
by
a1j9o94
2y ago
It's HR's entire job to set policies for hiring. They can say a candidate has to have a college degree. Why wouldn't they have the right to set this policy as well?
More ›