Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
aaronvg
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
aaronvg
2y ago
I doubt anyone would be confused with Traceloops the artist vs Traceloop the LLM Observability Platform
32.
▲
by
aaronvg
2y ago
[Another BAML creator here]. I agree this is an interesting direction! We have a "chat" feature on our roadmap to do this right in the VSCode playground, where an AI agent will have context on your prompt, schema, (and baml test r
33.
▲
by
aaronvg
2y ago
I think it depends on what you value as well, like DX. A large portion of our users switch to BAML because they actually "just want to see the damn prompt".
34.
▲
by
aaronvg
2y ago
My bad, I think I didnt explain correctly. Basically you have two options when a "," is missing (amongst other issues) in an LLM output which causes a parsing issue: - retry the request, which may take 30+ secs (if your LLM output
35.
▲
by
aaronvg
2y ago
[Other BAML creator here!] one time we told a customer to do this to fix small json mistakes but turns out their customers don't tolerate a +20-30s increase in latency for regenerating a long json structure. We instead had to write a p
36.
▲
by
aaronvg
3y ago
Not sure it's related to function calling. GPT4 can do function calling without using the specific function-calling API just by injecting the schema you want into the prompt with directions and asking it to return JSON. It works like &
37.
▲
by
aaronvg
3y ago
We felt this 100% so we built a DSL (called BAML) to solve the "prompt transparency" problem (amongst other issues). We have a VSCode playground that always shows you the full prompt -- kinda like a markdown preview works. We are
38.
▲
by
aaronvg
3y ago
(Other author of this blog post here) We actually do CPU inference. The SBERT models have a pretty small memory footprint -- you can fit a couple models on a t2.medium instance. On a C6.Large you can get 75ms inference. T2.medium is more ar
39.
▲
Ask HN: YouTube's website search isn't useful anymore, is there an alternative?
73 points
by
aaronvg
3y ago
|
31 comments
40.
▲
by
aaronvg
4y ago
sorry but I think the features you listed as being stolen seem pretty generic. For example "mobile IDE (ipad pro with keyboard)" doesn't tell me what specific IDE features you'd implement. In this case, Replit made the &
41.
▲
by
aaronvg
4y ago
Yep we encourage people to turn off our app (Gloo) a few days a week similarly to having "no meeting days". In the end, people can always decide how often they want to be reachable for.
42.
▲
by
aaronvg
4y ago
It's easier to be creative when you can bounce-off an idea with a coworker immediately, rather than wait for the Zoom meeting or wait for them to answer your Slack message. We're actually building a virtual office to make having t