Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
parmesant
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
parmesant
1y ago
Based on the feedback, we could have done a much better job with these results (lessons for our next experiment). But yes, the models were tested against the same dataset which was aggregated over different granularities (1 minute, 1 hour,
2.
▲
by
parmesant
1y ago
We'll definitely include it in our next experiment (shaping up to be quite big!)
3.
▲
by
parmesant
1y ago
At the moment our focus is on observability, hence the narrow scope of our dataset. A pretty good benchmark for observability seems to be Datadog's BOOM- https://huggingface.co/datasets/Datadog/BOOM But for g
4.
▲
by
parmesant
1y ago
we're grateful for the honest feedback (and the awesome resource!), makes it easier to identify areas for improvement. Also, your point about using multiple metrics (based on use-cases, audience, etc) makes a lot of sense. Will incorpo
5.
▲
by
parmesant
1y ago
That's actually one of the use-cases that we set out to explore with these models. We'll release a head-to-head comparison soon!
6.
▲
by
parmesant
1y ago
This looks like a great benchmark! We've been thinking of doing a better and more detailed follow-up and this seems like the perfect dataset to do that with. Thanks!
7.
▲
by
parmesant
1y ago
Author here, we're trying these out for the first time for our use-cases so these are great points for us to improve upon!