Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
hagen1778
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
hagen1778
5mo ago
I work at VictoriaMetrics. Just to clarify: VictoriaMetrics doesn't use bots for HN or for any other media for promotion. I don't know the person who you responded to. Most of the activity you see is coming from community members
2.
▲
by
hagen1778
6mo ago
I am curious to see more tests on the reading path. The article mentions matching 500 series over 6h window with 1m step - and it takes 2s for warmed caches. That doesn't sound good at all. Especially nowadays, when metrics from k8s ra
3.
▲
by
hagen1778
6mo ago
I was under impression that problem of zero injection was solved with Start Timestamp from OpenMetrics 2.0 spec - see https://prometheus.io/docs/specs/om/open_metrics_spec_2_0/#s...
4.
▲
by
hagen1778
6mo ago
Comparing self-hosted prices with managed solutions isn't exactly apples to apples. But if you do compare, VictoriaMetrics cloud for 3Mil active series and twice higher ingestion rate (100K samples/s or 30s scrape interval) will c
5.
▲
by
hagen1778
6mo ago
What do you use instead of Prometheus?
6.
▲
by
hagen1778
7mo ago
Disclaimer: I am affiliated with VictoriaMetrics. > The test setup is divergent from the real world (single node k8s cluster) The setup was chosen to simplify the suite, so it can be easily run anywhere. In real world, log collectors are
7.
▲
Benchmarking Kubernetes Log Collectors: Vector, Fluent Bit, OpenTelemetry
(victoriametrics.com)
3 points
by
hagen1778
7mo ago
|
6 comments
8.
▲
by
hagen1778
7mo ago
The benchmark suite is available here: https://github.com/VictoriaMetrics/log-collectors-benchmark
9.
▲
by
hagen1778
11mo ago
Using "period" triggers me :) If Mimir is the only one, why Roblox, GrafanaLabs's customer, isn't using Mimir for monitoring? They're using VictoriaMetrics on approx scale of 5 Billion active time series. See https
10.
▲
by
hagen1778
11mo ago
Just to add, VictoriaMetrics covers all 3 signals: - VictoriaMetrics for metrics. With Prometheus API support, so it integrates with Grafana using Prometheus datasource. It has its own Grafana datasource with extra functionality too. - Vict
11.
▲
by
hagen1778
11mo ago
There are plenty of ways to scale Prometheus: - Thanos - Mimir - VictoriaMetrics All of them provide a way to scale monitoring to insane numbers. The difference is in architecture, maintainability and performance. But make your own choices
12.
▲
by
hagen1778
11mo ago
I think OTEL has made things worse for metrics. Prometheus was so simple and clean before the long journey toward OTEL support began. Now Prometheus is much more complicated: - all the delta-vs-cumulative counter confusion - push support fo
13.
▲
by
hagen1778
1y ago
My understanding is that with Prometheus+Grafana, and the rest of their stack, you can achieve the same functionality as Datadog (or even more) at much lower costs. But, it requires engineering time to set up these tools, monitor them, buil
14.
▲
by
hagen1778
1y ago
Ofc you need to monitor your monitoring, because you run it. Datadog runs their own systems and monitors them, that's why they charge you so much. I barely can imagine a criticial piece of software that I need to run and not monitor i
15.
▲
by
hagen1778
2y ago
> Our tests revealed that Prometheus v3.1 requires 500 GiB of RAM to handle this workload, despite claims from its developers that memory efficiency has improved in v3. AFAIK, starting from v3 Prometheus has `auto-gomemlimit` set by defa
16.
▲
by
hagen1778
2y ago
It usually comes with increase of active series and churn rate. Of course, you can scale Prometheus horizontally by adding more replicas and by sharding scrape targets. But at some point you'd like to achieve the following: 1. Global q
17.
▲
by
hagen1778
2y ago
What makes you think that about docs? Of course, it was written by developers, not tech writers. But anyway, what do you think can be improved?
18.
▲
by
hagen1778
2y ago
We use the same approach in time series database I'm working on. While file creation and fsync aren't atomic, rename [1] syscall is. So we create a temporary file, write the data, call fsync and if all is good - rename it atomical
19.
▲
by
hagen1778
2y ago
ClickHouse recently got the support of TimeSeries table Engine [1]. It is marked as experimental, so yes - early stage. This engine is quite interesting, the data can be ingested via Prometheus remote write protocol. And read back via Prom
20.
▲
by
hagen1778
2y ago
Here you go https://victoriametrics.com/blog/mimir-benchmark/ It is from Sep 2022, it would be great to get newer results.
21.
▲
by
hagen1778
3y ago
> And yet people use ClickHouse quite effectively for this very problem There is no doubt that ClickHouse is a super-fast database. No one stops you from using it for this very problem. My point is that specialized time series databases
22.
▲
by
hagen1778
3y ago
Storing telemetry efficiently is only part of what Monitoring is supposed to do. The other part is querying: ad-hoc queries, dashboards, alerting queries executed each 15s or so. For querying to work fast, there has to be an efficient index
23.
▲
by
hagen1778
3y ago
> Would it be possible to do this in Postgres as well? Of course! The question is only in your requirements. Keeping a simple counter with limited cardinality should work just great. But nowadays monitoring is much more serious than that
24.
▲
by
hagen1778
3y ago
> mimir because of scale & self-host options Have you looked at VictoriaMetrics [0] before opting for Mimir? [0] https://victoriametrics.com/blog/mimir-benchmark/
25.
▲
by
hagen1778
3y ago
> What is it that you think prometheus offers over other solutions? I like Prometheus and think this is a great piece of software. But even if we won't go into actual details, Prometheus is baked in into Kubernetes monitoring [0]. T
26.
▲
by
hagen1778
3y ago
> Prometheus only handle aggregated data, though. That's not true. You're referring to pull-based approach for metrics collection. It has its tradeoffs (like fixed interval scraping), but has a lot of benefits too (like higher
27.
▲
by
hagen1778
3y ago
That's only a matter of time when they hire younger engineers who are familiar with modern monitoring systems and eager to apply their knowledge in practice.
28.
▲
by
hagen1778
3y ago
If I'm reading this [0] right, there will be no a standalone OS influxdb 3.0 version. So there's no point in comparing. I also wonder if it would be allowed to publish benchmarks of ENT version by 3rd-parties. [0] https:/&
29.
▲
by
hagen1778
3y ago
Yes! Alerting and recording rules are supported by vmalert [0]. vmalert then integrates with alertmanager for sending alerts, and alertmanager then dispatches notifications. Besides this, vmalert has features of retro-active rules evaluatio
30.
▲
by
hagen1778
3y ago
> Our proof-of-concept trial showed dramatically reduced compute and storage costs, translating into a 10x lower AWS bill
More ›