4 ms·
> You can’t improve what you can’t measure. I can't stand this cliche. You can improve things you can't measure, we do this all the time, and the things worth
by cle 4y ago
> You can’t improve what you can’t measure.
I can't stand this cliche. You can improve things you can't measure, we do this all the time, and the things worth improving most are often hard to measure. Setting objective goals are important to keep us honest with ourselves, but in that same vein, we should continuously acknowledge that, unless we are working in hard sciences, we are usually measuring proxies to what we really care about, and it's often hard to pick a representative measure.
As an example:
> Measures can be as simple as number of experiments per person per team or something more complicated.
I've seen this exact scenario in a large company I used to work for. The outcome was that some teams would hit their goal by running garbage experiments. It was a net negative for the product, because garbage still occasionally shows statistical significance. Acknowledge that the measure is an imperfect proxy, identify in what ways the measure could fail to represent the true desired outcome, and control for those (in this case, e.g. some oversight on experiment quality).
- throw1234651234 4y agoIn the same vein, I never achieve goals I set. Working out is the usual example because it's so easy to see and measure results. But if I do something casually, it improves. If I set something as a goal and try to stick to a program, it always fails due to injury, overuse, or not enough load/volume.
- d0mine 4y agowhat does word "improve" mean if you do not measure anything? How would you know that something is better/worse? (no metric is perfect, there may be downsides but "we don't need no metrics" attitude makes no sense--it leads to burning witches in an attempt to improve the weather)
- mrcsd 4y agoSurely burning witches to improve the weather is exactly an example of metric use gone awry? To your point though, it is exceedingly common to know we are better at something without a clear metric. It is not controversial to suggest that someone may be better now at communicating than they were as a child, even though there is no clear way to define this.
- time_to_smile 4y agoI think you might be confusing "measure" with "observe". Measure in this discussion has a distinct quantitative implication. You've never had one steak you prefer to another without doing quantitative analysis on it? I love making cocktails and I can definitely tell better from worse, noting qualitative differences (a bit dry, too sweet, mouthfeel a bit thick) with absolutely no quantitative measurements involved. I can certainly tell the difference between a good and bad violinist without measuring anything. From a different perspective: many of my favorite movies have a 3.5 star rating on Amazon, in this case the measure does not correlate well with my sense of improvement.
- d0mine 4y agoTry a blind test for your cocktails (you could follow https://en.wikipedia.org/wiki/Lady_tasting_tea https://en.wikipedia.org/wiki/Lady_tasting_tea ). The results may surprise you.
- ip26 4y agoYou can improve things you can't measure, we do this all the time I would still consider an indirect measure or proxy measure of a thing to be "a measure of the thing".
- __alexs 4y agoYes if you interpret "measurement" strictly then you can't measure anything at all. I don't even have a ruler I can measure the length of something with that isn't in some way a proxy for the true length. The error bars of our measurements may be large, and the there may be many confounding variables to the measurement, but it is still a measurement. Having said that, I agree with the overall sentiment above. We often pick metrics that don't measure what we think they do or do it so inaccurately as to be useless. And at the same time we invest too little in finding the right metrics before they become ingrained in the org structure.
- nicbou 4y agoI improve my cooking all the time. That doesn't seem to fit here.
- time_to_smile 4y agoGoodhart's law[0] should be posted above every experimentation team. The idea that "When a measure becomes a target, it ceases to be a good measure" is something that every data scientist/statistician knows, but almost none heed in practice. The worse lead companies that I've worked for are the ones that claim to be "data driven". Countless dashboards showing various progress towards various targets without even a hint of understanding what the big picture might even be. One of the biggest insights I've had over a career working in data science is that the person solving a problem based on years of experience without any numbers backing their decisions almost always is making choices close enough to optimal that it isn't worth the extra energy to push it optimal. An example is that the person selling hot dogs at the park is probably pricing them nearly optimal. You could bring in a team of dynamic pricing experts, build a data center to mine costumer data, and I'm willing to bet a few hot dogs that difference between the model optimal price and what the hot dog vendor is selling is not enough to justify the cost of figuring out the difference. I likewise would not be surprised if the real, long run benefit of A/B testing does not justify the cost of both employee time and especially the SaaS products that help manage these processes... but let's not do that analysis because my salary depends on no one checking this. 0. https://en.wikipedia.org/wiki/Goodhart%27s_law https://en.wikipedia.org/wiki/Goodhart%27s_law
- placidpanda 4y agoThe main reason that I think I make good decisions in the absence of data, is because of how much I've relied on whatever data is available to inform my decisions and learn going forward. This is a huge advantage to multivariate testing as a practice/culture. As a consequence, it's often very easy for me to pick out when readouts are giving a deceptive answer (i.e. oh, the scope of this uplift is too much, we need to double check if something happened to negatively impact the control). I'm not sure I'd agree that people are often operating "close enough to optimal", but I would definitely agree that integrating experimentation is hard enough that sometimes the effort (or mistakes) you can introduce will cause more problems than you're helping. But I think this is more a function of how poor people are at the mechanics and the mindset of running experiments than that they're doing good enough pricing hot dogs. Experiments in many places are looked at for either CYA or boasting about quarterly results and not to truly learn/grow/improve.