3 ms·
In my experience if you develop a "data science pipeline" forcing the data scientist to build - reproducible - validated - back-tested - easy to deploy mod
by NumberCruncher 9y ago
In my experience if you develop a "data science pipeline" forcing the data scientist to build
- reproducible
- validated
- back-tested
- easy to deploy
models, they are going to hate it. It just kills the fun and/or makes obvious if they made a mistake.
- CJefferson 9y agoI blame software. I don't understand why we couldn't have some system, perhaps using strace and friends, which tracks everything I ever do, and how every file was created. Then I could just say "how did I make X?"
- carlmr 9y agoMake it happen!
- gipp 9y agoSo we should sacrifice all the things that actually make a Data Scientist's work valuable in the name of fun and obscuring mistakes? Fun I almost get, obviously good for productivity (though I think you'd really be sacrificing productive output for non-productive output), but I just don't get where you're even coming from with the "making mistakes more obvious" angle.
- carlmr 9y agoI think he was being facetious. Of course we need all these things, but data science right now is still not that mature I guess.
- deleted 9y ago[deleted]