5 ms·
stan and edward dev here. happy to answer any questions. (shakir's blog posts are amazing; i recommend them all.)
by proditus 10y ago
stan and edward dev here. happy to answer any questions.
(shakir's blog posts are amazing; i recommend them all.)
- marmaduke 10y agoWhoa cool but the build is failing tisk tisk. A big issue I ran into with stan even with advi was scaling to large datasets since it (and Eigen) are single threaded. Would Edward answer all my prayers? When is Riemannian HMC going to arrive?
- proditus 10y agoi'm assuming you're referring to building edward? installation is a bit of a pain because tensorflow is not on pypi yet. please take a look here: http://edwardlib.org/troubleshooting http://edwardlib.org/troubleshooting edward should answer some of your prayers :) there's still some time until stan goes parallel/gpu, though there's lots of interest there. riemannin hmc is likely just around the corner!
- murbard2 10y agoVery cool, many questions 1) Why create a project distinct from Stan? Was it the prospect of benefiting of all the work going into TF and focus solely on the sampling procedures rather than autodiff or GPU integration? 2) Are you implementing NUTS? 3) Any plans to implement parallel tempering 4) Any plans to handle "tall" data using stochastic estimates of the likelihood?
- proditus 10y agogreat questions. 1. you touch upon the right strengths of TF; that was certainly one consideration. edward is designed to address two goals that complement stan. the first is to be a platform for inference research: as such, edward is primarily a tool for machine learning researchers. the second is to support a wider class of models than stan (at the cost of not offering a "works out of the box" solution). our recent whitepaper explains these goals in a bit more detail: https://arxiv.org/pdf/1610.09787.pdf https://arxiv.org/pdf/1610.09787.pdf 2) no immediate plans. but we have HMC and are looking for volunteers :) 3) same answer as above :) should be relatively easy to implement tempering. 4) this is already in the works! stay tuned!
- marmaduke 10y agoWhy TF instead of Theano as PyMC3 has done? Shouln't it be straightforward to port PyMC3 algos over TF? My main gripe with Theano is that OpenCL support is near non-existent, but this is also the case with TF.
- murbard2 10y ago4) which approach are you using? Generalized Poisson Estimator, or estimating the convexity effect of the exponential by looking at the sample variance of the log likelihood? The former is more pure, the latter may be more practical if ugly.
- proditus 10y agotheses are great insights. our first approach is the simplest: stochastic variational inference. consider a likelihood that factorizes over datapoints. stochastic variational inference then computes stochastic gradients of the variational objective function at each iteration by subsampling a "minibatch" of data at random. i reckon the techniques you suggest would work as we move forward!
- murbard2 10y agoEdit: ah never mind, variational inference, got it! I was thinking stochastic HMC! --- Ok but that will get an unbiased estimate of the log-likelihood. MCMC or HMC do work with noisy estimators, but they require unbiased estimates of the likelihood. At the very least, you need to do a convexity adjustment by measuring the variance inside your mini batch. Or you can use the Poisson technique which will get you unbiased estimates of exp(x) from unbiased estimates of x (albeit at the cost of introducing a lot of variance).