11 ms·
Facebook AI Research Team Open Source DeepMask and SharpMask
- spitfire 10y agoThis is very timely for me. Right at the moment I'm working on image segmentation for medical imaging. If I can draft off tech developed to sell more ads to do some good I'm all for it.
- dawson 10y agoI have a colleague working on something similar, drop me an email and I'll introduce you.
- dharma1 10y agoIt would be nice to see results compared to other SOTA semantic segmentation approaches like https://github.com/daijifeng001/MNC https://github.com/daijifeng001/MNC https://bitbucket.org/aquariusjay/deeplab-public-ver2 https://bitbucket.org/aquariusjay/deeplab-public-ver2 https://arxiv.org/abs/1607.07671 https://arxiv.org/abs/1607.07671
- Happpy 10y agoAdditional patent grand + bsd. Why not apache 2.0 or mpl 2.0?
- minimaxir 10y agoNB: The libraries use Torch. I'm still playing around with fasttext (which is amazing btw) officially announced last week so I'm surprised to see Facebook Research announce and release another project so soon.
- thewhitetulip 10y agoCould you share what you did with fasttext? It'd be great to see some examples around the amazing tools being open sourced! I read that there are no documentation comments, maybe they were removed?
- minimaxir 10y agoFor basic use, you can construct text vectors faster/better than word2vec/doc2vec. The Python package which serves as an API to fasttext (https://pypi.python.org/pypi/fasttext/0.7.2 https://pypi.python.org/pypi/fasttext/0.7.2 ) is well documented, is constantly being updated, and is easy to use.
- binarymax 10y agoDoes/can fastText use GPGPU at all? I use a CUDA cbow word2vec [1] that is really quick, and has served me quite well. Hard to know if switching to something more accurate would be worth giving up the speed :) [1] https://github.com/ChenglongChen/word2vec_cbow https://github.com/ChenglongChen/word2vec_cbow
- minimaxir 10y agoNo, fastText is CPU only.
- riyadparvez 10y agofastText is so fast, you may not need to use the GPU. That was their goal from the very beginning.
- iraphael 10y agoThese seem to be the papers for DeepMask [1] and SharpMask [2] [1] https://arxiv.org/abs/1506.06204 https://arxiv.org/abs/1506.06204 [2] https://arxiv.org/abs/1603.08695 https://arxiv.org/abs/1603.08695
- deleted 10y ago[deleted]
- said 10y agoIs there any way for those of us with average intelligence to contribute to tools like this? I know I'm a decent developer, but I feel entirely entirely inadequate to participate in this enormous, scary world of AI.
- magicalist 10y ago> Is there any way for those of us with average intelligence to contribute to tools like this? The people doing this aren't magical geniuses; they've just put the time and work into the subject and have been able to get themselves into a position they can do this all day surrounded by others they can collaborate with. As with most human endeavors, the trick is to just get started, and not get frustrated and give up when it turns out you don't know anything at first. Some people don't mind starting with a ton of abstract learning about the subject, others prefer trying to accomplish specific tasks, learning the theory along the way. For your specific question, as with all software, there's likely a lot to be done that has little to do with the main task of the tool and is just the everyday tasks of ease of use, interoperation with other tools, testing, etc. If on the other hand you want to get into the scary world of AI, like many others I'd recommend the Coursera machine learning course as a great place to start[1] [1] https://www.coursera.org/learn/machine-learning https://www.coursera.org/learn/machine-learning
- said 10y agoThank you for the link that course!
- bratsche 10y ago> The people doing this aren't magical geniuses Says the person named 'magicalist'. :)
- _audakel 10y agodouble upvote
- BOBOTWINSTON 10y ago
- wibr 10y agoLooks very promising. Self-driving cars will need algorithms like this to understand as much as possible of the environment that they see through the cameras. I think currently things like those ropes that the shepherd is holding in the picture are still very difficult to detect and classify correctly but they could just as well be a line across the road, maybe with a sign on it saying "closed".
- tectonic 10y agoFuture Facebook AR goggles will float people's names above their heads, replace billboard advertisements with live Facebook ads, and make people you 'mute' become invisible. It'll be a weird world.
- T-A 10y ago> make people you 'mute' become invisible That would be dangerous. They could follow the SL example and turn them into cardboard-like figures, or get more creative and let you dress them up as clowns, Barney the Dinosaur or some other avatar of your choice. :)
- deleted 10y ago[deleted]
- atombath 10y agoI can 'disappear' annoying people and relatives? sign me up
- azeirah 10y agoBlack mirror?
- timpark 10y agoFor those not familiar, specifically the "White Christmas" episode: http://www.imdb.com/title/tt3973198/?ref_=ttep_ep1 http://www.imdb.com/title/tt3973198/?ref_=ttep_ep1 For those who are familiar, new episodes are apparently arriving on Oct.21st. A hackathon project inspired by White Christmas: http://jonathandub.in/cognizance/ http://jonathandub.in/cognizance/ (Brand Killer)
- greenyouse 10y agoOr maybe more like Dennou Coil :) https://en.wikipedia.org/wiki/Denn%C5%8D_Coil https://en.wikipedia.org/wiki/Denn%C5%8D_Coil
- deleted 10y ago[deleted]
- guelo 10y agoI feel less and less willing to upload personal photos to social media sites where they will stick around forever attached to my identity while advancing computer vision techniques extract more and more information from them.
- dharma1 10y agoNot just social media sites - free searchable storage on Google Photos for instance (as nice as it is), is in exchange for Google to mine the shit out of your photos
- achr2 10y agoHow 'deep' are the networks used in something like DeepMask, and how does it compare with the number of layers of the human brain?
- nickparker 10y agoAs I understand it, you can't meaningfully ask about the number of layers in the human brain. Layers are an abstraction we use to make artificial neural networks dramatically faster to work with using linear algebra. By keeping each node interacting only with the layers above and below, we make the computations a lot nicer for our model of computation. In the brain, neurons link freely to other arbitrary neurons based on adaption processes we (or at least I) don't really understand. The brain also lacks a clear idea of direction to count the layers along, since it has innumerable different inputs coming in at all times, and the resulting signals interact all over the place. The most meaningful analog would probably be to ask "How many neuron firings typically occur between an external stimulus and a response to that stimulus?" Even that is extremely rough though, because through evolution a lot of 'short circuit' structures have formed in our bodies. The gag reflex is obviously triggered by sensory input, but it probably doesn't check with your frontal lobe before firing the appropriate muscles.
- dharma1 10y agoThe architecture of the brain is not really known and quite different to deep convolutional networks - but roughly 100 million neurons in a human brain, each one with thousands of synapses. It's several orders of magnitude more "neurons" and "connections" than even the largest ANN's
- sushirain 10y agoFor example, ResNet from 2015 had 152 layers. A real neuron takes in the order of 10ms to integrate and fire to the next neuron. Many subconscious reactions take less than 1 sec, which leaves time to a chain of length less than 100. Note that those neurons are not strictly arranged in layers. The human visual cortex has 10^12 synapses [1]. One popular 2015 deep learning net (ResNet 152-layers) used 10^12 FLOPs to classify objects in one image (but less weights.) In terms of depth, we're there. In terms of breadth, it will take several years. But the brain does things very differently. For example, it has top-down signals during "prediction." [1] http://www.ncbi.nlm.nih.gov/pubmed/7244322 http://www.ncbi.nlm.nih.gov/pubmed/7244322
- msie 10y agoWow, the code is in Lua! Just a little surprised. Also I just read some LinkedIn post on the dearth of Haskell education in schools. Always on the lookout for more companies using Haskell and hoping in the field of AI, machine-learning.
- avvakum 10y agoSharpMask looks very similar to a year-old "U-Net" http://arxiv.org/pdf/1505.04597 http://arxiv.org/pdf/1505.04597
- mtourne 10y agoI thought that too. Just finished a kaggle competition involving segmentation, like a lot of participant I used one form of U-net (my own implementation). You can probably find a lot of u-net implementations from this contest. One that performed really well [1]. It uses 'inception style' blocks feature extraction instead of vgg. But otherwise pretty similar. [1] https://github.com/EdwardTyantov/ultrasound-nerve-segmentation https://github.com/EdwardTyantov/ultrasound-nerve-segmentati...
- visarga 10y agoIs it safe to assume this release, as well as fastText are meant as PR for hiring?
- ChristianGeek 10y agoWhat's the performance like? Fast enough to handle live video with a high-end PC?
- swframe 10y agoIs there a link to get the code?
- pmontra 10y agoClick the links in the "We're making the code for DeepMask+SharpMask as well as MultiPathNet" line. They should have made them display more prominently. They require to login to FB, then redirect to GitHub https://github.com/facebookresearch/deepmask https://github.com/facebookresearch/deepmask https://github.com/facebookresearch/multipathnet https://github.com/facebookresearch/multipathnet