6 ms·
Well done you seem to have liberated an open model trained on open data for blind and visually impaired people. Paper: https://arxiv.org/pdf/2204.03738 https:/
by jonpo 2y ago
Well done you seem to have liberated an open model trained on open data for blind and visually impaired people.
Paper: https://arxiv.org/pdf/2204.03738 https://arxiv.org/pdf/2204.03738
Code: https://github.com/microsoft/banknote-net https://github.com/microsoft/banknote-net
Training data: https://raw.githubusercontent.com/microsoft/banknote-net/refs/heads/main/data/banknote_net.csv https://raw.githubusercontent.com/microsoft/banknote-net/ref...
model: https://github.com/microsoft/banknote-net/blob/main/models/banknote_net_encoder.h5 https://github.com/microsoft/banknote-net/blob/main/models/b...
Kinda easier to download it straight from github.
Its licenced under MIT and CDLA-Permissive-2.0 licenses.
But lets not let that get in the way of hating on AI shall we?
- cess11 2y ago[flagged]
- jonpo 2y agoYes nothing wrong with cool software or showing people how to use it for useful things. Sorry I'm just kind of sick of the whole 'kool aid', 'rage against AI' thing a lot of people seem to have going on and the way is presented in the post. I have family members with vision impairment helped by this particular app so its a bit personal. Nothing against opening stuff up and understanding how it works etc. I'd just rather see people build/train useful new models and stuff with the open datasets / models already available. I guess AI kind of does pay my bills in a round about way.
- a2128 2y agoSadly companies will hoard datasets and model research in the name of competitive advantage. Obviously with this specific model Microsoft chose to make it open, but this is not always the case, and it's not uncommon to read papers or technical reports saying they trained on an "internal dataset"
- jonpo 2y agoCompanies do have a lot of data, and some of that data might be useful for training AI. but >99% isn't. When companies do release a cool model or paper that doesn't have open data, (as you point out for competitive or other reasons privacy etc) people can then help build/collect similar open datasets. Unfortunately companies generally don't owe you their data, and if they are in the business of making models they probably won't share the model either, the situation is similar to source code for proprietary LoB applications. but fortunately the best AI researchers mostly do like to share their knowledge and because companies want to attract the best AI researchers they seem to generally allow researchers to publish if its not too commercially sensitive. It could be worse while the competitive situation has reduced some visibility of the cutting edge science, lots of datasets and papers are still published.
- cess11 2y agoIn my view there was almost nothing like that in this article, besides the first sentence it went right into the technical stuff, which I liked. Compared to a lot of articles linked here it felt almost free from the battles between "AI" fashions. It seems dang thinks I mistreated you somehow, if you agree I'm sorry, it wasn't my intention.
- dang 2y agoCan you please edit swipes out of your HN comments? Your post would be fine with just the first sentence. This is in the site guidelines: https://news.ycombinator.com/newsguidelines.html https://news.ycombinator.com/newsguidelines.html.
- cess11 2y agoWhat do you mean, "swipe"? The other person agreed they'd misjudged the article and apologised several hours before you wrote this.
- dang 2y ago"Does 'AI' pay your bills" was a gratuitous personal attack.
- cess11 2y agoIs it? How? In your mind, does it imply some particular humiliation or something?
- dang 2y agoIt's a variant of the "shill" argument, implying that the other person isn't posting in good faith.
- cess11 2y agoSorry, I don't follow. How do you arrive at that implication? Why would someone having a pecuniary interest in something necessarily make them insincere?
- dang 2y agoPerhaps one internet cliché can explain another: https://hn.algolia.com/?dateRange=all&page=0&prefix=true&query=difficult%20salary%20understand%20depends&sort=byDate&type=comment https://hn.algolia.com/?dateRange=all&page=0&prefix=true&que...
- llama_drama 2y agoIf this is exactly the same model then what's the point of encrypting it?
- TechDebtDevin 2y ago[flagged]
- timewizard 2y ago[flagged]
- TechDebtDevin 2y ago[flagged]
- rob_c 2y agoI am groot
- dang 2y agoPlease don't respond to a bad comment by breaking the site guidelines yourself. That only makes things worse. https://news.ycombinator.com/newsguidelines.html https://news.ycombinator.com/newsguidelines.html
- dang 2y agoPlease don't cross into personal attack or otherwise break the site guidelines when posting here. Your post would be fine with just the first sentence. https://news.ycombinator.com/newsguidelines.html https://news.ycombinator.com/newsguidelines.html
- DoctorOetker 2y agoDon't you think its intentional, so as not to demonstrate the technique on potentially copyrighted data?
- dang 2y ago> But lets not let that get in the way of hating on AI shall we? Can you please edit this kind of thing out of your HN comments? (This is in the site guidelines: https://news.ycombinator.com/newsguidelines.html https://news.ycombinator.com/newsguidelines.html.) It leads to a downward spiral, as one can see in the progression to https://news.ycombinator.com/item?id=42604422 https://news.ycombinator.com/item?id=42604422 and https://news.ycombinator.com/item?id=42604728 https://news.ycombinator.com/item?id=42604728. That's what we're trying to avoid here. Your post is informative and would be just fine without the last sentence (well, plus the snarky first two words).
- Lerc 2y agoCan you clarify this a bit. I presume you are talking about the tone more than the implied statement. If the last sentence were explicit rather than implied, for instance This article seems to be serving the growing prejudice against AI Is that better? It is still likely to be controversial and the accuracy debatable, but it is at least sincere and could be the start of a reasonable conversation, provided the responders behave accordingly. I would like people to talk about controversial things here if they do so in a considerate manner. I'd also like to personally acknowledge how much work you do to defuse situations on HN. You represent an excellent example of how to behave. Even when the people you are talking to assume bad faith you hold your composure.
- dang 2y agoSure, that would be better. It isn't snarky, and it makes fewer uncharitable assumptions.
- jonpo 2y agoI don't seem to be able to edit it, apologies I will try not to let this type of thing get to me in future. I would also like to point out that this is a fine tuned classifier vision model based on mobilenetv2 and not an LLM.
- rob_c 2y ago... Because if he did this with a model that's not open that's sure going to keep everyone happy and not result in lawsuit(s)... The same method/strategy applies to closed tools and models too, although you should probably be careful if you've handed over a credit card for a decryption key to a service and try this ;)
- biosboiii 2y agoAuthor here, it would be nice to claim that I did this on purpose but I really did not know it was open source. I was rather interested in the process of instrumenting of TF to make this "attack" scalable to other apps.