13 ms·
Deep learning on electronic medical records is doomed to fail
- fnbr 5y agoI agree with the conclusion. This is totally unsurprising to me as a ML engineer. If you put garbage data into the model, you get garbage predictions. That doesn’t strike me as particularly novel. The same is true for cooking, after all. However- this has been truly shocking to all of the non-technical stakeholders I’ve worked with. They take the stance that any large amount of data can be used to do ML on, presumably because they don’t know too much about what doing ML is like. So I’m convinced the author is right, and I’m also convinced that there will be many attempts to use ML on EMRs.
- Forgeties79 5y agoGarbage in -> Garbage out is basically a Newtonian law at this point haha
- momenti 5y agoThat's not entirely true. Neural networks are fairly robust to noisy training data (a.k.a. garbage).[0] Well, stochastic gradient descent has the noise in its name. More training data can compensate for noisy data to some extent.[1] I'm not sure know if model size can also compensate for noisy data though, but would not be surprised if it did. [0] https://arxiv.org/abs/1705.10694 https://arxiv.org/abs/1705.10694 [1] https://arxiv.org/abs/2202.01994 https://arxiv.org/abs/2202.01994
- midjji 5y agoThere are very specific conditions for this to hold, mostly that the incorrect sample is surrounded by correct ones, and that the model is small enough or the error vanishingly rare. Notably the reference you gave also shows horrendous generalization performance, so its really just showing how easy it is to overparametrize. Input errors can be accounted for to some extent, but also under specific circumstances, eg. https://openaccess.thecvf.com/content_CVPR_2020/html/Eldesokey_Uncertainty-Aware_CNNs_for_Depth_Completion_Uncertainty_from_Beginning_to_End_CVPR_2020_paper.html https://openaccess.thecvf.com/content_CVPR_2020/html/Eldesok...
- Forgeties79 5y agoI mean the argument there is basically if there is enough good quality data, the bad data is somewhat (or mostly) compensated for. To which I would argue that is no longer a “garbage in garbage out” situation as most people use it.
- jerf 5y agoThere is the old quote that we've all seen: "On two occasions I have been asked, 'Pray, Mr. Babbage, if you put into the machine wrong figures, will the right answers come out?' I am not able rightly to apprehend the kind of confusion of ideas that could provoke such a question." - Charles Babbage I will say, I have some good news for the late great Charles Babbage. For the most part, people now do indeed understand that if you put small amounts of wrong figures into a machine, the wrong answers will come out. If nothing else, pocket calculators and math class have given them the direct experience of this. However, it seems that people still expect that if you put gigabytes or petabytes worth of wrong figures into a machine that somehow the right answer will pop out. Ah well. The road never ends, you know.
- d1sxeyes 5y ago> However, it seems that people still expect that if you put gigabytes or petabytes worth of wrong figures into a machine that somehow the right answer will pop out. The interesting fact is that if you put in lots of correct figures, and only one, slightly wrong figure, then the answer may be correct to an acceptable degree. For example, take 100 values, all exactly one, and find the mean. You'll find 1. If you take 99 values of exactly one and one value of two, the mean of your sample will be 1.01, which is close enough to still be useful. In some interpretations, it may even be rounded to 1, meaning that in some circumstances, incorrect figures can indeed sometimes lead to correct answers. Or if you're trying to find out what adding 1 and 4 gives you, but accidentally you add 2 and 3, you get the correct answer despite incorrect inputs. I think people are assuming that if you put gigabytes or petabytes worth of data into a machine, the number of 'wrong figures' will be lost as noise.
- paulmd 5y agothe problem is that for medical coding, this translates to "a small number of procedures will be coded wrong", and that's not a meaningfully better situation than "a small number of procedures can't be coded", and in fact is probably worse. So you need a reasonably high confidence threshold, and really in most cases you probably want to have a human manually review the problem (or questionable) cases.
- AutumnCurtain 5y agoThe first point he raises is the most critical by far. The silverbacks of the industry deliberately stymie efforts for true interoperability because it goes directly against their primary goal, which is forcing everyone into their platform. Epic in particular has zero intention of allowing anyone else to take their market share by enabling easy sharing of data across platforms. It's far better for them from a business perspective to make interfacing so unreasonably difficult that you are forced to implement their full suite of applications, at which point they hold your organization's data hostage to induce other orgs to do the same. The larger their ecosystem grows, the less they need to worry about interoperability - improving patients' outcomes is not even an afterthought. Their vision of population health reporting is one in which every major healthcare org has been trapped inside their walled garden.
- colinmhayes 5y agoEpic controls 55% of the EMR market, and that number is only growing. It won't be long until this isn't a problem because the majority of the population has all their records with Epic.
- z3ugma 5y agoSome free advice from an ex-Epic: This is true when it's other vendors doing the data fetching. When it's a health system customer of Epic, they bend over backwards to help them extract the data properly and build cool clinical tools on top of the Epic platform. Health systems with big innovation arms like Atrium and Providence could be a good place to seek VC if your product idea relies on deep EMR access. Sometimes the left hand doesn't talk to the right in these health systems though - you'll need to get that innovation arm talking to the EMR analysts. Use the shibboleth "I want to talk to our Epic TS" for whatever speciality you work on. As for things like the App Orchard and Epic on FHIR https://fhir.epic.com/ https://fhir.epic.com/ : Epic is smart enough to realize that their future lies as the platform of the health system IT stack, in the Ben Thompson sense of a platform / aggregator. The hospitals are scared of open access, and Epic always does what's in the best interest of their customers, so they push against open access.
- 5y ago
- ZeroCool2u 5y agoHaving worked with data from EMR systems and having worked at a large EMR software development shop myself, and now using deep learning at work quite a bit for the past few years, I'm inclined to agree. This title is somewhat click bait though, because the fault is really with EMR systems and (esp) the American Healthcare system, not deep learning. The entire system is designed around billing and decisions are made my hospital and insurance executives that are generally not technical. There is no incentive to clean up the system or work on a well structured open protocol for interop the same way there is in say banking. Plus, the author gives some good examples like pulse ox%, doctors and nurses are not at all concerned with or trained to record data in a way that makes sense to use programmatically. They're typically thinking only as if they're recording it for another human to read. Deep learning could probably be quite useful in the medical field, but we won't know until someone comes along and disrupts the system top to bottom similar to how Tesla has done with not only manufacturing, but the sales process and shirking the dealership model. This would probably look something like Forward[1], but with a crazy amount of funding, so that insurance companies and billing codes could be ignored entirely. [1]: https://goforward.com/p/home https://goforward.com/p/home
- ulkesh 5y agoYes, exactly. Hospital data is atrocious, on every single level. Duplicate data are everywhere across multiple systems. Hospitals move excruciatingly slow to do anything technical. And, very few people seem to ever have any real understanding of what is going on with their systems, they tend to only know the user interface and have to rely on their vendor support (Epic, Cerner, etc) for anything beyond that. I work for a company that I was hoping would be such a disruption point for hospitals (at least in some small way), but instead they decided that it's just too difficult to get hospitals to do much of anything, so we effectively knuckle under -- creating numerous integration points, making more and more copies of data. The only way this will change is if a large enough player creates direct competition in the EHR/EMR business, with these kinds of data-oriented models in mind, all the while creating a system that, top-to-bottom, is better for both the user, the administrator, and the technical staff. Current players in this space have very little incentive to make their products better. And it reminds me of a quote from Tron Legacy: "Given the prices we charge to students and schools, what sort of improvements have been made in Flynn... I mean, um, ENCOM OS-12?"..."This year we put a "12" on the box."
- mercurywells 5y agoIs part of the solution going to be having ML figure out what data might be missing to make a better conclusion, then explicitly asking the patient (or the person gathering data from the patient) for that missing data or a clarification?
- paulsutter 5y agoBetter title, "ML on EMRs is very difficult". The article is much more reasonable than the clickbait title: "I don’t think deep learning models on EMRs are going to be useful any time soon .. Clinical expertise is absolutely necessary to ask the right questions, to set up the inputs to the model, and to sift through the findings. Following that, clinical research will be necessary to validate the discovery"
- BurningFrog 5y agoThis is very specific to the US medical system. Perhaps things are better in some of the many other health care systems around the world.
- Forgeties79 5y agoYou know, I have never really given EMR’s in other countries much thought and I’m now super curious how implementation has gone, what some good examples are, etc.
- AutumnCurtain 5y agoEpic expanded into the Nordics and their providers found it disastrous on the whole, mostly because it's designed for the US healthcare system and therefore cares almost exclusively about billing, which is completely irrelevant for them. See https://www.politico.com/story/2019/06/06/epic-denmark-health-1510223 https://www.politico.com/story/2019/06/06/epic-denmark-healt...
- sidlls 5y agoMaybe. Almost everyone follows the same standard (some version of ICD, https://www.cdc.gov/nchs/icd/icd10.htm https://www.cdc.gov/nchs/icd/icd10.htm) of coding. If it's much better outside the US I'd be surprised, though. The standard is a mess of strange, super-specific things (e.g. ICD10 code W56.52XA is "struck by other fish, initial encounter") that aren't useful for any purpose, and in practice is used mainly for billing (insurance or otherwise).
- AutumnCurtain 5y agoI'm fond of W59.22XA, "struck by turtle - initial encounter" aka the Aeschylus ICD code
- 6yyyyyy 5y agoY385X3A Terrorism involving nuclear weapons, terrorist injured, initial encounter
- RcouF1uZ4gsC 5y ago> The answer turns out to be rather mundane. Pap smears are not recommended for women older than 65, and heart failure onset is typically around age 65. The pap smear just turns out to be a really good age bucketing signal to the model. Pap smears are also a really good sex bucketing signal, and there are lots of diseases that are more prevalent in one sex, so I would expect Pap smears to be correlated and anti-correlated with a lot of other diseases as well.
- axg11 5y agoThe author is spot on - the current approach of applying ML to flawed and inconsistent data is doomed. However, on the bright side, I think they also highlight a possible path for data scientists to bring value to healthcare. All of the examples of relationships and correlations were spurious, but if you keep digging you will eventually unearth interesting relationships. A minority of these relationships will be useful for improving hospital administration. Examining the relationships might not necessarily improve health outcomes but there is a real chance to make hospitals more efficient using data science.
- wise0wl 5y agoI previously worked for a startup that did denormalization of healthcare data for the ostensible purpose of data "freedom" and interoperability, with a future focus on ML funsies. All the issues we had (besides ones we created) were around the healthcare providers fear of pissing off EPIC, Cerner, Meditech, Allscripts etc. They didn't like their EMRs---in fact they often hated them. The fact is though that there really is no viable alternative, and the data is kept behind a gate and essentially "owned" by the EMR. FHIR was supposed to solve the interoperability bit, but all the EMRs would still own the data, and their lawyers aren't keen to share.
- boas 5y ago> Life would be simpler if only these hospitals could set aside their arrogance and just go with the recommended workflow! This would be like asking programmers to standardize on the recommended programming language. we would love to just use the recommended workflow, if it worked for our hospital. There are differences in the patients, doctors, local regulations, existing systems, etc between hospitals. Patients: Top cancer hospital does a lot of clinical trials, so some of the forms require you to fill out clinical trial information for every patient. In a maternity ward, it would not be appropriate to ask about clinical trials for every patient. Doctors: Hospitals are staffed differently. If the hospital has residents, some of the work can be delegated to residents. If not, someone else has to do it. The workflow needs to account for who is actually available to do the work. Local regulations: Medicine is highly regulated, and each state and hospital has its own rules. Existing systems: Hospital computer systems have been around for decades, and usually it's not possible to migrate everything to a new system, so the new system needs to integrate with the old systems that couldn't be upgraded.
- tclancy 5y agoI think the rest of the paragraph clarifies they are joking.
- wutbrodo 5y agoLeaving aside the well-known dumpster fire of the healthcare system's operational incompetence, none of these are particularly novel, most of them are the author discovering basic best practices in statistics, and they don't come close to implying "DL for EMR is doomed". A health outcome is correlated with age? Ya don't say.... This is only a problem that literally every single economist and sociologist checks for _first_. Agents in causal graphs respond to their inputs? Shocker, that's only.... The definition of an agent. There are a ton of applications out there where a fairly naive model built by someone who took a Pytorch Coursera course can provide decent improvements. Healthcare is not one of them, and nobody serious ever thought it was. Bringing modern learning tools into healthcare is going to require a lot of smart people who know what they're doing, working cautiously to introduce these improvements into an operational quagmire with high stakes. But this article reads like: "I tried making healthcare 'smart' in a Jupyter notebook over a weekend and it didn't work: the effort is doomed".
- btown 5y agoIt strikes me that the decades of academic literature that establish "in order to get a result correlating X and Y, we needed to control for A, B, C, ..." would be a critical input into any system attempting to work with medical data. In a way, the plaintext of historical medical journals has encoded much of this expert knowledge, albeit with retractions and errors the further back you go. But that might change the conclusion of the OP into more of an "incredibly hard problem" rather than "doomed to fail."
- jcims 5y agoI'm running through a similar situation in risk management. There is so much domain and institutional knowledge that's encoded in rules that its nearly impossible to reason about them in a generic way. Add in operational that is also rife with data quality and coverage issues and it becomes quite difficult. I think we have a term for this in both areas: garbage in, garbage out.
- adultSwim 5y agoThese systems are primarily designed to support clinical care. Research is an after thought. Health care systems will have to decide it's a top level priority. There has been modest progress through wider efforts: - Standard vocabularies, eg LOINC codes for different kinds of lab tests - Mappings between vocabularies, eg OMOP - Semantically rich vocabularies, eg OBO's OBI
- midjji 5y agoDeep learning has great potential in medicine, in particular in radiology and tissue classification. Creating the datasets from scratch will take decades of careful deliberate highly costly effort however, and the current crap hospitals call records is utterly useless. It will truly have to be a bottom up approach, and in the process systematic studies to actually verify a lot of bullshit medical ideas will also have to be done. Basic questions like how many kinds of tissues are there have very dubious answers which are known to be coarse approximations, and some diseases are specifically deviation from the approximations. Its probably not quite as bad as linguistics where, but its really bad. Once datasets with millions of people followed and tested regularly throughout their lives for the specific purpose of generating the dataset, are available for training, it will be quite good. Shame we wont live to see it.
- lumost 5y agoThe challenge with ML and DL systems is that it's difficult to know a-priori what will and won't work. The math would indicate that there is nothing a suitable DL system cannot learn, however in practice certain neural architectures can only learn certain inputs. The cycle time to develop a new system is long, and data is unfortunately scarce. Developing a DL system to solve a problem involves guesswork as to the impact of innovating on any one of dozens of components. Which is to say, it is and will be difficult to create a business based on applying a novel DL method to a particular problem space. We're seeing a consistent trend that focusing on the other aspects of the problem such as data, tooling, or end to end services tends to be much more successful.
- ilaksh 5y agoWe need either A) competent, highly technologically sophisticated government (which seems very far away obviously) or B) something really similar to take it's place. So much of what government does (aside from the bombings etc.) is really about providing and enforcing a framework for people to work together. And in this era that needs to be a high tech framework. Actually, it needs to not only be very high tech, but also very cutting edge, decentralized, sufficiently holistic but also flexible enough to evolve. Which is incredibly hard, and we probably will not get due to greed, stupidity and politics, and that may be the actual reason that human civilization is superceded by AI civilization.
- katekoch 5y ago[flagged]
- stared 5y agoMost of the points are naive and confuse predictive power with interpretability (the latter is harder). For example: > Did you know that a blood oxygen saturation of 0% is highly correlated with healthy outcomes? No, I didn’t get the percentage backwards. A 0% reading is what you get when the nurse looks at you and decides you’re too obviously healthy to bother with putting the pulse oximeter on your finger. The empty field value gets saved as a 0, of course. Well, statistical methods (no matter if classical, Bayesian, Deep Learning, or anything, as long as it goes beyond linear methods) will perfectly capture the special case of 0% and predict accordingly. These methods are free of our biases and will take consistent approaches (e.g. empty columns, columns that de facto mean something else, common typos, etc). Sure, interpretability is problematic, and we often need to consider the knowledge of physicians and medical institutions rather than be based on raw data.
- methehack 5y agohttps://en.wikipedia.org/wiki/Carte_Vitale https://en.wikipedia.org/wiki/Carte_Vitale France's system is universal and private (private docs, insurance, and reference pricing -- a service has a single fee across coherently regulated payers). They have had a standardized medical record for 20+ years. One system and one medical record for everyone. WHO ranked #1 healthcare in the world (to US ~40th) at 1/2 the cost per capita of US healthcare. This is catastrophic legislative failure for a problem largely solved by lots of other people around the world. The legislative failure has created vast administrative overhead (10, or more, staff per doc at a hospital) and corrupt insurance companies. When an insurance company has to pay a claim, they call it a "medical loss" (they had the money and they lost it). They make their money on poor service and deceit -- charging wildly different prices for the same product where they can get away with it). In France, an insurance company, by law, has to pay a claim to a practice in a few days. Imagine the decreased capital needs for running a medical practice or a hospital. The hospitals are not blameless in all this, but the heart of it is the payer system/s.
- methehack 5y agoReplying to my own post, but to the OP's point -- we can't research very well on US healthcare data in aggregate because it is so error prone and inexact to aggregate it. Data scientists should imagine what we could do with 20 years of la carte vitale data...or better of data in a regulated medical record format that was designed to be researched in aggregate. That's the second order failure here.
- nradov 5y agoThat is simply false and misinformation. A "loss" is just a technical term in the insurance industry generally. It doesn't mean that a medical insurer had money and then lost it. The Affordable Care Act (Obamacare) imposed a minimum 80% medical loss ratio. On most policies the "insurance" companies aren't even providing insurance any more. They simply act as third-party administrators for self insured employers. So the insurance companies have no financial incentive to deny claims. In most cases where claim payment is delayed or denied it's because the provider organization failed to follow the rules for claim coding and attachments.
- 5y ago
- wistlo 5y agoA rule encompassing "pap smear" and "over 65" data fields is encoded right here in the OP comment. Aren't these kinds of rules and relationships what "deep learning" is supposed to suss out, automatically and without intervention? If not, I wouldn't call it "deep".
- deleted 5y ago[deleted]
- elij 5y agoI agree with all points with respect to static EMRs including NLP efforts. The cadence, uniformity and alignment of event stores to underlying pathways does pose an opportunity (but there's so little research and I suspect Brian didn't have access to this space). An EMR is a projection/point in time snapsot of these events (basically API calls). There's also inherent natural labels because inferences are evaluated against an actual pathway. Specifically speaking about intra-episodic inferences -- longitudinal predictions (over several episodes -- or a patient life time) becomes wildly inaccurate. There's a lot of research to demonstrate this no matter which models are used.
- apwheele 5y agoI've always taken part of my job as a data scientist is to articulate when data is not sufficient to meet a particular a goal (and to outline what data would be necessary). For some goals doing special data sampling and building a model on small (complete) data is better than using the fragmented overall database. I work mostly on claims data now, and building models to identify problematic claims I assure folks is feasible.
- idoh 5y agoI happen to know a lot of doctors, including, as an example, an OBGYN. As it was explained to me, for vaginal births: - at some point someone, without evidence, speculated that cervix dilation should proceed along some curve - cervix dilation is actually measured by hand - literally inserting fingers and having the doctor practice "so many fingers = so many centimeters". There's plastic sheets with holes in them so they can practice measuring the size of holes with fingers. - the OBGYN knows that the cervix dilation curve should look like, and kinda sorta maps their hand readings to what it should look like - the OBGYN has a general sense as to how labor is going, and will game the cervix dilation stats to match their expectation, e.g. if labor is going well but the cervix hasn't dilated then they'll kinda sorta report progress anyway Anyway, given the above it seems like the data around cervix dilation is suspect - the measurements are fitted to what the curve should look like, and then the data matches the curve, and that makes people more confident in the curve, and so on. The point is, can you really apply ML to the EMR of cervix dilation? Does it make sense, could you really draw conclusions from this?
- pc86 5y agoJust to clarify, this OB will report incorrect clinical data to support how they feel labor is going?
- idoh 5y agoIn some situations yes, in others, if they read a 5.5 or a 6, then they will pick the one that fits.
- calvano915 5y agoIn cases of subjective data, you will always see variances in reporting that may be construed as misrepresentation. It's not at all easy. As "objective" as the example might seem, I'd argue the clinician has too much leeway to truly be objective. There's a whole ton of subjective data involved in every patient course of care, much more than objective in many cases.
- YeGoblynQueenne 5y ago>> The point is, can you really apply ML to the EMR of cervix dilation? Yes, it's a perfect fit. If I may. Sorry, to clarify, the way most people do machine learning is what you describe: tweak a model until it fits the dataset. If the ability of the model to fit the dataset translates to anything beyond that, it's anybody's guess. You just put me off being a parent for life, btw.
- Dan_- 5y agoSetting aside the sensationalist headline, the entire premise of the article is flawed. It's a case of not even being wrong. Of course you're going to get spurious results using poor data. The author's attempt at using structured EMR data is the root cause. We have found that structured data, which the author attempted to use, is at best 35% accurate. Sure it's better than claims, but it does not reach the level of quality necessary to inform clinical decision-making. The reason for this is that almost everything clinically relevant is captured in freeform text fields--clinical notes. To build proper models from information in EMRs, you have to start with processing the narrative data, which is a hard problem. Training models to interpret clinical notes requires clinical expertise. Clinicians record facts differently in different locations, and there are many different ways to say the same things, and sometimes they skip underlying facts because some other fact implies the rest. Different specialties record things differently too. You really cannot just throw some data into a notebook and hope it works. Even with clinician input, we still find that high quality results require ensemble models with multiple techniques; plain NLP doesn't work either. Take for example, non-alcoholic steatohepatitis (NASH), the leading cause of liver failure requiring liver transplant. NASH is a complication of non-alcoholic fatty liver disease, in which your liver has unusually large deposits of fat. NAFLD is not coded in structured data. To identify it from unstructured data, you have to extract concepts related to liver cancer, pre-diabetes, alcohol use, liver fibrosis, cirrhosis, jaundice, fatigue, and loss of appetite. To make a long story short, you cannot do these things using structured data or naive NLP approaches. F1 is zero. So maybe his point, "Data encodes clinical expertise" is worthwhile, but the rest of the article...not so much. Source: My company, Verantos https://verantos.com https://verantos.com , specializes in the generation of high-validity evidence from data we abstract from EHRs using machine techniques.
- hyperbovine 5y ago"This problem is at least as hard as solving NLP" strikes me as supporting the author's claim, not refuting it.
- ialyos 5y agoThe article is simply wrong. I know this because I worked as an ML engineer at an extremely successful company that automated medical coding using deep learning. The confusion stems from conflating a "perfect solution" with a "human augmented" one. 90% of coding cases are trivial, have low value and can be done by a model. 10% are really subtle and need human expertise. That's fine. You can make a billion dollar company on low hanging fruit. I think it's best not to conflate the perfect solution with a very good solution.
- sjg007 5y agoMedical coding is just billing right? You match doctor notes to ICD-10 codes. Seems reasonable.
- nradov 5y agoMedical coding is mainly billing with ICD-10 codes for diagnoses and CPT + HCPCS codes for procedures. However, there is also non-billing clinical coding for things like LOINC, SNOMED CT, and RxNorm.
- not2b 5y agoWhat was was the criterion that you were optimizing? Currently hospitals try to assign codes in a way that maximizes payouts from insurance companies while avoiding straight up lying in a way that could cause them problems. So they'll handle that 10% by choosing the codes with the bigger payout.
- javadocmd 5y agoYou've not refuted the article so much as pointed out a corner case the author didn't address in which ML is a good fit. Your example, using ML to perform the medical coding function, is using a data source (in this case the EMR) for one of the purposes for which it was explicitly designed and for which it is (arguably) non-deficient. That is a realm not doomed to failure. The realm doomed to failure is using a data source for a completely oblique purpose for which it is horribly distorted. Namely, the purpose of optimizing individual and public health by discovering guidelines and treatments, diagnosing illness, and delivering optimal care. (Of course medical billing as an enterprise shouldn't even exist, but that is another topic.)
- teleforce 5y agoIt's kind of funny that the author mentioned about data fragmentation in the very first section and at the end went to provide Heart rate/ECG/Afib as one of the successful case studies for deep learning for EMR. ECG is the classic case of data fragmentation that he's talking about and if you want to export raw data from the popular ECG machine like Mortara, good luck with that. Heck even Apple currently does not support exporting raw data ECG out of the box that's crucial for deep learning, apart from the PDF image file for the ECG waveform [1]. Literally there are more than a few dozens open formats for the ECG [2],[3] and this obligatory XKCD comic comes to mind except that there are 39 standards instead of 15! [4]. Even if the ECG machine manufacturer is using one of the formats there are still serious interoperability issues down the road. Rambling aside, there's yearly cardiology global challenge competition organized by Computing and Cardiology Conference (CINC) and recently there are machine learning and deep learning techniques proposed for multi-lead ECG diagnostics but the results are not that great [5]. Hopefully it will start an impetus to deep learning in EMR similar to ImageNet Challenge that gave rise to the actual deep learning algorithm that drived the community past the winter AI. [1] Accessing the ECG Data of the Apple Watch and Accomplishing Interoperability Through FHIR: https://pubmed.ncbi.nlm.nih.gov/34042901/ https://pubmed.ncbi.nlm.nih.gov/34042901/ [2] A review of ECG storage formats: https://pubmed.ncbi.nlm.nih.gov/21775198/ https://pubmed.ncbi.nlm.nih.gov/21775198/ [3] A Review on Digital ECG Formats and the Relationships Between Them: http://diec.unizar.es/~imr/personal/docs/paper12IEEETITB1.pdf http://diec.unizar.es/~imr/personal/docs/paper12IEEETITB1.pd... [4] How standards proliferate: https://xkcd.com/927/ https://xkcd.com/927/ [5]The PhysioNet/CinC Challenge: https://cinc.org/physionet-cinc-challenge-awards/ https://cinc.org/physionet-cinc-challenge-awards/
- verisimi 5y agoWhat if I do need to go to a hospital, but don't want any information about me to be shared? Is that impossible? What are the ethical considerations for me, when I disagree entirely to any data collection? Must I be coerced against my will into providing to Google or whoever regardless? The point I am raising is that there is an embedded assumption that this info should even be available for deep learning to work over.
- sjg007 5y agoPay cash maybe and even then it might be in the terms of service of the clinic or hospital that they can use or share your data as needed. It may be deidentified as well. An alternative is to check in under an assumed name which many famous people do. There are concierge doctors as well that probably have stricter guidelines.
- paulmd 5y agoFor transmissible/infectious diseases, this is certainly impossible, since states ingest case reports and lab data surrounding this and send digests to CDC. The term of interest would be "NEDSS reporting". https://www.cdc.gov/nndss/about/nedss.html https://www.cdc.gov/nndss/about/nedss.html
- cptaj 5y ago>EMR software is widely hated by the nurses and doctors who have to use it. It’s slow, bloated, nonintuitive, requires workarounds, etc. etc. etc.. The root of this evil is that every hospital brings its own conceited and byzantine patchwork of procedures, checks, and rituals to the table. You just described every admin software for every industry I've worked on. The problem is individual orgs dictating software architecture. When each purchase is in the millions, you accommodate every whim no matter how absurd... and then you end up with these bloated, messy systems. For software systems to REALLY, shockingly improve efficiency in an organization, all the processes in the org need to change to accommodate a new overarching system design. Tailoring software to mirror legacy processes defeats the purpose almost entirely. I think there is a truly absurd competitive advantage in doing this right but you seldom see enough leverage to completely overhaul every department in order to implement software admin systems.
- csours 5y agoPeople don't hate Jira because of Jira, they hate the experience of using Jira with the rules that their organization has put into Jira. They hate the culture of how their org uses Jira.
- justin_oaks 5y agoThere's a lot of truth to that, but Jira (or more accurately Atlassian) does things that are worthy of hate too. - Some information is displayed in unexpected places - The application can be slow to load - Customizations are confusing (Do i need to edit the screen, the screen scheme, or the issue type screen scheme?) - Confusing options (Field configurations have both a "Configure" action and an "Edit" action.) - Basic features are often left out and require plugins to bridge the gap. And much more!
- csours 5y agoFair. I still think MOST of the hate is due to configuration and culture.
- tomlue 5y agoSpurious correlations are common in medical models not idiosyncratic like in other fields. Which makes it hard to use models in clinics. Doctors are smart about this. Every providier we spoke to about survival models based on biomarker data did not want supervised learning models. The explainability, trust, authorization just aren't there and the risk of misuse is too high. For now, research is a better use case, but less commercial funding. It is not hopeless, but your basic LSTM is probably not going to revolutionize medicine.
- mechanical_bear 5y agoThese sort of headlines are just cheap click bait.
- lepapillon 5y agoI don't really want to comment on whether or not DL is doomed to fail on the EMR, but coming from an EMR background, I can say he lays out very accurate points. I particularly like how he explains #3 concisely, and it's a point I use to criticize the private healthcare system. The continual war between hospitals having to opportunistically charge for their services vs. the insurance industry having to take a default stance of deflection creates the massive, meaty layer of coding and billing waste. Thousands upon thousands of jobs exist just for this purpose, and I think any inefficiency in a single-payer system is more than offset by getting rid of that layer and everyone benefits.
- dekhn 5y agoThere is something I've been developing over the years which I call "Konerding's Empirical Observation #7": every attempt to improve health care with technology in the US will only every increase the cost while also decreasing the quality of service (on average). This goes hand in hand with Konerding's 3rd empirical observation, which is that there is a never ending supply of "machine learning geniuses" who are naive about how health care works in the US, and spend some 20 years learning how hard it is to actually do anything actionable with health care data.
- jmugan 5y agoWe need to work on the hard problem of building causal models of the world, then we can build causal models of medicine on top, then we can do learning on medical records.
- deltarholamda 5y agoWhile a lot of the post is good info, there are some upsides, if we can get the EHR situation worked out. Many years ago, prior to anything like ML, Canada figured out that cystic fibrosis patients whose weight is higher than 50th percentile, had significantly better lung function. Nobody really understood why, but the correlation was so strong (.85 or something like that) it could not be ignored. Treatment protocols for CF changed to encourage weight gain, and lifespan outcomes have steadily improved over the years. What other oddball correlations are hiding in the depths of bloodwork, weight/height, etc. for patients? We've teased out all the easy ones, the ones that are left are combinations nobody thought to even measure. Regarding the EHR debacle, I'm optimistic that something could be worked out as a standard and implemented across the board. Expensive? Sure, but it's an investment that pays off pretty quickly, I think.
- nradov 5y agoWe're never going to get every provider organization using the same EHR, nor would that even be desirable. But almost all of them have passed ONC Health IT certification so they have similar functionality, including exposing at least some data in open industry standard formats.
- ltbarcly3 5y agoIf there is no meaningful statistical information in medical records, then we should stop keeping them? By hypothesis a doctor who opens a medical record can't gain any meaningful information from it for the same reasons listed in the article. I think this is sufficient to demonstrate the article is incorrect, there is significant information in medical records, and therefore it will be possible to train a model which can reproduce missing information to some degree.
- yold__ 5y agoBecause there is, and the person who wrote the original article has no domain experience. Fixing data sucks and it requires judgement. Public health and medical researchers derive an enormous amount of research benefit from anonymized health records. Medicare publishes a large dataset.
- deleted 5y ago[deleted]
- suifbwish 5y agoWhere deep learning will prevail in medical is in predictive medicine. When it finally becomes common to have personal genomic data available during checkups, the machine learning will be able to guess/order tests for individuals based on their age, specific genes they are carrying, known life age that diseases onset ect, diseases for that area. It will also be able to look across the population and detect statistical patterns in geographic incidences of diseases with environmental causes which occur outside of the normal expected distribution or in hotspots. The important thing to remember about AI is you need reliable data to train it as well as reliable data to test it. If you don’t the FDA will not allow you to employ it medically.
- nradov 5y agoThat seems doubtful. Outside of a few limited, specific cases like the BRCA genes, that personalized medicine data has mostly turned out to be a disappointment and not actionable. Like if I find out that due to genetics my lifetime risk of rotator cuff injury is 25% compared to 21% for the general population then what am I supposed to do with that information?
- suifbwish 5y agoIt’s not a disappointment as we have not been able to fully try it out yet. Personalized medicine is in its embryonic infancy. Naysayers can say what they will but the medicine/therapy of the future will likely be completely personalized AI generated RNA retroviral/dendrasome delivered cocktails that are based on an AIs comprehension of your entire genome and epigenome. I’m talking the next 100 years. One major problem although beneficial at times, the FDA approval process inhibits the development and shipment of new solutions to medical disorders.
- citizenpaul 5y agoYou may not realize that EMRs owe their existence to 1.billing 2. government mandates 3. billing 4.helping doctors keep track of their patients’ records just like how people think ADP is in the business of payroll. Theyare actually in the business off mitigating regulation, taxes and liability. Getting your wage to you is at best 2nd priority to all those.
- iancmceachern 5y agoYou see these same pitfalls in much of the medical device industry. Often hospitals and their weird political internal workings drive reasons for things, and not quality or efficiency of care, care for their workers, etc. You see this present itself af far up as to help choose a certain technology development path for an entire industry based on these non-real internal hospital dynamics and intrenched ways of doing things, and billing for them rather than working to improve quality of care or outcomes.
- myrryr 5y agoThis seems like a lot of very "USA" problems. A lot of countries don't have the same drivers.
- nikanj 5y agoBut not before a lot of companies make tons of money by promising deep learning will cut healthcare costs by percentage points!
- steve76 5y ago
- jrapdx3 5y agoI'm an American physician. I've practiced in a couple of specialties and know the insurance billing drill (or should I say game) pretty well. Medical record systems, as other comments point out, have been constructed for the benefit of administrators, recording clinical data is mainly to support administrative needs. In reality manifestations of a given illness vary continuously across a wide spectrum. It means patients with the same diagnosis have differing sets of symptoms and course of illness. As I like to say it, "no two patients have exactly the same disease." Official disease classification schemes embody the "splitter" model which attempts to fit continuous data into discrete categories. Of course sharp-edged distinctions suit purposes like billing and other management operations. However diagnosis is often ambiguous, mixed or multiple, but no matter what categorical diagnoses must be assigned. Unsurprisingly doctors are biased to select the choices with the greatest reimbursement, and "stretching" criteria to cover the patient's condition (or vice versa) is not at all uncommon. (Also, there's not assigning diagnoses where it might negatively impact payment.) EHR clinical data follows the discrete assumptions of the system design. This can create an impedance mismatch between clinician observations and data input. This may be troublesome, for example, in specialties (behavioral health) where data is complex with many overlapping subtle but meaningful variations. I've often received records after a patient has been hospitalized. More than not it's difficult to understand the course of treatment as there's no coherent summary or narrative description provided. The "pile of data" literally transcribed from the EHR isn't very useful to human readers. To be sure information like lab reports, etc., are good to have, but the marginalized human-to-human element is troublesome. EHR clinical data failing to serve ML purposes could only point to problematic EHR system design. Though I've been aware of EHR limitations, I wouldn't have guessed about ML issues. The article taught me something about the problem that exists. Now it remains to be seen what can or will be done about it.
- jesseryoung 5y agoI've worked in healthcare IT for my entire professional career - It's A LOT more complicated than most people think. For the last 5 years I've focused on the data side of healthcare and I think that deep learning is 100% possible - it's just not achievable by a single person and it's likely VERY expensive. There are so many facets to healthcare data that's it's just impossible for a single individual to achieve something meaningful by themselves without the help of teams of doctors, data analysts, data engineers and data scientists. Just dealing with data quality issues (such as the ones called out in this essay) require a team of people to determine if metrics you are trying to measure are legit or not. On billing: I'm convinced that the primary reason why healthcare (at least in the US) is so complex - is because of the dichotomy of saving people at all costs, while doing so fiscally responsibly. It is fairly common for large healthcare organizations to have ACTING doctors in their c-suite, who's primary goal is not to make money - it's to save lives. The people who care about saving money, reducing cost and increasing efficiency have no control over the organization. I'm not saying this is a bad thing, but IMO it's the largest contributing factor as to why healthcare billing is so complex, and healthcare costs get as high as they do (at least in the US).
- naveen99 5y agoIf anyone is interested in working with emr data, my lab is looking for phd students, postdocs. We have quality anonymized emr and dicom data and research protocols to work with along with domain expertise and deep learning infrastructure.
- aliu22 5y ago"Doomed to fail" is too strong IMO. All the problems the author brings up, while legitimate, are being worked on. For example, on the interoperability front, TEFCA is making big strides on government-mandated nationwide interoperability: https://www.healthit.gov/topic/interoperability/trusted-exchange-framework-and-common-agreement-tefca https://www.healthit.gov/topic/interoperability/trusted-exch... > In January 2022, ONC and the RCE announced the publication of the Trusted Exchange Framework and the Common Agreement (TEFCA). Entities will soon be able to apply and be designated as Qualified Health Information Networks. Google also has made significant strides on deep learning on EHRs: https://ai.googleblog.com/2018/05/deep-learning-for-electronic-health.html https://ai.googleblog.com/2018/05/deep-learning-for-electron...