An AI model predicting acute kidney injury works, but not without some tweaking

The model identified AKI 48 hours in advance, allowing ample time for clinicians to intervene and provide treatment.

10:39 AM

Author | Clarissa Piatek

scientist examining a kidney
Getty Images

In 2019, Google AI subsidiary DeepMind used a large dataset of patient records from the Veterans Health Administration to develop a predictive model for acute kidney injury—a potentially fatal condition whose prognosis improves the earlier a treatment intervention is administered. The DeepMind model purported to predict AKI 48 hours in advance, allowing ample lead time for clinicians to intervene and administer treatment.

The study team reviewed electronic health record data over a five-year period from more than 700,000 individuals.

“This is a phenomenal model because it can predict AKI up to 48 hours in advance, in a continuous manner, and has the best model performance compared to all previously published models,” said Jie Cao, a Ph.D., student in the Department of Computational Medicine and Bioinformatics. Cao is a researcher in the ML4LHS Lab, run by Karandeep Singh, M.D., MMSc., assistant professor in the Department of Learning Health Sciences and the Biomedical & Clinical Informatics Lab run by Kayvan Najarian, Ph.D., professor in the Department of Computational Medicine and Bioinformatics. All are of the University of Michigan. However, “concerns were raised about the generalizability of a model like this given the predominantly male [VA] population that it was trained on,” Cao added.

This led Cao and her colleagues to evaluate the model’s generalizability in a non-VA, more sex-balanced population. Their findings have been published in Nature Machine Intelligence. 

The researchers reconstructed aspects of the DeepMind model, then trained and validated this model on two cohorts: one comprising 278,813 VA hospitalizations (from 118 VA hospitals) and the other 165,359 U-M hospitalizations. Not surprisingly, given the 94% male population with which the original model was developed, the reconstructed model performed worse for female patients in both cohorts.

To mitigate the model’s sex-based discrepancies, researchers updated the model with data from U-M’s more sex-balanced patient population, which extended the original model from 160 decision trees to 170. This small extension improved performance in the U-M cohort both overall and between sexes.

Like Podcasts? Add the Michigan Medicine News Break on Spotify, Apple Podcasts or anywhere you listen to podcasts.

“The extended model was successful at U-M. It used the VA model as the backbone, added information from U-M, and the final product worked well for the U-M patient population,” said Cao, lead author of the paper. “When researchers would like to benefit from the rich information contained in the original model and do not want to build a new local model from scratch, our study is a good example of how ‘fine-tuning’ could work when the original model was not trained on a diverse population,” she explained.

When the extended model was applied to VA patients, however, the discrepancies in model performance between males and females actually worsened.

“This finding surprised us to some extent,” said Cao, “but it is also reasonable and helps us understand the problem better. Difference in patient characteristics is one common factor contributing to model performance discrepancy. By matching the female patients at two different health systems and still finding discrepant model performance, we actually show that difference in patient characteristics wasn’t the only reason contributing to model performance discrepancy.”

Lower performance of the extended model in female VA patients, then, was not a factor of patient characteristics or low sample size, but likely attributable to variables such as differences in practice patterns between males and females at the VA.

Overall, the study demonstrated the value of updating existing models with data from the population to which the model will be applied.  

“If a predictive model is to be taken out of one healthcare system and applied to another, the population the model was trained on is often different from the population it is going to be applied to. Even if the training population is diverse, we could observe a drop in model performance if nothing is done,” said Cao. “Our ‘extended model’ approach is to provide a solution to partially address this issue.”

To achieve peak performance, a model would, in theory, be applied only to a population matching the population it was trained on. But this is often not the case in practice.

“In the real world,” said Cao, this approach is “infeasible due to limited resources, time, expertise, etc. Our ‘extended model’ strategy is a workaround in these scenarios.”

This research is significant, said Cao, because it shows “the complexity of discrepancies in model performance in subgroups that cannot be explained simply on the basis of sample size.” It also offers “a potential strategy to mitigate the generalizability issue,” she said, and, finally, it demonstrates “the importance of reproducing and evaluating artificial intelligence studies.”

Live your healthiest life: Get tips from top experts weekly. Subscribe to the Michigan Health blog newsletter

Headlines from the frontlines: The power of scientific discovery harnessed and delivered to your inbox every week. Subscribe to the Michigan Health Lab blog newsletter


More Articles About:

All Research Topics Kidney Disease Kidney Failure Preventative health and wellness Future Think Hospitals & Centers Health Care Quality Health Tech
Health Lab word mark overlaying blue cells

Health Lab

Explore thousands of health news & research stories by visiting the Health Lab homepage for more.

Media Contact

University Hospital at U-M Health in the spring with flowering trees in foreground and Survival Flight helicopter visible

Public Relations

Department of Communication at Michigan Medicine

[email protected]

734-764-2220

Stay Informed

Want top health & research news weekly? Sign up for Health Lab’s newsletters today!

Subscribe

Featured News & Stories

pills close up some powder some round some long
Health Lab

What’s the latest on 7-OH and kratom availability and addiction care?

The United States Drug Enforcement Agency proposed temporary rescheduling of 7-OH products could remove synthetic opioids from the market, leaving natural leaf kratom available amid a shortage of addiction treatment providers.
pink person talking to one orange person with question mark above their head confused and then a yellow person talking to the orange person and a teal person talking to the orange person
Health Lab

Many young adults may not be ready to manage their own health care

Many young adults may be entering adulthood without the practical skills needed to manage their own health care, suggests a national C.S. Mott Children's Hospital poll.
e-bike handles on road close up
Health Lab

E-bikes are faster than ever. Are families prepared for the risks?

As e-bikes and e-scooters grow in popularity, Michigan Medicine experts are warning families that many riders and parents may not fully understand the speeds these devices can reach or the injuries that can follow when things go wrong.
cells floating green black background strings many floating in backdrop
Health Lab

Scientists discover that bacteria use an understudied polymer found in all life to protect cell functions during stress

New study discovers that polyphosphate, a molecule that exists in all lifeforms but has to date been hard to study, could be a fundamental element worth exploring, say Michigan Medicine researchers.
lab with microscope vials
Health Lab

The hidden life of parasites like Cyclospora

Michigan Medicine biologists and doctors explain how parasites infect people—and how to respond or reduce your risk
Health Lab

The quiet warning signs your kidneys may be signaling

Top nephrologist Julie Wright Nunes discusses red flags your kidneys might be signaling that you shouldn’t ignore.