bioRxiv Science⌕ Search

bioRxiv · 10.1101/2025.05.01.651780

Hybrid Population PK-Machine Learning Modeling to Predict Infliximab Pharmacokinetics in Pediatric and Young Adult Patients with Crohn's Disease

Abstract

Population pharmacokinetic (PK) model-based Bayesian estimation is widely used for dose individualization, particularly when sample availability is limited. However, its predictive accuracy can be compromised by factors such as misspecified prior information, intra-patient variability, and uncertainties in PK variations. In this study, we developed a hybrid approach that combines machine learning (ML) with population PK-based Bayesian methods to improve the prediction of infliximab concentrations in children with Crohns disease. We calculated prediction errors between Bayesian-estimated and observed infliximab concentrations from 292 measurements across 93 patients. Incorporating clinical patient features, we explored various ML algorithms, including linear regression, random forest, support vector regression, neural networks, and XGBoost to correct the Bayesian-based prediction errors. The predictive performance of these ML models was assessed using root mean square error (RMSE) and mean prediction error (MPE) with 5-fold cross-validation. For Bayesian estimation alone, the RMSE and MPE were 4.8 {micro}g/mL and -0.67 {micro}g/mL, respectively. Among the ML algorithms, the XGBoost model demonstrated the best performance, achieving an RMSE of 3.78 {+/-} 0.85 {micro}g/mL and an MPE of -0.03 {+/-} 0.69 {micro}g/mL in 5-fold cross-validation. The ML-corrected Bayesian estimation significantly reduced the absolute prediction error compared to Bayesian estimation alone. This hybrid population PK-ML approach provides a promising framework for improving the predictive performance of Bayesian estimation, with the potential for continuous learning from new clinical data to enhance dose individualization. Key pointsO_LIA new hybrid model combining population pharmacokinetic model-based Bayesian estimation and machine learning significantly improved the accuracy of infliximab concentration predictions in young adult and pediatric patients with Crohns disease. C_LIO_LIThe developed hybrid model can facilitate infliximab individualized dosing by accounting for changes in clinical conditions and patient-specific factors that the conventional Bayesian estimation approach may not address, and can be integrated into precision dosing dashboards, such as RoadMAB, for real-world clinical application. C_LIO_LIThis study indicates that model predictive accuracy can be enhanced by combining the Bayesian method with machine learning, even with a relatively small amount of clinical data. This is particularly encouraging for specific populations, such as pediatric patients, where obtaining rich clinical data is challenging. C_LI

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Irie, K., Phillip, M., Reifenberg, J., Brendan, B. M., Noe, J. D., Jeffrey, H., Mizuno, T.. 2025-05-07. Hybrid Population PK-Machine Learning Modeling to Predict Infliximab Pharmacokinetics in Pediatric and Young Adult Patients with Crohn's Disease. https://doi.org/10.1101/2025.05.01.651780

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Aquaporin-9 and aquaporin-10 but not aquaporin-3 confer susceptibility to dimethylarsinic acid genotoxicity in human cells

Human metabolism converts inorganic arsenic to the pentavalent methylated species MMA(V) and DMA(V), the forms most people excrete, and the forms long read as the end of a detoxification pathway. Whether a transporter sets how much of these metabolites reaches the genome has not been tested in a mammalian cell. We expressed human AQP3, AQP7, AQP9 or AQP10 in HEK293T and MRC5-SV40 cells and measured gamma-H2AX by flow cytometry across dose series of As(V), MMA(V) and DMA(V), pairing every aquaporin with a GFP-Tubulin control and an untransfected mock acquired in the same replicate. As(V) was inactive in HEK293T cells and only weakly active in MRC5-SV40 cells to 20 micromolar, and both methylated species damaged DNA only in the millimolar range, DMA(V) being the more potent of the two in both cell lines. Against that weak baseline, AQP9 and AQP10 raised DMA(V)-induced gamma-H2AX in HEK293T cells by roughly 17 percentage points over the matched control, more than doubling the damage the same exposure produced in control cells, whereas AQP3 and AQP7 changed it not at all. AQP9 alone remained active with MMA(V). The ranking held in MRC5-SV40 fibroblasts at one-sixth the size, and within single wells the damage rose with the amount of AQP9 a cell carried while the control was flat. Aquaglyceroporins therefore discriminate among arsenic species, and AQP9 and AQP10 turn a weakly genotoxic metabolite into a substantially more genotoxic one.

pharmacology and toxicology↗

Quantitative Systems Pharmacology Model for Trop-2 Targeting Antibody-Drug Conjugate in Triple-Negative Breast Cancer

TROP2-targeted antibody-drug conjugates (ADCs) have demonstrated promising clinical activity in triple-negative breast cancer (TNBC) as monotherapies; however, therapeutic benefit varies among patients. Combination strategies pairing TROP2-targeted ADCs with immune checkpoint inhibitors are also being investigated. Elucidating the mechanistic drivers of ADC monotherapy variability and enabling the rational development of combination regimens require computational frameworks that integrate ADC pharmacology with tumor-immune interactions. A quantitative systems pharmacology (QSP) model is presented that incorporates an ADC module into our established immuno-oncology model for TNBC. The module captures ADC and payload pharmacokinetics and pharmacodynamics. TNBC heterogeneity is represented by two tumor cell clones with high and low TROP2 expression, informed by prior characterizations, and differential sensitivity to the ADC payload is incorporated as an intrinsic property of each clone. Although generalizable, the model was applied to the TROP2-targeted ADC sacituzumab govitecan (SG, TRODELVY). A virtual patient cohort was generated using Latin hypercube sampling and calibrated against objective response rate (ORR) data from SG Phase I/II TNBC basket trial. The model predicted an ORR of 33.2% consistent with ASCENT study (NCT02574455). Simulations suggest TROP2-mediated delivery contributes modestly to SG efficacy with tumor exposure driven largely by systemically released SN-38 payload being sufficient to induce cytotoxicity. Tumor heterogeneity emerged as a key determinant of response with ORR increasing as the fraction of payload-sensitive clones increased. Overall, this QSP framework for TROP2-targeted ADCs accounts for TNBC heterogeneity and is extendable to other ADCs and targets enabling interrogation of ADC mechanisms of action in conjunction with tumor-immune interactions.

pharmacology and toxicology↗

Computer-Assisted Systematic Chemical-Space Mapping of a First-in-Class Peripherally Restricted α2AAR Agonist through Scaffold-Seeded Enumeration

CC10137 is a first-in-class peripherally restricted 2A-adrenergic receptor (2AAR) agonist with broad-spectrum analgesic efficacy and a favorable safety profile. Systematic exploration of the chemical space surrounding first-in-class leads is important for defining series boundaries and guiding continued optimization, but conventional analogue-by-analogue medicinal chemistry samples only a small fraction of the accessible structural space. Here, we used a scaffold-seeded enumeration strategy to expand the chemical space surrounding CC10137 from four SAR-informed seed compounds comprising CC10137 and three closely related structural variants. Application of predefined medicinal chemistry transformation rules in StarDrop generated a virtual library of 16,601,163 unique structures. Morgan fingerprint-based principal component analysis indicated that the library occupied a highly multidimensional structural space involving variation in scaffold substitution, peripheral functional groups, and side-chain composition. A retrospective comparison set of 43 compounds independently designed and experimentally characterized in the earlier CC10137 program represented only approximately 0.00026% of the 16.6-million-member library, yet all 43 were recovered as exact structural matches. Three compounds selected directly from the virtual library retained 2AAR binding affinity and agonist potency below 25 nM. Five representative compounds further showed significant anti-allodynic effects in the in vivo spared nerve injury model, with inhibition rates ranging from 39.3% to 55.7%. These findings support scaffold-seeded computational enumeration as a practical strategy for systematic chemical-space mapping around a first-in-class lead and for identifying additional pharmacologically active structural regions for further optimization.

pharmacology and toxicology↗