Urine Proteomics Identifies Biomarkers for Diabetic Kidney Disease at Different Stages | Research Square window.SnipcartSettings = { analytics: { enabled: false } }; (function() { var accessVector = localStorage.getItem('access_vector') || ''; window.dataLayer = window.dataLayer || []; if (accessVector) { window.dataLayer.push({ user: { profile: { profileInfo: { snid: accessVector } } } }); } })(); (function(w,d,s,l,i){w[l]=w[l]||[];w[l].push({'gtm.start':new Date().getTime(),event:'gtm.js'});var f=d.getElementsByTagName(s)[0],j=d.createElement(s),dl=l!='dataLayer'?'&l='+l:'';j.async=true;j.src='https://www.googletagmanager.com/gtm.js?id='+i+dl;f.parentNode.insertBefore(j,f);})(window,document,'script','dataLayer','GTM-K279D39R'); Browse Preprints In Review Journals COVID-19 Preprints AJE Video Bytes Research Tools Research Promotion AJE Professional Editing AJE Rubriq About Preprint Platform In Review Editorial Policies Our Team Advisory Board Help Center Sign In Submit a Preprint Cite Share Download PDF Research Urine Proteomics Identifies Biomarkers for Diabetic Kidney Disease at Different Stages Guanjie Fan, Tongqing Gong, Yuping Lin, Jianping Wang, Lu Sun, and 31 more This is a preprint; it has not been peer reviewed by a journal. https://doi.org/ 10.21203/rs.3.rs-966683/v1 This work is licensed under a CC BY 4.0 License Status: Published Journal Publication published 30 Nov, 2021 Read the published version in Clinical Proteomics → Version 1 posted 11 You are reading this latest preprint version Abstract Background Type 2 diabetic kidney disease is the most common cause of chronic kidney diseases (CKD) and end-stage renal diseases (ESRD). Although kidney biopsy is considered as the ‘gold standard’ for diabetic kidney disease (DKD) diagnosis, it is an invasive procedure, and the diagnosis can be influenced by sampling bias and personal judgement. It is desirable to establish a non-invasive procedure that can complement kidney biopsy in diagnosis and tracking the DKD progress. Methods In this cross-sectional study, we collected 252 urine samples, including 134 uncomplicated diabetes, 65 DKD, 40 CKD without diabetes and 13 follow-up diabetic sample, and analyzed the urine proteomes with liquid chromatography coupled with tandem mass spectrometry (LC-MS/MS). We built logistic regression models to distinguish uncomplicated diabetes, DKD and other CKDs. Results We quantified 559 ± 202 gene products (GPs) (Mean ± SD) on a single sample and 2,946 GPs in total. Based on logistic regression models, DKD patients could be differentiated from the uncomplicated diabetic patients with 2 urinary proteins (AUC = 0.928), and the stage 3 (DKD3) and stage 4 (DKD4) DKD patients with 3 urinary proteins (AUC = 0.949). These results were validated in an independent data set. Finally, a 4-protein classifier identified putative pre-DKD3 patients, who showed DKD3 proteomic features but were not diagnosed by clinical standards. Follow-up studies on 11 patients indicated that 2 putative pre-DKD patient have progressed to DKD3. Conclusions Our study demonstrated the potential for urinary proteomics as a noninvasive method for DKD diagnosis and identifying high-risk patients for progression monitoring. General Biochemistry Molecular Biology Urine Proteomics DKD Progression monitoring Figures Figure 1 Figure 2 Figure 3 Figure 4 Background In 2019, more than 463 million people worldwide were estimated to be living with diabetes, representing 9.3% of the global adult population (20ཞ79 years)[ 1 ]. Among these people, 20–40% will progress to DKD[ 2 ], which remains a leading cause of morbidity and mortality in people with type 2 diabetes[ 3 – 5 ]. DKD patients are at significant risk of progression to ESRD and cardiovascular diseases. Therefore, early detection, prevention, and treatment are of great importance for disease management[ 6 ]. Clinically, the diagnosis of DKD is based on the measurement of eGFR, albuminuria and urinary microalbumin creatinine ratio (UACR) along with other clinical features[ 7 ]. The “gold standard” for definitive diagnosis of DKD requires kidney biopsy with which an expert renal pathologist makes the diagnosis from histological tissue. However, the procedure is invasive and sometime dangerous, and the evaluation process can be biased by human judgment[ 8 ]. It is important to find a noninvasive test method that can complement or replace renal puncture. DKD is divided into five stages according to clinical guidelines[ 9 , 10 ]. The stage 1 (DKD1) and stage 2 (DKD2) DKDs are preclinical stages, and are characterized by increase in glomerular filtration rate (GFR), normal albuminuria or intermittent microalbuminuria. DKD3, characterized by persistent microalbuminuria, mild hypertension, and a normal or slight decline in GFR, is the onset of clinical stage[ 11 ]. Patients with DKD4 exhibit clinical symptoms of edema and hypertension, along with an increase in albuminuria, which is difficult to treat[ 12 ]. In the overt DKD4, as the glomerular filtration rate (GFR) declines, the albumin-to-creatinine ratio (ACR) would further increase (>300 mg/g). Under this circumstance, taking any medicine could increase the kidney burden thus exacerbate the disease. With proper clinical intervention, the DKD progression can often be delayed or even reversed before it progresses to stage 4. Therefore, monitoring the kidney function for DKD3 patients is critical to the medical treatments and delaying disease prognosis. In the past years, several noninvasive methods have been proposed, mostly based on the ‘omics' techniques for the evaluation of urine or serum biomarkers[ 13 , 14 ]. Among these approaches, urinary proteome analysis has gained popularity, and has the potential to be transitioned towards clinical implementation[ 15 , 16 ]. Measuring urine proteome is non-invasive, and quantitative determining the urinary proteins and peptides can be an enabling platform to distinguish disease stages or monitor response to therapies. Much efforts have been made in finding biomarkers for CKD, aiding risk assessment for DKD patients. Using high-resolution capillary electrophoresis coupled with electrospray-ionization mass spectrometry, a panel of 40 biomarkers were identified from the urinary peptides and they were reported to be able to differentiate healthy individuals from diabetic patients with persistent normal albuminuria, low-grade albuminuria, or nephropathy (PREDICTION)[ 17 ]. Furthermore, it could distinguish patients with DKD from patients with other CKDs. However, the exact protein identities of these peptide biomarkers were not clear. A metabolic study utilized normal phase liquid chromatography coupled with time-of-flight mass spectrometry (NPLC-TOF/MS) investigated the profile of the plasma phospholipids of type 2 diabetes and DKD[ 18 ] and found that 2 novel biomarkers, PI C18:0/22:6 and SM dC18:0/20:2, could be used to distinguish healthy individuals, type 2 diabetes cases and DKD cases from each other, but these two biomarkers cannot distinguish the stages of DKD. A CKD273 classifier was successful in using urine peptides to predict the development of DKD before patients developed microalbuminuria[ 19 ]. A work by Liao, W.-L. et al identified haptoglobin (HPT) and α-1-microglobulin/bikunin precursor (AMBP) as two biomarkers with the highest ability to distinguish between healthy individuals and patients with nephropathy, and between diabetic patients with or without DKD[ 20 ] in a Taiwanese population. However, the study population was limited to patients with an ACR of less than 300 mg/g, and none of the patients had stage 3 of CKD. In addition, the information on the use of insulin or other drugs (such as renin-angiotensin system antagonists) that may alter renal function was not provided. Currently, clinical DKD surveillance relies on the measurements of eGFR and UACR, along with other physical and clinical parameters. The accuracy of these tests is not ideal and often depends on the results of multiple tests[ 21 ]. Here, we report a streamlined urine proteomics workflow to monitor DKD at different stages using liquid chromatography coupled with tandem mass spectrometry (LC-MS/MS). Urine proteomes of diabetes without complications, DKD, and other nephropathy patients without diabetes were measured and used to build logistic regression models to differentiate DKD from uncomplicated diabetes and different stages of DKD. Materials And Methods Sample Collection A total of 239 urine samples from 236 patients, including diabetes without nephropathy (n = 134), diabetic kidney disease (n = 65), and nephropathy without diabetes (n = 40), were collected in Guangdong Provincial Hospital of Chinese Medicine, Guangzhou, China. The batch 1 dataset of 30 urine samples was collected in year 2018, including 21 cases of diabetes without nephropathy and 9 cases of DKD; the batch 2 dataset of 209 urine samples was collected in year 2019, including diabetes without nephropathy (n = 113), DKD (n = 56) and nephropathy without diabetes (n = 40). DKDs were separated into two stages as DKD3 (Urinary Albumin/Creatinine Ratio: 30-300 mg/g) and DKD4 (Urinary Albumin/Creatinine Ratio >300 mg/g). Sample Preparation and Nano HPLC-MS Analysis Midstream of the first-morning urine was obtained and stored at -80℃. One milliliter of all urine samples was centrifuged at 176,000g for 1 h and the pellet was collected. The pellet was resuspended with 160 µl of resuspension buffer (100 mM Tris, pH 7.4) with 50 mM dithiothreitol (DTT). The suspension was then heated at 65°C for 30 min and centrifuged at 176,000 g for 0.5 h. The pellet was resuspended with 30 µl NH 4 HCO 3, heated at 95℃ for 3 min and cooled to room temperature, then digested by trypsin or 12 h. The digested peptides were vacuum dried and re-dissolved in 0.1% formic acid and resolved on an UltiMate 3000 RSLCnano System (Thermo Fisher Scientific) operating on a 20 minutes linear gradient for batch 1 and 30 minutes for batch 2 (5-35% acetonitrile in 0.1% formic acid) at a flow rate of 600 nl/min. Tandem mass spectra were acquired on a QExactive HF mass spectrometer (Thermo Fisher Scientific) in the data-dependent mode. Quality control samples were made prepared from trypsin digests of 293T cells and were routinely analyzed to assess the LC-MS/MS sensitivity and reproducibility. Protein Identification and Label-free Quantification MS data were processed on the Firmiana platform[ 22 ]. Proteins were identified against the NCBI human RefSeq protein database (released on 04/07/2013, 32,015 entries) using the MASCOT search engine (Matrix Science, version 2.3.01). Mass tolerance was set as 20 ppm for precursor ions and 0.05 Da for product-ions, respectively. Up to one missed cleavage was allowed for trypsin digestion. Cysteine carbamidomethylation was considered as a fixed modification, N-terminal acetylation and methionine oxidation were considered as dynamic modifications. 1% FDR on both the peptide and protein levels estimated by searching a decoy database were allowed. Only identifications with ≥1 unique and strict peptides and ≥2 strict peptides (ion score > 20) or ≥3 strict peptides, which was comparable to 1 % FDR at the protein level, were used for subsequent analyses. For protein quantification, intensity based absolute quantification (iBAQ) algorithm[ 23 ] was used. To normalize the differences in sample amounts, iBAQ values were converted to iFOTs (fraction of total) calculated by dividing the iBAQ value of each protein by the total iBAQ of the sample followed by multiplying 10 5 for easy visualization. All missing values were substituted with zero. Statistical Analysis Pathway enrichment analysis was performed using reactome ( https://reactome.org/ )[ 24 ]. Protein differential excretion was assessed with Kruskal-Wallis test. The correlation analyses between clinical and proteomics data and the statistical analyses were calculated by scipy (version 1.3.0) package with Python. For high-dimensional MS data, dimension reduction was applied to avoid over-fitting. We first defined the differentially expressed proteins as the candidate biomarker set, then go through the potentially combinations of less than or equal to 4 proteins. At last, a logistic regression model was carried out to select the optimal panel of biomarker. A Logistic Regression Classifier (sklearn (0.21.2) package with Python) was built using batch1 data as training set and batch2 data as validation set. The performance of the model was evaluated by sensitivity and specificity computed based on the confusion matrix. The Receiver Operating Characteristic curve plotted sensitivity (True Positive Ratio) as the x axis and 1-specificity (False Positive Ratio) as the y axis. Feature selection was applied to select most important features that can make the model achieve a higher AUC (area under ROC curve). Principal component analysis (PCA) was performed by the sklearn (0.21.2) package with Python on validation dataset. Two components were used for data visualization. Risk of progression to DKD3 was assessed by a risk score, representing the probability of being DKD3 by Model 3. Results Urine Proteomics of uncomplicated Diabetes without nephropathy, DKD, and CKD without diabetes We employed a previously published procedure for measuring urine proteomes[ 25 ] with minor modifications. This streamlined workflow achieved high efficiency and batch-to-batch reproducibility. The stability and reproducibility of the MS platform was ensured by running 293T cell protein extracts as quality control (QC) samples during the data-collecting period. The median of the Pearson correlation coefficienct between QC samples was 0.93 (Fig. s1a). With this procedure, we measured 239 urine samples from 236 patients, including uncomplicated diabetes without nephropathy (diabetes herein for simplicity, 134 samples from 132 patients), DKD (65 samples from 64 patients), and chronic kidney disease without diabetes (CKD herein for simplicity, 40 samples from 40 patients) (Fig. 1a, Table s1). A total of 2,946 proteins were identified and quantified with high analytical confidence (Fig. s1b, Table s1) with a median of 571 proteins identified per urine sample (Fig. s1c). Principle component analysis (PCA) of all proteins in the dataset showed that, when all samples were projected onto the first two principal components, representing 79.6% and 13.5% of the variance respectively (Fig. 1b), diabetes and CKD can be clearly separated, whereas DKD appears to be in between. Next, we selected differentially excreted proteins (DEPs) in each group, using the following criteria: 1) the ratio of the mean of the highest group / second highest group >= 2, and 2) the p-value of Kruskal-Wallis test <0.05 (Table s2). The 509 DEPs whose expressions were at least two-fold higher in one disease than in any other disease were defined as disease-specific DEPs (323 DEPs for diabetes, 98 DEPs for DKD, 88 DEPs for CKD). Reactome analysis showed that proteins specific in diabetes were enriched in late endosomal microautophagy, regulation of IGF transport and uptake by IGFBPs. In contrast, innate immune system and platelet activation were significantly up-regulated in both DKD and CKD, indicating that the two diseases share similar pathology. Moreover, enrichment of these pathways was higher in statistical significance between CKD and diabetes than those between DKD and diabetes. Notably, the formation of fibrin clot (clotting cascade) was up-regulated in CKD, but not in DKD (Fig. 1c). To investigate whether the urinary proteome correlated with the clinical parameters assessed by serum tests, we performed correlation analysis for the 2,946 proteins and 13 routinely tested clinical or health indexes. We calculated the Pearson correlation coefficient for each protein with each index (Fig. 1d, Table s2). Strong correlations (Pearson correlation coefficient > 0.5 or < -0.5) were found between 4 kidney functional indexes (serum creatinine, albumin-to-creatinine ratio, blood urea nitrogen, and glomerular filtration rate) and 46 urinary proteins. For example, serum creatinine was strongly correlated with the urinary protein abundance of RBP4 (retinol binding protein 4) and C7 (complement C7), with Pearson correlation coefficient of 0.812 and 0.769, respectively. C7 abundance was also correlated with blood urea nitrogen (Pearson correlation coefficient: 0.729) (Fig. 1e). Consistently, 3 proteins (ALB, RBP4, C7) were negative correlated with glomerular filtration rate (Fig. s1d). These results indicate that serum clinical indexes may be reflected in the urine proteome, which could provide additional biological insights into kidney dysfunctions in these diseases. Distinguish DKD from Diabetes with a 2-protein Classifier We used the dataset collected from year 2018 as the discovery dataset and the dataset collected from year 2019 as the validation dataset (Fig. 1a). To identify urinary biomarkers that distinguish DKD from diabetes, we determined DEPs between diabetes and the DKD group (including both DKD3 and DKD4 samples) (Fig. 2a). From the discovery dataset, which contained 21 diabetes and 9 DKD, we found 177 DEPs (Mann–Whitney U test p values 3), among which 79 were up-regulated and 73 were down-regulated in DKD patients (Fig. 2b) Next, we selected high-abundant DEPs (mean iFOT > 1) in diabetes or DKD as potential biomarkers, resulting in 135 proteins (Table s3). After dimension reduction, we showed that using 2 proteins (ALB, AFM) could readily distinguish DKD from diabetes (AUC: 0.928, accuracy: 85.2%, specificity: 83.2%, sensitivity: 87.5%, Fig. 2c) in the validation set. Both ALB and AFM were significantly excreted at higher levels in the validation dataset (Fig. 2d). Reactome analysis revealed that the up-regulated DEPs were enriched in neutrophil degranulation, adaptive immune system, and complement cascade, while the down-regulated DEPs were enriched in membrane trafficking and cellular responses to stress (Fig. 2e). These results indicated the major functional differences between DKD and diabetics as revealed by the urine proteome. Close examination of the high abundance DEPs (top 50) revealed that liver- (29/50) and bone marrow (6/50)-specific proteins[ 26 ] were enriched in DKD, but low tissue specificity were found among those highly expressed in diabetes (Table s3). Distinguish DKD3 from DKD4 with a 3-protein classifier As DKD3 and DKD4 are managed differently in the clinics[ 27 ], a non-invasive test to distinguish the DKD stages could aid in disease management. To this end, we first used the DKD3 (n=3) and DKD4 (n=6) samples in the discovery set to find DEPs (Fig. 3a). Since the difference between DKD3 and DKD4 might be smaller than that between diabetes and DKD, we applied a more relaxed constraints for protein filtering (Mann–Whitney U test p 3 and detection frequency > 50%). We identified 104 DEPs with 45 up-regulated and 59 down-regulated proteins in DKD4 compared to DKD3 (Table s4). An unsupervised hierarchical clustering analysis of the validation set using the 104 DEPs showed that DKD3 and DKD4 can be well separated (Fig. 3b), suggesting that significant differences have occurred when DKD3 progressed to DKD4. Using the same protein filtering criteria in the validation set, 32 up-regulated and 468 down-regulated proteins were found in DKD4 compared to DKD3 (Table s4). A Venn diagram showed that 7 of the 45 (24.4%) up-regulated DEPs in the discovery set and 30 of the 59 (50.8%) down-regulated DEPs were also found in the validation set (Fig. 3c). After performing a dimension reduction, a 3-protein classifier could distinguish DKD3 from DKD4. Applying the prediction model to the validation set resulted in an ROC of 0.949 (accuracy: 85.7%, specificity: 86.8%, sensitivity: 83.3%, Fig. 3d). Reactome analysis revealed that the up-regulated DEPs in DKD3, including all 7 overlapping ones, were enriched in hemostasis and immune system, especially in complement cascade (Fig. 3e-f). Monitor the transition to DKD with a 4-protein classifier Since our goal is early diagnosis for DKD, we re-analyzed DKD3 and diabetes datasets to identify proteins that could implicate the transition from diabetes to DKD (Fig. 4a). As small sample size would result in poor estimation and low statistical power, in the discovery set, we performed DEP analysis between diabetes group (n=21) and the DKD group (n=3) with the following 2 criteria: 1) the mean iFOT in DKD group was 2 times higher than that in diabetes group; 2) detected in at least 2 samples in the DKD group, resulting in 203 DEPs. Next, we identified DEPs (Mann–Whitney U test p 2 and detection frequency > 50%) in the validation dataset, resulting in 73 DEPs. After performing the dimension reduction from the 24 overlapping DEPs from discovery and validation datasets, a 4-protein classifier (SERPINA5, VPS4A, CP, TF) could distinguish diabetes from DKD3 (Fig. 4b). Applying this classifier to the validation dataset resulted in an area under the ROC curve of 0.952 (accuracy: 89.4%, specificity: 88.5%, sensitivity: 92.1%, Fig. 4c). We noted that 13 of the 113 diabetic samples were classified as DKD3 (Fig. 4d). This seemingly false-positive result may indicate the inaccuracy of the classifier; alternatively, it could also indicate clinical misdiagnosis, as progression from diabetes to early stage of DKD (pre-DKD3) is difficult to detect. We then explored the possibility of inferring the uncomplicated diabetes to “pre-DKD3” progression based on the difference between diabetes and these 13 false-positive DKD3 samples. To this end, we identified DEPs between the 13 false-positive DKD3 and the rest of the diabetes samples in the validation set. Using the criteria described (Wilcoxon p value =2 or 0.6), we found 35 DEPs (16 upregulated and 19 downregulated, Table s6), which include all 4 classifier proteins, as expected. As shown in Fig. 4e, the levels of the 4 classifier proteins, as well as several other DEPs also showed significant changes at other stages of the disease, and the changes were more significant as the severity of the disease progress. Together these data indicate that the dysfunctional pathways in different stages of the diseases share similarity and DKD is a progressive disease. To test the predictive power of our model, we analyzed longitudinal urine samples collected from 11 diabetic patients, among which 6 were predicted as the high-risk “pre-DKD” patients. We assess the progression to DKD by calculating the risk scores for each urine sample using model 3. Four of the 6 pre-DKD patients were classified as DKD3 once again(Fig. s2). Importantly, one of the patients (clinical ID: 0359150) was diagnosed as DKD3 by clinical index (Fig. 4f, Table s7). Moreover, one patient (clinical ID: 8073510), who was classified as diabetes in the validation dataset and re-classified as DKD3 in 2020, was admitted to hospital in March 2021 for DKD treatment, and then recovered in July 2021 after Valsartan Capsules treatment. These results suggest urine proteomics has the potential to monitor DKD progression, especially for high-risk diabetes patients. Discussion In this study, we conducted urinary proteomic studies on 236 patients. This streamlined, sedimentation-based method allowed us to detect 559 proteins/sample on average and a total of 2,946 proteins, requiring 0.5 hour/sample of processing time. The data allowed us to build models that could distinguish DKD from diabetes (AUC = 0.928) with 2 proteins (ALB, AFM), and distinguish DKD3 from DKD4 (AUC = 0.949) with 3 proteins (ANXA7, APOD, C9). Moreover, a 4-protein (SERPINA5, VPS4A, CP, TF) classifier was built to predict early stage DKD3 from diabetes (Fig. 4g). Importantly, a small-scale follow-up studies on 11 patients showed that 2 high-risk pre-DKD patients predicted by our model had progressed to clinical DKD3, confirming the ability of our approach as a noninvasive method to monitor progression of renal dysfunctions, particularly for patients who have been classified as high-risk. Previous studies showed that inactivation of similar biological processes may contribute to the disease progression. Samples in our dataset allowed us to analyze the different stages of the disease in a quantitative manner. We found that similar biological processes are indeed altered at different stages. Importantly, activation of complement cascades is unique for DKD4 and CKD. Our results also suggested that CKD caused by other renal dysfunction can be readily differentiated from DKD, and for the most part, appeared to have caused more severe injury to the kidney. However, these results should be interpreted with caution as the CKD cases in our dataset have complex etiology. Larger sample size and more careful data analysis could reveal the underlying mechanism of the changes. Abnormal urine albumin excretion is a hallmark for incipient nephropathy. This has been attributed to dysfunction in glomerular permselectivity and tubular reabsorption, among other causes. Our data showed that, besides albumin, which is the 2nd most abundant urinary excreted proteins identified in our assay, other small serum proteins also showed an increase in excretion, consistent with the notion of renal dysfunction as an underlying cause. Recently, Van and colleague reviewed published DKD urinary proteomic/peptidomic literatures, and compiled a list of 75 most robust candidate markers at each stage of diabetic kidney disease and highlighted their roles in biological processes that may contribute to progression[ 28 ]. Of the 75 biomarkers, 55 (73.3%) can be detected in our dataset and 23 of them (41.8%) were among the top 100 abundant proteins in our dataset (Table s1). Interestingly, many small serum proteins, including SERPINA1, TF, AMBP, APOA1, that showed increased excretion were expressed specifically or enriched in liver and bone marrow. On the other hand, proteins that showed decreased excretion, such as AHSG, CUBN, MASP2, showed low tissue specificity and were highly expressed in more diverse tissues. Notably, expressions of several decreased proteins have been reported in the literatures or Human Protein Atlas as enriched or specific in the tubular segment and play important functional roles in endocytosis. Recent technical advancement in mass spectrometry empowered high throughput unbiased biomarker discoveries. For example, a urinary peptide-based classifier consisting of 273 naturally occurring urinary peptides was obtained from 3,600 individuals analyzed by capillary electrophoresis coupled to MS. The so-called “CKD273” first proposed by Good et al[ 29 ] can differentiate patients of chronical kidney diseases with different etiology including those caused by diabetes. Recently, the CKD273 classifier was applied in a trial for early detection of DKD (PRIORITY). This multicenter, prospective, observational study showed that a high-risk CKD273 score was associated with an increased risk of progression to microalbuminuria over a median of 2.5 years, independent of clinical characteristics. However, since the CE-MS detects naturally occurring urinary peptides without enrichment, the variety of proteins detected is often limited. In fact, the CKD273 peptides were derived from a total of 30 proteins and the collagen fragments constitute more than 66% of the 273 peptides. This low protein diversity provides limited biological insights into disease etiology and offered few treatment recommendations. Alternatively, analysis of urinary proteins, particularly with the enrichment of excreted vesicles, allows the detection and quantification of thousands of proteins from 5 ml of a urine sample, and the streamlined workflow also enables improved quantitative measurement. Interestingly, collagens were not detected in high abundance as excreted proteins in our dataset. This is in contrast to results from urine peptidomics analyses represented by the CKD 273 panel, as Collagen a-1 (I) (126 fragments) and (III) (55 fragments) chains constitute >66% of the 273 markers. Since the CKD273 detects naturally occurring peptides generated by endogenous proteases, these small proteins or peptides were likely lost in our sedimentation process. Together these results suggest that proteomics and peptidomics analyses are complementary assays that could provide more accurate diagnosis with biological insights. In summary, the urinary proteomic cross-sectional study allowed the differentiation of uncomplicated diabetes and DKDs at different stages, and the identification of biomarkers. Disease progression of the high-risk patients identified by the pre-DKD classifier was partly validated in a small-scale follow-up study. These results warrant the the design of perspective studies to test the predictive power of our model for monitoring high-risk patients. Conclusions In this study, we have established a urinary proteomic workflow to conduct a cross-sectional investigation of uncomplicated diabetes, DKD, and CKD patients. We identified biomarkers that could distinguish DKD from uncomplicated diabetes, and stage 4 DKD from stage 3 DKD. Logistic regression models were built upon the discovery dataset, and the classification were validated in an independent dataset. Bioinformatics analysis suggests that (upregulation of DKD4) complement cascade is an indication of DKD progression. Moreover, an algorithm was established to identify earlier stage DKD patients. Follow-up studies on 11 patients indicated that 2 putative pre-DKD patients have progressed to DKD3. This study demonstrated the potential for urine proteomics as a noninvasive method for DKD diagnosis and identifying high-risk patients for progression monitoring. The pathways identified from differentially excreted proteins provides important clues on biological basis during disease progression. Declarations Acknowledgments The authors thank all the colleagues involved in this project, including clinicians and nurses from Guangdong Provincial Hospital of Chinese Medicine (Guangzhou, China), students and engineers from National Center for Protein Sciences (Beijing, China). The authors acknowledge the support of Guangdong Provincial Hospital of Chinese Medicine (Guangzhou, China) and National Center for Protein Sciences (Beijing, China) for providing study participants and case samples and laboratory for this study. Author’s C ontribut ors G.F., Y.W., and J.Q. conceived and designed the study. T.G., X.N. and J.W. had primary responsibility for writing of the manuscript with the support of Y.W. and J.Q. Y.L., H.W., Z.L., L.Z., H.H., Y.W., X.T. and J.Q. were responsible for study implementation and administration. X.Y., H.L., and X.L. performed the experiments; T.G., X.N., J.W., L.S. and L.L. contributed to urine sample testing and bioinformatics analysis with the guidance of Y.W. and J.Q.. All other authors collected urine sample from subjects. G.F. and J.Q. are the guarantors of this work and, as such, had full access to all the data in the study and takes responsibility for the integrity of the data and the accuracy of the data analysis. Funding This work was supported by grants from the Guangdong Provincial Hospital of Chinese Medicine (2018DB03, YN2021DB02), National Key Research and Development Program of China (2017YFC0908404, 2018YFA0507503, 2017YFA0505102), State Key Laboratory of Proteomics (No. SKLP-K201705). Availability of d ata and materials The MS raw data generated in this study have been submitted to ProteomeXchange database (www.proteomexchange.org) via the iProX partner repository[30] under accession number IPX0002858000. Ethics approval and consent to participate The study was approved by the medical ethics committee of the Guangdong Provincial Hospital of Chinese Medicine. Consent for publication All authors consent to the publication of this manuscript. Competing interests Dr. Jun Qin is the cofounder and co-owner of the Beijing Pineal Health Management Co., Ltd.. Dr. Xiaotian Ni, Tongqing Gong., Xing Yang., Haibo Liu., Xinliang Li., Lifeng Liu, are employees of Beijing Pineal Health Management Co., Ltd.. All other authors declare no competing interests. Authors' information 1 The Second Affiliated Hospital of Guangzhou University of Chinese Medicine, Guangzhou, 510120 China 2 The Second Clinical College of Guangzhou University of Chinese Medicine, Guangzhou, 510120 China 3 Guangdong Provincial Hospital of Chinese Medicine, Guangzhou, 510120 China 4 Guangdong Provincial Academy of Chinese Medical Sciences, Guangzhou, 510120 China 5 Beijing Pineal Health Management Co., Ltd., Beijing 102206, China 6 State Key Laboratory of Proteomics, Beijing Proteome Research Center, National Center for Protein Sciences, Institute of Lifeomics, Beijing, 102206, China 7 Chongqing Key Laboratory of Big Data for Bio Intelligence, School of Bioinformation, Chongqing University of Posts and Telecommunications, Chongqing, 400065 China References Saeedi P, Petersohn I, Salpea P, Malanda B, Karuranga S, Unwin N, et al. Global and regional diabetes prevalence estimates for 2019 and projections for 2030 and 2045: Results from the International Diabetes Federation Diabetes Atlas, 9th edition. Diabetes Res Clin Pract 2019;157:107843. https://doi.org/10.1016/j.diabres.2019.107843. Gheith O, Farouk N, Nampoory N, Halim MA, Al-Otaibi T. Diabetic kidney disease: world wide difference of prevalence and risk factors. J Nephropharmacol 2016;5:49–56. Maahs DM, Rewers M. Editorial: Mortality and renal disease in type 1 diabetes mellitus--progress made, more to be done. J Clin Endocrinol Metab 2006;91:3757–9. https://doi.org/10.1210/jc.2006-1730. Orchard TJ, Secrest AM, Miller RG, Costacou T. In the absence of renal disease, 20 year mortality risk in type 1 diabetes is comparable to that of the general population: a report from the Pittsburgh Epidemiology of Diabetes Complications Study. Diabetologia 2010;53:2312–9. https://doi.org/10.1007/s00125-010-1860-3. Collins AJ, Foley RN, Chavers B, Gilbertson D, Herzog C, Johansen K, et al. ’United States Renal Data System 2011 Annual Data Report: Atlas of chronic kidney disease & end-stage renal disease in the United States. Am J Kidney Dis 2012;59:A7, e1-420. https://doi.org/10.1053/j.ajkd.2011.11.015. Macisaac RJ, Ekinci EI, Jerums G. Markers of and risk factors for the development and progression of diabetic kidney disease. Am J Kidney Dis 2014;63:S39-62. https://doi.org/10.1053/j.ajkd.2013.10.048. Tuttle KR, Bakris GL, Bilous RW, Chiang JL, de Boer IH, Goldstein-Fuchs J, et al. Diabetic kidney disease: a report from an ADA Consensus Conference. Diabetes Care 2014;37:2864–83. https://doi.org/10.2337/dc14-1296. Mischak H. Pro: urine proteomics as a liquid kidney biopsy: no more kidney punctures! Nephrol Dial Transplant 2015;30:532–7. https://doi.org/10.1093/ndt/gfv046. Vinik AI, Nevoret M-L, Casellini C, Parson H. Diabetic neuropathy. Endocrinol Metab Clin North Am 2013;42:747–87. https://doi.org/10.1016/j.ecl.2013.06.001. National Kidney Foundation. KDOQI Clinical Practice Guideline for Diabetes and CKD: 2012 Update. Am J Kidney Dis 2012;60:850–86. https://doi.org/10.1053/j.ajkd.2012.07.005. Chen Y, Lee K, Ni Z, He JC. Diabetic Kidney Disease: Challenges, Advances, and Opportunities. Kidney Dis (Basel) 2020;6:215–25. https://doi.org/10.1159/000506634. Smart NA, Dieberg G, Ladhani M, Titus T. Early referral to specialist nephrology services for preventing the progression to end-stage kidney disease. Cochrane Database Syst Rev 2014:CD007333. https://doi.org/10.1002/14651858.CD007333.pub2. Mullen W, Delles C, Mischak H, EuroKUP COST action. Urinary proteomics in the assessment of chronic kidney disease. Curr Opin Nephrol Hypertens 2011;20:654–61. https://doi.org/10.1097/MNH.0b013e32834b7ffa. Li J, Zhang Z, Rosenzweig J, Wang YY, Chan DW. Proteomics and bioinformatics approaches for identification of serum biomarkers to detect breast cancer. Clin Chem 2002;48:1296–304. Fliser D, Novak J, Thongboonkerd V, Argilés A, Jankowski V, Girolami MA, et al. Advances in urinary proteome analysis and biomarker discovery. J Am Soc Nephrol 2007;18:1057–71. https://doi.org/10.1681/ASN.2006090956. Gao Y. Urine is a better biomarker source than blood especially for kidney diseases. Adv Exp Med Biol 2015;845:3–12. https://doi.org/10.1007/978-94-017-9523-4_1. Rossing K, Mischak H, Dakna M, Zürbig P, Novak J, Julian BA, et al. Urinary proteomics in diabetes and CKD. J Am Soc Nephrol 2008;19:1283–90. https://doi.org/10.1681/ASN.2007091025. Zhu C, Liang Q, Hu P, Wang Y, Luo G. Phospholipidomic identification of potential plasma biomarkers associated with type 2 diabetes mellitus and diabetic nephropathy. Talanta 2011;85:1711–20. https://doi.org/10.1016/j.talanta.2011.05.036. Zürbig P, Jerums G, Hovind P, Macisaac RJ, Mischak H, Nielsen SE, et al. Urinary proteomics for early diagnosis in diabetic nephropathy. Diabetes 2012;61:3304–13. https://doi.org/10.2337/db12-0348. Liao W-L, Chang C-T, Chen C-C, Lee W-J, Lin S-Y, Liao H-Y, et al. Urinary Proteomics for the Early Diagnosis of Diabetic Nephropathy in Taiwanese Patients. J Clin Med 2018;7. https://doi.org/10.3390/jcm7120483. Kwiendacz H, Nabrdalik K, Stompór T, Gumprecht J. What do we know about biomarkers in diabetic kidney disease? Endokrynol Pol 2020;71:545–50. https://doi.org/10.5603/EP.a2020.0077. Feng J, Ding C, Qiu N, Ni X, Zhan D, Liu W, et al. Firmiana: towards a one-stop proteomic cloud platform for data processing and analysis. Nat Biotechnol 2017;35:409–12. https://doi.org/10.1038/nbt.3825. Schwanhäusser B, Busse D, Li N, Dittmar G, Schuchhardt J, Wolf J, et al. Global quantification of mammalian gene expression control. Nature 2011;473:337–42. https://doi.org/10.1038/nature10098. Jassal B, Matthews L, Viteri G, Gong C, Lorente P, Fabregat A, et al. The reactome pathway knowledgebase. Nucleic Acids Res 2020;48:D498–503. https://doi.org/10.1093/nar/gkz1031. Leng W, Ni X, Sun C, Lu T, Malovannaya A, Jung SY, et al. Proof-of-Concept Workflow for Establishing Reference Intervals of Human Urine Proteome for Monitoring Physiological and Pathological Changes. EBioMedicine 2017;18:300–10. https://doi.org/10.1016/j.ebiom.2017.03.028. Uhlén M, Fagerberg L, Hallström BM, Lindskog C, Oksvold P, Mardinoglu A, et al. Proteomics. Tissue-based map of the human proteome. Science 2015;347:1260419. https://doi.org/10.1126/science.1260419. Selby NM, Taal MW. An updated overview of diabetic nephropathy: Diagnosis, prognosis, treatment goals and latest guidelines. Diabetes Obes Metab 2020;22 Suppl 1:3–15. https://doi.org/10.1111/dom.14007. Van JAD, Scholey JW, Konvalinka A. Insights into Diabetic Kidney Disease Using Urinary Proteomics and Bioinformatics. J Am Soc Nephrol 2017;28:1050–61. https://doi.org/10.1681/ASN.2016091018. Good DM, Zürbig P, Argilés A, Bauer HW, Behrens G, Coon JJ, et al. Naturally occurring human urinary peptides for use in diagnosis of chronic kidney disease. Mol Cell Proteomics 2010;9:2424–37. https://doi.org/10.1074/mcp.M110.001917. Ma J, Chen T, Wu S, Yang C, Bai M, Shu K, et al. iProX: an integrated proteome resource. Nucleic Acids Res 2019;47:D1211–7. https://doi.org/10.1093/nar/gky869. Supplementary Files Fig.s1.pdf Fig. s1 a. LC-MS/MS analysis of 293T cells as quality control samples. b. The dynamic range of the proteomics data. c. Number of GPs quantified in every urinary sample. d. Scatter plots showing negative correlation coefficients between the urinary proteins and glomerular filtration rate. Fig.s2.pdf Fig. s2 a. The risk scores and the 4-marker expression of the high-risk “pre-DKD” patients. b. The risk scores and the 4-marker expression of the diabetic patients. Tables1.xlsx Tables2.xlsx Tables3.xlsx Tables4.xlsx Tables5.xlsx Tables6.xlsx Tables7.xlsx Cite Share Download PDF Status: Published Journal Publication published 30 Nov, 2021 Read the published version in Clinical Proteomics → Version 1 posted Reviewer # 2 agreed at journal 02 Nov, 2021 Review # 1 received at journal 02 Nov, 2021 Review # 2 received at journal 02 Nov, 2021 Editorial decision: Minor revision 02 Nov, 2021 Reviews received at journal 18 Oct, 2021 Reviewers invited by journal 17 Oct, 2021 Reviewer # 1 agreed at journal 17 Oct, 2021 Editor assigned by journal 16 Oct, 2021 Submission checks completed at journal 15 Oct, 2021 Editor invited by journal 15 Oct, 2021 First submitted to journal 14 Oct, 2021 You are reading this latest preprint version Research Square lets you share your work early, gain feedback from the community, and start making changes to your manuscript prior to peer review in a journal. As a division of Research Square Company, we’re committed to making research communication faster, fairer, and more useful. We do this by developing innovative software and high quality services for the global research community. Our growing team is made up of researchers and industry professionals working together to solve the most critical problems facing scientific publishing. Also discoverable on Platform About Our Team In Review Editorial Policies Advisory Board Help Center Resources Author Services Accessibility API Access RSS feed Manage Cookie Preferences © Research Square 2026 | ISSN 2693-5015 (online) Privacy Policy Terms of Service Do Not Sell My Personal Information {"props":{"pageProps":{"initialData":{"identity":"rs-966683","acceptedTermsAndConditions":true,"allowDirectSubmit":false,"archivedVersions":[],"articleType":"Research","associatedPublications":[],"authors":[{"id":57053921,"identity":"6974047c-389f-4a23-b26b-a2e14a1ae353","order_by":0,"name":"Guanjie Fan","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Guanjie","middleName":"","lastName":"Fan","suffix":""},{"id":57053922,"identity":"1013c53f-db0b-40f0-817d-519cdd05a020","order_by":1,"name":"Tongqing Gong","email":"","orcid":"","institution":"Beijing Pineal Health Management Co., Ltd.","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Tongqing","middleName":"","lastName":"Gong","suffix":""},{"id":57053923,"identity":"9c427311-6b65-4b76-99ca-ec26f648f2b9","order_by":2,"name":"Yuping Lin","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Yuping","middleName":"","lastName":"Lin","suffix":""},{"id":57053924,"identity":"491cc5b2-f957-4270-895e-47c085e02e41","order_by":3,"name":"Jianping Wang","email":"","orcid":"","institution":"Beijing Proteome Research Center","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Jianping","middleName":"","lastName":"Wang","suffix":""},{"id":57053925,"identity":"5d2628ec-f492-4f99-b9cb-482ab4af6273","order_by":4,"name":"Lu Sun","email":"","orcid":"","institution":"Second Clinical Medical College of Guangzhou University of Chinese Medicine: The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Lu","middleName":"","lastName":"Sun","suffix":""},{"id":57053926,"identity":"777b35c6-68a9-4f3a-84c6-5316b590bdf0","order_by":5,"name":"Hua Wei","email":"","orcid":"","institution":"Second Clinical Medical College of Guangzhou University of Chinese Medicine: The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Hua","middleName":"","lastName":"Wei","suffix":""},{"id":57053927,"identity":"6b29c4f5-5c4a-49fc-adb7-22f597571c5a","order_by":6,"name":"Xing Yang","email":"","orcid":"","institution":"Beijing Pineal Health Management Co., Ltd.","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Xing","middleName":"","lastName":"Yang","suffix":""},{"id":57053928,"identity":"a138774c-6668-41d8-919f-6680c15562ca","order_by":7,"name":"Zhenjie Liu","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Zhenjie","middleName":"","lastName":"Liu","suffix":""},{"id":57053929,"identity":"1f5cb7f4-4cfb-4fad-8a15-4b88564b9d8f","order_by":8,"name":"Xinliang Li","email":"","orcid":"","institution":"Beijing Pineal Health Management Co., Ltd.","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Xinliang","middleName":"","lastName":"Li","suffix":""},{"id":57053930,"identity":"fc8743cd-e277-4133-acf2-137910ac1005","order_by":9,"name":"Ling Zhao","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Ling","middleName":"","lastName":"Zhao","suffix":""},{"id":57053931,"identity":"19911d8e-2d70-4f6c-b4c5-8d53c45b3f35","order_by":10,"name":"Lan Song","email":"","orcid":"","institution":"Beijing Proteome Research Center","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Lan","middleName":"","lastName":"Song","suffix":""},{"id":57053932,"identity":"342f479c-cd08-4b80-be55-6686552aae78","order_by":11,"name":"Jiali He","email":"","orcid":"","institution":"Second Clinical Medical College of Guangzhou University of Chinese Medicine: The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Jiali","middleName":"","lastName":"He","suffix":""},{"id":57053933,"identity":"aa59d3aa-98bc-4891-937a-a5671597ccff","order_by":12,"name":"Haibo Liu","email":"","orcid":"","institution":"Beijing Pineal Health Management Co., Ltd.","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Haibo","middleName":"","lastName":"Liu","suffix":""},{"id":57053934,"identity":"7ce8b668-5c05-44dc-bd98-a43c58b28cda","order_by":13,"name":"Xiuming Li","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Xiuming","middleName":"","lastName":"Li","suffix":""},{"id":57053935,"identity":"475b0eea-a939-45ff-8929-e9e47a006cfa","order_by":14,"name":"Lifeng Liu","email":"","orcid":"","institution":"Beijing Pineal Health Management Co., Ltd.","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Lifeng","middleName":"","lastName":"Liu","suffix":""},{"id":57053936,"identity":"d3c089a6-4b48-426c-bc23-b69c0f675394","order_by":15,"name":"Anxiang Li","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Anxiang","middleName":"","lastName":"Li","suffix":""},{"id":57053937,"identity":"9176ee0c-76e9-423d-9587-165302abc84f","order_by":16,"name":"Qiyun Lu","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Qiyun","middleName":"","lastName":"Lu","suffix":""},{"id":57053938,"identity":"de557ea7-0629-42d2-b397-fa504a6ee54c","order_by":17,"name":"Dongyin Zou","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Dongyin","middleName":"","lastName":"Zou","suffix":""},{"id":57053939,"identity":"9688c274-e7f5-4ab0-a830-9bf5fc0b27eb","order_by":18,"name":"Jianxuan Wen","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Jianxuan","middleName":"","lastName":"Wen","suffix":""},{"id":57053940,"identity":"c3757077-0150-4ec2-9b57-cd55ebd7652b","order_by":19,"name":"Yaqing Xia","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Yaqing","middleName":"","lastName":"Xia","suffix":""},{"id":57053941,"identity":"631a10a0-3f1c-4037-9894-209c4a85c19b","order_by":20,"name":"Liyan Wu","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Liyan","middleName":"","lastName":"Wu","suffix":""},{"id":57053942,"identity":"831577fa-18d8-43ec-99a3-c9ac1cf1d608","order_by":21,"name":"Haoyue Huang","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Haoyue","middleName":"","lastName":"Huang","suffix":""},{"id":57053943,"identity":"74a78418-3f02-4d77-83c8-e489d4a05594","order_by":22,"name":"Yuan Zhang","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Yuan","middleName":"","lastName":"Zhang","suffix":""},{"id":57053944,"identity":"ca2d0807-f05d-4ec2-bcff-bb1732a22c4f","order_by":23,"name":"Wenwen Xie","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Wenwen","middleName":"","lastName":"Xie","suffix":""},{"id":57053945,"identity":"6d55804e-f253-4b70-895b-cf75a361b309","order_by":24,"name":"Jinzhu Huang","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Jinzhu","middleName":"","lastName":"Huang","suffix":""},{"id":57053946,"identity":"c2b94be8-5140-4fe8-bfe2-2c2a83080061","order_by":25,"name":"Lulu Luo","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Lulu","middleName":"","lastName":"Luo","suffix":""},{"id":57053947,"identity":"a0520bf1-68c7-4b91-9c0a-b4355ae18b6f","order_by":26,"name":"Lulu Wu","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Lulu","middleName":"","lastName":"Wu","suffix":""},{"id":57053948,"identity":"71ce6230-cd0c-4e69-936d-790b3c5ca608","order_by":27,"name":"Liu He","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Liu","middleName":"","lastName":"He","suffix":""},{"id":57053949,"identity":"ff2d3deb-222d-4b31-9038-f7f53b61cbd0","order_by":28,"name":"Qingshun Liang","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Qingshun","middleName":"","lastName":"Liang","suffix":""},{"id":57053950,"identity":"8b388f16-e040-4051-be64-36f5084d1e1d","order_by":29,"name":"Qubo Chen","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Qubo","middleName":"","lastName":"Chen","suffix":""},{"id":57053951,"identity":"46c03e17-ed4a-472c-93f6-c1b971e45532","order_by":30,"name":"Guowei Chen","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Guowei","middleName":"","lastName":"Chen","suffix":""},{"id":57053952,"identity":"af7bc5cd-d080-4298-8cb0-8a2f9901450c","order_by":31,"name":"Mingze Bai","email":"","orcid":"","institution":"Chongqing University of Posts and Telecommunications","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Mingze","middleName":"","lastName":"Bai","suffix":""},{"id":57053953,"identity":"f2621be6-11b9-4cc3-ae67-3faada237fa7","order_by":32,"name":"Jun Qin","email":"","orcid":"","institution":"Beijing Proteome Research Center","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Jun","middleName":"","lastName":"Qin","suffix":""},{"id":57053954,"identity":"c0d85b64-3b54-43bd-aa19-a69b5a90bfd8","order_by":33,"name":"Xiaotian Ni","email":"","orcid":"","institution":"Beijing Pineal Health Management Co., Ltd.","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Xiaotian","middleName":"","lastName":"Ni","suffix":""},{"id":57053955,"identity":"be49a2c8-cdc7-4354-aa3b-70c0c62e9bfc","order_by":34,"name":"Xianyu Tang","email":"","orcid":"","institution":"The Second Affiliated Hospital of Guangzhou University of Chinese Medicine","correspondingAuthor":false,"submittingAuthor":false,"prefix":"","firstName":"Xianyu","middleName":"","lastName":"Tang","suffix":""},{"id":57053956,"identity":"90865ce7-7a89-4fc4-b3f8-b20b47760b8d","order_by":35,"name":"Yi Wang","email":"data:image/png;base64,iVBORw0KGgoAAAANSUhEUgAAAZAAAAAyAQMAAABI0h/eAAAABlBMVEX///8AAABVwtN+AAAACXBIWXMAAA7EAAAOxAGVKw4bAAAA6klEQVRIiWNgGAWjYBACNvbmAwcSKiTk7NubDz5IqKghrIWP51jigQ9nLIwNeI4lGzw4c4ywFjmJHOODM9sqEjdI+JhJPmxhJsJhPAcMDvO2STBul2BLq0hsYGPgb+9OIOCXhoTDPOckmC1nNx+7kbhDhkHizNkNhGw5cJinTIKN4c6xtBuJZ9gYDCRyCWiRSGw4zMMmwcNwI8esILGNmRgtyQwHZ7RJSBgAtTAQp4XnGAMwkCUMJHuOJUsknDnGQ9Av8u39nz8kVNTV97M3H/z4o6JGjr+9F78WDMBDmvJRMApGwSgYBVgBACcIUEr/G9w5AAAAAElFTkSuQmCC","orcid":"","institution":"Beijing Proteome Research Center","correspondingAuthor":true,"submittingAuthor":false,"prefix":"","firstName":"Yi","middleName":"","lastName":"Wang","suffix":""}],"badges":[],"createdAt":"2021-10-13 10:29:56","currentVersionCode":1,"declarations":"","doi":"10.21203/rs.3.rs-966683/v1","doiUrl":"https://doi.org/10.21203/rs.3.rs-966683/v1","draftVersion":[],"editorialEvents":[{"content":"https://doi.org/10.1186/s12014-021-09338-6","type":"published","date":"2021-12-01T03:48:28+00:00"}],"editorialNote":"","failedWorkflow":false,"files":[{"id":14669716,"identity":"c985ddff-d42b-4e02-a37b-dac9413a08cb","added_by":"auto","created_at":"2021-10-19 14:28:14","extension":"png","order_by":1,"title":"Figure 1","display":"","copyAsset":false,"role":"figure","size":397106,"visible":true,"origin":"","legend":"Urine proteomic analysis of Diabetes, DKD, and CKD\na. Clinical stages and the number of samples used in the discovery and validation datasets. b. Principal component analysis of the proteomics data. Each dot represents an urinary sample. Blue: Diabetes, Yellow: DKD, Red: CKD. c. Reactome pathway analysis of the disease-specific DEPs (323 DEPs for diabetes, 98 DEPs for DKD, 88 DEPs for CKD). d. Pearson correlation coefficients between the abundance of the 2,946 urine proteins and the 13 routinely tested clinical or health indexes. e. Scatter plots of selected high abundance urine proteins and kidney function indexes with strong positive correlations.\n","description":"","filename":"1.png","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/e29cbc2e2c01ae695121f6c6.png"},{"id":14670230,"identity":"a8d19b25-3c3f-4c7e-9c3d-af7d6e079aec","added_by":"auto","created_at":"2021-10-19 14:31:14","extension":"png","order_by":2,"title":"Figure 2","display":"","copyAsset":false,"role":"figure","size":329473,"visible":true,"origin":"","legend":"A classifier for distinguishing DKD from Diabetes\na. A bioinformatic analysis workflow to find candidate biomarkers between the Diabetes and the DKD group. n: number of samples used in the analyses. b. Volcano plot displaying the differentially expressed proteins between Diabetes and DKD. Red and Blue indicated proteins that were significantly enriched in DKD and Diabetes, respectively (p values \u003c 0.05, more than 3-fold change). Other proteins were colored in grey. c. ROC curve of distinguishing the Diabetes and DKD samples predicted with the 2-protein classifier . d. Boxplots showing the ALB and AFM abundance in Diabetes and DKD in the two datasets (center line: median, bounds of box: 25th and 75th percentiles, and whiskers: from Q1-1.5*IQR to Q3+1.5*IQR, p-value calculated by Mann–Whitney U test). e. Reactome pathway analysis of the up-regulated (red) and down-regulated (blue) DEPs in DKD samples.\n","description":"","filename":"2.png","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/d4178e4861b96dbf9722363c.png"},{"id":14669717,"identity":"90d9dc33-e479-4165-8190-e89d084ec3c1","added_by":"auto","created_at":"2021-10-19 14:28:14","extension":"png","order_by":3,"title":"Figure 3","display":"","copyAsset":false,"role":"figure","size":310309,"visible":true,"origin":"","legend":"A classifier for distinguishing DKD4 from DKD3\na. A bioinformatic analysis workflow to find candidate biomarkers between DKD3 and DKD4 samples. n: number of samples used in the analyses. b. Hierarchical clustering of DKD3 and DKD4 DEPs using complete linkage. Protein expression values were normalized by z-scores. c. Venn diagram indicating the overlap of DEPs between discovery and validation datasets. d. ROC curve of distinguishing DKD3 and DKD4 predicted by a 3-protein classifier. e. Reactome pathway analysis of the up-regulated DEPs in DKD4 samples. f. Boxplots displaying the abundance of the complement component proteins at different stages of DKD and CKD in the two datasets.\n","description":"","filename":"3.png","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/d6b236bdd322f6df47361d91.png"},{"id":14670782,"identity":"0f0b071f-faee-40a6-83da-039de1c46387","added_by":"auto","created_at":"2021-10-19 14:34:14","extension":"png","order_by":4,"title":"Figure 4","display":"","copyAsset":false,"role":"figure","size":273940,"visible":true,"origin":"","legend":"A classifier for monitoring early transition to DKD \na. A bioinformatic analysis workflow to find candidate biomarkers between DKD3 and Diabetes samples. b. Venn diagram showing the overlap of DEPs in the discovery and the validation set. c. ROC curve of distinguishing DKD3 and Diabetes predicted by a 4-protein classifier. d. Dotplot indicating the Diabetes and DKD samples and their predicted stages by the classifer. Each point represent one urine sample. Blue and green dots presents Diabetes and DKD, respectively. The x-axis indicates the predicted DKD stage 3. Blue dots positioned on the right side of the Prediction line were predicted incorrectly as DKD3 by the model and were used as putative pre-DKD in the subsequent analysis. e. Boxplots displaying the iFOT intensities of the 4 biomarkers (CP, TF, SERPINA5, VPS4A) in Diabetes, DKD3, DKD4, and CKD samples. f. The risk scores(left panel) and the 4-marker expression (right panel) of the patient (clinical ID: 8073510). g. A schematic summary of bioinformatic analysis workflow that derived the 3 prediction models in this study.\n","description":"","filename":"4.png","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/f99ed3ee21c68cbcc1839661.png"},{"id":16819042,"identity":"c26ebcb2-319b-4d3d-9c8f-56cda43f2838","added_by":"auto","created_at":"2021-12-29 03:48:31","extension":"pdf","order_by":0,"title":"","display":"","copyAsset":false,"role":"manuscript-pdf","size":1558186,"visible":true,"origin":"","legend":"","description":"","filename":"manuscript.pdf","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/b0f2fd26-6550-405c-a4b2-1aa32e85c728.pdf"},{"id":14670227,"identity":"550f3689-2112-464c-8c2e-900d152b2ebc","added_by":"auto","created_at":"2021-10-19 14:31:14","extension":"pdf","order_by":1,"title":"","display":"","copyAsset":false,"role":"supplement","size":1366774,"visible":true,"origin":"","legend":"Fig. s1 a. LC-MS/MS analysis of 293T cells as quality control samples. b. The dynamic range of the proteomics data. c. Number of GPs quantified in every urinary sample. d. Scatter plots showing negative correlation coefficients between the urinary proteins and glomerular filtration rate.","description":"","filename":"Fig.s1.pdf","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/eee3b0388aa287e419a0de66.pdf"},{"id":14670226,"identity":"e8e0ed3c-ccec-4585-8711-e8d84bf5a130","added_by":"auto","created_at":"2021-10-19 14:31:14","extension":"pdf","order_by":2,"title":"","display":"","copyAsset":false,"role":"supplement","size":254484,"visible":true,"origin":"","legend":"Fig. s2 a. The risk scores and the 4-marker expression of the high-risk “pre-DKD” patients. b. The risk scores and the 4-marker expression of the diabetic patients.","description":"","filename":"Fig.s2.pdf","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/071bd843dd49b152676d2e83.pdf"},{"id":14669723,"identity":"ab0a9a6a-c20c-444d-a6e7-5e05a10510c0","added_by":"auto","created_at":"2021-10-19 14:28:14","extension":"xlsx","order_by":3,"title":"","display":"","copyAsset":false,"role":"supplement","size":3078428,"visible":true,"origin":"","legend":"","description":"","filename":"Tables1.xlsx","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/1463e0561935f9a8bee5e370.xlsx"},{"id":14669728,"identity":"a4296d87-84d4-472a-b804-c481262f3e35","added_by":"auto","created_at":"2021-10-19 14:28:15","extension":"xlsx","order_by":4,"title":"","display":"","copyAsset":false,"role":"supplement","size":25690026,"visible":true,"origin":"","legend":"","description":"","filename":"Tables2.xlsx","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/f731fc2cbbf71a3fbd002666.xlsx"},{"id":14669719,"identity":"70ca0d9b-36b6-4217-8222-709f70bb1bcf","added_by":"auto","created_at":"2021-10-19 14:28:14","extension":"xlsx","order_by":5,"title":"","display":"","copyAsset":false,"role":"supplement","size":41519,"visible":true,"origin":"","legend":"","description":"","filename":"Tables3.xlsx","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/3745a9f25ec996c14edae715.xlsx"},{"id":14669720,"identity":"df267134-1ba5-4b9e-95e6-54df56a11e79","added_by":"auto","created_at":"2021-10-19 14:28:14","extension":"xlsx","order_by":6,"title":"","display":"","copyAsset":false,"role":"supplement","size":51906,"visible":true,"origin":"","legend":"","description":"","filename":"Tables4.xlsx","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/9071d34668b1ca72eda4ac40.xlsx"},{"id":14670228,"identity":"475a3873-bfc5-4f3e-8fd3-db79e68f8813","added_by":"auto","created_at":"2021-10-19 14:31:14","extension":"xlsx","order_by":7,"title":"","display":"","copyAsset":false,"role":"supplement","size":27632,"visible":true,"origin":"","legend":"","description":"","filename":"Tables5.xlsx","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/2739c19ceb64decbd616e1b1.xlsx"},{"id":14669727,"identity":"e5108d5b-7297-4f5c-9750-0fdbedf2c695","added_by":"auto","created_at":"2021-10-19 14:28:14","extension":"xlsx","order_by":8,"title":"","display":"","copyAsset":false,"role":"supplement","size":11871,"visible":true,"origin":"","legend":"","description":"","filename":"Tables6.xlsx","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/432695e9c4f1e92b8d0aeed3.xlsx"},{"id":14669725,"identity":"eec49c24-954f-45ac-8023-3e6643536a89","added_by":"auto","created_at":"2021-10-19 14:28:14","extension":"xlsx","order_by":9,"title":"","display":"","copyAsset":false,"role":"supplement","size":226075,"visible":true,"origin":"","legend":"","description":"","filename":"Tables7.xlsx","url":"https://assets-eu.researchsquare.com/files/rs-966683/v1/f84d8850a727ced37837ab23.xlsx"}],"financialInterests":"","formattedTitle":"\u003cp\u003eUrine Proteomics Identifies Biomarkers for Diabetic Kidney Disease at Different Stages\u003c/p\u003e","fulltext":[{"header":"Background","content":"\u003cp\u003eIn 2019, more than 463 million people worldwide were estimated to be living with diabetes, representing 9.3% of the global adult population (20ཞ79 years)[\u003cspan citationid=\"CR1\" class=\"CitationRef\"\u003e1\u003c/span\u003e]. Among these people, 20\u0026ndash;40% will progress to DKD[\u003cspan citationid=\"CR2\" class=\"CitationRef\"\u003e2\u003c/span\u003e], which remains a leading cause of morbidity and mortality in people with type 2 diabetes[\u003cspan additionalcitationids=\"CR4\" citationid=\"CR3\" class=\"CitationRef\"\u003e3\u003c/span\u003e\u0026ndash;\u003cspan citationid=\"CR5\" class=\"CitationRef\"\u003e5\u003c/span\u003e]. DKD patients are at significant risk of progression to ESRD and cardiovascular diseases. Therefore, early detection, prevention, and treatment are of great importance for disease management[\u003cspan citationid=\"CR6\" class=\"CitationRef\"\u003e6\u003c/span\u003e]. Clinically, the diagnosis of DKD is based on the measurement of eGFR, albuminuria and urinary microalbumin creatinine ratio (UACR) along with other clinical features[\u003cspan citationid=\"CR7\" class=\"CitationRef\"\u003e7\u003c/span\u003e]. The \u0026ldquo;gold standard\u0026rdquo; for definitive diagnosis of DKD requires kidney biopsy with which an expert renal pathologist makes the diagnosis from histological tissue. However, the procedure is invasive and sometime dangerous, and the evaluation process can be biased by human judgment[\u003cspan citationid=\"CR8\" class=\"CitationRef\"\u003e8\u003c/span\u003e]. It is important to find a noninvasive test method that can complement or replace renal puncture.\u003c/p\u003e \u003cp\u003eDKD is divided into five stages according to clinical guidelines[\u003cspan citationid=\"CR9\" class=\"CitationRef\"\u003e9\u003c/span\u003e, \u003cspan citationid=\"CR10\" class=\"CitationRef\"\u003e10\u003c/span\u003e]. The stage 1 (DKD1) and stage 2 (DKD2) DKDs are preclinical stages, and are characterized by increase in glomerular filtration rate (GFR), normal albuminuria or intermittent microalbuminuria. DKD3, characterized by persistent microalbuminuria, mild hypertension, and a normal or slight decline in GFR, is the onset of clinical stage[\u003cspan citationid=\"CR11\" class=\"CitationRef\"\u003e11\u003c/span\u003e]. Patients with DKD4 exhibit clinical symptoms of edema and hypertension, along with an increase in albuminuria, which is difficult to treat[\u003cspan citationid=\"CR12\" class=\"CitationRef\"\u003e12\u003c/span\u003e]. In the overt DKD4, as the glomerular filtration rate (GFR) declines, the albumin-to-creatinine ratio (ACR) would further increase (\u0026gt;300 mg/g). Under this circumstance, taking any medicine could increase the kidney burden thus exacerbate the disease. With proper clinical intervention, the DKD progression can often be delayed or even reversed before it progresses to stage 4. Therefore, monitoring the kidney function for DKD3 patients is critical to the medical treatments and delaying disease prognosis.\u003c/p\u003e \u003cp\u003eIn the past years, several noninvasive methods have been proposed, mostly based on the \u0026lsquo;omics' techniques for the evaluation of urine or serum biomarkers[\u003cspan citationid=\"CR13\" class=\"CitationRef\"\u003e13\u003c/span\u003e, \u003cspan citationid=\"CR14\" class=\"CitationRef\"\u003e14\u003c/span\u003e]. Among these approaches, urinary proteome analysis has gained popularity, and has the potential to be transitioned towards clinical implementation[\u003cspan citationid=\"CR15\" class=\"CitationRef\"\u003e15\u003c/span\u003e, \u003cspan citationid=\"CR16\" class=\"CitationRef\"\u003e16\u003c/span\u003e]. Measuring urine proteome is non-invasive, and quantitative determining the urinary proteins and peptides can be an enabling platform to distinguish disease stages or monitor response to therapies. Much efforts have been made in finding biomarkers for CKD, aiding risk assessment for DKD patients. Using high-resolution capillary electrophoresis coupled with electrospray-ionization mass spectrometry, a panel of 40 biomarkers were identified from the urinary peptides and they were reported to be able to differentiate healthy individuals from diabetic patients with persistent normal albuminuria, low-grade albuminuria, or nephropathy (PREDICTION)[\u003cspan citationid=\"CR17\" class=\"CitationRef\"\u003e17\u003c/span\u003e]. Furthermore, it could distinguish patients with DKD from patients with other CKDs. However, the exact protein identities of these peptide biomarkers were not clear. A metabolic study utilized normal phase liquid chromatography coupled with time-of-flight mass spectrometry (NPLC-TOF/MS) investigated the profile of the plasma phospholipids of type 2 diabetes and DKD[\u003cspan citationid=\"CR18\" class=\"CitationRef\"\u003e18\u003c/span\u003e] and found that 2 novel biomarkers, PI C18:0/22:6 and SM dC18:0/20:2, could be used to distinguish healthy individuals, type 2 diabetes cases and DKD cases from each other, but these two biomarkers cannot distinguish the stages of DKD.\u003c/p\u003e \u003cp\u003eA CKD273 classifier was successful in using urine peptides to predict the development of DKD before patients developed microalbuminuria[\u003cspan citationid=\"CR19\" class=\"CitationRef\"\u003e19\u003c/span\u003e]. A work by Liao, W.-L. et al identified haptoglobin (HPT) and α-1-microglobulin/bikunin precursor (AMBP) as two biomarkers with the highest ability to distinguish between healthy individuals and patients with nephropathy, and between diabetic patients with or without DKD[\u003cspan citationid=\"CR20\" class=\"CitationRef\"\u003e20\u003c/span\u003e] in a Taiwanese population. However, the study population was limited to patients with an ACR of less than 300 mg/g, and none of the patients had stage 3 of CKD. In addition, the information on the use of insulin or other drugs (such as renin-angiotensin system antagonists) that may alter renal function was not provided.\u003c/p\u003e \u003cp\u003eCurrently, clinical DKD surveillance relies on the measurements of eGFR and UACR, along with other physical and clinical parameters. The accuracy of these tests is not ideal and often depends on the results of multiple tests[\u003cspan citationid=\"CR21\" class=\"CitationRef\"\u003e21\u003c/span\u003e]. Here, we report a streamlined urine proteomics workflow to monitor DKD at different stages using liquid chromatography coupled with tandem mass spectrometry (LC-MS/MS). Urine proteomes of diabetes without complications, DKD, and other nephropathy patients without diabetes were measured and used to build logistic regression models to differentiate DKD from uncomplicated diabetes and different stages of DKD.\u003c/p\u003e"},{"header":"Materials And Methods","content":"\u003ch2\u003eSample Collection\u003c/h2\u003e\n\u003cp\u003eA total of 239 urine samples from 236 patients, including diabetes without nephropathy (n = 134), diabetic kidney disease (n = 65), and nephropathy without diabetes (n = 40), were collected in Guangdong Provincial Hospital of Chinese Medicine, Guangzhou, China. The batch 1 dataset of 30 urine samples was collected in year 2018, including 21 cases of diabetes without nephropathy and 9 cases of DKD; the batch 2 dataset of 209 urine samples was collected in year 2019, including diabetes without nephropathy (n = 113), DKD (n = 56) and nephropathy without diabetes (n = 40). DKDs were separated into two stages as DKD3 (Urinary Albumin/Creatinine Ratio: 30-300 mg/g) and DKD4 (Urinary Albumin/Creatinine Ratio \u0026gt;300 mg/g).\u003c/p\u003e\n\u003ch2\u003eSample Preparation and Nano HPLC-MS Analysis\u003c/h2\u003e\n\u003cp\u003eMidstream of the first-morning urine was obtained and stored at -80℃. One milliliter of all urine samples was centrifuged at 176,000g for 1 h and the pellet was collected. The pellet was resuspended with 160 \u0026micro;l of resuspension buffer (100 mM Tris, pH 7.4) with 50 mM dithiothreitol (DTT). The suspension was then heated at 65\u0026deg;C for 30 min and centrifuged at 176,000 g for 0.5 h. The pellet was resuspended with 30 \u0026micro;l NH\u003csub\u003e4\u003c/sub\u003eHCO\u003csub\u003e3,\u003c/sub\u003e heated at 95℃ for 3 min and cooled to room temperature, then digested by trypsin or 12 h.\u003c/p\u003e\n\u003cp\u003eThe digested peptides were vacuum dried and re-dissolved in 0.1% formic acid and resolved on an UltiMate 3000 RSLCnano System (Thermo Fisher Scientific) operating on a 20 minutes linear gradient for batch 1 and 30 minutes for batch 2 (5-35% acetonitrile in 0.1% formic acid) at a flow rate of 600 nl/min. Tandem mass spectra were acquired on a QExactive HF mass spectrometer (Thermo Fisher Scientific) in the data-dependent mode. Quality control samples were made prepared from trypsin digests of 293T cells and were routinely analyzed to assess the LC-MS/MS sensitivity and reproducibility.\u003c/p\u003e\n\u003ch2\u003eProtein Identification and Label-free Quantification\u003c/h2\u003e\n\u003cp\u003eMS data were processed on the Firmiana platform[\u003cspan class=\"CitationRef\"\u003e22\u003c/span\u003e]. Proteins were identified against the NCBI human RefSeq protein database (released on 04/07/2013, 32,015 entries) using the MASCOT search engine (Matrix Science, version 2.3.01). Mass tolerance was set as 20 ppm for precursor ions and 0.05 Da for product-ions, respectively. Up to one missed cleavage was allowed for trypsin digestion. Cysteine carbamidomethylation was considered as a fixed modification, N-terminal acetylation and methionine oxidation were considered as dynamic modifications. 1% FDR on both the peptide and protein levels estimated by searching a decoy database were allowed. Only identifications with \u0026ge;1 unique and strict peptides and \u0026ge;2 strict peptides (ion score \u0026gt; 20) or \u0026ge;3 strict peptides, which was comparable to 1 % FDR at the protein level, were used for subsequent analyses. For protein quantification, intensity based absolute quantification (iBAQ) algorithm[\u003cspan class=\"CitationRef\"\u003e23\u003c/span\u003e] was used. To normalize the differences in sample amounts, iBAQ values were converted to iFOTs (fraction of total) calculated by dividing the iBAQ value of each protein by the total iBAQ of the sample followed by multiplying 10\u003csup\u003e5\u003c/sup\u003e for easy visualization. All missing values were substituted with zero.\u003c/p\u003e\n\u003cdiv class=\"Section2\" id=\"Sec3\"\u003e\n \u003ch2\u003eStatistical Analysis\u003c/h2\u003e\n \u003cp\u003ePathway enrichment analysis was performed using reactome (\u003cspan class=\"ExternalRef\"\u003e\u003cspan class=\"RefSource\"\u003ehttps://reactome.org/\u003c/span\u003e\u003c/span\u003e)[\u003cspan class=\"CitationRef\"\u003e24\u003c/span\u003e]. Protein differential excretion was assessed with Kruskal-Wallis test. The correlation analyses between clinical and proteomics data and the statistical analyses were calculated by scipy (version 1.3.0) package with Python. For high-dimensional MS data, dimension reduction was applied to avoid over-fitting. We first defined the differentially expressed proteins as the candidate biomarker set, then go through the potentially combinations of less than or equal to 4 proteins. At last, a logistic regression model was carried out to select the optimal panel of biomarker.\u003c/p\u003e\n \u003cp\u003eA Logistic Regression Classifier (sklearn (0.21.2) package with Python) was built using batch1 data as training set and batch2 data as validation set. The performance of the model was evaluated by sensitivity and specificity computed based on the confusion matrix. The Receiver Operating Characteristic curve plotted sensitivity (True Positive Ratio) as the x axis and 1-specificity (False Positive Ratio) as the y axis. Feature selection was applied to select most important features that can make the model achieve a higher AUC (area under ROC curve). Principal component analysis (PCA) was performed by the sklearn (0.21.2) package with Python on validation dataset. Two components were used for data visualization. Risk of progression to DKD3 was assessed by a risk score, representing the probability of being DKD3 by Model 3.\u003c/p\u003e\n\u003c/div\u003e"},{"header":"Results","content":"\u003cdiv id=\"Sec5\" class=\"Section2\"\u003e \u003ch2\u003eUrine Proteomics of uncomplicated Diabetes without nephropathy, DKD, and CKD without diabetes\u003c/h2\u003e \u003cp\u003eWe employed a previously published procedure for measuring urine proteomes[\u003cspan citationid=\"CR25\" class=\"CitationRef\"\u003e25\u003c/span\u003e] with minor modifications. This streamlined workflow achieved high efficiency and batch-to-batch reproducibility. The stability and reproducibility of the MS platform was ensured by running 293T cell protein extracts as quality control (QC) samples during the data-collecting period. The median of the Pearson correlation coefficienct between QC samples was 0.93 (Fig. s1a). With this procedure, we measured 239 urine samples from 236 patients, including uncomplicated diabetes without nephropathy (diabetes herein for simplicity, 134 samples from 132 patients), DKD (65 samples from 64 patients), and chronic kidney disease without diabetes (CKD herein for simplicity, 40 samples from 40 patients) (Fig.\u0026nbsp;1a, Table s1). A total of 2,946 proteins were identified and quantified with high analytical confidence (Fig. s1b, Table s1) with a median of 571 proteins identified per urine sample (Fig. s1c).\u003c/p\u003e \u003cp\u003ePrinciple component analysis (PCA) of all proteins in the dataset showed that, when all samples were projected onto the first two principal components, representing 79.6% and 13.5% of the variance respectively (Fig.\u0026nbsp;1b), diabetes and CKD can be clearly separated, whereas DKD appears to be in between. Next, we selected differentially excreted proteins (DEPs) in each group, using the following criteria: 1) the ratio of the mean of the highest group / second highest group \u0026gt;= 2, and 2) the p-value of Kruskal-Wallis test \u0026lt;0.05 (Table s2). The 509 DEPs whose expressions were at least two-fold higher in one disease than in any other disease were defined as disease-specific DEPs (323 DEPs for diabetes, 98 DEPs for DKD, 88 DEPs for CKD). Reactome analysis showed that proteins specific in diabetes were enriched in late endosomal microautophagy, regulation of IGF transport and uptake by IGFBPs. In contrast, innate immune system and platelet activation were significantly up-regulated in both DKD and CKD, indicating that the two diseases share similar pathology. Moreover, enrichment of these pathways was higher in statistical significance between CKD and diabetes than those between DKD and diabetes. Notably, the formation of fibrin clot (clotting cascade) was up-regulated in CKD, but not in DKD (Fig.\u0026nbsp;1c).\u003c/p\u003e \u003cp\u003eTo investigate whether the urinary proteome correlated with the clinical parameters assessed by serum tests, we performed correlation analysis for the 2,946 proteins and 13 routinely tested clinical or health indexes. We calculated the Pearson correlation coefficient for each protein with each index (Fig.\u0026nbsp;1d, Table s2). Strong correlations (Pearson correlation coefficient \u0026gt; 0.5 or \u0026lt; -0.5) were found between 4 kidney functional indexes (serum creatinine, albumin-to-creatinine ratio, blood urea nitrogen, and glomerular filtration rate) and 46 urinary proteins. For example, serum creatinine was strongly correlated with the urinary protein abundance of RBP4 (retinol binding protein 4) and C7 (complement C7), with Pearson correlation coefficient of 0.812 and 0.769, respectively. C7 abundance was also correlated with blood urea nitrogen (Pearson correlation coefficient: 0.729) (Fig.\u0026nbsp;1e). Consistently, 3 proteins (ALB, RBP4, C7) were negative correlated with glomerular filtration rate (Fig. s1d). These results indicate that serum clinical indexes may be reflected in the urine proteome, which could provide additional biological insights into kidney dysfunctions in these diseases.\u003c/p\u003e \u003c/div\u003e \u003cdiv id=\"Sec6\" class=\"Section2\"\u003e \u003ch2\u003eDistinguish DKD from Diabetes with a 2-protein Classifier\u003c/h2\u003e \u003cp\u003eWe used the dataset collected from year 2018 as the discovery dataset and the dataset collected from year 2019 as the validation dataset (Fig.\u0026nbsp;1a). To identify urinary biomarkers that distinguish DKD from diabetes, we determined DEPs between diabetes and the DKD group (including both DKD3 and DKD4 samples) (Fig.\u0026nbsp;2a). From the discovery dataset, which contained 21 diabetes and 9 DKD, we found 177 DEPs (Mann\u0026ndash;Whitney U test p values \u0026lt; 0.05, fold-change of means \u0026gt;3), among which 79 were up-regulated and 73 were down-regulated in DKD patients (Fig.\u0026nbsp;2b) Next, we selected high-abundant DEPs (mean iFOT \u0026gt; 1) in diabetes or DKD as potential biomarkers, resulting in 135 proteins (Table s3). After dimension reduction, we showed that using 2 proteins (ALB, AFM) could readily distinguish DKD from diabetes (AUC: 0.928, accuracy: 85.2%, specificity: 83.2%, sensitivity: 87.5%, Fig.\u0026nbsp;2c) in the validation set. Both ALB and AFM were significantly excreted at higher levels in the validation dataset (Fig.\u0026nbsp;2d).\u003c/p\u003e \u003cp\u003eReactome analysis revealed that the up-regulated DEPs were enriched in neutrophil degranulation, adaptive immune system, and complement cascade, while the down-regulated DEPs were enriched in membrane trafficking and cellular responses to stress (Fig.\u0026nbsp;2e). These results indicated the major functional differences between DKD and diabetics as revealed by the urine proteome. Close examination of the high abundance DEPs (top 50) revealed that liver- (29/50) and bone marrow (6/50)-specific proteins[\u003cspan citationid=\"CR26\" class=\"CitationRef\"\u003e26\u003c/span\u003e] were enriched in DKD, but low tissue specificity were found among those highly expressed in diabetes (Table s3).\u003c/p\u003e \u003c/div\u003e \u003cdiv id=\"Sec7\" class=\"Section2\"\u003e \u003ch2\u003eDistinguish DKD3 from DKD4 with a 3-protein classifier\u003c/h2\u003e \u003cp\u003eAs DKD3 and DKD4 are managed differently in the clinics[\u003cspan citationid=\"CR27\" class=\"CitationRef\"\u003e27\u003c/span\u003e], a non-invasive test to distinguish the DKD stages could aid in disease management. To this end, we first used the DKD3 (n=3) and DKD4 (n=6) samples in the discovery set to find DEPs (Fig.\u0026nbsp;3a). Since the difference between DKD3 and DKD4 might be smaller than that between diabetes and DKD, we applied a more relaxed constraints for protein filtering (Mann\u0026ndash;Whitney U test p \u0026lt; 0.05 or fold-change \u0026gt; 3 and detection frequency \u0026gt; 50%). We identified 104 DEPs with 45 up-regulated and 59 down-regulated proteins in DKD4 compared to DKD3 (Table s4). An unsupervised hierarchical clustering analysis of the validation set using the 104 DEPs showed that DKD3 and DKD4 can be well separated (Fig.\u0026nbsp;3b), suggesting that significant differences have occurred when DKD3 progressed to DKD4. Using the same protein filtering criteria in the validation set, 32 up-regulated and 468 down-regulated proteins were found in DKD4 compared to DKD3 (Table s4). A Venn diagram showed that 7 of the 45 (24.4%) up-regulated DEPs in the discovery set and 30 of the 59 (50.8%) down-regulated DEPs were also found in the validation set (Fig.\u0026nbsp;3c). After performing a dimension reduction, a 3-protein classifier could distinguish DKD3 from DKD4. Applying the prediction model to the validation set resulted in an ROC of 0.949 (accuracy: 85.7%, specificity: 86.8%, sensitivity: 83.3%, Fig.\u0026nbsp;3d). Reactome analysis revealed that the up-regulated DEPs in DKD3, including all 7 overlapping ones, were enriched in hemostasis and immune system, especially in complement cascade (Fig.\u0026nbsp;3e-f).\u003c/p\u003e \u003c/div\u003e \u003cdiv id=\"Sec8\" class=\"Section2\"\u003e \u003ch2\u003eMonitor the transition to DKD with a 4-protein classifier\u003c/h2\u003e \u003cp\u003eSince our goal is early diagnosis for DKD, we re-analyzed DKD3 and diabetes datasets to identify proteins that could implicate the transition from diabetes to DKD (Fig.\u0026nbsp;4a). As small sample size would result in poor estimation and low statistical power, in the discovery set, we performed DEP analysis between diabetes group (n=21) and the DKD group (n=3) with the following 2 criteria: 1) the mean iFOT in DKD group was 2 times higher than that in diabetes group; 2) detected in at least 2 samples in the DKD group, resulting in 203 DEPs. Next, we identified DEPs (Mann\u0026ndash;Whitney U test p \u0026lt; 0.05, fold change \u0026gt; 2 and detection frequency \u0026gt; 50%) in the validation dataset, resulting in 73 DEPs. After performing the dimension reduction from the 24 overlapping DEPs from discovery and validation datasets, a 4-protein classifier (SERPINA5, VPS4A, CP, TF) could distinguish diabetes from DKD3 (Fig.\u0026nbsp;4b). Applying this classifier to the validation dataset resulted in an area under the ROC curve of 0.952 (accuracy: 89.4%, specificity: 88.5%, sensitivity: 92.1%, Fig.\u0026nbsp;4c).\u003c/p\u003e \u003cp\u003eWe noted that 13 of the 113 diabetic samples were classified as DKD3 (Fig.\u0026nbsp;4d). This seemingly false-positive result may indicate the inaccuracy of the classifier; alternatively, it could also indicate clinical misdiagnosis, as progression from diabetes to early stage of DKD (pre-DKD3) is difficult to detect. We then explored the possibility of inferring the uncomplicated diabetes to \u0026ldquo;pre-DKD3\u0026rdquo; progression based on the difference between diabetes and these 13 false-positive DKD3 samples. To this end, we identified DEPs between the 13 false-positive DKD3 and the rest of the diabetes samples in the validation set. Using the criteria described (Wilcoxon p value \u0026lt; 0.05, ratio of means\u0026gt;=2 or \u0026lt;= 0.5, and frequency \u0026gt; 0.6), we found 35 DEPs (16 upregulated and 19 downregulated, Table s6), which include all 4 classifier proteins, as expected. As shown in Fig.\u0026nbsp;4e, the levels of the 4 classifier proteins, as well as several other DEPs also showed significant changes at other stages of the disease, and the changes were more significant as the severity of the disease progress. Together these data indicate that the dysfunctional pathways in different stages of the diseases share similarity and DKD is a progressive disease.\u003c/p\u003e \u003cp\u003eTo test the predictive power of our model, we analyzed longitudinal urine samples collected from 11 diabetic patients, among which 6 were predicted as the high-risk \u0026ldquo;pre-DKD\u0026rdquo; patients. We assess the progression to DKD by calculating the risk scores for each urine sample using model 3. Four of the 6 pre-DKD patients were classified as DKD3 once again(Fig. s2). Importantly, one of the patients (clinical ID: 0359150) was diagnosed as DKD3 by clinical index (Fig.\u0026nbsp;4f, Table s7). Moreover, one patient (clinical ID: 8073510), who was classified as diabetes in the validation dataset and re-classified as DKD3 in 2020, was admitted to hospital in March 2021 for DKD treatment, and then recovered in July 2021 after Valsartan Capsules treatment. These results suggest urine proteomics has the potential to monitor DKD progression, especially for high-risk diabetes patients.\u003c/p\u003e \u003c/div\u003e"},{"header":"Discussion","content":"\u003cp\u003eIn this study, we conducted urinary proteomic studies on 236 patients. This streamlined, sedimentation-based method allowed us to detect 559 proteins/sample on average and a total of 2,946 proteins, requiring 0.5 hour/sample of processing time. The data allowed us to build models that could distinguish DKD from diabetes (AUC = 0.928) with 2 proteins (ALB, AFM), and distinguish DKD3 from DKD4 (AUC = 0.949) with 3 proteins (ANXA7, APOD, C9). Moreover, a 4-protein (SERPINA5, VPS4A, CP, TF) classifier was built to predict early stage DKD3 from diabetes (Fig.\u0026nbsp;4g). Importantly, a small-scale follow-up studies on 11 patients showed that 2 high-risk pre-DKD patients predicted by our model had progressed to clinical DKD3, confirming the ability of our approach as a noninvasive method to monitor progression of renal dysfunctions, particularly for patients who have been classified as high-risk.\u003c/p\u003e \u003cp\u003ePrevious studies showed that inactivation of similar biological processes may contribute to the disease progression. Samples in our dataset allowed us to analyze the different stages of the disease in a quantitative manner. We found that similar biological processes are indeed altered at different stages. Importantly, activation of complement cascades is unique for DKD4 and CKD. Our results also suggested that CKD caused by other renal dysfunction can be readily differentiated from DKD, and for the most part, appeared to have caused more severe injury to the kidney. However, these results should be interpreted with caution as the CKD cases in our dataset have complex etiology. Larger sample size and more careful data analysis could reveal the underlying mechanism of the changes.\u003c/p\u003e \u003cp\u003eAbnormal urine albumin excretion is a hallmark for incipient nephropathy. This has been attributed to dysfunction in glomerular permselectivity and tubular reabsorption, among other causes. Our data showed that, besides albumin, which is the 2nd most abundant urinary excreted proteins identified in our assay, other small serum proteins also showed an increase in excretion, consistent with the notion of renal dysfunction as an underlying cause. Recently, Van and colleague reviewed published DKD urinary proteomic/peptidomic literatures, and compiled a list of 75 most robust candidate markers at each stage of diabetic kidney disease and highlighted their roles in biological processes that may contribute to progression[\u003cspan citationid=\"CR28\" class=\"CitationRef\"\u003e28\u003c/span\u003e]. Of the 75 biomarkers, 55 (73.3%) can be detected in our dataset and 23 of them (41.8%) were among the top 100 abundant proteins in our dataset (Table s1). Interestingly, many small serum proteins, including SERPINA1, TF, AMBP, APOA1, that showed increased excretion were expressed specifically or enriched in liver and bone marrow. On the other hand, proteins that showed decreased excretion, such as AHSG, CUBN, MASP2, showed low tissue specificity and were highly expressed in more diverse tissues. Notably, expressions of several decreased proteins have been reported in the literatures or Human Protein Atlas as enriched or specific in the tubular segment and play important functional roles in endocytosis.\u003c/p\u003e \u003cp\u003eRecent technical advancement in mass spectrometry empowered high throughput unbiased biomarker discoveries. For example, a urinary peptide-based classifier consisting of 273 naturally occurring urinary peptides was obtained from 3,600 individuals analyzed by capillary electrophoresis coupled to MS. The so-called \u0026ldquo;CKD273\u0026rdquo; first proposed by Good et al[\u003cspan citationid=\"CR29\" class=\"CitationRef\"\u003e29\u003c/span\u003e] can differentiate patients of chronical kidney diseases with different etiology including those caused by diabetes. Recently, the CKD273 classifier was applied in a trial for early detection of DKD (PRIORITY). This multicenter, prospective, observational study showed that a high-risk CKD273 score was associated with an increased risk of progression to microalbuminuria over a median of 2.5 years, independent of clinical characteristics. However, since the CE-MS detects naturally occurring urinary peptides without enrichment, the variety of proteins detected is often limited. In fact, the CKD273 peptides were derived from a total of 30 proteins and the collagen fragments constitute more than 66% of the 273 peptides. This low protein diversity provides limited biological insights into disease etiology and offered few treatment recommendations. Alternatively, analysis of urinary proteins, particularly with the enrichment of excreted vesicles, allows the detection and quantification of thousands of proteins from 5 ml of a urine sample, and the streamlined workflow also enables improved quantitative measurement.\u003c/p\u003e \u003cp\u003eInterestingly, collagens were not detected in high abundance as excreted proteins in our dataset. This is in contrast to results from urine peptidomics analyses represented by the CKD 273 panel, as Collagen a-1 (I) (126 fragments) and (III) (55 fragments) chains constitute \u0026gt;66% of the 273 markers. Since the CKD273 detects naturally occurring peptides generated by endogenous proteases, these small proteins or peptides were likely lost in our sedimentation process. Together these results suggest that proteomics and peptidomics analyses are complementary assays that could provide more accurate diagnosis with biological insights.\u003c/p\u003e \u003cp\u003eIn summary, the urinary proteomic cross-sectional study allowed the differentiation of uncomplicated diabetes and DKDs at different stages, and the identification of biomarkers. Disease progression of the high-risk patients identified by the pre-DKD classifier was partly validated in a small-scale follow-up study. These results warrant the the design of perspective studies to test the predictive power of our model for monitoring high-risk patients.\u003c/p\u003e"},{"header":"Conclusions","content":"\u003cp\u003eIn this study, we have established a urinary proteomic workflow to conduct a cross-sectional investigation of uncomplicated diabetes, DKD, and CKD patients. We identified biomarkers that could distinguish DKD from uncomplicated diabetes, and stage 4 DKD from stage 3 DKD. Logistic regression models were built upon the discovery dataset, and the classification were validated in an independent dataset. Bioinformatics analysis suggests that (upregulation of DKD4) complement cascade is an indication of DKD progression. Moreover, an algorithm was established to identify earlier stage DKD patients. Follow-up studies on 11 patients indicated that 2 putative pre-DKD patients have progressed to DKD3.\u003c/p\u003e \u003cp\u003eThis study demonstrated the potential for urine proteomics as a noninvasive method for DKD diagnosis and identifying high-risk patients for progression monitoring. The pathways identified from differentially excreted proteins provides important clues on biological basis during disease progression.\u003c/p\u003e"},{"header":"Declarations","content":"\u003cp\u003e\u003cstrong\u003eAcknowledgments\u003c/strong\u003e\u003c/p\u003e\n\u003cp\u003eThe authors thank all the colleagues involved in this project, including clinicians and nurses from Guangdong Provincial Hospital of Chinese Medicine (Guangzhou, China), students and engineers from National Center for Protein Sciences (Beijing, China). The authors acknowledge the support of Guangdong Provincial Hospital of Chinese Medicine (Guangzhou, China) and National Center for Protein Sciences (Beijing, China) for providing study participants and case samples and laboratory for this study.\u003c/p\u003e\n\u003cp\u003e\u003cstrong\u003eAuthor\u0026rsquo;s C\u003c/strong\u003e\u003cstrong\u003eontribut\u003c/strong\u003e\u003cstrong\u003eors\u003c/strong\u003e\u003c/p\u003e\n\u003cp\u003eG.F., Y.W., and J.Q. conceived and designed the study. T.G., X.N. and J.W. had primary responsibility for writing of the manuscript with the support of Y.W. and J.Q. Y.L., H.W., Z.L., L.Z., H.H., Y.W., X.T. and J.Q. were responsible for study implementation and administration. X.Y., H.L., and X.L. performed the experiments; T.G., X.N., J.W., L.S. and L.L. contributed to urine sample testing and bioinformatics analysis with the guidance of Y.W. and J.Q.. All other authors collected urine sample from subjects. G.F. and J.Q. are the guarantors of this work and, as such, had full access to all the data in the study and takes responsibility for the integrity of the data and the accuracy of the data analysis.\u003c/p\u003e\n\u003cp\u003e\u003cstrong\u003eFunding\u003c/strong\u003e\u003c/p\u003e\n\u003cp\u003eThis work was supported by grants from the Guangdong Provincial Hospital of Chinese Medicine (2018DB03, YN2021DB02), National Key Research and Development Program of China (2017YFC0908404, 2018YFA0507503, 2017YFA0505102), State Key Laboratory of Proteomics (No. SKLP-K201705).\u003c/p\u003e\n\u003cp\u003e\u003cstrong\u003eAvailability of d\u003c/strong\u003e\u003cstrong\u003eata and materials\u003c/strong\u003e\u003c/p\u003e\n\u003cp\u003eThe MS raw data generated in this study have been submitted to ProteomeXchange database (www.proteomexchange.org) via the iProX partner repository[30] under accession number IPX0002858000.\u003c/p\u003e\n\u003cp\u003e\u003cstrong\u003eEthics approval and consent to participate\u003c/strong\u003e\u003c/p\u003e\n\u003cp\u003eThe study was approved by the medical ethics committee of the\u0026nbsp;Guangdong Provincial Hospital of Chinese Medicine.\u003c/p\u003e\n\u003cp\u003e\u003cstrong\u003eConsent for publication\u003c/strong\u003e\u003c/p\u003e\n\u003cp\u003eAll authors consent to the publication of this manuscript.\u003c/p\u003e\n\u003cp\u003e\u003cstrong\u003eCompeting interests\u003c/strong\u003e\u003c/p\u003e\n\u003cp\u003eDr. Jun Qin is the cofounder and co-owner of the Beijing Pineal Health Management Co., Ltd.. Dr. Xiaotian Ni, Tongqing Gong., Xing Yang., Haibo Liu., Xinliang Li., Lifeng Liu, are employees of Beijing Pineal Health Management Co., Ltd.. All other authors declare no competing interests.\u003c/p\u003e\n\u003cp\u003e\u003cstrong\u003eAuthors\u0026apos; information\u003c/strong\u003e\u003c/p\u003e\n\u003cp\u003e\u003csup\u003e1\u003c/sup\u003eThe Second Affiliated Hospital of Guangzhou University of Chinese Medicine, Guangzhou, 510120 China\u0026nbsp;\u003c/p\u003e\n\u003cp\u003e\u003csup\u003e2\u003c/sup\u003eThe Second Clinical College of Guangzhou University of Chinese Medicine, Guangzhou, 510120 China\u0026nbsp;\u003c/p\u003e\n\u003cp\u003e\u003csup\u003e3\u003c/sup\u003eGuangdong Provincial Hospital of Chinese Medicine, Guangzhou, 510120 China\u0026nbsp;\u003c/p\u003e\n\u003cp\u003e\u003csup\u003e4\u003c/sup\u003eGuangdong Provincial Academy of Chinese Medical Sciences, Guangzhou, 510120 China\u003c/p\u003e\n\u003cp\u003e\u003csup\u003e5\u003c/sup\u003eBeijing Pineal Health Management Co., Ltd., Beijing 102206, China\u003c/p\u003e\n\u003cp\u003e\u003csup\u003e6\u003c/sup\u003eState Key Laboratory of Proteomics, Beijing Proteome Research Center, National Center for Protein Sciences, Institute of Lifeomics, Beijing, 102206, China\u003c/p\u003e\n\u003cp\u003e\u003csup\u003e7\u003c/sup\u003eChongqing Key Laboratory of Big Data for Bio Intelligence, School of Bioinformation, Chongqing University of Posts and Telecommunications, Chongqing, 400065 China\u003c/p\u003e"},{"header":"References","content":"\u003col\u003e\n\u003cli\u003eSaeedi P, Petersohn I, Salpea P, Malanda B, Karuranga S, Unwin N, et al. Global and regional diabetes prevalence estimates for 2019 and projections for 2030 and 2045: Results from the International Diabetes Federation Diabetes Atlas, 9th edition. Diabetes Res Clin Pract 2019;157:107843. https://doi.org/10.1016/j.diabres.2019.107843.\u003c/li\u003e\n\u003cli\u003eGheith O, Farouk N, Nampoory N, Halim MA, Al-Otaibi T. Diabetic kidney disease: world wide difference of prevalence and risk factors. J Nephropharmacol 2016;5:49\u0026ndash;56.\u003c/li\u003e\n\u003cli\u003eMaahs DM, Rewers M. Editorial: Mortality and renal disease in type 1 diabetes mellitus--progress made, more to be done. J Clin Endocrinol Metab 2006;91:3757\u0026ndash;9. https://doi.org/10.1210/jc.2006-1730.\u003c/li\u003e\n\u003cli\u003eOrchard TJ, Secrest AM, Miller RG, Costacou T. In the absence of renal disease, 20 year mortality risk in type 1 diabetes is comparable to that of the general population: a report from the Pittsburgh Epidemiology of Diabetes Complications Study. Diabetologia 2010;53:2312\u0026ndash;9. https://doi.org/10.1007/s00125-010-1860-3.\u003c/li\u003e\n\u003cli\u003eCollins AJ, Foley RN, Chavers B, Gilbertson D, Herzog C, Johansen K, et al. \u0026rsquo;United States Renal Data System 2011 Annual Data Report: Atlas of chronic kidney disease \u0026amp; end-stage renal disease in the United States. Am J Kidney Dis 2012;59:A7, e1-420. https://doi.org/10.1053/j.ajkd.2011.11.015.\u003c/li\u003e\n\u003cli\u003eMacisaac RJ, Ekinci EI, Jerums G. Markers of and risk factors for the development and progression of diabetic kidney disease. Am J Kidney Dis 2014;63:S39-62. https://doi.org/10.1053/j.ajkd.2013.10.048.\u003c/li\u003e\n\u003cli\u003eTuttle KR, Bakris GL, Bilous RW, Chiang JL, de Boer IH, Goldstein-Fuchs J, et al. Diabetic kidney disease: a report from an ADA Consensus Conference. Diabetes Care 2014;37:2864\u0026ndash;83. https://doi.org/10.2337/dc14-1296.\u003c/li\u003e\n\u003cli\u003eMischak H. Pro: urine proteomics as a liquid kidney biopsy: no more kidney punctures! Nephrol Dial Transplant 2015;30:532\u0026ndash;7. https://doi.org/10.1093/ndt/gfv046.\u003c/li\u003e\n\u003cli\u003eVinik AI, Nevoret M-L, Casellini C, Parson H. Diabetic neuropathy. Endocrinol Metab Clin North Am 2013;42:747\u0026ndash;87. https://doi.org/10.1016/j.ecl.2013.06.001.\u003c/li\u003e\n\u003cli\u003eNational Kidney Foundation. KDOQI Clinical Practice Guideline for Diabetes and CKD: 2012 Update. Am J Kidney Dis 2012;60:850\u0026ndash;86. https://doi.org/10.1053/j.ajkd.2012.07.005.\u003c/li\u003e\n\u003cli\u003eChen Y, Lee K, Ni Z, He JC. Diabetic Kidney Disease: Challenges, Advances, and Opportunities. Kidney Dis (Basel) 2020;6:215\u0026ndash;25. https://doi.org/10.1159/000506634.\u003c/li\u003e\n\u003cli\u003eSmart NA, Dieberg G, Ladhani M, Titus T. Early referral to specialist nephrology services for preventing the progression to end-stage kidney disease. Cochrane Database Syst Rev 2014:CD007333. https://doi.org/10.1002/14651858.CD007333.pub2.\u003c/li\u003e\n\u003cli\u003eMullen W, Delles C, Mischak H, EuroKUP COST action. Urinary proteomics in the assessment of chronic kidney disease. Curr Opin Nephrol Hypertens 2011;20:654\u0026ndash;61. https://doi.org/10.1097/MNH.0b013e32834b7ffa.\u003c/li\u003e\n\u003cli\u003eLi J, Zhang Z, Rosenzweig J, Wang YY, Chan DW. Proteomics and bioinformatics approaches for identification of serum biomarkers to detect breast cancer. Clin Chem 2002;48:1296\u0026ndash;304.\u003c/li\u003e\n\u003cli\u003eFliser D, Novak J, Thongboonkerd V, Argil\u0026eacute;s A, Jankowski V, Girolami MA, et al. Advances in urinary proteome analysis and biomarker discovery. J Am Soc Nephrol 2007;18:1057\u0026ndash;71. https://doi.org/10.1681/ASN.2006090956.\u003c/li\u003e\n\u003cli\u003eGao Y. Urine is a better biomarker source than blood especially for kidney diseases. Adv Exp Med Biol 2015;845:3\u0026ndash;12. https://doi.org/10.1007/978-94-017-9523-4_1.\u003c/li\u003e\n\u003cli\u003eRossing K, Mischak H, Dakna M, Z\u0026uuml;rbig P, Novak J, Julian BA, et al. Urinary proteomics in diabetes and CKD. J Am Soc Nephrol 2008;19:1283\u0026ndash;90. https://doi.org/10.1681/ASN.2007091025.\u003c/li\u003e\n\u003cli\u003eZhu C, Liang Q, Hu P, Wang Y, Luo G. Phospholipidomic identification of potential plasma biomarkers associated with type 2 diabetes mellitus and diabetic nephropathy. Talanta 2011;85:1711\u0026ndash;20. https://doi.org/10.1016/j.talanta.2011.05.036.\u003c/li\u003e\n\u003cli\u003eZ\u0026uuml;rbig P, Jerums G, Hovind P, Macisaac RJ, Mischak H, Nielsen SE, et al. Urinary proteomics for early diagnosis in diabetic nephropathy. Diabetes 2012;61:3304\u0026ndash;13. https://doi.org/10.2337/db12-0348.\u003c/li\u003e\n\u003cli\u003eLiao W-L, Chang C-T, Chen C-C, Lee W-J, Lin S-Y, Liao H-Y, et al. Urinary Proteomics for the Early Diagnosis of Diabetic Nephropathy in Taiwanese Patients. J Clin Med 2018;7. https://doi.org/10.3390/jcm7120483.\u003c/li\u003e\n\u003cli\u003eKwiendacz H, Nabrdalik K, Stomp\u0026oacute;r T, Gumprecht J. What do we know about biomarkers in diabetic kidney disease? Endokrynol Pol 2020;71:545\u0026ndash;50. https://doi.org/10.5603/EP.a2020.0077.\u003c/li\u003e\n\u003cli\u003eFeng J, Ding C, Qiu N, Ni X, Zhan D, Liu W, et al. Firmiana: towards a one-stop proteomic cloud platform for data processing and analysis. Nat Biotechnol 2017;35:409\u0026ndash;12. https://doi.org/10.1038/nbt.3825.\u003c/li\u003e\n\u003cli\u003eSchwanh\u0026auml;usser B, Busse D, Li N, Dittmar G, Schuchhardt J, Wolf J, et al. Global quantification of mammalian gene expression control. Nature 2011;473:337\u0026ndash;42. https://doi.org/10.1038/nature10098.\u003c/li\u003e\n\u003cli\u003eJassal B, Matthews L, Viteri G, Gong C, Lorente P, Fabregat A, et al. The reactome pathway knowledgebase. Nucleic Acids Res 2020;48:D498\u0026ndash;503. https://doi.org/10.1093/nar/gkz1031.\u003c/li\u003e\n\u003cli\u003eLeng W, Ni X, Sun C, Lu T, Malovannaya A, Jung SY, et al. Proof-of-Concept Workflow for Establishing Reference Intervals of Human Urine Proteome for Monitoring Physiological and Pathological Changes. EBioMedicine 2017;18:300\u0026ndash;10. https://doi.org/10.1016/j.ebiom.2017.03.028.\u003c/li\u003e\n\u003cli\u003eUhl\u0026eacute;n M, Fagerberg L, Hallstr\u0026ouml;m BM, Lindskog C, Oksvold P, Mardinoglu A, et al. Proteomics. Tissue-based map of the human proteome. Science 2015;347:1260419. https://doi.org/10.1126/science.1260419.\u003c/li\u003e\n\u003cli\u003eSelby NM, Taal MW. An updated overview of diabetic nephropathy: Diagnosis, prognosis, treatment goals and latest guidelines. Diabetes Obes Metab 2020;22 Suppl 1:3\u0026ndash;15. https://doi.org/10.1111/dom.14007.\u003c/li\u003e\n\u003cli\u003eVan JAD, Scholey JW, Konvalinka A. Insights into Diabetic Kidney Disease Using Urinary Proteomics and Bioinformatics. J Am Soc Nephrol 2017;28:1050\u0026ndash;61. https://doi.org/10.1681/ASN.2016091018.\u003c/li\u003e\n\u003cli\u003eGood DM, Z\u0026uuml;rbig P, Argil\u0026eacute;s A, Bauer HW, Behrens G, Coon JJ, et al. Naturally occurring human urinary peptides for use in diagnosis of chronic kidney disease. Mol Cell Proteomics 2010;9:2424\u0026ndash;37. https://doi.org/10.1074/mcp.M110.001917.\u003c/li\u003e\n\u003cli\u003eMa J, Chen T, Wu S, Yang C, Bai M, Shu K, et al. iProX: an integrated proteome resource. Nucleic Acids Res 2019;47:D1211\u0026ndash;7. https://doi.org/10.1093/nar/gky869.\u003c/li\u003e\n\u003c/ol\u003e"}],"fulltextSource":"","fullText":"","funders":[],"hasAdminPriorityOnWorkflow":false,"hasManuscriptDocX":true,"hasOptedInToPreprint":true,"hasPassedJournalQc":"","hasAnyPriority":false,"hideJournal":false,"highlight":"","institution":"","isAcceptedByJournal":true,"isAuthorSuppliedPdf":false,"isDeskRejected":"","isHiddenFromSearch":false,"isInQc":false,"isInWorkflow":false,"isPdf":false,"isPdfUpToDate":true,"isWithdrawnOrRetracted":false,"journal":{"display":true,"email":"
[email protected]","identity":"clinical-proteomics","isNatureJournal":false,"hasQc":true,"allowDirectSubmit":false,"externalIdentity":"clip","sideBox":"Learn more about [Clinical Proteomics](http://clinicalproteomicsjournal.biomedcentral.com/)","snPcode":"12014","submissionUrl":"https://submission.nature.com/new-submission/12014/3","title":"Clinical Proteomics","twitterHandle":"@BioMedCentral","acdcEnabled":true,"dfaEnabled":true,"editorialSystem":"em","reportingPortfolio":"BMC/SO AJ","inReviewEnabled":true,"inReviewRevisionsEnabled":true},"keywords":"Urine, Proteomics, DKD, Progression monitoring","lastPublishedDoi":"10.21203/rs.3.rs-966683/v1","lastPublishedDoiUrl":"https://doi.org/10.21203/rs.3.rs-966683/v1","license":{"name":"CC BY 4.0","url":"https://creativecommons.org/licenses/by/4.0/"},"manuscriptAbstract":"\u003ch2\u003eBackground\u003c/h2\u003e \u003cp\u003eType 2 diabetic kidney disease is the most common cause of chronic kidney diseases (CKD) and end-stage renal diseases (ESRD). Although kidney biopsy is considered as the \u0026lsquo;gold standard\u0026rsquo; for diabetic kidney disease (DKD) diagnosis, it is an invasive procedure, and the diagnosis can be influenced by sampling bias and personal judgement. It is desirable to establish a non-invasive procedure that can complement kidney biopsy in diagnosis and tracking the DKD progress.\u003c/p\u003e\u003ch2\u003eMethods\u003c/h2\u003e \u003cp\u003eIn this cross-sectional study, we collected 252 urine samples, including 134 uncomplicated diabetes, 65 DKD, 40 CKD without diabetes and 13 follow-up diabetic sample, and analyzed the urine proteomes with liquid chromatography coupled with tandem mass spectrometry (LC-MS/MS). We built logistic regression models to distinguish uncomplicated diabetes, DKD and other CKDs.\u003c/p\u003e\u003ch2\u003eResults\u003c/h2\u003e \u003cp\u003eWe quantified 559 \u0026plusmn; 202 gene products (GPs) (Mean \u0026plusmn; SD) on a single sample and 2,946 GPs in total. Based on logistic regression models, DKD patients could be differentiated from the uncomplicated diabetic patients with 2 urinary proteins (AUC = 0.928), and the stage 3 (DKD3) and stage 4 (DKD4) DKD patients with 3 urinary proteins (AUC = 0.949). These results were validated in an independent data set. Finally, a 4-protein classifier identified putative pre-DKD3 patients, who showed DKD3 proteomic features but were not diagnosed by clinical standards. Follow-up studies on 11 patients indicated that 2 putative pre-DKD patient have progressed to DKD3.\u003c/p\u003e\u003ch2\u003eConclusions\u003c/h2\u003e \u003cp\u003eOur study demonstrated the potential for urinary proteomics as a noninvasive method for DKD diagnosis and identifying high-risk patients for progression monitoring.\u003c/p\u003e","manuscriptTitle":"Urine Proteomics Identifies Biomarkers for Diabetic Kidney Disease at Different Stages","msid":"","msnumber":"","nonDraftVersions":[{"code":1,"date":"2021-10-19 14:28:11","doi":"10.21203/rs.3.rs-966683/v1","editorialEvents":[{"type":"communityComments","content":0},{"type":"reviewerAgreed","content":"","date":"2021-11-03T00:00:00+00:00","index":2,"fulltext":""},{"type":"editorInvitedReview","content":"","date":"2021-11-03T00:00:00+00:00","index":1,"fulltext":"Recommendation: Reviewer's comments unavailable due to the journal's policy.\n"},{"type":"editorInvitedReview","content":"","date":"2021-11-03T00:00:00+00:00","index":2,"fulltext":"Recommendation: Reviewer's comments unavailable due to the journal's policy.\n"},{"type":"decision","content":"Minor revision","date":"2021-11-03T00:00:00+00:00","index":"","fulltext":""},{"type":"editorInvitedReview","content":"","date":"2021-10-18T16:16:30+00:00","index":0,"fulltext":""},{"type":"reviewersInvited","content":"","date":"2021-10-18T03:11:38+00:00","index":"","fulltext":""},{"type":"reviewerAgreed","content":"","date":"2021-10-18T00:00:00+00:00","index":1,"fulltext":""},{"type":"editorAssigned","content":"","date":"2021-10-16T11:47:31+00:00","index":"","fulltext":""},{"type":"checksComplete","content":"","date":"2021-10-15T23:00:00+00:00","index":"","fulltext":""},{"type":"editorInvited","content":"","date":"2021-10-15T23:00:00+00:00","index":"","fulltext":""},{"type":"submitted","content":"Clinical Proteomics","date":"2021-10-14T08:21:43+00:00","index":"","fulltext":""}],"status":"published","journal":{"display":true,"email":"
[email protected]","identity":"clinical-proteomics","isNatureJournal":false,"hasQc":true,"allowDirectSubmit":false,"externalIdentity":"clip","sideBox":"Learn more about [Clinical Proteomics](http://clinicalproteomicsjournal.biomedcentral.com/)","snPcode":"12014","submissionUrl":"https://submission.nature.com/new-submission/12014/3","title":"Clinical Proteomics","twitterHandle":"@BioMedCentral","acdcEnabled":true,"dfaEnabled":true,"editorialSystem":"em","reportingPortfolio":"BMC/SO AJ","inReviewEnabled":true,"inReviewRevisionsEnabled":true}}],"origin":"","ownerIdentity":"cd4a4d65-530e-4ee1-a762-abdab7122f48","owner":[],"postedDate":"October 19th, 2021","published":true,"recentEditorialEvents":[],"rejectedJournal":[],"revision":"","amendment":"","status":"published-in-journal","subjectAreas":[{"id":7955056,"name":"General Biochemistry"},{"id":7955057,"name":"Molecular Biology"}],"tags":[],"updatedAt":"2021-12-29T03:48:28+00:00","versionOfRecord":{"articleIdentity":"rs-966683","link":"https://doi.org/10.1186/s12014-021-09338-6","journal":{"identity":"clinical-proteomics","isVorOnly":false,"title":"Clinical Proteomics"},"publishedOn":"2021-12-01 03:48:28","publishedOnDateReadable":"December 1st, 2021"},"versionCreatedAt":"2021-10-19 14:28:11","video":"","vorDoi":"10.1186/s12014-021-09338-6","vorDoiUrl":"https://doi.org/10.1186/s12014-021-09338-6","workflowStages":[]},"version":"v1","identity":"rs-966683","journalConfig":"researchsquare"},"__N_SSP":true},"page":"/article/[identity]/[[...version]]","query":{"redirect":"/article/rs-966683","identity":"rs-966683","version":["v1"]},"buildId":"rHA-KDH7Qsr4HCuvH75dn","isFallback":false,"isExperimentalCompile":false,"dynamicIds":[84888],"gssp":true,"scriptLoader":[]}
Text is read by the "Ask this paper" AI Q&A widget below.
Extraction quality varies by source — PMC NXML preserves structure
cleanly, OA-HTML may include some navigation residue, and OA-PDF can
have broken hyphenation. The publisher copy
(via DOI)
is the canonical version.