EventsMOL2NET'15, Conference on Molecular, Biomed., Comput. & Network Science and Engineering, 1st ed.
Published
This submission belongs to the session 03. USEDAT-01: USA-Europe Data Analysis Training Congress, Cambridge, UK-Bilbao, Spain-Miami, USA, 2015 of the event MOL2NET'15, Conference on Molecular, Biomed., Comput. & Network Science and Engineering, 1st ed.
Published date
02 Dec, 2015
Citation
Diana Maria Herrera-Ibatá, Ricardo Alfredo Orbegozo-Medina, Complex networks of anti-HIV drugs activity vs. prevalence of AIDS in US Counties using symmetry information indices , in Proceedings of MOL2NET'15, Conference on Molecular, Biomed., Comput. & Network Science and Engineering, 1st ed., 5 December–15 December 2015, MDPI: Basel, Switzerland, doi: 10.3390/MOL2NET-1-b002
Share
Email
Facebook
Twitter
LinkedIn

Complex networks of anti-HIV drugs activity vs. prevalence of AIDS in US Counties using symmetry information indices

Ricardo Alfredo Orbegozo-Medina 2
1. Department of Information and Communication Technologies, University of A Coruña UDC, 15071, A Coruña, Spain.
2. Department of Microbiology and Parasitology, University of Santiago de Compostela (USC), 15782, Santiago de Compostela, A Coruña, Spain.
Abstract

Different aspects about the epidemiology, drugs, targets, chem-bioinformatics, and systems biology methods, related to AIDS/HIV have been reviewed. Next, we developed a new model to predict complex networks of the AIDS prevalence in U.S. counties taking into consideration the Gini coefficient (income inequality) and activity/structure data of anti-HIV drugs in preclinical assays. First, we trained different Artificial Neural Networks (ANNs) using as input Markov and Symmetry information indices of social networks and of molecular graphs, respectively. We obtained the data about AIDS prevalence and Gini coefficient from the AIDSVu database of the Rollins School of Public Health at Emory University and the data about anti-HIV compounds from ChEMBL database. To train/validate the model and predict the complex network we needed to analyze 43,249 data points including values of AIDS prevalence in 2310 US counties vs. ChEMBL results for 21,582 unique drugs, 9 viral or human protein targets, 4856 protocols, and 10 possible experimental measures. The best model found was a Linear Neural Network (LNN) with Accuracy, Specificity, Sensitivity, and AUROC above 0.72-0.73 in training and external validation series. The new linear equation was shown to be useful to generate complex network maps of drug activity vs. AIDS/HIV epidemiology in U.S. at county level.

Keywords
anti-HIV drugs
Gini coefficient
AIDS prevalence
neighborhood symmetry indices
complex networks
Manuscript
Chemoinformatics Profiling of Ionic Liquids Cytotoxicity—From Machine Learning to Network-Like Similarity Graphs
Kinetic study of activated carbon synthesis from Marabou Wood