DOI: 10.1021/acs.jcim.6c01937 ISSN: 1549-9596

DeepKa Protein p K a Database: Identifying pH Dependence in the Protein Data Bank

Baorong Yang, Xiangxiang Lu, Zhitao Cai, Jiawen Sun, Yandong Huang

Abstract

pH plays a central role in many biological events. To explore the underlying molecular mechanisms, it is of importance to determine pKa values of ionizable residues in proteins. In this work, DeepKa, a deep learning-based protein pKa predictor, was employed to create the pKa database DeepKa DB (http://computbiophys.com/DeepKa/database), which contains pKa data of over 30 million residues in about 200 thousand soluble proteins. The database has been integrated into the existing DeepKa web server, from which users can obtain the data of interest freely. In addition to the data presentation, four representative case studies were set up to demonstrate how pKa’s extracted from the database could be applied to the investigation of pH-dependent processes in proteins.