Dissertation > Excellent graduate degree dissertation topics show
Computational Methods for Drug Target Identification Based on Data Mining
Author: GongJiaZuo
Tutor: GaoDaQi
School: East China University of Science and Technology
Course: Applied Computer Technology
Keywords: Target Identification Network Pharmacology Polypharmacology Data Mining Molecular Similarity
CLC: TQ460.1
Type: PhD thesis
Year: 2013
Downloads: 82
Quote: 0
Read: Download Dissertation
Abstract
|
The process of new drug discovery and development is time-consuming, costly and risky. The average cycle for new drugs development generally reaches15years, which consumes more than800million US dollars. However, due to the low efficiency and side effects, the drug development often fails in the clinical trial phase, causing huge lost. As the origin of drug development, drug target discovery and identification is critical to the success rate of drug development. With the development of bioinformatics technology as well as fast growing of proteomics and chemical genomics data, the combination of computational chemical biology and traditional experimental techniques provides information support to drug target discovery and novel methods for drug target prediction. Aiming at drug target discovery, this thesis will construct drug target database and pesticide target database by using computational methods to perform data mining and integration on the existing drug target database, and developed several methods and tools on drug target discovery. This thesis includes the following five parts:1. Based on the existing drugs and indication for drugs under development, we constructed a drug target database called TargetBank by means of text mining and manual validation. This database included4,357records of target proteins, covering23therapy domains and more than600clinical indications. This database also included detailed function annotation for each protein record, providing useful reference for drug development researchers.2. We constructed the first database of the interaction network for pesticides and targets, named PTID. This database contained1,342records of pesticides by means of data integration, which was divided into22categories by function, and provided with annotation on environmental fate and ecotoxicology. By means of text mining,4,245records of protein targets interacted with the pesticide are collected, and an interaction network between pesticides and targets was constructed with detailed annotation on protein sequence and function. This database also provided computational tools such as similarity search and sequence alignment, offering data support on pesticide research.3. We designed and implemented a molecular solvent-accessible surface generation program based on iso-surface theory. This method adopted a rapid one dimensional recursive Gaussian filter calculation, constructed iso-surface in the discrete three dimensional spaces and obtained the triangle representation of molecular solvent-accessible surface by using iso-surface extraction technique "Marching Cubes". Central difference method was applied to calculating the normal vector of triangles for better render effect. The test result indicated that this method was both fast and accurate, and could provide drug target surface representation for structure-based target identification. This program has been integrated into the D3Pharma software package.4. We developed a random-walk-based polypharmacology method, combining the molecular similarity methods and network pharmacology concepts. This method introduced the target compound into an existing drug-target network using molecular similarity, then discovered its polypharmacology effect and predicted its new targets or side effects by bipartite and random walk based network inference algorithm. Through prediction of recent reported off-targets events, this method was proved to be capable of predicting polypharmacology effect.5. Based on a molecular similarity evaluation program SHAFTS developed by our research group, this thesis implemented a molecular similarity based target discovery and virtual screening platform ChemMapper. This platform collected the structure information for more than3.5million compounds, with400thousands of them possessing target annotation. This platform can perform target or side effect prediction against a query structure based on three dimensional molecular similarity methods; it can also perform virtual screening or scaffold hopping research on commercial compound databases. Through the precise prediction of toxicity for the existing drug Astemizole, the lead compound discovery and scaffold hopping validation for EGFR, ChemMapper could facilitate the research on genomics, target discovery, polypharmacology and virtual screening.
|
Related Dissertations
- A Study on Healthcare Product Marketing Based on Data Mining Technology,F426.72
- Gao Zhong-ying academic thought and experience and use of Bufei Decoction treatment of common diseases of the respiratory system drug law,R249.2
- Bing- thick academic thought and clinical experience and empirical studies apply to turtle soups treatment of chronic kidney disease,R249.2
- The Design and Implementation of Bicluster Data Analyzing Software,TP311.52
- Research on Clustering Algorithm Based on Mutation Particle Swarm Optimization,TP18
- Research on Fuzzy C-Mean Clustering Algorithm Based on Particle Swarm Optimization and Shuffled Frog Leaping Algorithm,TP18
- Research on Clustering Algorithm Based on Genetic Algorithm and Rough Set Theory,TP18
- Based on data mining research tax audit case selection,F812.42
- Community-oriented education, personalized learning system and its implementation,TP391.6
- Association rule mining based Intrusion Detection System Research and Implementation,TP393.08
- Data warehouse technology in the banking customer management systems research and implementation,TP315
- Investments Projects Evaluation Supporting System Based on Optimizing Model of Industrial Parameters,F283
- The Research and Application of Data Mart in the Telecommunication Business Analysis,TP311.13
- Application of Data Mining in the Analysis of Higher Vocational Colleges’ Achievement,TP311.13
- Traffic Classification Using Stream Data Mining Algorithm,TP393.06
- Research and Design of Individualized Online Teaching-aided System Based on Data Mining,TP311.13
- Design and Implementation of Course Assessment and Analysis of Decision System Based on Data Mining,TP311.13
- Design to E-learning System in Senior Vocational School Base on Moodle,TP311.52
- Design and Development of Teaching Quality Assessment System Based on Data Mining,TP311.13
- Data Mining of Application in the School Management and Training Students,TP311.13
- Application Research in Crm of Bank Based on the Improved K-means Cluster Algorithm,TP311.13
CLC: > Industrial Technology > Chemical Industry > Pharmaceutical chemical industry > General issues > Basic theory
© 2012 www.DissertationTopic.Net Mobile
|