Добавил:
kiopkiopkiop18@yandex.ru t.me/Prokururor I Вовсе не секретарь, но почту проверяю Опубликованный материал нарушает ваши авторские права? Сообщите нам.
Вуз: Предмет: Файл:
Ординатура / Хирургия / Библиотека им академика М.И. Перельмана / Книга_5586_Библиотеки_им_академика_М_И_Перельмана.pdf
Скачиваний:
0
Добавлен:
31.08.2026
Размер:
35 Мб
Скачать
Introduction to Pharmaceutical Biotechnology, Volume 2 (Second Edition)
more comprehensive representation of these intricate systems [95]. Multiscale modeling is emerging as a signicant discipline within computational biomedicine. While numerous promising multiscale models are in existence and are currently under development, it is worth noting the need for a unied multiscale modeling methodology, indicating a potential area for further research and development to enhance the multiscale modeling foundation [96].
9.5.11 Computational genomics
Computational genomics is a eld that utilizes computational and statistical techniques to deduce biological information from genomic data. With the advent of high-throughput sequencing technologies, there has been an exponential increase in genomic data available, necessitating the development of computational methods to analyze this data effectively. In the eld of genomics, various essential processes and techniques are employed for data handling and analysis. Database construction and management involve creating and maintaining databases to store, organize, and manage genomic data efciently. Data retrieval systems are developed to facilitate the swift retrieval of data from these genomic databases. Sequence alignment techniques are crucial in aligning DNA, RNA, or protein sequences to pinpoint regions of similarity. Methods for detecting motifs and domains are used to locate regions of DNA or RNA that are conserved. Prediction of 3D structures is a computer approach for predicting the 3D structures of proteins and nucleic acids. Structural elements are assigned biological roles based on their 3D structures using the functional annotation of structural elements. One may create phylogenetic trees showing evolutionary links between different species using phylogenetic analysis. To understand evolutionary linkages and processes, molecular evolutionary analysis examines molecular sequences. Genome-wide association studies (GWAS) use statistical methods to nd gene variations linked to certain diseases. This informa­tion may then be used to understand the underlying biology of such conditions better. Biological consequences may be predicted with the help of predictive models generated using ML algorithms from genetic data. Deep learning techniques are used to examine large genetic datasets and discover hidden patterns and correlations [97]. Essential for controling gene expression, regulatory elements are the subject of regulatory element identication. These elements include promoters, enhancers, and silencers. Building and analyzing regulatory networks to learn more about how genes are controlled is the goal of regulatory network construction. Understanding DNA methylation and histone modications, two examples of epigenetic marks may be gained by the computational process known as epigenetic mark identication. In order to analyze the effects of gene regulation on cellular processes, it is necessary to combine genomic and transcriptomic data with epigenetic data [98].
9.5.12 Functional genomics
The goal of functional genomics is to determine how genetics affects behavior. Methods from several disciplines are used to show the complete process of how genes function in living organisms. When we talk about sequencing, we mean doing
9-23
Introduction to Pharmaceutical Biotechnology, Volume 2 (Second Edition)
a whole-genome sequence from scratch [99]. In order to nd variants in a genome that has previously been sequenced, genome resequencing is used. Pan-genomic analyses take into account all of a species or a set of related speciesgenomic information, which helps to understand the functional components and evolutionary dynamics of the genomes [99]. The RNA-seq and metabolomics analyses cover the RNA-seq which is utilized for transcriptome proling to determine the continuously altering cellular transcriptome. Metabolomics provides a more comprehensive picture of the organisms metabolic prole by analyzing all of the substrates and metabolites in a biological sample. Uncovering the molecular pathways that underlie different biological activities requires a thorough knowledge of novel genes, including their roles and interactions [99, 100].
9.5.13 Comparative genomics
Genomes from various species are analyzed and compared as part of comparative genomics. Comparative genomics uses principles of structure, function, and genomic evolution. Understanding gene function and evolution is aided by gene family analyses, which look at the evolution and function of gene families across many species. Genomic data from different plant species may be compared using plant evolutionary analyses, which can reveal hidden connections and histories between them [101]. Gene evolution and function may be better understood when orthologs and paralogs are identied and compared across species and within genomes, respectively. Understanding the evolution and function of genomes depends on the identication of conserved areas within genomes, which is made possible by the use of colinearity analysis in comparative genomics [102].
9.5.14 Epigenomics
Epigenomics explores modications of genetic materials and associated proteins, affecting gene expression without altering the DNA sequence. It inuences high­throughput technologies to map the epigenome, enhancing understanding of differ­ent genomic outputs [103]. Single-cell epigenomics focused on understanding non­coding disease-associated human variants, single-cell epigenomics provides detailed maps of candidate cis-regulatory elements in heterogeneous human tissues, aiding in interpreting the genetic basis of common traits and diseases [103, 104]. High­throughput technologies have propelled epigenomics into a multi-omics era, where powerful tools describe and record diverse layers of genomic output, allowing for a thorough interrogation of cellular components [104].
9.5.15 Metagenomics
Metagenomics encompasses the analysis of genetic material directly sourced from environmental samples, extending the scope of microbial genomics beyond cultured organisms [105]. In the eld of metagenomics, two fundamental method­ologies play a pivotal role in understanding microbial communitiesmarker gene studies and whole-genome shotgun (WGS) metagenomics. These approaches are the foundation of our understanding of diverse ecosystems. Marker gene studies
9-24
Introduction to Pharmaceutical Biotechnology, Volume 2 (Second Edition)
focus on specic genes, often 16S ribosomal RNA genes, to provide insights into the taxonomic composition of microbial communities. Conversely, WGS meta­genomics takes a comprehensive approach, searching through the entire genomic material present within a sample, offering a comprehensive snapshot of microbial diversity and potential functional capabilities. These methods, while distinct, work together to enhance our com prehen sion of microbial ecosystems [106]. Metagenomics empowered by NGS, has brought signicant advancements in our ability to investigate this complex ecosystem. The gut microbiome exerts a profound inuence on host health, and metagenomi cs provides us with the tool s needed to interpret its role with unprecedented precision. By decoding the genetic makeup of the many microorganisms residing within our intestines, we gain valuable insights into their functional contributions to processes such as digestion, metabolism, and even immune system modulation. This novel understanding not only deepens our appreciation of the gut microbiomessignificance but also paves the way for novel therapeutic interventions aimed at maintaining and restoring the delicate balance within this microbial community, ultimately boosting host health and well-being [107].

9.6 Bioinformatics in precision medicine

Bioinformatics plays a pivotal role in precision medicine, which aims to tailor medical treatment to individual characteristics of each patient. It is an interdiscipli­nary eld that utilizes bioinformatics to analyze molecular data to identify genetic, epigenomic, and other molecular predictors of drug response and disease suscept­ibility. Genomic medicine is an emerging medical discipline that involves using genomic information about an individual as part of their clinical care and the health outcomes and policy implications of that clinical use. Pharmacogenomics is a crucial part of genomic medicine and precision medicine at large. It is the study of how genes affect a persons response to drugs. This eld aims to develop rational means to optimize drug therapy, with respect to the patients genotype, to ensure maximum efcacy with minimal adverse effects. By understanding an individuals genetic makeup, doctors can prescribe medications and doses personalized to the genetic prole of the individual [108, 109]. For instance, pharmacogenomics helps in predicting drug dose, like the dose of thiopurines based on variation in the thiopurine methyltransferase (TPMT). Despite its potential, the understanding of pharmacogenomic guidelines in clinical practice has been slow but is gradually improving. Genomic predictors of disease refer to genetic variants or patterns that are associated with the risk of developing particular diseases. Bioinformatics tools and techniques are employed to analyze large datasets to identify these predictors which can then be used for risk assessment, early detection, or personalized treatment plans. Genomic counseling extends beyond traditional genetic counseling to encompass a broader range of genetic and genomic information. It involves discussing with individuals or families about their genomic information and what it means for their health, disease risk, and other life aspects [110]. Key points in genomic counseling include:
9-25
Introduction to Pharmaceutical Biotechnology, Volume 2 (Second Edition)
The number and/or type of diseases for which testing is available and discussed.
The purpose of testing.
Intervention and clinical utility.
Access to testing [111].
Transition from traditional genetic counseling to genomic counseling involves
a move from reactive genetic testing for diagnosis of primarily single-gene diseases to proactive genome-based testing for multiple conditions [111].
Personal genomics pertains to the use of genomic information at an individual level to guide healthcare decisions includes:
Advanced diagnostics.
Tumor proling.
Genomic risk assessments [112].
Personal genomics has the potential to signicantly inuence healthcare by enabling personalized medicine, which tailors medical interventions to an individuals unique genomic prole. It is also a subject of genomics education, emphasizing the importance and benets of understanding personal genomic information in society, including in classrooms, clinics, and the public sphere [113].

9.7 Translational bioinformatics

Translational bioinformatics (TBI) is a crucial eld that bridges the gap between biomedical data science and informatics, determined for a seamless translation of scientic discoveries into clinical practice. TBI is in the forefront in areas like the clinical utility of polygenic risk scores, data integration, and using articial intelligence (AI) and ML for enhanced healthcare delivery [114]. Clinical decision support systems (CDSSs) are integral tools in TBI that assist healthcare profes­sionals in making informed clinical decisions. They analyze data from various sources to provide evidence-based recommendations tailored to individual patient conditions. By integrating with electronic health records (EHRs), these systems ensure a seamless ow of patient data, thereby facilitating better clinical decisions, improving patient outcomes, and reducing healthcare costs. The integration of EHRs is central to TBI as it facilitates the utilization of rich patient data for better clinical decision-making. By convoluting EHRs with clinical decision support, healthcare differences can be addressed more effectively, especially in low-resource settings, thereby improving health equity [115]. Moreover, the real-world integration of genomic data into EHRs is an effort to operationalize guidelines for precision medicine delivery, enhancing the EHRs capability for such advanced healthcare provision. A signicant aspect of a CDSS is its integration or interoperability with EHRs to enhance healthcare for populations facing discrepancies. This integration facilitates the seamless ow of patient-specic information, enabling personalized clinical recommendations. Various studies have been conducted on clinicians facing CDSS integrated or interoperable with EHR, revealing that most of these were
9-26
Introduction to Pharmaceutical Biotechnology, Volume 2 (Second Edition)
implemented in primary care outpatient settings for screening or treatment purposes. The types of clinical decision support (CDS) tools include point-of-care alerts, order facilitators, workow support, relevant information display, expert systems, and medication dosing support. The potential of CDS tools lies in improving health equity and outcomes, especially for patients who face disparities. The user-centered design is crucial during the planning, development, and implementation phases of CDSS. It ensures that the system is adapted to meet the needs and preferences of its users, thereby enhancing its usability and effectiveness. Implementation strategies most frequently employed include education and consensus facilitation, which are vital for the successful deployment and adoption of CDSSs [115].
Patient stratication is a crucial process in personalized medicine, which involves categorizing patients into different subgroups based on their risk of developing certain medical conditions or their response to specic treatments. This process aids in tailoring medical treatment to individual patients or groups of patients, thereby enhancing the effectiveness and efciency of healthcare delivery [116]. ML models play a signicant role in patient risk stratication by predicting future disease states based on a patients current clinical state and historical data. These models can help identify which applications of predictions will be sound by distinguishing between predictions based on physician behavior and those based on a constant representa­tion of patient physiology [117]. Disease modeling is essential for understanding the dynamics of disease spread and the impact of various factors on disease outcomes. It is a multidimensional approach encompassing biological, social, and behavioral factors. The integration of social and behavioral dynamics in disease models is particularly signicant as it inuences the emergence, spread, and containment of diseases. Recent global health threats like the Ebola and COVID-19 pandemics have highlighted the importance of developing disease models that incorporate these dynamics for more effective response measures and policies [118]. Disease modeling makes use of a wide range of approaches, including computational systems biology and the creation of patient-specic cell lines from induced pluripotent stem cells (iPSCs), to highlight on disease causes, medication responses, and prospective therapeutic treatments [119, 120].

9.8 Bioinformatics in drug discovery and development

The eld of bioinformatics has become more important in the creation of new medicines in recent years. It has several uses and is a major factor in the success of modern drug research. Understanding the molecular basis of diseases and ther­apeutic targets requires the management and analysis of massive volumes of biological data produced by elds such as genomics, transcriptomics, proteomics, population genetics, and molecular phylogenetics [121]. Predictive models may be developed to understand the behavior of bioactive chemicals using bioinformatics tools and methodologies, which can help in the design and development of novel medications. Bioinformatics can anticipate interactions between medications and proteins, examine the inuence on biological processes and activities, and uncover genetic variations that may modify drug response. This is critical for comprehending
9-27
Introduction to Pharmaceutical Biotechnology, Volume 2 (Second Edition)
the workings of possible medications and their side effects [122]. Bioinformatics may greatly speed up drug target selection, drug candidate screening, and renement despite the difculties involved in drug development. This helps to overcome the lengthy and costly nature of conventional drug discovery procedures [123]. Potential targets must successfully move from the discovery stage to clinical trials and ultimately the market, and bioinformatics plays a critical role in this process by assisting with the prediction, analysis, or interpretation of clinical data. To better understand and solve problems in drug discovery and development, chemoinfor­matics, combines chemistry and computer science. To better comprehend chemical characteristics and molecular interactions in biological systems, it uses computa­tional methods to examine chemical data. When it comes to the computer study of molecules, molecular representation plays a crucial role by providing a means for the systematic storage and retrieval of chemical data in both 2D and 3D formats. Chemical databases are mined for relevant patterns and insights using data mining methods such structural similarity matrices, classication algorithms, and descriptor computations. Additionally, chemoinformatics and quantitative structure–activity relationship (QSAR) modeling work in combination to improve molecular design prediction modeling and provide insightful information about the behavior of bioactive substances. Together, these fundamental applications enhance our knowl­edge of chemical structures and how they interact, advancing materials research and medication development. Chemoinformatics supports the production of compre­hensive databases, predictive models, and analytical tools required for current drug discovery procedures, expediting the journey from molecular idea to drug develop­ment [124].
New therapeutic indications for previously authorized or experimental medica­tions may be found via a process known as drug repositioning, repurposing, redirecting, or reproling. Because of the high prices, severe hazards, and poor speed of conventional drug development techniques, this method is gaining popular­ity. A crucial component of pharmaceutical research is drug repositioning, which entails a number of essential components. In this process, computational techniques play a key role by analyzing current data to nd novel potential uses for well­established medications. These methods use ML and computational tools to nd correlations between medications and diseases based on chemical, genomic, or phenotypic data. In addition, drug repositioning is well-known for its nancial viability, as it provides a cost-effective alternative to conventional drug discovery by making use of already-existing data and resources, thus shortening the time and lowering the costs normally associated with bringing drugs to market [125].
In the early phases of drug development, virtual screening (VS) is used to lter through chemical libraries in search of structures with the highest propensity to bind to therapeutic targets. It is crucial for cutting down on the time and money needed to nd potential medication candidates. Consensus virtual screening (CVS) is a recent development in SBVS (structure-based virtual screening) with the goal of improving SBVS accuracy and decreasing false positives in these kinds of research. Having access to the 3D structure of the target protein is necessary for SBVS to be used. The process of predicting drug–target interactions (DTIs) is a cornerstone of the
9-28
Introduction to Pharmaceutical Biotechnology, Volume 2 (Second Edition)
pharmaceutical industry. Drug mechanism of action, impact prediction, and side effect avoidance all rely on this process of discovering possible drug interactions with biological targets. Predicting DTIs is important for several reasons in the pharma­ceutical industry, including virtual screening, drug repurposing, and the detection of adverse drug reactions. It has been reported that using deep learning and ML models in virtual screening might enhance DTI prediction, emphasizing the use of computa­tional methods in contemporary drug development procedures [126].
9.8.1 Articial intelligence (AI) and machine learning in bioinformatics
The use of deep learning, a type of ML and AI, in the analysis of big biological data, prediction, and the discovery of new biological insights has shown great promise in bioinformatics. The analysis of sequences, the identication of motifs, and the prediction of the function or structure of molecules have all been performed using deep learning models like CNNs and RNNs. Understanding the function of proteins and other biomolecules requires knowing their 3D structures, which may be predicted. Predicting molecular structures using deep learning, and in particular CNNs, has led to signicant advances in the eld, with tools like AlphaFold highlighting the potential of deep learning in structural biology. Learning how genes work and what inuences their expression, such as epigenetics, is made easier with the help of deep learning. Deep learning may be used to investigate massive genomic datasets in order to better understand the regulatory networks that govern cellular activity. Deep learning aids in the comprehension of complex biological systems by combining data at many biological scales. Multi-omics analysis, which involves characterizing and quantifying groups of biological molecules, falls under this category [127]. Different methods used in AI are listed in gure 9.6.
9.8.2 AI-driven drug discovery
Modern drug development procedures greatly benet from AI, especially ML and deep learning, which speed up the conversion of data into useful ideas. Big datasets of chemical substances may be processed by AI algorithms, which can then forecast their possible toxicity and effectiveness. This speeds up the drug development processs rst screening step considerably. With the analysis of vast datasets, AI may discover novel therapeutic applications for already-approved medications and potentially connect pharmacological mechanisms with disease pathways. For pharmaceuticals to be safe, it is essential to anticipate how they may interact with one another or with diseases. AI may help in the creation of safer medications by analyzing large, complicated information to predict potential interactions. The identication of biomarkers is essential for tracking the effectiveness of treatments and making diagnosis. AI can select through massive biological datasets to uncover potential biomarkers. AI can help design more efcient clinical trials, predict outcomes, and identify the most suitable candidates for trials. These speed up the drug development process and bring new therapies to patients faster. Predictive modeling encompasses a variety of statistical and ML techniques to predict outcomes based on data. It is a crucial tool in bioinformatics and drug discovery,
9-29
Introduction to Pharmaceutical Biotechnology, Volume 2 (Second Edition)
Figure 9.6. Commonly employed machine learning algorithms in bioinformatics investigations are outlined. Each algorithm is accompanied by an example of its application and the corresponding input data, as depicted on the right side. Abbreviations include SVM for support vector machines, KNN for K-nearest neighbors, CNN for convolutional neural networks, RNN for recurrent neural networks, PCA for principal component analysis, t-SNE for t-distributed stochastic neighbor embedding, and NMF for non-negative matrix factorization. Reproduced from [
128] CC BY 4.0.
where it enables the forecasting of biological and chemical phenomena based on existing data [129]. Regression analysis is used to predict continuous outcomes. For instance, it can be used to predict the binding afnity of a drug to a particular target based on certain molecular descriptors. Classication techniques are used to predict categorical outcomes. In bioinformatics, it might be used to classify genes into different functional categories based on expression data. Time-series analysis can predict outcomes over time, essential in studying biological processes that change
9-30
Introduction to Pharmaceutical Biotechnology, Volume 2 (Second Edition)
over time like cell growth or gene expression levels over a developmental timeline. Essential for handling high-dimensional data common in bioinformatics and drug discovery. It identies the most informative features or reduces the dimensionality of the data without losing essential information, and ensures the reliability and robustness of predictive models through techniques like cross-validation [130].
Advancements in AI, particularly deep learning, have signicantly impacted image analysis and microscopy in biological research, enabling automated and more accurate analysis. Deep learning algorithms, especially CNNs, have empowered automated image segmentation, object detection, and classication in biological and medical imaging. Deep learning can enhance the resolution of microscopy images through super-resolution techniques, enabling finer examination of biological structures. Analyzing time-lapse microscopy images using AI can reveal dynamic biological processes, tracking cells and understanding cellular dynamics over time. High-content screening (HCS) often involves analyzing large sets of microscopy images to understand cellular phenotypes. AI accelerates this analysis, identifying phenotypic changes in response to various treatments. 3D image analysis is crucial for understanding complex biological structures. Deep learning algorithms can segment, classify, and analyze 3D images, providing insights into 3D cellular structures and tissues. Comprehensive understanding of biological processes may be gained by combining data from several imaging modalities or picture data with other forms of data (such as genomic data) [131].

9.9 CRISPR and genome editing in bioinformatics

Genome editing has been given a major boost by the CRISPR-Cas (Clustered Regularly Interspaced Short Palindromic Repeats and CRISPR-associated proteins) technology. Bioinformaticsability to streamline the planning, analysis, and interpretation of studies is crucial to expanding and directing CRISPR-Cas applications. Designing good guide RNAs (gRNAs) is critical for efcient CRISPR-Cas genome editing. In order to pick gRNA sequences with optimal on­target activity and minimal off-target effects, bioinformatics methods are useful. Using sequence similarity and other genetic markers, bioinformatics systems may identify probable off-target regions. Algorithms and ML models have been created to predict the on-target effectiveness of gRNAs, facilitating in the selection of the most successful gRNAs for genome editing. Numerous options for genome editing are made possible by the variety of CRISPR-Cas systems [132]. Bioinformatics aids in dening and cataloging these variations, understanding their causes, and predicting their genome-editing capabilities. The collection of data from CRISPR­Cas research has led to the formation of databases and libraries. These tools are useful for the community, allowing the exchange of gRNA sequences, efciency ratings, and off-target predictions. Genes implicated in certain traits may be discovered using CRISPR screening. The high-throughput data produced by these screens, the identication of relevant hits, and the interpretation of the ndings all need the use of bioinformatics tools. The integration of CRISPR-Cas bioinformatics tools with other bioinformatics resources may give more complete insights. The
9-31
Introduction to Pharmaceutical Biotechnology, Volume 2 (Second Edition)
whole extent to which genome editing affects biological processes, for instance, may be shown by integrating data on genome editing with those on gene expression, proteomics, or pathway analysis [133].
Computational tools are essential in genome editing because they guarantee accuracy and productivity of the editing procedure. The development of guide RNAs (gRNAs), the anticipation of off-target effects, and the analysis of genome editing experiment results. Software tools like CRISPR Design and Benchling aid in designing gRNAs with high on-target efciency and minimal off-target effects. They often provide a user-friendly interface for selecting target sequences and designing gRNAs. Tools like Cas-OFFinder and CRISPRscan help in predicting potential off­target sites based on sequence similarity, aiding in the design of gRNAs with higher specicity. Certain tools offer predictions on the efciency of designed gRNAs, integrating factors like sequence composition and genomic context to provide a score indicative of the likely editing efciency. For projects requiring multiplexed editing, tools that assist in designing multiple gRNAs and organizing complex editing strategies are crucial.
Genome editing, especially in the context of human genomes, brings out plenty of ethical considerations. Editing the germline potentially alters the genetic makeup of future generations, raising concerns about unintended consequences and the ethics of altering human evolution. Obtaining informed consent, especially in germline editing, is a complex issue, given the long-term and often unknown implications of genome editing. There is concern over genome editing technologies becoming a means of extending existing social inequalities if access is restricted to afuent individuals or communities. Establishing robust regulatory frameworks is crucial to ensure the responsible use of genome editing technologies [134].
Functional genomics aims to understand the relationship between the genome and the phenotype. CRISPR-based screens are powerful tools in functional genomics. Systematically interrogating gene function on a genome-wide scale by utilizing libraries of gRNAs targeting every gene in the genome. Following genome editing, phenotypic changes can be assessed using a variety of assays, elucidating the function of individual genes or gene networks. Bioinformatics tools are essential for analyzing the vast datasets generated from CRISPR screens, identifying signicant genes, and interpreting the biological implications. Combining CRISPR screen data with other omics data (transcriptomics and proteomics) can provide a more broader understanding of gene function and cellular processes [135].

9.10 Integrative and multi-omics analysis

The advent of high-throughput technologies has facilitated the generation of large­scale biological data across different molecular levels. Integrative and multi-omics analyses seek to combine these data to provide a more comprehensive understanding of biological systems. Integrative genomics aims to combine data from different genomic datasets to elucidate the underlying genetic architectures of traits and diseases. Integrative genomics is the harmonization of different genomic datasets, ensuring data representation and analysis consistency. Meta-analysis techniques are
9-32