Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 119 for “"protein sequences"”.

  1. Folding patterns in protein sequences

    … fruitful genome projects, a growing number of protein sequences remain uncharacterised. Generally, computational techniques have not been developed to distinguish between regions that are conserved in proteins owing to evolutionary pressures to maintain structure and those that are conserved in …

    westminster Repository record for Folding patterns in protein sequences (opens in a new tab)

  2. Improving Profile Similarity Search and Alignment of Protein Sequences

    Protein function prediction is one of the most important problems in the field of computational biology. The most reliable method to predict protein function is to detect homologs. Homologous proteins tend to possess conserved sequence motifs, the same structure folds, and similar functional sites. …

    utswmed Repository record for Improving Profile Similarity Search and Alignment of Protein Sequences (opens in a new tab)

  3. Modeling protein sequences, structures and functions with deep neural networks

    Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2023-12-04 without embargo terms

    uiuc Repository record for Modeling protein sequences, structures and functions with deep neural networks (opens in a new tab)

  4. Enabling the discovery of novel bioactives and enzymes from billions of protein sequences

    … used platforms for the analysis of metagenomic sequences. Since 2018, MGnify has assembled and analysed over 57,000 shotgun metagenomics datasets. From these assemblies, over 2.4 billion non-redundant protein sequences have been identified. This protein database outscales other protein

    cambridge Repository record for Enabling the discovery of novel bioactives and enzymes from billions of protein sequences (opens in a new tab)

  5. Prediction of parallel in-register amyloidogenic beta-structures In highly beta-rich protein sequences by pairwise propensity analysis

    Amyloids and prion proteins are clinically and biologically important beta-structures, whose supersecondary structures are difficult to determine by standard experimental or computational means. In addition, significant conformational heterogeneity is known or suspected to exist in many amyloid …

    mit Repository record for Prediction of parallel in-register amyloidogenic beta-structures In highly beta-rich protein sequences by pairwise propensity analysis (opens in a new tab)

  6. ChaperoNet: Distillation of Language Model Semantics to Folded Three-Dimensional Protein Structures

    Determining the structure of proteins has been a long-standing goal in biology. Lan- guage models have been recently deployed to capture the evolutionary semantics of protein sequences, and as an emergent property, were found to be structural learn- ers. Enriched with multiple sequence alignments …

    mit Repository record for ChaperoNet: Distillation of Language Model Semantics to Folded Three-Dimensional Protein Structures (opens in a new tab)

  7. Integrating Functional Knowledge into Protein Design: A Novel Approach to Tokenization and Noise Injection for Function-Aware Protein Language Models

    Designing novel proteins with specific biological functions remains a fundamental challenge in computational biology. While recent advances in protein language models have enabled powerful sequence-based representations, most models, including state-of-the-art systems like ESM3, fall short in …

    mit Repository record for Integrating Functional Knowledge into Protein Design: A Novel Approach to Tokenization and Noise Injection for Function-Aware Protein Language Models (opens in a new tab)

  8. Calculating the structure-based phylogenetic relationship of distantly related homologous proteins utilizing maximum likelihood structural alignment combinatorics and a novel structural molecular clock hypothesis

    … relationships and homology of species, proteins, or genes. Homology modeling, ligand binding, and pharmaceutical testing all depend upon the homology ascertained by dendrograms. Regardless of the specific algorithm, all dendrograms that ascertain protein evolutionary homology are …

    umkc Repository record for Calculating the structure-based phylogenetic relationship of distantly related homologous proteins utilizing maximum likelihood structural alignment combinatorics and a novel structural molecular clock hypothesis (opens in a new tab)

  9. Assessing the Role of Clusters Derived from Large Sequence Similarity Networks for Gene Function Predictions

    … utilize various forms of data, such as DNA/RNA/protein sequences, protein structures, interaction networks, literature mining, and a combination of these data sources. However, these methods do not always produce precise results as the underlying data sets used for training or modeling are quite …

    vt Repository record for Assessing the Role of Clusters Derived from Large Sequence Similarity Networks for Gene Function Predictions (opens in a new tab)

  10. Fast learning optimized prediction methodology for protein secondary structure prediction, relative solvent accessibility prediction and phosphorylation prediction

    … and the large disparity between the number of sequences and the number of structures. There has been an exponential growth in the number of available protein sequences and a slower growth in the number of structures. There is therefore an urgent need to develop computed structures and identify …

    iastate Repository record for Fast learning optimized prediction methodology for protein secondary structure prediction, relative solvent accessibility prediction and phosphorylation prediction (opens in a new tab)

  11. Molecular Evolution of Viruses

    … of evolution at the scale of DNA, RNA and proteins. Our goal was to study molecular evolution of viruses with special reference to Influenza A Virus.</p> <p>The recent Influenza A/H1N1(2009) outbreak has been the focus of intense research because of its high level of infectivity across the …

    south-carolina Repository record for Molecular Evolution of Viruses (opens in a new tab)

  12. Intrinsically Disordered Protein Polymer Libraries as Tools to Understand Protein Hydrophobicity

    <p>Intrinsically disordered protein polymers (IDPPs) are repetitive biopolymers that, when enriched with prolines, glycines, and aliphatic amino acids, have observable lower critical solution temperature (LCST) phase transition behavior at physiologically relevant temperature and concentration …

    duke Repository record for Intrinsically Disordered Protein Polymer Libraries as Tools to Understand Protein Hydrophobicity (opens in a new tab)

  13. Improving the genome assembly and annotation of the white-tailed deer (Odocoileus virginianus borealis)

    … the BRAKER annotation pipeline using RNA and protein sequences as extrinsic evidence. The final assembly was highly contiguous, with 90% of the total length represented by 134 contigs. The largest contig was 108 million base pairs. Functional annotation was performed using reciprocal best hits …

    uiuc Repository record for Improving the genome assembly and annotation of the white-tailed deer (Odocoileus virginianus borealis) (opens in a new tab)

  14. Studies of the structure, evolution, and TonB-dependence of the colicin I receptor of Escherichia coli

    The colicin I receptor (Cir) is an outer membrane protein, the expression of which is iron regulated. The DNA downstream of the cloned cir gene, when expressed in a maxicell system, encoded a protein of molecular weight 32,000. The expression of this protein was determined to be directed toward the …

    uiuc Repository record for Studies of the structure, evolution, and TonB-dependence of the colicin I receptor of Escherichia coli (opens in a new tab)

  15. PASSS: Protein Active Site Structure Search

    The Protein Structure Initiative, a project on the scale of the Human Genome Project aimed at protein structure determination, has successfully identified the structure of multiple human proteins. Unfortunately, knowledge of structure alone provides little insight into a protein's function within …

    wfu Repository record for PASSS: Protein Active Site Structure Search (opens in a new tab)

  16. Structural and functional relationship of photosynthetic bacterial reaction centers

    … Rhodopseudomonas viridis. By comparison of the protein sequences and structures of RCs of three species, it was suggested that Cys$\sp{\rm L108}$ in Rb. sphaeroides (Cys$\sp{\rm L109}$ in Rb. capsulatus) is the best candidate for the mercurial reagents to inhibit Q$\sb{\rm B}$ function. And a …

    uiuc Repository record for Structural and functional relationship of photosynthetic bacterial reaction centers (opens in a new tab)

  17. Predicting the beta-trefoil fold from protein sequence data

    … level, to predict the beta-structural motifs in protein sequences. A program called Wrap-and-Pack implements this method, and is shown to recognize β-trefoils, an important class of globular β-structures, in the Protein Data Bank with 92% specificity and 92.3% sensitivity in cross-validation. It …

    mit Repository record for Predicting the beta-trefoil fold from protein sequence data (opens in a new tab)

  18. Statistics and algorithms for peptide mass fingerprinting

    … (PMF). In a PMF experiment, a purified protein sample is digested by a protease using an enzymatic cleavage reaction, the masses of the resulting peptides are measured by mass spectrometry, yielding the mass fingerprint, and compared to predicted mass fingerprints of reference protein

    bielefeld Repository record for Statistics and algorithms for peptide mass fingerprinting (opens in a new tab)

Page 1 of 6