首页 | 本学科首页   官方微博 | 高级检索  
相似文献
 共查询到20条相似文献,搜索用时 31 毫秒
1.
Mathematical tools developed in the context of Shannon information theory were used to analyze the meaning of the BLOSUM score, which was split into three components termed as the BLOSUM spectrum (or BLOSpectrum). These relate respectively to the sequence convergence (the stochastic similarity of the two protein sequences), to the background frequency divergence (typicality of the amino acid probability distribution in each sequence), and to the target frequency divergence (compliance of the amino acid variations between the two sequences to the protein model implicit in the BLOCKS database). This treatment sharpens the protein sequence comparison, providing a rationale for the biological significance of the obtained score, and helps to identify weakly related sequences. Moreover, the BLOSpectrum can guide the choice of the most appropriate scoring matrix, tailoring it to the evolutionary divergence associated with the two sequences, or indicate if a compositionally adjusted matrix could perform better.[1,2,3,4,5,6,7,8,9,10,11,12,13,14,15,16,17,18,19,20,21,22,23,24,25,26,27,28,29]  相似文献   

2.
3.
4.
5.
6.
A decoding algorithm is tested that mechanistically models the progressive alignments that arise as the mRNA moves past the rRNA tail during translation elongation. Each of these alignments provides an opportunity for hybridization between the single-stranded, -terminal nucleotides of the 16S rRNA and the spatially accessible window of mRNA sequence, from which a free energy value can be calculated. Using this algorithm we show that a periodic, energetic pattern of frequency 1/3 is revealed. This periodic signal exists in the majority of coding regions of eubacterial genes, but not in the non-coding regions encoding the 16S and 23S rRNAs. Signal analysis reveals that the population of coding regions of each bacterial species has a mean phase that is correlated in a statistically significant way with species () content. These results suggest that the periodic signal could function as a synchronization signal for the maintenance of reading frame and that codon usage provides a mechanism for manipulation of signal phase.[1,2,3,4,5,6,7,8,9,10,11,12,13,14,15,16,17,18,19,20,21,22,23,24,25,26,27,28,29,30,31,32]  相似文献   

7.
8.
9.
10.
Glycoprotein structure determination and quantification by MS requires efficient isolation of glycopeptides from a proteolytic digest of complex protein mixtures. Here we describe that the use of acids as ion-pairing reagents in normal-phase chromatography (IP-NPLC) considerably increases the hydrophobicity differences between non-glycopeptides and glycopeptides, thereby resulting in the reproducible isolation of N-linked high mannose type and sialylated glycopeptides from the tryptic digest of a ribonuclease B and fetuin mixture. The elution order of non-glycopeptides relative to glycopeptides in IP-NPLC is predictable by their hydrophobicity values calculated using the Wimley-White water/octanol hydrophobicity scale. O-linked glycopeptides can be efficiently isolated from fetuin tryptic digests using IP-NPLC when N-glycans are first removed with PNGase. IP-NPLC recovers close to 100% of bacterial N-linked glycopeptides modified with non-sialylated heptasaccharides from tryptic digests of periplasmic protein extracts from Campylobacter jejuni 11168 and its pglD mutant. Label-free nano-flow reversed-phase LC-MS is used for quantification of differentially expressed glycopeptides from the C. jejuni wild-type and pglD mutant followed by identification of these glycoproteins using multiple stage tandem MS. This method further confirms the acetyltransferase activity of PglD and demonstrates for the first time that heptasaccharides containing monoacetylated bacillosamine are transferred to proteins in both the wild-type and mutant strains. We believe that IP-NPLC will be a useful tool for quantitative glycoproteomics.Protein glycosylation is a biologically significant and complex post-translational modification, involved in cell-cell and receptor-ligand interactions (14). In fact, clinical biomarkers and therapeutic targets are often glycoproteins (59). Comprehensive glycoprotein characterization, involving glycosylation site identification, glycan structure determination, site occupancy, and glycan isoform distribution, is a technical challenge particularly for quantitative profiling of complex protein mixtures (1013). Both N- and O-glycans are structurally heterogeneous (i.e. a single site may have different glycans attached or be only partially occupied). Therefore, the MS1 signals from glycopeptides originating from a glycoprotein are often weaker than from non-glycopeptides. In addition, the ionization efficiency of glycopeptides is low compared with that of non-glycopeptides and is often suppressed in the presence of non-glycopeptides (1113). When the MS signals of glycopeptides are relatively high in simple protein digests then diagnostic sugar oxonium ion fragments produced by, for example, front-end collisional activation can be used to detect them. However, when peptides and glycopeptides co-elute, parent ion scanning is required to selectively detect the glycopeptides (14). This can be problematic in terms of sensitivity, especially for detecting glycopeptides in digests of complex protein extracts.Isolation of glycopeptides from proteolytic digests of complex protein mixtures can greatly enhance the MS signals of glycopeptides using reversed-phase LC-ESI-MS (RPLC-ESI-MS) or MALDI-MS (1524). Hydrazide chemistry is used to isolate, identify, and quantify N-linked glycopeptides effectively, but this method involves lengthy chemical procedures and does not preserve the glycan moieties thereby losing valuable information on glycan structure and site occupancy (1517). Capturing glycopeptides with lectins has been widely used, but restricted specificities and unspecific binding are major drawbacks of this method (1821). Under reversed-phase LC conditions, glycopeptides from tryptic digests of gel-separated glycoproteins have been enriched using graphite powder medium (22). In this case, however, a second digestion with proteinase K is required for trimming down the peptide moieties of tryptic glycopeptides so that the glycopeptides (typically <5 amino acid residues) essentially resemble the glycans with respect to hydrophilicity for subsequent separation. Moreover, the short peptide sequences of the proteinase K digest are often inadequate for de novo sequencing of the glycopeptides.Glycopeptide enrichment under normal-phase LC (NPLC) conditions has been demonstrated using various hydrophilic media and different capture and elution conditions (2328). NPLC allows either direct enrichment of peptides modified by various N-linked glycan structures using a ZIC®-HILIC column (2327) or targeting sialylated glycopeptides using a titanium dioxide micro-column (28). However, NPLC is neither effective for enriching less hydrophilic glycopeptides, e.g. the five high mannose type glycopeptides modified by 7–11 monosaccharide units from a tryptic digest of ribonuclease b (RNase B), nor for enriching O-linked glycopeptides of bovine fetuin using a ZIC-HILIC column (23). The use of Sepharose medium for enriching glycopeptides yielded only modest recovery of glycopeptides (28). In addition, binding of hydrophilic non-glycopeptides with these hydrophilic media contaminates the enriched glycopeptides (23, 28).We have recently developed an ion-pairing normal-phase LC (IP-NPLC) method to enrich glycopeptides from complex tryptic digests using Sepharose medium and salts or bases as ion-pairing reagents (29). Though reasonably effective the technique still left room for significant improvement. For example, the method demonstrated relatively modest glycopeptide selectivity, providing only 16% recovery for high mannose type glycopeptides (29). Here we report on a new IP-NPLC method using acids as ion-pairing reagents and polyhydroxyethyl aspartamide (A) as the stationary phase for the effective isolation of tryptic glycopeptides. The method was developed and evaluated using a tryptic digest of RNase B and fetuin mixture. In addition, we demonstrate that O-linked glycopeptides can be effectively isolated from a fetuin tryptic digest by IP-NPLC after removal of the N-linked glycans by PNGase F.The new IP-NPLC method was used to enrich N-linked glycopeptides from the tryptic digests of protein extracts of wild-type (wt) and PglD mutant strains of Campylobacter jejuni NCTC 11168. C. jejuni has a unique N-glycosylation system that glycosylates periplasmic and inner membrane proteins containing the extended N-linked sequon, D/E-X-N-X-S/T, where X is any amino acid other than proline (3032). The N-linked glycan of C. jejuni has been previously determined to be GalNAc-α1,4-GalNAc-α1,4-[Glcβ1,3]-GalNAc-α1,4-GalNAc-α1,4-GalNAc-α1,3-Bac-β1 (BacGalNAc5Glc residue mass: 1406 Da), where Bac is 2,4-diacetamido-2,4,6-trideoxyglucopyranose (30). In addition, the glycan structure of C. jejuni is conserved, unlike in eukaryotic systems (3032). IP-NPLC recovered close to 100% of the bacterial N-linked glycopeptides with virtually no contamination of non-glycopeptides. Furthermore, we demonstrate for the first time that acetylation of bacillosamine is incomplete in the wt using IP-NPLC and label-free MS.  相似文献   

11.
12.
13.
SPA2 encodes a yeast protein that is one of the first proteins to localize to sites of polarized growth, such as the shmoo tip and the incipient bud. The dynamics and requirements for Spa2p localization in living cells are examined using Spa2p green fluorescent protein fusions. Spa2p localizes to one edge of unbudded cells and subsequently is observable in the bud tip. Finally, during cytokinesis Spa2p is present as a ring at the mother–daughter bud neck. The bud emergence mutants bem1 and bem2 and mutants defective in the septins do not affect Spa2p localization to the bud tip. Strikingly, a small domain of Spa2p comprised of 150 amino acids is necessary and sufficient for localization to sites of polarized growth. This localization domain and the amino terminus of Spa2p are essential for its function in mating. Searching the yeast genome database revealed a previously uncharacterized protein which we name, Sph1p (Spa2p homolog), with significant homology to the localization domain and amino terminus of Spa2p. This protein also localizes to sites of polarized growth in budding and mating cells. SPH1, which is similar to SPA2, is required for bipolar budding and plays a role in shmoo formation. Overexpression of either Spa2p or Sph1p can block the localization of either protein fused to green fluorescent protein, suggesting that both Spa2p and Sph1p bind to and are localized by the same component. The identification of a 150–amino acid domain necessary and sufficient for localization of Spa2p to sites of polarized growth and the existence of this domain in another yeast protein Sph1p suggest that the early localization of these proteins may be mediated by a receptor that recognizes this small domain.Polarized cell growth and division are essential cellular processes that play a crucial role in the development of eukaryotic organisms. Cell fate can be determined by cell asymmetry during cell division (Horvitz and Herskowitz, 1992; Cohen and Hyman, 1994; Rhyu and Knoblich, 1995). Consequently, the molecules involved in the generation and maintenance of cell asymmetry are important in the process of cell fate determination. Polarized growth can occur in response to external signals such as growth towards a nutrient (Rodriguez-Boulan and Nelson, 1989; Eaton and Simons, 1995) or hormone (Jackson and Hartwell, 1990a , b ; Segall, 1993; Keynes and Cook, 1995) and in response to internal signals as in Caenorhabditis elegans (Goldstein et al., 1993; Kimble, 1994; Priess, 1994) and Drosophila melanogaster (St Johnston and Nusslein-Volhard, 1992; Anderson, 1995) early development. Saccharomyces cerevisiae undergo polarized growth towards an external cue during mating and to an internal cue during budding. Polarization towards a mating partner (shmoo formation) and towards a new bud site requires a number of proteins (Chenevert, 1994; Chant, 1996; Drubin and Nelson, 1996). Many of these proteins are necessary for both processes and are localized to sites of polarized growth, identified by the insertion of new cell wall material (Tkacz and Lampen, 1972; Farkas et al., 1974; Lew and Reed, 1993) to the shmoo tip, bud tip, and mother–daughter bud neck. In yeast, proteins localized to growth sites include cytoskeletal proteins (Adams and Pringle, 1984; Kilmartin and Adams, 1984; Ford, S.K., and J.R. Pringle. 1986. Yeast. 2:S114; Drubin et al., 1988; Snyder, 1989; Snyder et al., 1991; Amatruda and Cooper, 1992; Lew and Reed, 1993; Waddle et al., 1996), neck filament components (septins) (Byers and Goetsch, 1976; Kim et al., 1991; Ford and Pringle, 1991; Haarer and Pringle, 1987; Longtine et al., 1996), motor proteins (Lillie and Brown, 1994), G-proteins (Ziman, 1993; Yamochi et al., 1994; Qadota et al., 1996), and two membrane proteins (Halme et al., 1996; Roemer et al., 1996; Qadota et al., 1996). Septins, actin, and actin-associated proteins localize early in the cell cycle, before a bud or shmoo tip is recognizable. How this group of proteins is localized to and maintained at sites of cell growth remains unclear.Spa2p is one of the first proteins involved in bud formation to localize to the incipient bud site before a bud is recognizable (Snyder, 1989; Snyder et al., 1991; Chant, 1996). Spa2p has been localized to where a new bud will form at approximately the same time as actin patches concentrate at this region (Snyder et al., 1991). An understanding of how Spa2p localizes to incipient bud sites will shed light on the very early stages of cell polarization. Later in the cell cycle, Spa2p is also found at the mother–daughter bud neck in cells undergoing cytokinesis. Spa2p, a nonessential protein, has been shown to be involved in bud site selection (Snyder, 1989; Zahner et al., 1996), shmoo formation (Gehrung and Snyder, 1990), and mating (Gehrung and Snyder, 1990; Chenevert et al., 1994; Yorihuzi and Ohsumi, 1994; Dorer et al., 1995). Genetic studies also suggest that Spa2p has a role in cytokinesis (Flescher et al., 1993), yet little is known about how this protein is localized to sites of polarized growth.We have used Spa2p green fluorescent protein (GFP)1 fusions to investigate the early localization of Spa2p to sites of polarized growth in living cells. Our results demonstrate that a small domain of ∼150 amino acids of this large 1,466-residue protein is sufficient for targeting to sites of polarized growth and is necessary for Spa2p function. Furthermore, we have identified and characterized a novel yeast protein, Sph1p, which has homology to both the Spa2p amino terminus and the Spa2p localization domain. Sph1p localizes to similar regions of polarized growth and sph1 mutants have similar phenotypes as spa2 mutants.  相似文献   

14.
15.
A Boolean network is a model used to study the interactions between different genes in genetic regulatory networks. In this paper, we present several algorithms using gene ordering and feedback vertex sets to identify singleton attractors and small attractors in Boolean networks. We analyze the average case time complexities of some of the proposed algorithms. For instance, it is shown that the outdegree-based ordering algorithm for finding singleton attractors works in time for , which is much faster than the naive time algorithm, where is the number of genes and is the maximum indegree. We performed extensive computational experiments on these algorithms, which resulted in good agreement with theoretical results. In contrast, we give a simple and complete proof for showing that finding an attractor with the shortest period is NP-hard.[1,2,3,4,5,6,7,8,9,10,11,12,13,14,15,16,17,18,19,20,21,22,23,24,25,26,27,28,29,30,31,32]  相似文献   

16.
17.
A complete understanding of the biological functions of large signaling peptides (>4 kDa) requires comprehensive characterization of their amino acid sequences and post-translational modifications, which presents significant analytical challenges. In the past decade, there has been great success with mass spectrometry-based de novo sequencing of small neuropeptides. However, these approaches are less applicable to larger neuropeptides because of the inefficient fragmentation of peptides larger than 4 kDa and their lower endogenous abundance. The conventional proteomics approach focuses on large-scale determination of protein identities via database searching, lacking the ability for in-depth elucidation of individual amino acid residues. Here, we present a multifaceted MS approach for identification and characterization of large crustacean hyperglycemic hormone (CHH)-family neuropeptides, a class of peptide hormones that play central roles in the regulation of many important physiological processes of crustaceans. Six crustacean CHH-family neuropeptides (8–9.5 kDa), including two novel peptides with extensive disulfide linkages and PTMs, were fully sequenced without reference to genomic databases. High-definition de novo sequencing was achieved by a combination of bottom-up, off-line top-down, and on-line top-down tandem MS methods. Statistical evaluation indicated that these methods provided complementary information for sequence interpretation and increased the local identification confidence of each amino acid. Further investigations by MALDI imaging MS mapped the spatial distribution and colocalization patterns of various CHH-family neuropeptides in the neuroendocrine organs, revealing that two CHH-subfamilies are involved in distinct signaling pathways.Neuropeptides and hormones comprise a diverse class of signaling molecules involved in numerous essential physiological processes, including analgesia, reward, food intake, learning and memory (1). Disorders of the neurosecretory and neuroendocrine systems influence many pathological processes. For example, obesity results from failure of energy homeostasis in association with endocrine alterations (2, 3). Previous work from our lab used crustaceans as model organisms found that multiple neuropeptides were implicated in control of food intake, including RFamides, tachykinin related peptides, RYamides, and pyrokinins (46).Crustacean hyperglycemic hormone (CHH)1 family neuropeptides play a central role in energy homeostasis of crustaceans (717). Hyperglycemic response of the CHHs was first reported after injection of crude eyestalk extract in crustaceans. Based on their preprohormone organization, the CHH family can be grouped into two sub-families: subfamily-I containing CHH, and subfamily-II containing molt-inhibiting hormone (MIH) and mandibular organ-inhibiting hormone (MOIH). The preprohormones of the subfamily-I have a CHH precursor related peptide (CPRP) that is cleaved off during processing; and preprohormones of the subfamily-II lack the CPRP (9). Uncovering their physiological functions will provide new insights into neuroendocrine regulation of energy homeostasis.Characterization of CHH-family neuropeptides is challenging. They are comprised of more than 70 amino acids and often contain multiple post-translational modifications (PTMs) and complex disulfide bridge connections (7). In addition, physiological concentrations of these peptide hormones are typically below picomolar level, and most crustacean species do not have available genome and proteome databases to assist MS-based sequencing.MS-based neuropeptidomics provides a powerful tool for rapid discovery and analysis of a large number of endogenous peptides from the brain and the central nervous system. Our group and others have greatly expanded the peptidomes of many model organisms (3, 1833). For example, we have discovered more than 200 neuropeptides with several neuropeptide families consisting of as many as 20–40 members in a simple crustacean model system (5, 6, 2531, 34). However, a majority of these neuropeptides are small peptides with 5–15 amino acid residues long, leaving a gap of identifying larger signaling peptides from organisms without sequenced genome. The observed lack of larger size peptide hormones can be attributed to the lack of effective de novo sequencing strategies for neuropeptides larger than 4 kDa, which are inherently more difficult to fragment using conventional techniques (3437). Although classical proteomics studies examine larger proteins, these tools are limited to identification based on database searching with one or more peptides matching without complete amino acid sequence coverage (36, 38).Large populations of neuropeptides from 4–10 kDa exist in the nervous systems of both vertebrates and invertebrates (9, 39, 40). Understanding their functional roles requires sufficient molecular knowledge and a unique analytical approach. Therefore, developing effective and reliable methods for de novo sequencing of large neuropeptides at the individual amino acid residue level is an urgent gap to fill in neurobiology. In this study, we present a multifaceted MS strategy aimed at high-definition de novo sequencing and comprehensive characterization of the CHH-family neuropeptides in crustacean central nervous system. The high-definition de novo sequencing was achieved by a combination of three methods: (1) enzymatic digestion and LC-tandem mass spectrometry (MS/MS) bottom-up analysis to generate detailed sequences of proteolytic peptides; (2) off-line LC fractionation and subsequent top-down MS/MS to obtain high-quality fragmentation maps of intact peptides; and (3) on-line LC coupled to top-down MS/MS to allow rapid sequence analysis of low abundance peptides. Combining the three methods overcomes the limitations of each, and thus offers complementary and high-confidence determination of amino acid residues. We report the complete sequence analysis of six CHH-family neuropeptides including the discovery of two novel peptides. With the accurate molecular information, MALDI imaging and ion mobility MS were conducted for the first time to explore their anatomical distribution and biochemical properties.  相似文献   

18.
19.
20.
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号