Burkholderia Pseudomallei strain K96243 was a clinical isolate from Thailand. The genome of B. pseudomallei is 7 247 547 bp in size and the G+C content is 68.06%. It has 2 chromosome of 4 074 542 bp and 3 173 005 bp. The genome of B. pseudomallei carries many genomic islands as compared to its related organism, B. mallei, suggesting extensive horizontal transfer.
There are 3 different type III secretion systems (TTSS) are encoded on the chromosomes of this organism. Two of it are similar to plant pathogenic TTSSs, while the third is similar to the Salmonella pathogenicity island, all of which may contribute to pathogenicity. This organism also carries a number of small sequence repeats which may promoter antigenic variation, similar to what was found with the B. mallei genome.
The lineage of the B. pseudomallei is:
Bacteria > Proteobacteria > Betaproteobacteria > Burkholderiales > Burkholderiaceae > Burkholderia > Pseudomallei group > Burkholderia pseudomallei > Burkholderia pseudomallei K96243
reference:
1) http://www.ncbi.nlm.nih.gov/sites/entrez?db=genomeprj&cmd=Retrieve&dopt=Overview&list_uids=178
Wednesday, May 26, 2010
Sunday, May 23, 2010
AA & DNA sequence of P1
AA sequence of P1
MVKTFYITAAPVGAVPKFLDPLEPKFIPHALLELLPADAREATTQALEANGWEAVPAGGI
VREYGYDAPIDLTDYDGAQASASVQDALRNTGWTPCGTVWHRTQTSPSLAQPPLITRTTL
ERLSSVDLVRQIVLQLTTFGWTATEDGSLTWTHERIHSYLSPDFVERMRADKAAVLESLF
DNGWRVCGAGYWQPGKARSPYLPITADGIVDASREALREGAAVVHLHTRATDDQATLAIP
GLNTPIGIGSQRNHIVLDDYDRIVPTMLDLEPSAILNLSTSARGDRRASQSPLRRAHLKR
YGHAQLAPDVASFSPGPVVFQAGGGYDNPNAFLADQLAHFAEVGVRPEIEVFNHTIVENS
VTLYQSPLVKAGVPVLFMLVAAVDQYHRDPVSGDTSDDSLIDVPTRKAIAKLLQAGTDDA
HEKAVELAATQLRPTVDKLRDNFPSCKISLLLPGPFQALLVDVAIALDLDGIRVGLEDAL
NVFDARVPGGVRKACGTGDQVRWLRLELERRGIGIVDAEALRDELGMSRPDVALFRQAEA
ALAHYPADERLVSADTILDALRPIVDTYRKVEDRLATHLASAEALPADPAALAEHVLTAA
RSFGVTIRSFVEELDRYEDHEYLVARYIQVPQALNFARELLVPRGYSIDAYDRALEDYAR
PGKTVTREHASYSVRVDQFKPLPLRCLEYLVGIPCRYNGDYSNVVNLGLRQSPRYSATMA
LLYHALRELTLELRERSNASRKTCGPVWTVLETSANASEPPVRRDIAPDALTAAIDGVDW
VVLPSTPTTNYPLGLKLANGMAQLFHGFVAQIAADPTLRPSRQTHRDTPLRLLAITHSGR
RDDGETVIEASMLHNRFALNADPSGIYFSEESQLIYERLILPRLVDKPAKLAYNERQLVR
RDTAGFPLYQDGSRARRIKAEQIERLPFLKCFAHSSGIATAQQLDVQACRDGERLGLTAD
ELRAFFDRALLVSFGSAADIHLDWLGTSVVDVTAFNDVRSLAGTTSRHYLIQPGEHADVL
QHCLVHTQPADYRYDHATPVWQEGRQGKVVARLTGVFLLDDHARLDDGHSIRRYLAASPL
WLRQWIARFHDAPADAGAHAILRELQASMTDYRSSANQTTRRALA
DNA sequence of P1
ATGGTCAAGACCTTCTACATCACCGCGGCGCCCGTCGGCGCTGTCCCGAAGTTCCTCGAT
CCGCTGGAGCCGAAGTTCATCCCGCACGCATTGCTCGAGCTGCTGCCCGCCGACGCGCGC
GAGGCGACGACCCAGGCCCTCGAGGCCAACGGATGGGAGGCTGTGCCCGCAGGCGGCATC
GTGCGCGAATACGGCTACGACGCGCCGATCGATCTGACCGATTACGACGGCGCGCAAGCG
TCGGCGTCCGTGCAGGATGCGCTGCGCAACACCGGCTGGACACCGTGCGGCACGGTCTGG
CATCGCACGCAGACCTCGCCGTCGCTCGCGCAGCCGCCGCTGATCACGCGCACCACGCTC
GAGCGGCTATCGTCGGTCGATCTGGTCCGCCAGATCGTCCTGCAGCTCACGACGTTCGGC
TGGACCGCCACCGAAGACGGCAGCCTGACCTGGACGCACGAACGAATCCACTCCTATCTG
TCGCCGGACTTCGTCGAACGAATGCGCGCCGACAAAGCGGCCGTGCTCGAATCGCTGTTC
GACAACGGCTGGCGCGTGTGCGGCGCCGGCTACTGGCAGCCGGGCAAGGCGCGCTCGCCC
TATCTGCCGATCACCGCGGACGGCATCGTCGACGCATCGCGCGAGGCGCTGCGCGAAGGC
GCCGCGGTGGTCCACCTGCACACGCGCGCGACCGACGATCAGGCCACGCTCGCGATCCCG
GGGCTGAACACGCCGATCGGCATCGGCTCGCAGCGCAATCACATCGTGCTCGACGATTAC
GACCGAATCGTGCCGACGATGCTCGATCTGGAACCATCCGCCATCCTGAACCTGTCGACG
AGCGCGCGCGGCGATCGCCGCGCGTCGCAAAGCCCGCTGCGGCGTGCGCACCTGAAGCGC
TACGGCCACGCGCAGCTCGCGCCCGACGTCGCATCGTTCAGCCCCGGCCCCGTCGTGTTC
CAGGCGGGCGGCGGCTACGACAATCCGAACGCGTTCCTCGCGGATCAACTCGCTCACTTC
GCGGAAGTCGGCGTACGGCCCGAGATCGAGGTGTTCAACCATACGATCGTCGAGAACTCG
GTCACGCTCTATCAATCGCCGCTCGTGAAGGCCGGCGTTCCGGTGCTGTTCATGCTCGTC
GCCGCGGTCGACCAATACCACCGCGATCCCGTGAGCGGCGACACGAGCGACGATTCGCTG
ATCGACGTGCCCACCCGCAAGGCGATCGCGAAGCTGCTGCAGGCGGGCACCGACGACGCG
CACGAGAAGGCCGTCGAGCTCGCCGCGACGCAGTTGCGCCCGACCGTCGACAAGCTGCGC
GACAACTTCCCGTCGTGCAAGATCTCGCTGCTGCTGCCGGGCCCGTTCCAGGCGCTGCTC
GTCGACGTGGCGATCGCGCTCGATCTCGACGGCATTCGCGTCGGGCTCGAGGACGCGCTG
AACGTCTTCGACGCGCGCGTGCCGGGCGGCGTGCGCAAGGCGTGCGGCACCGGCGATCAG
GTGCGCTGGCTGCGGCTCGAGCTCGAGCGCCGCGGCATCGGCATCGTCGACGCCGAAGCG
CTGCGCGACGAACTCGGCATGTCGCGGCCCGACGTCGCGCTGTTCCGTCAAGCCGAAGCC
GCGCTCGCGCATTATCCGGCCGACGAGCGGCTCGTATCGGCCGACACGATCCTCGACGCG
CTGCGGCCGATCGTCGACACGTACCGCAAGGTCGAGGATCGACTCGCCACCCACCTCGCG
AGCGCCGAAGCGCTGCCGGCGGACCCCGCCGCGCTCGCCGAGCACGTACTGACGGCCGCG
CGCAGCTTCGGCGTGACGATCCGCTCGTTCGTCGAAGAACTCGATCGCTACGAGGACCAC
GAGTATCTGGTCGCGCGCTACATTCAGGTTCCGCAGGCGCTGAACTTCGCGCGCGAGCTG
CTCGTGCCGCGCGGCTATTCGATCGATGCGTACGACCGCGCGCTCGAGGACTACGCGCGC
CCGGGCAAGACCGTGACGCGCGAGCACGCGAGCTACAGCGTGCGCGTCGACCAGTTCAAG
CCGCTGCCGCTGCGCTGCCTCGAATATCTGGTGGGAATTCCTTGCCGCTACAACGGCGAC
TACAGCAATGTCGTCAATCTCGGCTTGCGCCAGAGCCCGCGCTACAGCGCGACGATGGCG
CTGCTCTATCACGCGCTGCGCGAGCTCACGCTCGAGTTGCGCGAGCGCTCGAATGCGTCG
CGCAAGACATGCGGCCCCGTGTGGACCGTGCTCGAGACCTCGGCGAACGCAAGCGAGCCG
CCCGTGCGCCGCGATATCGCGCCCGACGCGCTGACAGCCGCGATCGACGGCGTCGACTGG
GTCGTGCTGCCGAGCACGCCGACCACCAACTACCCGCTCGGCCTCAAGCTGGCGAATGGG
ATGGCGCAGCTGTTCCACGGCTTCGTCGCGCAGATCGCCGCCGATCCGACGCTGCGCCCG
TCGCGGCAGACGCACCGCGACACGCCGCTGCGCCTGCTCGCGATCACGCATTCGGGCCGC
CGCGACGACGGCGAAACGGTGATCGAAGCCAGCATGCTGCACAACCGCTTTGCGCTGAAC
GCGGATCCGTCGGGCATCTACTTCAGCGAGGAGTCGCAACTGATCTACGAGCGGCTGATC
CTGCCGCGCCTCGTCGACAAGCCCGCCAAGCTCGCATACAACGAGCGGCAACTGGTGCGC
CGGGACACGGCCGGCTTCCCGCTGTATCAGGACGGTTCGCGCGCGCGCCGCATCAAGGCC
GAGCAGATCGAGCGGCTGCCGTTTCTCAAGTGCTTCGCGCACAGCTCGGGGATCGCCACC
GCGCAGCAGCTCGACGTGCAGGCATGCCGCGACGGCGAACGGCTCGGCCTCACGGCCGAC
GAACTCCGCGCGTTCTTCGATCGCGCGCTGCTGGTGTCGTTCGGCTCCGCCGCGGACATC
CACCTCGACTGGCTCGGCACCTCGGTCGTCGACGTGACCGCGTTCAACGACGTGCGCAGC
CTCGCCGGCACGACGAGCCGCCATTACTTGATCCAGCCGGGCGAGCATGCGGACGTGCTG
CAGCACTGCCTCGTGCACACGCAGCCCGCCGACTATCGCTACGATCACGCGACGCCGGTC
TGGCAGGAAGGCCGGCAGGGCAAGGTCGTCGCGCGGCTCACGGGCGTGTTCCTGCTCGAC
GATCATGCGCGGCTCGACGACGGCCACTCGATTCGCCGCTACCTCGCCGCGAGCCCGCTG
TGGCTGCGTCAATGGATCGCGCGTTTTCACGACGCGCCCGCCGACGCCGGCGCACACGCG
ATTCTGCGCGAACTGCAGGCGTCGATGACCGACTACCGGTCGAGCGCGAACCAGACGACG
CGGCGAGCACTCGCGTAA
Thursday, May 20, 2010
The accession number of BP
Accession number can be written based on two sources; one is from GenBank and the other is from RefSeq.
The accession number is the unique identifier assigned to the entire sequence record when the record is submitted to GenBank. The GenBank accession number is a combination of letters and numbers that are usually in the format of one letter followed by five digits (e.g., M12345) or two letters followed by six digits (e.g., AC123456). The accession number for a particular record will not change even if the author submits a request to change some of the information in the record. GenBank is an annotated collection of all publicly available DNA sequences.
This accession number is the unique identification number for a complete RefSeq sequence record. RefSeq accession numbers are written in the following format: two letters followed by an underscore and six digits (e.g., NT_123456). Refseq are derived from GenBank records but differ in that each RefSeq is a synthesis of information,not an archived unit of primary research data.
The accession number for the chromosome 1 of burkholderia pseudomallei is BX571965 and for the chromosome 2, the accession number is BX571966.
The P1 protein, the accession number is YP_111366. The sequences of P1 protein is found to be on chromosome 2 with the accession number of NC_006351 from DBsource.
REFERENCES:
1. http://www.ornl.gov/sci/techresources/Human_Genome/posters/chromosome/genejargon.shtml
2. http://www.ncbi.nlm.nih.gov/bookshelf/br.fcgi?book=handbook&part=ch18
3. http://www.ncbi.nlm.nih.gov/genbank/
The accession number is the unique identifier assigned to the entire sequence record when the record is submitted to GenBank. The GenBank accession number is a combination of letters and numbers that are usually in the format of one letter followed by five digits (e.g., M12345) or two letters followed by six digits (e.g., AC123456). The accession number for a particular record will not change even if the author submits a request to change some of the information in the record. GenBank is an annotated collection of all publicly available DNA sequences.
This accession number is the unique identification number for a complete RefSeq sequence record. RefSeq accession numbers are written in the following format: two letters followed by an underscore and six digits (e.g., NT_123456). Refseq are derived from GenBank records but differ in that each RefSeq is a synthesis of information,not an archived unit of primary research data.
The accession number for the chromosome 1 of burkholderia pseudomallei is BX571965 and for the chromosome 2, the accession number is BX571966.
The P1 protein, the accession number is YP_111366. The sequences of P1 protein is found to be on chromosome 2 with the accession number of NC_006351 from DBsource.
REFERENCES:
1. http://www.ornl.gov/sci/techresources/Human_Genome/posters/chromosome/genejargon.shtml
2. http://www.ncbi.nlm.nih.gov/bookshelf/br.fcgi?book=handbook&part=ch18
3. http://www.ncbi.nlm.nih.gov/genbank/
Wednesday, May 19, 2010
Post all info & knowledge on P1 here
Every little scrap of info will be centralise here. The biology of P1 is unknown - we are the 10 blind men with an "elephant" in the room.
Subscribe to:
Posts (Atom)