PAI Gene Information


Name : sigA (S4824)
Accession : NP_838462.1
PAI name : SHI-1
PAI accession : NC_004741_P1
Strain : Shigella flexneri 2002017
Virulence or Resistance: Virulence
Product : serine protease
Function : -
Note : residues 1 to 1285 of 1285 are 100.00 pct identical to residues 1 to 1285 of 1285 from NRprotein_Feb03 : >ref|NP_708742.1| exported serine protease SigA [Shigella flexneri 2a str. 301] gb|AAF67320.1|AF200692_1 exported serine protease SigA [Shigella flexn
Homologs in the searched genomes :   30 hits    ( 30 protein-level )  
Publication :
    -Wei,J., Goldberg,M.B., Burland,V., Venkatesan,M.M., Deng,W., Fournier,G., Mayhew,G.F., Plunkett,G. III, Rose,D.J., Darling,A., Mau,B., Perna,N.T., Payne,S.M., Runyen-Janecky,L.J., Zhou,S., Schwartz,D.C. and Blattner,F.R., "Complete genome sequence and comparative genomics of Shigella flexneri serotype 2a strain 2457T", Infect. Immun. 71 (5), 2775-2786 (2003) PUBMED 12704152 REMARK Erratum:[Infect Immun. 2003 Jul;71(7):4223].

    -Wei,J., Goldberg,M.B., Burland,V., Venkatesan,M.M., Deng,W., Fournier,G., Mayhew,G.F., Plunkett,G. III, Rose,D.J., Darling,A., Mau,B., Perna,N.T., Payne,S.M., Runyen-Janecky,L.J., Zhou,S., Schwartz,D.C. and Blattner,F.R., "Direct Submission", Submitted (23-APR-2003) National Center for Biotechnology Information, NIH, Bethesda, MD 20894, USA.

    -Wei,J., Goldberg,M.B., Burland,V., Venkatesan,M.M., Deng,W., Fournier,G., Mayhew,G.F., Plunkett,G. III, Rose,D.J., Darling,A., Mau,B., Perna,N.T., Payne,S.M., Runyen-Janecky,L.J., Zhou,S., Schwartz,D.C. and Blattner,F.R., "Direct Submission", Submitted (13-JUN-2002) Genetics Laboratory, University of Wisconsin - Madison, 445 Henry Mall, Madison, WI 53706, USA.


DNA sequence :
ATGAATAAAATTTATTCACTGAAATATAGTCATATTACAGGTGGATTAGTTGCTGTTTCTGAACTGACCCGGAAAGTTAG
TGTCGGTACATCAAGAAAGAAAGTTATCCTCGGTATTATTTTATCCTCAATATATGGAAGTTATGGCGAAACAGCATTTG
CAGCAATGCTGGATATAAATAATATATGGACCCGCGATTATCTTGACCTTGCTCAAAACAGAGGAGAGTTCAGACCGGGT
GCAACAAATGTTCAATTAATGATGAAAGATGGAAAGATATTTCATTTTCCAGAACTACCTGTACCTGATTTTTCTGCTGT
TTCCAACAAAGGTGCAACAACATCAATTGGAGGTGCGTACAGTGTTACTGCGACTCATAACGGTACACAGCATCATGCAA
TAACAACACAGTCATGGGATCAGACAGCATATAAAGCAAGTAACAGAGTATCATCTGGCGACTTTTCGGTTCATCGTCTG
AATAAATTCGTCGTGGAAACAACAGGGGTTACGGAGAGTGCCGACTTCTCACTTTCTCCCGAAGATGCGATGAAAAGATA
TGGCGTAAACTACAACGGTAAGGAACAAATAATTGGCTTCAGAGCAGGTGCCGGAACAACCTCAACGATATTAAACGGCA
AACAATATCTGTTTGGACAAAACTATAATCCCGACTTGTTAAGCGCAAGTCTTTTTAATCTGGACTGGAAAAACAAGAGT
TACATTTATACCAACAGAACCCCTTTTAAAAACTCACCAATTTTTGGCGATAGTGGTTCTGGTTCTTATCTATATGATAA
AGAACAACAAAAATGGGTTTTCCATGGTGTTACCAGTACAGTTGGTTTTATCAGTAGTACCAATATAGCCTGGACAAACT
ACTCGTTATTTAATAATATTCTGGTAAACAATTTAAAAAAGAATTTCACAAACACTATGCAGCTGGATGGTAAAAAACAA
GAGTTATCATCGATTATAAAAGATAAGGACCTGTCTGTCTCAGGAGGAGGGGTATTAACGCTCAAGCAGGATACCGATCT
TGGCATTGGCGGGCTTATATTCGATAAGAACCAGACATATAAAGTGTACGGAAAAGATAAGTCTTATAAAGGTGCCGGGA
TAGATATTGATAATAATACCACCGTTGAATGGAATGTTAAGGGCGTTGCCGGAGATAATCTGCATAAAATAGGTAGTGGT
ACTCTGGATGTAAAAATAGCACAGGGAAATAACCTTAAAATAGGTAATGGGACTGTCATCCTTAGTGCTGAAAAAGCCTT
CAATAAAATTTACATGGCCGGAGGTAAAGGTACGGTAAAAATAAATGCCAAAGACGCTTTAAGCGAAAGCGGTAATGGCG
AAATCTATTTTACCAGAAATGGCGGAACACTGGATCTAAACGGCTATGACCAGTCATTTCAGAAAATCGCAGCAACAGAT
GCGGGAACAACCGTAACGAACTCAAACGTGAAGCAATCAACATTATCACTTACTAATACTGATGCATATATGTACCATGG
GAATGTATCAGGTAATATAAGCATAAATCATATTATCAATACTACCCAGCAACATAACAATAATGCCAATCTGATCTTTG
ATGGCTCAGTCGATATCAAAAACGATATCTCTGTCCGGAATGCACAGTTAACATTACAAGGACATGCGACAGAACATGCC
ATATTTAAAGAAGGCAATAACAACTGTCCAATTCCTTTTTTATGTCAAAAAGACTATTCTGCTGCCATAAAGGACCAGGA
AAGCACTGTAAATAAACGTTACAATACGGAATATAAGTCCAACAATCAGATAGCCTCTTTTTCCCAGCCCGACTGGGAAA
GTCGTAAATTTAATTTCCGGAAATTAAATTTAGAAAACGCAACCCTGAGTATAGGCCGGGATGCTAATGTAAAAGGACAC
ATAGAGGCTAAAAACTCTCAAATTGTTCTGGGAAATAAAACTGCATACATTGACATGTTCTCAGGAAGAAACATTACTGG
CGAAGGTTTTGGATTCAGACAACAGCTTCGCTCCGGGGATTCAGCAGGCGAAAGTAGTTTCAACGGCAGTCTGAGTGCTC
AAAACAGCAAAATAACTGTTGGTGATAAATCAACTGTTACTATGACTGGTGCATTATCCTTAATTAATACAGACCTGATT
ATCAACAAAGGAGCTACTGTTACCGCCCAGGGAAAAATGTATGTAGATAAAGCTATTGAACTGGCCGGAACCCTGACATT
AACAGGCACCCCTACAGAAAATAATAAATACAGCCCGGCAATCTATATGTCAGATGGATATAATATGACAGAAGATGGTG
CCACGTTAAAGGCTCAAAATTATGCCTGGGTCAATGGTAATATAAAATCAGACAAAAAAGCATCTATTCTGTTTGGTGTT
GACCAGTATAAAGAAGATAACCTGGACAAAACCACACACACACCGCTGGCTACAGGTTTGCTGGGTGGCTTTGATACTTC
TTATACCGGAGGTATTGATGCTCCTGCAGCCTCAGCCAGCATGTATAACACCTTATGGAGAGTAAACGGACAGTCAGCCC
TGCAATCATTAAAAACCCGCGACAGTCTTTTGTTGTTTAGTAACATAGAGAATTCGGGTTTCCATACTGTGACAGTAAAC
ACACTGGATGCCACTAATACTGCTGTGATTATGCGGGCTGATCTGAGCCAGTCTGTAAATCAATCGGATAAACTCATTGT
TAAAAATCAGTTAACCGGAAGCAATAACAGTCTGTCGGTCGATATACAGAAAGTGGGAAATAATAACTCAGGATTAAACG
TTGACCTGATAACAGCCCCAAAAGGAAGCAATAAAGAGATATTTAAAGCCAGTACTCAGGCCATAGGTTTCAGCAACATA
TCTCCTGTGATCAGCACGAAAGAGGATCAGGAACATACCACGTGGACCCTGACCGGATATAAGGTGGCTGAAAATACAGC
ATCTTCCGGTGCAGCAAAATCGTATATGTCCGGTAATTACAAAGCCTTCCTGACAGAAGTCAACAACCTGAATAAACGAA
TGGGGGATCTGCGTGACACCAATGGCGAGGCCGGTGCATGGGCCCGCATCATGAGCGGAGCAGGTTCAGCTTCTGGTGGA
TACAGTGACAACTACACCCATGTGCAGATTGGTGTGGATAAAAAACATGAGCTGGATGGACTTGACCTTTTCACTGGTCT
GACTATGACGTATACCGACAGTCATGCCAGCAGTAATGCATTCAGTGGCAAGACGAAGTCCGTCGGGGCAGGTCTGTATG
CTTCCGCTATATTTGACTCTGGTGCCTATATCGACCTGATTAGTAAGTATGTTCACCATGATAATGAGTACTCGGCGACC
TTTGCTGGACTCGGAACAAAAGACTACAGTTCTCATTCCTTGTATGTGGGTGCTGAAGCAGGCTACCGCTATCATGTAAC
AGAAGACTCCTGGATTGAGCCGCAGGCAGAACTGGTTTATGGGGCCGTATCAGGTAAACGGTTCGACTGGCAGGATCGCG
GAATGAGCGTGACCATGAAGGATAAGGACTTTAATCCGCTGATTGGGCGTACCGGTGTTGATGTGGGTAAATCCTTCTCC
GGTAAGGACTGGAAAGTCACAGCCCGCGCCGGCCTTGGCTACCAGTTTGACCTGTTTGCCAACGGTGAAACCGTACTGCG
TGATGCGTCCGGTGAGAAACGTATCAAAGGTGAAAAAGACGGTCGTATTCTCATGAATGTTGGTCTCAACGCCGAAATTC
GCGATAATCTTCGCTTCGGTCTTGAGTTTGAGAAATCGGCATTTGGTAAATACAACGTGGATAACGCGATCAACGCCAAC
TTCCGTTACTCTTTCTGA

Protein sequence :
MNKIYSLKYSHITGGLVAVSELTRKVSVGTSRKKVILGIILSSIYGSYGETAFAAMLDINNIWTRDYLDLAQNRGEFRPG
ATNVQLMMKDGKIFHFPELPVPDFSAVSNKGATTSIGGAYSVTATHNGTQHHAITTQSWDQTAYKASNRVSSGDFSVHRL
NKFVVETTGVTESADFSLSPEDAMKRYGVNYNGKEQIIGFRAGAGTTSTILNGKQYLFGQNYNPDLLSASLFNLDWKNKS
YIYTNRTPFKNSPIFGDSGSGSYLYDKEQQKWVFHGVTSTVGFISSTNIAWTNYSLFNNILVNNLKKNFTNTMQLDGKKQ
ELSSIIKDKDLSVSGGGVLTLKQDTDLGIGGLIFDKNQTYKVYGKDKSYKGAGIDIDNNTTVEWNVKGVAGDNLHKIGSG
TLDVKIAQGNNLKIGNGTVILSAEKAFNKIYMAGGKGTVKINAKDALSESGNGEIYFTRNGGTLDLNGYDQSFQKIAATD
AGTTVTNSNVKQSTLSLTNTDAYMYHGNVSGNISINHIINTTQQHNNNANLIFDGSVDIKNDISVRNAQLTLQGHATEHA
IFKEGNNNCPIPFLCQKDYSAAIKDQESTVNKRYNTEYKSNNQIASFSQPDWESRKFNFRKLNLENATLSIGRDANVKGH
IEAKNSQIVLGNKTAYIDMFSGRNITGEGFGFRQQLRSGDSAGESSFNGSLSAQNSKITVGDKSTVTMTGALSLINTDLI
INKGATVTAQGKMYVDKAIELAGTLTLTGTPTENNKYSPAIYMSDGYNMTEDGATLKAQNYAWVNGNIKSDKKASILFGV
DQYKEDNLDKTTHTPLATGLLGGFDTSYTGGIDAPAASASMYNTLWRVNGQSALQSLKTRDSLLLFSNIENSGFHTVTVN
TLDATNTAVIMRADLSQSVNQSDKLIVKNQLTGSNNSLSVDIQKVGNNNSGLNVDLITAPKGSNKEIFKASTQAIGFSNI
SPVISTKEDQEHTTWTLTGYKVAENTASSGAAKSYMSGNYKAFLTEVNNLNKRMGDLRDTNGEAGAWARIMSGAGSASGG
YSDNYTHVQIGVDKKHELDGLDLFTGLTMTYTDSHASSNAFSGKTKSVGAGLYASAIFDSGAYIDLISKYVHHDNEYSAT
FAGLGTKDYSSHSLYVGAEAGYRYHVTEDSWIEPQAELVYGAVSGKRFDWQDRGMSVTMKDKDFNPLIGRTGVDVGKSFS
GKDWKVTARAGLGYQFDLFANGETVLRDASGEKRIKGEKDGRILMNVGLNAEIRDNLRFGLEFEKSAFGKYNVDNAINAN
FRYSF