Gene Information

Name : ECP_0345 (ECP_0345)
Accession : YP_668278.1
Strain : Escherichia coli 536
Genome accession: NC_008253
Putative virulence/resistance : Virulence
Product : protein SepA
Function : -
COG functional category : M : Cell wall/membrane/envelope biogenesis
COG ID : COG3468
EC number : -
Position : 364050 - 368180 bp
Length : 4131 bp
Strand : -
Note : extracellular serine protease of the IgA1 protease, secreted by a C-terminal autotransporter domain

DNA sequence :
ATGAATAAAATATACGCTCTAAAATATTGTTATATTACTAACACAGTAAAGGTTGTCTCTGAACTAGCCCGAAGGGTATG
TAAAGGGAGTACCCGCAGAGGAAAAAGACTTTCAGTACTTACCTCTCTGGCACTATCTGCATTACTCCCAACCGTTGCTG
GTGCATCAACGGTTGGTGGCAACAATCCTTACCAGACATACCGCGACTTTGCAGAAAACAAAGGGCAGTTTCAGGCTGGC
GCAACAAACATTCCTATTTTTAATAATAAAGGGGAATTAGTAGGACATCTTGATAAAGCGCCCATGGTTGATTTTAGCAG
TGTGAATGTAAGCTCAAATCCCGGCGTTGCAACATTAATTAACCCGCAATATATAGCCAGTGTAAAACATAATAAAGGAT
ATCAGAGCGTCAGCTTCGGTGATGGTCAGAACAGTTACCATATTGTGGATCGTAATGAACACAGTTCATCTGATCTCCAC
ACACCAAGACTTGATAAGCTCGTAACTGAGGTTGCTCCGGCTACCGTAACCAGCTCATCAACAGCTGATATATTGACCCC
TTCAAAATACTCGGCATTCTACAGGGCTGGTTCGGGAAGTCAGTATATTCAGGATAGTCAGGGTAAGCGACATTGGGTAA
CAGGTGGGTATGGTTATCTGACAGGAGGAATACTCCCGACATCATTCTTTTATCACGGCTCAGACGGCATTCAGCTGTAT
ATGGGGGGCAACATACATGATCATAGCATCCTGCCCTCTTTTGGAGAGGCCGGCGACAGTGGTTCTCCATTATTTGGCTG
GAATACGGCCAAAGGGCAGTGGGAACTGGTCGGTGTTTACTCGGGAGTAGGAGGGGGGACCAATTTGATATATTCTCTTA
TTCCTCAGAGTTTTCTCTCACAGATCTATTCAGAGGATAATGACGCTCCCGTCTTTTTTAATGCCTCATCCGGCGCCCCC
CTGCAATGGAAATTTGACAGCAGCACCGGCACTGGCTCTCTGAAACAGGGTTCCGATGAATATGCCATGCACGGGCAAAA
AGGTTCTGACCTGAACGCAGGTAAAAATCTGACATTCCTGGGACATAATGGTCAGATTGACCTGGAAAACTCTGTCACGC
AGGGTGCCGGTTCACTGACATTTACTGATGACTACACTGTCACCACTTCAAACGGAAGTACCTGGACCGGGGCCGGTATT
ATTGTGGACAAGGATGCCTCCGTAAACTGGCAGGTTAATGGTGTGAAAGGTGACAACCTGCATAAAATCGGCGAAGGAAC
CCTGGTTGTACAGGGAACCGGTGTTAATGAGGGCGGCCTGAAAGTCGGGGATGGGACCGTTGTCCTCAATCAGCAGGCTG
ACAGTTCAGGACACGTTCAGGCATTCAGTAGCGTGAATATTGCCAGCGGCCGCCCGACAGTCGTGCTGGCAGACAACCAG
CAGGTTAATCCGGACAATATATCCTGGGGCTACCGGGGGGGTGTTCTGGATGTTAACGGGAATGACCTGACATTTCATAA
GCTGAATGCCGCCGATTATGGCGCAACTCTCGGTAACAGCAGTGATAAAACGGCTAATATCACTCTGGATTATCAGACGC
ATCCGGCAGACGTAAAAGTTAATGAATGGTCATCATCAAACAGGGGAACAGTAGGTTCATTATATATTTATAATAATCCC
TATACTCATACCGTCGATTATTTTATCCTGAAAACAAGTAGTTATGGCTGGTTCCCTACCGGTCAGGTCAGTAACGAGCA
CTGGGAATATGTCGGACATGACCAGAACAGTGCACAGGCGCTGCTTGCAAACAGAATTAATAATAAAGGGTATCTGTATC
ATGGCAAGTTGCTGGGAAATATTAATTTCTCAAATAAAGCAACCCCGGGTACAACCGGCGCATTGGTTATGGACGGCTCA
GCGAATATGTCCGGTACATTTACTCAGGAAAACGGTCGTCTGACCATTCAGGGCCACCCGGTTATCCATGCTTCAACGTC
TCAGAGTATTGCAAATACAGTCTCGTCTCTGGGCGACAATTCCGTTCTGACACAGCCCACCTCATTTACACAGGATGACT
GGGAGAACAGGACGTTCAGCTTTGGTTCGCTCGTGTTAAAAGATACAGACTTTGGTCTGGGCCGCAATGCCACACTGAAC
ACAACCATCCAGGCAGATAACTCCAGCGTCACGCTGGGCGACAGTCGGGTATTTATCGACAAAAAAGATGGCCAGGGAAC
AGCCTTTACCCTTGAAGAAGGCACATCTGTTGCAACTAAAGATGCAGATAAAAGTGTCTTCAACGGCACCGTCAACCTGG
ATAATCAGTCAGTGCTGAATATCAATGATATATTCAATGGCGGAATACAGGCGAACAACAGTACCGTGAATATCTCCTCA
GACAGTGCCATTCTGGGGAACTCAACGCTGACCAGTACCGCCCTGAATCTGAACAAGGGAGCAAATGCTCTGGCCAGTCA
GAGTTTTGTTTCTGACGGTCCAGTGAATATTTCTGATGCCACCCTGAGTCTGAACAGCCGTCCTGATGAGGTATCTCACA
CACTTTTACCTGTATACGATTATGCCGGTTCATGGAACCTGAAGGGAGACGATGCCCGCCTGAACGTGGGGCCGTACAGT
ATGTTGTCAGGTAATATCAATGTTCAGGATAAAGGGACTGTCACCCTCGGAGGGGAAGGGGAACTGAGTCCTGACCTGAC
TCTTCAGAATCAGATGTTGTACAGCCTGTTTAACGGGTACCGCAATACCTGGAGCGGGAGCCTGAATGCACCGGATGCCA
CCGTCAGCATGACAGACACCCAGTGGTCGATGAACGGAAACTCCACGGCAGGAAATATGAAACTTAACCGGACAATAGTC
GGTTTTAACGGGGGAACATCATCGTTCACGACACTGACAACAGATAATCTGGACGCGGTTCAGTCAGCATTTGTCATGCG
TACAGACCTTAACAAGGCAGACAAACTGGTGATAAACAAGTCGGCAACAGGTCATGACAACAGCATCTGGGTTAACTTCC
TGAAAAAACCCTCTGACAAGGACACGCTTGATATTCCACTGGTCAGCGCACCTGAAGCGACAGCTGATAATCTGTTCAGG
GCATCAACACGGGTTGTGGGATTCAGTGATGTCACCCCCACCCTTAGTGTCAGAAAAGAGGACGGGAAAAAAGAGTGGGT
CCTCGATGGTTACCAGGTTGCACGTAACGACGGCCAGGGTAAGGCTGCCGCCACATTCATGCACATCAGCTATAACAACT
TCATCACTGAAGTTAACAACCTGAACAAACGCATGGGCGATTTGAGGGATATTAACGGCGAAGCCGGTACGTGGGTGCGT
CTGCTGAACGGTTCCGGCTCTGCTGATGGCGGTTTCACTGACCACTATACCCTGCTGCAGATGGGGGCTGACCGTAAGCA
CGAACTGGGAAGTATGGACCTGTTTACCGGCGTGATGGCCACCTACACTGACACAGATGCGTCAGCAGGCCTGTACAGCG
GTAAAACAAAATCATGGGGTGGTGGTTTCTATGCCAGTGGTCTGTTCCGGTCCGGCGCTTACTTTGATTTGATTGCCAAA
TATATTCACAATGAAAACAAATATGACCTGAACTTTGCCGGAGCTGGTAAACAGAACTTCCGCAGCCATTCACTGTATGC
AGGTGCAGAAGTCGGATACCGTTATCATCTGACAGATACGACGTTTGTTGAACCTCAGGCGGAACTGGTCTGGGGAAGAC
TGCAGGGCCAAACATTTAACTGGAACGACAGTGGAATGGATGTCTCAATGCGTCGTAACAGCGTTAATCCTCTGGTAGGC
AGAACCGGCGTTGTTTCCGGTAAAACCTTCAGTGGTAAGGACTGGAGTCTGACAGCCCGTGCCGGCCTACATTATGAGTT
CGATCTGACGGACAGTGCTGACGTTCACCTGAAGGATGCAGCGGGAGAACATCAGATTAATGGGAGAAAAGACGGTCGTA
TGCTTTACGGTGTGGGGTTAAATGCCCGGTTTGGCGACAATACGCGTCTGGGGCTGGAAGTTGAACGCTCTGCATTCGGT
AAATACAACACAGATGATGCGATAAACGCTAACATTCGTTATTCATTCTGA

Protein sequence :
MNKIYALKYCYITNTVKVVSELARRVCKGSTRRGKRLSVLTSLALSALLPTVAGASTVGGNNPYQTYRDFAENKGQFQAG
ATNIPIFNNKGELVGHLDKAPMVDFSSVNVSSNPGVATLINPQYIASVKHNKGYQSVSFGDGQNSYHIVDRNEHSSSDLH
TPRLDKLVTEVAPATVTSSSTADILTPSKYSAFYRAGSGSQYIQDSQGKRHWVTGGYGYLTGGILPTSFFYHGSDGIQLY
MGGNIHDHSILPSFGEAGDSGSPLFGWNTAKGQWELVGVYSGVGGGTNLIYSLIPQSFLSQIYSEDNDAPVFFNASSGAP
LQWKFDSSTGTGSLKQGSDEYAMHGQKGSDLNAGKNLTFLGHNGQIDLENSVTQGAGSLTFTDDYTVTTSNGSTWTGAGI
IVDKDASVNWQVNGVKGDNLHKIGEGTLVVQGTGVNEGGLKVGDGTVVLNQQADSSGHVQAFSSVNIASGRPTVVLADNQ
QVNPDNISWGYRGGVLDVNGNDLTFHKLNAADYGATLGNSSDKTANITLDYQTHPADVKVNEWSSSNRGTVGSLYIYNNP
YTHTVDYFILKTSSYGWFPTGQVSNEHWEYVGHDQNSAQALLANRINNKGYLYHGKLLGNINFSNKATPGTTGALVMDGS
ANMSGTFTQENGRLTIQGHPVIHASTSQSIANTVSSLGDNSVLTQPTSFTQDDWENRTFSFGSLVLKDTDFGLGRNATLN
TTIQADNSSVTLGDSRVFIDKKDGQGTAFTLEEGTSVATKDADKSVFNGTVNLDNQSVLNINDIFNGGIQANNSTVNISS
DSAILGNSTLTSTALNLNKGANALASQSFVSDGPVNISDATLSLNSRPDEVSHTLLPVYDYAGSWNLKGDDARLNVGPYS
MLSGNINVQDKGTVTLGGEGELSPDLTLQNQMLYSLFNGYRNTWSGSLNAPDATVSMTDTQWSMNGNSTAGNMKLNRTIV
GFNGGTSSFTTLTTDNLDAVQSAFVMRTDLNKADKLVINKSATGHDNSIWVNFLKKPSDKDTLDIPLVSAPEATADNLFR
ASTRVVGFSDVTPTLSVRKEDGKKEWVLDGYQVARNDGQGKAAATFMHISYNNFITEVNNLNKRMGDLRDINGEAGTWVR
LLNGSGSADGGFTDHYTLLQMGADRKHELGSMDLFTGVMATYTDTDASAGLYSGKTKSWGGGFYASGLFRSGAYFDLIAK
YIHNENKYDLNFAGAGKQNFRSHSLYAGAEVGYRYHLTDTTFVEPQAELVWGRLQGQTFNWNDSGMDVSMRRNSVNPLVG
RTGVVSGKTFSGKDWSLTARAGLHYEFDLTDSADVHLKDAAGEHQINGRKDGRMLYGVGLNARFGDNTRLGLEVERSAFG
KYNTDDAINANIRYSF

• Homologs from PAI DB

GeneGenBank Accn Product Virulance or Resistance PAI or REI Alignment Type E-val Identity
unnamed CAD66214.1 putative hemoglobin protease Not tested PAI III 536 Protein 0.0 100
vat YP_851472.1 vacuolating autotransporter Not tested PAI III APEC-O1 Protein 0.0 99
vat AAO21903.1 vacuolating autotransporter toxin Virulence Not named Protein 0.0 97
pic NP_838464.1 serine protease precurser Virulence SHI-1 Protein 0.0 49
pic NP_708747.3 serine protease Not tested SHI-1 Protein 0.0 48
she AAB58244.1 mucinase Virulence SHI-1 Protein 0.0 48
pic AAK00464.1 Pic Virulence SHI-1 Protein 0.0 48
unnamed CAC39286.1 hypothetical protein Not tested LPA Protein 0.0 44

• Homologs from VFDB (virulence genes)

GeneGenBank Accn Product ID of source DB Alignment Type E-val Identity
ECP_0345 YP_668278.1 protein SepA VFG1689 Protein 0.0 100
ECP_0345 YP_668278.1 protein SepA VFG0904 Protein 0.0 99
ECP_0345 YP_668278.1 protein SepA VFG0635 Protein 0.0 48
ECP_0345 YP_668278.1 protein SepA VFG0861 Protein 0.0 48
ECP_0345 YP_668278.1 protein SepA VFG0903 Protein 0.0 48