Gene Information

Name : ECABU_c03670 (ECABU_c03670)
Accession : YP_006104502.1
Strain : Escherichia coli ABU 83972
Genome accession: NC_017631
Putative virulence/resistance : Virulence
Product : IgA-specific serine endopeptidase precursor
Function : -
COG functional category : -
COG ID : -
EC number : -
Position : 381696 - 385826 bp
Length : 4131 bp
Strand : -
Note : -

DNA sequence :
ATGAATAAAATATACGCTCTAAAATATTGTTATATTACTAACACAGTAAAGGTTGTCTCTGAACTAGCCCGAAGGGTATG
TAAAGGGAGTACCCGCAGAGGAAAAAGACTTTCAGTACTTACCTCTCTGGCACTATCTGCATTACTCCCAACCGTTGCTG
GTGCATCAACGGTTGGTGGCAACAATCCTTACCAGACATACCGCGACTTTGCAGAAAACAAAGGGCAGTTTCAGGCTGGC
GCAACAAACATTCCTATTTTTAATAATAAAGGGGAATTAGTAGGACATCTTGATAAAGCGCCCATGGTTGATTTTAGCAG
TGTGAATGTAAGCTCAAATCCCGGCGTTGCAACATTAATTAACCCGCAATATATAGCCAGTGTAAAACATAATAAAGGAT
ATCAGAGCGTCAGCTTCGGTGATGGTCAGAACAGTTACCATATTGTGGATCGTAATGAACACAGTTCATCTGATCTCCAC
ACACCAAGACTTGATAAGCTCGTAACTGAGGTTGCTCCGGCTACCGTAACCAGCTCATCAACAGCTGATATATTGAACCC
TTCAAAATACTCGGCATTCTACAGGGCTGGTTCGGGAAGTCAGTATATTCAGGATAGTCAGGGTAAGCGACATTGGGTAA
CAGGTGGGTATGGTTATCTGACAGGAGGAATACTCCCGACATCATTCTTTTATCACGGCTCAGACGGCATTCAGCTGTAT
ATGGGGGGCAACATACATGATCATAGCATCCTGCCCTCTTTTGGAGAGGCCGGCGACAGTGGTTCTCCATTATTTGGCTG
GAATACGGCCAAAGGGCAGTGGGAACTGGTCGGTGTTTACTCGGGAGTAGGAGGGGGGACCAATTTGATATATTCTCTTA
TTCCTCAGAGTTTTCTCTCTCAGATCTATTCAGAGGATAATGACGCTCCCGTCTTTTTTAATGCCTCATCCGGCGCCCCC
CTGCAATGGAAATTTGACAGCAGCACCGGCACTGGCTCTCTGAAACAGGGTTCCGATGAATATGCCATGCACGGGCAAAA
AGGTTCTGACCTGAACGCAGGTAAAAATCTGACATTCCTGGGACATAATGGTCAGATTGACCTGGAAAACTCTGTCACGC
AGGGTGCCGGTTCACTGACATTTACTGATGACTACACTGTCACCACTTCAAACGGAAGTACCTGGACCGGGGCCGGTATT
ATTGTGGACAAGGATGCCTCCGTAAACTGGCAGGTTAATGGTGTGAAAGGTGACAACCTGCATAAAATCGGCGAAGGAAC
CCTGGTTGTACAGGGAACCGGTGTTAATGAGGGCGGCCTGAAAGTCGGGGATGGGACCGTTGTCCTCAATCAGCAGGCTG
ACAGTTCAGGACACGTTCAGGCATTCAGTAGCGTGAATATTGCCAGCGGCCGCCCGACAGTCGTGCTGGCAGACAACCAG
CAGGTTAATCCGGACAATATATCCTGGGGCTACCGGGGGGGGGTTCTGGATGTTAACGGGAATGACCTGACATTTCATAA
GCTGAATGCCGCCGATTATGGCGCAACTCTCGGTAACAGCAGTGATAAAACGGCTAATATCACTCTGGATTATCAGACGC
GTCCGGCAGACGTAAAAGTTAATGAATGGTCATCATCAAACAGGGGAACAGTAGGTTCATTATATATTTATAATAATCCC
TATACTCATACCGTCGATTATTTTATCCTGAAAACAAGTAGTTATGGCTGGTTCCCTACCGGTCAGGTCAGTAACGAGCA
CTGGGAATATGTCGGACATGACCAGAACAGTGCACAGGCACTGCTTGCAAACAGAATTAATAATAAAGGGTATCTGTATC
ATGGCAAGTTGCTGGGAAATATTAATTTCTCAAATAAAGCAACCCCGGGTACAACCGGCGCATTGGTTATGGACGGCTCA
GCGAATATGTCCGGTACATTTACTCAGGAAAACGGTCGTCTGACCATTCAGGGCCACCCGGTTATCCATGCTTCAACGTC
TCAGAGTATTGCAAATACAGTCTCGTCTCTGGGCGACAATTCCGTTCTGACACAGCCCACCTCATTTACACAGGATGACT
GGGAGAACAGGACGTTCAGCTTTGGTTCGCTCGTGTTAAAAGATACAGACTTTGGTCTGGGCCGCAATGCCACACTGAAC
ACAACCATCCAGGCAGATAACTCCAGCGTCACGCTGGGCGACAGTCGGGTATTTATCGACAAAAAAGATGGCCAGGGAAC
AGCATTTACCCTTGAAGAAGGCACATCTGTTGCAACTAAAGATGCAGATAAAAGCGTCTTCAACGGCACCGTCAACCTGG
ATAATCAGTCAGTGCTGAATATCAATGAGATATTCAATGGCGGAATACAGGCGAACAACAGTACCGTGAATATCTCCTCA
GACAGTGCCGTTCTGGAGAACTCAACGCTGACCAGTACCGCCCTGAATCTGAACAAGGGAGCAAATGTTCTGGCCAGTCA
GAGTTTTGTTTCTGACGGTCCGGTGAATATTTCTGATGCCACCCTGAGTCTGAACAGCCGTCCTGATGAGGTATCTCACA
CACTTTTACCTGTATACGATTATGCCGGTTCATGGAACCTGAAGGGAGACGATGCCCGCCTGAACGTGGGGCCGTACAGT
ATGTTGTCAGGTAATATCAATGTTCAGGATAAAGGGACTGTCACCCTCGGAGGGGAAGGGGAACTGAGTCCTGACCTGAC
TCTTCAGAATCAGATGTTGTACAGCCTGTTTAACGGGTACCGCAATACCTGGAGCGGGAGCCTGAATGCACCGGATGCCA
CCGTCAGCATGACAGACACCCAGTGGTCGATGAACGGAAACTCCACGGCAGGAAATATGAAACTTAACCGGACAATAGTC
GGTTTTAACGGGGGAACATCATCGTTCACGACACTGACAACAGATAATCTGGACGCGGTTCAGTCAGCATTTGTCATGCG
TACAGACCTTAACAAGGCAGACAAACTGGTGATAAACAAGTCGGCAACAGGTCATGACAACAGCATCTGGGTTAACTTCC
TGAAAAAACCCTCTGACAAGGACACGCTTGATATTCCACTGGTCAGCGCACCTGAAGCGACAGCTGATAATCTGTTCAGG
GCATCAACACGGGTTGTGGGATTCAGTGATGTCACCCCCACCCTTAGTGTCAGAAAAGAGGACGGGAAAAAAGAGTGGGT
CCTCGATGGTTACCAGGTTGCACGTAACGACGGCCAGGGTAAGGCTGCCGCCACATTCATGCACATCAGCTATAACAACT
TCATCACTGAAGTTAACAACCTGAACAAACGCATGGGCGATTTGAGGGATATTAACGGCGAAGCCGGTACGTGGGTGCGT
CTGCTGAACGGTTCCGGCTCTGCTGATGGCGGTTTCACTGACCACTATACCCTGCTGCAGATGGGGGCTGACCGTAAGCA
CGAACTGGGAAGTATGGACCTGTTTACCGGCGTGATGGCCACCTACACTGACACAGATGCGTCAGCAGGCCTGTACAGCG
GTAAAACAAAATCATGGGGTGGTGGTTTCTATGCCAGTGGTCTGTTCCGGTCCGGCGCTTACTTTGATTTGATTGCCAAA
TATATTCACAATGAAAACAAATATGACCTGAACTTTGCCGGAGCTGGTAAACAGAACTTCCGCAGCCATTCACTGTATGC
AGGTGCAGAAGTCGGATACCGTTATCATCTGACAGATACGACGTTTGTTGAACCTCAGGCGGAACTGGTCTGGGGAAGAC
TGCAGGGCCAAACATTTAACTGGAACGACAGTGGAATGGATGTCTCAATGCGTCGTAACAGCGTTAATCCTCTGGTAGGC
AGAACCGGCGTTGTTTCCGGTAAAACCTTCAGTGGTAAGGACTGGAGTCTGACAGCCCGTGCCGGCCTGCATTATGAGTT
CGATCTGACGGACAGTGCTGACGTTCACCTGAAGGATGCAGCGGGAGAACATCAGATTAATGGCAGAAAAGACGGTCGTA
TGCTTTACGGTGTGGGGTTAAATGCCCGGTTTGGCGACAATACGCGTCTGGGGCTGGAAGTTGAACGCTCTGCATTCGGT
AAATACAACACAGATGATGCGATAAACGCTAATATTCGTTATTCATTCTGA

Protein sequence :
MNKIYALKYCYITNTVKVVSELARRVCKGSTRRGKRLSVLTSLALSALLPTVAGASTVGGNNPYQTYRDFAENKGQFQAG
ATNIPIFNNKGELVGHLDKAPMVDFSSVNVSSNPGVATLINPQYIASVKHNKGYQSVSFGDGQNSYHIVDRNEHSSSDLH
TPRLDKLVTEVAPATVTSSSTADILNPSKYSAFYRAGSGSQYIQDSQGKRHWVTGGYGYLTGGILPTSFFYHGSDGIQLY
MGGNIHDHSILPSFGEAGDSGSPLFGWNTAKGQWELVGVYSGVGGGTNLIYSLIPQSFLSQIYSEDNDAPVFFNASSGAP
LQWKFDSSTGTGSLKQGSDEYAMHGQKGSDLNAGKNLTFLGHNGQIDLENSVTQGAGSLTFTDDYTVTTSNGSTWTGAGI
IVDKDASVNWQVNGVKGDNLHKIGEGTLVVQGTGVNEGGLKVGDGTVVLNQQADSSGHVQAFSSVNIASGRPTVVLADNQ
QVNPDNISWGYRGGVLDVNGNDLTFHKLNAADYGATLGNSSDKTANITLDYQTRPADVKVNEWSSSNRGTVGSLYIYNNP
YTHTVDYFILKTSSYGWFPTGQVSNEHWEYVGHDQNSAQALLANRINNKGYLYHGKLLGNINFSNKATPGTTGALVMDGS
ANMSGTFTQENGRLTIQGHPVIHASTSQSIANTVSSLGDNSVLTQPTSFTQDDWENRTFSFGSLVLKDTDFGLGRNATLN
TTIQADNSSVTLGDSRVFIDKKDGQGTAFTLEEGTSVATKDADKSVFNGTVNLDNQSVLNINEIFNGGIQANNSTVNISS
DSAVLENSTLTSTALNLNKGANVLASQSFVSDGPVNISDATLSLNSRPDEVSHTLLPVYDYAGSWNLKGDDARLNVGPYS
MLSGNINVQDKGTVTLGGEGELSPDLTLQNQMLYSLFNGYRNTWSGSLNAPDATVSMTDTQWSMNGNSTAGNMKLNRTIV
GFNGGTSSFTTLTTDNLDAVQSAFVMRTDLNKADKLVINKSATGHDNSIWVNFLKKPSDKDTLDIPLVSAPEATADNLFR
ASTRVVGFSDVTPTLSVRKEDGKKEWVLDGYQVARNDGQGKAAATFMHISYNNFITEVNNLNKRMGDLRDINGEAGTWVR
LLNGSGSADGGFTDHYTLLQMGADRKHELGSMDLFTGVMATYTDTDASAGLYSGKTKSWGGGFYASGLFRSGAYFDLIAK
YIHNENKYDLNFAGAGKQNFRSHSLYAGAEVGYRYHLTDTTFVEPQAELVWGRLQGQTFNWNDSGMDVSMRRNSVNPLVG
RTGVVSGKTFSGKDWSLTARAGLHYEFDLTDSADVHLKDAAGEHQINGRKDGRMLYGVGLNARFGDNTRLGLEVERSAFG
KYNTDDAINANIRYSF

• Homologs from PAI DB

GeneGenBank Accn Product Virulance or Resistance PAI or REI Alignment Type E-val Identity
unnamed CAD66214.1 putative hemoglobin protease Not tested PAI III 536 Protein 0.0 99
vat YP_851472.1 vacuolating autotransporter Not tested PAI III APEC-O1 Protein 0.0 99
vat AAO21903.1 vacuolating autotransporter toxin Virulence Not named Protein 0.0 98
pic NP_838464.1 serine protease precurser Virulence SHI-1 Protein 0.0 49
pic NP_708747.3 serine protease Not tested SHI-1 Protein 0.0 48
she AAB58244.1 mucinase Virulence SHI-1 Protein 0.0 48
pic AAK00464.1 Pic Virulence SHI-1 Protein 0.0 48
unnamed CAC39286.1 hypothetical protein Not tested LPA Protein 0.0 44

• Homologs from VFDB (virulence genes)

GeneGenBank Accn Product ID of source DB Alignment Type E-val Identity
ECABU_c03670 YP_006104502.1 IgA-specific serine endopeptidase precursor VFG0904 Protein 0.0 100
ECABU_c03670 YP_006104502.1 IgA-specific serine endopeptidase precursor VFG1689 Protein 0.0 99
ECABU_c03670 YP_006104502.1 IgA-specific serine endopeptidase precursor VFG0635 Protein 0.0 48
ECABU_c03670 YP_006104502.1 IgA-specific serine endopeptidase precursor VFG0861 Protein 0.0 48
ECABU_c03670 YP_006104502.1 IgA-specific serine endopeptidase precursor VFG0903 Protein 0.0 48