Gene Information

Name : clpB (Erum6400)
Accession : YP_180504.1
Strain : Ehrlichia ruminantium Welgevonden
Genome accession: NC_005295
Putative virulence/resistance : Virulence
Product : ClpA-type chaperone
Function : -
COG functional category : O : Posttranslational modification, protein turnover, chaperones
COG ID : COG0542
EC number : -
Position : 1074846 - 1077425 bp
Length : 2580 bp
Strand : -
Note : clpB | heat shock protein ClpB | len: 859 aa | Highly similar to many e.g. CLPB_ECOLI P03815 ClpB protein (Heat shock protein F84.1) (857 aa) from Escherichia coli, fasta scores: E(): 7.5e-156, 55.064% identity in 859 aa overlap | Contains 2 Pfam matches

DNA sequence :
ATGGATTTAAATAAATTTACTGATATATCAAAGAATTTCATAGTGCAAGCGCAAACTTCAGCTGTTGCATTAGGGCATCA
GTCTTTAGTACCTGAGCATTTACTTAAGGTAATGTTGGATGATAAAGATGAGATAGTTGAAGTTTTGCTTACTTCTTGTG
GGTGTAATGTAGAGACGTTACGTAATGATGTTATATCGGCTTTAAATAAATTACCAGTTGTAAGTGGTCCAGGTAGTGGT
CATATACATTTATCAAAGGAAATGGCGCAAGTTTTACAAGAGGCTGTTAATCTTGCAAAAAGACATCAGGACTCTTATGT
TACTGTTGAAAGATTACTGCAAGCTCTGACAATAATAAAGGACAGTAATGTTTCTAGGATATTAATTGCACATGGTGTGA
CTCCTCAGAAGTTGGAGTCATTAATAGTAAACATGCGTAATGGTGCTAGAGCTGATAGTGTAAATTCTGAGCAAAAGTTT
AATGCACTAAAAAAATATGCTAAAGATGTGACTGAAGTTGCTAGAGCAGGAAAATTAGATCCAGTAATTGGAAGAGATGA
GGAAATTAGACGTACAATACAGGTATTATTGAGAAGAACAAAAAATAATCCTGTATTAATTGGAGAGCCTGGTGTTGGTA
AAACAGCAATTATTGAAGGGTTAGCACATAAGATAGTGAAGGGAGATGTTCCAATTGGGTTGCGAGATATGAGAATAATG
TCATTAGATCTTGGTATGCTTGTTGCTGGGACTAAATATAGAGGTGAATTTGAAGAAAGGTTGAAAGCTGTAGTTAATGA
AATTGTTTCTTCAAATGGTAGTATTATATTATTCATTGATGAGTTACATACATTAGTTGGTGCTGGTGCAACAGATGGAG
CAATGGATGCATCAAATTTGTTGAAGCCAGCATTAGCTAGAGGTGAAATACATTGTATAGGTGCAACAACATTGGATGAA
TATAGAAAGCATATAGAAAAAGATGTAGCACTTGCTAGAAGATTTCAAACTATATTTATTTCTGAGCCAACTTGTGATGA
TACAATTTCTATGTTACGTGGGTTAAAGGAAAGATATGAAGGACATCATGGTATAGATATTCCTGACAGATCGATAATTG
CTGCTGTAGCTTTATCGCAGCGTTATATTACGGATAGGTATTTACCAGATAAAGCTATAGACCTTATTGATGAAGCAGCG
AGTCGTGCGAGAATGGAGATTGATAGTAAACCTGAGGTTATTGATAAGTTAGATAGAAAGATAATGCAGCTAAAAATCGA
GATAGGAGTATTAGAAAAAGAAAGTGATGAATCCTCAAAACAGAGGTTAATGAAGTTAAAAGATGAACTAGAAAAACTAA
ATGTTCAGTCTGCTGAGCTAAGTAGTAAATGGCAAGCGGAAAAAATGAAAATGTCAAAGATGAAAGCATGTAAGGAAAAG
CTTGATATTGCTAGAAGTGATTTAGAAAGAGCACAAAGATCTGGTGATTTGGCAAAAGCTGGTGAGTTAATGTATGGTGT
AATACCAGAAATTGAGAAAGAGTTAAAAGAACATGAAAAATTTACAAGTAGCCTTTTTAAGAAGGAAATTACAGAACATG
ACATAGCAAGTATTGTATCAAAATGGACTGGTATTCCTATTGAGAACATAATGAGTAGTGAAAGAGAAAAACTACTGCGT
ATGGAGGAGGAGATAGGCAAAACAGTTATTGGTCAGGATAGTGCTGTAAAAGCAGTAAGTGATGCTGTCAGGAGATCACG
TGCAGGGGTACAGGATGCACAGAAACCATTGGGGTCTTTTTTATTTCTTGGGCCAACTGGAGTAGGTAAAACTGAGTTGG
TTAAAACATTAGCTGAGTTTTTATTTTGTGATAAGTCTGCACTTTTAAGATTTGACATGTCAGAATTTATGGAAAAGCAT
GCTGTTTCACGATTAATAGGAGCTCCTCCAGGATATGTTGGATATGACCAAGGTGGTGCATTAACTGAAGCTGTGAGGAG
AAGGCCTTATCAAGTAATATTATTTGATGAAATTGAAAAAGCACATGGAGATATTTTCAATATTTTATTGCAAGTATTAG
ATGAAGGAAGATTGACTGATAATCATGGTAAGTTAGTGGATTTCCGTAATACAATACTGGTATTAACTTCAAATTTAGGG
CAAGAAATATTAATGAACAATGAATCTGGAAATATCAATGAAGAGTCAGTTAAAGAGTCTGTTACTAATGTGTTGCGTAG
TCATTTTCGGCCAGAATTTTTAAATAGATTGGATGAAATTATTATATTTCATAGGTTAACTAAAGAACATATTGAAAGAA
TTATTGATGTGCAATTTTCTATATTACAAAAAATTGTTGCTCAAAGAAAATTAGAGATTACTTTATCTTCAGATGCAAAA
ACATGGTTGATAAATAATGGCTATGATCCTTTATATGGGGCAAGACCTTTAAAGAGGTTAATACAACAGCAAATACAGAA
TAACTTGGCAAAATTAATACTTGCTAATCAGGTAGCTGAAGGTAATAAATTAAGGGTAGATTTATTAGATGATAATCTTG
TTATTCATAAGATTAGTTAA

Protein sequence :
MDLNKFTDISKNFIVQAQTSAVALGHQSLVPEHLLKVMLDDKDEIVEVLLTSCGCNVETLRNDVISALNKLPVVSGPGSG
HIHLSKEMAQVLQEAVNLAKRHQDSYVTVERLLQALTIIKDSNVSRILIAHGVTPQKLESLIVNMRNGARADSVNSEQKF
NALKKYAKDVTEVARAGKLDPVIGRDEEIRRTIQVLLRRTKNNPVLIGEPGVGKTAIIEGLAHKIVKGDVPIGLRDMRIM
SLDLGMLVAGTKYRGEFEERLKAVVNEIVSSNGSIILFIDELHTLVGAGATDGAMDASNLLKPALARGEIHCIGATTLDE
YRKHIEKDVALARRFQTIFISEPTCDDTISMLRGLKERYEGHHGIDIPDRSIIAAVALSQRYITDRYLPDKAIDLIDEAA
SRARMEIDSKPEVIDKLDRKIMQLKIEIGVLEKESDESSKQRLMKLKDELEKLNVQSAELSSKWQAEKMKMSKMKACKEK
LDIARSDLERAQRSGDLAKAGELMYGVIPEIEKELKEHEKFTSSLFKKEITEHDIASIVSKWTGIPIENIMSSEREKLLR
MEEEIGKTVIGQDSAVKAVSDAVRRSRAGVQDAQKPLGSFLFLGPTGVGKTELVKTLAEFLFCDKSALLRFDMSEFMEKH
AVSRLIGAPPGYVGYDQGGALTEAVRRRPYQVILFDEIEKAHGDIFNILLQVLDEGRLTDNHGKLVDFRNTILVLTSNLG
QEILMNNESGNINEESVKESVTNVLRSHFRPEFLNRLDEIIIFHRLTKEHIERIIDVQFSILQKIVAQRKLEITLSSDAK
TWLINNGYDPLYGARPLKRLIQQQIQNNLAKLILANQVAEGNKLRVDLLDDNLVIHKIS

• Homologs from PAI DB

GeneGenBank Accn Product Virulance or Resistance PAI or REI Alignment Type E-val Identity
clpC YP_005163377.1 ATP-dependent Clp protease ATP-binding subunit Not tested Not named Protein 1e-172 43

• Homologs from VFDB (virulence genes)

GeneGenBank Accn Product ID of source DB Alignment Type E-val Identity
clpB YP_180504.1 ClpA-type chaperone VFG2076 Protein 3e-112 41