ID JF909130; SV 1; circular; genomic DNA; STD; VRL; 2798 BP. XX AC JF909130; XX DT 21-JUN-2012 (Rel. 113, Created) DT 05-DEC-2012 (Rel. 115, Last updated, Version 3) XX DE East African cassava mosaic Kenya virus isolate Comoros:Moheli:MO05BD3:2009 DE segment DNA-A, complete sequence. XX KW . XX OS East African cassava mosaic Kenya virus OC Viruses; Geminiviridae; Begomovirus. XX RN [1] RC Publication Status: Online-Only RP 1-2798 RX DOI; 10.1186/1471-2148-12-228. RX PUBMED; 23186303. RA De Bruyn A., Villemot J., Lefeuvre P., Villar E., Hoareau M., RA Harimalala M., Abdoul-Karime A.L., Abdou-Chakour C., Reynaud B., RA Harkins G.W., Varsani A., Martin D.P., Lett J.M.; RT "East African cassava mosaic-like viruses from Africa to Indian ocean RT islands: molecular diversity, evolutionary history and geographical RT dissemination of a bipartite begomovirus"; RL BMC Evol. Biol. 12(1):228-228(2012). XX RN [2] RP 1-2798 RA Villemot J., Lefeuvre P., Villar E., Hoareau M., Harimalala M., RA Abdoul-Karime A.L., Abdou-Chakour C., Reynaud B., Varsani A., Martin D.P., RA Lett J.-M.; RT ; RL Submitted (24-MAR-2011) to the INSDC. RL UMR PVBMT, CIRAD, 7, chemin de l'IRAT, Saint-Pierre, Reunion 97410, France XX DR MD5; ea09d914f65ef36f9ac10f4db4df467e. XX FH Key Location/Qualifiers FH FT source 1..2798 FT /organism="East African cassava mosaic Kenya virus" FT /segment="DNA-A" FT /host="Manihot esculenta (cassava)" FT /isolate="Comoros:Moheli:MO05BD3:2009" FT /mol_type="genomic DNA" FT /country="Comoros:Moheli" FT /lat_lon="12.29 S 43.75 E" FT /collection_date="2009" FT /db_xref="taxon:393599" FT gene 174..530 FT /gene="AV2" FT CDS 174..530 FT /codon_start=1 FT /gene="AV2" FT /product="movement protein" FT /db_xref="GOA:I6LXA7" FT /db_xref="InterPro:IPR002511" FT /db_xref="InterPro:IPR005159" FT /db_xref="UniProtKB/TrEMBL:I6LXA7" FT /protein_id="AEG90114.1" FT /translation="MWDPLLNDFPETVHGFRSMLAVKYLLHLEQEYDRGTVGAEYIRDL FT IGVLRCKSYVEATRRYNNLNTRIQGAEEAELRQPIHEPCCCPHCPRHQKQNMGQQAHVS FT EAQDVQNVSKPRCS" FT gene 334..1107 FT /gene="AV1" FT CDS 334..1107 FT /codon_start=1 FT /gene="AV1" FT /product="coat protein" FT /db_xref="GOA:I6LY56" FT /db_xref="InterPro:IPR000263" FT /db_xref="InterPro:IPR000650" FT /db_xref="UniProtKB/TrEMBL:I6LY56" FT /protein_id="AEG90113.1" FT /translation="MSKRPGDIIISTPVSKVRRRLNFDSPYTNRVVAPTVRVTRSKIWA FT NRPMYRKPKMYRMYRSPDVPKGCEGPCKVQSYEQRDDVKHTGTVRCVSDVTRGSGITHR FT VGKRFCVKSIYILGKIWMDDNIKKQNHTNHVMFFLVRDRRPYGQSPQEFGQVFNMFDNE FT PTTATVKNDLRDRYQVLRKFYATVVGGPSGMKEQALVKRFFRINNHVVYNHQEQAKYEN FT HTENALLLYMACTHASNPVYATLKIRIYFYDAVTN" FT gene complement(1104..1508) FT /gene="AC3" FT CDS complement(1104..1508) FT /codon_start=1 FT /gene="AC3" FT /product="replication enhancer" FT /db_xref="GOA:I6LY54" FT /db_xref="InterPro:IPR000657" FT /db_xref="UniProtKB/TrEMBL:I6LY54" FT /protein_id="AEG90117.1" FT /translation="MDSRTGELITAPQAKNGVFTWEITNPLYFDITNHDRRPGNMNHDI FT ITFQIRFNHNLRKALGIHKCFLNFKVWTTLQPPTGLFLKVFKYQVLKYLDMIGVISINT FT VIQAVDHVLYNVLLNTLQVTEHHAIKFNLY" FT gene complement(1249..1656) FT /gene="AC2" FT CDS complement(1249..1656) FT /codon_start=1 FT /gene="AC2" FT /product="transcription activator protein" FT /db_xref="GOA:I6LY59" FT /db_xref="InterPro:IPR000942" FT /db_xref="UniProtKB/TrEMBL:I6LY59" FT /protein_id="AEG90116.1" FT /translation="MPPSSPSTSHCSQVPIKVQHRTAKTRAVRRRRVDLECGCSFYLHI FT DCINHGFSHRGTHHCASSKEWRFYLGNNKSPLFRHHQPRQETREHEPRHHHIPDTVQPQ FT PPEGIGDSQVFSQLQGLDDLTASDWSFLKSI" FT gene complement(1580..2644) FT /gene="AC1" FT CDS complement(1580..2644) FT /codon_start=1 FT /gene="AC1" FT /product="replication associated protein" FT /db_xref="GOA:I6LY52" FT /db_xref="InterPro:IPR001191" FT /db_xref="InterPro:IPR001301" FT /db_xref="InterPro:IPR022690" FT /db_xref="InterPro:IPR022692" FT /db_xref="UniProtKB/TrEMBL:I6LY52" FT /protein_id="AEG90115.1" FT /translation="MPRAGRFSIKAKNYFLTYPKCSLSKEEALNQLRQLQTPTNKLFIK FT ICRELHENGEPHLHALIQFEGKYNCTNQRFFDLISPSRSAHFHPNIQGAKSSSDVKSYL FT DKDGDTIQWGEFQIDGRSARGGQQSANDAYAKALNSANKSEALNVIRELAPKDFVLQFH FT NLISNLERIFQEPLTPYISPFLSSSFTNVPEELEAWVSENVMGSAARPWRPSSIVIEGD FT SRTGKTMWARSLGPHNYLCGHLDLSPKVYRNDAWYNVIDDVDPHYLKHFKEFMGAQRDW FT QSNTKYGKPIQIKGGIPTIFLCNPGPTSSYKEFLDEEKNQSLKAWALKNATFITLHEPL FT FSSAHQSPTPHSED" FT gene complement(2197..2493) FT /gene="AC4" FT CDS complement(2197..2493) FT /codon_start=1 FT /gene="AC4" FT /product="C4 protein" FT /db_xref="InterPro:IPR002488" FT /db_xref="UniProtKB/TrEMBL:I6LY61" FT /protein_id="AEG90118.1" FT /translation="MKMGNLICMPSFSSRASTIVPTNDSSTSYPLPGPPISTQIFRELN FT QAPTSSPIWIRTETPSNGASFRSTDDLLEADNNPPMTLTPRLLTQQISQRLLM" XX SQ Sequence 2798 BP; 729 A; 555 C; 725 G; 789 T; 0 other; accggatggc cgcgcccgaa aaagcaggtg gaccccacaa gatggccgcg cccgttaaag 60 aaagtggtcc ccgcgcactt gtgttggtcg gccagtcata ttcacgcgtg aaagtctaga 120 tatttgttgt ttgtcattat agacttcgtc gcgaagtaga tgagcgcgtc aacatgtggg 180 atccattgtt gaacgatttt cccgaaaccg ttcacggttt ccgttctatg cttgctgtta 240 aatacctgtt acatctggaa caggaatacg atcgcggtac agtcggggct gagtatatac 300 gtgatttaat aggggttcta cggtgtaaga gttatgtcga agcgaccagg agatataata 360 atctcaacac ccgtatccaa ggtgcggagg aggctgaact tcgacagccc atacacgaac 420 cgtgttgttg cccccactgt ccgcgtcacc agaagcaaaa tatgggccaa caggcccatg 480 tatcggaagc ccaagatgta cagaatgtat cgaagcccag atgttcctaa gggctgtgaa 540 ggcccatgta aggttcagtc ctatgaacag agggatgatg tgaagcacac tggtacggtc 600 cgatgtgtca gtgatgtaac tcgtggatca ggcattaccc atagagtcgg gaagaggttt 660 tgtgtgaagt ccatatatat attgggcaag atttggatgg atgataatat caagaagcaa 720 aatcatacga atcatgttat gttcttcctt gttcgagata gaaggcctta tggtcagagt 780 cctcaagagt ttggacaagt gttcaacatg tttgataatg aacctactac ggcaactgtg 840 aagaatgatc ttagggaccg atatcaggtg ttacgtaaat tctatgcgac tgttgttggt 900 ggaccctctg ggatgaagga acaagctttg gttaagaggt tcttcaggat caataatcat 960 gtagtgtata atcatcagga acaggccaag tatgagaatc atactgagaa tgcgttgtta 1020 ttgtatatgg catgtacaca tgcctcgaat cctgtgtacg ctacgctgaa aatacgcatc 1080 tatttctatg atgcagtgac aaattaataa aggttgaatt ttattgcatg gtgctccgta 1140 acttggagtg tgtttagtaa tacattgtac agaacatgat caacagcttg aattacagtg 1200 ttaatggaaa taacgcctat catatctaaa tacttgagca cttgatattt aaatactttt 1260 aagaaaagac cagtcggagg ctgtaaggtc gtccagacct tgaagttgag aaaacacttg 1320 tgaatcccca atgccttccg gaggttgtgg ttgaaccgta tctggaatgt gatgatgtcg 1380 tggttcatgt tccctggtct cctgtcgtgg ttggtgatgt cgaaatagag gggatttgtt 1440 atttcccagg taaaaacgcc attctttgct tgaggcgcag tgatgagttc ccctgtgcga 1500 gaatccatga ttgatgcagt cgatatggag atagaacgag caaccgcatt cgaggtctac 1560 ccgcctacgt ctgacggccc tagtcttcgc tgtgcggtgt tggactttga tgggcacttg 1620 agaacaatgg ctcgtggagg gtgatgaagg tggcattctt taaagcccag gctttaaggg 1680 actggttctt ttcctcgtcc agaaactctt tatatgatga tgttggtcct ggattgcata 1740 ggaagatagt gggaatgccg cctttaattt gaattggctt cccgtatttt gtattgcttt 1800 gccagtccct ttgggccccc atgaattctt tgaaatgctt gaggtagtgg gggtcgacgt 1860 catcaatgac gttgtaccat gcgtcgttgc ggtatacctt tggactgaga tccaggtgtc 1920 cacacaagta gttatgtggt cccaaagagc gagcccacat tgtcttccct gtcctactat 1980 cgccctcgat tacgatacta ctaggtctcc atggccgcgc agcggaaccc atcacgttct 2040 cggaaaccca ggcttcaagt tcctcaggaa cgttagtgaa agaagaagaa agaaagggag 2100 aaatataagg agtgagaggc tcttgaaaaa tcctctctaa attgctaatt aaattatgaa 2160 actgtaaaac aaaatctttt ggggctagtt cccgtattac attaagagcc tctgacttat 2220 ttgctgagtt aagagccttg gcgtaagcgt cattggcgga ttgttgtccg cctcgagcag 2280 atcgtccgtc gatctgaaac tcgccccatt ggatggtgtc tccgtcctta tccagatagg 2340 acttgacgtc ggagcttgat ttagctccct gaatatttgg gtggaaatgg gcggaccggg 2400 aaggggatat gaggtcgaag aatcgttggt tggtacaatt gtacttgccc tcgaactgaa 2460 tgagggcatg cagatgaggt tccccatttt catggagctc tctgcagatc ttgatgaaca 2520 atttatttgt tggggtttgg agttgtcgga gctgattcaa ggcctcttct ttcgatagag 2580 aacatttggg atatgtgagg aaatagtttt tggctttgat gctaaaacga ccagcccttg 2640 gcattttcgc tgtcgtatag caatcggggg gcactcaaag tctgtagcaa tcgggggaat 2700 gggggggcaa tttatatgat gccccccaaa tggcatttat gtaatatcct catgaaattt 2760 gaattgcaaa tgtggaaagc ggccatccgt ataatatt 2798 //