Query 029214
Match_columns 197
No_of_seqs 120 out of 187
Neff 3.3
Searched_HMMs 46136
Date Fri Mar 29 09:20:02 2013
Command hhsearch -i /work/01045/syshi/csienesis_hhblits_a3m/029214.a3m -d /work/01045/syshi/HHdatabase/Cdd.hhm -o /work/01045/syshi/hhsearch_cdd/029214hhsearch_cdd -cpu 12 -v 0
No Hit Prob E-value P-value Score SS Cols Query HMM Template HMM
1 PF04690 YABBY: YABBY protein; 100.0 5.2E-78 1.1E-82 499.9 11.9 157 7-170 4-170 (170)
2 PF09011 HMG_box_2: HMG-box do 98.3 8E-07 1.7E-11 62.9 4.9 46 119-164 1-47 (73)
3 cd01390 HMGB-UBF_HMG-box HMGB- 97.9 2.8E-05 6E-10 52.4 5.1 42 123-164 2-43 (66)
4 cd00084 HMG-box High Mobility 97.8 4.3E-05 9.4E-10 50.8 5.2 43 123-165 2-44 (66)
5 cd01388 SOX-TCF_HMG-box SOX-TC 97.8 3.6E-05 7.9E-10 54.3 5.1 42 123-164 3-44 (72)
6 PF00505 HMG_box: HMG (high mo 97.8 3.9E-05 8.4E-10 52.2 4.5 42 123-164 2-43 (69)
7 smart00398 HMG high mobility g 97.8 6E-05 1.3E-09 50.7 5.3 43 122-164 2-44 (70)
8 cd01389 MATA_HMG-box MATA_HMG- 97.7 7.9E-05 1.7E-09 53.0 5.4 43 123-165 3-45 (77)
9 PTZ00199 high mobility group p 97.7 8.6E-05 1.9E-09 55.7 5.8 49 117-165 18-68 (94)
10 PF06244 DUF1014: Protein of u 97.0 0.0011 2.4E-08 53.4 4.3 51 117-169 70-120 (122)
11 KOG0381 HMG box-containing pro 96.8 0.0031 6.7E-08 45.9 5.3 46 120-165 21-66 (96)
12 KOG3223 Uncharacterized conser 94.7 0.022 4.8E-07 50.0 2.4 55 115-171 160-214 (221)
13 PF11331 DUF3133: Protein of u 93.0 0.057 1.2E-06 37.2 1.5 38 15-54 6-46 (46)
14 TIGR02098 MJ0042_CXXC MJ0042 f 80.3 1.5 3.4E-05 27.4 2.1 34 16-51 3-37 (38)
15 PF13719 zinc_ribbon_5: zinc-r 79.3 1.4 3E-05 28.3 1.7 34 16-51 3-37 (37)
16 PRK14892 putative transcriptio 73.9 2.6 5.7E-05 32.9 2.2 44 14-60 20-63 (99)
17 PF08073 CHDNT: CHDNT (NUC034) 73.3 6.3 0.00014 28.2 3.8 33 128-163 18-50 (55)
18 TIGR01562 FdhE formate dehydro 72.4 2 4.4E-05 39.2 1.5 27 11-47 206-232 (305)
19 PF04032 Rpr2: RNAse P Rpr2/Rp 71.6 1.9 4.2E-05 30.7 0.9 32 16-47 47-85 (85)
20 PF05129 Elf1: Transcription e 71.3 2.3 4.9E-05 31.8 1.3 42 15-56 22-63 (81)
21 PRK03954 ribonuclease P protei 70.8 2.4 5.2E-05 34.2 1.4 32 18-49 67-103 (121)
22 PF13717 zinc_ribbon_4: zinc-r 69.4 5.5 0.00012 25.5 2.6 30 16-49 3-35 (36)
23 PF10122 Mu-like_Com: Mu-like 67.5 3.2 7E-05 29.4 1.3 32 16-51 5-36 (51)
24 PRK03564 formate dehydrogenase 64.6 3.8 8.2E-05 37.7 1.6 27 11-47 208-234 (309)
25 PF09788 Tmemb_55A: Transmembr 58.8 7.3 0.00016 35.4 2.3 39 11-53 153-191 (256)
26 KOG4684 Uncharacterized conser 57.2 4.2 9.2E-05 36.8 0.5 16 37-52 168-183 (275)
27 COG3058 FdhE Uncharacterized p 54.6 2 4.4E-05 39.7 -1.9 34 10-53 206-239 (308)
28 COG5648 NHP6B Chromatin-associ 52.5 28 0.0006 30.9 4.8 45 119-163 68-112 (211)
29 PF04216 FdhE: Protein involve 52.0 4.7 0.0001 35.4 -0.0 32 13-54 195-226 (290)
30 PF11331 DUF3133: Protein of u 49.2 10 0.00022 26.2 1.3 18 13-30 29-46 (46)
31 PF06382 DUF1074: Protein of u 48.8 26 0.00056 30.6 3.9 35 127-165 84-118 (183)
32 KOG4684 Uncharacterized conser 44.3 11 0.00023 34.3 1.0 33 13-53 168-203 (275)
33 TIGR01053 LSD1 zinc finger dom 40.9 26 0.00056 22.2 2.1 25 16-46 2-26 (31)
34 KOG0526 Nucleosome-binding fac 38.7 42 0.0009 33.9 4.1 44 118-163 532-575 (615)
35 PF03811 Zn_Tnp_IS1: InsA N-te 37.5 23 0.0005 23.0 1.5 17 35-51 1-17 (36)
36 COG4416 Com Mu-like prophage p 37.4 7.9 0.00017 28.2 -0.7 14 37-50 2-15 (60)
37 PF10963 DUF2765: Protein of u 33.1 65 0.0014 24.6 3.5 14 134-147 48-61 (83)
38 PF05047 L51_S25_CI-B8: Mitoch 31.6 40 0.00086 22.2 2.0 18 130-147 2-19 (52)
39 PF01020 Ribosomal_L40e: Ribos 31.3 18 0.00039 25.8 0.3 9 41-49 38-46 (52)
40 COG4888 Uncharacterized Zn rib 30.4 37 0.00079 27.3 1.9 45 15-59 22-66 (104)
41 PF12876 Cellulase-like: Sugar 30.4 73 0.0016 23.1 3.3 31 118-148 29-59 (88)
42 KOG4715 SWI/SNF-related matrix 29.7 68 0.0015 30.8 3.8 46 117-165 63-108 (410)
43 PF14599 zinc_ribbon_6: Zinc-r 28.7 35 0.00076 24.6 1.4 30 12-47 27-56 (61)
44 PF10159 MMtag: Kinase phospho 28.5 52 0.0011 25.1 2.3 13 132-144 57-69 (78)
45 PF05180 zf-DNL: DNL zinc fing 28.0 33 0.00071 25.3 1.2 32 17-48 6-38 (66)
46 PF02892 zf-BED: BED zinc fing 27.9 27 0.00059 22.1 0.6 17 12-28 13-29 (45)
47 COG3712 FecR Fe2+-dicitrate se 26.7 58 0.0013 30.3 2.8 31 132-164 32-62 (322)
48 PF00527 E7: E7 protein, Early 26.5 26 0.00057 26.7 0.5 17 16-32 53-69 (92)
49 PF02723 NS3_envE: Non-structu 25.8 15 0.00032 28.2 -1.0 16 11-26 37-52 (82)
50 PF05164 ZapA: Cell division p 24.8 1.1E+02 0.0023 21.6 3.3 32 131-162 28-59 (89)
51 PF04769 MAT_Alpha1: Mating-ty 24.2 1.9E+02 0.004 25.2 5.3 44 118-165 40-83 (201)
52 PF13408 Zn_ribbon_recom: Reco 24.1 37 0.00079 22.1 0.7 13 39-51 5-17 (58)
53 KOG0527 HMG-box transcription 23.8 96 0.0021 28.9 3.6 62 117-186 58-119 (331)
54 PRK09774 fec operon regulator 22.7 72 0.0016 28.5 2.6 29 134-164 33-61 (319)
55 PF04420 CHD5: CHD5-like prote 22.4 60 0.0013 26.6 1.8 37 124-160 36-72 (161)
56 smart00614 ZnF_BED BED zinc fi 22.4 49 0.0011 21.9 1.1 14 14-27 17-30 (50)
57 PF09788 Tmemb_55A: Transmembr 21.8 81 0.0018 28.8 2.7 17 36-52 154-170 (256)
58 PF09102 Exotox-A_target: Exot 21.5 48 0.001 27.6 1.1 40 135-175 77-116 (143)
59 TIGR02147 Fsuc_second hypothet 21.5 88 0.0019 28.1 2.8 26 127-152 8-33 (271)
60 COG4357 Zinc finger domain con 21.2 27 0.00058 28.1 -0.4 39 15-53 26-76 (105)
61 PRK12336 translation initiatio 20.8 1E+02 0.0022 26.3 3.0 36 15-56 98-136 (201)
62 KOG4520 Predicted coiled-coil 20.8 1.2E+02 0.0025 27.4 3.4 15 130-144 62-76 (238)
63 PF04690 YABBY: YABBY protein; 20.1 50 0.0011 28.2 1.0 19 15-33 36-54 (170)
No 1
>PF04690 YABBY: YABBY protein; InterPro: IPR006780 YABBY proteins are a group of plant-specific transcription factors involved in the specification of abaxial polarity in lateral organs such as leaves and floral organs [, ].
Probab=100.00 E-value=5.2e-78 Score=499.90 Aligned_cols=157 Identities=61% Similarity=0.948 Sum_probs=116.9
Q ss_pred CCCCCceeeecCCCcceeeeccccCCCccceeeeecCCCCCccccccchhccCCCcchhcc--ccCC---CC--CCCccc
Q 029214 7 DVAPEQLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMAAAFQSLSWQDVHH--HQAP---SY--ASPECR 79 (197)
Q Consensus 7 ~~~~E~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr~llqsl~~~~~~~--~~~~---~~--~~~~~~ 79 (197)
+.++|||||||||||||||||||||+|||+|||||||||+|||||||++++++++.+++.. +..+ .. ..+...
T Consensus 4 ~~~sE~lCYVhCnFC~TiLaVsVP~ssL~~~VTVRCGHCtNLLSVNm~~~~~~~~~~~~l~~~~~~~~~~~~~~~~~~~~ 83 (170)
T PF04690_consen 4 FSPSEQLCYVHCNFCNTILAVSVPCSSLLKTVTVRCGHCTNLLSVNMRALLQPLPSQDHLQHSLLPPQSQELQFQPENFG 83 (170)
T ss_pred cCCCCcEEEEEcCCcCeEEEEecchhhhhhhhceeccCccceeeeeccccccCCCcccchhccccccccccccccccccc
Confidence 4468999999999999999999999999999999999999999999999998887665411 0000 00 001111
Q ss_pred ccCCCC-Cccccccc-ccCCCCCccccccc-cccCCCCCCCCCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHH
Q 029214 80 IDLGSS-SKCNNKIS-AMRTPTNKATEERV-VNRRESPHSTTPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTA 156 (197)
Q Consensus 80 ~~~~ss-s~~~~~~~-~~~~~~~~~~~~~~-~~k~~~~~~kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~a 156 (197)
....++ +.++.... .+....+++.|+.+ ++| |||||||+|||||+||||||||||++||||+|||||++|
T Consensus 84 ~~~~~~~~~~~~~~~~~~~~~~~~~~pr~~~v~k-------PPEKRqR~psaYn~f~k~ei~rik~~~p~ishkeaFs~a 156 (170)
T PF04690_consen 84 SNSSSSSSSSSSSSSSSMSFSEEEEIPRAPPVNK-------PPEKRQRVPSAYNRFMKEEIQRIKAENPDISHKEAFSAA 156 (170)
T ss_pred cccccCCCccccccccccCccccccccccccccC-------CccccCCCchhHHHHHHHHHHHHHhcCCCCCHHHHHHHH
Confidence 111111 11111100 11112233455443 345 999999999999999999999999999999999999999
Q ss_pred HHhhccCCCccccc
Q 029214 157 AKNWAHFPHIHFGL 170 (197)
Q Consensus 157 AknW~~~Phihfgl 170 (197)
||||||+|||||||
T Consensus 157 AknW~h~phihfgl 170 (170)
T PF04690_consen 157 AKNWAHFPHIHFGL 170 (170)
T ss_pred HHhhhhCcccccCC
Confidence 99999999999997
No 2
>PF09011 HMG_box_2: HMG-box domain; InterPro: IPR015101 This domain is predominantly found in Maelstrom homologue proteins. It has no known function. ; GO: 0005634 nucleus; PDB: 2EQZ_A 1V64_A 2CTO_A 1H5P_A 3TQ6_A 3FGH_A 3TMM_A 1J3X_A 2YRQ_A 1AAB_A ....
Probab=98.34 E-value=8e-07 Score=62.88 Aligned_cols=46 Identities=35% Similarity=0.634 Sum_probs=38.6
Q ss_pred CcccCCCchhhhHHHHHHHHHHHhh-CCCCCHHHHHHHHHHhhccCC
Q 029214 119 PEKRQRVPSAYNQFIKEEIQRIKAN-NPDISHREAFSTAAKNWAHFP 164 (197)
Q Consensus 119 PEKRQR~PSaYN~FmK~ei~riK~~-~P~i~hkEaFs~aAknW~~~P 164 (197)
|.|..|.+|||+.||++.+.++++. .+.++++|++..++..|+..+
T Consensus 1 p~kpK~~~say~lF~~~~~~~~k~~G~~~~~~~e~~k~~~~~Wk~Ls 47 (73)
T PF09011_consen 1 PKKPKRPPSAYNLFMKEMRKEVKEEGGQKQSFREVMKEISERWKSLS 47 (73)
T ss_dssp SSS--SSSSHHHHHHHHHHHHHHHHT-T-SSHHHHHHHHHHHHHHS-
T ss_pred CcCCCCCCCHHHHHHHHHHHHHHHhcccCCCHHHHHHHHHHHHHhcC
Confidence 6677899999999999999999999 888999999999999999754
No 3
>cd01390 HMGB-UBF_HMG-box HMGB-UBF_HMG-box, class II and III members of the HMG-box superfamily of DNA-binding proteins. These proteins bind the minor groove of DNA in a non-sequence specific fashion and contain two or more tandem HMG boxes. Class II members include non-histone chromosomal proteins, HMG1 and HMG2, which bind to bent or distorted DNA such as four-way DNA junctions, synthetic DNA cruciforms, kinked cisplatin-modified DNA, DNA bulges, cross-overs in supercoiled DNA, and can cause looping of linear DNA. Class III members include nucleolar and mitochondrial transcription factors, UBF and mtTF1, which bind four-way DNA junctions.
Probab=97.90 E-value=2.8e-05 Score=52.39 Aligned_cols=42 Identities=31% Similarity=0.498 Sum_probs=39.4
Q ss_pred CCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214 123 QRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFP 164 (197)
Q Consensus 123 QR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~P 164 (197)
.|-+|||..|+++....+++.+|+++..|....++..|+..+
T Consensus 2 krp~saf~~f~~~~r~~~~~~~p~~~~~~i~~~~~~~W~~ls 43 (66)
T cd01390 2 KRPLSAYFLFSQEQRPKLKKENPDASVTEVTKILGEKWKELS 43 (66)
T ss_pred CCCCcHHHHHHHHHHHHHHHHCcCCCHHHHHHHHHHHHHhCC
Confidence 467899999999999999999999999999999999999754
No 4
>cd00084 HMG-box High Mobility Group (HMG)-box is found in a variety of eukaryotic chromosomal proteins and transcription factors. HMGs bind to the minor groove of DNA and have been classified by DNA binding preferences. Two phylogenically distinct groups of Class I proteins bind DNA in a sequence specific fashion and contain a single HMG box. One group (SOX-TCF) includes transcription factors, TCF-1, -3, -4; and also SRY and LEF-1, which bind four-way DNA junctions and duplex DNA targets. The second group (MATA) includes fungal mating type gene products MC, MATA1 and Ste11. Class II and III proteins (HMGB-UBF) bind DNA in a non-sequence specific fashion and contain two or more tandem HMG boxes. Class II members include non-histone chromosomal proteins, HMG1 and HMG2, which bind to bent or distorted DNA such as four-way DNA junctions, synthetic DNA cruciforms, kinked cisplatin-modified DNA, DNA bulges, cross-overs in supercoiled DNA, and can cause looping of linear DNA. Class III member
Probab=97.84 E-value=4.3e-05 Score=50.84 Aligned_cols=43 Identities=30% Similarity=0.432 Sum_probs=39.8
Q ss_pred CCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214 123 QRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH 165 (197)
Q Consensus 123 QR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph 165 (197)
.|-+|+|..|++++...+++.+|+++..|....+++.|+..+.
T Consensus 2 krp~~af~~f~~~~~~~~~~~~~~~~~~~i~~~~~~~W~~l~~ 44 (66)
T cd00084 2 KRPLSAYFLFSQEHRAEVKAENPGLSVGEISKILGEMWKSLSE 44 (66)
T ss_pred CCCCcHHHHHHHHHHHHHHHHCcCCCHHHHHHHHHHHHHhCCH
Confidence 4678999999999999999999999999999999999997653
No 5
>cd01388 SOX-TCF_HMG-box SOX-TCF_HMG-box, class I member of the HMG-box superfamily of DNA-binding proteins. These proteins contain a single HMG box, and bind the minor groove of DNA in a highly sequence-specific manner. Members include SRY and its homologs in insects and vertebrates, and transcription factor-like proteins, TCF-1, -3, -4, and LEF-1. They appear to bind the minor groove of the A/T C A A A G/C-motif.
Probab=97.83 E-value=3.6e-05 Score=54.27 Aligned_cols=42 Identities=17% Similarity=0.294 Sum_probs=39.3
Q ss_pred CCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214 123 QRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFP 164 (197)
Q Consensus 123 QR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~P 164 (197)
.|-|+||..|+++.-.+|+++||+++..|.-+.++..|+..+
T Consensus 3 KrP~naf~~F~~~~r~~~~~~~p~~~~~eisk~l~~~Wk~ls 44 (72)
T cd01388 3 KRPMNAFMLFSKRHRRKVLQEYPLKENRAISKILGDRWKALS 44 (72)
T ss_pred CCCCcHHHHHHHHHHHHHHHHCCCCCHHHHHHHHHHHHHcCC
Confidence 478999999999999999999999999999999999999754
No 6
>PF00505 HMG_box: HMG (high mobility group) box; InterPro: IPR000910 High mobility group (HMG or HMGB) proteins are a family of relatively low molecular weight non-histone components in chromatin. HMG1 (also called HMG-T in fish) and HMG2 are two highly related proteins that bind single-stranded DNA preferentially and unwind double-stranded DNA. Although they have no sequence specificity, they have a high affinity for bent or distorted DNA, and bend linear DNA. HMG1 and HMG2 contain two DNA-binding HMG-box domains (A and B) that show structural and functional differences, and have a long acidic C-terminal domain rich in aspartic and glutamic acid residues. The acidic tail modulates the affinity of the tandem HMG boxes in HMG1 and 2 for a variety of DNA targets. HMG1 and 2 appear to play important architectural roles in the assembly of nucleoprotein complexes in a variety of biological processes, for example V(D)J recombination, the initiation of transcription, and DNA repair []. The profile in this entry describing the HMG-domains is much more general than the signature. In addition to the HMG1 and HMG2 proteins, HMG-domains occur in single or multiple copies in the following protein classes; the SOX family of transcription factors; SRY sex determining region Y protein and related proteins []; LEF1 lymphoid enhancer binding factor 1 []; SSRP recombination signal recognition protein; MTF1 mitochondrial transcription factor 1; UBF1/2 nucleolar transcription factors; Abf2 yeast ARS-binding factor []; and Saccharomyces cerevisiae transcription factors Ixr1, Rox1, Nhp6a, Nhp6b and Spp41.; GO: 0003677 DNA binding; PDB: 1I11_A 1J3C_A 1J3D_A 1WZ6_A 1WGF_A 2D7L_A 1GT0_D 3U2B_C 2CRJ_A 2CS1_A ....
Probab=97.79 E-value=3.9e-05 Score=52.23 Aligned_cols=42 Identities=33% Similarity=0.635 Sum_probs=37.5
Q ss_pred CCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214 123 QRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFP 164 (197)
Q Consensus 123 QR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~P 164 (197)
.|-++||..|+++....|++.+|+++..|.-..+++.|+..+
T Consensus 2 krP~~af~lf~~~~~~~~k~~~p~~~~~~i~~~~~~~W~~l~ 43 (69)
T PF00505_consen 2 KRPPNAFMLFCKEKRAKLKEENPDLSNKEISKILAQMWKNLS 43 (69)
T ss_dssp SSS--HHHHHHHHHHHHHHHHSTTSTHHHHHHHHHHHHHCSH
T ss_pred cCCCCHHHHHHHHHHHHHHHHhcccccccchhhHHHHHhcCC
Confidence 578999999999999999999999999999999999999753
No 7
>smart00398 HMG high mobility group.
Probab=97.78 E-value=6e-05 Score=50.67 Aligned_cols=43 Identities=33% Similarity=0.511 Sum_probs=39.9
Q ss_pred cCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214 122 RQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFP 164 (197)
Q Consensus 122 RQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~P 164 (197)
..|-+|+|..|+++....+++.+|+++..|....++..|+..+
T Consensus 2 pkrp~~~y~~f~~~~r~~~~~~~~~~~~~~i~~~~~~~W~~l~ 44 (70)
T smart00398 2 PKRPMSAFMLFSQENRAKIKAENPDLSNAEISKKLGERWKLLS 44 (70)
T ss_pred cCCCCcHHHHHHHHHHHHHHHHCcCCCHHHHHHHHHHHHHcCC
Confidence 4578999999999999999999999999999999999999754
No 8
>cd01389 MATA_HMG-box MATA_HMG-box, class I member of the HMG-box superfamily of DNA-binding proteins. These proteins contain a single HMG box, and bind the minor groove of DNA in a highly sequence-specific manner. Members include the fungal mating type gene products MC, MATA1 and Ste11.
Probab=97.72 E-value=7.9e-05 Score=53.03 Aligned_cols=43 Identities=16% Similarity=0.358 Sum_probs=40.4
Q ss_pred CCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214 123 QRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH 165 (197)
Q Consensus 123 QR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph 165 (197)
.|-|+||-.|+++....|+++||++++.|.-..++..|+..+.
T Consensus 3 kRP~naf~lf~~~~r~~~~~~~p~~~~~eisk~~g~~Wk~ls~ 45 (77)
T cd01389 3 PRPRNAFILYRQDKHAQLKTENPGLTNNEISRIIGRMWRSESP 45 (77)
T ss_pred CCCCcHHHHHHHHHHHHHHHHCCCCCHHHHHHHHHHHHhhCCH
Confidence 5889999999999999999999999999999999999998654
No 9
>PTZ00199 high mobility group protein; Provisional
Probab=97.72 E-value=8.6e-05 Score=55.70 Aligned_cols=49 Identities=27% Similarity=0.441 Sum_probs=43.2
Q ss_pred CCCcccCCCchhhhHHHHHHHHHHHhhCCCCC--HHHHHHHHHHhhccCCC
Q 029214 117 TTPEKRQRVPSAYNQFIKEEIQRIKANNPDIS--HREAFSTAAKNWAHFPH 165 (197)
Q Consensus 117 kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~--hkEaFs~aAknW~~~Ph 165 (197)
+.|.+..|-+|||..|+++.-..|+++||+++ ..|....++..|+..+.
T Consensus 18 kdp~~PKrP~sAY~~F~~~~R~~i~~~~P~~~~~~~evsk~ige~Wk~ls~ 68 (94)
T PTZ00199 18 KDPNAPKRALSAYMFFAKEKRAEIIAENPELAKDVAAVGKMVGEAWNKLSE 68 (94)
T ss_pred CCCCCCCCCCcHHHHHHHHHHHHHHHHCcCCcccHHHHHHHHHHHHHcCCH
Confidence 36777889999999999999999999999986 67888999999998653
No 10
>PF06244 DUF1014: Protein of unknown function (DUF1014); InterPro: IPR010422 This family consists of several hypothetical eukaryotic proteins of unknown function.
Probab=96.96 E-value=0.0011 Score=53.41 Aligned_cols=51 Identities=25% Similarity=0.513 Sum_probs=46.5
Q ss_pred CCCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCCcccc
Q 029214 117 TTPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPHIHFG 169 (197)
Q Consensus 117 kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Phihfg 169 (197)
+.||||-+ -||..|--.+|.+||++||++.+-+---..=|.|..+|..+|-
T Consensus 70 rHPErR~K--AAy~afeE~~Lp~lK~E~PgLrlsQ~kq~l~K~w~KSPeNP~N 120 (122)
T PF06244_consen 70 RHPERRMK--AAYKAFEERRLPELKEENPGLRLSQYKQMLWKEWQKSPENPFN 120 (122)
T ss_pred CCcchhHH--HHHHHHHHHHhHHHHhhCCCchHHHHHHHHHHHHhcCCCCCcc
Confidence 38999875 5999999999999999999999999888999999999998874
No 11
>KOG0381 consensus HMG box-containing protein [General function prediction only]
Probab=96.80 E-value=0.0031 Score=45.87 Aligned_cols=46 Identities=28% Similarity=0.441 Sum_probs=41.4
Q ss_pred cccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214 120 EKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH 165 (197)
Q Consensus 120 EKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph 165 (197)
...+|-+|||..|+.+.-.+||++||+++..|.-+.+..+|.....
T Consensus 21 ~~pkrp~sa~~~f~~~~~~~~k~~~p~~~~~~v~k~~g~~W~~l~~ 66 (96)
T KOG0381|consen 21 QAPKRPLSAFFLFSSEQRSKIKAENPGLSVGEVAKALGEMWKNLAE 66 (96)
T ss_pred CCCCCCCcHHHHHHHHHHHHHHHhCCCCCHHHHHHHHHHHHhcCCH
Confidence 3567889999999999999999999999999999999999987543
No 12
>KOG3223 consensus Uncharacterized conserved protein [Function unknown]
Probab=94.66 E-value=0.022 Score=49.98 Aligned_cols=55 Identities=27% Similarity=0.496 Sum_probs=48.8
Q ss_pred CCCCCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCCcccccc
Q 029214 115 HSTTPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPHIHFGLM 171 (197)
Q Consensus 115 ~~kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Phihfgl~ 171 (197)
+.+-||||-|+ ||--|=..++-|||.+||+++|-+-=-..-|.|...|..+|.-+
T Consensus 160 ddrHPEkRmrA--A~~afEe~~LPrLK~e~P~lrlsQ~Kqll~Kew~KsPDNP~Nq~ 214 (221)
T KOG3223|consen 160 DDRHPEKRMRA--AFKAFEEARLPRLKKENPGLRLSQYKQLLKKEWQKSPDNPFNQA 214 (221)
T ss_pred cccChHHHHHH--HHHHHHHhhchhhhhcCCCccHHHHHHHHHHHHhhCCCChhhHH
Confidence 33689999885 89999999999999999999999888888899999999998643
No 13
>PF11331 DUF3133: Protein of unknown function (DUF3133); InterPro: IPR021480 This eukaryotic family of proteins has no known function.
Probab=93.02 E-value=0.057 Score=37.23 Aligned_cols=38 Identities=29% Similarity=0.697 Sum_probs=29.3
Q ss_pred eecCCCcceeeeccccCCCccc---eeeeecCCCCCccccccc
Q 029214 15 YIPCNFCNIVLAVSVPCSSLLD---IVTVRCGHCSNLWSVNMA 54 (197)
Q Consensus 15 YV~CnfC~TILaVsVPcssL~~---tVTVRCGHCtnLlSVNmr 54 (197)
||-|..|..+|- +|-+.+.. .-.+|||.|+.++++.++
T Consensus 6 Fv~C~~C~~lLq--lP~~~~~~~k~~~klrCGaCs~vl~~s~~ 46 (46)
T PF11331_consen 6 FVVCSSCFELLQ--LPAKFSLSKKNQQKLRCGACSEVLSFSLP 46 (46)
T ss_pred EeECccHHHHHc--CCCccCCCccceeEEeCCCCceeEEEecC
Confidence 788999977665 47765443 568999999999988653
No 14
>TIGR02098 MJ0042_CXXC MJ0042 family finger-like domain. This domain contains a CXXCX(19)CXXC motif suggestive of both zinc fingers and thioredoxin, usually found at the N-terminus of prokaryotic proteins. One partially characterized gene, agmX, is among a large set in Myxococcus whose interruption affects adventurous gliding motility.
Probab=80.27 E-value=1.5 Score=27.40 Aligned_cols=34 Identities=35% Similarity=0.721 Sum_probs=24.0
Q ss_pred ecCCCcceeeeccccCCCcc-ceeeeecCCCCCcccc
Q 029214 16 IPCNFCNIVLAVSVPCSSLL-DIVTVRCGHCSNLWSV 51 (197)
Q Consensus 16 V~CnfC~TILaVsVPcssL~-~tVTVRCGHCtnLlSV 51 (197)
+.|..|.+..-|.. +.+- +...|+|++|.+.+.+
T Consensus 3 ~~CP~C~~~~~v~~--~~~~~~~~~v~C~~C~~~~~~ 37 (38)
T TIGR02098 3 IQCPNCKTSFRVVD--SQLGANGGKVRCGKCGHVWYA 37 (38)
T ss_pred EECCCCCCEEEeCH--HHcCCCCCEEECCCCCCEEEe
Confidence 67999999876653 2221 3347999999987764
No 15
>PF13719 zinc_ribbon_5: zinc-ribbon domain
Probab=79.33 E-value=1.4 Score=28.29 Aligned_cols=34 Identities=32% Similarity=0.665 Sum_probs=25.4
Q ss_pred ecCCCcceeeeccccCCCc-cceeeeecCCCCCcccc
Q 029214 16 IPCNFCNIVLAVSVPCSSL-LDIVTVRCGHCSNLWSV 51 (197)
Q Consensus 16 V~CnfC~TILaVsVPcssL-~~tVTVRCGHCtnLlSV 51 (197)
++|--|.|.+.| |=+.| -....|||++|.+.+.|
T Consensus 3 i~CP~C~~~f~v--~~~~l~~~~~~vrC~~C~~~f~v 37 (37)
T PF13719_consen 3 ITCPNCQTRFRV--PDDKLPAGGRKVRCPKCGHVFRV 37 (37)
T ss_pred EECCCCCceEEc--CHHHcccCCcEEECCCCCcEeeC
Confidence 689999998865 44443 34679999999987654
No 16
>PRK14892 putative transcription elongation factor Elf1; Provisional
Probab=73.92 E-value=2.6 Score=32.90 Aligned_cols=44 Identities=18% Similarity=0.530 Sum_probs=34.8
Q ss_pred eeecCCCcceeeeccccCCCccceeeeecCCCCCccccccchhccCC
Q 029214 14 CYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMAAAFQSL 60 (197)
Q Consensus 14 CYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr~llqsl 60 (197)
=++.|.||+. ..|+||... .+..+.|..|.---.-.+..|.++.
T Consensus 20 t~f~CP~Cge-~~v~v~~~k--~~~h~~C~~CG~y~~~~V~~l~epI 63 (99)
T PRK14892 20 KIFECPRCGK-VSISVKIKK--NIAIITCGNCGLYTEFEVPSVYDEV 63 (99)
T ss_pred cEeECCCCCC-eEeeeecCC--CcceEECCCCCCccCEECCccccch
Confidence 3789999995 688888888 7999999999877665566555543
No 17
>PF08073 CHDNT: CHDNT (NUC034) domain; InterPro: IPR012958 The CHD N-terminal domain is found in PHD/RING fingers and chromo domain-associated helicases [].; GO: 0003677 DNA binding, 0005524 ATP binding, 0008270 zinc ion binding, 0016818 hydrolase activity, acting on acid anhydrides, in phosphorus-containing anhydrides, 0006355 regulation of transcription, DNA-dependent, 0005634 nucleus
Probab=73.30 E-value=6.3 Score=28.22 Aligned_cols=33 Identities=18% Similarity=0.456 Sum_probs=26.0
Q ss_pred hhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccC
Q 029214 128 AYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHF 163 (197)
Q Consensus 128 aYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~ 163 (197)
+|..+|+ .-|-++||++.+-.-+...+..|++|
T Consensus 18 ~Fsq~vR---P~l~~~NPk~~~sKl~~l~~AKwrEF 50 (55)
T PF08073_consen 18 AFSQHVR---PLLAKANPKAPMSKLMMLLQAKWREF 50 (55)
T ss_pred HHHHHHH---HHHHHHCCCCcHHHHHHHHHHHHHHH
Confidence 3444444 34567999999999999999999986
No 18
>TIGR01562 FdhE formate dehydrogenase accessory protein FdhE. The only sequence scoring between trusted and noise is that from Aquifex aeolicus, which shows certain structural differences from the proteobacterial forms in the alignment. However it is notable that A. aeolicus also has a sequence scoring above trusted to the alpha subunit of formate dehydrogenase (TIGR01553).
Probab=72.45 E-value=2 Score=39.24 Aligned_cols=27 Identities=33% Similarity=0.897 Sum_probs=23.1
Q ss_pred CceeeecCCCcceeeeccccCCCccceeeeecCCCCC
Q 029214 11 EQLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSN 47 (197)
Q Consensus 11 E~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtn 47 (197)
+..=|.||++|.| -...|-++|.+|.|
T Consensus 206 ~G~RyL~CslC~t----------eW~~~R~~C~~Cg~ 232 (305)
T TIGR01562 206 TGLRYLSCSLCAT----------EWHYVRVKCSHCEE 232 (305)
T ss_pred CCceEEEcCCCCC----------cccccCccCCCCCC
Confidence 4566999999977 57889999999998
No 19
>PF04032 Rpr2: RNAse P Rpr2/Rpp21/SNM1 subunit domain; InterPro: IPR007175 This family contains a ribonuclease P subunit of human and yeast. Other members of the family include the probable archaeal homologues. This subunit possibly binds the precursor tRNA [].; PDB: 2K3R_A 2KI7_B 2ZAE_B 1X0T_A.
Probab=71.62 E-value=1.9 Score=30.69 Aligned_cols=32 Identities=25% Similarity=0.580 Sum_probs=20.2
Q ss_pred ecCCCcceeeeccccCCC-------ccceeeeecCCCCC
Q 029214 16 IPCNFCNIVLAVSVPCSS-------LLDIVTVRCGHCSN 47 (197)
Q Consensus 16 V~CnfC~TILaVsVPcss-------L~~tVTVRCGHCtn 47 (197)
.-|.-|.++|.-|+-|+- .-+.|.++|..|.+
T Consensus 47 ~~Ck~C~~~liPG~~~~vri~~~~~~~~~l~~~C~~C~~ 85 (85)
T PF04032_consen 47 TICKKCGSLLIPGVNCSVRIRKKKKKKNFLVYTCLNCGH 85 (85)
T ss_dssp TB-TTT--B--CTTTEEEEEE---SSS-EEEEEETTTTE
T ss_pred ccccCCCCEEeCCCccEEEEEecCCCCCEEEEEccccCC
Confidence 458999999999998863 24688999999963
No 20
>PF05129 Elf1: Transcription elongation factor Elf1 like; InterPro: IPR007808 This family of uncharacterised, mostly short, proteins contain a putative zinc binding domain with four conserved cysteines.; PDB: 1WII_A.
Probab=71.33 E-value=2.3 Score=31.81 Aligned_cols=42 Identities=24% Similarity=0.453 Sum_probs=26.2
Q ss_pred eecCCCcceeeeccccCCCccceeeeecCCCCCccccccchh
Q 029214 15 YIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMAAA 56 (197)
Q Consensus 15 YV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr~l 56 (197)
+-.|-||+.--+|+|=.+.-..+-++.||-|.--...++..|
T Consensus 22 ~F~CPfC~~~~sV~v~idkk~~~~~~~C~~Cg~~~~~~i~~L 63 (81)
T PF05129_consen 22 VFDCPFCNHEKSVSVKIDKKEGIGILSCRVCGESFQTKINPL 63 (81)
T ss_dssp ----TTT--SS-EEEEEETTTTEEEEEESSS--EEEEE--SS
T ss_pred eEcCCcCCCCCeEEEEEEccCCEEEEEecCCCCeEEEccCcc
Confidence 457999999999999999999999999999954444444433
No 21
>PRK03954 ribonuclease P protein component 4; Validated
Probab=70.85 E-value=2.4 Score=34.25 Aligned_cols=32 Identities=25% Similarity=0.569 Sum_probs=24.6
Q ss_pred CCCcceeeeccccCC-Cccc----eeeeecCCCCCcc
Q 029214 18 CNFCNIVLAVSVPCS-SLLD----IVTVRCGHCSNLW 49 (197)
Q Consensus 18 CnfC~TILaVsVPcs-sL~~----tVTVRCGHCtnLl 49 (197)
|-.|+|.|.-||-|. ++-+ -|.|+|..|...-
T Consensus 67 CK~C~t~LiPG~n~~vRi~~~~~~~vvitCl~CG~~k 103 (121)
T PRK03954 67 CKRCHSFLVPGVNARVRLRQKRMPHVVITCLECGHIM 103 (121)
T ss_pred hhcCCCeeecCCceEEEEecCCcceEEEECccCCCEE
Confidence 889999999888765 2222 5899999998753
No 22
>PF13717 zinc_ribbon_4: zinc-ribbon domain
Probab=69.44 E-value=5.5 Score=25.54 Aligned_cols=30 Identities=27% Similarity=0.739 Sum_probs=23.3
Q ss_pred ecCCCcceeeecc---ccCCCccceeeeecCCCCCcc
Q 029214 16 IPCNFCNIVLAVS---VPCSSLLDIVTVRCGHCSNLW 49 (197)
Q Consensus 16 V~CnfC~TILaVs---VPcssL~~tVTVRCGHCtnLl 49 (197)
+.|.-|.+...|. || =+.+.|||+.|.+.+
T Consensus 3 i~Cp~C~~~y~i~d~~ip----~~g~~v~C~~C~~~f 35 (36)
T PF13717_consen 3 ITCPNCQAKYEIDDEKIP----PKGRKVRCSKCGHVF 35 (36)
T ss_pred EECCCCCCEEeCCHHHCC----CCCcEEECCCCCCEe
Confidence 6799999988765 44 245799999999865
No 23
>PF10122 Mu-like_Com: Mu-like prophage protein Com; InterPro: IPR019294 Members of this entry belong to the Com family of proteins that act as translational regulators of mom [, ].
Probab=67.51 E-value=3.2 Score=29.42 Aligned_cols=32 Identities=28% Similarity=0.598 Sum_probs=23.4
Q ss_pred ecCCCcceeeeccccCCCccceeeeecCCCCCcccc
Q 029214 16 IPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSV 51 (197)
Q Consensus 16 V~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSV 51 (197)
+||..|+..||-+- -+..+.++|..|..+-.|
T Consensus 5 iRC~~CnklLa~~g----~~~~leIKCpRC~tiN~~ 36 (51)
T PF10122_consen 5 IRCGHCNKLLAKAG----EVIELEIKCPRCKTINHV 36 (51)
T ss_pred eeccchhHHHhhhc----CccEEEEECCCCCccceE
Confidence 68888888888642 344788888888876554
No 24
>PRK03564 formate dehydrogenase accessory protein FdhE; Provisional
Probab=64.62 E-value=3.8 Score=37.66 Aligned_cols=27 Identities=33% Similarity=0.922 Sum_probs=23.0
Q ss_pred CceeeecCCCcceeeeccccCCCccceeeeecCCCCC
Q 029214 11 EQLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSN 47 (197)
Q Consensus 11 E~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtn 47 (197)
+-.=|.||++|.| ....|-++|.+|.|
T Consensus 208 ~G~RyL~CslC~t----------eW~~~R~~C~~Cg~ 234 (309)
T PRK03564 208 QGLRYLHCNLCES----------EWHVVRVKCSNCEQ 234 (309)
T ss_pred CCceEEEcCCCCC----------cccccCccCCCCCC
Confidence 5567999999977 57889999999997
No 25
>PF09788 Tmemb_55A: Transmembrane protein 55A; InterPro: IPR019178 Members of this family catalyse the hydrolysis of the 4-position phosphate of phosphatidylinositol 4,5-bisphosphate, in the reaction: 1-phosphatidyl-myo-inositol 4,5-bisphosphate + H(2)O = 1-phosphatidyl-1D-myo-inositol 5-phosphate + phosphate.
Probab=58.76 E-value=7.3 Score=35.37 Aligned_cols=39 Identities=26% Similarity=0.534 Sum_probs=25.3
Q ss_pred CceeeecCCCcceeeeccccCCCccceeeeecCCCCCcccccc
Q 029214 11 EQLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNM 53 (197)
Q Consensus 11 E~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNm 53 (197)
..-|=|.|+.|+...+...+-++- -.||-||-.++||.-
T Consensus 153 p~~~rv~CghC~~~Fl~~~~~~~t----lARCPHCrKvSSVG~ 191 (256)
T PF09788_consen 153 PGSCRVICGHCSNTFLFNTLTSNT----LARCPHCRKVSSVGP 191 (256)
T ss_pred CCceeEECCCCCCcEeccCCCCCc----cccCCCCceeccccc
Confidence 345777888887776665544221 267888888887753
No 26
>KOG4684 consensus Uncharacterized conserved protein, contains C4-type Zn-finger [General function prediction only]
Probab=57.18 E-value=4.2 Score=36.82 Aligned_cols=16 Identities=38% Similarity=0.893 Sum_probs=10.6
Q ss_pred eeeeecCCCCCccccc
Q 029214 37 IVTVRCGHCSNLWSVN 52 (197)
Q Consensus 37 tVTVRCGHCtnLlSVN 52 (197)
.+-|+||||++..--|
T Consensus 168 gcRV~CgHC~~tFLfn 183 (275)
T KOG4684|consen 168 GCRVKCGHCNETFLFN 183 (275)
T ss_pred ceEEEecCccceeehh
Confidence 3778888887764433
No 27
>COG3058 FdhE Uncharacterized protein involved in formate dehydrogenase formation [Posttranslational modification, protein turnover, chaperones]
Probab=54.61 E-value=2 Score=39.72 Aligned_cols=34 Identities=24% Similarity=0.647 Sum_probs=26.8
Q ss_pred CCceeeecCCCcceeeeccccCCCccceeeeecCCCCCcccccc
Q 029214 10 PEQLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNM 53 (197)
Q Consensus 10 ~E~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNm 53 (197)
.+-+=|.|||.|-| ....|-|+|-+|.+---+++
T Consensus 206 ~~GlRYL~CslC~t----------eW~~VR~KC~nC~~t~~l~y 239 (308)
T COG3058 206 EQGLRYLHCSLCET----------EWHYVRVKCSNCEQSKKLHY 239 (308)
T ss_pred cccchhhhhhhHHH----------HHHHHHHHhccccccCCccc
Confidence 46788999999987 56789999999998544433
No 28
>COG5648 NHP6B Chromatin-associated proteins containing the HMG domain [Chromatin structure and dynamics]
Probab=52.52 E-value=28 Score=30.93 Aligned_cols=45 Identities=24% Similarity=0.448 Sum_probs=39.7
Q ss_pred CcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccC
Q 029214 119 PEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHF 163 (197)
Q Consensus 119 PEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~ 163 (197)
|-=..|--|||-.|..+.=.+|...+|+++.-|.=..+++.|+..
T Consensus 68 pN~PKRp~sayf~y~~~~R~ei~~~~p~l~~~e~~k~~~e~WK~L 112 (211)
T COG5648 68 PNGPKRPLSAYFLYSAENRDEIRKENPKLTFGEVGKLLSEKWKEL 112 (211)
T ss_pred CCCCCCchhHHHHHHHHHHHHHHHhCCCCChHHHHHHHHHHHHhc
Confidence 333457889999999999999999999999999999999999975
No 29
>PF04216 FdhE: Protein involved in formate dehydrogenase formation; InterPro: IPR006452 This family of sequences describe an accessory protein required for the assembly of formate dehydrogenase of certain proteobacteria although not present in the final complex []. The exact nature of the function of FdhE in the assembly of the complex is unknown, but considering the presence of selenocysteine, molybdopterin, iron-sulphur clusters and cytochrome b556, it is likely to be involved in the insertion of cofactors. ; GO: 0005737 cytoplasm; PDB: 2FIY_B.
Probab=51.96 E-value=4.7 Score=35.40 Aligned_cols=32 Identities=22% Similarity=0.588 Sum_probs=17.3
Q ss_pred eeeecCCCcceeeeccccCCCccceeeeecCCCCCccccccc
Q 029214 13 LCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMA 54 (197)
Q Consensus 13 lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr 54 (197)
.=|-+|.+|.|- ...+-++|.+|.|--...+.
T Consensus 195 ~R~L~Cs~C~t~----------W~~~R~~Cp~Cg~~~~~~l~ 226 (290)
T PF04216_consen 195 KRYLHCSLCGTE----------WRFVRIKCPYCGNTDHEKLE 226 (290)
T ss_dssp EEEEEETTT--E----------EE--TTS-TTT---SS-EEE
T ss_pred cEEEEcCCCCCe----------eeecCCCCcCCCCCCCccee
Confidence 468899999874 56778899999986555444
No 30
>PF11331 DUF3133: Protein of unknown function (DUF3133); InterPro: IPR021480 This eukaryotic family of proteins has no known function.
Probab=49.24 E-value=10 Score=26.21 Aligned_cols=18 Identities=33% Similarity=0.654 Sum_probs=15.9
Q ss_pred eeeecCCCcceeeecccc
Q 029214 13 LCYIPCNFCNIVLAVSVP 30 (197)
Q Consensus 13 lCYV~CnfC~TILaVsVP 30 (197)
.==+||+-|..||-+++|
T Consensus 29 ~~klrCGaCs~vl~~s~~ 46 (46)
T PF11331_consen 29 QQKLRCGACSEVLSFSLP 46 (46)
T ss_pred eeEEeCCCCceeEEEecC
Confidence 556899999999999987
No 31
>PF06382 DUF1074: Protein of unknown function (DUF1074); InterPro: IPR024460 This family consists of several proteins which appear to be specific to Insecta. The function of this family is unknown.
Probab=48.84 E-value=26 Score=30.56 Aligned_cols=35 Identities=23% Similarity=0.574 Sum_probs=30.1
Q ss_pred hhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214 127 SAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH 165 (197)
Q Consensus 127 SaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph 165 (197)
.+|=.||.+ .+..|.++..+|....||+.|...+.
T Consensus 84 naYLNFLRe----FRrkh~~L~p~dlI~~AAraW~rLSe 118 (183)
T PF06382_consen 84 NAYLNFLRE----FRRKHCGLSPQDLIQRAARAWCRLSE 118 (183)
T ss_pred hHHHHHHHH----HHHHccCCCHHHHHHHHHHHHHhCCH
Confidence 578888764 77899999999999999999987654
No 32
>KOG4684 consensus Uncharacterized conserved protein, contains C4-type Zn-finger [General function prediction only]
Probab=44.28 E-value=11 Score=34.29 Aligned_cols=33 Identities=36% Similarity=0.828 Sum_probs=28.5
Q ss_pred eeeecCCCcceeeeccccCCCccceee---eecCCCCCcccccc
Q 029214 13 LCYIPCNFCNIVLAVSVPCSSLLDIVT---VRCGHCSNLWSVNM 53 (197)
Q Consensus 13 lCYV~CnfC~TILaVsVPcssL~~tVT---VRCGHCtnLlSVNm 53 (197)
-|-|-|+.|+.+. ||+|.| -||-||-..+||--
T Consensus 168 gcRV~CgHC~~tF--------Lfnt~tnaLArCPHCrKvSsvGs 203 (275)
T KOG4684|consen 168 GCRVKCGHCNETF--------LFNTLTNALARCPHCRKVSSVGS 203 (275)
T ss_pred ceEEEecCcccee--------ehhhHHHHHhcCCcccchhhhhh
Confidence 4899999998876 789998 59999999999844
No 33
>TIGR01053 LSD1 zinc finger domain, LSD1 subclass. This model describes a putative zinc finger domain found in three closely spaced copies in Arabidopsis protein LSD1 and in two copies in other proteins from the same species. The motif resembles CxxCRxxLMYxxGASxVxCxxC
Probab=40.87 E-value=26 Score=22.18 Aligned_cols=25 Identities=28% Similarity=0.681 Sum_probs=13.8
Q ss_pred ecCCCcceeeeccccCCCccceeeeecCCCC
Q 029214 16 IPCNFCNIVLAVSVPCSSLLDIVTVRCGHCS 46 (197)
Q Consensus 16 V~CnfC~TILaVsVPcssL~~tVTVRCGHCt 46 (197)
|.|+-|.|+|+.- ...-.|||--|.
T Consensus 2 ~~C~~C~t~L~yP------~gA~~vrCs~C~ 26 (31)
T TIGR01053 2 VVCGGCRTLLMYP------RGASSVRCALCQ 26 (31)
T ss_pred cCcCCCCcEeecC------CCCCeEECCCCC
Confidence 4566666666542 233456666664
No 34
>KOG0526 consensus Nucleosome-binding factor SPN, POB3 subunit [Transcription; Replication, recombination and repair; Chromatin structure and dynamics]
Probab=38.70 E-value=42 Score=33.90 Aligned_cols=44 Identities=25% Similarity=0.483 Sum_probs=39.6
Q ss_pred CCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccC
Q 029214 118 TPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHF 163 (197)
Q Consensus 118 PPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~ 163 (197)
-|-+..|+-|||=.|...+-..||+. +|+.-|.=+.+...|+.-
T Consensus 532 dpnapkra~sa~m~w~~~~r~~ik~d--gi~~~dv~kk~g~~wk~m 575 (615)
T KOG0526|consen 532 DPNAPKRATSAYMLWLNASRESIKED--GISVGDVAKKAGEKWKQM 575 (615)
T ss_pred CCCCCccchhHHHHHHHhhhhhHhhc--CchHHHHHHHHhHHHhhh
Confidence 45566899999999999999999999 999999999999999963
No 35
>PF03811 Zn_Tnp_IS1: InsA N-terminal domain; InterPro: IPR003220 Insertion elements are mobile elements in DNA, usually encoding proteins required for transposition, for example transposases. Protein InsA is absolutely required for transposition of insertion element 1. This entry represents a short zinc binding domain found in IS1 InsA family protein. It is found at the N terminus of the protein and may be a DNA-binding domain.; GO: 0006313 transposition, DNA-mediated
Probab=37.54 E-value=23 Score=23.01 Aligned_cols=17 Identities=24% Similarity=0.565 Sum_probs=13.8
Q ss_pred cceeeeecCCCCCcccc
Q 029214 35 LDIVTVRCGHCSNLWSV 51 (197)
Q Consensus 35 ~~tVTVRCGHCtnLlSV 51 (197)
|.+|+|.|-+|.+-.+|
T Consensus 1 Ma~i~v~CP~C~s~~~v 17 (36)
T PF03811_consen 1 MAKIDVHCPRCQSTEGV 17 (36)
T ss_pred CCcEeeeCCCCCCCCcc
Confidence 46899999999887654
No 36
>COG4416 Com Mu-like prophage protein Com [General function prediction only]
Probab=37.44 E-value=7.9 Score=28.23 Aligned_cols=14 Identities=36% Similarity=0.987 Sum_probs=11.3
Q ss_pred eeeeecCCCCCccc
Q 029214 37 IVTVRCGHCSNLWS 50 (197)
Q Consensus 37 tVTVRCGHCtnLlS 50 (197)
+-|.||.||+-||-
T Consensus 2 ~~tiRC~~CnKlLa 15 (60)
T COG4416 2 MQTIRCAKCNKLLA 15 (60)
T ss_pred ceeeehHHHhHHHH
Confidence 45899999998764
No 37
>PF10963 DUF2765: Protein of unknown function (DUF2765); InterPro: IPR024406 This family of proteins with no known function is found in phages and suspected prophages.
Probab=33.12 E-value=65 Score=24.63 Aligned_cols=14 Identities=29% Similarity=0.567 Sum_probs=10.0
Q ss_pred HHHHHHHHhhCCCC
Q 029214 134 KEEIQRIKANNPDI 147 (197)
Q Consensus 134 K~ei~riK~~~P~i 147 (197)
|+.+..+-+++||.
T Consensus 48 KeaL~~lle~~PGa 61 (83)
T PF10963_consen 48 KEALKELLEENPGA 61 (83)
T ss_pred HHHHHHHHHHCCCH
Confidence 56677777777776
No 38
>PF05047 L51_S25_CI-B8: Mitochondrial ribosomal protein L51 / S25 / CI-B8 domain ; InterPro: IPR007741 Proteins containing this domain are located in the mitochondrion and include ribosomal protein L51, and S25. This domain is also found in mitochondrial NADH-ubiquinone oxidoreductase B8 subunit (CI-B8) 1.6.5.3 from EC. It is not known whether all members of this family form part of the NADH-ubiquinone oxidoreductase and whether they are also all ribosomal proteins.; PDB: 1S3A_A.
Probab=31.61 E-value=40 Score=22.21 Aligned_cols=18 Identities=28% Similarity=0.719 Sum_probs=15.5
Q ss_pred hHHHHHHHHHHHhhCCCC
Q 029214 130 NQFIKEEIQRIKANNPDI 147 (197)
Q Consensus 130 N~FmK~ei~riK~~~P~i 147 (197)
..|+++.+..|+..||++
T Consensus 2 R~F~~~~lp~l~~~NP~v 19 (52)
T PF05047_consen 2 RDFLKNNLPTLKYHNPQV 19 (52)
T ss_dssp HHHHHHTHHHHHHHSTT-
T ss_pred HhHHHHhHHHHHHHCCCc
Confidence 369999999999999986
No 39
>PF01020 Ribosomal_L40e: Ribosomal L40e family; InterPro: IPR001975 Ribosomes are the particles that catalyse mRNA-directed protein synthesis in all organisms. The codons of the mRNA are exposed on the ribosome to allow tRNA binding. This leads to the incorporation of amino acids into the growing polypeptide chain in accordance with the genetic information. Incoming amino acid monomers enter the ribosomal A site in the form of aminoacyl-tRNAs complexed with elongation factor Tu (EF-Tu) and GTP. The growing polypeptide chain, situated in the P site as peptidyl-tRNA, is then transferred to aminoacyl-tRNA and the new peptidyl-tRNA, extended by one residue, is translocated to the P site with the aid the elongation factor G (EF-G) and GTP as the deacylated tRNA is released from the ribosome through one or more exit sites [, ]. About 2/3 of the mass of the ribosome consists of RNA and 1/3 of protein. The proteins are named in accordance with the subunit of the ribosome which they belong to - the small (S1 to S31) and the large (L1 to L44). Usually they decorate the rRNA cores of the subunits. Many ribosomal proteins, particularly those of the large subunit, are composed of a globular, surfaced-exposed domain with long finger-like projections that extend into the rRNA core to stabilise its structure. Most of the proteins interact with multiple RNA elements, often from different domains. In the large subunit, about 1/3 of the 23S rRNA nucleotides are at least in van der Waal's contact with protein, and L22 interacts with all six domains of the 23S rRNA. Proteins S4 and S7, which initiate assembly of the 16S rRNA, are located at junctions of five and four RNA helices, respectively. In this way proteins serve to organise and stabilise the rRNA tertiary structure. While the crucial activities of decoding and peptide transfer are RNA based, proteins play an active role in functions that may have evolved to streamline the process of protein synthesis. In addition to their function in the ribosome, many ribosomal proteins have some function 'outside' the ribosome [, ]. This family contains the L40 ribosomal protein from both archaea and eukaryotes. Bovine ribosomal protein L40 has been identified as a secondary RNA binding protein []. L40 is fused to a ubiquitin protein [].; GO: 0003735 structural constituent of ribosome, 0006412 translation, 0005840 ribosome; PDB: 3IZS_p 3IZR_p 2AYJ_A 4A1B_K 4A19_K 4A18_K 4A1D_K.
Probab=31.31 E-value=18 Score=25.80 Aligned_cols=9 Identities=56% Similarity=1.258 Sum_probs=5.5
Q ss_pred ecCCCCCcc
Q 029214 41 RCGHCSNLW 49 (197)
Q Consensus 41 RCGHCtnLl 49 (197)
+|||++||-
T Consensus 38 kCGhsn~LR 46 (52)
T PF01020_consen 38 KCGHSNNLR 46 (52)
T ss_dssp SCTS-S-EE
T ss_pred cCCCCcccC
Confidence 399999874
No 40
>COG4888 Uncharacterized Zn ribbon-containing protein [General function prediction only]
Probab=30.43 E-value=37 Score=27.31 Aligned_cols=45 Identities=18% Similarity=0.362 Sum_probs=33.0
Q ss_pred eecCCCcceeeeccccCCCccceeeeecCCCCCccccccchhccC
Q 029214 15 YIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMAAAFQS 59 (197)
Q Consensus 15 YV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr~llqs 59 (197)
|--|-||+-...|+--.+--.++-|+-||-|.--.-+-.++++++
T Consensus 22 ~FtCp~Cghe~vs~ctvkk~~~~g~~~Cg~CGls~e~ev~~l~~~ 66 (104)
T COG4888 22 TFTCPRCGHEKVSSCTVKKTVNIGTAVCGNCGLSFECEVPELSEP 66 (104)
T ss_pred eEecCccCCeeeeEEEEEecCceeEEEcccCcceEEEeccccccc
Confidence 667999999999887767777888999999974433444444443
No 41
>PF12876 Cellulase-like: Sugar-binding cellulase-like; InterPro: IPR024778 O-Glycosyl hydrolases 3.2.1. from EC are a widespread group of enzymes that hydrolyse the glycosidic bond between two or more carbohydrates, or between a carbohydrate and a non-carbohydrate moiety. A classification system for glycosyl hydrolases, based on sequence similarity, has led to the definition of 85 different families [, ]. This classification is available on the CAZy (CArbohydrate-Active EnZymes) web site. This entry represents a family of putative cellulase enzymes.; PDB: 3GYC_B.
Probab=30.35 E-value=73 Score=23.06 Aligned_cols=31 Identities=26% Similarity=0.428 Sum_probs=22.8
Q ss_pred CCcccCCCchhhhHHHHHHHHHHHhhCCCCC
Q 029214 118 TPEKRQRVPSAYNQFIKEEIQRIKANNPDIS 148 (197)
Q Consensus 118 PPEKRQR~PSaYN~FmK~ei~riK~~~P~i~ 148 (197)
|.+.......+|-.|+++-++.||+.+|+.+
T Consensus 29 ~~~~~~~~~~~~~~~l~~~~~~iR~~dP~~p 59 (88)
T PF12876_consen 29 PAEWGDPKAEAYAEWLKEAFRWIRAVDPSQP 59 (88)
T ss_dssp TT-TT-TTSHHHHHHHHHHHHHHHTT-TTS-
T ss_pred cccccchhHHHHHHHHHHHHHHHHHhCCCCc
Confidence 3344555778999999999999999999764
No 42
>KOG4715 consensus SWI/SNF-related matrix-associated actin-dependent regulator of chromatin [Chromatin structure and dynamics]
Probab=29.70 E-value=68 Score=30.76 Aligned_cols=46 Identities=24% Similarity=0.460 Sum_probs=39.8
Q ss_pred CCCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214 117 TTPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH 165 (197)
Q Consensus 117 kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph 165 (197)
|||||. .-.|=+|-+.=-..+|+.||++--=|.=+.+++.|.+.|.
T Consensus 63 kppekp---l~pymrySrkvWd~VkA~nPe~kLWeiGK~Ig~mW~dLpd 108 (410)
T KOG4715|consen 63 KPPEKP---LMPYMRYSRKVWDQVKASNPELKLWEIGKIIGGMWLDLPD 108 (410)
T ss_pred CCCCcc---cchhhHHhhhhhhhhhccCcchHHHHHHHHHHHHHhhCcc
Confidence 477764 5679999888899999999999999999999999998875
No 43
>PF14599 zinc_ribbon_6: Zinc-ribbon; PDB: 2K2D_A.
Probab=28.68 E-value=35 Score=24.60 Aligned_cols=30 Identities=30% Similarity=0.703 Sum_probs=15.9
Q ss_pred ceeeecCCCcceeeeccccCCCccceeeeecCCCCC
Q 029214 12 QLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSN 47 (197)
Q Consensus 12 ~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtn 47 (197)
..+.|-||=|..-=- | -|-++..||+||.+
T Consensus 27 ~~v~IlCNDC~~~s~--v----~fH~lg~KC~~C~S 56 (61)
T PF14599_consen 27 KKVWILCNDCNAKSE--V----PFHFLGHKCSHCGS 56 (61)
T ss_dssp -EEEEEESSS--EEE--E----E--TT----TTTS-
T ss_pred CEEEEECCCCCCccc--e----eeeHhhhcCCCCCC
Confidence 579999999987543 3 37788899999975
No 44
>PF10159 MMtag: Kinase phosphorylation protein; InterPro: IPR019315 This entry represents a glycine-rich domain that is the most highly conserved region of a family of proteins that, in vertebrates, are associated with tumours in multiple myelomas. The region may contain phosphorylation sites for several protein kinases, as well as N-myristoylation sites and nuclear localisation signals, so it might act as a signal molecule in the nucleus [].
Probab=28.50 E-value=52 Score=25.13 Aligned_cols=13 Identities=54% Similarity=0.567 Sum_probs=11.6
Q ss_pred HHHHHHHHHHhhC
Q 029214 132 FIKEEIQRIKANN 144 (197)
Q Consensus 132 FmK~ei~riK~~~ 144 (197)
=.++||++||+..
T Consensus 57 ~~~eE~~~iK~~E 69 (78)
T PF10159_consen 57 ERKEEIRRIKEAE 69 (78)
T ss_pred hHHHHHHHHHHHH
Confidence 6799999999986
No 45
>PF05180 zf-DNL: DNL zinc finger; InterPro: IPR007853 Zinc finger (Znf) domains are relatively small protein motifs which contain multiple finger-like protrusions that make tandem contacts with their target molecule. Some of these domains bind zinc, but many do not; instead binding other metals such as iron, or no metal at all. For example, some family members form salt bridges to stabilise the finger-like folds. They were first identified as a DNA-binding motif in transcription factor TFIIIA from Xenopus laevis (African clawed frog), however they are now recognised to bind DNA, RNA, protein and/or lipid substrates [, , , , ]. Their binding properties depend on the amino acid sequence of the finger domains and of the linker between fingers, as well as on the higher-order structures and the number of fingers. Znf domains are often found in clusters, where fingers can have different binding specificities. There are many superfamilies of Znf motifs, varying in both sequence and structure. They display considerable versatility in binding modes, even between members of the same class (e.g. some bind DNA, others protein), suggesting that Znf motifs are stable scaffolds that have evolved specialised functions. For example, Znf-containing proteins function in gene transcription, translation, mRNA trafficking, cytoskeleton organisation, epithelial development, cell adhesion, protein folding, chromatin remodelling and zinc sensing, to name but a few []. Zinc-binding motifs are stable structures, and they rarely undergo conformational changes upon binding their target. The DNL-type zinc finger is found in Tim15, a zinc finger protein essential for protein import into mitochondria. Mitochondrial functions rely on the correct transport of resident proteins synthesized in the cytosol to mitochondria. Protein import into mitochondria is mediated by membrane protein complexes, protein translocators, in the outer and inner mitochondrial membranes, in cooperation with their assistant proteins in the cytosol, intermembrane space and matrix. Proteins destined to the mitochondrial matrix cross the outer membrane with the aid of the outer membrane translocator, the tOM40 complex, and then the inner membrane with the aid of the inner membrane translocator, the TIM23 complex, and mitochondrial motor and chaperone (MMC) proteins including mitochondrial heat- shock protein 70 (mtHsp70), and translocase in the inner mitochondrial membrane (Tim)15. Tim15 is also known as zinc finger motif (Zim)17 or mtHsp70 escort protein (Hep)1. Tim15 contains a zinc-finger motif (CXXC and CXXC) of ~100 residues, which has been named DNL after a short C-terminal motif of D(N/H)L [, , ]. The DNL-type zinc finger is an L-shaped molecule. The two CXXC motifs are located at the end of the L, and are sandwiched by two- stranded antiparallel beta-sheets. Two short alpha-helices constitute another leg of the L. The outer (convex) face of the L has a large acidic groove, which is lined with five acidic residues, whereas the inner (concave) face of the L has two positively charged residues, next to the CXXC motifs []. This entry represents the DNL-type zinc finger.; GO: 0008270 zinc ion binding; PDB: 2E2Z_A.
Probab=28.02 E-value=33 Score=25.27 Aligned_cols=32 Identities=28% Similarity=0.482 Sum_probs=15.9
Q ss_pred cCCCcceeeeccccCCC-ccceeeeecCCCCCc
Q 029214 17 PCNFCNIVLAVSVPCSS-LLDIVTVRCGHCSNL 48 (197)
Q Consensus 17 ~CnfC~TILaVsVPcss-L~~tVTVRCGHCtnL 48 (197)
-|+-|+|-=+-.+-=.+ =--+|-|||+.|.|.
T Consensus 6 TC~~C~~Rs~~~~sk~aY~~GvViv~C~gC~~~ 38 (66)
T PF05180_consen 6 TCNKCGTRSAKMFSKQAYHKGVVIVQCPGCKNR 38 (66)
T ss_dssp EETTTTEEEEEEEEHHHHHTSEEEEE-TTS--E
T ss_pred EcCCCCCccceeeCHHHHhCCeEEEECCCCcce
Confidence 36666665442221111 124799999999985
No 46
>PF02892 zf-BED: BED zinc finger; InterPro: IPR003656 Zinc finger (Znf) domains are relatively small protein motifs which contain multiple finger-like protrusions that make tandem contacts with their target molecule. Some of these domains bind zinc, but many do not; instead binding other metals such as iron, or no metal at all. For example, some family members form salt bridges to stabilise the finger-like folds. They were first identified as a DNA-binding motif in transcription factor TFIIIA from Xenopus laevis (African clawed frog), however they are now recognised to bind DNA, RNA, protein and/or lipid substrates [, , , , ]. Their binding properties depend on the amino acid sequence of the finger domains and of the linker between fingers, as well as on the higher-order structures and the number of fingers. Znf domains are often found in clusters, where fingers can have different binding specificities. There are many superfamilies of Znf motifs, varying in both sequence and structure. They display considerable versatility in binding modes, even between members of the same class (e.g. some bind DNA, others protein), suggesting that Znf motifs are stable scaffolds that have evolved specialised functions. For example, Znf-containing proteins function in gene transcription, translation, mRNA trafficking, cytoskeleton organisation, epithelial development, cell adhesion, protein folding, chromatin remodelling and zinc sensing, to name but a few []. Zinc-binding motifs are stable structures, and they rarely undergo conformational changes upon binding their target. This entry represents predicted BED-type zinc finger domains. The BED finger which was named after the Drosophila proteins BEAF and DREF, is found in one or more copies in cellular regulatory factors and transposases from plants, animals and fungi. The BED finger is an about 50 to 60 amino acid residues domain that contains a characteristic motif with two highly conserved aromatic positions, as well as a shared pattern of cysteines and histidines that is predicted to form a zinc finger. As diverse BED fingers are able to bind DNA, it has been suggested that DNA-binding is the general function of this domain []. Some proteins known to contain a BED domain include animal, plant and fungi AC1 and Hobo-like transposases; Caenorhabditis elegans Dpy-20 protein, a predicted cuticular gene transcriptional regulator; Drosophila BEAF (boundary element-associated factor), thought to be involved in chromatin insulation; Drosophila DREF, a transcriptional regulator for S-phase genes; and tobacco 3AF1 and tomato E4/E8-BP1, light- and ethylene-regulated DNA binding proteins that contain two BED fingers. More information about these proteins can be found at Protein of the Month: Zinc Fingers [].; GO: 0003677 DNA binding; PDB: 2DJR_A 2CT5_A.
Probab=27.89 E-value=27 Score=22.13 Aligned_cols=17 Identities=24% Similarity=0.522 Sum_probs=10.1
Q ss_pred ceeeecCCCcceeeecc
Q 029214 12 QLCYIPCNFCNIVLAVS 28 (197)
Q Consensus 12 ~lCYV~CnfC~TILaVs 28 (197)
..-++.|.+|..++..+
T Consensus 13 ~~~~a~C~~C~~~~~~~ 29 (45)
T PF02892_consen 13 DKKKAKCKYCGKVIKYS 29 (45)
T ss_dssp CSS-EEETTTTEE----
T ss_pred CcCeEEeCCCCeEEeeC
Confidence 45689999999888765
No 47
>COG3712 FecR Fe2+-dicitrate sensor, membrane component [Inorganic ion transport and metabolism / Signal transduction mechanisms]
Probab=26.70 E-value=58 Score=30.33 Aligned_cols=31 Identities=23% Similarity=0.492 Sum_probs=27.8
Q ss_pred HHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214 132 FIKEEIQRIKANNPDISHREAFSTAAKNWAHFP 164 (197)
Q Consensus 132 FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~P 164 (197)
-..+|.++|++..|+ |.+||..+..-|....
T Consensus 32 ~~r~af~~W~~~~p~--H~~A~~~~e~lw~~l~ 62 (322)
T COG3712 32 ADRAAFERWRAASPE--HARAWERAERLWQALG 62 (322)
T ss_pred HHHHHHHHHHhcCHH--HHHHHHHHHHHHhhhc
Confidence 357899999999996 9999999999999865
No 48
>PF00527 E7: E7 protein, Early protein; InterPro: IPR000148 This family includes the E7 oncoprotein from various papillomaviruses []. Along with E5 and E6 their activities seem to be especially important for viral oncogenesis. E5 is located at the cell surface and reduces cell gap-gap junction communication. In cervical cancer E5 is expressed in earlier stages of neoplastic transformation of the cervical epithelium during viral infection. The role of E7 is less well understood but it has been shown to impede growth arrest signals in both NIH 3T3 cells and HFKs and that this correlates with elevated cdc25A gene expression. This deregulation of cdc25A is linked to disruption of cell cycle arrest [].; GO: 0003677 DNA binding, 0003700 sequence-specific DNA binding transcription factor activity, 0006355 regulation of transcription, DNA-dependent, 0005622 intracellular; PDB: 2F8B_A 2EWL_A 2B9D_A.
Probab=26.55 E-value=26 Score=26.67 Aligned_cols=17 Identities=29% Similarity=0.419 Sum_probs=7.1
Q ss_pred ecCCCcceeeeccccCC
Q 029214 16 IPCNFCNIVLAVSVPCS 32 (197)
Q Consensus 16 V~CnfC~TILaVsVPcs 32 (197)
+.|.+|+..|-+.|=++
T Consensus 53 t~C~~C~~~lrl~V~as 69 (92)
T PF00527_consen 53 TCCGRCGKRLRLVVVAS 69 (92)
T ss_dssp EEBTTT--EEEEEEEC-
T ss_pred eECCCCCCEEEEEEEeC
Confidence 34555555555544444
No 49
>PF02723 NS3_envE: Non-structural protein NS3/Small envelope protein E; InterPro: IPR003873 This is a family of small nonstructural proteins, well conserved among Coronavirus strains. This protein is also found in Murine hepatitis virus as small envelope protein E.; GO: 0016020 membrane
Probab=25.84 E-value=15 Score=28.20 Aligned_cols=16 Identities=38% Similarity=0.910 Sum_probs=13.2
Q ss_pred CceeeecCCCcceeee
Q 029214 11 EQLCYIPCNFCNIVLA 26 (197)
Q Consensus 11 E~lCYV~CnfC~TILa 26 (197)
=|||..=|+||||++.
T Consensus 37 IqLC~~cc~~~n~~v~ 52 (82)
T PF02723_consen 37 IQLCFQCCRLCNTTVY 52 (82)
T ss_pred HHHHHHHhhhhcceEe
Confidence 3789999999998875
No 50
>PF05164 ZapA: Cell division protein ZapA; InterPro: IPR007838 This entry a structural domain found in the cell division protein ZapA, as well as in related proteins. This domain has a core structure consisting of two layers alpha/beta, and has a long C-terminal helix that forms dimeric parallel and tetrameric antiparallel coiled coils []. ZapA interacts with FtsZ, where FtsZ is part of a mid-cell cytokinetic structure termed the Z-ring that recruits a hierarchy of fission related proteins early in the bacterial cell cycle. ZapA drives the polymerisation and filament bundling of FtsZ, thereby contributing to the spatio-temporal tuning of the Z-ring.; PDB: 1T3U_B 1W2E_B 3HNW_A.
Probab=24.78 E-value=1.1e+02 Score=21.64 Aligned_cols=32 Identities=34% Similarity=0.420 Sum_probs=28.4
Q ss_pred HHHHHHHHHHHhhCCCCCHHHHHHHHHHhhcc
Q 029214 131 QFIKEEIQRIKANNPDISHREAFSTAAKNWAH 162 (197)
Q Consensus 131 ~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~ 162 (197)
.++.+.|..++...|.++...++..||-|-++
T Consensus 28 ~~i~~~i~~~~~~~~~~~~~~~~vlaaLnla~ 59 (89)
T PF05164_consen 28 ELINEKINEIKKKYPKLSPERLAVLAALNLAD 59 (89)
T ss_dssp HHHHHHHHHHCTTCCTSSHHHHHHHHHHHHHH
T ss_pred HHHHHHHHHHHHHcCCCCHHHHHHHHHHHHHH
Confidence 57889999999999999999999999987653
No 51
>PF04769 MAT_Alpha1: Mating-type protein MAT alpha 1; InterPro: IPR006856 This family includes Saccharomyces cerevisiae (Baker's yeast) mating type protein alpha 1 (P01365 from SWISSPROT). MAT alpha 1 is a transcription activator that activates mating-type alpha-specific genes with the help of the MADS-box containing MCM1 transcription factor, which together bind cooperatively to PQ elements upstream of alpha-specific genes. The MCM1-MATalpha1 complex is required for the proper DNA-bending that is needed for transcriptional activation []. Alpha 1 interacts in vivo with STE12, linking expression of alpha-specific genes to the alpha-pheromone (IPR006742 from INTERPRO) response pathway [].; GO: 0000772 mating pheromone activity, 0003677 DNA binding, 0045895 positive regulation of transcription, mating-type specific, 0005634 nucleus
Probab=24.22 E-value=1.9e+02 Score=25.17 Aligned_cols=44 Identities=23% Similarity=0.366 Sum_probs=37.1
Q ss_pred CCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214 118 TPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH 165 (197)
Q Consensus 118 PPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph 165 (197)
.++++.|.-.+|+.|+-= .+...|+.+.|++-...++.|...|+
T Consensus 40 ~~~~~kr~lN~Fm~FRsy----y~~~~~~~~Qk~~S~~l~~lW~~dp~ 83 (201)
T PF04769_consen 40 SPEKAKRPLNGFMAFRSY----YSPIFPPLPQKELSGILTKLWEKDPF 83 (201)
T ss_pred cccccccchhHHHHHHHH----HHhhcCCcCHHHHHHHHHHHHhCCcc
Confidence 678888888899998764 33677889999999999999999886
No 52
>PF13408 Zn_ribbon_recom: Recombinase zinc beta ribbon domain
Probab=24.08 E-value=37 Score=22.10 Aligned_cols=13 Identities=38% Similarity=0.986 Sum_probs=10.8
Q ss_pred eeecCCCCCcccc
Q 029214 39 TVRCGHCSNLWSV 51 (197)
Q Consensus 39 TVRCGHCtnLlSV 51 (197)
.|+||+|..-+..
T Consensus 5 ~l~C~~CG~~m~~ 17 (58)
T PF13408_consen 5 LLRCGHCGSKMTR 17 (58)
T ss_pred cEEcccCCcEeEE
Confidence 4799999988775
No 53
>KOG0527 consensus HMG-box transcription factor [Transcription]
Probab=23.77 E-value=96 Score=28.94 Aligned_cols=62 Identities=18% Similarity=0.372 Sum_probs=50.0
Q ss_pred CCCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCCccccccccCCCCCCCcchhhc
Q 029214 117 TTPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPHIHFGLMLEANNQPKLDDASGN 186 (197)
Q Consensus 117 kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Phihfgl~~d~~~~~~~~~~~~~ 186 (197)
+..++-.|--.||=-|=+.|=++|-++||+|---|.=+...+.|+. .-|..|..=+|+++--
T Consensus 58 ~~~~hIKRPMNAFMVWSq~~RRkma~qnP~mHNSEISK~LG~~WK~--------Lse~EKrPFi~EAeRL 119 (331)
T KOG0527|consen 58 TSTDRIKRPMNAFMVWSQGQRRKLAKQNPKMHNSEISKRLGAEWKL--------LSEEEKRPFVDEAERL 119 (331)
T ss_pred CCccccCCCcchhhhhhHHHHHHHHHhCcchhhHHHHHHHHHHHhh--------cCHhhhccHHHHHHHH
Confidence 4677778899999999999999999999999889999999999984 2345555555555433
No 54
>PRK09774 fec operon regulator FecR; Reviewed
Probab=22.74 E-value=72 Score=28.45 Aligned_cols=29 Identities=10% Similarity=0.109 Sum_probs=25.2
Q ss_pred HHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214 134 KEEIQRIKANNPDISHREAFSTAAKNWAHFP 164 (197)
Q Consensus 134 K~ei~riK~~~P~i~hkEaFs~aAknW~~~P 164 (197)
+++.++|.+++|+ |++||..+..-|....
T Consensus 33 ~~~f~~Wl~a~p~--H~~A~~~~~~lw~~~~ 61 (319)
T PRK09774 33 EARWQQWYEQDQD--NQWAWQQVENLRNQMG 61 (319)
T ss_pred HHHHHHHHhCCHH--HHHHHHHHHHHHHHhh
Confidence 4678999999997 9999999999997754
No 55
>PF04420 CHD5: CHD5-like protein; InterPro: IPR007514 Members of this family are probably coiled-coil proteins that are similar to the CHD5 (Congenital heart disease 5) protein. The exact molecular function of these eukaryotic proteins is unknown.; PDB: 3SJA_H 3SJC_D 3SJB_D 3ZS8_D 3VLC_E.
Probab=22.45 E-value=60 Score=26.65 Aligned_cols=37 Identities=24% Similarity=0.242 Sum_probs=30.7
Q ss_pred CCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhh
Q 029214 124 RVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNW 160 (197)
Q Consensus 124 R~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW 160 (197)
..++.-.+=++.||+.+|++.-.++..|-|...||+=
T Consensus 36 ~~~~~~~~~l~~Ei~~l~~E~~~iS~qDeFAkwaKl~ 72 (161)
T PF04420_consen 36 SKSSKEQRQLRKEILQLKRELNAISAQDEFAKWAKLN 72 (161)
T ss_dssp -HHHHHHHHHHHHHHHHHHHHTTS-TTTSHHHHHHHH
T ss_pred ccccHHHHHHHHHHHHHHHHHHcCCcHHHHHHHHHHH
Confidence 3566677789999999999999999999999998863
No 56
>smart00614 ZnF_BED BED zinc finger. DNA-binding domain in chromatin-boundary-element-binding proteins and transposases
Probab=22.44 E-value=49 Score=21.92 Aligned_cols=14 Identities=21% Similarity=0.548 Sum_probs=11.6
Q ss_pred eeecCCCcceeeec
Q 029214 14 CYIPCNFCNIVLAV 27 (197)
Q Consensus 14 CYV~CnfC~TILaV 27 (197)
-++.|++|..+|..
T Consensus 17 ~~a~C~~C~~~l~~ 30 (50)
T smart00614 17 QRAKCKYCGKKLSR 30 (50)
T ss_pred eEEEecCCCCEeee
Confidence 58999999998863
No 57
>PF09788 Tmemb_55A: Transmembrane protein 55A; InterPro: IPR019178 Members of this family catalyse the hydrolysis of the 4-position phosphate of phosphatidylinositol 4,5-bisphosphate, in the reaction: 1-phosphatidyl-myo-inositol 4,5-bisphosphate + H(2)O = 1-phosphatidyl-1D-myo-inositol 5-phosphate + phosphate.
Probab=21.75 E-value=81 Score=28.80 Aligned_cols=17 Identities=47% Similarity=0.878 Sum_probs=13.3
Q ss_pred ceeeeecCCCCCccccc
Q 029214 36 DIVTVRCGHCSNLWSVN 52 (197)
Q Consensus 36 ~tVTVRCGHCtnLlSVN 52 (197)
...-|.||||.+-..-|
T Consensus 154 ~~~rv~CghC~~~Fl~~ 170 (256)
T PF09788_consen 154 GSCRVICGHCSNTFLFN 170 (256)
T ss_pred CceeEECCCCCCcEecc
Confidence 45779999999976654
No 58
>PF09102 Exotox-A_target: Exotoxin A, targeting; InterPro: IPR015186 This domain, found in Pseudomonas aeruginosa exotoxin A, is responsible for transmembrane targeting of the toxin, as well as transmembrane translocation of the catalytic domain into the cytoplasmic compartment. A furin cleavage site is present within the domain: cleavage generates a 37 kDa carboxy-terminal fragment, which includes the enzymatic domain, which is then is translocated into the cytoplasm. It adopts a helical structure, with six alpha-helices forming a bundle []. ; PDB: 1IKP_A 1IKQ_A 2Q5T_A 3Q9O_A.
Probab=21.48 E-value=48 Score=27.63 Aligned_cols=40 Identities=25% Similarity=0.475 Sum_probs=33.4
Q ss_pred HHHHHHHhhCCCCCHHHHHHHHHHhhccCCCccccccccCC
Q 029214 135 EEIQRIKANNPDISHREAFSTAAKNWAHFPHIHFGLMLEAN 175 (197)
Q Consensus 135 ~ei~riK~~~P~i~hkEaFs~aAknW~~~Phihfgl~~d~~ 175 (197)
..+.||.+++|++ .+.+.+.|+.-..++-.-|-||.+++.
T Consensus 77 ~DL~~~~~~~P~~-~~~~LT~A~~~~~~~V~~~~Gltpe~~ 116 (143)
T PF09102_consen 77 SDLRRINENNPGM-VTQVLTVARQIYNDYVTHHPGLTPEQT 116 (143)
T ss_dssp HHHHHHHHHSCCH-HHHHHHHHHHHHHHHHCCSTT--HHHH
T ss_pred hHHHHHHhcCchH-HHHHHHHHHHHHHHHHHcCCCCCcccc
Confidence 5789999999996 789999999999999888999988743
No 59
>TIGR02147 Fsuc_second hypothetical protein, TIGR02147. This family consists of the 40 members of a paralogous protein family in the rumen anaerobe Fibrobacter succinogenes S85. Member proteins are about 270 residues long and appear to lack signal sequences and transmembrane helices. The only perfectly conserved residue is a glycine in an otherwise poorly conserved region, suggesting members are not enzymes. The family is not characterized.
Probab=21.46 E-value=88 Score=28.06 Aligned_cols=26 Identities=19% Similarity=0.438 Sum_probs=23.6
Q ss_pred hhhhHHHHHHHHHHHhhCCCCCHHHH
Q 029214 127 SAYNQFIKEEIQRIKANNPDISHREA 152 (197)
Q Consensus 127 SaYN~FmK~ei~riK~~~P~i~hkEa 152 (197)
..|..||++...+-|+.+|..|.|+-
T Consensus 8 ~dYR~fl~d~ye~rk~~~p~fS~R~f 33 (271)
T TIGR02147 8 TDYRKYLRDYYEERKKTDPAFSWRFF 33 (271)
T ss_pred hhHHHHHHHHHHHHhccCcCcCHHHH
Confidence 46999999999999999999999873
No 60
>COG4357 Zinc finger domain containing protein (CHY type) [Function unknown]
Probab=21.18 E-value=27 Score=28.08 Aligned_cols=39 Identities=21% Similarity=0.477 Sum_probs=27.6
Q ss_pred eecCCCcceeeec--------cccCC----CccceeeeecCCCCCcccccc
Q 029214 15 YIPCNFCNIVLAV--------SVPCS----SLLDIVTVRCGHCSNLWSVNM 53 (197)
Q Consensus 15 YV~CnfC~TILaV--------sVPcs----sL~~tVTVRCGHCtnLlSVNm 53 (197)
=.+|.-|++--|- .-|.. ..++.-.|.||+|-++|+++=
T Consensus 26 alkc~~C~kyYaCy~CHdel~~Hpf~p~~~~~~~~~~iiCGvC~~~LT~~E 76 (105)
T COG4357 26 ALKCKCCQKYYACYHCHDELEDHPFEPWGLQEFNPKAIICGVCRKLLTRAE 76 (105)
T ss_pred eeeechhhhhhhHHHHHhHHhcCCCccCChhhcCCccEEhhhhhhhhhHHH
Confidence 3577777776652 22332 577788899999999999753
No 61
>PRK12336 translation initiation factor IF-2 subunit beta; Provisional
Probab=20.83 E-value=1e+02 Score=26.33 Aligned_cols=36 Identities=25% Similarity=0.602 Sum_probs=27.8
Q ss_pred eecCCCc---ceeeeccccCCCccceeeeecCCCCCccccccchh
Q 029214 15 YIPCNFC---NIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMAAA 56 (197)
Q Consensus 15 YV~CnfC---~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr~l 56 (197)
||.|.-| +|.|.+. . .+...+|.-|..-.+|.-...
T Consensus 98 yV~C~~C~~pdT~l~k~---~---~~~~l~C~aCGa~~~v~~~~~ 136 (201)
T PRK12336 98 YVICSECGLPDTRLVKE---D---RVLMLRCDACGAHRPVKKRKA 136 (201)
T ss_pred eEECCCCCCCCcEEEEc---C---CeEEEEcccCCCCcccccccc
Confidence 9999999 4777664 2 466789999999998865543
No 62
>KOG4520 consensus Predicted coiled-coil protein [General function prediction only]
Probab=20.77 E-value=1.2e+02 Score=27.36 Aligned_cols=15 Identities=33% Similarity=0.428 Sum_probs=12.2
Q ss_pred hHHHHHHHHHHHhhC
Q 029214 130 NQFIKEEIQRIKANN 144 (197)
Q Consensus 130 N~FmK~ei~riK~~~ 144 (197)
---|||||++||+..
T Consensus 62 ~~~~keEi~~vkE~E 76 (238)
T KOG4520|consen 62 KEKYKEEILEVKERE 76 (238)
T ss_pred hHHHHHHHHHHHHHH
Confidence 345899999999874
No 63
>PF04690 YABBY: YABBY protein; InterPro: IPR006780 YABBY proteins are a group of plant-specific transcription factors involved in the specification of abaxial polarity in lateral organs such as leaves and floral organs [, ].
Probab=20.07 E-value=50 Score=28.24 Aligned_cols=19 Identities=21% Similarity=0.518 Sum_probs=15.9
Q ss_pred eecCCCcceeeeccccCCC
Q 029214 15 YIPCNFCNIVLAVSVPCSS 33 (197)
Q Consensus 15 YV~CnfC~TILaVsVPcss 33 (197)
=|+|+-|+.+|-|.+.-..
T Consensus 36 TVRCGHCtNLLSVNm~~~~ 54 (170)
T PF04690_consen 36 TVRCGHCTNLLSVNMRALL 54 (170)
T ss_pred ceeccCccceeeeeccccc
Confidence 4999999999999886544
Done!