Query         029214
Match_columns 197
No_of_seqs    120 out of 187
Neff          3.3 
Searched_HMMs 46136
Date          Fri Mar 29 09:20:02 2013
Command       hhsearch -i /work/01045/syshi/csienesis_hhblits_a3m/029214.a3m -d /work/01045/syshi/HHdatabase/Cdd.hhm -o /work/01045/syshi/hhsearch_cdd/029214hhsearch_cdd -cpu 12 -v 0 

 No Hit                             Prob E-value P-value  Score    SS Cols Query HMM  Template HMM
  1 PF04690 YABBY:  YABBY protein; 100.0 5.2E-78 1.1E-82  499.9  11.9  157    7-170     4-170 (170)
  2 PF09011 HMG_box_2:  HMG-box do  98.3   8E-07 1.7E-11   62.9   4.9   46  119-164     1-47  (73)
  3 cd01390 HMGB-UBF_HMG-box HMGB-  97.9 2.8E-05   6E-10   52.4   5.1   42  123-164     2-43  (66)
  4 cd00084 HMG-box High Mobility   97.8 4.3E-05 9.4E-10   50.8   5.2   43  123-165     2-44  (66)
  5 cd01388 SOX-TCF_HMG-box SOX-TC  97.8 3.6E-05 7.9E-10   54.3   5.1   42  123-164     3-44  (72)
  6 PF00505 HMG_box:  HMG (high mo  97.8 3.9E-05 8.4E-10   52.2   4.5   42  123-164     2-43  (69)
  7 smart00398 HMG high mobility g  97.8   6E-05 1.3E-09   50.7   5.3   43  122-164     2-44  (70)
  8 cd01389 MATA_HMG-box MATA_HMG-  97.7 7.9E-05 1.7E-09   53.0   5.4   43  123-165     3-45  (77)
  9 PTZ00199 high mobility group p  97.7 8.6E-05 1.9E-09   55.7   5.8   49  117-165    18-68  (94)
 10 PF06244 DUF1014:  Protein of u  97.0  0.0011 2.4E-08   53.4   4.3   51  117-169    70-120 (122)
 11 KOG0381 HMG box-containing pro  96.8  0.0031 6.7E-08   45.9   5.3   46  120-165    21-66  (96)
 12 KOG3223 Uncharacterized conser  94.7   0.022 4.8E-07   50.0   2.4   55  115-171   160-214 (221)
 13 PF11331 DUF3133:  Protein of u  93.0   0.057 1.2E-06   37.2   1.5   38   15-54      6-46  (46)
 14 TIGR02098 MJ0042_CXXC MJ0042 f  80.3     1.5 3.4E-05   27.4   2.1   34   16-51      3-37  (38)
 15 PF13719 zinc_ribbon_5:  zinc-r  79.3     1.4   3E-05   28.3   1.7   34   16-51      3-37  (37)
 16 PRK14892 putative transcriptio  73.9     2.6 5.7E-05   32.9   2.2   44   14-60     20-63  (99)
 17 PF08073 CHDNT:  CHDNT (NUC034)  73.3     6.3 0.00014   28.2   3.8   33  128-163    18-50  (55)
 18 TIGR01562 FdhE formate dehydro  72.4       2 4.4E-05   39.2   1.5   27   11-47    206-232 (305)
 19 PF04032 Rpr2:  RNAse P Rpr2/Rp  71.6     1.9 4.2E-05   30.7   0.9   32   16-47     47-85  (85)
 20 PF05129 Elf1:  Transcription e  71.3     2.3 4.9E-05   31.8   1.3   42   15-56     22-63  (81)
 21 PRK03954 ribonuclease P protei  70.8     2.4 5.2E-05   34.2   1.4   32   18-49     67-103 (121)
 22 PF13717 zinc_ribbon_4:  zinc-r  69.4     5.5 0.00012   25.5   2.6   30   16-49      3-35  (36)
 23 PF10122 Mu-like_Com:  Mu-like   67.5     3.2   7E-05   29.4   1.3   32   16-51      5-36  (51)
 24 PRK03564 formate dehydrogenase  64.6     3.8 8.2E-05   37.7   1.6   27   11-47    208-234 (309)
 25 PF09788 Tmemb_55A:  Transmembr  58.8     7.3 0.00016   35.4   2.3   39   11-53    153-191 (256)
 26 KOG4684 Uncharacterized conser  57.2     4.2 9.2E-05   36.8   0.5   16   37-52    168-183 (275)
 27 COG3058 FdhE Uncharacterized p  54.6       2 4.4E-05   39.7  -1.9   34   10-53    206-239 (308)
 28 COG5648 NHP6B Chromatin-associ  52.5      28  0.0006   30.9   4.8   45  119-163    68-112 (211)
 29 PF04216 FdhE:  Protein involve  52.0     4.7  0.0001   35.4  -0.0   32   13-54    195-226 (290)
 30 PF11331 DUF3133:  Protein of u  49.2      10 0.00022   26.2   1.3   18   13-30     29-46  (46)
 31 PF06382 DUF1074:  Protein of u  48.8      26 0.00056   30.6   3.9   35  127-165    84-118 (183)
 32 KOG4684 Uncharacterized conser  44.3      11 0.00023   34.3   1.0   33   13-53    168-203 (275)
 33 TIGR01053 LSD1 zinc finger dom  40.9      26 0.00056   22.2   2.1   25   16-46      2-26  (31)
 34 KOG0526 Nucleosome-binding fac  38.7      42  0.0009   33.9   4.1   44  118-163   532-575 (615)
 35 PF03811 Zn_Tnp_IS1:  InsA N-te  37.5      23  0.0005   23.0   1.5   17   35-51      1-17  (36)
 36 COG4416 Com Mu-like prophage p  37.4     7.9 0.00017   28.2  -0.7   14   37-50      2-15  (60)
 37 PF10963 DUF2765:  Protein of u  33.1      65  0.0014   24.6   3.5   14  134-147    48-61  (83)
 38 PF05047 L51_S25_CI-B8:  Mitoch  31.6      40 0.00086   22.2   2.0   18  130-147     2-19  (52)
 39 PF01020 Ribosomal_L40e:  Ribos  31.3      18 0.00039   25.8   0.3    9   41-49     38-46  (52)
 40 COG4888 Uncharacterized Zn rib  30.4      37 0.00079   27.3   1.9   45   15-59     22-66  (104)
 41 PF12876 Cellulase-like:  Sugar  30.4      73  0.0016   23.1   3.3   31  118-148    29-59  (88)
 42 KOG4715 SWI/SNF-related matrix  29.7      68  0.0015   30.8   3.8   46  117-165    63-108 (410)
 43 PF14599 zinc_ribbon_6:  Zinc-r  28.7      35 0.00076   24.6   1.4   30   12-47     27-56  (61)
 44 PF10159 MMtag:  Kinase phospho  28.5      52  0.0011   25.1   2.3   13  132-144    57-69  (78)
 45 PF05180 zf-DNL:  DNL zinc fing  28.0      33 0.00071   25.3   1.2   32   17-48      6-38  (66)
 46 PF02892 zf-BED:  BED zinc fing  27.9      27 0.00059   22.1   0.6   17   12-28     13-29  (45)
 47 COG3712 FecR Fe2+-dicitrate se  26.7      58  0.0013   30.3   2.8   31  132-164    32-62  (322)
 48 PF00527 E7:  E7 protein, Early  26.5      26 0.00057   26.7   0.5   17   16-32     53-69  (92)
 49 PF02723 NS3_envE:  Non-structu  25.8      15 0.00032   28.2  -1.0   16   11-26     37-52  (82)
 50 PF05164 ZapA:  Cell division p  24.8 1.1E+02  0.0023   21.6   3.3   32  131-162    28-59  (89)
 51 PF04769 MAT_Alpha1:  Mating-ty  24.2 1.9E+02   0.004   25.2   5.3   44  118-165    40-83  (201)
 52 PF13408 Zn_ribbon_recom:  Reco  24.1      37 0.00079   22.1   0.7   13   39-51      5-17  (58)
 53 KOG0527 HMG-box transcription   23.8      96  0.0021   28.9   3.6   62  117-186    58-119 (331)
 54 PRK09774 fec operon regulator   22.7      72  0.0016   28.5   2.6   29  134-164    33-61  (319)
 55 PF04420 CHD5:  CHD5-like prote  22.4      60  0.0013   26.6   1.8   37  124-160    36-72  (161)
 56 smart00614 ZnF_BED BED zinc fi  22.4      49  0.0011   21.9   1.1   14   14-27     17-30  (50)
 57 PF09788 Tmemb_55A:  Transmembr  21.8      81  0.0018   28.8   2.7   17   36-52    154-170 (256)
 58 PF09102 Exotox-A_target:  Exot  21.5      48   0.001   27.6   1.1   40  135-175    77-116 (143)
 59 TIGR02147 Fsuc_second hypothet  21.5      88  0.0019   28.1   2.8   26  127-152     8-33  (271)
 60 COG4357 Zinc finger domain con  21.2      27 0.00058   28.1  -0.4   39   15-53     26-76  (105)
 61 PRK12336 translation initiatio  20.8   1E+02  0.0022   26.3   3.0   36   15-56     98-136 (201)
 62 KOG4520 Predicted coiled-coil   20.8 1.2E+02  0.0025   27.4   3.4   15  130-144    62-76  (238)
 63 PF04690 YABBY:  YABBY protein;  20.1      50  0.0011   28.2   1.0   19   15-33     36-54  (170)

No 1  
>PF04690 YABBY:  YABBY protein;  InterPro: IPR006780 YABBY proteins are a group of plant-specific transcription factors involved in the specification of abaxial polarity in lateral organs such as leaves and floral organs [, ].
Probab=100.00  E-value=5.2e-78  Score=499.90  Aligned_cols=157  Identities=61%  Similarity=0.948  Sum_probs=116.9

Q ss_pred             CCCCCceeeecCCCcceeeeccccCCCccceeeeecCCCCCccccccchhccCCCcchhcc--ccCC---CC--CCCccc
Q 029214            7 DVAPEQLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMAAAFQSLSWQDVHH--HQAP---SY--ASPECR   79 (197)
Q Consensus         7 ~~~~E~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr~llqsl~~~~~~~--~~~~---~~--~~~~~~   79 (197)
                      +.++|||||||||||||||||||||+|||+|||||||||+|||||||++++++++.+++..  +..+   ..  ..+...
T Consensus         4 ~~~sE~lCYVhCnFC~TiLaVsVP~ssL~~~VTVRCGHCtNLLSVNm~~~~~~~~~~~~l~~~~~~~~~~~~~~~~~~~~   83 (170)
T PF04690_consen    4 FSPSEQLCYVHCNFCNTILAVSVPCSSLLKTVTVRCGHCTNLLSVNMRALLQPLPSQDHLQHSLLPPQSQELQFQPENFG   83 (170)
T ss_pred             cCCCCcEEEEEcCCcCeEEEEecchhhhhhhhceeccCccceeeeeccccccCCCcccchhccccccccccccccccccc
Confidence            4468999999999999999999999999999999999999999999999998887665411  0000   00  001111


Q ss_pred             ccCCCC-Cccccccc-ccCCCCCccccccc-cccCCCCCCCCCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHH
Q 029214           80 IDLGSS-SKCNNKIS-AMRTPTNKATEERV-VNRRESPHSTTPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTA  156 (197)
Q Consensus        80 ~~~~ss-s~~~~~~~-~~~~~~~~~~~~~~-~~k~~~~~~kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~a  156 (197)
                      ....++ +.++.... .+....+++.|+.+ ++|       |||||||+|||||+||||||||||++||||+|||||++|
T Consensus        84 ~~~~~~~~~~~~~~~~~~~~~~~~~~pr~~~v~k-------PPEKRqR~psaYn~f~k~ei~rik~~~p~ishkeaFs~a  156 (170)
T PF04690_consen   84 SNSSSSSSSSSSSSSSSMSFSEEEEIPRAPPVNK-------PPEKRQRVPSAYNRFMKEEIQRIKAENPDISHKEAFSAA  156 (170)
T ss_pred             cccccCCCccccccccccCccccccccccccccC-------CccccCCCchhHHHHHHHHHHHHHhcCCCCCHHHHHHHH
Confidence            111111 11111100 11112233455443 345       999999999999999999999999999999999999999


Q ss_pred             HHhhccCCCccccc
Q 029214          157 AKNWAHFPHIHFGL  170 (197)
Q Consensus       157 AknW~~~Phihfgl  170 (197)
                      ||||||+|||||||
T Consensus       157 AknW~h~phihfgl  170 (170)
T PF04690_consen  157 AKNWAHFPHIHFGL  170 (170)
T ss_pred             HHhhhhCcccccCC
Confidence            99999999999997


No 2  
>PF09011 HMG_box_2:  HMG-box domain;  InterPro: IPR015101 This domain is predominantly found in Maelstrom homologue proteins. It has no known function. ; GO: 0005634 nucleus; PDB: 2EQZ_A 1V64_A 2CTO_A 1H5P_A 3TQ6_A 3FGH_A 3TMM_A 1J3X_A 2YRQ_A 1AAB_A ....
Probab=98.34  E-value=8e-07  Score=62.88  Aligned_cols=46  Identities=35%  Similarity=0.634  Sum_probs=38.6

Q ss_pred             CcccCCCchhhhHHHHHHHHHHHhh-CCCCCHHHHHHHHHHhhccCC
Q 029214          119 PEKRQRVPSAYNQFIKEEIQRIKAN-NPDISHREAFSTAAKNWAHFP  164 (197)
Q Consensus       119 PEKRQR~PSaYN~FmK~ei~riK~~-~P~i~hkEaFs~aAknW~~~P  164 (197)
                      |.|..|.+|||+.||++.+.++++. .+.++++|++..++..|+..+
T Consensus         1 p~kpK~~~say~lF~~~~~~~~k~~G~~~~~~~e~~k~~~~~Wk~Ls   47 (73)
T PF09011_consen    1 PKKPKRPPSAYNLFMKEMRKEVKEEGGQKQSFREVMKEISERWKSLS   47 (73)
T ss_dssp             SSS--SSSSHHHHHHHHHHHHHHHHT-T-SSHHHHHHHHHHHHHHS-
T ss_pred             CcCCCCCCCHHHHHHHHHHHHHHHhcccCCCHHHHHHHHHHHHHhcC
Confidence            6677899999999999999999999 888999999999999999754


No 3  
>cd01390 HMGB-UBF_HMG-box HMGB-UBF_HMG-box, class II and III members of the HMG-box superfamily of DNA-binding proteins. These proteins bind the minor groove of DNA in a non-sequence specific fashion and contain two or more tandem HMG boxes. Class II members include non-histone chromosomal proteins, HMG1 and HMG2, which bind to bent or distorted DNA such as four-way DNA junctions, synthetic DNA cruciforms, kinked cisplatin-modified DNA, DNA bulges, cross-overs in supercoiled DNA, and can cause looping of linear DNA. Class III members include nucleolar and mitochondrial transcription factors, UBF and mtTF1, which bind four-way DNA junctions.
Probab=97.90  E-value=2.8e-05  Score=52.39  Aligned_cols=42  Identities=31%  Similarity=0.498  Sum_probs=39.4

Q ss_pred             CCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214          123 QRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFP  164 (197)
Q Consensus       123 QR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~P  164 (197)
                      .|-+|||..|+++....+++.+|+++..|....++..|+..+
T Consensus         2 krp~saf~~f~~~~r~~~~~~~p~~~~~~i~~~~~~~W~~ls   43 (66)
T cd01390           2 KRPLSAYFLFSQEQRPKLKKENPDASVTEVTKILGEKWKELS   43 (66)
T ss_pred             CCCCcHHHHHHHHHHHHHHHHCcCCCHHHHHHHHHHHHHhCC
Confidence            467899999999999999999999999999999999999754


No 4  
>cd00084 HMG-box High Mobility Group (HMG)-box is found in a variety of eukaryotic chromosomal proteins and transcription factors. HMGs bind to the minor groove of DNA and have been classified by DNA binding preferences. Two phylogenically distinct groups of Class I proteins bind DNA in a sequence specific fashion and contain a single HMG box. One group (SOX-TCF) includes transcription factors, TCF-1, -3, -4; and also SRY and LEF-1, which bind four-way DNA junctions and duplex DNA targets. The second group (MATA) includes fungal mating type gene products MC, MATA1 and Ste11. Class II and III proteins (HMGB-UBF) bind DNA in a non-sequence specific fashion and contain two or more tandem HMG boxes. Class II members include non-histone chromosomal proteins, HMG1 and HMG2, which bind to bent or distorted DNA such as four-way DNA junctions, synthetic DNA cruciforms, kinked cisplatin-modified DNA, DNA bulges, cross-overs in supercoiled DNA, and can cause looping of linear DNA. Class III member
Probab=97.84  E-value=4.3e-05  Score=50.84  Aligned_cols=43  Identities=30%  Similarity=0.432  Sum_probs=39.8

Q ss_pred             CCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214          123 QRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH  165 (197)
Q Consensus       123 QR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph  165 (197)
                      .|-+|+|..|++++...+++.+|+++..|....+++.|+..+.
T Consensus         2 krp~~af~~f~~~~~~~~~~~~~~~~~~~i~~~~~~~W~~l~~   44 (66)
T cd00084           2 KRPLSAYFLFSQEHRAEVKAENPGLSVGEISKILGEMWKSLSE   44 (66)
T ss_pred             CCCCcHHHHHHHHHHHHHHHHCcCCCHHHHHHHHHHHHHhCCH
Confidence            4678999999999999999999999999999999999997653


No 5  
>cd01388 SOX-TCF_HMG-box SOX-TCF_HMG-box, class I member of the HMG-box superfamily of DNA-binding proteins. These proteins contain a single HMG box, and bind the minor groove of DNA in a highly sequence-specific manner. Members include SRY and its homologs in insects and vertebrates, and transcription factor-like proteins, TCF-1, -3, -4, and LEF-1. They appear to bind the minor groove of the A/T C A A A G/C-motif.
Probab=97.83  E-value=3.6e-05  Score=54.27  Aligned_cols=42  Identities=17%  Similarity=0.294  Sum_probs=39.3

Q ss_pred             CCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214          123 QRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFP  164 (197)
Q Consensus       123 QR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~P  164 (197)
                      .|-|+||..|+++.-.+|+++||+++..|.-+.++..|+..+
T Consensus         3 KrP~naf~~F~~~~r~~~~~~~p~~~~~eisk~l~~~Wk~ls   44 (72)
T cd01388           3 KRPMNAFMLFSKRHRRKVLQEYPLKENRAISKILGDRWKALS   44 (72)
T ss_pred             CCCCcHHHHHHHHHHHHHHHHCCCCCHHHHHHHHHHHHHcCC
Confidence            478999999999999999999999999999999999999754


No 6  
>PF00505 HMG_box:  HMG (high mobility group) box;  InterPro: IPR000910 High mobility group (HMG or HMGB) proteins are a family of relatively low molecular weight non-histone components in chromatin. HMG1 (also called HMG-T in fish) and HMG2 are two highly related proteins that bind single-stranded DNA preferentially and unwind double-stranded DNA. Although they have no sequence specificity, they have a high affinity for bent or distorted DNA, and bend linear DNA. HMG1 and HMG2 contain two DNA-binding HMG-box domains (A and B) that show structural and functional differences, and have a long acidic C-terminal domain rich in aspartic and glutamic acid residues. The acidic tail modulates the affinity of the tandem HMG boxes in HMG1 and 2 for a variety of DNA targets. HMG1 and 2 appear to play important architectural roles in the assembly of nucleoprotein complexes in a variety of biological processes, for example V(D)J recombination, the initiation of transcription, and DNA repair []. The profile in this entry describing the HMG-domains is much more general than the signature. In addition to the HMG1 and HMG2 proteins, HMG-domains occur in single or multiple copies in the following protein classes; the SOX family of transcription factors; SRY sex determining region Y protein and related proteins []; LEF1 lymphoid enhancer binding factor 1 []; SSRP recombination signal recognition protein; MTF1 mitochondrial transcription factor 1; UBF1/2 nucleolar transcription factors; Abf2 yeast ARS-binding factor []; and Saccharomyces cerevisiae transcription factors Ixr1, Rox1, Nhp6a, Nhp6b and Spp41.; GO: 0003677 DNA binding; PDB: 1I11_A 1J3C_A 1J3D_A 1WZ6_A 1WGF_A 2D7L_A 1GT0_D 3U2B_C 2CRJ_A 2CS1_A ....
Probab=97.79  E-value=3.9e-05  Score=52.23  Aligned_cols=42  Identities=33%  Similarity=0.635  Sum_probs=37.5

Q ss_pred             CCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214          123 QRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFP  164 (197)
Q Consensus       123 QR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~P  164 (197)
                      .|-++||..|+++....|++.+|+++..|.-..+++.|+..+
T Consensus         2 krP~~af~lf~~~~~~~~k~~~p~~~~~~i~~~~~~~W~~l~   43 (69)
T PF00505_consen    2 KRPPNAFMLFCKEKRAKLKEENPDLSNKEISKILAQMWKNLS   43 (69)
T ss_dssp             SSS--HHHHHHHHHHHHHHHHSTTSTHHHHHHHHHHHHHCSH
T ss_pred             cCCCCHHHHHHHHHHHHHHHHhcccccccchhhHHHHHhcCC
Confidence            578999999999999999999999999999999999999753


No 7  
>smart00398 HMG high mobility group.
Probab=97.78  E-value=6e-05  Score=50.67  Aligned_cols=43  Identities=33%  Similarity=0.511  Sum_probs=39.9

Q ss_pred             cCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214          122 RQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFP  164 (197)
Q Consensus       122 RQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~P  164 (197)
                      ..|-+|+|..|+++....+++.+|+++..|....++..|+..+
T Consensus         2 pkrp~~~y~~f~~~~r~~~~~~~~~~~~~~i~~~~~~~W~~l~   44 (70)
T smart00398        2 PKRPMSAFMLFSQENRAKIKAENPDLSNAEISKKLGERWKLLS   44 (70)
T ss_pred             cCCCCcHHHHHHHHHHHHHHHHCcCCCHHHHHHHHHHHHHcCC
Confidence            4578999999999999999999999999999999999999754


No 8  
>cd01389 MATA_HMG-box MATA_HMG-box, class I member of the HMG-box superfamily of DNA-binding proteins. These proteins contain a single HMG box, and bind the minor groove of DNA in a highly sequence-specific manner. Members include the fungal mating type gene products MC, MATA1 and Ste11.
Probab=97.72  E-value=7.9e-05  Score=53.03  Aligned_cols=43  Identities=16%  Similarity=0.358  Sum_probs=40.4

Q ss_pred             CCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214          123 QRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH  165 (197)
Q Consensus       123 QR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph  165 (197)
                      .|-|+||-.|+++....|+++||++++.|.-..++..|+..+.
T Consensus         3 kRP~naf~lf~~~~r~~~~~~~p~~~~~eisk~~g~~Wk~ls~   45 (77)
T cd01389           3 PRPRNAFILYRQDKHAQLKTENPGLTNNEISRIIGRMWRSESP   45 (77)
T ss_pred             CCCCcHHHHHHHHHHHHHHHHCCCCCHHHHHHHHHHHHhhCCH
Confidence            5889999999999999999999999999999999999998654


No 9  
>PTZ00199 high mobility group protein; Provisional
Probab=97.72  E-value=8.6e-05  Score=55.70  Aligned_cols=49  Identities=27%  Similarity=0.441  Sum_probs=43.2

Q ss_pred             CCCcccCCCchhhhHHHHHHHHHHHhhCCCCC--HHHHHHHHHHhhccCCC
Q 029214          117 TTPEKRQRVPSAYNQFIKEEIQRIKANNPDIS--HREAFSTAAKNWAHFPH  165 (197)
Q Consensus       117 kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~--hkEaFs~aAknW~~~Ph  165 (197)
                      +.|.+..|-+|||..|+++.-..|+++||+++  ..|....++..|+..+.
T Consensus        18 kdp~~PKrP~sAY~~F~~~~R~~i~~~~P~~~~~~~evsk~ige~Wk~ls~   68 (94)
T PTZ00199         18 KDPNAPKRALSAYMFFAKEKRAEIIAENPELAKDVAAVGKMVGEAWNKLSE   68 (94)
T ss_pred             CCCCCCCCCCcHHHHHHHHHHHHHHHHCcCCcccHHHHHHHHHHHHHcCCH
Confidence            36777889999999999999999999999986  67888999999998653


No 10 
>PF06244 DUF1014:  Protein of unknown function (DUF1014);  InterPro: IPR010422 This family consists of several hypothetical eukaryotic proteins of unknown function.
Probab=96.96  E-value=0.0011  Score=53.41  Aligned_cols=51  Identities=25%  Similarity=0.513  Sum_probs=46.5

Q ss_pred             CCCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCCcccc
Q 029214          117 TTPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPHIHFG  169 (197)
Q Consensus       117 kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Phihfg  169 (197)
                      +.||||-+  -||..|--.+|.+||++||++.+-+---..=|.|..+|..+|-
T Consensus        70 rHPErR~K--AAy~afeE~~Lp~lK~E~PgLrlsQ~kq~l~K~w~KSPeNP~N  120 (122)
T PF06244_consen   70 RHPERRMK--AAYKAFEERRLPELKEENPGLRLSQYKQMLWKEWQKSPENPFN  120 (122)
T ss_pred             CCcchhHH--HHHHHHHHHHhHHHHhhCCCchHHHHHHHHHHHHhcCCCCCcc
Confidence            38999875  5999999999999999999999999888999999999998874


No 11 
>KOG0381 consensus HMG box-containing protein [General function prediction only]
Probab=96.80  E-value=0.0031  Score=45.87  Aligned_cols=46  Identities=28%  Similarity=0.441  Sum_probs=41.4

Q ss_pred             cccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214          120 EKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH  165 (197)
Q Consensus       120 EKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph  165 (197)
                      ...+|-+|||..|+.+.-.+||++||+++..|.-+.+..+|.....
T Consensus        21 ~~pkrp~sa~~~f~~~~~~~~k~~~p~~~~~~v~k~~g~~W~~l~~   66 (96)
T KOG0381|consen   21 QAPKRPLSAFFLFSSEQRSKIKAENPGLSVGEVAKALGEMWKNLAE   66 (96)
T ss_pred             CCCCCCCcHHHHHHHHHHHHHHHhCCCCCHHHHHHHHHHHHhcCCH
Confidence            3567889999999999999999999999999999999999987543


No 12 
>KOG3223 consensus Uncharacterized conserved protein [Function unknown]
Probab=94.66  E-value=0.022  Score=49.98  Aligned_cols=55  Identities=27%  Similarity=0.496  Sum_probs=48.8

Q ss_pred             CCCCCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCCcccccc
Q 029214          115 HSTTPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPHIHFGLM  171 (197)
Q Consensus       115 ~~kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Phihfgl~  171 (197)
                      +.+-||||-|+  ||--|=..++-|||.+||+++|-+-=-..-|.|...|..+|.-+
T Consensus       160 ddrHPEkRmrA--A~~afEe~~LPrLK~e~P~lrlsQ~Kqll~Kew~KsPDNP~Nq~  214 (221)
T KOG3223|consen  160 DDRHPEKRMRA--AFKAFEEARLPRLKKENPGLRLSQYKQLLKKEWQKSPDNPFNQA  214 (221)
T ss_pred             cccChHHHHHH--HHHHHHHhhchhhhhcCCCccHHHHHHHHHHHHhhCCCChhhHH
Confidence            33689999885  89999999999999999999999888888899999999998643


No 13 
>PF11331 DUF3133:  Protein of unknown function (DUF3133);  InterPro: IPR021480  This eukaryotic family of proteins has no known function. 
Probab=93.02  E-value=0.057  Score=37.23  Aligned_cols=38  Identities=29%  Similarity=0.697  Sum_probs=29.3

Q ss_pred             eecCCCcceeeeccccCCCccc---eeeeecCCCCCccccccc
Q 029214           15 YIPCNFCNIVLAVSVPCSSLLD---IVTVRCGHCSNLWSVNMA   54 (197)
Q Consensus        15 YV~CnfC~TILaVsVPcssL~~---tVTVRCGHCtnLlSVNmr   54 (197)
                      ||-|..|..+|-  +|-+.+..   .-.+|||.|+.++++.++
T Consensus         6 Fv~C~~C~~lLq--lP~~~~~~~k~~~klrCGaCs~vl~~s~~   46 (46)
T PF11331_consen    6 FVVCSSCFELLQ--LPAKFSLSKKNQQKLRCGACSEVLSFSLP   46 (46)
T ss_pred             EeECccHHHHHc--CCCccCCCccceeEEeCCCCceeEEEecC
Confidence            788999977665  47765443   568999999999988653


No 14 
>TIGR02098 MJ0042_CXXC MJ0042 family finger-like domain. This domain contains a CXXCX(19)CXXC motif suggestive of both zinc fingers and thioredoxin, usually found at the N-terminus of prokaryotic proteins. One partially characterized gene, agmX, is among a large set in Myxococcus whose interruption affects adventurous gliding motility.
Probab=80.27  E-value=1.5  Score=27.40  Aligned_cols=34  Identities=35%  Similarity=0.721  Sum_probs=24.0

Q ss_pred             ecCCCcceeeeccccCCCcc-ceeeeecCCCCCcccc
Q 029214           16 IPCNFCNIVLAVSVPCSSLL-DIVTVRCGHCSNLWSV   51 (197)
Q Consensus        16 V~CnfC~TILaVsVPcssL~-~tVTVRCGHCtnLlSV   51 (197)
                      +.|..|.+..-|..  +.+- +...|+|++|.+.+.+
T Consensus         3 ~~CP~C~~~~~v~~--~~~~~~~~~v~C~~C~~~~~~   37 (38)
T TIGR02098         3 IQCPNCKTSFRVVD--SQLGANGGKVRCGKCGHVWYA   37 (38)
T ss_pred             EECCCCCCEEEeCH--HHcCCCCCEEECCCCCCEEEe
Confidence            67999999876653  2221 3347999999987764


No 15 
>PF13719 zinc_ribbon_5:  zinc-ribbon domain
Probab=79.33  E-value=1.4  Score=28.29  Aligned_cols=34  Identities=32%  Similarity=0.665  Sum_probs=25.4

Q ss_pred             ecCCCcceeeeccccCCCc-cceeeeecCCCCCcccc
Q 029214           16 IPCNFCNIVLAVSVPCSSL-LDIVTVRCGHCSNLWSV   51 (197)
Q Consensus        16 V~CnfC~TILaVsVPcssL-~~tVTVRCGHCtnLlSV   51 (197)
                      ++|--|.|.+.|  |=+.| -....|||++|.+.+.|
T Consensus         3 i~CP~C~~~f~v--~~~~l~~~~~~vrC~~C~~~f~v   37 (37)
T PF13719_consen    3 ITCPNCQTRFRV--PDDKLPAGGRKVRCPKCGHVFRV   37 (37)
T ss_pred             EECCCCCceEEc--CHHHcccCCcEEECCCCCcEeeC
Confidence            689999998865  44443 34679999999987654


No 16 
>PRK14892 putative transcription elongation factor Elf1; Provisional
Probab=73.92  E-value=2.6  Score=32.90  Aligned_cols=44  Identities=18%  Similarity=0.530  Sum_probs=34.8

Q ss_pred             eeecCCCcceeeeccccCCCccceeeeecCCCCCccccccchhccCC
Q 029214           14 CYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMAAAFQSL   60 (197)
Q Consensus        14 CYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr~llqsl   60 (197)
                      =++.|.||+. ..|+||...  .+..+.|..|.---.-.+..|.++.
T Consensus        20 t~f~CP~Cge-~~v~v~~~k--~~~h~~C~~CG~y~~~~V~~l~epI   63 (99)
T PRK14892         20 KIFECPRCGK-VSISVKIKK--NIAIITCGNCGLYTEFEVPSVYDEV   63 (99)
T ss_pred             cEeECCCCCC-eEeeeecCC--CcceEECCCCCCccCEECCccccch
Confidence            3789999995 688888888  7999999999877665566555543


No 17 
>PF08073 CHDNT:  CHDNT (NUC034) domain;  InterPro: IPR012958 The CHD N-terminal domain is found in PHD/RING fingers and chromo domain-associated helicases [].; GO: 0003677 DNA binding, 0005524 ATP binding, 0008270 zinc ion binding, 0016818 hydrolase activity, acting on acid anhydrides, in phosphorus-containing anhydrides, 0006355 regulation of transcription, DNA-dependent, 0005634 nucleus
Probab=73.30  E-value=6.3  Score=28.22  Aligned_cols=33  Identities=18%  Similarity=0.456  Sum_probs=26.0

Q ss_pred             hhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccC
Q 029214          128 AYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHF  163 (197)
Q Consensus       128 aYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~  163 (197)
                      +|..+|+   .-|-++||++.+-.-+...+..|++|
T Consensus        18 ~Fsq~vR---P~l~~~NPk~~~sKl~~l~~AKwrEF   50 (55)
T PF08073_consen   18 AFSQHVR---PLLAKANPKAPMSKLMMLLQAKWREF   50 (55)
T ss_pred             HHHHHHH---HHHHHHCCCCcHHHHHHHHHHHHHHH
Confidence            3444444   34567999999999999999999986


No 18 
>TIGR01562 FdhE formate dehydrogenase accessory protein FdhE. The only sequence scoring between trusted and noise is that from Aquifex aeolicus, which shows certain structural differences from the proteobacterial forms in the alignment. However it is notable that A. aeolicus also has a sequence scoring above trusted to the alpha subunit of formate dehydrogenase (TIGR01553).
Probab=72.45  E-value=2  Score=39.24  Aligned_cols=27  Identities=33%  Similarity=0.897  Sum_probs=23.1

Q ss_pred             CceeeecCCCcceeeeccccCCCccceeeeecCCCCC
Q 029214           11 EQLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSN   47 (197)
Q Consensus        11 E~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtn   47 (197)
                      +..=|.||++|.|          -...|-++|.+|.|
T Consensus       206 ~G~RyL~CslC~t----------eW~~~R~~C~~Cg~  232 (305)
T TIGR01562       206 TGLRYLSCSLCAT----------EWHYVRVKCSHCEE  232 (305)
T ss_pred             CCceEEEcCCCCC----------cccccCccCCCCCC
Confidence            4566999999977          57889999999998


No 19 
>PF04032 Rpr2:  RNAse P Rpr2/Rpp21/SNM1 subunit domain;  InterPro: IPR007175 This family contains a ribonuclease P subunit of human and yeast. Other members of the family include the probable archaeal homologues. This subunit possibly binds the precursor tRNA [].; PDB: 2K3R_A 2KI7_B 2ZAE_B 1X0T_A.
Probab=71.62  E-value=1.9  Score=30.69  Aligned_cols=32  Identities=25%  Similarity=0.580  Sum_probs=20.2

Q ss_pred             ecCCCcceeeeccccCCC-------ccceeeeecCCCCC
Q 029214           16 IPCNFCNIVLAVSVPCSS-------LLDIVTVRCGHCSN   47 (197)
Q Consensus        16 V~CnfC~TILaVsVPcss-------L~~tVTVRCGHCtn   47 (197)
                      .-|.-|.++|.-|+-|+-       .-+.|.++|..|.+
T Consensus        47 ~~Ck~C~~~liPG~~~~vri~~~~~~~~~l~~~C~~C~~   85 (85)
T PF04032_consen   47 TICKKCGSLLIPGVNCSVRIRKKKKKKNFLVYTCLNCGH   85 (85)
T ss_dssp             TB-TTT--B--CTTTEEEEEE---SSS-EEEEEETTTTE
T ss_pred             ccccCCCCEEeCCCccEEEEEecCCCCCEEEEEccccCC
Confidence            458999999999998863       24688999999963


No 20 
>PF05129 Elf1:  Transcription elongation factor Elf1 like;  InterPro: IPR007808 This family of uncharacterised, mostly short, proteins contain a putative zinc binding domain with four conserved cysteines.; PDB: 1WII_A.
Probab=71.33  E-value=2.3  Score=31.81  Aligned_cols=42  Identities=24%  Similarity=0.453  Sum_probs=26.2

Q ss_pred             eecCCCcceeeeccccCCCccceeeeecCCCCCccccccchh
Q 029214           15 YIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMAAA   56 (197)
Q Consensus        15 YV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr~l   56 (197)
                      +-.|-||+.--+|+|=.+.-..+-++.||-|.--...++..|
T Consensus        22 ~F~CPfC~~~~sV~v~idkk~~~~~~~C~~Cg~~~~~~i~~L   63 (81)
T PF05129_consen   22 VFDCPFCNHEKSVSVKIDKKEGIGILSCRVCGESFQTKINPL   63 (81)
T ss_dssp             ----TTT--SS-EEEEEETTTTEEEEEESSS--EEEEE--SS
T ss_pred             eEcCCcCCCCCeEEEEEEccCCEEEEEecCCCCeEEEccCcc
Confidence            457999999999999999999999999999954444444433


No 21 
>PRK03954 ribonuclease P protein component 4; Validated
Probab=70.85  E-value=2.4  Score=34.25  Aligned_cols=32  Identities=25%  Similarity=0.569  Sum_probs=24.6

Q ss_pred             CCCcceeeeccccCC-Cccc----eeeeecCCCCCcc
Q 029214           18 CNFCNIVLAVSVPCS-SLLD----IVTVRCGHCSNLW   49 (197)
Q Consensus        18 CnfC~TILaVsVPcs-sL~~----tVTVRCGHCtnLl   49 (197)
                      |-.|+|.|.-||-|. ++-+    -|.|+|..|...-
T Consensus        67 CK~C~t~LiPG~n~~vRi~~~~~~~vvitCl~CG~~k  103 (121)
T PRK03954         67 CKRCHSFLVPGVNARVRLRQKRMPHVVITCLECGHIM  103 (121)
T ss_pred             hhcCCCeeecCCceEEEEecCCcceEEEECccCCCEE
Confidence            889999999888765 2222    5899999998753


No 22 
>PF13717 zinc_ribbon_4:  zinc-ribbon domain
Probab=69.44  E-value=5.5  Score=25.54  Aligned_cols=30  Identities=27%  Similarity=0.739  Sum_probs=23.3

Q ss_pred             ecCCCcceeeecc---ccCCCccceeeeecCCCCCcc
Q 029214           16 IPCNFCNIVLAVS---VPCSSLLDIVTVRCGHCSNLW   49 (197)
Q Consensus        16 V~CnfC~TILaVs---VPcssL~~tVTVRCGHCtnLl   49 (197)
                      +.|.-|.+...|.   ||    =+.+.|||+.|.+.+
T Consensus         3 i~Cp~C~~~y~i~d~~ip----~~g~~v~C~~C~~~f   35 (36)
T PF13717_consen    3 ITCPNCQAKYEIDDEKIP----PKGRKVRCSKCGHVF   35 (36)
T ss_pred             EECCCCCCEEeCCHHHCC----CCCcEEECCCCCCEe
Confidence            6799999988765   44    245799999999865


No 23 
>PF10122 Mu-like_Com:  Mu-like prophage protein Com;  InterPro: IPR019294  Members of this entry belong to the Com family of proteins that act as translational regulators of mom [, ]. 
Probab=67.51  E-value=3.2  Score=29.42  Aligned_cols=32  Identities=28%  Similarity=0.598  Sum_probs=23.4

Q ss_pred             ecCCCcceeeeccccCCCccceeeeecCCCCCcccc
Q 029214           16 IPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSV   51 (197)
Q Consensus        16 V~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSV   51 (197)
                      +||..|+..||-+-    -+..+.++|..|..+-.|
T Consensus         5 iRC~~CnklLa~~g----~~~~leIKCpRC~tiN~~   36 (51)
T PF10122_consen    5 IRCGHCNKLLAKAG----EVIELEIKCPRCKTINHV   36 (51)
T ss_pred             eeccchhHHHhhhc----CccEEEEECCCCCccceE
Confidence            68888888888642    344788888888876554


No 24 
>PRK03564 formate dehydrogenase accessory protein FdhE; Provisional
Probab=64.62  E-value=3.8  Score=37.66  Aligned_cols=27  Identities=33%  Similarity=0.922  Sum_probs=23.0

Q ss_pred             CceeeecCCCcceeeeccccCCCccceeeeecCCCCC
Q 029214           11 EQLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSN   47 (197)
Q Consensus        11 E~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtn   47 (197)
                      +-.=|.||++|.|          ....|-++|.+|.|
T Consensus       208 ~G~RyL~CslC~t----------eW~~~R~~C~~Cg~  234 (309)
T PRK03564        208 QGLRYLHCNLCES----------EWHVVRVKCSNCEQ  234 (309)
T ss_pred             CCceEEEcCCCCC----------cccccCccCCCCCC
Confidence            5567999999977          57889999999997


No 25 
>PF09788 Tmemb_55A:  Transmembrane protein 55A;  InterPro: IPR019178  Members of this family catalyse the hydrolysis of the 4-position phosphate of phosphatidylinositol 4,5-bisphosphate, in the reaction:  1-phosphatidyl-myo-inositol 4,5-bisphosphate + H(2)O = 1-phosphatidyl-1D-myo-inositol 5-phosphate + phosphate.  
Probab=58.76  E-value=7.3  Score=35.37  Aligned_cols=39  Identities=26%  Similarity=0.534  Sum_probs=25.3

Q ss_pred             CceeeecCCCcceeeeccccCCCccceeeeecCCCCCcccccc
Q 029214           11 EQLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNM   53 (197)
Q Consensus        11 E~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNm   53 (197)
                      ..-|=|.|+.|+...+...+-++-    -.||-||-.++||.-
T Consensus       153 p~~~rv~CghC~~~Fl~~~~~~~t----lARCPHCrKvSSVG~  191 (256)
T PF09788_consen  153 PGSCRVICGHCSNTFLFNTLTSNT----LARCPHCRKVSSVGP  191 (256)
T ss_pred             CCceeEECCCCCCcEeccCCCCCc----cccCCCCceeccccc
Confidence            345777888887776665544221    267888888887753


No 26 
>KOG4684 consensus Uncharacterized conserved protein, contains C4-type Zn-finger [General function prediction only]
Probab=57.18  E-value=4.2  Score=36.82  Aligned_cols=16  Identities=38%  Similarity=0.893  Sum_probs=10.6

Q ss_pred             eeeeecCCCCCccccc
Q 029214           37 IVTVRCGHCSNLWSVN   52 (197)
Q Consensus        37 tVTVRCGHCtnLlSVN   52 (197)
                      .+-|+||||++..--|
T Consensus       168 gcRV~CgHC~~tFLfn  183 (275)
T KOG4684|consen  168 GCRVKCGHCNETFLFN  183 (275)
T ss_pred             ceEEEecCccceeehh
Confidence            3778888887764433


No 27 
>COG3058 FdhE Uncharacterized protein involved in formate dehydrogenase formation [Posttranslational modification, protein turnover, chaperones]
Probab=54.61  E-value=2  Score=39.72  Aligned_cols=34  Identities=24%  Similarity=0.647  Sum_probs=26.8

Q ss_pred             CCceeeecCCCcceeeeccccCCCccceeeeecCCCCCcccccc
Q 029214           10 PEQLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNM   53 (197)
Q Consensus        10 ~E~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNm   53 (197)
                      .+-+=|.|||.|-|          ....|-|+|-+|.+---+++
T Consensus       206 ~~GlRYL~CslC~t----------eW~~VR~KC~nC~~t~~l~y  239 (308)
T COG3058         206 EQGLRYLHCSLCET----------EWHYVRVKCSNCEQSKKLHY  239 (308)
T ss_pred             cccchhhhhhhHHH----------HHHHHHHHhccccccCCccc
Confidence            46788999999987          56789999999998544433


No 28 
>COG5648 NHP6B Chromatin-associated proteins containing the HMG domain [Chromatin structure and dynamics]
Probab=52.52  E-value=28  Score=30.93  Aligned_cols=45  Identities=24%  Similarity=0.448  Sum_probs=39.7

Q ss_pred             CcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccC
Q 029214          119 PEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHF  163 (197)
Q Consensus       119 PEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~  163 (197)
                      |-=..|--|||-.|..+.=.+|...+|+++.-|.=..+++.|+..
T Consensus        68 pN~PKRp~sayf~y~~~~R~ei~~~~p~l~~~e~~k~~~e~WK~L  112 (211)
T COG5648          68 PNGPKRPLSAYFLYSAENRDEIRKENPKLTFGEVGKLLSEKWKEL  112 (211)
T ss_pred             CCCCCCchhHHHHHHHHHHHHHHHhCCCCChHHHHHHHHHHHHhc
Confidence            333457889999999999999999999999999999999999975


No 29 
>PF04216 FdhE:  Protein involved in formate dehydrogenase formation;  InterPro: IPR006452 This family of sequences describe an accessory protein required for the assembly of formate dehydrogenase of certain proteobacteria although not present in the final complex []. The exact nature of the function of FdhE in the assembly of the complex is unknown, but considering the presence of selenocysteine, molybdopterin, iron-sulphur clusters and cytochrome b556, it is likely to be involved in the insertion of cofactors. ; GO: 0005737 cytoplasm; PDB: 2FIY_B.
Probab=51.96  E-value=4.7  Score=35.40  Aligned_cols=32  Identities=22%  Similarity=0.588  Sum_probs=17.3

Q ss_pred             eeeecCCCcceeeeccccCCCccceeeeecCCCCCccccccc
Q 029214           13 LCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMA   54 (197)
Q Consensus        13 lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr   54 (197)
                      .=|-+|.+|.|-          ...+-++|.+|.|--...+.
T Consensus       195 ~R~L~Cs~C~t~----------W~~~R~~Cp~Cg~~~~~~l~  226 (290)
T PF04216_consen  195 KRYLHCSLCGTE----------WRFVRIKCPYCGNTDHEKLE  226 (290)
T ss_dssp             EEEEEETTT--E----------EE--TTS-TTT---SS-EEE
T ss_pred             cEEEEcCCCCCe----------eeecCCCCcCCCCCCCccee
Confidence            468899999874          56778899999986555444


No 30 
>PF11331 DUF3133:  Protein of unknown function (DUF3133);  InterPro: IPR021480  This eukaryotic family of proteins has no known function. 
Probab=49.24  E-value=10  Score=26.21  Aligned_cols=18  Identities=33%  Similarity=0.654  Sum_probs=15.9

Q ss_pred             eeeecCCCcceeeecccc
Q 029214           13 LCYIPCNFCNIVLAVSVP   30 (197)
Q Consensus        13 lCYV~CnfC~TILaVsVP   30 (197)
                      .==+||+-|..||-+++|
T Consensus        29 ~~klrCGaCs~vl~~s~~   46 (46)
T PF11331_consen   29 QQKLRCGACSEVLSFSLP   46 (46)
T ss_pred             eeEEeCCCCceeEEEecC
Confidence            556899999999999987


No 31 
>PF06382 DUF1074:  Protein of unknown function (DUF1074);  InterPro: IPR024460 This family consists of several proteins which appear to be specific to Insecta. The function of this family is unknown.
Probab=48.84  E-value=26  Score=30.56  Aligned_cols=35  Identities=23%  Similarity=0.574  Sum_probs=30.1

Q ss_pred             hhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214          127 SAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH  165 (197)
Q Consensus       127 SaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph  165 (197)
                      .+|=.||.+    .+..|.++..+|....||+.|...+.
T Consensus        84 naYLNFLRe----FRrkh~~L~p~dlI~~AAraW~rLSe  118 (183)
T PF06382_consen   84 NAYLNFLRE----FRRKHCGLSPQDLIQRAARAWCRLSE  118 (183)
T ss_pred             hHHHHHHHH----HHHHccCCCHHHHHHHHHHHHHhCCH
Confidence            578888764    77899999999999999999987654


No 32 
>KOG4684 consensus Uncharacterized conserved protein, contains C4-type Zn-finger [General function prediction only]
Probab=44.28  E-value=11  Score=34.29  Aligned_cols=33  Identities=36%  Similarity=0.828  Sum_probs=28.5

Q ss_pred             eeeecCCCcceeeeccccCCCccceee---eecCCCCCcccccc
Q 029214           13 LCYIPCNFCNIVLAVSVPCSSLLDIVT---VRCGHCSNLWSVNM   53 (197)
Q Consensus        13 lCYV~CnfC~TILaVsVPcssL~~tVT---VRCGHCtnLlSVNm   53 (197)
                      -|-|-|+.|+.+.        ||+|.|   -||-||-..+||--
T Consensus       168 gcRV~CgHC~~tF--------Lfnt~tnaLArCPHCrKvSsvGs  203 (275)
T KOG4684|consen  168 GCRVKCGHCNETF--------LFNTLTNALARCPHCRKVSSVGS  203 (275)
T ss_pred             ceEEEecCcccee--------ehhhHHHHHhcCCcccchhhhhh
Confidence            4899999998876        789998   59999999999844


No 33 
>TIGR01053 LSD1 zinc finger domain, LSD1 subclass. This model describes a putative zinc finger domain found in three closely spaced copies in Arabidopsis protein LSD1 and in two copies in other proteins from the same species. The motif resembles CxxCRxxLMYxxGASxVxCxxC
Probab=40.87  E-value=26  Score=22.18  Aligned_cols=25  Identities=28%  Similarity=0.681  Sum_probs=13.8

Q ss_pred             ecCCCcceeeeccccCCCccceeeeecCCCC
Q 029214           16 IPCNFCNIVLAVSVPCSSLLDIVTVRCGHCS   46 (197)
Q Consensus        16 V~CnfC~TILaVsVPcssL~~tVTVRCGHCt   46 (197)
                      |.|+-|.|+|+.-      ...-.|||--|.
T Consensus         2 ~~C~~C~t~L~yP------~gA~~vrCs~C~   26 (31)
T TIGR01053         2 VVCGGCRTLLMYP------RGASSVRCALCQ   26 (31)
T ss_pred             cCcCCCCcEeecC------CCCCeEECCCCC
Confidence            4566666666542      233456666664


No 34 
>KOG0526 consensus Nucleosome-binding factor SPN, POB3 subunit [Transcription; Replication, recombination and repair; Chromatin structure and dynamics]
Probab=38.70  E-value=42  Score=33.90  Aligned_cols=44  Identities=25%  Similarity=0.483  Sum_probs=39.6

Q ss_pred             CCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccC
Q 029214          118 TPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHF  163 (197)
Q Consensus       118 PPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~  163 (197)
                      -|-+..|+-|||=.|...+-..||+.  +|+.-|.=+.+...|+.-
T Consensus       532 dpnapkra~sa~m~w~~~~r~~ik~d--gi~~~dv~kk~g~~wk~m  575 (615)
T KOG0526|consen  532 DPNAPKRATSAYMLWLNASRESIKED--GISVGDVAKKAGEKWKQM  575 (615)
T ss_pred             CCCCCccchhHHHHHHHhhhhhHhhc--CchHHHHHHHHhHHHhhh
Confidence            45566899999999999999999999  999999999999999963


No 35 
>PF03811 Zn_Tnp_IS1:  InsA N-terminal domain;  InterPro: IPR003220 Insertion elements are mobile elements in DNA, usually encoding proteins required for transposition, for example transposases. Protein InsA is absolutely required for transposition of insertion element 1. This entry represents a short zinc binding domain found in IS1 InsA family protein. It is found at the N terminus of the protein and may be a DNA-binding domain.; GO: 0006313 transposition, DNA-mediated
Probab=37.54  E-value=23  Score=23.01  Aligned_cols=17  Identities=24%  Similarity=0.565  Sum_probs=13.8

Q ss_pred             cceeeeecCCCCCcccc
Q 029214           35 LDIVTVRCGHCSNLWSV   51 (197)
Q Consensus        35 ~~tVTVRCGHCtnLlSV   51 (197)
                      |.+|+|.|-+|.+-.+|
T Consensus         1 Ma~i~v~CP~C~s~~~v   17 (36)
T PF03811_consen    1 MAKIDVHCPRCQSTEGV   17 (36)
T ss_pred             CCcEeeeCCCCCCCCcc
Confidence            46899999999887654


No 36 
>COG4416 Com Mu-like prophage protein Com [General function prediction only]
Probab=37.44  E-value=7.9  Score=28.23  Aligned_cols=14  Identities=36%  Similarity=0.987  Sum_probs=11.3

Q ss_pred             eeeeecCCCCCccc
Q 029214           37 IVTVRCGHCSNLWS   50 (197)
Q Consensus        37 tVTVRCGHCtnLlS   50 (197)
                      +-|.||.||+-||-
T Consensus         2 ~~tiRC~~CnKlLa   15 (60)
T COG4416           2 MQTIRCAKCNKLLA   15 (60)
T ss_pred             ceeeehHHHhHHHH
Confidence            45899999998764


No 37 
>PF10963 DUF2765:  Protein of unknown function (DUF2765);  InterPro: IPR024406 This family of proteins with no known function is found in phages and suspected prophages.
Probab=33.12  E-value=65  Score=24.63  Aligned_cols=14  Identities=29%  Similarity=0.567  Sum_probs=10.0

Q ss_pred             HHHHHHHHhhCCCC
Q 029214          134 KEEIQRIKANNPDI  147 (197)
Q Consensus       134 K~ei~riK~~~P~i  147 (197)
                      |+.+..+-+++||.
T Consensus        48 KeaL~~lle~~PGa   61 (83)
T PF10963_consen   48 KEALKELLEENPGA   61 (83)
T ss_pred             HHHHHHHHHHCCCH
Confidence            56677777777776


No 38 
>PF05047 L51_S25_CI-B8:  Mitochondrial ribosomal protein L51 / S25 / CI-B8 domain ;  InterPro: IPR007741 Proteins containing this domain are located in the mitochondrion and include ribosomal protein L51, and S25. This domain is also found in mitochondrial NADH-ubiquinone oxidoreductase B8 subunit (CI-B8) 1.6.5.3 from EC. It is not known whether all members of this family form part of the NADH-ubiquinone oxidoreductase and whether they are also all ribosomal proteins.; PDB: 1S3A_A.
Probab=31.61  E-value=40  Score=22.21  Aligned_cols=18  Identities=28%  Similarity=0.719  Sum_probs=15.5

Q ss_pred             hHHHHHHHHHHHhhCCCC
Q 029214          130 NQFIKEEIQRIKANNPDI  147 (197)
Q Consensus       130 N~FmK~ei~riK~~~P~i  147 (197)
                      ..|+++.+..|+..||++
T Consensus         2 R~F~~~~lp~l~~~NP~v   19 (52)
T PF05047_consen    2 RDFLKNNLPTLKYHNPQV   19 (52)
T ss_dssp             HHHHHHTHHHHHHHSTT-
T ss_pred             HhHHHHhHHHHHHHCCCc
Confidence            369999999999999986


No 39 
>PF01020 Ribosomal_L40e:  Ribosomal L40e family;  InterPro: IPR001975 Ribosomes are the particles that catalyse mRNA-directed protein synthesis in all organisms. The codons of the mRNA are exposed on the ribosome to allow tRNA binding. This leads to the incorporation of amino acids into the growing polypeptide chain in accordance with the genetic information. Incoming amino acid monomers enter the ribosomal A site in the form of aminoacyl-tRNAs complexed with elongation factor Tu (EF-Tu) and GTP. The growing polypeptide chain, situated in the P site as peptidyl-tRNA, is then transferred to aminoacyl-tRNA and the new peptidyl-tRNA, extended by one residue, is translocated to the P site with the aid the elongation factor G (EF-G) and GTP as the deacylated tRNA is released from the ribosome through one or more exit sites [, ]. About 2/3 of the mass of the ribosome consists of RNA and 1/3 of protein. The proteins are named in accordance with the subunit of the ribosome which they belong to - the small (S1 to S31) and the large (L1 to L44). Usually they decorate the rRNA cores of the subunits.  Many ribosomal proteins, particularly those of the large subunit, are composed of a globular, surfaced-exposed domain with long finger-like projections that extend into the rRNA core to stabilise its structure. Most of the proteins interact with multiple RNA elements, often from different domains. In the large subunit, about 1/3 of the 23S rRNA nucleotides are at least in van der Waal's contact with protein, and L22 interacts with all six domains of the 23S rRNA. Proteins S4 and S7, which initiate assembly of the 16S rRNA, are located at junctions of five and four RNA helices, respectively. In this way proteins serve to organise and stabilise the rRNA tertiary structure. While the crucial activities of decoding and peptide transfer are RNA based, proteins play an active role in functions that may have evolved to streamline the process of protein synthesis. In addition to their function in the ribosome, many ribosomal proteins have some function 'outside' the ribosome [, ]. This family contains the L40 ribosomal protein from both archaea and eukaryotes. Bovine ribosomal protein L40 has been identified as a secondary RNA binding protein []. L40 is fused to a ubiquitin protein [].; GO: 0003735 structural constituent of ribosome, 0006412 translation, 0005840 ribosome; PDB: 3IZS_p 3IZR_p 2AYJ_A 4A1B_K 4A19_K 4A18_K 4A1D_K.
Probab=31.31  E-value=18  Score=25.80  Aligned_cols=9  Identities=56%  Similarity=1.258  Sum_probs=5.5

Q ss_pred             ecCCCCCcc
Q 029214           41 RCGHCSNLW   49 (197)
Q Consensus        41 RCGHCtnLl   49 (197)
                      +|||++||-
T Consensus        38 kCGhsn~LR   46 (52)
T PF01020_consen   38 KCGHSNNLR   46 (52)
T ss_dssp             SCTS-S-EE
T ss_pred             cCCCCcccC
Confidence            399999874


No 40 
>COG4888 Uncharacterized Zn ribbon-containing protein [General function prediction only]
Probab=30.43  E-value=37  Score=27.31  Aligned_cols=45  Identities=18%  Similarity=0.362  Sum_probs=33.0

Q ss_pred             eecCCCcceeeeccccCCCccceeeeecCCCCCccccccchhccC
Q 029214           15 YIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMAAAFQS   59 (197)
Q Consensus        15 YV~CnfC~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr~llqs   59 (197)
                      |--|-||+-...|+--.+--.++-|+-||-|.--.-+-.++++++
T Consensus        22 ~FtCp~Cghe~vs~ctvkk~~~~g~~~Cg~CGls~e~ev~~l~~~   66 (104)
T COG4888          22 TFTCPRCGHEKVSSCTVKKTVNIGTAVCGNCGLSFECEVPELSEP   66 (104)
T ss_pred             eEecCccCCeeeeEEEEEecCceeEEEcccCcceEEEeccccccc
Confidence            667999999999887767777888999999974433444444443


No 41 
>PF12876 Cellulase-like:  Sugar-binding cellulase-like;  InterPro: IPR024778 O-Glycosyl hydrolases 3.2.1. from EC are a widespread group of enzymes that hydrolyse the glycosidic bond between two or more carbohydrates, or between a carbohydrate and a non-carbohydrate moiety. A classification system for glycosyl hydrolases, based on sequence similarity, has led to the definition of 85 different families [, ]. This classification is available on the CAZy (CArbohydrate-Active EnZymes) web site. This entry represents a family of putative cellulase enzymes.; PDB: 3GYC_B.
Probab=30.35  E-value=73  Score=23.06  Aligned_cols=31  Identities=26%  Similarity=0.428  Sum_probs=22.8

Q ss_pred             CCcccCCCchhhhHHHHHHHHHHHhhCCCCC
Q 029214          118 TPEKRQRVPSAYNQFIKEEIQRIKANNPDIS  148 (197)
Q Consensus       118 PPEKRQR~PSaYN~FmK~ei~riK~~~P~i~  148 (197)
                      |.+.......+|-.|+++-++.||+.+|+.+
T Consensus        29 ~~~~~~~~~~~~~~~l~~~~~~iR~~dP~~p   59 (88)
T PF12876_consen   29 PAEWGDPKAEAYAEWLKEAFRWIRAVDPSQP   59 (88)
T ss_dssp             TT-TT-TTSHHHHHHHHHHHHHHHTT-TTS-
T ss_pred             cccccchhHHHHHHHHHHHHHHHHHhCCCCc
Confidence            3344555778999999999999999999764


No 42 
>KOG4715 consensus SWI/SNF-related matrix-associated actin-dependent regulator of chromatin  [Chromatin structure and dynamics]
Probab=29.70  E-value=68  Score=30.76  Aligned_cols=46  Identities=24%  Similarity=0.460  Sum_probs=39.8

Q ss_pred             CCCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214          117 TTPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH  165 (197)
Q Consensus       117 kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph  165 (197)
                      |||||.   .-.|=+|-+.=-..+|+.||++--=|.=+.+++.|.+.|.
T Consensus        63 kppekp---l~pymrySrkvWd~VkA~nPe~kLWeiGK~Ig~mW~dLpd  108 (410)
T KOG4715|consen   63 KPPEKP---LMPYMRYSRKVWDQVKASNPELKLWEIGKIIGGMWLDLPD  108 (410)
T ss_pred             CCCCcc---cchhhHHhhhhhhhhhccCcchHHHHHHHHHHHHHhhCcc
Confidence            477764   5679999888899999999999999999999999998875


No 43 
>PF14599 zinc_ribbon_6:  Zinc-ribbon; PDB: 2K2D_A.
Probab=28.68  E-value=35  Score=24.60  Aligned_cols=30  Identities=30%  Similarity=0.703  Sum_probs=15.9

Q ss_pred             ceeeecCCCcceeeeccccCCCccceeeeecCCCCC
Q 029214           12 QLCYIPCNFCNIVLAVSVPCSSLLDIVTVRCGHCSN   47 (197)
Q Consensus        12 ~lCYV~CnfC~TILaVsVPcssL~~tVTVRCGHCtn   47 (197)
                      ..+.|-||=|..-=-  |    -|-++..||+||.+
T Consensus        27 ~~v~IlCNDC~~~s~--v----~fH~lg~KC~~C~S   56 (61)
T PF14599_consen   27 KKVWILCNDCNAKSE--V----PFHFLGHKCSHCGS   56 (61)
T ss_dssp             -EEEEEESSS--EEE--E----E--TT----TTTS-
T ss_pred             CEEEEECCCCCCccc--e----eeeHhhhcCCCCCC
Confidence            579999999987543  3    37788899999975


No 44 
>PF10159 MMtag:  Kinase phosphorylation protein;  InterPro: IPR019315  This entry represents a glycine-rich domain that is the most highly conserved region of a family of proteins that, in vertebrates, are associated with tumours in multiple myelomas. The region may contain phosphorylation sites for several protein kinases, as well as N-myristoylation sites and nuclear localisation signals, so it might act as a signal molecule in the nucleus []. 
Probab=28.50  E-value=52  Score=25.13  Aligned_cols=13  Identities=54%  Similarity=0.567  Sum_probs=11.6

Q ss_pred             HHHHHHHHHHhhC
Q 029214          132 FIKEEIQRIKANN  144 (197)
Q Consensus       132 FmK~ei~riK~~~  144 (197)
                      =.++||++||+..
T Consensus        57 ~~~eE~~~iK~~E   69 (78)
T PF10159_consen   57 ERKEEIRRIKEAE   69 (78)
T ss_pred             hHHHHHHHHHHHH
Confidence            6799999999986


No 45 
>PF05180 zf-DNL:  DNL zinc finger;  InterPro: IPR007853 Zinc finger (Znf) domains are relatively small protein motifs which contain multiple finger-like protrusions that make tandem contacts with their target molecule. Some of these domains bind zinc, but many do not; instead binding other metals such as iron, or no metal at all. For example, some family members form salt bridges to stabilise the finger-like folds. They were first identified as a DNA-binding motif in transcription factor TFIIIA from Xenopus laevis (African clawed frog), however they are now recognised to bind DNA, RNA, protein and/or lipid substrates [, , , , ]. Their binding properties depend on the amino acid sequence of the finger domains and of the linker between fingers, as well as on the higher-order structures and the number of fingers. Znf domains are often found in clusters, where fingers can have different binding specificities. There are many superfamilies of Znf motifs, varying in both sequence and structure. They display considerable versatility in binding modes, even between members of the same class (e.g. some bind DNA, others protein), suggesting that Znf motifs are stable scaffolds that have evolved specialised functions. For example, Znf-containing proteins function in gene transcription, translation, mRNA trafficking, cytoskeleton organisation, epithelial development, cell adhesion, protein folding, chromatin remodelling and zinc sensing, to name but a few []. Zinc-binding motifs are stable structures, and they rarely undergo conformational changes upon binding their target.  The DNL-type zinc finger is found in Tim15, a zinc finger protein essential for protein import into mitochondria. Mitochondrial functions rely on the correct transport of resident proteins synthesized in the cytosol to mitochondria. Protein import into mitochondria is mediated by membrane protein complexes, protein translocators, in the outer and inner mitochondrial membranes, in cooperation with their assistant proteins in the cytosol, intermembrane space and matrix. Proteins destined to the mitochondrial matrix cross the outer membrane with the aid of the outer membrane translocator, the tOM40 complex, and then the inner membrane with the aid of the inner membrane translocator, the TIM23 complex, and mitochondrial motor and chaperone (MMC) proteins including mitochondrial heat- shock protein 70 (mtHsp70), and translocase in the inner mitochondrial membrane (Tim)15. Tim15 is also known as zinc finger motif (Zim)17 or mtHsp70 escort protein (Hep)1. Tim15 contains a zinc-finger motif (CXXC and CXXC) of ~100 residues, which has been named DNL after a short C-terminal motif of D(N/H)L [, , ]. The DNL-type zinc finger is an L-shaped molecule. The two CXXC motifs are located at the end of the L, and are sandwiched by two- stranded antiparallel beta-sheets. Two short alpha-helices constitute another leg of the L. The outer (convex) face of the L has a large acidic groove, which is lined with five acidic residues, whereas the inner (concave) face of the L has two positively charged residues, next to the CXXC motifs []. This entry represents the DNL-type zinc finger.; GO: 0008270 zinc ion binding; PDB: 2E2Z_A.
Probab=28.02  E-value=33  Score=25.27  Aligned_cols=32  Identities=28%  Similarity=0.482  Sum_probs=15.9

Q ss_pred             cCCCcceeeeccccCCC-ccceeeeecCCCCCc
Q 029214           17 PCNFCNIVLAVSVPCSS-LLDIVTVRCGHCSNL   48 (197)
Q Consensus        17 ~CnfC~TILaVsVPcss-L~~tVTVRCGHCtnL   48 (197)
                      -|+-|+|-=+-.+-=.+ =--+|-|||+.|.|.
T Consensus         6 TC~~C~~Rs~~~~sk~aY~~GvViv~C~gC~~~   38 (66)
T PF05180_consen    6 TCNKCGTRSAKMFSKQAYHKGVVIVQCPGCKNR   38 (66)
T ss_dssp             EETTTTEEEEEEEEHHHHHTSEEEEE-TTS--E
T ss_pred             EcCCCCCccceeeCHHHHhCCeEEEECCCCcce
Confidence            36666665442221111 124799999999985


No 46 
>PF02892 zf-BED:  BED zinc finger;  InterPro: IPR003656 Zinc finger (Znf) domains are relatively small protein motifs which contain multiple finger-like protrusions that make tandem contacts with their target molecule. Some of these domains bind zinc, but many do not; instead binding other metals such as iron, or no metal at all. For example, some family members form salt bridges to stabilise the finger-like folds. They were first identified as a DNA-binding motif in transcription factor TFIIIA from Xenopus laevis (African clawed frog), however they are now recognised to bind DNA, RNA, protein and/or lipid substrates [, , , , ]. Their binding properties depend on the amino acid sequence of the finger domains and of the linker between fingers, as well as on the higher-order structures and the number of fingers. Znf domains are often found in clusters, where fingers can have different binding specificities. There are many superfamilies of Znf motifs, varying in both sequence and structure. They display considerable versatility in binding modes, even between members of the same class (e.g. some bind DNA, others protein), suggesting that Znf motifs are stable scaffolds that have evolved specialised functions. For example, Znf-containing proteins function in gene transcription, translation, mRNA trafficking, cytoskeleton organisation, epithelial development, cell adhesion, protein folding, chromatin remodelling and zinc sensing, to name but a few []. Zinc-binding motifs are stable structures, and they rarely undergo conformational changes upon binding their target.  This entry represents predicted BED-type zinc finger domains. The BED finger which was named after the Drosophila proteins BEAF and DREF, is found in one or more copies in cellular regulatory factors and transposases from plants, animals and fungi. The BED finger is an about 50 to 60 amino acid residues domain that contains a characteristic motif with two highly conserved aromatic positions, as well as a shared pattern of cysteines and histidines that is predicted to form a zinc finger. As diverse BED fingers are able to bind DNA, it has been suggested that DNA-binding is the general function of this domain []. Some proteins known to contain a BED domain include animal, plant and fungi AC1 and Hobo-like transposases; Caenorhabditis elegans Dpy-20 protein, a predicted cuticular gene transcriptional regulator; Drosophila BEAF (boundary element-associated factor), thought to be involved in chromatin insulation; Drosophila DREF, a transcriptional regulator for S-phase genes; and tobacco 3AF1 and tomato E4/E8-BP1, light- and ethylene-regulated DNA binding proteins that contain two BED fingers. More information about these proteins can be found at Protein of the Month: Zinc Fingers [].; GO: 0003677 DNA binding; PDB: 2DJR_A 2CT5_A.
Probab=27.89  E-value=27  Score=22.13  Aligned_cols=17  Identities=24%  Similarity=0.522  Sum_probs=10.1

Q ss_pred             ceeeecCCCcceeeecc
Q 029214           12 QLCYIPCNFCNIVLAVS   28 (197)
Q Consensus        12 ~lCYV~CnfC~TILaVs   28 (197)
                      ..-++.|.+|..++..+
T Consensus        13 ~~~~a~C~~C~~~~~~~   29 (45)
T PF02892_consen   13 DKKKAKCKYCGKVIKYS   29 (45)
T ss_dssp             CSS-EEETTTTEE----
T ss_pred             CcCeEEeCCCCeEEeeC
Confidence            45689999999888765


No 47 
>COG3712 FecR Fe2+-dicitrate sensor, membrane component [Inorganic ion transport and metabolism / Signal transduction mechanisms]
Probab=26.70  E-value=58  Score=30.33  Aligned_cols=31  Identities=23%  Similarity=0.492  Sum_probs=27.8

Q ss_pred             HHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214          132 FIKEEIQRIKANNPDISHREAFSTAAKNWAHFP  164 (197)
Q Consensus       132 FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~P  164 (197)
                      -..+|.++|++..|+  |.+||..+..-|....
T Consensus        32 ~~r~af~~W~~~~p~--H~~A~~~~e~lw~~l~   62 (322)
T COG3712          32 ADRAAFERWRAASPE--HARAWERAERLWQALG   62 (322)
T ss_pred             HHHHHHHHHHhcCHH--HHHHHHHHHHHHhhhc
Confidence            357899999999996  9999999999999865


No 48 
>PF00527 E7:  E7 protein, Early protein;  InterPro: IPR000148 This family includes the E7 oncoprotein from various papillomaviruses []. Along with E5 and E6 their activities seem to be especially important for viral oncogenesis. E5 is located at the cell surface and reduces cell gap-gap junction communication. In cervical cancer E5 is expressed in earlier stages of neoplastic transformation of the cervical epithelium during viral infection. The role of E7 is less well understood but it has been shown to impede growth arrest signals in both NIH 3T3 cells and HFKs and that this correlates with elevated cdc25A gene expression. This deregulation of cdc25A is linked to disruption of cell cycle arrest [].; GO: 0003677 DNA binding, 0003700 sequence-specific DNA binding transcription factor activity, 0006355 regulation of transcription, DNA-dependent, 0005622 intracellular; PDB: 2F8B_A 2EWL_A 2B9D_A.
Probab=26.55  E-value=26  Score=26.67  Aligned_cols=17  Identities=29%  Similarity=0.419  Sum_probs=7.1

Q ss_pred             ecCCCcceeeeccccCC
Q 029214           16 IPCNFCNIVLAVSVPCS   32 (197)
Q Consensus        16 V~CnfC~TILaVsVPcs   32 (197)
                      +.|.+|+..|-+.|=++
T Consensus        53 t~C~~C~~~lrl~V~as   69 (92)
T PF00527_consen   53 TCCGRCGKRLRLVVVAS   69 (92)
T ss_dssp             EEBTTT--EEEEEEEC-
T ss_pred             eECCCCCCEEEEEEEeC
Confidence            34555555555544444


No 49 
>PF02723 NS3_envE:  Non-structural protein NS3/Small envelope protein E;  InterPro: IPR003873 This is a family of small nonstructural proteins, well conserved among Coronavirus strains. This protein is also found in Murine hepatitis virus as small envelope protein E.; GO: 0016020 membrane
Probab=25.84  E-value=15  Score=28.20  Aligned_cols=16  Identities=38%  Similarity=0.910  Sum_probs=13.2

Q ss_pred             CceeeecCCCcceeee
Q 029214           11 EQLCYIPCNFCNIVLA   26 (197)
Q Consensus        11 E~lCYV~CnfC~TILa   26 (197)
                      =|||..=|+||||++.
T Consensus        37 IqLC~~cc~~~n~~v~   52 (82)
T PF02723_consen   37 IQLCFQCCRLCNTTVY   52 (82)
T ss_pred             HHHHHHHhhhhcceEe
Confidence            3789999999998875


No 50 
>PF05164 ZapA:  Cell division protein ZapA;  InterPro: IPR007838 This entry a structural domain found in the cell division protein ZapA, as well as in related proteins. This domain has a core structure consisting of two layers alpha/beta, and has a long C-terminal helix that forms dimeric parallel and tetrameric antiparallel coiled coils []. ZapA interacts with FtsZ, where FtsZ is part of a mid-cell cytokinetic structure termed the Z-ring that recruits a hierarchy of fission related proteins early in the bacterial cell cycle. ZapA drives the polymerisation and filament bundling of FtsZ, thereby contributing to the spatio-temporal tuning of the Z-ring.; PDB: 1T3U_B 1W2E_B 3HNW_A.
Probab=24.78  E-value=1.1e+02  Score=21.64  Aligned_cols=32  Identities=34%  Similarity=0.420  Sum_probs=28.4

Q ss_pred             HHHHHHHHHHHhhCCCCCHHHHHHHHHHhhcc
Q 029214          131 QFIKEEIQRIKANNPDISHREAFSTAAKNWAH  162 (197)
Q Consensus       131 ~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~  162 (197)
                      .++.+.|..++...|.++...++..||-|-++
T Consensus        28 ~~i~~~i~~~~~~~~~~~~~~~~vlaaLnla~   59 (89)
T PF05164_consen   28 ELINEKINEIKKKYPKLSPERLAVLAALNLAD   59 (89)
T ss_dssp             HHHHHHHHHHCTTCCTSSHHHHHHHHHHHHHH
T ss_pred             HHHHHHHHHHHHHcCCCCHHHHHHHHHHHHHH
Confidence            57889999999999999999999999987653


No 51 
>PF04769 MAT_Alpha1:  Mating-type protein MAT alpha 1;  InterPro: IPR006856 This family includes Saccharomyces cerevisiae (Baker's yeast) mating type protein alpha 1 (P01365 from SWISSPROT). MAT alpha 1 is a transcription activator that activates mating-type alpha-specific genes with the help of the MADS-box containing MCM1 transcription factor, which together bind cooperatively to PQ elements upstream of alpha-specific genes. The MCM1-MATalpha1 complex is required for the proper DNA-bending that is needed for transcriptional activation []. Alpha 1 interacts in vivo with STE12, linking expression of alpha-specific genes to the alpha-pheromone (IPR006742 from INTERPRO) response pathway [].; GO: 0000772 mating pheromone activity, 0003677 DNA binding, 0045895 positive regulation of transcription, mating-type specific, 0005634 nucleus
Probab=24.22  E-value=1.9e+02  Score=25.17  Aligned_cols=44  Identities=23%  Similarity=0.366  Sum_probs=37.1

Q ss_pred             CCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCC
Q 029214          118 TPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPH  165 (197)
Q Consensus       118 PPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Ph  165 (197)
                      .++++.|.-.+|+.|+-=    .+...|+.+.|++-...++.|...|+
T Consensus        40 ~~~~~kr~lN~Fm~FRsy----y~~~~~~~~Qk~~S~~l~~lW~~dp~   83 (201)
T PF04769_consen   40 SPEKAKRPLNGFMAFRSY----YSPIFPPLPQKELSGILTKLWEKDPF   83 (201)
T ss_pred             cccccccchhHHHHHHHH----HHhhcCCcCHHHHHHHHHHHHhCCcc
Confidence            678888888899998764    33677889999999999999999886


No 52 
>PF13408 Zn_ribbon_recom:  Recombinase zinc beta ribbon domain
Probab=24.08  E-value=37  Score=22.10  Aligned_cols=13  Identities=38%  Similarity=0.986  Sum_probs=10.8

Q ss_pred             eeecCCCCCcccc
Q 029214           39 TVRCGHCSNLWSV   51 (197)
Q Consensus        39 TVRCGHCtnLlSV   51 (197)
                      .|+||+|..-+..
T Consensus         5 ~l~C~~CG~~m~~   17 (58)
T PF13408_consen    5 LLRCGHCGSKMTR   17 (58)
T ss_pred             cEEcccCCcEeEE
Confidence            4799999988775


No 53 
>KOG0527 consensus HMG-box transcription factor [Transcription]
Probab=23.77  E-value=96  Score=28.94  Aligned_cols=62  Identities=18%  Similarity=0.372  Sum_probs=50.0

Q ss_pred             CCCcccCCCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhhccCCCccccccccCCCCCCCcchhhc
Q 029214          117 TTPEKRQRVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNWAHFPHIHFGLMLEANNQPKLDDASGN  186 (197)
Q Consensus       117 kPPEKRQR~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW~~~Phihfgl~~d~~~~~~~~~~~~~  186 (197)
                      +..++-.|--.||=-|=+.|=++|-++||+|---|.=+...+.|+.        .-|..|..=+|+++--
T Consensus        58 ~~~~hIKRPMNAFMVWSq~~RRkma~qnP~mHNSEISK~LG~~WK~--------Lse~EKrPFi~EAeRL  119 (331)
T KOG0527|consen   58 TSTDRIKRPMNAFMVWSQGQRRKLAKQNPKMHNSEISKRLGAEWKL--------LSEEEKRPFVDEAERL  119 (331)
T ss_pred             CCccccCCCcchhhhhhHHHHHHHHHhCcchhhHHHHHHHHHHHhh--------cCHhhhccHHHHHHHH
Confidence            4677778899999999999999999999999889999999999984        2345555555555433


No 54 
>PRK09774 fec operon regulator FecR; Reviewed
Probab=22.74  E-value=72  Score=28.45  Aligned_cols=29  Identities=10%  Similarity=0.109  Sum_probs=25.2

Q ss_pred             HHHHHHHHhhCCCCCHHHHHHHHHHhhccCC
Q 029214          134 KEEIQRIKANNPDISHREAFSTAAKNWAHFP  164 (197)
Q Consensus       134 K~ei~riK~~~P~i~hkEaFs~aAknW~~~P  164 (197)
                      +++.++|.+++|+  |++||..+..-|....
T Consensus        33 ~~~f~~Wl~a~p~--H~~A~~~~~~lw~~~~   61 (319)
T PRK09774         33 EARWQQWYEQDQD--NQWAWQQVENLRNQMG   61 (319)
T ss_pred             HHHHHHHHhCCHH--HHHHHHHHHHHHHHhh
Confidence            4678999999997  9999999999997754


No 55 
>PF04420 CHD5:  CHD5-like protein;  InterPro: IPR007514 Members of this family are probably coiled-coil proteins that are similar to the CHD5 (Congenital heart disease 5) protein. The exact molecular function of these eukaryotic proteins is unknown.; PDB: 3SJA_H 3SJC_D 3SJB_D 3ZS8_D 3VLC_E.
Probab=22.45  E-value=60  Score=26.65  Aligned_cols=37  Identities=24%  Similarity=0.242  Sum_probs=30.7

Q ss_pred             CCchhhhHHHHHHHHHHHhhCCCCCHHHHHHHHHHhh
Q 029214          124 RVPSAYNQFIKEEIQRIKANNPDISHREAFSTAAKNW  160 (197)
Q Consensus       124 R~PSaYN~FmK~ei~riK~~~P~i~hkEaFs~aAknW  160 (197)
                      ..++.-.+=++.||+.+|++.-.++..|-|...||+=
T Consensus        36 ~~~~~~~~~l~~Ei~~l~~E~~~iS~qDeFAkwaKl~   72 (161)
T PF04420_consen   36 SKSSKEQRQLRKEILQLKRELNAISAQDEFAKWAKLN   72 (161)
T ss_dssp             -HHHHHHHHHHHHHHHHHHHHTTS-TTTSHHHHHHHH
T ss_pred             ccccHHHHHHHHHHHHHHHHHHcCCcHHHHHHHHHHH
Confidence            3566677789999999999999999999999998863


No 56 
>smart00614 ZnF_BED BED zinc finger. DNA-binding domain in chromatin-boundary-element-binding proteins and transposases
Probab=22.44  E-value=49  Score=21.92  Aligned_cols=14  Identities=21%  Similarity=0.548  Sum_probs=11.6

Q ss_pred             eeecCCCcceeeec
Q 029214           14 CYIPCNFCNIVLAV   27 (197)
Q Consensus        14 CYV~CnfC~TILaV   27 (197)
                      -++.|++|..+|..
T Consensus        17 ~~a~C~~C~~~l~~   30 (50)
T smart00614       17 QRAKCKYCGKKLSR   30 (50)
T ss_pred             eEEEecCCCCEeee
Confidence            58999999998863


No 57 
>PF09788 Tmemb_55A:  Transmembrane protein 55A;  InterPro: IPR019178  Members of this family catalyse the hydrolysis of the 4-position phosphate of phosphatidylinositol 4,5-bisphosphate, in the reaction:  1-phosphatidyl-myo-inositol 4,5-bisphosphate + H(2)O = 1-phosphatidyl-1D-myo-inositol 5-phosphate + phosphate.  
Probab=21.75  E-value=81  Score=28.80  Aligned_cols=17  Identities=47%  Similarity=0.878  Sum_probs=13.3

Q ss_pred             ceeeeecCCCCCccccc
Q 029214           36 DIVTVRCGHCSNLWSVN   52 (197)
Q Consensus        36 ~tVTVRCGHCtnLlSVN   52 (197)
                      ...-|.||||.+-..-|
T Consensus       154 ~~~rv~CghC~~~Fl~~  170 (256)
T PF09788_consen  154 GSCRVICGHCSNTFLFN  170 (256)
T ss_pred             CceeEECCCCCCcEecc
Confidence            45779999999976654


No 58 
>PF09102 Exotox-A_target:  Exotoxin A, targeting;  InterPro: IPR015186 This domain, found in Pseudomonas aeruginosa exotoxin A, is responsible for transmembrane targeting of the toxin, as well as transmembrane translocation of the catalytic domain into the cytoplasmic compartment. A furin cleavage site is present within the domain: cleavage generates a 37 kDa carboxy-terminal fragment, which includes the enzymatic domain, which is then is translocated into the cytoplasm. It adopts a helical structure, with six alpha-helices forming a bundle []. ; PDB: 1IKP_A 1IKQ_A 2Q5T_A 3Q9O_A.
Probab=21.48  E-value=48  Score=27.63  Aligned_cols=40  Identities=25%  Similarity=0.475  Sum_probs=33.4

Q ss_pred             HHHHHHHhhCCCCCHHHHHHHHHHhhccCCCccccccccCC
Q 029214          135 EEIQRIKANNPDISHREAFSTAAKNWAHFPHIHFGLMLEAN  175 (197)
Q Consensus       135 ~ei~riK~~~P~i~hkEaFs~aAknW~~~Phihfgl~~d~~  175 (197)
                      ..+.||.+++|++ .+.+.+.|+.-..++-.-|-||.+++.
T Consensus        77 ~DL~~~~~~~P~~-~~~~LT~A~~~~~~~V~~~~Gltpe~~  116 (143)
T PF09102_consen   77 SDLRRINENNPGM-VTQVLTVARQIYNDYVTHHPGLTPEQT  116 (143)
T ss_dssp             HHHHHHHHHSCCH-HHHHHHHHHHHHHHHHCCSTT--HHHH
T ss_pred             hHHHHHHhcCchH-HHHHHHHHHHHHHHHHHcCCCCCcccc
Confidence            5789999999996 789999999999999888999988743


No 59 
>TIGR02147 Fsuc_second hypothetical protein, TIGR02147. This family consists of the 40 members of a paralogous protein family in the rumen anaerobe Fibrobacter succinogenes S85. Member proteins are about 270 residues long and appear to lack signal sequences and transmembrane helices. The only perfectly conserved residue is a glycine in an otherwise poorly conserved region, suggesting members are not enzymes. The family is not characterized.
Probab=21.46  E-value=88  Score=28.06  Aligned_cols=26  Identities=19%  Similarity=0.438  Sum_probs=23.6

Q ss_pred             hhhhHHHHHHHHHHHhhCCCCCHHHH
Q 029214          127 SAYNQFIKEEIQRIKANNPDISHREA  152 (197)
Q Consensus       127 SaYN~FmK~ei~riK~~~P~i~hkEa  152 (197)
                      ..|..||++...+-|+.+|..|.|+-
T Consensus         8 ~dYR~fl~d~ye~rk~~~p~fS~R~f   33 (271)
T TIGR02147         8 TDYRKYLRDYYEERKKTDPAFSWRFF   33 (271)
T ss_pred             hhHHHHHHHHHHHHhccCcCcCHHHH
Confidence            46999999999999999999999873


No 60 
>COG4357 Zinc finger domain containing protein (CHY type) [Function unknown]
Probab=21.18  E-value=27  Score=28.08  Aligned_cols=39  Identities=21%  Similarity=0.477  Sum_probs=27.6

Q ss_pred             eecCCCcceeeec--------cccCC----CccceeeeecCCCCCcccccc
Q 029214           15 YIPCNFCNIVLAV--------SVPCS----SLLDIVTVRCGHCSNLWSVNM   53 (197)
Q Consensus        15 YV~CnfC~TILaV--------sVPcs----sL~~tVTVRCGHCtnLlSVNm   53 (197)
                      =.+|.-|++--|-        .-|..    ..++.-.|.||+|-++|+++=
T Consensus        26 alkc~~C~kyYaCy~CHdel~~Hpf~p~~~~~~~~~~iiCGvC~~~LT~~E   76 (105)
T COG4357          26 ALKCKCCQKYYACYHCHDELEDHPFEPWGLQEFNPKAIICGVCRKLLTRAE   76 (105)
T ss_pred             eeeechhhhhhhHHHHHhHHhcCCCccCChhhcCCccEEhhhhhhhhhHHH
Confidence            3577777776652        22332    577788899999999999753


No 61 
>PRK12336 translation initiation factor IF-2 subunit beta; Provisional
Probab=20.83  E-value=1e+02  Score=26.33  Aligned_cols=36  Identities=25%  Similarity=0.602  Sum_probs=27.8

Q ss_pred             eecCCCc---ceeeeccccCCCccceeeeecCCCCCccccccchh
Q 029214           15 YIPCNFC---NIVLAVSVPCSSLLDIVTVRCGHCSNLWSVNMAAA   56 (197)
Q Consensus        15 YV~CnfC---~TILaVsVPcssL~~tVTVRCGHCtnLlSVNmr~l   56 (197)
                      ||.|.-|   +|.|.+.   .   .+...+|.-|..-.+|.-...
T Consensus        98 yV~C~~C~~pdT~l~k~---~---~~~~l~C~aCGa~~~v~~~~~  136 (201)
T PRK12336         98 YVICSECGLPDTRLVKE---D---RVLMLRCDACGAHRPVKKRKA  136 (201)
T ss_pred             eEECCCCCCCCcEEEEc---C---CeEEEEcccCCCCcccccccc
Confidence            9999999   4777664   2   466789999999998865543


No 62 
>KOG4520 consensus Predicted coiled-coil protein [General function prediction only]
Probab=20.77  E-value=1.2e+02  Score=27.36  Aligned_cols=15  Identities=33%  Similarity=0.428  Sum_probs=12.2

Q ss_pred             hHHHHHHHHHHHhhC
Q 029214          130 NQFIKEEIQRIKANN  144 (197)
Q Consensus       130 N~FmK~ei~riK~~~  144 (197)
                      ---|||||++||+..
T Consensus        62 ~~~~keEi~~vkE~E   76 (238)
T KOG4520|consen   62 KEKYKEEILEVKERE   76 (238)
T ss_pred             hHHHHHHHHHHHHHH
Confidence            345899999999874


No 63 
>PF04690 YABBY:  YABBY protein;  InterPro: IPR006780 YABBY proteins are a group of plant-specific transcription factors involved in the specification of abaxial polarity in lateral organs such as leaves and floral organs [, ].
Probab=20.07  E-value=50  Score=28.24  Aligned_cols=19  Identities=21%  Similarity=0.518  Sum_probs=15.9

Q ss_pred             eecCCCcceeeeccccCCC
Q 029214           15 YIPCNFCNIVLAVSVPCSS   33 (197)
Q Consensus        15 YV~CnfC~TILaVsVPcss   33 (197)
                      =|+|+-|+.+|-|.+.-..
T Consensus        36 TVRCGHCtNLLSVNm~~~~   54 (170)
T PF04690_consen   36 TVRCGHCTNLLSVNMRALL   54 (170)
T ss_pred             ceeccCccceeeeeccccc
Confidence            4999999999999886544


Done!