Query         008451
Match_columns 565
No_of_seqs    44 out of 46
Neff          2.5 
Searched_HMMs 46136
Date          Thu Mar 28 12:19:31 2013
Command       hhsearch -i /work/01045/syshi/csienesis_hhblits_a3m/008451.a3m -d /work/01045/syshi/HHdatabase/Cdd.hhm -o /work/01045/syshi/hhsearch_cdd/008451hhsearch_cdd -cpu 12 -v 0 

 No Hit                             Prob E-value P-value  Score    SS Cols Query HMM  Template HMM
  1 KOG3583 Uncharacterized conser 100.0 2.6E-32 5.6E-37  264.5  11.7  174   37-234     8-202 (279)
  2 PF10232 Med8:  Mediator of RNA 100.0 1.5E-31 3.3E-36  255.4   2.4  170   39-228    11-192 (226)
  3 PF12128 DUF3584:  Protein of u  79.5       8 0.00017   46.0   9.0  132   39-183   511-653 (1201)
  4 KOG2235 Uncharacterized conser  75.4     9.7 0.00021   43.8   7.8  184   41-243   536-720 (776)
  5 PF15011 CK2S:  Casein Kinase 2  74.4      13 0.00029   35.3   7.4   30  153-182    73-102 (168)
  6 KOG2129 Uncharacterized conser  61.3 2.9E+02  0.0062   31.3  15.0  104   32-173   172-275 (552)
  7 PF05983 Med7:  MED7 protein;    57.2      90   0.002   29.7   9.3   89   45-174    71-161 (162)
  8 PF06248 Zw10:  Centromere/kine  55.5 1.5E+02  0.0032   32.8  11.8  127   25-183     6-140 (593)
  9 PF04253 TFR_dimer:  Transferri  55.3      24 0.00052   31.0   4.8   60   96-175    64-123 (125)
 10 TIGR03017 EpsF chain length de  54.8   1E+02  0.0023   31.9  10.1  136   40-181   172-313 (444)
 11 TIGR00634 recN DNA repair prot  49.5      39 0.00083   37.0   6.3   31  151-181   346-376 (563)
 12 PF11336 DUF3138:  Protein of u  49.0      32  0.0007   38.3   5.5   75  150-224    24-118 (514)
 13 PRK09039 hypothetical protein;  48.5   1E+02  0.0022   32.4   8.9   68  150-225   143-218 (343)
 14 KOG3091 Nuclear pore complex,   47.0      91   0.002   35.2   8.6   27   40-66    377-403 (508)
 15 TIGR01834 PHA_synth_III_E poly  45.6      54  0.0012   34.8   6.4  107   52-179   207-317 (320)
 16 PF10146 zf-C4H2:  Zinc finger-  44.5 3.6E+02  0.0077   27.4  11.8   27   40-66      2-28  (230)
 17 PF08580 KAR9:  Yeast cortical   44.3 1.2E+02  0.0026   35.1   9.2   21  148-168   210-230 (683)
 18 PRK04863 mukB cell division pr  43.1 1.2E+02  0.0027   37.9   9.7   46  136-183  1075-1120(1486)
 19 KOG3598 Thyroid hormone recept  40.8      78  0.0017   40.1   7.4   22   97-118  1768-1789(2220)
 20 PF08654 DASH_Dad2:  DASH compl  40.7      23 0.00051   31.8   2.5   64   30-93      5-73  (103)
 21 PF07106 TBPIP:  Tat binding pr  37.9 2.3E+02  0.0051   26.3   8.7  127   72-213    17-160 (169)
 22 PF05377 FlaC_arch:  Flagella a  37.4      55  0.0012   27.1   3.9   46  139-186     4-49  (55)
 23 COG0497 RecN ATPase involved i  37.4 1.2E+02  0.0027   34.5   7.9   88   21-110   200-300 (557)
 24 cd00176 SPEC Spectrin repeats,  35.5 1.8E+02  0.0039   25.4   7.2   28  154-181   182-209 (213)
 25 PF15445 ATS:  acidic terminal   35.5      57  0.0012   35.8   4.8   42  147-188   281-325 (437)
 26 PF11887 DUF3407:  Protein of u  35.2      57  0.0012   32.9   4.6   29   76-104    56-84  (267)
 27 PRK04778 septation ring format  34.3 1.1E+02  0.0024   33.8   7.0   63   37-104   273-342 (569)
 28 cd03204 GST_C_GDAP1 GST_C fami  33.9 1.4E+02  0.0029   26.7   6.2   68  128-199     2-70  (111)
 29 PRK15048 methyl-accepting chem  32.7      69  0.0015   34.2   4.9   26  201-226   367-400 (553)
 30 PRK12425 fumarate hydratase; P  32.5 5.1E+02   0.011   28.5  11.4   33   76-108   274-312 (464)
 31 PF06160 EzrA:  Septation ring   32.4 1.2E+02  0.0025   33.8   6.7   68   37-104   269-338 (560)
 32 PF15237 PTRF_SDPR:  PTRF/SDPR   31.6 4.5E+02  0.0098   27.5  10.2  133   39-204    72-220 (246)
 33 PF05873 Mt_ATP-synt_D:  ATP sy  31.1 3.6E+02  0.0077   25.8   8.8  102   37-172    23-124 (161)
 34 cd07595 BAR_RhoGAP_Rich-like T  30.9 1.4E+02   0.003   30.2   6.4   30   40-69     16-45  (244)
 35 PRK09367 histidine ammonia-lya  30.6      83  0.0018   35.0   5.2   66  147-218    24-90  (500)
 36 PRK14145 heat shock protein Gr  30.6      67  0.0014   31.9   4.1   69   21-102    40-108 (196)
 37 PF02074 Peptidase_M32:  Carbox  29.9 2.4E+02  0.0052   31.5   8.5   62   40-104     9-72  (494)
 38 PF11640 TAN:  Telomere-length   29.1      75  0.0016   29.2   3.9   77  162-246    45-121 (155)
 39 PF08549 SWI-SNF_Ssr4:  Fungal   28.3 1.3E+02  0.0028   35.0   6.3   82   91-177   322-410 (669)
 40 PRK05771 V-type ATP synthase s  27.3 1.9E+02   0.004   32.4   7.2   22   40-61     94-115 (646)
 41 PF08656 DASH_Dad3:  DASH compl  27.1 1.4E+02   0.003   26.2   4.9   68   75-183     6-74  (78)
 42 TIGR03832 Tyr_2_3_mutase tyros  27.0      94   0.002   34.7   4.9   57  160-218    28-85  (507)
 43 KOG3598 Thyroid hormone recept  26.2      77  0.0017   40.2   4.3   15  184-198  1842-1856(2220)
 44 PF08580 KAR9:  Yeast cortical   25.8      85  0.0019   36.2   4.4  127   40-181   200-336 (683)
 45 PRK08776 cystathionine gamma-s  25.7 1.9E+02  0.0041   30.5   6.6   92   83-180   307-402 (405)
 46 COG1937 Uncharacterized protei  25.7 1.4E+02  0.0031   26.5   4.8   50   40-95      7-56  (89)
 47 KOG1412 Aspartate aminotransfe  25.7      97  0.0021   33.9   4.5   61   39-107   316-377 (410)
 48 PF10444 Nbl1_Borealin_N:  Nbl1  24.9      77  0.0017   25.4   2.8   43   38-81      9-59  (59)
 49 KOG1655 Protein involved in va  24.9 1.7E+02  0.0037   30.0   5.7   57  152-210    13-75  (218)
 50 TIGR00293 prefoldin, archaeal   24.9 3.7E+02   0.008   23.6   7.3   30  153-182    88-117 (126)
 51 PF10732 DUF2524:  Protein of u  24.8 1.2E+02  0.0026   27.1   4.2   38  162-199     6-43  (84)
 52 cd07606 BAR_SFC_plant The Bin/  24.7 2.1E+02  0.0046   28.3   6.3   29   38-66      7-35  (202)
 53 PF12795 MscS_porin:  Mechanose  24.2 4.4E+02  0.0095   25.8   8.4   65  132-199    15-82  (240)
 54 cd00332 PAL-HAL Phenylalanine   24.0 1.2E+02  0.0026   33.2   5.0   57  160-218    25-82  (444)
 55 PF10239 DUF2465:  Protein of u  24.0 1.9E+02  0.0041   30.5   6.2  152   50-211    32-201 (318)
 56 PF15601 Imm42:  Immunity prote  23.5 2.4E+02  0.0052   26.8   6.2   55   48-109    15-77  (134)
 57 TIGR01225 hutH histidine ammon  23.2 1.2E+02  0.0025   34.0   4.7   56  161-218    31-87  (506)
 58 PF08172 CASP_C:  CASP C termin  22.4 3.6E+02  0.0078   27.6   7.6   28   39-66      6-33  (248)
 59 PF15368 BioT2:  Spermatogenesi  22.3 2.1E+02  0.0046   28.4   5.7   64  114-188    46-109 (170)
 60 PF05546 She9_MDM33:  She9 / Md  21.9 1.5E+02  0.0032   30.2   4.7   42  142-183    23-64  (207)
 61 PF09712 PHA_synth_III_E:  Poly  21.8 1.1E+02  0.0023   31.7   3.8   98   52-170   190-291 (293)
 62 PF04949 Transcrip_act:  Transc  21.5 2.2E+02  0.0047   28.1   5.5   36  146-181    79-114 (159)
 63 PF08397 IMD:  IRSp53/MIM homol  21.4      52  0.0011   31.9   1.5  145   28-183     5-178 (219)
 64 PF00221 Lyase_aromatic:  Aroma  21.4      91   0.002   34.3   3.4   65  148-218    22-87  (473)
 65 PTZ00365 60S ribosomal protein  21.2 9.4E+02    0.02   25.5  10.3  102   83-185   132-237 (266)
 66 PRK14154 heat shock protein Gr  21.2      78  0.0017   31.8   2.7   65   26-103    52-116 (208)
 67 PF07464 ApoLp-III:  Apolipopho  20.8   1E+02  0.0023   29.5   3.3   52  126-179    83-150 (155)
 68 PF05960 DUF885:  Bacterial pro  20.6 4.6E+02  0.0099   28.2   8.3   97   39-147    20-116 (549)
 69 COG4453 Uncharacterized protei  20.6      88  0.0019   28.3   2.6   17  162-178    40-56  (95)
 70 PLN03097 FHY3 Protein FAR-RED   20.5 2.1E+02  0.0046   34.1   6.3   43   73-132   417-459 (846)
 71 KOG2196 Nuclear porin [Nuclear  20.5 2.5E+02  0.0054   29.4   6.1   56   40-102   135-201 (254)
 72 PRK10869 recombination and rep  20.5 3.2E+02   0.007   30.4   7.4   34   73-106   262-295 (553)
 73 PF02050 FliJ:  Flagellar FliJ   20.5 2.1E+02  0.0046   23.2   4.7   29  154-182    55-83  (123)
 74 PRK10280 dipeptidyl carboxypep  20.5   1E+03   0.022   27.6  11.4   20   71-90     52-71  (681)
 75 PF05794 Tcp11:  T-complex prot  20.3 3.1E+02  0.0067   28.8   6.9   35   85-129    51-85  (441)
 76 PF05597 Phasin:  Poly(hydroxya  20.3 1.2E+02  0.0025   28.5   3.4   26  146-171   104-129 (132)
 77 smart00721 BAR BAR domain.      20.3 2.2E+02  0.0048   26.5   5.3   26   41-66     29-54  (239)

No 1  
>KOG3583 consensus Uncharacterized conserved protein [Function unknown]
Probab=99.97  E-value=2.6e-32  Score=264.48  Aligned_cols=174  Identities=21%  Similarity=0.256  Sum_probs=158.4

Q ss_pred             hhhcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhH----HHHhhhhcceEeeeccCCCC
Q 008451           37 QQLNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIV----DEIRKVSKAFVVHPKNVNAE  112 (565)
Q Consensus        37 ~QlNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLv----EEIkPvLr~FvVlPlnVnae  112 (565)
                      ....++-|..|++|+|+.|..+|.+||.+|   +++ ||.+|++|+.|+.++.+|.    +|-+|.||+.||+|..|.-|
T Consensus         8 ~~~~~d~~ikr~~d~k~~i~~llq~ldlq~---~~~-wp~~le~fs~las~ms~l~~~~~k~~~p~lr~~~~~~~~~~~e   83 (279)
T KOG3583|consen    8 IAQATDMMIKRVTDAKKIIEELLQMLDLQE---KCP-WPLMLEKFSTLASFMSSLQSSVRKSGMPHLRSHVLVTQRLQYE   83 (279)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHhhhh---cCc-cHHHHHHHHHHHHHHHHHHHHHHHccCCccccchhhhhhhhcC
Confidence            345678999999999999999999999998   677 9999999999999998887    67778999999999999855


Q ss_pred             ----------------CCccchhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHH
Q 008451          113 ----------------NATILPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLAD  176 (565)
Q Consensus       113 ----------------Na~IVPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~  176 (565)
                                      ||++||||||||++||||.++.+    ++.++.++    ..|++.|+|..|||+|+++.+.|++
T Consensus        84 ~detl~r~TeGRVpvfsH~lVPdyLRTkPdPe~E~~e~q----l~~~aa~~----saDaa~kQI~~yNK~is~ll~~lsk  155 (279)
T KOG3583|consen   84 PDETLQRATEGRVPVFSHALVPDYLRTKPDPEMENEEGQ----LDGEAAAK----SADAAVKQIAAYNKNISGLLNHLSK  155 (279)
T ss_pred             chHHHHHHhcCcccccccccchHhhccCCChhhHHHHhh----hhhHHhhh----hhHHHHHHHHHHHHHHHHHHHHHHH
Confidence                            99999999999999999999999    78899999    9999999999999999999999999


Q ss_pred             HHHhh-hhccCCCCCCCCccChhHHHHHHHHHHHHHHHHhcCCCcccCCCcCCCCCCCc
Q 008451          177 TRKAY-CFGTRQGPQILPTLDKGQALKIQEQENLLRAAVNSGEGLRLPGDQRQMTPALP  234 (565)
Q Consensus       177 aRk~~-e~gtRqGp~~~pT~dkadaaki~eqt~lL~AAVn~GeGLr~p~dqr~~~~~lp  234 (565)
                      .|++| |++.|.+  +.+|.+.+|       |++|||||.+|||||   .||.++.+=|
T Consensus       156 ~~re~tEs~~~~p--iqQT~n~~d-------T~~lVaaV~~GkGl~---~~r~~~~~gP  202 (279)
T KOG3583|consen  156 VDREHTESAIEKP--IQQTYNRDD-------TAKLVAAVLTGKGLR---SQRTMAPAGP  202 (279)
T ss_pred             HHHHHHHhhhcCc--cccccChhH-------HHHHHHHHHhccccc---cccccCCCCC
Confidence            99997 4488888  999999999       999999999999998   4666554433


No 2  
>PF10232 Med8:  Mediator of RNA polymerase II transcription complex subunit 8;  InterPro: IPR019364 The Mediator complex is a coactivator involved in the regulated transcription of nearly all RNA polymerase II-dependent genes. Mediator functions as a bridge to convey information from gene-specific regulatory proteins to the basal RNA polymerase II transcription machinery. The Mediator complex, having a compact conformation in its free form, is recruited to promoters by direct interactions with regulatory proteins and serves for the assembly of a functional preinitiation complex with RNA polymerase II and the general transcription factors. On recruitment the Mediator complex unfolds to an extended conformation and partially surrounds RNA polymerase II, specifically interacting with the unphosphorylated form of the C-terminal domain (CTD) of RNA polymerase II. The Mediator complex dissociates from the RNA polymerase II holoenzyme and stays at the promoter when transcriptional elongation begins.  The Mediator complex is composed of at least 31 subunits: MED1, MED4, MED6, MED7, MED8, MED9, MED10, MED11, MED12, MED13, MED13L, MED14, MED15, MED16, MED17, MED18, MED19, MED20, MED21, MED22, MED23, MED24, MED25, MED26, MED27, MED29, MED30, MED31, CCNC, CDK8 and CDC2L6/CDK11.  The subunits form at least three structurally distinct submodules. The head and the middle modules interact directly with RNA polymerase II, whereas the elongated tail module interacts with gene-specific regulatory proteins. Mediator containing the CDK8 module is less active than Mediator lacking this module in supporting transcriptional activation.   The head module contains: MED6, MED8, MED11, SRB4/MED17, SRB5/MED18, ROX3/MED19, SRB2/MED20 and SRB6/MED22.  The middle module contains: MED1, MED4, NUT1/MED5, MED7, CSE2/MED9, NUT2/MED10, SRB7/MED21 and SOH1/MED31. CSE2/MED9 interacts directly with MED4.  The tail module contains: MED2, PGD1/MED3, RGR1/MED14, GAL11/MED15 and SIN4/MED16.  The CDK8 module contains: MED12, MED13, CCNC and CDK8.   Individual preparations of the Mediator complex lacking one or more distinct subunits have been variously termed ARC, CRSP, DRIP, PC2, SMCC and TRAP.  Arc32, or Med8, is one of the subunits of the Mediator complex of RNA polymerase II. The region conserved contains two alpha helices putatively necessary for binding to other subunits within the core of the Mediator complex. The N terminus of Med8 binds to the essential core Head part of Mediator and the C terminus hinges to Med18 on the non-essential part of the Head that also includes Med20 []. ; GO: 0001104 RNA polymerase II transcription cofactor activity, 0006357 regulation of transcription from RNA polymerase II promoter, 0016592 mediator complex; PDB: 3C0T_B 3RJ1_J 2HZS_I.
Probab=99.96  E-value=1.5e-31  Score=255.39  Aligned_cols=170  Identities=20%  Similarity=0.293  Sum_probs=108.4

Q ss_pred             hcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhHHHHh---hhhcceEeeeccCCCC--C
Q 008451           39 LNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIVDEIR---KVSKAFVVHPKNVNAE--N  113 (565)
Q Consensus        39 lNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLvEEIk---PvLr~FvVlPlnVnae--N  113 (565)
                      .+||+||.|+.+|+++|..|+.+|+..+   +.+.|++||++|+||+++|.+|.+.++   ++|+++||||+.+.|+  .
T Consensus        11 ~aLe~ir~Rl~qL~~SL~~l~~~L~~~~---~lp~W~slq~qf~il~~qL~sL~~~L~~~~~~L~~~vv~P~~~fP~~~~   87 (226)
T PF10232_consen   11 KALEAIRQRLAQLKHSLQSLIDKLEQSQ---PLPPWPSLQDQFAILSSQLSSLSKTLQHNKPLLRNTVVYPLPLFPGRDE   87 (226)
T ss_dssp             TTTSTTTHHHHHHHHHHHHHHHHHT-T----SS---HHHHHHHHHHHHHHHHHHHHTTTSSTTTTTS--TTS---S--TT
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHccC---CCCCcHHHHHHHHHHHHHHHHHHHHHHhccccccceeeecCCCCCCcCh
Confidence            6899999999999999999999999877   899999999999999999999995554   8999999999999876  6


Q ss_pred             CccchhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHH-----HHHHHHHHHhHHHHHHHHHHhhhh--ccC
Q 008451          114 ATILPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSR-----IDMIGAACESAEKVLADTRKAYCF--GTR  186 (565)
Q Consensus       114 a~IVPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQ-----Id~iNkacE~aekvIa~aRk~~e~--gtR  186 (565)
                      .++||+||||||+||||..+.+.    ...+.++    ..+...++     |..+|+.|++++++|.+.|++|+.  ..+
T Consensus        88 e~ll~~lLRtKl~PeVE~~~~~~----~~~~~~~----~~~~~~~q~~~~~i~~~~k~~~~~~~~~t~lree~e~~~~~~  159 (226)
T PF10232_consen   88 EDLLPDLLRTKLDPEVEEWEAQL----REEAANA----TPDAAQKQQAELAIAQYNKRISNWLDVVTGLREEWEFEDFSD  159 (226)
T ss_dssp             GGGTTHHHHH----GGGTTTSTT----T---TTS-----S--SSS-SSTTTHHHHHHHHHHH--TTTTTHHH---HTTT-
T ss_pred             HHHHHHHHhCCCCChHHHHHHHH----HHHHHhc----CcCHHHhhhhhhHHHHHHHHHHHHHHHHHHHHHHHhhhhccc
Confidence            68999999999999999999994    4445555    55666677     999999999999999999999998  555


Q ss_pred             CCCCCCCccChhHHHHHHHHHHHHHHHHhcCCCcccCCCcCC
Q 008451          187 QGPQILPTLDKGQALKIQEQENLLRAAVNSGEGLRLPGDQRQ  228 (565)
Q Consensus       187 qGp~~~pT~dkadaaki~eqt~lL~AAVn~GeGLr~p~dqr~  228 (565)
                      .+  ..+|++.+|       ++.|++||.+|+||+-..++..
T Consensus       160 ~~--~~~~~~~~d-------t~~lv~~~~~g~gl~~~~~~~~  192 (226)
T PF10232_consen  160 RE--EEQTSEEED-------TEALVAAVGFGKGLKPQFEQQI  192 (226)
T ss_dssp             --------------------HHHHHH------SS-HH-----
T ss_pred             cc--cccccChhH-------HHHHHHHhhcchhcccccchhh
Confidence            55  889999999       9999999999999995555443


No 3  
>PF12128 DUF3584:  Protein of unknown function (DUF3584);  InterPro: IPR021979  This family consist of uncharacterised bacterial proteins. 
Probab=79.51  E-value=8  Score=46.01  Aligned_cols=132  Identities=14%  Similarity=0.221  Sum_probs=80.1

Q ss_pred             hcHHHHHHHHHHHHHHHHHHHHhhhhccc------cCCCCChhhhhhhhhhcchhhhhhH-HHHhhhhcce----Eeeec
Q 008451           39 LNLEAVKTRAISLFKAISRILEDFDAYAR------TNTTPKWQDILGQYSMVNLELFNIV-DEIRKVSKAF----VVHPK  107 (565)
Q Consensus        39 lNLEAVraRA~DLkkaIsriI~~LE~e~~------tN~t~kWpDVLdqFSVIS~QL~nLv-EEIkPvLr~F----vVlPl  107 (565)
                      .-|...+.++.+++..|..+-.-|+....      -...+.|.+-||+  ||+-+|  |. .|+.|.+..-    .+|-+
T Consensus       511 ~~l~~~~~~~~~~~~~~~~l~~~L~p~~gSL~~fL~~~~p~We~tIGK--Vid~eL--L~r~dL~P~l~~~~~~dslyGl  586 (1201)
T PF12128_consen  511 EELRQARRELEELRAQIAELQRQLDPQKGSLLEFLRKNKPGWEQTIGK--VIDEEL--LYRTDLEPQLVEDSGSDSLYGL  586 (1201)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHhhCCCCCcHHHHHHhCCCcHHHHhHh--hCCHHH--hcCCCCCCeecCCCccccccee
Confidence            34566777777777777777666653322      1258999999986  888775  32 6777755422    46666


Q ss_pred             cCCCCCCccchhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhh
Q 008451          108 NVNAENATILPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCF  183 (565)
Q Consensus       108 nVnaeNa~IVPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~  183 (565)
                      .++=++= .+|+|..++-  ++|.+...+.+++...      ....+.+.+++..+++.++.+.+-+.+++-+++.
T Consensus       587 ~LdL~~I-~~pd~~~~ee--~L~~~l~~~~~~l~~~------~~~~~~~e~~l~~~~~~~~~~~~~~~~~~~~~~~  653 (1201)
T PF12128_consen  587 SLDLSAI-DVPDYAASEE--ELRERLEQAEDQLQSA------EERQEELEKQLKQINKKIEELKREITQAEQELKQ  653 (1201)
T ss_pred             Eeehhhc-CCchhhcChH--HHHHHHHHHHHHHHHH------HHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHh
Confidence            6654432 4677776644  4444433333332221      2245667777777888877777777666655544


No 4  
>KOG2235 consensus Uncharacterized conserved protein [Function unknown]
Probab=75.40  E-value=9.7  Score=43.76  Aligned_cols=184  Identities=17%  Similarity=0.182  Sum_probs=114.7

Q ss_pred             HHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhHHHHhhhhcceEeeeccCCCCCCccchhh
Q 008451           41 LEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIVDEIRKVSKAFVVHPKNVNAENATILPVM  120 (565)
Q Consensus        41 LEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLvEEIkPvLr~FvVlPlnVnaeNa~IVPdm  120 (565)
                      +.+|..|...|+..|.-+...+.+.++   .  -...|.+|=+=+     +-.||...+.+|+---.+.+-+|| .+-.-
T Consensus       536 i~aiqdk~~~ly~nirlyEkalklF~d---d--tq~~L~k~LLkt-----v~neI~n~l~nyvase~~~tvdn~-~L~s~  604 (776)
T KOG2235|consen  536 ISAIQDKCRQLYDNIRLYEKALKLFAD---D--TQSDLRKYLLKT-----VGNEIANALLNYVASEESKTVDNH-QLKSK  604 (776)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHhccC---c--hHHHHHHHHHHH-----HHHHHHHHHHHHHHHHhhhhhhhh-hccHH
Confidence            445666666666666666666666552   1  455666665433     558999999999999899888984 78888


Q ss_pred             hhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhh-hhccCCCCCCCCccChhH
Q 008451          121 LSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAY-CFGTRQGPQILPTLDKGQ  199 (565)
Q Consensus       121 LRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~-e~gtRqGp~~~pT~dkad  199 (565)
                      -|+|+.-+.+..-+.++-.+....++-    .+|+++.-++-..++|+-+.|.+.+-++.- ...-|..- ..+-++-.|
T Consensus       605 qR~kla~nl~~~lr~all~l~~aLn~k----siDdF~~a~~saaea~sl~lKKvDKK~er~ll~~~rk~L-~eQl~~~~e  679 (776)
T KOG2235|consen  605 QREKLAENLPEMLRDALLSLFAALNSK----SIDDFHDAVYSAAEACSLALKKVDKKGERELLAKHRKEL-HEQLCSQTE  679 (776)
T ss_pred             HHHHHHHhhhHHHHHHHHHHHHHhccc----chHHHHHHHHHHHHhhhHHHHHhhhHHHHHHHHHHHHHH-HHHHhcccc
Confidence            899998888777777655555554444    888888888888899997776654433321 11222210 111111122


Q ss_pred             HHHHHHHHHHHHHHHhcCCCcccCCCcCCCCCCCchhhhhcccc
Q 008451          200 ALKIQEQENLLRAAVNSGEGLRLPGDQRQMTPALPMHLVDLLPV  243 (565)
Q Consensus       200 aaki~eqt~lL~AAVn~GeGLr~p~dqr~~~~~lp~hl~~~l~~  243 (565)
                      -|-|.-=.-+|.=+--.|+-|.-||.   .-+++=.||-|-|+-
T Consensus       680 PallL~l~vllLf~ki~~s~lhA~Gk---~Vsaiiahik~kl~E  720 (776)
T KOG2235|consen  680 PALLLHLSVLLLFAKITNSPLHASGK---FVSAIIAHIKDKLPE  720 (776)
T ss_pred             hHHHHHHHHHHHHHHHcCCcccCccc---hHHHHHHHHHhhCCh
Confidence            22233334455556667888887774   224455677666654


No 5  
>PF15011 CK2S:  Casein Kinase 2 substrate
Probab=74.35  E-value=13  Score=35.34  Aligned_cols=30  Identities=17%  Similarity=0.317  Sum_probs=24.8

Q ss_pred             HHHHHHHHHHHHHHHHhHHHHHHHHHHhhh
Q 008451          153 IEKLKSRIDMIGAACESAEKVLADTRKAYC  182 (565)
Q Consensus       153 iEklqKQId~iNkacE~aekvIa~aRk~~e  182 (565)
                      ..++.+.++.+++|++.+++...+...-|+
T Consensus        73 l~~L~e~l~~l~~v~~~l~~~~~~~~~l~~  102 (168)
T PF15011_consen   73 LAKLRETLEELQKVRDSLSRQVRDVFQLYE  102 (168)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            455678888889999888888888888888


No 6  
>KOG2129 consensus Uncharacterized conserved protein H4 [Function unknown]
Probab=61.26  E-value=2.9e+02  Score=31.31  Aligned_cols=104  Identities=22%  Similarity=0.418  Sum_probs=54.6

Q ss_pred             cHHHHhhhcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhHHHHhhhhcceEeeeccCCC
Q 008451           32 NQAVVQQLNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIVDEIRKVSKAFVVHPKNVNA  111 (565)
Q Consensus        32 n~~V~~QlNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLvEEIkPvLr~FvVlPlnVna  111 (565)
                      |.-...|.+||.+|--+.+|.+++..=-+.|-.       .=|-.    ..       -|-.|++-+-+.+ =.|-    
T Consensus       172 n~t~~kq~~leQLRre~V~lentlEQEqEalvN-------~LwKr----md-------kLe~ekr~Lq~Kl-Dqpv----  228 (552)
T KOG2129|consen  172 NKTLLKQNTLEQLRREAVQLENTLEQEQEALVN-------SLWKR----MD-------KLEQEKRYLQKKL-DQPV----  228 (552)
T ss_pred             hhhHHhhhhHHHHHHHHHHHhhHHHHHHHHHHH-------HHHHH----HH-------HHHHHHHHHHHHh-cCcc----
Confidence            666778888888888888888877432221111       11211    11       1222333222222 1111    


Q ss_pred             CCCccchhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHH
Q 008451          112 ENATILPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKV  173 (565)
Q Consensus       112 eNa~IVPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekv  173 (565)
                       +...+|--. +|. |.|+.++.+.        ..+    .|++|+..|+-|.+-|..|++-
T Consensus       229 -s~p~~prdi-a~~-~~~~gD~a~~--------~~~----hi~~l~~EveRlrt~l~~Aqk~  275 (552)
T KOG2129|consen  229 -STPSLPRDI-AKI-PDVHGDEAAA--------EKL----HIDKLQAEVERLRTYLSRAQKS  275 (552)
T ss_pred             -cCCCchhhh-hcC-ccccCchHHH--------HHH----HHHHHHHHHHHHHHHHHHHHHH
Confidence             111111111 455 7888887772        122    7888888888888888776653


No 7  
>PF05983 Med7:  MED7 protein;  InterPro: IPR009244 The Mediator complex is a coactivator involved in the regulated transcription of nearly all RNA polymerase II-dependent genes. Mediator functions as a bridge to convey information from gene-specific regulatory proteins to the basal RNA polymerase II transcription machinery. The Mediator complex, having a compact conformation in its free form, is recruited to promoters by direct interactions with regulatory proteins and serves for the assembly of a functional preinitiation complex with RNA polymerase II and the general transcription factors. On recruitment the Mediator complex unfolds to an extended conformation and partially surrounds RNA polymerase II, specifically interacting with the unphosphorylated form of the C-terminal domain (CTD) of RNA polymerase II. The Mediator complex dissociates from the RNA polymerase II holoenzyme and stays at the promoter when transcriptional elongation begins.  The Mediator complex is composed of at least 31 subunits: MED1, MED4, MED6, MED7, MED8, MED9, MED10, MED11, MED12, MED13, MED13L, MED14, MED15, MED16, MED17, MED18, MED19, MED20, MED21, MED22, MED23, MED24, MED25, MED26, MED27, MED29, MED30, MED31, CCNC, CDK8 and CDC2L6/CDK11.  The subunits form at least three structurally distinct submodules. The head and the middle modules interact directly with RNA polymerase II, whereas the elongated tail module interacts with gene-specific regulatory proteins. Mediator containing the CDK8 module is less active than Mediator lacking this module in supporting transcriptional activation.   The head module contains: MED6, MED8, MED11, SRB4/MED17, SRB5/MED18, ROX3/MED19, SRB2/MED20 and SRB6/MED22.  The middle module contains: MED1, MED4, NUT1/MED5, MED7, CSE2/MED9, NUT2/MED10, SRB7/MED21 and SOH1/MED31. CSE2/MED9 interacts directly with MED4.  The tail module contains: MED2, PGD1/MED3, RGR1/MED14, GAL11/MED15 and SIN4/MED16.  The CDK8 module contains: MED12, MED13, CCNC and CDK8.   Individual preparations of the Mediator complex lacking one or more distinct subunits have been variously termed ARC, CRSP, DRIP, PC2, SMCC and TRAP. This family consists of several eukaryotic proteins, which are homologues of the yeast MED7 protein. Activation of gene transcription in metazoans is a multistep process that is triggered by factors that recognise transcriptional enhancer sites in DNA. These factors work with co-activators such as MED7 to direct transcriptional initiation by the RNA polymerase II apparatus [].; GO: 0001104 RNA polymerase II transcription cofactor activity, 0006357 regulation of transcription from RNA polymerase II promoter, 0016592 mediator complex; PDB: 3FBI_C 3FBN_A 1YKH_A 1YKE_A.
Probab=57.23  E-value=90  Score=29.71  Aligned_cols=89  Identities=22%  Similarity=0.387  Sum_probs=55.5

Q ss_pred             HHHHHHHHHHHHHHHHhhhhccc--cCCCCChhhhhhhhhhcchhhhhhHHHHhhhhcceEeeeccCCCCCCccchhhhh
Q 008451           45 KTRAISLFKAISRILEDFDAYAR--TNTTPKWQDILGQYSMVNLELFNIVDEIRKVSKAFVVHPKNVNAENATILPVMLS  122 (565)
Q Consensus        45 raRA~DLkkaIsriI~~LE~e~~--tN~t~kWpDVLdqFSVIS~QL~nLvEEIkPvLr~FvVlPlnVnaeNa~IVPdmLR  122 (565)
                      ..|..+||+....++..+-..-+  ...-..|..-++...+|-..+..|..|.+|.                        
T Consensus        71 ~d~~~eLkkL~~sll~nfleLl~~l~~~P~~~~~ki~~i~~L~~NmhhllNeyRPh------------------------  126 (162)
T PF05983_consen   71 VDRKKELKKLNKSLLLNFLELLDILSKNPSQYERKIEDIRLLFINMHHLLNEYRPH------------------------  126 (162)
T ss_dssp             HHHHHHHHHHHHHHHHHHHHHTTSS---CCCHHHHHHHHHHHHHHHHHHHHHTHHH------------------------
T ss_pred             chHHHHHHHHHHHHHHHHHHHHHHHHhCCccHHHHHHHHHHHHHHHHHHHHHhCHH------------------------
Confidence            66777777777766655433322  2233466777777777777777777776652                        


Q ss_pred             cccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHH
Q 008451          123 SKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVL  174 (565)
Q Consensus       123 TKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvI  174 (565)
                                  +.|+.+..-|     ..++|.-+..|+.|.++|+.|+++|
T Consensus       127 ------------QARetLi~~m-----e~Ql~~kr~~i~~i~~~~~~~~~~l  161 (162)
T PF05983_consen  127 ------------QARETLIMMM-----EEQLEEKREEIEEIRKVCEKAREVL  161 (162)
T ss_dssp             ------------HHHHHHHHHH-----HHHHHHHHHHHHHHHHHHHHHHHHH
T ss_pred             ------------HHHHHHHHHH-----HHHHHHHHHHHHHHHHHHHHHHHHh
Confidence                        2222222221     2267777888999999999998887


No 8  
>PF06248 Zw10:  Centromere/kinetochore Zw10;  InterPro: IPR009361 Zeste white 10 (ZW10) was initially identified as a mitotic checkpoint protein involved in chromosome segregation, and then implicated in targeting cytoplasmic dynein and dynactin to mitotic kinetochores, but it is also important in non-dividing cells. These include cytoplasmic dynein targeting to Golgi and other membranes, and SNARE-mediated ER-Golgi trafficking [, ]. Dominant-negative ZW10, anti-ZW10 antibody, and ZW10 RNA interference (RNAi) cause Golgi dispersal. ZW10 RNAi also disperse endosomes and lysosomes []. Drosophila kinetochore components Rough deal (Rod) and Zw10 are required for the proper functioning of the metaphase checkpoint in flies []. The eukaryotic spindle assembly checkpoint (SAC) monitors microtubule attachment to kinetochores and prevents anaphase onset until all kinetochores are aligned on the metaphase plate. It is an essential surveillance mechanism that ensures high fidelity chromosome segregation during mitosis. In higher eukaryotes, cytoplasmic dynein is involved in silencing the SAC by removing the checkpoint proteins Mad2 and the Rod-Zw10-Zwilch complex (RZZ) from aligned kinetochores [, , ].; GO: 0007067 mitosis, 0000775 chromosome, centromeric region, 0005634 nucleus
Probab=55.48  E-value=1.5e+02  Score=32.76  Aligned_cols=127  Identities=16%  Similarity=0.231  Sum_probs=71.2

Q ss_pred             cchhhhhcHHHHhhhcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhh---hhcchhhhhhHHHHhhhhcc
Q 008451           25 PVAVERLNQAVVQQLNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQY---SMVNLELFNIVDEIRKVSKA  101 (565)
Q Consensus        25 ~~~~e~ln~~V~~QlNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqF---SVIS~QL~nLvEEIkPvLr~  101 (565)
                      |...|-|+..+.      .|..|++++|..|..+|.+           +|.||+..+   .-+-..+..|.+||..+++.
T Consensus         6 ~l~~edl~~~I~------~L~~~i~~~k~eV~~~I~~-----------~y~df~~~~~~~~~L~~~~~~l~~eI~d~l~~   68 (593)
T PF06248_consen    6 PLSKEDLRKSIS------RLSRRIEELKEEVHSMINK-----------KYSDFSPSLQSAKDLIERSKSLAREINDLLQS   68 (593)
T ss_pred             CCCHhHHHHHHH------HHHHHHHHHHHHHHHHHHH-----------HHHHHHHHHHhHHHHHHHHHHHHHHHHHHHHh
Confidence            344444444444      7777777777777766653           334444333   33445566777888777765


Q ss_pred             eEeeeccCCCCCCccchhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHH-----hHHHHHHH
Q 008451          102 FVVHPKNVNAENATILPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACE-----SAEKVLAD  176 (565)
Q Consensus       102 FvVlPlnVnaeNa~IVPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE-----~aekvIa~  176 (565)
                      -+        ++. +.+..-      +...+...++.++.....-+-+-..+.++.++|+.++.+++     .|-+.+.+
T Consensus        69 ~~--------~~~-i~~~l~------~a~~e~~~L~~eL~~~~~~l~~L~~L~~i~~~l~~~~~al~~~~~~~Aa~~L~~  133 (593)
T PF06248_consen   69 EI--------ENE-IQPQLR------DAAEELQELKRELEENEQLLEVLEQLQEIDELLEEVEEALKEGNYLDAADLLEE  133 (593)
T ss_pred             hc--------cch-hHHHHH------HHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHhcCCHHHHHHHHHH
Confidence            22        111 222222      23444555555555554444444566677777777776643     45566666


Q ss_pred             HHHhhhh
Q 008451          177 TRKAYCF  183 (565)
Q Consensus       177 aRk~~e~  183 (565)
                      +++....
T Consensus       134 ~~~~L~~  140 (593)
T PF06248_consen  134 LKSLLDD  140 (593)
T ss_pred             HHHHHHh
Confidence            6666655


No 9  
>PF04253 TFR_dimer:  Transferrin receptor-like dimerisation domain;  InterPro: IPR007365 This entry represents the dimerisation domain found in the transferrin receptor, as well as in a number of other proteins including glutamate carboxypeptidase II and N-acetylated-alpha-linked acidic dipeptidase like protein. The transferrin receptor (TfR) assists iron uptake into vertebrate cells through a cycle of endo- and exocytosis of the iron transport protein transferrin (Tf). TfR binds iron-loaded (diferric) Tf at the cell surface and carries it to the endosome, where the iron dissociates from Tf. The apo-Tf remains bound to TfR until it reaches the cell surface, where apo-Tf is replaced by diferric Tf from the serum to begin the cycle again. Human TfR is a homodimeric type II transmembrane protein. The crystal structure of a TfR monomer reveals a 3-domain structure: a protease-like domain that closely resembles carboxy- and amino-peptidases; an apical domain consisting of a beta-sandwich; and a helical dimerisation domain. The dimerisation domain consists of a 4-helical bundle that makes contact with each of the three domains in the dimer partner [].; PDB: 3FF3_A 3FEC_A 3FED_A 3FEE_A 3BXM_A 2C6P_A 1Z8L_C 3SJF_A 3BHX_A 2C6G_A ....
Probab=55.27  E-value=24  Score=30.98  Aligned_cols=60  Identities=18%  Similarity=0.222  Sum_probs=38.3

Q ss_pred             hhhhcceEeeeccCCCCCCccchhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHH
Q 008451           96 RKVSKAFVVHPKNVNAENATILPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLA  175 (565)
Q Consensus        96 kPvLr~FvVlPlnVnaeNa~IVPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa  175 (565)
                      +|-.|+.++-|-.-+....         ...|.+       +|.+..+-.+.    .++.++++|+.+..++++|-++|+
T Consensus        64 r~~~kHvifap~~~~~y~~---------~~fPgI-------~dai~~~~~~~----~~~~~~~~i~~v~~~i~~Aa~~L~  123 (125)
T PF04253_consen   64 RPWYKHVIFAPGRWNGYAS---------WTFPGI-------RDAIEDKDSSK----DWEEAQKQISRVAKAIQNAANTLS  123 (125)
T ss_dssp             BTT--BSSEEEETTEEEEE---------EESHHH-------HHHHTTGGGTS----THHHHHHHHHHHHHHHHHHHHHCS
T ss_pred             CcccceeeeCCCCCCCCcC---------cccHHH-------HHHHHhcccCc----hHHHHHHHHHHHHHHHHHHHHHhc
Confidence            5566666666665543322         445554       34455553333    399999999999999999988764


No 10 
>TIGR03017 EpsF chain length determinant protein EpsF. Sequences in this family of proteins are members of the chain length determinant family (pfam02706) which includes the wzc protein from E.coli. This family of proteins are homologous to the EpsF protein of the methanolan biosynthesis operon of Methylobacillus species strain 12S. The distribution of this protein appears to be restricted to a subset of exopolysaccharide operons containing a syntenic grouping of genes including a variant of the EpsH exosortase protein. Exosortase has been proposed to be involved in the targetting and processing of proteins containing the PEP-CTERM domain to the exopolysaccharide layer.
Probab=54.83  E-value=1e+02  Score=31.92  Aligned_cols=136  Identities=12%  Similarity=0.140  Sum_probs=68.4

Q ss_pred             cHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhh----hhhhhhhcchhhhhhHHHHhhhhcceEeeeccCCCCCCc
Q 008451           40 NLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQD----ILGQYSMVNLELFNIVDEIRKVSKAFVVHPKNVNAENAT  115 (565)
Q Consensus        40 NLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpD----VLdqFSVIS~QL~nLvEEIkPvLr~FvVlPlnVnaeNa~  115 (565)
                      .++-+..|+..+++.+.+...+|+.+-+-|....+.+    ...+..-++.+|..+..+.......+   -   ..+..+
T Consensus       172 ~~~fl~~ql~~~~~~l~~ae~~l~~fr~~~~i~~~~~~~~~~~~~l~~l~~~l~~~~~~~~~~~~~~---~---~~~~~~  245 (444)
T TIGR03017       172 AALWFVQQIAALREDLARAQSKLSAYQQEKGIVSSDERLDVERARLNELSAQLVAAQAQVMDASSKE---G---GSSGKD  245 (444)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHHHHcCCcccCcccchHHHHHHHHHHHHHHHHHHHHHHHHHH---h---ccCCcc
Confidence            4566777788888888888777777765455444432    12333344444444443222111100   0   112234


Q ss_pred             cchhhhhcccCcchhhhhhHHHHHHHhhcCC-CCCchhHHHHHHHHHHHHHHHH-hHHHHHHHHHHhh
Q 008451          116 ILPVMLSSKLLPEMEIDDNSKREQLLLGMQN-LPIPSQIEKLKSRIDMIGAACE-SAEKVLADTRKAY  181 (565)
Q Consensus       116 IVPdmLRTKLlPEmEtee~q~~~ql~~kA~n-LP~~~qiEklqKQId~iNkacE-~aekvIa~aRk~~  181 (565)
                      .+|........-++..+..++..++..-... .|-+..+-.++++|+.+.+.+. .+.+++...++++
T Consensus       246 ~~~~~~~~~~i~~l~~~l~~le~~l~~l~~~y~~~hP~v~~l~~~i~~l~~~l~~e~~~~~~~~~~~~  313 (444)
T TIGR03017       246 ALPEVIANPIIQNLKTDIARAESKLAELSQRLGPNHPQYKRAQAEINSLKSQLNAEIKKVTSSVGTNS  313 (444)
T ss_pred             cchhhhcChHHHHHHHHHHHHHHHHHHHHHHhCCCCcHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            4565544433334444444443443333332 2667777778888887777654 2333433333333


No 11 
>TIGR00634 recN DNA repair protein RecN. All proteins in this family for which functions are known are ATP binding proteins involved in the initiation of recombination and recombinational repair.
Probab=49.54  E-value=39  Score=37.02  Aligned_cols=31  Identities=19%  Similarity=0.303  Sum_probs=16.6

Q ss_pred             hhHHHHHHHHHHHHHHHHhHHHHHHHHHHhh
Q 008451          151 SQIEKLKSRIDMIGAACESAEKVLADTRKAY  181 (565)
Q Consensus       151 ~qiEklqKQId~iNkacE~aekvIa~aRk~~  181 (565)
                      ..++.+.++++.+.+.++.+-+.|++.|+.+
T Consensus       346 ~~le~L~~el~~l~~~l~~~a~~Ls~~R~~~  376 (563)
T TIGR00634       346 ESLEALEEEVDKLEEELDKAAVALSLIRRKA  376 (563)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            3455555555555555555555555555543


No 12 
>PF11336 DUF3138:  Protein of unknown function (DUF3138);  InterPro: IPR021485  This family of proteins with unknown function appear to be restricted to Proteobacteria. 
Probab=48.99  E-value=32  Score=38.35  Aligned_cols=75  Identities=24%  Similarity=0.327  Sum_probs=47.9

Q ss_pred             chhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhhccCCCCC-----------CCCccChhHH-------HHHHHHHHHHH
Q 008451          150 PSQIEKLKSRIDMIGAACESAEKVLADTRKAYCFGTRQGPQ-----------ILPTLDKGQA-------LKIQEQENLLR  211 (565)
Q Consensus       150 ~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~gtRqGp~-----------~~pT~dkada-------aki~eqt~lL~  211 (565)
                      ..+||.|++++..+.+-|..+++.|+..-.+-..|+..+|.           ..+++..+|+       |.++-.+..|.
T Consensus        24 a~~i~~L~~ql~aLq~~v~eL~~~laa~~~aa~~gA~~~~~~~a~~~aP~~~a~~~~T~d~~~~~~qqiAn~~lKv~~l~  103 (514)
T PF11336_consen   24 ADQIKALQAQLQALQDQVNELRAKLAAKPAAAPGGAAIGPAATAAAAAPSSDAQAGLTNDDATEMRQQIANAQLKVESLE  103 (514)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHhcCCCCCCccccccccccccccCCCcccccccChHHHHHHHHHHHhhhhhHHHHh
Confidence            34788889999888888888887776544433333322221           2334555555       33444567788


Q ss_pred             HHHhcC--CCcccCC
Q 008451          212 AAVNSG--EGLRLPG  224 (565)
Q Consensus       212 AAVn~G--eGLr~p~  224 (565)
                      .|...|  |||+|.+
T Consensus       104 da~~t~~~kGLsITG  118 (514)
T PF11336_consen  104 DAAETGGFKGLSITG  118 (514)
T ss_pred             hHHhcCCcccceEee
Confidence            888777  7999876


No 13 
>PRK09039 hypothetical protein; Validated
Probab=48.46  E-value=1e+02  Score=32.36  Aligned_cols=68  Identities=29%  Similarity=0.444  Sum_probs=42.3

Q ss_pred             chhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhh-ccCCCCCCCCccChhHHHHHHHHHH-------HHHHHHhcCCCcc
Q 008451          150 PSQIEKLKSRIDMIGAACESAEKVLADTRKAYCF-GTRQGPQILPTLDKGQALKIQEQEN-------LLRAAVNSGEGLR  221 (565)
Q Consensus       150 ~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~-gtRqGp~~~pT~dkadaaki~eqt~-------lL~AAVn~GeGLr  221 (565)
                      ..+|++|+++++.+..+++.+++-.++.++..+. +.+        +..+=|.|+.|=+.       -|++..+.-+|++
T Consensus       143 ~~qI~aLr~Qla~le~~L~~ae~~~~~~~~~i~~L~~~--------L~~a~~~~~~~l~~~~~~~~~~l~~~~~~~~~ir  214 (343)
T PRK09039        143 NQQIAALRRQLAALEAALDASEKRDRESQAKIADLGRR--------LNVALAQRVQELNRYRSEFFGRLREILGDREGIR  214 (343)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHH--------HHHHHHHHHHHHHHhHHHHHHHHHHHhCCCCCcE
Confidence            4467777777777777777777777766666655 332        33444444444333       3567777777887


Q ss_pred             cCCC
Q 008451          222 LPGD  225 (565)
Q Consensus       222 ~p~d  225 (565)
                      |-+|
T Consensus       215 i~g~  218 (343)
T PRK09039        215 IVGD  218 (343)
T ss_pred             EECC
Confidence            7655


No 14 
>KOG3091 consensus Nuclear pore complex, p54 component (sc Nup57) [Nuclear structure; Intracellular trafficking, secretion, and vesicular transport]
Probab=47.03  E-value=91  Score=35.19  Aligned_cols=27  Identities=22%  Similarity=0.352  Sum_probs=21.5

Q ss_pred             cHHHHHHHHHHHHHHHHHHHHhhhhcc
Q 008451           40 NLEAVKTRAISLFKAISRILEDFDAYA   66 (565)
Q Consensus        40 NLEAVraRA~DLkkaIsriI~~LE~e~   66 (565)
                      -++-.|.|..+|-+-|.||+-+.|...
T Consensus       377 KI~~~k~r~~~Ls~RiLRv~ikqeilr  403 (508)
T KOG3091|consen  377 KIEEAKNRHVELSHRILRVMIKQEILR  403 (508)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHh
Confidence            356778888999999999988877654


No 15 
>TIGR01834 PHA_synth_III_E poly(R)-hydroxyalkanoic acid synthase, class III, PhaE subunit. This model represents the PhaE subunit of the heterodimeric class (class III) of polymerase for poly(R)-hydroxyalkanoic acids (PHAs), carbon and energy storage polymers of many bacteria. The most common PHA is polyhydroxybutyrate but about 150 different constituent hydroxyalkanoic acids (HAs) have been identified in various species. This model must be designated subfamily to indicate the heterogeneity of PHAs.
Probab=45.65  E-value=54  Score=34.84  Aligned_cols=107  Identities=26%  Similarity=0.367  Sum_probs=66.0

Q ss_pred             HHHHHHHHHhhhhccccCCCCC-hhhhhhhhhhcchhhhhhH---HHHhhhhcceEeeeccCCCCCCccchhhhhcccCc
Q 008451           52 FKAISRILEDFDAYARTNTTPK-WQDILGQYSMVNLELFNIV---DEIRKVSKAFVVHPKNVNAENATILPVMLSSKLLP  127 (565)
Q Consensus        52 kkaIsriI~~LE~e~~tN~t~k-WpDVLdqFSVIS~QL~nLv---EEIkPvLr~FvVlPlnVnaeNa~IVPdmLRTKLlP  127 (565)
                      .+++.++..+|....+-.+.++ |.++.|.+.-+..+.+.-+   +|..++...+      ||+-      .-|+..+..
T Consensus       207 ~ks~e~~~~~l~~~~~~g~~v~s~re~~d~W~~~ae~~~~e~~~S~efak~~G~l------vna~------m~lr~~~qe  274 (320)
T TIGR01834       207 YKSFAALMSDLLARAKSGKPVKTAKALYDLWVIAAEEAYAEVFASEENAKVHGKF------INAL------MRLRIQQQE  274 (320)
T ss_pred             HHHHHHHHHHHHhccccCCCchhHHHHHHHHHHHHHHHHHHHHcCHHHHHHHHHH------HHHH------HHHHHHHHH
Confidence            4677777777777554334455 9999998876554443333   4444433322      1111      112222222


Q ss_pred             chhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHH
Q 008451          128 EMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRK  179 (565)
Q Consensus       128 EmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk  179 (565)
                      .|        |. .-+..|||--..+|.+.+||..|.+-+-.++|-|.+..+
T Consensus       275 ~~--------e~-~L~~LnlPTRsElDe~~krL~ELrR~vr~L~k~l~~l~~  317 (320)
T TIGR01834       275 IV--------EA-LLKMLNLPTRSELDEAHQRIQQLRREVKSLKKRLGDLEA  317 (320)
T ss_pred             HH--------HH-HHHhCCCCCHHHHHHHHHHHHHHHHHHHHHHHHHHHhhh
Confidence            22        22 223468999999999999999999999988888876554


No 16 
>PF10146 zf-C4H2:  Zinc finger-containing protein ;  InterPro: IPR018482 Zinc finger (Znf) domains are relatively small protein motifs which contain multiple finger-like protrusions that make tandem contacts with their target molecule. Some of these domains bind zinc, but many do not; instead binding other metals such as iron, or no metal at all. For example, some family members form salt bridges to stabilise the finger-like folds. They were first identified as a DNA-binding motif in transcription factor TFIIIA from Xenopus laevis (African clawed frog), however they are now recognised to bind DNA, RNA, protein and/or lipid substrates [, , , , ]. Their binding properties depend on the amino acid sequence of the finger domains and of the linker between fingers, as well as on the higher-order structures and the number of fingers. Znf domains are often found in clusters, where fingers can have different binding specificities. There are many superfamilies of Znf motifs, varying in both sequence and structure. They display considerable versatility in binding modes, even between members of the same class (e.g. some bind DNA, others protein), suggesting that Znf motifs are stable scaffolds that have evolved specialised functions. For example, Znf-containing proteins function in gene transcription, translation, mRNA trafficking, cytoskeleton organisation, epithelial development, cell adhesion, protein folding, chromatin remodelling and zinc sensing, to name but a few []. Zinc-binding motifs are stable structures, and they rarely undergo conformational changes upon binding their target.  This entry represents a family of proteins which appears to have a highly conserved zinc finger domain at the C-terminal end, described as -C-X2-CH-X3-H-X5-C-X2-C-. The structure is predicted to contain a coiled coil. Members of this family are annotated as being tumour-associated antigen HCA127 in humans, but this could not be confirmed.
Probab=44.48  E-value=3.6e+02  Score=27.40  Aligned_cols=27  Identities=15%  Similarity=0.433  Sum_probs=22.6

Q ss_pred             cHHHHHHHHHHHHHHHHHHHHhhhhcc
Q 008451           40 NLEAVKTRAISLFKAISRILEDFDAYA   66 (565)
Q Consensus        40 NLEAVraRA~DLkkaIsriI~~LE~e~   66 (565)
                      +|..||.++.+|.+.-.+|+..++..-
T Consensus         2 ~i~~ir~K~~~lek~k~~i~~e~~~~e   28 (230)
T PF10146_consen    2 KIKEIRNKTLELEKLKNEILQEVESLE   28 (230)
T ss_pred             cHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            466899999999999999988887754


No 17 
>PF08580 KAR9:  Yeast cortical protein KAR9;  InterPro: IPR013889  The KAR9 protein in Saccharomyces cerevisiae (Baker's yeast) is a cytoskeletal protein required for karyogamy, correct positioning of the mitotic spindle and for orientation of cytoplasmic microtubules []. KAR9 localises at the shmoo tip in mating cells and at the tip of the growing bud in anaphase []. 
Probab=44.28  E-value=1.2e+02  Score=35.08  Aligned_cols=21  Identities=24%  Similarity=0.461  Sum_probs=16.7

Q ss_pred             CCchhHHHHHHHHHHHHHHHH
Q 008451          148 PIPSQIEKLKSRIDMIGAACE  168 (565)
Q Consensus       148 P~~~qiEklqKQId~iNkacE  168 (565)
                      |+....|=|=.||++++..|+
T Consensus       210 PLraSLdfLP~Ri~~F~~ra~  230 (683)
T PF08580_consen  210 PLRASLDFLPMRIEEFQSRAE  230 (683)
T ss_pred             hHHHHHHHHHHHHHHHHHHHH
Confidence            777788888888888887765


No 18 
>PRK04863 mukB cell division protein MukB; Provisional
Probab=43.15  E-value=1.2e+02  Score=37.87  Aligned_cols=46  Identities=11%  Similarity=0.328  Sum_probs=36.1

Q ss_pred             HHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhh
Q 008451          136 KREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCF  183 (565)
Q Consensus       136 ~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~  183 (565)
                      .+++++......  ...++.|.++|+...+-++..++.|...++.|+.
T Consensus      1075 ~~~~~~~~~~~r--e~EIe~L~kkL~~~~~e~~~~re~I~~aK~~W~~ 1120 (1486)
T PRK04863       1075 RRNQLEKQLTFC--EAEMDNLTKKLRKLERDYHEMREQVVNAKAGWCA 1120 (1486)
T ss_pred             HHHHHHHHHHHH--HHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            344444443332  5688999999999999999999999999999998


No 19 
>KOG3598 consensus Thyroid hormone receptor-associated protein complex, subunit TRAP230 [Transcription]
Probab=40.83  E-value=78  Score=40.09  Aligned_cols=22  Identities=9%  Similarity=0.218  Sum_probs=12.3

Q ss_pred             hhhcceEeeeccCCCCCCccch
Q 008451           97 KVSKAFVVHPKNVNAENATILP  118 (565)
Q Consensus        97 PvLr~FvVlPlnVnaeNa~IVP  118 (565)
                      |--|.|-+-|+-|.||+-+--|
T Consensus      1768 p~pR~yyL~PlPlPpedEEe~~ 1789 (2220)
T KOG3598|consen 1768 PFPRDYYLAPLPLPPEDEEEAK 1789 (2220)
T ss_pred             CCcchhhccCCCCCcccccCCC
Confidence            3345566666667666544433


No 20 
>PF08654 DASH_Dad2:  DASH complex subunit Dad2;  InterPro: IPR013963  The DASH complex is a ~10 subunit microtubule-binding complex that is transferred to the kinetochore prior to mitosis []. In Saccharomyces cerevisiae (Baker's yeast) DASH forms both rings and spiral structures on microtubules in vitro [, ]. 
Probab=40.66  E-value=23  Score=31.78  Aligned_cols=64  Identities=17%  Similarity=0.334  Sum_probs=49.1

Q ss_pred             hhcHHHHhhhcHHHHHHHHHHHHHHHHHHHHhhhhccccC-----CCCChhhhhhhhhhcchhhhhhHH
Q 008451           30 RLNQAVVQQLNLEAVKTRAISLFKAISRILEDFDAYARTN-----TTPKWQDILGQYSMVNLELFNIVD   93 (565)
Q Consensus        30 ~ln~~V~~QlNLEAVraRA~DLkkaIsriI~~LE~e~~tN-----~t~kWpDVLdqFSVIS~QL~nLvE   93 (565)
                      ||...-.-=.+|..+|.=..+|..-+..|-.+|+.-.+-.     ---||+.|+...++.|..|....+
T Consensus         5 ri~eKk~ELe~L~~l~~lS~~L~~qle~L~~kl~~m~dg~e~Va~Vl~NW~nV~r~Is~AS~~l~~~~~   73 (103)
T PF08654_consen    5 RIAEKKAELEALKQLRDLSADLASQLEALSEKLETMADGAEAVASVLANWQNVFRAISMASLSLAKYSE   73 (103)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHhccHHHHHHHHhHHHHHHHHHHHHhhhhhccc
Confidence            3444444445778888899999999999999998876522     235999999999999988887764


No 21 
>PF07106 TBPIP:  Tat binding protein 1(TBP-1)-interacting protein (TBPIP);  InterPro: IPR010776 This family consists of several eukaryotic TBP-1 interacting protein (TBPIP) sequences. TBP-1 has been demonstrated to interact with the human immunodeficiency virus type 1 (HIV-1) viral protein Tat, then modulate the essential replication process of HIV. In addition, TBP-1 has been shown to be a component of the 26S proteasome, a basic multiprotein complex that degrades ubiquitinated proteins in an ATP-dependent fashion. Human TBPIP interacts with human TBP-1 then modulates the inhibitory action of human TBP-1 on HIV-Tat-mediated transactivation [].
Probab=37.87  E-value=2.3e+02  Score=26.25  Aligned_cols=127  Identities=19%  Similarity=0.166  Sum_probs=72.6

Q ss_pred             CChhhhhhhh------hhcchhhhhhHHHHhhhhc----ceEeeeccCCCC--CCccchhhhhc-----ccCcchhhhhh
Q 008451           72 PKWQDILGQY------SMVNLELFNIVDEIRKVSK----AFVVHPKNVNAE--NATILPVMLSS-----KLLPEMEIDDN  134 (565)
Q Consensus        72 ~kWpDVLdqF------SVIS~QL~nLvEEIkPvLr----~FvVlPlnVnae--Na~IVPdmLRT-----KLlPEmEtee~  134 (565)
                      -+=.||.+++      +.|-.-|..|+++-+=+.|    .-++++..-..+  +.+-+..|=..     .-+-+.+.+..
T Consensus        17 ys~~di~~nL~~~~~K~~v~k~Ld~L~~~g~i~~K~~GKqkiY~~~Q~~~~~~s~eel~~ld~ei~~L~~el~~l~~~~k   96 (169)
T PF07106_consen   17 YSAQDIFDNLHNKVGKTAVQKALDSLVEEGKIVEKEYGKQKIYFANQDELEVPSPEELAELDAEIKELREELAELKKEVK   96 (169)
T ss_pred             CcHHHHHHHHHhhccHHHHHHHHHHHHhCCCeeeeeecceEEEeeCccccCCCCchhHHHHHHHHHHHHHHHHHHHHHHH
Confidence            3445666665      5566667777755433333    335555544333  23333322221     11122333333


Q ss_pred             HHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhhccCCCCCCCCccChhHHHHHHHHHHHHHHH
Q 008451          135 SKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCFGTRQGPQILPTLDKGQALKIQEQENLLRAA  213 (565)
Q Consensus       135 q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~gtRqGp~~~pT~dkadaaki~eqt~lL~AA  213 (565)
                      .++.++-.-...+    ..+.+.+.|+.+.+-|+..++-|...|..|..           +++.+..+|...-.-++..
T Consensus        97 ~l~~eL~~L~~~~----t~~el~~~i~~l~~e~~~l~~kL~~l~~~~~~-----------vs~ee~~~~~~~~~~~~k~  160 (169)
T PF07106_consen   97 SLEAELASLSSEP----TNEELREEIEELEEEIEELEEKLEKLRSGSKP-----------VSPEEKEKLEKEYKKWRKE  160 (169)
T ss_pred             HHHHHHHHHhcCC----CHHHHHHHHHHHHHHHHHHHHHHHHHHhCCCC-----------CCHHHHHHHHHHHHHHHHH
Confidence            3333333333334    78889999999999999999999888864432           7788888888766665544


No 22 
>PF05377 FlaC_arch:  Flagella accessory protein C (FlaC);  InterPro: IPR008039 Although archaeal flagella appear superficially similar to those of bacteria, they are quite distinct []. In several archaea, the flagellin genes are followed immediately by the flagellar accessory genes flaCDEFGHIJ. The gene products may have a role in translocation, secretion, or assembly of the flagellum. FlaC is a protein whose exact role is unknown but it has been shown to be membrane-associated (by immuno-blotting fractionated cells) [].
Probab=37.45  E-value=55  Score=27.08  Aligned_cols=46  Identities=20%  Similarity=0.226  Sum_probs=36.8

Q ss_pred             HHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhhccC
Q 008451          139 QLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCFGTR  186 (565)
Q Consensus       139 ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~gtR  186 (565)
                      .++.++..+  ...++++++.++.|.+.+|.+++-|.+.=+-||.=||
T Consensus         4 elEn~~~~~--~~~i~tvk~en~~i~~~ve~i~envk~ll~lYE~Vs~   49 (55)
T PF05377_consen    4 ELENELPRI--ESSINTVKKENEEISESVEKIEENVKDLLSLYEVVSN   49 (55)
T ss_pred             HHHHHHHHH--HHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHc
Confidence            345555444  5678899999999999999999999998888988554


No 23 
>COG0497 RecN ATPase involved in DNA repair [DNA replication, recombination, and repair]
Probab=37.43  E-value=1.2e+02  Score=34.47  Aligned_cols=88  Identities=17%  Similarity=0.209  Sum_probs=51.9

Q ss_pred             CCCCcchhhhhcHHHHhhhcHHHHHHHHHHHHHHH-------------HHHHHhhhhccccCCCCChhhhhhhhhhcchh
Q 008451           21 VPPQPVAVERLNQAVVQQLNLEAVKTRAISLFKAI-------------SRILEDFDAYARTNTTPKWQDILGQYSMVNLE   87 (565)
Q Consensus        21 ~~~~~~~~e~ln~~V~~QlNLEAVraRA~DLkkaI-------------sriI~~LE~e~~tN~t~kWpDVLdqFSVIS~Q   87 (565)
                      .-|+|--.|+|..--..-+|.|.+..-+.+....|             .+.++.|+...  +-..+..++....+=.--+
T Consensus       200 ~~l~~gE~e~L~~e~~rLsn~ekl~~~~~~a~~~L~ge~~~~~~~~~l~~a~~~l~~~~--~~d~~l~~~~~~l~ea~~~  277 (557)
T COG0497         200 LNLQPGEDEELEEERKRLSNSEKLAEAIQNALELLSGEDDTVSALSLLGRALEALEDLS--EYDGKLSELAELLEEALYE  277 (557)
T ss_pred             cCCCCchHHHHHHHHHHHhhHHHHHHHHHHHHHHHhCCCCchhHHHHHHHHHHHHHHhh--ccChhHHHHHHHHHHHHHH
Confidence            45666677777766666666665554443333322             22233333211  1244666666666655566


Q ss_pred             hhhhHHHHhhhhcceEeeeccCC
Q 008451           88 LFNIVDEIRKVSKAFVVHPKNVN  110 (565)
Q Consensus        88 L~nLvEEIkPvLr~FvVlPlnVn  110 (565)
                      |..+.+|++..++.+-+-|..+.
T Consensus       278 l~ea~~el~~~~~~le~Dp~~L~  300 (557)
T COG0497         278 LEEASEELRAYLDELEFDPNRLE  300 (557)
T ss_pred             HHHHHHHHHHHHhcCCCCHHHHH
Confidence            67777888888888888777764


No 24 
>cd00176 SPEC Spectrin repeats, found in several proteins involved in cytoskeletal structure; family members include spectrin, alpha-actinin and dystrophin; the spectrin repeat forms a three helix bundle with the second helix interrupted by proline in some sequences; the repeats are independent folding units; tandem repeats are found in differing numbers and arrange in an antiparallel manner to form dimers; the repeats are defined by a characteristic tryptophan (W) residue in helix A and a leucine (L) at the carboxyl end of helix C and separated by a linker of 5 residues; two copies of the repeat are present here
Probab=35.52  E-value=1.8e+02  Score=25.38  Aligned_cols=28  Identities=11%  Similarity=0.302  Sum_probs=17.2

Q ss_pred             HHHHHHHHHHHHHHHhHHHHHHHHHHhh
Q 008451          154 EKLKSRIDMIGAACESAEKVLADTRKAY  181 (565)
Q Consensus       154 EklqKQId~iNkacE~aekvIa~aRk~~  181 (565)
                      ....++++.|+.-++.+...+.+.++..
T Consensus       182 ~~~~~~l~~l~~~~~~l~~~~~~~~~~L  209 (213)
T cd00176         182 EEIEEKLEELNERWEELLELAEERQKKL  209 (213)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            4556666666666666666665555443


No 25 
>PF15445 ATS:  acidic terminal segments, variant surface antigen of PfEMP1
Probab=35.45  E-value=57  Score=35.75  Aligned_cols=42  Identities=21%  Similarity=0.385  Sum_probs=38.0

Q ss_pred             CCCchhHHHHHHHHHHHHHHHH---hHHHHHHHHHHhhhhccCCC
Q 008451          147 LPIPSQIEKLKSRIDMIGAACE---SAEKVLADTRKAYCFGTRQG  188 (565)
Q Consensus       147 LP~~~qiEklqKQId~iNkacE---~aekvIa~aRk~~e~gtRqG  188 (565)
                      =|+.-+++-++|-+|.+..+||   +-+++|.+..|+|+.-+-.|
T Consensus       281 DpI~nQl~LfHkWLDRHRdmCekw~~kee~L~KLkEeW~~e~~sg  325 (437)
T PF15445_consen  281 DPIHNQLNLFHKWLDRHRDMCEKWNNKEELLDKLKEEWEKENHSG  325 (437)
T ss_pred             CchhhhHHHHHHHHHHHHHHHHHhccHHHHHHHHHHHHhcccCCC
Confidence            4889999999999999999999   67899999999999966666


No 26 
>PF11887 DUF3407:  Protein of unknown function (DUF3407);  InterPro: IPR024516 This entry represents a domain of unknown function found at the C terminus of many proteins in the mammalian cell entry family. 
Probab=35.22  E-value=57  Score=32.88  Aligned_cols=29  Identities=7%  Similarity=0.028  Sum_probs=14.7

Q ss_pred             hhhhhhhhcchhhhhhHHHHhhhhcceEe
Q 008451           76 DILGQYSMVNLELFNIVDEIRKVSKAFVV  104 (565)
Q Consensus        76 DVLdqFSVIS~QL~nLvEEIkPvLr~FvV  104 (565)
                      +.|+.++-|+.-|..-..||...++++++
T Consensus        56 ~~l~~l~~v~~~~a~aapdL~~~l~~~~~   84 (267)
T PF11887_consen   56 EDLRNLADVADTYADAAPDLLDALDNLTT   84 (267)
T ss_pred             HHHHHHHHHHHHHHHhhhHHHHHHHHHHH
Confidence            34555555555555555555555554443


No 27 
>PRK04778 septation ring formation regulator EzrA; Provisional
Probab=34.26  E-value=1.1e+02  Score=33.77  Aligned_cols=63  Identities=10%  Similarity=0.245  Sum_probs=45.1

Q ss_pred             hhhcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhh-------cchhhhhhHHHHhhhhcceEe
Q 008451           37 QQLNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSM-------VNLELFNIVDEIRKVSKAFVV  104 (565)
Q Consensus        37 ~QlNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSV-------IS~QL~nLvEEIkPvLr~FvV  104 (565)
                      ..+.|+.+.....+|..-|+.+-..|+.+..     -...|-.+...       +..+...|..||..+-.+|.+
T Consensus       273 ~~l~l~~~~~~~~~i~~~Id~Lyd~lekE~~-----A~~~vek~~~~l~~~l~~~~e~~~~l~~Ei~~l~~sY~l  342 (569)
T PRK04778        273 EELDLDEAEEKNEEIQERIDQLYDILEREVK-----ARKYVEKNSDTLPDFLEHAKEQNKELKEEIDRVKQSYTL  342 (569)
T ss_pred             HhcChHHHHHHHHHHHHHHHHHHHHHHHHHH-----HHHHHHHhhHHHHHHHHHHHHHHHHHHHHHHHHHHcccc
Confidence            4788999999999999999999999988763     23333333333       444455566788888888765


No 28 
>cd03204 GST_C_GDAP1 GST_C family, Ganglioside-induced differentiation-associated protein 1 (GDAP1) subfamily; GDAP1 was originally identified as a highly expressed gene at the differentiated stage of GD3 synthase-transfected cells. More recently, mutations in GDAP1 have been reported to cause both axonal and demyelinating autosomal-recessive Charcot-Marie-Tooth (CMT) type 4A neuropathy. CMT is characterized by slow and progressive weakness and atrophy of muscles. Sequence analysis of GDAP1 shows similarities and differences with GSTs; it appears to contain both N-terminal thioredoxin-fold and C-terminal alpha helical domains of GSTs, however, it also contains additional C-terminal transmembrane domains unlike GSTs. GDAP1 is mainly expressed in neuronal cells and is localized in the mitochondria through its transmembrane domains. It does not exhibit GST activity using standard substrates.
Probab=33.88  E-value=1.4e+02  Score=26.73  Aligned_cols=68  Identities=13%  Similarity=0.126  Sum_probs=40.9

Q ss_pred             chhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhhccCCCC-CCCCccChhH
Q 008451          128 EMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCFGTRQGP-QILPTLDKGQ  199 (565)
Q Consensus       128 EmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~gtRqGp-~~~pT~dkad  199 (565)
                      +|-+.-..|...++++...   +-+.+.+.+.+..++.++..+|+.+.+..++|| ++.-+| -.+-+++.||
T Consensus         2 ~~~~~~~~~~~~~~~~~~~---~~~~~~i~~~~~~l~~~l~~LE~~L~~~~~~~~-~~~~~~yL~Gd~~TlAD   70 (111)
T cd03204           2 DLATAYIAKQKKLKSKLLD---HDNVEYLKKILDELEMVLDQVEQELQRRKEETE-EQKCQLWLCGDTFTLAD   70 (111)
T ss_pred             cHHHHHHHHHHHHHHHHHh---cccHHHHHHHHHHHHHHHHHHHHHHHcCCcccc-cccCCCccCCCCCCHHH
Confidence            3333344444445555433   235667778888888888888888875444555 333223 2335788999


No 29 
>PRK15048 methyl-accepting chemotaxis protein II; Provisional
Probab=32.69  E-value=69  Score=34.20  Aligned_cols=26  Identities=35%  Similarity=0.448  Sum_probs=19.2

Q ss_pred             HHHHHHHHHH--HHHH------hcCCCcccCCCc
Q 008451          201 LKIQEQENLL--RAAV------NSGEGLRLPGDQ  226 (565)
Q Consensus       201 aki~eqt~lL--~AAV------n~GeGLr~p~dq  226 (565)
                      ..|.||||||  -||+      -.|+|+-|+.|.
T Consensus       367 ~~Ia~QTNLLALNAaIEAARAGE~GrGFAVVA~E  400 (553)
T PRK15048        367 DGIAFQTNILALNAAVEAARAGEQGRGFAVVAGE  400 (553)
T ss_pred             HHHHHHHHHHHHHHHHHHhccccCCCCChhHHHH
Confidence            3588999996  2333      279999988875


No 30 
>PRK12425 fumarate hydratase; Provisional
Probab=32.49  E-value=5.1e+02  Score=28.47  Aligned_cols=33  Identities=24%  Similarity=0.457  Sum_probs=28.0

Q ss_pred             hhhhhhhhcchhhhhhHHHHhhhh---c---ceEeeecc
Q 008451           76 DILGQYSMVNLELFNIVDEIRKVS---K---AFVVHPKN  108 (565)
Q Consensus        76 DVLdqFSVIS~QL~nLvEEIkPvL---r---~FvVlPln  108 (565)
                      ++++-.++|...|..|.+||.-..   +   .++.+|..
T Consensus       274 e~~~~l~~la~~L~kia~Dl~llsS~p~~g~~ei~lp~~  312 (464)
T PRK12425        274 SLSGALKTLAVALMKIANDLRLLGSGPRAGLAEVRLPAN  312 (464)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHccCccCCceEEECCCC
Confidence            788899999999999999999987   4   35688854


No 31 
>PF06160 EzrA:  Septation ring formation regulator, EzrA ;  InterPro: IPR010379 During the bacterial cell cycle, the tubulin-like cell-division protein FtsZ polymerises into a ring structure that establishes the location of the nascent division site. EzrA modulates the frequency and position of FtsZ ring formation [].; GO: 0000921 septin ring assembly, 0005940 septin ring, 0016021 integral to membrane
Probab=32.42  E-value=1.2e+02  Score=33.81  Aligned_cols=68  Identities=10%  Similarity=0.247  Sum_probs=56.1

Q ss_pred             hhhcHHHHHHHHHHHHHHHHHHHHhhhhcccc--CCCCChhhhhhhhhhcchhhhhhHHHHhhhhcceEe
Q 008451           37 QQLNLEAVKTRAISLFKAISRILEDFDAYART--NTTPKWQDILGQYSMVNLELFNIVDEIRKVSKAFVV  104 (565)
Q Consensus        37 ~QlNLEAVraRA~DLkkaIsriI~~LE~e~~t--N~t~kWpDVLdqFSVIS~QL~nLvEEIkPvLr~FvV  104 (565)
                      .+++|+.++....+|..-|+.+-..||.+.++  .-.-+|+.+.+...-+..+...|..|+..+..+|++
T Consensus       269 ~~l~l~~~~~~~~~i~~~Id~lYd~le~E~~Ak~~V~~~~~~l~~~l~~~~~~~~~l~~e~~~v~~sY~L  338 (560)
T PF06160_consen  269 KNLELDEVEEENEEIEERIDQLYDILEKEVEAKKYVEKNLKELYEYLEHAKEQNKELKEELERVSQSYTL  338 (560)
T ss_pred             HcCCHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHhHHHHHHHHHHHHHHHHHHHHHHHHHHHhcCC
Confidence            57899999999999999999999999988752  133467777777777777778888999999999964


No 32 
>PF15237 PTRF_SDPR:  PTRF/SDPR family
Probab=31.60  E-value=4.5e+02  Score=27.51  Aligned_cols=133  Identities=19%  Similarity=0.264  Sum_probs=80.6

Q ss_pred             hcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhh--hhhhhcchhhhhhHHHHhhhhcceEeeeccCC------
Q 008451           39 LNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDIL--GQYSMVNLELFNIVDEIRKVSKAFVVHPKNVN------  110 (565)
Q Consensus        39 lNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVL--dqFSVIS~QL~nLvEEIkPvLr~FvVlPlnVn------  110 (565)
                      .|...||.|++-=-.-|    .+||.        |-..+|  ++|.|+=     .-||.+=+.+.|+.-|..+.      
T Consensus        72 ~~vk~Vr~r~ekQ~~qV----kklE~--------n~~eLL~Rn~FkVlI-----~Qee~eiPa~~~~k~~~~~~~~~~~~  134 (246)
T PF15237_consen   72 VNVKEVRERLEKQAAQV----KKLEA--------NHAELLKRNKFKVLI-----FQEENEIPASVFVKEPEPLPSEGSEA  134 (246)
T ss_pred             hhHHHHHHHHHHHHHHH----hhhhc--------cHHHHhhccCceEEe-----ccccccCCCcccccCCcccccccccc
Confidence            45667888876544433    56665        334455  4677764     44888878888888777772      


Q ss_pred             -------CCCCccchhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhh
Q 008451          111 -------AENATILPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCF  183 (565)
Q Consensus       111 -------aeNa~IVPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~  183 (565)
                             .+..+..|+-|++--+=++|.+.-+      ++|..|     --+..++||-|-+|++  -+-+.++|...|-
T Consensus       135 ~~~~~~~~~~e~~~~~~lsSDEe~~v~e~~ee------SRA~ri-----KRSgLkrVdsLKKAFS--kenm~KTR~n~~k  201 (246)
T PF15237_consen  135 GEEDEEKEEEEFLEPIDLSSDEEYEVEEEIEE------SRAERI-----KRSGLKRVDSLKKAFS--KENMEKTRQNIEK  201 (246)
T ss_pred             cccccccCCccccCCCCCCCccccchhhhhhH------hHHHHH-----HHHHHHHHHHHHHHHH--HHHHHHHHHHHHh
Confidence                   1245567777887555445333222      222222     1123579999999998  3446678887776


Q ss_pred             -ccCCCCCCCCccChhHHHHHH
Q 008451          184 -GTRQGPQILPTLDKGQALKIQ  204 (565)
Q Consensus       184 -gtRqGp~~~pT~dkadaaki~  204 (565)
                       ..+.|..|   +.+.--.||.
T Consensus       202 Kmnk~gTri---V~pERREKir  220 (246)
T PF15237_consen  202 KMNKLGTRI---VTPERREKIR  220 (246)
T ss_pred             hccccCCCc---CChHHhhhHh
Confidence             66666333   5555555665


No 33 
>PF05873 Mt_ATP-synt_D:  ATP synthase D chain, mitochondrial (ATP5H);  InterPro: IPR008689 ATPases (or ATP synthases) are membrane-bound enzyme complexes/ion transporters that combine ATP synthesis and/or hydrolysis with the transport of protons across a membrane. ATPases can harness the energy from a proton gradient, using the flux of ions across the membrane via the ATPase proton channel to drive the synthesis of ATP. Some ATPases work in reverse, using the energy from the hydrolysis of ATP to create a proton gradient. There are different types of ATPases, which can differ in function (ATP synthesis and/or hydrolysis), structure (e.g., F-, V- and A-ATPases, which contain rotary motors) and in the type of ions they transport [, ]. The different types include:   F-ATPases (F1F0-ATPases), which are found in mitochondria, chloroplasts and bacterial plasma membranes where they are the prime producers of ATP, using the proton gradient generated by oxidative phosphorylation (mitochondria) or photosynthesis (chloroplasts). V-ATPases (V1V0-ATPases), which are primarily found in eukaryotic vacuoles and catalyse ATP hydrolysis to transport solutes and lower pH in organelles. A-ATPases (A1A0-ATPases), which are found in Archaea and function like F-ATPases (though with respect to their structure and some inhibitor responses, A-ATPases are more closely related to the V-ATPases). P-ATPases (E1E2-ATPases), which are found in bacteria and in eukaryotic plasma membranes and organelles, and function to transport a variety of different ions across membranes. E-ATPases, which are cell-surface enzymes that hydrolyse a range of NTPs, including extracellular ATP.   F-ATPases (also known as F1F0-ATPase, or H(+)-transporting two-sector ATPase) (3.6.3.14 from EC) are composed of two linked complexes: the F1 ATPase complex is the catalytic core and is composed of 5 subunits (alpha, beta, gamma, delta, epsilon), while the F0 ATPase complex is the membrane-embedded proton channel that is composed of at least 3 subunits (A-C), nine in mitochondria (A-G, F6, F8). Both the F1 and F0 complexes are rotary motors that are coupled back-to-back. In the F1 complex, the central gamma subunit forms the rotor inside the cylinder made of the alpha(3)beta(3) subunits, while in the F0 complex, the ring-shaped C subunits forms the rotor. The two rotors rotate in opposite directions, but the F0 rotor is usually stronger, using the force from the proton gradient to push the F1 rotor in reverse in order to drive ATP synthesis []. These ATPases can also work in reverse to hydrolyse ATP to create a proton gradient. This entry represents subunit D from the F0 complex in F-ATPases found in mitochondria. The D subunit is part of the peripheral stalk that links the F1 and F0 complexes together, and which acts as a stator to prevent certain subunits from rotating with the central rotary element. The peripheral stalk differs in subunit composition between mitochondrial, chloroplast and bacterial F-ATPases. In mitochondria, the peripheral stalk is composed of one copy each of subunits OSCP (oligomycin sensitivity conferral protein), F6, B and D []. There is no homologue of subunit D in bacterial or chloroplast F-ATPase, whose peripheral stalks are composed of one copy of the delta subunit (homologous to OSCP), and two copies of subunit B in bacteria, or one copy each of subunits B and B' in chloroplasts and photosynthetic bacteria.  More information about this protein can be found at Protein of the Month: ATP Synthases [].; GO: 0015078 hydrogen ion transmembrane transporter activity, 0015986 ATP synthesis coupled proton transport, 0000276 mitochondrial proton-transporting ATP synthase complex, coupling factor F(o); PDB: 2CLY_E 2WSS_U.
Probab=31.11  E-value=3.6e+02  Score=25.81  Aligned_cols=102  Identities=19%  Similarity=0.255  Sum_probs=53.4

Q ss_pred             hhhcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhHHHHhhhhcceEeeeccCCCCCCcc
Q 008451           37 QQLNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIVDEIRKVSKAFVVHPKNVNAENATI  116 (565)
Q Consensus        37 ~QlNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLvEEIkPvLr~FvVlPlnVnaeNa~I  116 (565)
                      +-..|.++|.|-.+++..+..       +.+.-+++.|.--=..   |. +-.+||+++.+-++.|-| |.-        
T Consensus        23 ~~~~~~afk~r~d~~~~~v~~-------~pe~pp~IDwa~Yk~~---l~-~~~~lVD~feK~y~s~ki-p~p--------   82 (161)
T PF05873_consen   23 QKAQFQAFKKRSDEYKRRVSK-------LPEQPPKIDWAHYKSV---LK-ENPGLVDEFEKQYESFKI-PYP--------   82 (161)
T ss_dssp             GHHHHHHHHHHHHHHHHHHHH-------S-SS-----HHHHHHC----S--STTHHHHHHHHHCC---------------
T ss_pred             HHHHHHHHHHHHHHHHHHHHh-------CcCCCCCCCHHHHHHH---hh-hhHHHHHHHHHHHhccCC-CCC--------
Confidence            345677888888888777752       3333478888654332   33 556699999999999874 422        


Q ss_pred             chhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHH
Q 008451          117 LPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEK  172 (565)
Q Consensus       117 VPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aek  172 (565)
                           .++.+.++|..+..+.+...         ...+...+||+.+.+.+++.+.
T Consensus        83 -----~d~~~~~i~~~e~~~~~~~~---------~~~~~s~~~i~~l~keL~~i~~  124 (161)
T PF05873_consen   83 -----VDKQTKEIDAQEKEAIKEAK---------EFEAESKKRIAELEKELANIES  124 (161)
T ss_dssp             -------TTTTHHHHHHHHHHHCHH---------HHHHHHHHHHHHHHHHHHHHT-
T ss_pred             -----hHHHHHHHHHHHHHHHHHHH---------HHHHHHHHHHHHHHHHHHHHHc
Confidence                 25677778777766422211         1334445566655555554444


No 34 
>cd07595 BAR_RhoGAP_Rich-like The Bin/Amphiphysin/Rvs (BAR) domain of Rich-like Rho GTPase Activating Proteins. BAR domains are dimerization, lipid binding and curvature sensing modules found in many different proteins with diverse functions. This subfamily is composed of Rho and Rac GTPase activating proteins (GAPs) with similarity to GAP interacting with CIP4 homologs proteins (Rich). Members contain an N-terminal BAR domain, followed by a Rho GAP domain, and a C-terminal prolin-rich region. Vertebrates harbor at least three Rho GAPs in this subfamily including Rich1, Rich2, and SH3-domain binding protein 1 (SH3BP1). Rich1 and Rich2 play complementary roles in the establishment and maintenance of cell polarity. Rich1 is a Cdc42- and Rac-specific GAP that binds to polarity proteins through the scaffold protein angiomotin and plays a role in maintaining the integrity of tight junctions. Rich2 is a Rac GAP that interacts with CD317 and plays a role in actin cytoskeleton organization and 
Probab=30.91  E-value=1.4e+02  Score=30.21  Aligned_cols=30  Identities=13%  Similarity=0.144  Sum_probs=24.3

Q ss_pred             cHHHHHHHHHHHHHHHHHHHHhhhhccccC
Q 008451           40 NLEAVKTRAISLFKAISRILEDFDAYARTN   69 (565)
Q Consensus        40 NLEAVraRA~DLkkaIsriI~~LE~e~~tN   69 (565)
                      .|..|-.|++.+++++..|+.++..+..-|
T Consensus        16 ~~~~lE~~~d~~k~~~~~~~k~~~~~lq~n   45 (244)
T cd07595          16 ELLQIEKRVEAVKDACQNIHKKLISCLQGQ   45 (244)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHhhHHhcCCC
Confidence            456788999999999999999888776544


No 35 
>PRK09367 histidine ammonia-lyase; Provisional
Probab=30.57  E-value=83  Score=34.96  Aligned_cols=66  Identities=21%  Similarity=0.251  Sum_probs=47.0

Q ss_pred             CCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhhccCCCCCCCCccChhHHHHHHHHHHHHHH-HHhcCC
Q 008451          147 LPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCFGTRQGPQILPTLDKGQALKIQEQENLLRA-AVNSGE  218 (565)
Q Consensus       147 LP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~gtRqGp~~~pT~dkadaaki~eqt~lL~A-AVn~Ge  218 (565)
                      .++....|.    ++.+.+..+.+++.+++.+..|++.|=-|+....+++.++.+.+  |.++|++ |++.|+
T Consensus        24 ~~v~ls~~~----~~ri~~sr~~l~~~~~~~~~iYGvnTG~G~~~~~~i~~~~~~~l--q~nLi~sha~GvG~   90 (500)
T PRK09367         24 AKVELDPSA----RAAIAASRAVVERIVAEGRPVYGINTGFGKLASVRIAPEDLEQL--QRNLVLSHAAGVGE   90 (500)
T ss_pred             CceeeCHHH----HHHHHHHHHHHHHHHhcCCcccccccCCccccCcccCHHHHHHH--HHHHHHHHHcCCCC
Confidence            444445554    34556677777788999999999988888777777777775554  5888875 666666


No 36 
>PRK14145 heat shock protein GrpE; Provisional
Probab=30.57  E-value=67  Score=31.92  Aligned_cols=69  Identities=20%  Similarity=0.311  Sum_probs=47.9

Q ss_pred             CCCCcchhhhhcHHHHhhhcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhHHHHhhhhc
Q 008451           21 VPPQPVAVERLNQAVVQQLNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIVDEIRKVSK  100 (565)
Q Consensus        21 ~~~~~~~~e~ln~~V~~QlNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLvEEIkPvLr  100 (565)
                      ....+..++.|...      |+.++.++.+|+....|..-+||.|-+  -+-+=-+-+..|++-+     ++.++-|+++
T Consensus        40 ~~~~~~e~~~l~~~------l~~le~e~~el~d~~lR~~AEfeN~rk--R~~kE~e~~~~~a~e~-----~~~~LLpV~D  106 (196)
T PRK14145         40 QQQTVDEIEELKQK------LQQKEVEAQEYLDIAQRLKAEFENYRK--RTEKEKSEMVEYGKEQ-----VILELLPVMD  106 (196)
T ss_pred             ccCchhHHHHHHHH------HHHHHHHHHHHHHHHHHHHHHHHHHHH--HHHHHHHHHHHHHHHH-----HHHHHHhHHh
Confidence            33444556666554      458999999999999999999999873  1112234455566554     8888888888


Q ss_pred             ce
Q 008451          101 AF  102 (565)
Q Consensus       101 ~F  102 (565)
                      +|
T Consensus       107 nL  108 (196)
T PRK14145        107 NF  108 (196)
T ss_pred             HH
Confidence            76


No 37 
>PF02074 Peptidase_M32:  Carboxypeptidase Taq (M32) metallopeptidase;  InterPro: IPR001333 In the MEROPS database peptidases and peptidase homologues are grouped into clans and families. Clans are groups of families for which there is evidence of common ancestry based on a common structural fold:  Each clan is identified with two letters, the first representing the catalytic type of the families included in the clan (with the letter 'P' being used for a clan containing families of more than one of the catalytic types serine, threonine and cysteine). Some families cannot yet be assigned to clans, and when a formal assignment is required, such a family is described as belonging to clan A-, C-, M-, N-, S-, T- or U-, according to the catalytic type. Some clans are divided into subclans because there is evidence of a very ancient divergence within the clan, for example MA(E), the gluzincins, and MA(M), the metzincins. Peptidase families are grouped by their catalytic type, the first character representing the catalytic type: A, aspartic; C, cysteine; G, glutamic acid; M, metallo; N, asparagine; S, serine; T, threonine; and U, unknown. The serine, threonine and cysteine peptidases utilise the amino acid as a nucleophile and form an acyl intermediate - these peptidases can also readily act as transferases. In the case of aspartic, glutamic and metallopeptidases, the nucleophile is an activated water molecule. In the case of the asparagine endopeptidases, the nucleophile is asparagine and all are self-processing endopeptidases.   In many instances the structural protein fold that characterises the clan or family may have lost its catalytic activity, yet retain its function in protein recognition and binding.  Metalloproteases are the most diverse of the four main types of protease, with more than 50 families identified to date. In these enzymes, a divalent cation, usually zinc, activates the water molecule. The metal ion is held in place by amino acid ligands, usually three in number. The known metal ligands are His, Glu, Asp or Lys and at least one other residue is required for catalysis, which may play an electrophillic role. Of the known metalloproteases, around half contain an HEXXH motif, which has been shown in crystallographic studies to form part of the metal-binding site []. The HEXXH motif is relatively common, but can be more stringently defined for metalloproteases as 'abXHEbbHbc', where 'a' is most often valine or threonine and forms part of the S1' subsite in thermolysin and neprilysin, 'b' is an uncharged residue, and 'c' a hydrophobic residue. Proline is never found in this site, possibly because it would break the helical structure adopted by this motif in metalloproteases []. This group of metallopeptidases belong to MEROPS peptidase family M32 (carboxypeptidase Taq family, clan MA(E)). The predicted active site residues for members of this family and thermolysin, the type example for clan MA, occur in the motif HEXXH.  Carboxypeptidase Taq is a zinc-containing thermostable metallopeptidase. It was originally discovered and purified from Thermus aquaticus; optimal enzymatic activity occurs at 80 celcius. Although very little is known about this enzyme, it is thought either to be associated with a membrane or to be particle bound.; GO: 0004181 metallocarboxypeptidase activity, 0006508 proteolysis; PDB: 1K9X_A 1KA4_A 1KA2_A 3DWC_A 1WGZ_A 3HQ2_A 3HOA_B.
Probab=29.91  E-value=2.4e+02  Score=31.55  Aligned_cols=62  Identities=13%  Similarity=0.051  Sum_probs=34.8

Q ss_pred             cHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhH--HHHhhhhcceEe
Q 008451           40 NLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIV--DEIRKVSKAFVV  104 (565)
Q Consensus        40 NLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLv--EEIkPvLr~FvV  104 (565)
                      .|.....|+.+|.++++-+--+.++..   +.--=+.=-.+-+.+++..+.+.  ++++.+|+.+.-
T Consensus         9 ~l~~~~~~i~~l~~a~slL~WD~~T~m---P~~g~~~Raeqla~Ls~~~hel~T~~~~~elL~~l~~   72 (494)
T PF02074_consen    9 ELKEHLREISALEHASSLLYWDQETMM---PKGGAEARAEQLATLSGLIHELLTSPEIGELLEELEE   72 (494)
T ss_dssp             HHHHHHHHHHHHHHHHHHHHHHHHCT-----GGGHHHHHHHHHHHHHHHHHHHTSHHHHHHHHHHHC
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHhCC---CcccHHHHHHHHHHHHHHHHHHHcCHHHHHHHHHHhc
Confidence            355566666777766665555555543   22222333345566677777776  667777665543


No 38 
>PF11640 TAN:  Telomere-length maintenance and DNA damage repair;  InterPro: IPR021668  ATM is a large protein kinase, in humans, critical for responding to DNA double-strand breaks (DSBs). Tel1, the orthologue from budding yeast, also regulates responses to DSBs. Tel1 is important for maintaining viability and for phosphorylation of the DNA damage signal transducer kinase Rad53 (an orthologue of mammalian CHK2). In addition to functioning in the response to DSBs, numerous findings indicate that Tel1/ATM regulates telomeres. The overall domain structure of Tel1/ATM is shared by proteins of the phosphatidylinositol 3-kinase (PI3K)-related kinase (PIKK) family, but this family carries a unique and functionally important TAN sequence motif, near its N-terminal, LxxxKxxE/DRxxxL. which is conserved specifically in the Tel1/ATM subclass of the PIKKs. The TAN motif is essential for both telomere length maintenance and Tel1 action in response to DNA damage []. It is classified as an 2.7.11.1 from EC. ; GO: 0004674 protein serine/threonine kinase activity
Probab=29.09  E-value=75  Score=29.22  Aligned_cols=77  Identities=23%  Similarity=0.285  Sum_probs=52.9

Q ss_pred             HHHHHHHhHHHHHHHHHHhhhhccCCCCCCCCccChhHHHHHHHHHHHHHHHHhcCCCcccCCCcCCCCCCCchhhhhcc
Q 008451          162 MIGAACESAEKVLADTRKAYCFGTRQGPQILPTLDKGQALKIQEQENLLRAAVNSGEGLRLPGDQRQMTPALPMHLVDLL  241 (565)
Q Consensus       162 ~iNkacE~aekvIa~aRk~~e~gtRqGp~~~pT~dkadaaki~eqt~lL~AAVn~GeGLr~p~dqr~~~~~lp~hl~~~l  241 (565)
                      .+.+++|.+.+.|.+.++.|..+...   ...|... -+.|+++=..+||-+|..|-- +++   |.--..|-.|+.|+|
T Consensus        45 ~~~~ifeaL~~~i~~Ek~~y~~~~~~---~~s~~~~-~~~RL~~~a~~lR~~ve~~~~-~~k---~kt~~~Ll~hI~~~l  116 (155)
T PF11640_consen   45 QWHSIFEALFRCIEKEKEAYSRKKSS---SASTATT-AESRLSSCASALRLFVEKSNS-RLK---RKTVKALLDHITDLL  116 (155)
T ss_pred             hHHHHHHHHHHHHHHHHHHHhcCCCc---ccchHHH-HHHHHHHHHHHHHHHHHHHHh-hcc---cchHHHHHHHHHHHh
Confidence            57889999999999999999322111   1222222 246888888999999976643 222   333356788999999


Q ss_pred             ccCCC
Q 008451          242 PVGDG  246 (565)
Q Consensus       242 ~~~dg  246 (565)
                      ...||
T Consensus       117 ~~~~~  121 (155)
T PF11640_consen  117 PDPDD  121 (155)
T ss_pred             hCCch
Confidence            98873


No 39 
>PF08549 SWI-SNF_Ssr4:  Fungal domain of unknown function (DUF1750);  InterPro: IPR013859  This is a fungal protein of unknown function. 
Probab=28.30  E-value=1.3e+02  Score=35.03  Aligned_cols=82  Identities=20%  Similarity=0.259  Sum_probs=49.4

Q ss_pred             hHHHHhhhhcceEeeeccCCCCCCccchhhhh-cccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHH----HHHHHH-
Q 008451           91 IVDEIRKVSKAFVVHPKNVNAENATILPVMLS-SKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKS----RIDMIG-  164 (565)
Q Consensus        91 LvEEIkPvLr~FvVlPlnVnaeNa~IVPdmLR-TKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqK----QId~iN-  164 (565)
                      +.-|+..+-+.|---|.--..+...-.|.-.. .||+||.-.|=+   +++..+...+  ...|||+|+    +++.++ 
T Consensus       322 rKGEL~sLTeGiF~ap~~~~~~~~~~~~~~~~~gkLdp~~aeeF~---kRV~~~ia~~--~AEIekmK~~Hak~m~k~k~  396 (669)
T PF08549_consen  322 RKGELESLTEGIFEAPGGQSGDASKEGPKKPYVGKLDPGKAEEFR---KRVAKKIADM--NAEIEKMKARHAKRMAKFKR  396 (669)
T ss_pred             cccchhhhhcccccCCCCCCccccccCCCcccccCCCHHHHHHHH---HHHHHHHHHH--HHHHHHHHHHHHHHHHHHhh
Confidence            34678888888877776666555455554444 899999744433   3444444433  557888776    455553 


Q ss_pred             -HHHHhHHHHHHHH
Q 008451          165 -AACESAEKVLADT  177 (565)
Q Consensus       165 -kacE~aekvIa~a  177 (565)
                       +++..+|+-|.++
T Consensus       397 ~s~lk~AE~~LR~a  410 (669)
T PF08549_consen  397 NSLLKDAEKELRDA  410 (669)
T ss_pred             ccHHHHHHHHHHhc
Confidence             3456666655443


No 40 
>PRK05771 V-type ATP synthase subunit I; Validated
Probab=27.30  E-value=1.9e+02  Score=32.43  Aligned_cols=22  Identities=23%  Similarity=0.330  Sum_probs=11.5

Q ss_pred             cHHHHHHHHHHHHHHHHHHHHh
Q 008451           40 NLEAVKTRAISLFKAISRILED   61 (565)
Q Consensus        40 NLEAVraRA~DLkkaIsriI~~   61 (565)
                      .++.+..++.+|.+.+.++...
T Consensus        94 ~~~~~~~~i~~l~~~~~~L~~~  115 (646)
T PRK05771         94 ELEKIEKEIKELEEEISELENE  115 (646)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHH
Confidence            4555555555555555444443


No 41 
>PF08656 DASH_Dad3:  DASH complex subunit Dad3;  InterPro: IPR013965  The DASH complex is a ~10 subunit microtubule-binding complex that is transferred to the kinetochore prior to mitosis []. In Saccharomyces cerevisiae (Baker's yeast) DASH forms both rings and spiral structures on microtubules in vitro [, ]. 
Probab=27.09  E-value=1.4e+02  Score=26.18  Aligned_cols=68  Identities=16%  Similarity=0.298  Sum_probs=45.5

Q ss_pred             hhhhhhhhhcchhhhhhHHHHhhhhcceEeeeccCCCCCCccchhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHH
Q 008451           75 QDILGQYSMVNLELFNIVDEIRKVSKAFVVHPKNVNAENATILPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIE  154 (565)
Q Consensus        75 pDVLdqFSVIS~QL~nLvEEIkPvLr~FvVlPlnVnaeNa~IVPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiE  154 (565)
                      .+||+.|.-++..|.+|.++++.+.          +.++++.                  .    ++.++..|       
T Consensus         6 q~VL~eY~~La~~L~~L~~~l~~L~----------~~~~~~~------------------~----lL~~LR~L-------   46 (78)
T PF08656_consen    6 QEVLDEYQRLADNLKTLSDTLKDLN----------SSNSPSE------------------E----LLDGLREL-------   46 (78)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHH----------ccCCChH------------------H----HHHHHHHH-------
Confidence            5799999999999999999999871          0111100                  1    22222222       


Q ss_pred             HHHHHHHHHHHHHH-hHHHHHHHHHHhhhh
Q 008451          155 KLKSRIDMIGAACE-SAEKVLADTRKAYCF  183 (565)
Q Consensus       155 klqKQId~iNkacE-~aekvIa~aRk~~e~  183 (565)
                        .+.|-.+.-.+. +|+.+|..-..+|+.
T Consensus        47 --E~K~glV~TL~KaSVYslvlq~~~~~e~   74 (78)
T PF08656_consen   47 --ERKIGLVYTLFKASVYSLVLQQEIDNEE   74 (78)
T ss_pred             --HHHHHHHHHHHHHHHHHHHHHHHHHhhh
Confidence              556666666665 788899888877764


No 42 
>TIGR03832 Tyr_2_3_mutase tyrosine 2,3-aminomutase. Members of this protein family are tyrosine 2,3-aminomutase. It is variable from member to member as to whether the (R)-beta-Tyr or (S)-beta-Tyr is the preferred product from L-Tyr. This enzyme tends to occur in secondary metabolite biosynthesis systems, as in the production of chondramides in Chondromyces crocatus.
Probab=26.99  E-value=94  Score=34.66  Aligned_cols=57  Identities=21%  Similarity=0.283  Sum_probs=42.8

Q ss_pred             HHHHHHHHHhHHHHHHHHHHhhhhccCCCCCCCCccChhHHHHHHHHHHHHHH-HHhcCC
Q 008451          160 IDMIGAACESAEKVLADTRKAYCFGTRQGPQILPTLDKGQALKIQEQENLLRA-AVNSGE  218 (565)
Q Consensus       160 Id~iNkacE~aekvIa~aRk~~e~gtRqGp~~~pT~dkadaaki~eqt~lL~A-AVn~Ge  218 (565)
                      ++.+.+..+-+++.+++.+-.|++.|=-|...-.+++.++.++.  |.++|++ |++.|+
T Consensus        28 ~~ri~~sr~~l~~~~~~g~~iYGvnTGfG~~~d~~i~~~~~~~l--Q~nLi~sha~GvG~   85 (507)
T TIGR03832        28 LAKAAKSRAIFEGIAEQNVPIYGVTTGYGEMIYMLVDKEHEVEL--QTNLVRSHSAGVGP   85 (507)
T ss_pred             HHHHHHHHHHHHHHHhcCCceeeecCCCCCccCccCCHHHHHHH--HHHHHHHHhcCCCC
Confidence            34556677777789999999999988888766677788887655  4888875 555554


No 43 
>KOG3598 consensus Thyroid hormone receptor-associated protein complex, subunit TRAP230 [Transcription]
Probab=26.18  E-value=77  Score=40.16  Aligned_cols=15  Identities=27%  Similarity=0.206  Sum_probs=8.3

Q ss_pred             ccCCCCCCCCccChh
Q 008451          184 GTRQGPQILPTLDKG  198 (565)
Q Consensus       184 gtRqGp~~~pT~dka  198 (565)
                      -.|.+|.--+|+.+-
T Consensus      1842 ~~r~~p~~~~~s~~~ 1856 (2220)
T KOG3598|consen 1842 CKRASPKDDVTSEKN 1856 (2220)
T ss_pred             cccCCCCCCCCChHh
Confidence            456666555555443


No 44 
>PF08580 KAR9:  Yeast cortical protein KAR9;  InterPro: IPR013889  The KAR9 protein in Saccharomyces cerevisiae (Baker's yeast) is a cytoskeletal protein required for karyogamy, correct positioning of the mitotic spindle and for orientation of cytoplasmic microtubules []. KAR9 localises at the shmoo tip in mating cells and at the tip of the growing bud in anaphase []. 
Probab=25.76  E-value=85  Score=36.16  Aligned_cols=127  Identities=12%  Similarity=0.100  Sum_probs=63.5

Q ss_pred             cHHHHHHHHHHHHHHHHHHHHhhhhcc---ccC-C------CCChhhhhhhhhhcchhhhhhHHHHhhhhcceEeeeccC
Q 008451           40 NLEAVKTRAISLFKAISRILEDFDAYA---RTN-T------TPKWQDILGQYSMVNLELFNIVDEIRKVSKAFVVHPKNV  109 (565)
Q Consensus        40 NLEAVraRA~DLkkaIsriI~~LE~e~---~tN-~------t~kWpDVLdqFSVIS~QL~nLvEEIkPvLr~FvVlPlnV  109 (565)
                      +|=++.+|+.=|+.+|+=+=.+|+...   +.+ +      .-+|..+.++|..+-.++..|-+|+..-=-  +++=+++
T Consensus       200 ~ll~L~arm~PLraSLdfLP~Ri~~F~~ra~~~fp~a~e~L~~r~~~L~~k~~~L~~e~~~LK~ELiedRW--~~vFr~l  277 (683)
T PF08580_consen  200 SLLALFARMQPLRASLDFLPMRIEEFQSRAESIFPSACEELEDRYERLEKKWKKLEKEAESLKKELIEDRW--NIVFRNL  277 (683)
T ss_pred             HHHHHHhccchHHHHHHHHHHHHHHHHHHHHHhhHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHhhhhhH--HHHHHHH
Confidence            455777777777777743333333221   100 0      124555555665555555555555443111  1111222


Q ss_pred             CCCCCccchhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhh
Q 008451          110 NAENATILPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAY  181 (565)
Q Consensus       110 naeNa~IVPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~  181 (565)
                      +-+   +      .|..=++|.....+.+..+.++..-    ..+++.|+|+.+-+.|.|--.+|-++|+.=
T Consensus       278 ~~q---~------~~m~esver~~~kl~~~~~~~~~~~----~~~~l~~~i~s~~~k~~~~~~~I~ka~~~s  336 (683)
T PF08580_consen  278 GRQ---A------QKMCESVERSLSKLQEAIDSGIHLD----NPSKLSKQIESKEKKKSHYFPAIYKARVLS  336 (683)
T ss_pred             HHH---H------HHHHHHHHHHHHHhhcccccccccc----chHHHHHHHHHHHHHHhccHHHHHHHHHHH
Confidence            211   0      1112233333333333322222222    457788999999999998888887777654


No 45 
>PRK08776 cystathionine gamma-synthase; Provisional
Probab=25.73  E-value=1.9e+02  Score=30.50  Aligned_cols=92  Identities=16%  Similarity=0.195  Sum_probs=54.7

Q ss_pred             hcchhhhhhHHHHhhhhcceEeeeccCC-CCCCccc--hhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHH
Q 008451           83 MVNLELFNIVDEIRKVSKAFVVHPKNVN-AENATIL--PVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSR  159 (565)
Q Consensus        83 VIS~QL~nLvEEIkPvLr~FvVlPlnVn-aeNa~IV--PdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQ  159 (565)
                      |||-+|..=.++.+++++++-+..+-++ .+..++|  |......-+++-|-+..-    +..+...  ++..+|....-
T Consensus       307 ~~s~~~~~~~~~~~~f~~~l~l~~~~~s~G~~~sl~~~p~~~~h~~~~~~~~~~~g----i~~~liR--~svGlE~~~dl  380 (405)
T PRK08776        307 MLSFELEGGEAAVRAFVDGLRYFTLAESLGGVESLIAHPASMTHAAMTAEARAAAG----ISDGLLR--LSVGIESAEDL  380 (405)
T ss_pred             EEEEEEcCCHHHHHHHHHhCCcceEccCCCCCceEEECCcccccccCCHHHHHhcC----CCCCeEE--EEeCcCCHHHH
Confidence            7777775435777889999888887777 4455555  655554444431111111    2222333  35566777777


Q ss_pred             HHHHHHHHHhHHHHHH-HHHHh
Q 008451          160 IDMIGAACESAEKVLA-DTRKA  180 (565)
Q Consensus       160 Id~iNkacE~aekvIa-~aRk~  180 (565)
                      |+-|..+.+.++.++. .+||.
T Consensus       381 i~dl~~al~~~~~~~~~~~~~~  402 (405)
T PRK08776        381 LIDLRAGLARAEAVLTAAARKK  402 (405)
T ss_pred             HHHHHHHHHHhHHHHHHHhhhh
Confidence            7777777777777664 44553


No 46 
>COG1937 Uncharacterized protein conserved in bacteria [Function unknown]
Probab=25.73  E-value=1.4e+02  Score=26.50  Aligned_cols=50  Identities=16%  Similarity=0.151  Sum_probs=42.2

Q ss_pred             cHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhHHHH
Q 008451           40 NLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIVDEI   95 (565)
Q Consensus        40 NLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLvEEI   95 (565)
                      ..+.+++|+.-++.-|..|..-+|-..      --.|||-+++=|.+-|.++..+|
T Consensus         7 ~kkkl~~RlrRi~GQv~gI~rMlEe~~------~C~dVl~QIaAVr~Al~~~~~~v   56 (89)
T COG1937           7 EKKKLLNRLRRIEGQVRGIERMLEEDR------DCIDVLQQIAAVRGALNGLMREV   56 (89)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHhCCC------cHHHHHHHHHHHHHHHHHHHHHH
Confidence            457899999999999999988888743      56899999999999998888554


No 47 
>KOG1412 consensus Aspartate aminotransferase/Glutamic oxaloacetic transaminase AAT2/GOT1 [Amino acid transport and metabolism]
Probab=25.72  E-value=97  Score=33.85  Aligned_cols=61  Identities=13%  Similarity=0.210  Sum_probs=48.7

Q ss_pred             hcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhH-HHHhhhhcceEeeec
Q 008451           39 LNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIV-DEIRKVSKAFVVHPK  107 (565)
Q Consensus        39 lNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLv-EEIkPvLr~FvVlPl  107 (565)
                      -++..+-+|+...|+++.+-|.+|.+-      .+|+-|.++.-|.|  +++|. ..++-+.++..||-+
T Consensus       316 ~sik~MssRI~~MR~aLrd~L~aL~TP------GtWDHI~~QiGMFS--yTGLtp~qV~~li~~h~vyLl  377 (410)
T KOG1412|consen  316 QSIKTMSSRIKKMRTALRDHLVALKTP------GTWDHITQQIGMFS--YTGLTPAQVDHLIENHKVYLL  377 (410)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHhcCCC------CcHHHHHhhcccee--ecCCCHHHHHHHHHhceEEEe
Confidence            456688899999999999888888884      49999999999999  77777 455557777777643


No 48 
>PF10444 Nbl1_Borealin_N:  Nbl1 / Borealin N terminal;  InterPro: IPR018851 This entry represents the N-terminal domain of borealin, and is also found in the N-terminal-Borealin-like (NBL; YHR199C-A) protein from Saccharomyces cerevisiae (Baker's yeast). NBL is a subunit of the conserved chromosomal passenger complex (CPC; Ipl1p-Sli15p-Bir1p-Nbl1p), which regulates mitotic chromosome segregation. It is not required for the kinase activity of the complex and it mediates the interaction of Sli15p and Bir1p [].; PDB: 2RAW_B 2RAX_Y 2QFA_B.
Probab=24.90  E-value=77  Score=25.43  Aligned_cols=43  Identities=21%  Similarity=0.332  Sum_probs=29.6

Q ss_pred             hhcHHHHHHHHHHHHHHHHHHHHhhhhccccC--------CCCChhhhhhhh
Q 008451           38 QLNLEAVKTRAISLFKAISRILEDFDAYARTN--------TTPKWQDILGQY   81 (565)
Q Consensus        38 QlNLEAVraRA~DLkkaIsriI~~LE~e~~tN--------~t~kWpDVLdqF   81 (565)
                      ..++| |..|+..|+.....++.+++..++..        -.-+|-||+.+|
T Consensus         9 ~fd~E-v~~r~~~lr~~~~~~~~~~~~~~~~~l~riP~~vR~m~~~d~~~~y   59 (59)
T PF10444_consen    9 NFDLE-VEERIRRLRAQYENLLQSLRNRLEMELLRIPKAVRKMTMRDFLEKY   59 (59)
T ss_dssp             HHHHH-HHHHHHHHHHHHHHHHHHHHHHHHHHHHHS-HHHHTSBHHHHHH--
T ss_pred             HHHHH-HHHHHHHHHHHHHHHHHHHHHHHHHHHHHcCHHHHhCCHHHHhhcC
Confidence            34454 78888888888888888888776511        256888888776


No 49 
>KOG1655 consensus Protein involved in vacuolar protein sorting [Intracellular trafficking, secretion, and vesicular transport]
Probab=24.88  E-value=1.7e+02  Score=29.98  Aligned_cols=57  Identities=25%  Similarity=0.350  Sum_probs=43.1

Q ss_pred             hHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhh------ccCCCCCCCCccChhHHHHHHHHHHHH
Q 008451          152 QIEKLKSRIDMIGAACESAEKVLADTRKAYCF------GTRQGPQILPTLDKGQALKIQEQENLL  210 (565)
Q Consensus       152 qiEklqKQId~iNkacE~aekvIa~aRk~~e~------gtRqGp~~~pT~dkadaaki~eqt~lL  210 (565)
                      ....|..-|+-+|+.-+++++.|++.--+.+-      -+|.|  ++.++-|..|-++..|.+.+
T Consensus        13 p~psL~dai~~v~~r~dSve~KIskLDaeL~k~~~Qi~k~R~g--paq~~~KqrAlrVLkQKK~y   75 (218)
T KOG1655|consen   13 PPPSLQDAIDSVNKRSDSVEKKISKLDAELCKYKDQIKKTRPG--PAQNALKQRALRVLKQKKMY   75 (218)
T ss_pred             CChhHHHHHHHHHHhhhhHHHHHHHHHHHHHHHHHHHHhcCCC--cchhHHHHHHHHHHHHHHHH
Confidence            44567777889999999999888766555544      78999  77788888888887766544


No 50 
>TIGR00293 prefoldin, archaeal alpha subunit/eukaryotic subunit 5. This model finds a set of small proteins from the Archaea and from Aquifex aeolicus that may represent two orthologous groups. The proteins are predicted to be mostly coiled coil, and may hit large numbers of proteins that contain coiled coil regions.
Probab=24.87  E-value=3.7e+02  Score=23.58  Aligned_cols=30  Identities=33%  Similarity=0.351  Sum_probs=19.2

Q ss_pred             HHHHHHHHHHHHHHHHhHHHHHHHHHHhhh
Q 008451          153 IEKLKSRIDMIGAACESAEKVLADTRKAYC  182 (565)
Q Consensus       153 iEklqKQId~iNkacE~aekvIa~aRk~~e  182 (565)
                      .+-+.+||+.+++..+..++.|.+.++.+.
T Consensus        88 ~~~l~~~~~~l~~~~~~l~~~l~~l~~~~~  117 (126)
T TIGR00293        88 IEFLKKRIEELEKAIEKLQEALAELASRAQ  117 (126)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            345566777777777777666666666554


No 51 
>PF10732 DUF2524:  Protein of unknown function (DUF2524);  InterPro: IPR019668  This entry represents proteins with unknown function, and appear to be restricted to the Bacillaceae. 
Probab=24.83  E-value=1.2e+02  Score=27.15  Aligned_cols=38  Identities=13%  Similarity=0.121  Sum_probs=30.6

Q ss_pred             HHHHHHHhHHHHHHHHHHhhhhccCCCCCCCCccChhH
Q 008451          162 MIGAACESAEKVLADTRKAYCFGTRQGPQILPTLDKGQ  199 (565)
Q Consensus       162 ~iNkacE~aekvIa~aRk~~e~gtRqGp~~~pT~dkad  199 (565)
                      .+...++.++++|..+.|-|+.|.|++.--.--.+.|+
T Consensus         6 s~~~~lq~~e~~i~~a~eQ~~~~~rqehynd~eYt~Aq   43 (84)
T PF10732_consen    6 SVDEFLQQCEQAIRFAQEQFEEGSRQEHYNDEEYTEAQ   43 (84)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHhhcccchHHHHHHH
Confidence            45677888999999999999999999965555555555


No 52 
>cd07606 BAR_SFC_plant The Bin/Amphiphysin/Rvs (BAR) domain of the plant protein SCARFACE (SFC). BAR domains are dimerization, lipid binding and curvature sensing modules found in many different proteins with diverse functions including organelle biogenesis, membrane trafficking or remodeling, and cell division and migration. The plant protein SCARFACE (SFC), also called VAscular Network 3 (VAN3), is a plant ACAP (ArfGAP with Coiled-coil, ANK repeat and PH domain containing protein), an Arf GTPase Activating Protein (GAP) that plays a role in the trafficking of auxin efflux regulators from the plasma membrane to the endosome. It is required for the normal vein patterning in leaves. SCF contains an N-terminal BAR domain, followed by a Pleckstrin Homology (PH) domain, an Arf GAP domain, and C-terminal ankyrin (ANK) repeats. BAR domains form dimers that bind to membranes, induce membrane bending and curvature, and may also be involved in protein-protein interactions.
Probab=24.69  E-value=2.1e+02  Score=28.30  Aligned_cols=29  Identities=10%  Similarity=0.281  Sum_probs=25.0

Q ss_pred             hhcHHHHHHHHHHHHHHHHHHHHhhhhcc
Q 008451           38 QLNLEAVKTRAISLFKAISRILEDFDAYA   66 (565)
Q Consensus        38 QlNLEAVraRA~DLkkaIsriI~~LE~e~   66 (565)
                      +-+.+.+++|++-|.|.-.+++..+..++
T Consensus         7 E~~~~~l~~~~~Kl~K~~~~~~~a~~~~~   35 (202)
T cd07606           7 EGSADELRDRSLKLYKGCRKYRDALGEAY   35 (202)
T ss_pred             HhhHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            34677999999999999999888888876


No 53 
>PF12795 MscS_porin:  Mechanosensitive ion channel porin domain
Probab=24.18  E-value=4.4e+02  Score=25.84  Aligned_cols=65  Identities=18%  Similarity=0.209  Sum_probs=48.0

Q ss_pred             hhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhhc-cC--CCCCCCCccChhH
Q 008451          132 DDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCFG-TR--QGPQILPTLDKGQ  199 (565)
Q Consensus       132 ee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~g-tR--qGp~~~pT~dkad  199 (565)
                      +...+++.+......|   ..+++.+++++.|.++++.+-+.|.+.|++.+.= .-  ..+.+....+..+
T Consensus        15 ~~~~~i~~l~~al~~L---~~~~~~~~~~~~~~~~i~~aP~~~~~l~~~l~~l~~~~~~~~~~~~~~s~~e   82 (240)
T PF12795_consen   15 EQKALIQDLQQALSFL---DEIKKQKKRAAEYQKQIDQAPKEIRELQKELEALKSQDAPSKEILANLSLEE   82 (240)
T ss_pred             hhHHHHHHHHHHHHHH---HHHHHHHHHHHHHHHHHHHhHHHHHHHHHHHHhhhccccccccCcccCCHHH
Confidence            3555566666666666   7899999999999999999999999999999872 22  3334444555544


No 54 
>cd00332 PAL-HAL Phenylalanine ammonia-lyase (PAL) and histidine ammonia-lyase (HAL). PAL and HAL are members of the Lyase class I_like superfamily of enzymes that, catalyze similar beta-elimination reactions and are active as homotetramers. The four active sites of the homotetrameric enzyme are each formed by residues from three different subunits. PAL, present in plants and fungi, catalyzes the conversion of L-phenylalanine to E-cinnamic acid. HAL, found in several bacteria and animals, catalyzes the conversion of L-histidine to E-urocanic acid. Both PAL and HAL contain the cofactor 3, 5-dihydro-5-methylidene-4H-imidazol-4-one (MIO) which is formed by autocatalytic excision/cyclization of the internal tripeptide, Ala-Ser-Gly. PAL is being explored as enzyme substitution therapy for Phenylketonuria (PKU), a disorder which involves an inability to metabolize phenylalanine. HAL failure in humans results in the disease histidinemia.
Probab=24.02  E-value=1.2e+02  Score=33.19  Aligned_cols=57  Identities=25%  Similarity=0.316  Sum_probs=43.7

Q ss_pred             HHHHHHHHHhHHHHHHHHHHhhhhccCCCCCCCCccChhHHHHHHHHHHHHHH-HHhcCC
Q 008451          160 IDMIGAACESAEKVLADTRKAYCFGTRQGPQILPTLDKGQALKIQEQENLLRA-AVNSGE  218 (565)
Q Consensus       160 Id~iNkacE~aekvIa~aRk~~e~gtRqGp~~~pT~dkadaaki~eqt~lL~A-AVn~Ge  218 (565)
                      ++.+++..+.+++.+++.+..|++-|=-|+....+++.++.+..+  .++|++ |++.|+
T Consensus        25 ~~ri~~s~~~l~~~~~~~~~iYGvnTG~G~~~d~~i~~~~~~~~q--~nLi~sha~GvG~   82 (444)
T cd00332          25 RERVDASRAALEEAAAEGKPVYGVNTGFGALADVRIDDADLRALQ--RNLLRSHAAGVGP   82 (444)
T ss_pred             HHHHHHHHHHHHHHHhcCCceeeeCCCCCCcCCcccCHHHHHHHH--HHHHHHHhcCCCC
Confidence            455667777788888889999999888887777778887766665  888775 566665


No 55 
>PF10239 DUF2465:  Protein of unknown function (DUF2465);  InterPro: IPR018797 FAM98A, B and C are glycine-rich proteins found from worms to humans whose function is unknown.
Probab=24.01  E-value=1.9e+02  Score=30.49  Aligned_cols=152  Identities=16%  Similarity=0.166  Sum_probs=75.5

Q ss_pred             HHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhHHHHhhhhcceE---eeeccCCCCCCccchhhhhccc-
Q 008451           50 SLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIVDEIRKVSKAFV---VHPKNVNAENATILPVMLSSKL-  125 (565)
Q Consensus        50 DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLvEEIkPvLr~Fv---VlPlnVnaeNa~IVPdmLRTKL-  125 (565)
                      |+.+.|..+...|..++.+++++.=+|=   +..+-.||.+++.|+.=+++.++   |-.+..+.+|.-.|=.||.|-| 
T Consensus        32 ~f~~L~~wL~~EL~~l~~leE~v~~~dd---~~~f~~Els~~L~El~CPy~~L~~G~~~~rl~~~~~~l~LL~fL~sELq  108 (318)
T PF10239_consen   32 EFRELCAWLASELKTLCKLEESVSSPDD---AESFLLELSGFLKELGCPYSALTSGDISDRLQSKEDRLLLLEFLCSELQ  108 (318)
T ss_pred             HHHHHHHHHHHHHHHHhccccccCCCch---HHHHHHHHHHHHHhcCCCcHHHcCCcchhhhcCHHHHHHHHHHHHHHHH
Confidence            4666666677777777766666663333   23345666667766654444333   2455566667777777776532 


Q ss_pred             ----------Cc-chhh-hhhHHHHHHHh--hcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhhccCCCCCC
Q 008451          126 ----------LP-EMEI-DDNSKREQLLL--GMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCFGTRQGPQI  191 (565)
Q Consensus       126 ----------lP-EmEt-ee~q~~~ql~~--kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~gtRqGp~~  191 (565)
                                .+ +.+. ++......+..  .+.++|.+...-...   ..++++...+++++++.    ..+..-.|-.
T Consensus       109 aarl~~~k~~~~~~~~~~~~s~~~~~l~~i~~~L~l~~p~~~i~~~---~lf~~i~~ki~~~L~~l----p~~~~~~PLl  181 (318)
T PF10239_consen  109 AARLLAKKKPEEPEQEEEKESEVAQELKAICQALGLPKPPPNITAS---QLFSKIEAKIEELLSKL----PPGHMGKPLL  181 (318)
T ss_pred             HHHHHHhccCCccccccccccHHHHHHHHHHHHhCCCCCCCCCCHH---HHHHHHHHHHHHHHHhc----CccccCCCCc
Confidence                      11 1111 02222222222  234443331111111   34455555555555432    2233444555


Q ss_pred             CCccChhHHHHHHHHHHHHH
Q 008451          192 LPTLDKGQALKIQEQENLLR  211 (565)
Q Consensus       192 ~pT~dkadaaki~eqt~lL~  211 (565)
                      ...++.++-++|++-...|.
T Consensus       182 ~~~L~~~Qw~~Le~i~~~L~  201 (318)
T PF10239_consen  182 KKSLTDEQWEKLEKINQALS  201 (318)
T ss_pred             CCCCCHHHHHHHHHHHHHHH
Confidence            55677777777766555554


No 56 
>PF15601 Imm42:  Immunity protein 42
Probab=23.45  E-value=2.4e+02  Score=26.76  Aligned_cols=55  Identities=24%  Similarity=0.196  Sum_probs=35.9

Q ss_pred             HHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhh--hh------hHHHHhhhhcceEeeeccC
Q 008451           48 AISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLEL--FN------IVDEIRKVSKAFVVHPKNV  109 (565)
Q Consensus        48 A~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL--~n------LvEEIkPvLr~FvVlPlnV  109 (565)
                      ..+|...-+-|.+.||.+.   .-.+||-+++.+  -.+.|  ..      .+++|+..|+.+-  |.-|
T Consensus        15 ~dfl~sFFsti~~~lE~~~---wGskfP~Lm~~L--Y~g~L~~~~~~~A~~eL~~I~~~l~~~~--p~~V   77 (134)
T PF15601_consen   15 PDFLHSFFSTISYRLENEG---WGSKFPLLMNEL--YRGYLRYEELEKALKELEEIRKELKKFP--PSEV   77 (134)
T ss_pred             HHHHHHHHHHHHHHhhccC---CCCcchHHHHHH--HcCCCCHHHHHHHHHHHHHHHHHHhcCC--hhhh
Confidence            4566677777788888876   566999999987  33333  22      2366666666654  4444


No 57 
>TIGR01225 hutH histidine ammonia-lyase. This enzyme deaminates histidine to urocanic acid, the first step in histidine degradation. It is closely related to phenylalanine ammonia-lyase.
Probab=23.24  E-value=1.2e+02  Score=33.98  Aligned_cols=56  Identities=21%  Similarity=0.388  Sum_probs=41.4

Q ss_pred             HHHHHHHHhHHHHHHHHHHhhhhccCCCCCCCCccChhHHHHHHHHHHHHHH-HHhcCC
Q 008451          161 DMIGAACESAEKVLADTRKAYCFGTRQGPQILPTLDKGQALKIQEQENLLRA-AVNSGE  218 (565)
Q Consensus       161 d~iNkacE~aekvIa~aRk~~e~gtRqGp~~~pT~dkadaaki~eqt~lL~A-AVn~Ge  218 (565)
                      +.+.+..+.+++++++.+..|+..|=-|+....+++.++.++++  ++||++ |++.|+
T Consensus        31 ~~i~~s~~~l~~~~~~g~~iYGvnTGfG~~~d~~i~~~~~~~lq--~nLi~sha~GvG~   87 (506)
T TIGR01225        31 EAVAKSRAAIEQIIAGDETVYGINTGFGKLASTRIDSEDLAELQ--RNLVRSHAAGVGD   87 (506)
T ss_pred             HHHHHHHHHHHHHHhcCCceeeecCCCCCccCcccCHHHHHHHH--HHHHHHHhcCCCC
Confidence            44556666677788899999999888887776777888766554  888875 566655


No 58 
>PF08172 CASP_C:  CASP C terminal;  InterPro: IPR012955 This domain is the C-terminal region of the CASP family of proteins. These are Golgi membrane proteins which are thought to have a role in vesicle transport [].; GO: 0006891 intra-Golgi vesicle-mediated transport, 0030173 integral to Golgi membrane
Probab=22.37  E-value=3.6e+02  Score=27.57  Aligned_cols=28  Identities=7%  Similarity=0.089  Sum_probs=24.1

Q ss_pred             hcHHHHHHHHHHHHHHHHHHHHhhhhcc
Q 008451           39 LNLEAVKTRAISLFKAISRILEDFDAYA   66 (565)
Q Consensus        39 lNLEAVraRA~DLkkaIsriI~~LE~e~   66 (565)
                      -.|+.+.+++.+.+..|.++..+|+...
T Consensus         6 ~~l~~l~~~~~~~~~L~~kLE~DL~~~~   33 (248)
T PF08172_consen    6 KELSELEAKLEEQKELNAKLENDLAKVQ   33 (248)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHHh
Confidence            4566899999999999999999998865


No 59 
>PF15368 BioT2:  Spermatogenesis family BioT2
Probab=22.29  E-value=2.1e+02  Score=28.35  Aligned_cols=64  Identities=22%  Similarity=0.310  Sum_probs=45.9

Q ss_pred             CccchhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhhccCCC
Q 008451          114 ATILPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCFGTRQG  188 (565)
Q Consensus       114 a~IVPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~gtRqG  188 (565)
                      -.|=|+-||+  -|-.|+        ++.-|.-+|... ++.+-..-++|.++..|+.-|++..-|+|+...--|
T Consensus        46 dkiEpMVLrs--PPTgES--------ivryALPIPssk-tkell~~de~irkitkhLkmvVstLEeTyG~~~~~g  109 (170)
T PF15368_consen   46 DKIEPMVLRS--PPTGES--------IVRYALPIPSSK-TKELLSEDEMIRKITKHLKMVVSTLEETYGLDIQNG  109 (170)
T ss_pred             cccccccccC--CCCchh--------HHHhhcCCCchh-hhhhhhHHHHHHHHHHHHHHHHHHHHHHhCcccccc
Confidence            3567888888  444444        333355554443 444567789999999999999999999999976666


No 60 
>PF05546 She9_MDM33:  She9 / Mdm33 family;  InterPro: IPR008839 Members of this family are mitochondrial inner membrane proteins with a role in inner mitochondrial membrane organisation and biogenesis []. The yeast Mdm33 protein assembles into an oligomeric complex in the inner membrane where it performs homotypic protein-protein interactions. It has been suggested that Mdm33 plays a distinct role, possibly involved in fission of the mitochondrial inner membrane [].
Probab=21.90  E-value=1.5e+02  Score=30.16  Aligned_cols=42  Identities=26%  Similarity=0.335  Sum_probs=35.5

Q ss_pred             hhcCCCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhh
Q 008451          142 LGMQNLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCF  183 (565)
Q Consensus       142 ~kA~nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~  183 (565)
                      ++.+.+.=-..||+|++.|+.....++.+-+-+.++|++|..
T Consensus        23 ~~lNd~TGYs~Ie~LK~~i~~~E~~l~~~r~~~~~aK~~Y~~   64 (207)
T PF05546_consen   23 QALNDVTGYSEIEKLKKSIEELEDELEAARQEVREAKAAYDD   64 (207)
T ss_pred             HHHHhccChHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            334445445689999999999999999999999999999987


No 61 
>PF09712 PHA_synth_III_E:  Poly(R)-hydroxyalkanoic acid synthase subunit (PHA_synth_III_E)
Probab=21.76  E-value=1.1e+02  Score=31.71  Aligned_cols=98  Identities=20%  Similarity=0.329  Sum_probs=53.4

Q ss_pred             HHHHHHHHHhhhhccccCCCC-Chhhhhhhhhhcchh-hhhhH--HHHhhhhcceEeeeccCCCCCCccchhhhhcccCc
Q 008451           52 FKAISRILEDFDAYARTNTTP-KWQDILGQYSMVNLE-LFNIV--DEIRKVSKAFVVHPKNVNAENATILPVMLSSKLLP  127 (565)
Q Consensus        52 kkaIsriI~~LE~e~~tN~t~-kWpDVLdqFSVIS~Q-L~nLv--EEIkPvLr~FvVlPlnVnaeNa~IVPdmLRTKLlP  127 (565)
                      .+++.++..+|.-..+.++++ .|.+|.|-+.-+.-+ +..++  +|-..+...++      ++.      .-|+.    
T Consensus       190 ~~a~~~~~~~l~~~~~~g~~~~s~re~~d~Wi~~ae~~~~~~~~S~ef~~~~g~~~------~a~------m~~r~----  253 (293)
T PF09712_consen  190 MKAFERMMEKLQERAEEGEQIKSWREFYDIWIDAAEEAYEELFRSEEFAQAYGQLV------NAL------MDLRK----  253 (293)
T ss_pred             HHHHHHHHHHHHHhhccCCCCcCHHHHHHHHHHHHHHHHHHHHCCHHHHHHHHHHH------HHH------HHHHH----
Confidence            677778888884333233444 588888876544322 22222  33333332221      110      01111    


Q ss_pred             chhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHHHHHHHHHhH
Q 008451          128 EMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRIDMIGAACESA  170 (565)
Q Consensus       128 EmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId~iNkacE~a  170 (565)
                          ...++.|. .-+..|||-...+|.+.|||..|.+-+..+
T Consensus       254 ----~~~~~~e~-~L~~l~lPTr~evd~l~k~l~eLrre~r~L  291 (293)
T PF09712_consen  254 ----QQQEVVEE-YLRSLNLPTRSEVDELYKRLHELRREVRAL  291 (293)
T ss_pred             ----HHHHHHHH-HHHHCCCCCHHHHHHHHHHHHHHHHHHHHh
Confidence                11222222 233468999999999999999998877654


No 62 
>PF04949 Transcrip_act:  Transcriptional activator;  InterPro: IPR007033 Golgins are a family of coiled-coil proteins associated with the Golgi apparatus necessary for tethering events in membrane fusion and as structural supports for Golgi cisternae []. This entry represents proteins annotated as RAB6-interacting golgins.
Probab=21.50  E-value=2.2e+02  Score=28.10  Aligned_cols=36  Identities=17%  Similarity=0.357  Sum_probs=25.3

Q ss_pred             CCCCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhh
Q 008451          146 NLPIPSQIEKLKSRIDMIGAACESAEKVLADTRKAY  181 (565)
Q Consensus       146 nLP~~~qiEklqKQId~iNkacE~aekvIa~aRk~~  181 (565)
                      .-|.-..++.+.|+||++|.-...+-...-+.-++|
T Consensus        79 ~dP~RkEv~~vRkkID~vNreLkpl~~~cqKKEkEy  114 (159)
T PF04949_consen   79 ADPMRKEVEMVRKKIDSVNRELKPLGQSCQKKEKEY  114 (159)
T ss_pred             ccchHHHHHHHHHHHHHHHHHhhHHHHHHHHHHHHH
Confidence            348889999999999999987555544444444444


No 63 
>PF08397 IMD:  IRSp53/MIM homology domain;  InterPro: IPR013606 The IMD (IRSp53 and MIM (missing in metastases) homology) domain is a BAR-like domain of approximately 250 amino acids found at the N-terminal in the insulin receptor tyrosine kinase substrate p53 (IRSp53) and in the evolutionarily related IRSp53/MIM family. In IRSp53, a ubiquitous regulator o the actin cytoskeleton, the IMD domain acts as conserved F-actin bundling domain involved in filopodium formation. Filopodium-inducing IMD activity is regulated by Cdc42 and Rac1 (Rho-family GTPases) and is SH3-independent [, , ]. The IRSp53/MIM family is a novel F-actin bundling protein family that includes invertebrate relatives:    Vertebrate MIM (missing in metastasis), an actin-binding scaffold protein that may be involved in cancer metastasis.  Vertebrate ABBA-1, a MIM-related protein. Vertebrate brain-specific angiogenesis inhibitor 1-associated protein 2 (BAI1-associated protein 2) or insulin receptor tyrosine kinase substrate p53 (IRSp53), a multifunctional adaptor protein that links Rac1 with a Wiskott-Aldrich syndrome family verprolin-homologous protein 2 (WAVE2) to induce lamellipodia or Cdc42 with Mena to induce filopodia [].  Vertebrate brain-specific angiogenesis inhibitor 1-associated protein 2-like proteins 1 and 2 (BAI1-associated protein 2-like proteins 1 and 2).  Drosophila melanogaster (Fruit fly) CG32082-PA.  Caenorhabditis elegans M04F3.5 protein.   The vertebrate IRSp53/MIM family is divided into two major groups: the IRSp53 subfamily and the MIM/ABBA subfamily. The putative invertebrate homologues are positioned between them. The IRSp53 subfamily members contain an SH3 domain, and the MIM/ABBA subfamily proteins contain a WH2 (WASP-homology 2) domain. The vertebrate SH3-containing subfamily is further divided into three groups according to the presence or absence of the WWB and the half-CRIB motif. The IMD domain can bind to and bundle actin filaments, bind to membranes and interact with the small GTPase Rac [, ].  The IMD domain folds as a coiled coil of three extended alpha-helices and a shorter C-terminal helix. Helix 4 packs tightly against the other three helices, and thus represents an integral part of the domain. The fold of the IMD domain closely resembles that of the BAR (Bin-Amphiphysin-RVS) domain, a functional module serving both as a sensor and inducer of membrane curvature []. The WH2 domain performs a scaffolding function [].; GO: 0008093 cytoskeletal adaptor activity, 0017124 SH3 domain binding, 0007165 signal transduction, 0046847 filopodium assembly; PDB: 2D1L_A 3OK8_B 1WDZ_B 1Y2O_A 2YKT_A.
Probab=21.40  E-value=52  Score=31.85  Aligned_cols=145  Identities=19%  Similarity=0.260  Sum_probs=73.0

Q ss_pred             hhhhcHHHHhhhcHHHHHHHHHHHHHHHHHHHHhhhhccc-----------cCCCCChhhhhhhhhhcchhhhhhHHHHh
Q 008451           28 VERLNQAVVQQLNLEAVKTRAISLFKAISRILEDFDAYAR-----------TNTTPKWQDILGQYSMVNLELFNIVDEIR   96 (565)
Q Consensus        28 ~e~ln~~V~~QlNLEAVraRA~DLkkaIsriI~~LE~e~~-----------tN~t~kWpDVLdqFSVIS~QL~nLvEEIk   96 (565)
                      +|.+||+..      .+.+.+..+.+++..++.....+.+           +.++-.=.+.|-+++.+=..+.+-.+++.
T Consensus         5 ~~~~~P~~e------~lv~~~~kY~~al~~~~~a~~~f~dal~ki~~~A~~s~~s~~lG~~L~~~s~~~r~i~~~~~~~~   78 (219)
T PF08397_consen    5 MEDFNPAWE------NLVSLGKKYQKALRAMSQAAAAFFDALQKIGDMASNSRGSKELGDALMQISEVHRRIENELEEVF   78 (219)
T ss_dssp             HHTHHHHHH------HHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHTSSSHHHHHHHHHHHHHHHHHHHHHHHHHH
T ss_pred             HhhcCHHHH------HHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHhccCCCccccHHHHHHHHHHHHHHHHHHHHHHH
Confidence            566675544      7777777777777766665555443           12233344566666666666655555555


Q ss_pred             hhhcceEeeeccCCCC-CCccchhhhhcccCcchhhhhhHHHHHHHh----------hcCC-CC-----CchhHHHHHHH
Q 008451           97 KVSKAFVVHPKNVNAE-NATILPVMLSSKLLPEMEIDDNSKREQLLL----------GMQN-LP-----IPSQIEKLKSR  159 (565)
Q Consensus        97 PvLr~FvVlPlnVnae-Na~IVPdmLRTKLlPEmEtee~q~~~ql~~----------kA~n-LP-----~~~qiEklqKQ  159 (565)
                      ..+..-.+.|+.-+-| ....|..     ++=+-+.+-+.+++.++-          +... -+     ....++.+..+
T Consensus        79 ~~~~~~li~pLe~~~e~d~k~i~~-----~~K~y~ke~k~~~~~l~K~~se~~Kl~KK~~kgk~~~~~~~~~~~~~v~~~  153 (219)
T PF08397_consen   79 KAFHSELIQPLEKKLEEDKKYITQ-----LEKDYEKEYKRKRDELKKAESELKKLRKKSRKGKDDQKYELKEALQDVTER  153 (219)
T ss_dssp             HHHHHHTHHHHHHHHHHHHHHHHH-----HHHHHHHHHHHHHHHHHHHHHHHHHHHCCCCCCTSCHHHHHHHHHHHHHHH
T ss_pred             HHHHHHHHHHHHHHHHHHHHHhhH-----HHHHHHHHHHHHHHHHHHHHHHHHHHhhcccCCCccccHHHHHHHHHHHHH
Confidence            5544444555543222 1111111     111122222333333332          2221 11     11124444555


Q ss_pred             HHHHHHHHH-hHHHHHHHHHHhhhh
Q 008451          160 IDMIGAACE-SAEKVLADTRKAYCF  183 (565)
Q Consensus       160 Id~iNkacE-~aekvIa~aRk~~e~  183 (565)
                      ...|...|. ++.+.+-+.|+-||+
T Consensus       154 ~~ele~~~~~~~r~al~EERrRyc~  178 (219)
T PF08397_consen  154 QSELEEFEKQSLREALLEERRRYCF  178 (219)
T ss_dssp             HHHHHHHHHHHHHHHHHHHHHHHHH
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            555666555 677788899999997


No 64 
>PF00221 Lyase_aromatic:  Aromatic amino acid lyase;  InterPro: IPR001106 This entry represents phenylalanine ammonia-lyase (PAL; 4.3.1.24 from EC) and the mechanistically related protein histidine ammonia lyase (HAL; 4.3.1.3 from EC). Both contain a catalytic Ala-Ser-Gly triad that is post-translationally cyclised []. PAL is a key biosynthetic catalyst in phenylpropanoid assembly in plants and fungi, and is involved in the biosynthesis of a wide variety of secondary metabolites such as flavanoids, furanocoumarin phytoalexins and cell wall components. These compounds are important for normal growth and in responses to environmental stress. HAL catalyses the first step in histidine degradation, the removal of an ammonia group from histidine to produce urocanic acid. The core domain in PAL and Hal share about 30% sequence identity, with PAL containing an additional approximately 160 residues extending from the common fold [].; GO: 0016841 ammonia-lyase activity, 0009058 biosynthetic process; PDB: 2RJR_A 2RJS_A 2OHY_B 3KDZ_B 2QVE_A 3KDY_A 2NYF_A 2YII_B 1Y2M_B 1T6P_B ....
Probab=21.38  E-value=91  Score=34.32  Aligned_cols=65  Identities=26%  Similarity=0.300  Sum_probs=42.8

Q ss_pred             CCchhHHHHHHHHHHHHHHHHhHHHHHHHHHHhhhhccCCCCCCCCccChhHHHHHHHHHHHHHH-HHhcCC
Q 008451          148 PIPSQIEKLKSRIDMIGAACESAEKVLADTRKAYCFGTRQGPQILPTLDKGQALKIQEQENLLRA-AVNSGE  218 (565)
Q Consensus       148 P~~~qiEklqKQId~iNkacE~aekvIa~aRk~~e~gtRqGp~~~pT~dkadaaki~eqt~lL~A-AVn~Ge  218 (565)
                      ++....|. .+||   .+..+-+++.+.+-+..|++-|=-|+....++++.+....+  .+||+. |++.|+
T Consensus        22 ~v~l~~~a-~~ri---~~sr~~l~~~~~~~~~iYGvnTG~G~~~~~~i~~~~~~~~q--~nll~~h~~gvG~   87 (473)
T PF00221_consen   22 KVELSPEA-RERI---EASRAFLEDILASGKPIYGVNTGFGALKDVRIPPEELAELQ--RNLLRSHAAGVGP   87 (473)
T ss_dssp             EEEE-HHH-HHHH---HHHHHHHHHHHHTTCTCTTTSBSSGGGTTSBC-GHHHHHHH--HHHHHHH---EEE
T ss_pred             cEEECHHH-HHHH---HHHHHHHHHHHhcCCceeccccCCccccCCcCCHHHHHHHH--HHHHHhhcccccc
Confidence            34445333 4554   45566677788888889999888887777777777765555  899988 888887


No 65 
>PTZ00365 60S ribosomal protein L7Ae-like; Provisional
Probab=21.22  E-value=9.4e+02  Score=25.51  Aligned_cols=102  Identities=17%  Similarity=0.238  Sum_probs=63.7

Q ss_pred             hcchhhhhhHHHHhhhhcceEeeeccCCCCCCcc-chhhhhcccCcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHHH
Q 008451           83 MVNLELFNIVDEIRKVSKAFVVHPKNVNAENATI-LPVMLSSKLLPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRID  161 (565)
Q Consensus        83 VIS~QL~nLvEEIkPvLr~FvVlPlnVnaeNa~I-VPdmLRTKLlPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQId  161 (565)
                      +|-..+..+...|+.---.+||+--.|+|..--. +|++++.+=.|-......+.+-....+-...-+... +--...-.
T Consensus       132 ~vk~Gin~VtklIekkKAkLVIIA~DVsP~t~kk~LP~LC~k~~VPY~iv~sK~eLG~AIGkktraVVAIt-dV~~EDk~  210 (266)
T PTZ00365        132 MLKYGLNHVTDLVEYKKAKLVVIAHDVDPIELVCFLPALCRKKEVPYCIIKGKSRLGKLVHQKTAAVVAID-NVRKEDQA  210 (266)
T ss_pred             HHHhhhHHHHHHHHhCCccEEEEeCCCCHHHHHHHHHHHHhccCCCEEEECCHHHHHHHhCCCCceEEEec-ccCHHHHH
Confidence            4556677777888887788999998888775544 599999999998888877743333332110000000 11112334


Q ss_pred             HHHHHHHhHHHHH---HHHHHhhhhcc
Q 008451          162 MIGAACESAEKVL---ADTRKAYCFGT  185 (565)
Q Consensus       162 ~iNkacE~aekvI---a~aRk~~e~gt  185 (565)
                      .++++|+.+...+   .+.|+.|+-|.
T Consensus       211 ~l~~lv~~~~~~~nd~~e~rr~wGG~~  237 (266)
T PTZ00365        211 EFDNLCKNFRAMFNDNSELRRRWGGGI  237 (266)
T ss_pred             HHHHHHHHHHHhccccHhhhhhcCCCc
Confidence            5666666666555   57788897654


No 66 
>PRK14154 heat shock protein GrpE; Provisional
Probab=21.21  E-value=78  Score=31.76  Aligned_cols=65  Identities=9%  Similarity=0.189  Sum_probs=44.8

Q ss_pred             chhhhhcHHHHhhhcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhHHHHhhhhcceE
Q 008451           26 VAVERLNQAVVQQLNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIVDEIRKVSKAFV  103 (565)
Q Consensus        26 ~~~e~ln~~V~~QlNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLvEEIkPvLr~Fv  103 (565)
                      |+++.|...      |+.++.++.+|+....|+.-+||.|-+  -+-+=-.-+..|++-+     +++++-|++++|-
T Consensus        52 ~~~~~l~~e------l~~le~e~~elkd~~lRl~ADfeNyRK--R~~kE~e~~~~~a~e~-----~~~~LLpVlDnLe  116 (208)
T PRK14154         52 PSREKLEGQ------LTRMERKVDEYKTQYLRAQAEMDNLRK--RIEREKADIIKFGSKQ-----LITDLLPVADSLI  116 (208)
T ss_pred             cchhhHHHH------HHHHHHHHHHHHHHHHHHHHHHHHHHH--HHHHHHHHHHHHHHHH-----HHHHHhhHHhHHH
Confidence            456666544      568999999999999999999999863  1111223345555544     8888888887763


No 67 
>PF07464 ApoLp-III:  Apolipophorin-III precursor (apoLp-III);  InterPro: IPR010009 This family consists of several insect apolipoprotein-III sequences. Exchangeable apolipoproteins constitute a functionally important family of proteins that play critical roles in lipid transport and lipoprotein metabolism. Apolipophorin III (apoLp-III) is a prototypical exchangeable apolipoprotein found in many insect species that functions in transport of diacylglycerol (DAG) from the fat body lipid storage depot to flight muscles in the adult life stage [].; GO: 0008289 lipid binding, 0006869 lipid transport, 0005576 extracellular region; PDB: 1EQ1_A.
Probab=20.81  E-value=1e+02  Score=29.49  Aligned_cols=52  Identities=27%  Similarity=0.353  Sum_probs=30.2

Q ss_pred             CcchhhhhhHHHHHHHhhcCCCCCchhHHHHHHHH----------------HHHHHHHHhHHHHHHHHHH
Q 008451          126 LPEMEIDDNSKREQLLLGMQNLPIPSQIEKLKSRI----------------DMIGAACESAEKVLADTRK  179 (565)
Q Consensus       126 lPEmEtee~q~~~ql~~kA~nLP~~~qiEklqKQI----------------d~iNkacE~aekvIa~aRk  179 (565)
                      -||+|....++++.+-++..+|  -...+++.+.|                .+|..++++++++.....+
T Consensus        83 ~Pev~~qa~~l~e~lQ~~vq~l--~~E~qk~~k~v~~~~~~~~e~l~~~~K~~~D~~~k~~~~~~~~l~~  150 (155)
T PF07464_consen   83 NPEVEKQANELQEKLQSAVQSL--VQESQKLAKEVSENSEGANEKLQPAIKQAYDDAVKAAQKVQKQLHE  150 (155)
T ss_dssp             SHHHHHT-SSSHHHHHHHHHHH--HHHHHHHHHHHHS---SS-GGGHHHHHHHHHHHHHHHHHHHHHHHH
T ss_pred             ChHHHHHHHHHHHHHHHHHHHH--HHHHHHHHHHHHHHHHhhhHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            5888888888888877776655  33333333332                3455555555555554443


No 68 
>PF05960 DUF885:  Bacterial protein of unknown function (DUF885);  InterPro: IPR010281 This family consists of hypothetical bacterial proteins.; PDB: 3O0Y_B 3U24_A 3IUK_A.
Probab=20.60  E-value=4.6e+02  Score=28.17  Aligned_cols=97  Identities=21%  Similarity=0.140  Sum_probs=48.8

Q ss_pred             hcHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhhhhhcchhhhhhHHHHhhhhcceEeeeccCCCCCCccch
Q 008451           39 LNLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQYSMVNLELFNIVDEIRKVSKAFVVHPKNVNAENATILP  118 (565)
Q Consensus        39 lNLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdqFSVIS~QL~nLvEEIkPvLr~FvVlPlnVnaeNa~IVP  118 (565)
                      ...+++..++..+++.+.++ ..++.     ..++=.+-++--.+.. .+....+...--.. .-.+|.+-...-+.-+|
T Consensus        20 ~S~~~~~~~~~~~~~~l~~L-~~id~-----~~Ls~~~~~~~~~l~~-~~~~~~~~~~~~~~-~~~~p~~~~~g~~~~l~   91 (549)
T PF05960_consen   20 LSPEAIEQRLAELRKLLKRL-EAIDR-----ASLSPEQQIDYDILRY-LLERLLEGDEFEYW-PYRYPLNYLSGAQAQLP   91 (549)
T ss_dssp             -SHHHHHHHHHHHHHHHHHH-HTS-G-----SC-SHHHHHHHHHHHH-HHHHHHHHHTCSHC-ECCGCS-SSSSHHHHHH
T ss_pred             CCHHHHHHHHHHHHHHHHHH-hCcCc-----ccCCHHHHHHHHHHHH-HHHHHHHHhhcccc-cccCCccccccHHHHHH
Confidence            35677888888888877665 55555     3445455555444444 33333322221111 24555555555566677


Q ss_pred             hhhhcccCcchhhhhhHHHHHHHhhcCCC
Q 008451          119 VMLSSKLLPEMEIDDNSKREQLLLGMQNL  147 (565)
Q Consensus       119 dmLRTKLlPEmEtee~q~~~ql~~kA~nL  147 (565)
                      ..|....-++-+.+...    +...+..+
T Consensus        92 ~~l~~~~~~~~~~d~~~----~~~rL~~i  116 (549)
T PF05960_consen   92 SLLSEYHPFETEEDAED----YLARLAAI  116 (549)
T ss_dssp             HHHCHTS-HSSHHHHHH----HHHHHTTH
T ss_pred             HHHHhcCCCCCHHHHHH----HHHHHHHH
Confidence            77766534444444444    44444444


No 69 
>COG4453 Uncharacterized protein conserved in bacteria [Function unknown]
Probab=20.57  E-value=88  Score=28.27  Aligned_cols=17  Identities=47%  Similarity=0.685  Sum_probs=15.5

Q ss_pred             HHHHHHHhHHHHHHHHH
Q 008451          162 MIGAACESAEKVLADTR  178 (565)
Q Consensus       162 ~iNkacE~aekvIa~aR  178 (565)
                      ++++||++|++||.+.|
T Consensus        40 vl~aA~~~A~~vi~~~~   56 (95)
T COG4453          40 VLSAALEAAEDVIEDQR   56 (95)
T ss_pred             HHHHHHHHHHHHHHhhH
Confidence            68999999999999877


No 70 
>PLN03097 FHY3 Protein FAR-RED ELONGATED HYPOCOTYL 3; Provisional
Probab=20.51  E-value=2.1e+02  Score=34.06  Aligned_cols=43  Identities=23%  Similarity=0.411  Sum_probs=34.8

Q ss_pred             ChhhhhhhhhhcchhhhhhHHHHhhhhcceEeeeccCCCCCCccchhhhhcccCcchhhh
Q 008451           73 KWQDILGQYSMVNLELFNIVDEIRKVSKAFVVHPKNVNAENATILPVMLSSKLLPEMEID  132 (565)
Q Consensus        73 kWpDVLdqFSVIS~QL~nLvEEIkPvLr~FvVlPlnVnaeNa~IVPdmLRTKLlPEmEte  132 (565)
                      .|.++|+.|.+-..+....+.|++.-                 .+|.|++....+.|-+.
T Consensus       417 ~W~~mi~ky~L~~n~WL~~LY~~Rek-----------------WapaY~k~~F~agm~sT  459 (846)
T PLN03097        417 RWWKILDRFELKEDEWMQSLYEDRKQ-----------------WVPTYMRDAFLAGMSTV  459 (846)
T ss_pred             HHHHHHHhhcccccHHHHHHHHhHhh-----------------hhHHHhcccccCCcccc
Confidence            89999999999988887777777643                 48889998888888443


No 71 
>KOG2196 consensus Nuclear porin [Nuclear structure]
Probab=20.49  E-value=2.5e+02  Score=29.43  Aligned_cols=56  Identities=11%  Similarity=0.230  Sum_probs=38.1

Q ss_pred             cHHHHHHHHHHHHHHHHHHHHhhhhccccCCCCChhhhhhh-----------hhhcchhhhhhHHHHhhhhcce
Q 008451           40 NLEAVKTRAISLFKAISRILEDFDAYARTNTTPKWQDILGQ-----------YSMVNLELFNIVDEIRKVSKAF  102 (565)
Q Consensus        40 NLEAVraRA~DLkkaIsriI~~LE~e~~tN~t~kWpDVLdq-----------FSVIS~QL~nLvEEIkPvLr~F  102 (565)
                      +|+-|-.--++|.+.+..+..+++.-       -|..+|..           -.-|..+|..+.+|++.+.++.
T Consensus       135 ~L~~I~sqQ~ELE~~L~~lE~k~~~~-------~g~~~~~~~D~eR~qty~~a~nidsqLk~l~~dL~~ii~~l  201 (254)
T KOG2196|consen  135 ELEFILSQQQELEDLLDPLETKLELQ-------SGHTYLSRADVEREQTYKMAENIDSQLKRLSEDLKQIIKSL  201 (254)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHhcc-------ccchhhhhhhHHHHHHHHHHHHHHHHHHHHHhhHHHHHHHH
Confidence            57788899999999999888888762       34333333           2335566677777777666553


No 72 
>PRK10869 recombination and repair protein; Provisional
Probab=20.49  E-value=3.2e+02  Score=30.43  Aligned_cols=34  Identities=18%  Similarity=0.297  Sum_probs=17.0

Q ss_pred             ChhhhhhhhhhcchhhhhhHHHHhhhhcceEeee
Q 008451           73 KWQDILGQYSMVNLELFNIVDEIRKVSKAFVVHP  106 (565)
Q Consensus        73 kWpDVLdqFSVIS~QL~nLvEEIkPvLr~FvVlP  106 (565)
                      .+.++++...=+.-+|..+.+|+...+..+.+-|
T Consensus       262 ~~~~~~~~l~~~~~~l~~~~~~l~~~~~~~~~dp  295 (553)
T PRK10869        262 KLSGVLDMLEEALIQIQEASDELRHYLDRLDLDP  295 (553)
T ss_pred             hHHHHHHHHHHHHHHHHHHHHHHHHHHhhcCCCH
Confidence            4444555555555555555555555555443333


No 73 
>PF02050 FliJ:  Flagellar FliJ protein;  InterPro: IPR012823 Many flagellar proteins are exported by a flagellum-specific export pathway. Attempts have been made to characterise the apparatus responsible for this process, by designing assays to screen for mutants with export defects []. Experiments involving filament removal from temperature-sensitive flagellar mutants of Salmonella typhimurium have shown that, while most mutants were able to regrow filaments, flhA, fliH, fliI and fliN mutants showed no or greatly reduced regrowth. This suggests that the corresponding gene products are involved in the process of flagellum-specific export. The sequences of fliH, fliI and the adjacent gene, fliJ, have been deduced. FliJ was shown to encode a protein of molecular mass 17,302 Da []. It is a membrane-associated protein that affects chemotactic events, mutations in FliJ result in failure to respond to chemotactic stimuli.; GO: 0003774 motor activity, 0001539 ciliary or flagellar motility, 0006935 chemotaxis, 0009288 bacterial-type flagellum, 0016020 membrane, 0044461 bacterial-type flagellum part; PDB: 3AJW_A.
Probab=20.46  E-value=2.1e+02  Score=23.21  Aligned_cols=29  Identities=17%  Similarity=0.187  Sum_probs=14.4

Q ss_pred             HHHHHHHHHHHHHHHhHHHHHHHHHHhhh
Q 008451          154 EKLKSRIDMIGAACESAEKVLADTRKAYC  182 (565)
Q Consensus       154 EklqKQId~iNkacE~aekvIa~aRk~~e  182 (565)
                      +.+...|+...+.++.+++.+..+|+.|-
T Consensus        55 ~~l~~~i~~~~~~~~~~~~~~~~~r~~l~   83 (123)
T PF02050_consen   55 SALEQAIQQQQQELERLEQEVEQAREELQ   83 (123)
T ss_dssp             HHHHHHHHHHHHHHHHHHHHHHHHHHHHH
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            33444555555555555555555555543


No 74 
>PRK10280 dipeptidyl carboxypeptidase II; Provisional
Probab=20.45  E-value=1e+03  Score=27.58  Aligned_cols=20  Identities=10%  Similarity=0.134  Sum_probs=14.4

Q ss_pred             CCChhhhhhhhhhcchhhhh
Q 008451           71 TPKWQDILGQYSMVNLELFN   90 (565)
Q Consensus        71 t~kWpDVLdqFSVIS~QL~n   90 (565)
                      .++|.+++.-+.-++..|..
T Consensus        52 ~~t~~n~i~~ld~~~~~l~~   71 (681)
T PRK10280         52 APDFNNTILALEQSGELLTR   71 (681)
T ss_pred             CCCHHHHHHHHHHHHHHHHH
Confidence            45899998888777755533


No 75 
>PF05794 Tcp11:  T-complex protein 11;  InterPro: IPR008862 This family consists of several eukaryotic T-complex protein 11 (Tcp11) related sequences. Tcp11 is only expressed in fertile adult mammalian testes and is thought to be important in sperm function and fertility. The family also contains the Saccharomyces cerevisiae Sok1 protein which is known to suppress cyclic AMP-dependent protein kinase mutants [].
Probab=20.33  E-value=3.1e+02  Score=28.82  Aligned_cols=35  Identities=17%  Similarity=0.240  Sum_probs=23.7

Q ss_pred             chhhhhhHHHHhhhhcceEeeeccCCCCCCccchhhhhcccCcch
Q 008451           85 NLELFNIVDEIRKVSKAFVVHPKNVNAENATILPVMLSSKLLPEM  129 (565)
Q Consensus        85 S~QL~nLvEEIkPvLr~FvVlPlnVnaeNa~IVPdmLRTKLlPEm  129 (565)
                      -..+..|.+||+..|..++          ++-.-.....-||+|+
T Consensus        51 ~~~~~~Ll~~ike~L~~ll----------~~~~~~~I~e~LD~~l   85 (441)
T PF05794_consen   51 YSRLPQLLEEIKEILLSLL----------PSRLRQEIEEVLDLEL   85 (441)
T ss_pred             hhHHHHHHHHHHHHHHHhc----------CHHHHHHHHHHCChHH
Confidence            3456678899999988887          2333455666666665


No 76 
>PF05597 Phasin:  Poly(hydroxyalcanoate) granule associated protein (phasin);  InterPro: IPR008769 Polyhydroxyalkanoates (PHAs) are storage polyesters synthesised by various bacteria as intracellular carbon and energy reserve material. PHAs are accumulated as water-insoluble inclusions within the cells. This family consists of the phasins PhaF and PhaI which act as a transcriptional regulator of PHA biosynthesis genes. PhaF has been proposed to repress expression of the phaC1 gene and the phaIF operon.
Probab=20.30  E-value=1.2e+02  Score=28.53  Aligned_cols=26  Identities=31%  Similarity=0.489  Sum_probs=22.9

Q ss_pred             CCCCchhHHHHHHHHHHHHHHHHhHH
Q 008451          146 NLPIPSQIEKLKSRIDMIGAACESAE  171 (565)
Q Consensus       146 nLP~~~qiEklqKQId~iNkacE~ae  171 (565)
                      ++|-...+|+|.+|||.|++.++.+.
T Consensus       104 gvPs~~dv~~L~~rId~L~~~v~~l~  129 (132)
T PF05597_consen  104 GVPSRKDVEALSARIDQLTAQVERLA  129 (132)
T ss_pred             CCCCHHHHHHHHHHHHHHHHHHHHHh
Confidence            56888899999999999999988764


No 77 
>smart00721 BAR BAR domain.
Probab=20.28  E-value=2.2e+02  Score=26.52  Aligned_cols=26  Identities=19%  Similarity=0.266  Sum_probs=17.2

Q ss_pred             HHHHHHHHHHHHHHHHHHHHhhhhcc
Q 008451           41 LEAVKTRAISLFKAISRILEDFDAYA   66 (565)
Q Consensus        41 LEAVraRA~DLkkaIsriI~~LE~e~   66 (565)
                      ++.+..|++++++.+.+|+.+.+.|-
T Consensus        29 f~~le~~~~~~~~~~~kl~k~~~~y~   54 (239)
T smart00721       29 FEELERRFDTTEAEIEKLQKDTKLYL   54 (239)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHc
Confidence            55666666777777776666666664


Done!