Query         psy16860
Match_columns 139
No_of_seqs    123 out of 1420
Neff          9.1 
Searched_HMMs 46136
Date          Fri Aug 16 23:55:19 2013
Command       hhsearch -i /work/01045/syshi/Psyhhblits/psy16860.a3m -d /work/01045/syshi/HHdatabase/Cdd.hhm -o /work/01045/syshi/hhsearch_cdd/16860hhsearch_cdd -cpu 12 -v 0 

 No Hit                             Prob E-value P-value  Score    SS Cols Query HMM  Template HMM
  1 KOG4219|consensus               99.8 4.2E-19 9.2E-24  134.2   6.4  103   36-139    30-132 (423)
  2 PHA03234 DNA packaging protein  99.8 1.3E-17 2.8E-22  126.6  11.5   99   38-139    29-129 (338)
  3 PHA02834 chemokine receptor-li  99.7 1.3E-15 2.9E-20  114.8  10.8   95   41-139    28-122 (323)
  4 PHA02638 CC chemokine receptor  99.6 6.6E-14 1.4E-18  109.1  12.8   96   40-139    97-192 (417)
  5 PHA03235 DNA packaging protein  99.6 3.7E-14 8.1E-19  110.2  11.2   99   38-139    29-129 (409)
  6 KOG4220|consensus               99.5 6.2E-16 1.3E-20  117.6  -0.4   98   41-139    30-127 (503)
  7 PHA03087 G protein-coupled che  99.5 4.9E-14 1.1E-18  106.4   9.6   98   39-139    38-135 (335)
  8 PF00001 7tm_1:  7 transmembran  99.3 2.4E-12 5.2E-17   91.8   6.8   81   58-139     1-81  (257)
  9 PF10320 7TM_GPCR_Srsx:  Serpen  98.5 7.1E-08 1.5E-12   70.9   3.5   75   53-129     2-76  (257)
 10 KOG2087|consensus               98.2 1.1E-06 2.4E-11   66.6   3.1   92   42-134    25-124 (363)
 11 PF05296 TAS2R:  Mammalian tast  97.9 0.00044 9.5E-09   52.1  11.5   95   40-134     5-102 (303)
 12 PF05462 Dicty_CAR:  Slime mold  97.8  0.0003 6.6E-09   53.0   9.5   87   43-133     8-94  (303)
 13 PF11710 Git3:  G protein-coupl  97.7 0.00064 1.4E-08   48.3   9.8   60   70-129    30-89  (201)
 14 PF10328 7TM_GPCR_Srx:  Serpent  96.8   0.012 2.6E-07   43.4   8.4   45   51-95      3-47  (274)
 15 PF10324 7TM_GPCR_Srw:  Serpent  96.4    0.02 4.4E-07   42.9   7.4   51   51-102     6-57  (318)
 16 PF10321 7TM_GPCR_Srt:  Serpent  96.1   0.065 1.4E-06   40.7   8.7   75   39-119    30-105 (313)
 17 PF03402 V1R:  Vomeronasal orga  96.0   0.018   4E-07   42.6   5.3   61   70-132     5-66  (265)
 18 PF10317 7TM_GPCR_Srd:  Serpent  94.6    0.27 5.8E-06   36.7   7.9   50   47-96      4-54  (292)
 19 PF00002 7tm_2:  7 transmembran  87.9    0.28 6.2E-06   35.2   1.4   75   53-130    12-88  (242)
 20 PF10292 7TM_GPCR_Srab:  Serpen  84.1      16 0.00035   27.7   9.8   97   42-139    17-120 (324)
 21 PF02101 Ocular_alb:  Ocular al  78.9      19 0.00042   28.3   8.0   95   43-137    28-140 (405)
 22 KOG4564|consensus               75.2      43 0.00093   27.2  12.2   84   48-131   151-246 (473)
 23 PF10327 7TM_GPCR_Sri:  Serpent  67.6      21 0.00045   26.9   5.8   64   42-105     9-76  (303)
 24 PF09882 DUF2109:  Predicted me  66.7      23  0.0005   21.2   4.6   47   52-98      4-50  (78)
 25 PF10316 7TM_GPCR_Srbc:  Serpen  66.6      32 0.00069   25.7   6.5   61   42-102     6-66  (273)
 26 PF01102 Glycophorin_A:  Glycop  57.2      15 0.00032   24.1   2.9   11   61-71     83-93  (122)
 27 PF10323 7TM_GPCR_Srv:  Serpent  57.2      46   0.001   24.7   6.0   41   56-96      9-53  (283)
 28 PF11446 DUF2897:  Protein of u  53.7      20 0.00042   20.0   2.7   22   46-67      6-27  (55)
 29 TIGR01477 RIFIN variant surfac  52.8      20 0.00044   27.8   3.4   31   44-74    310-340 (353)
 30 PTZ00046 rifin; Provisional     50.0      24 0.00052   27.5   3.5   30   45-74    316-345 (358)
 31 PF06024 DUF912:  Nucleopolyhed  46.0      29 0.00064   21.7   2.9   11   63-73     83-93  (101)
 32 PF02009 Rifin_STEVOR:  Rifin/s  45.6      23  0.0005   26.9   2.8   27   47-73    259-285 (299)
 33 PF02532 PsbI:  Photosystem II   42.4      42 0.00091   16.9   2.5   18   42-59      9-26  (36)
 34 PF15330 SIT:  SHP2-interacting  42.2      53  0.0011   21.0   3.7   28   44-71      3-30  (107)
 35 KOG4193|consensus               40.6   2E+02  0.0043   24.3   7.6   63   60-130   338-402 (610)
 36 KOG2927|consensus               39.9      95  0.0021   24.3   5.3    9  105-113   255-263 (372)
 37 PF10326 7TM_GPCR_Str:  Serpent  39.8      28  0.0006   25.9   2.5   48   49-96      6-54  (307)
 38 COG1230 CzcD Co/Zn/Cd efflux s  37.7 1.8E+02   0.004   22.1   9.3   70   47-119   126-196 (296)
 39 PRK03557 zinc transporter ZitB  36.8 1.9E+02   0.004   21.9  10.0   73   50-123   126-198 (312)
 40 PF08114 PMP1_2:  ATPase proteo  35.6      13 0.00029   19.4   0.1   24   48-71     11-34  (43)
 41 cd07912 Tweety_N N-terminal do  34.5 2.4E+02  0.0053   22.6   8.4   15   76-90     79-93  (418)
 42 PF12606 RELT:  Tumour necrosis  33.2      51  0.0011   18.0   2.2   23   47-71      6-28  (50)
 43 PF10873 DUF2668:  Protein of u  30.1      59  0.0013   22.0   2.5   32   43-74     63-94  (155)
 44 KOG1482|consensus               29.9 1.2E+02  0.0026   23.9   4.5   78   53-130   184-279 (379)
 45 PF10319 7TM_GPCR_Srj:  Serpent  28.6 1.2E+02  0.0025   23.4   4.2   48   47-94     10-58  (310)
 46 KOG4349|consensus               26.1   2E+02  0.0043   19.0   7.8   21  108-128   114-134 (143)
 47 PF05545 FixQ:  Cbb3-type cytoc  26.1 1.1E+02  0.0025   16.1   3.8    9   61-69     25-33  (49)
 48 PF05961 Chordopox_A13L:  Chord  25.9 1.1E+02  0.0024   17.8   2.9   13   61-73     17-29  (68)
 49 CHL00024 psbI photosystem II p  24.5      34 0.00074   17.3   0.5   17   43-59     10-26  (36)
 50 PF06679 DUF1180:  Protein of u  24.1 2.5E+02  0.0054   19.4   4.9   26   42-67     93-118 (163)
 51 PRK13664 hypothetical protein;  24.0 1.1E+02  0.0023   17.3   2.5   17   47-63     10-26  (62)
 52 PRK02655 psbI photosystem II r  22.1      38 0.00083   17.2   0.4   16   43-58     10-25  (38)
 53 PF10329 DUF2417:  Region of un  21.4 2.2E+02  0.0047   20.9   4.3   15   76-90    104-118 (232)
 54 PHA03049 IMV membrane protein;  20.8 1.7E+02  0.0037   17.0   2.9   17   53-71     11-27  (68)

No 1  
>KOG4219|consensus
Probab=99.77  E-value=4.2e-19  Score=134.25  Aligned_cols=103  Identities=31%  Similarity=0.627  Sum_probs=96.3

Q ss_pred             cchHHHHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccch
Q psy16860         36 IHIFPFIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCD  115 (139)
Q Consensus        36 ~~~~~~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~  115 (139)
                      ...+.+.+...++.++.+++++||.+|+|++..+|++|+.+|+|++|||++|++.+++..|+.........|.+|...|+
T Consensus        30 lp~~~~~~wai~yg~l~~vAv~GN~iVlwIil~hrrMRtvtnyfL~NLAfADl~~s~Fn~~f~f~yal~~~W~~G~f~C~  109 (423)
T KOG4219|consen   30 LPAWQQALWAIAYGLLVFVAVVGNLIVLWIILAHRRMRTVTNYFLVNLAFADLSMSIFNTVFNFQYALHQEWYFGSFYCR  109 (423)
T ss_pred             CCHHHHHHHHHHHHHHHHHHHhcCceEEEEEeehhehhhhHHHHHHHHHHHHHHHHHHhhHHHHHHHHHhccccccceee
Confidence            34555889999999999999999999999999999999999999999999999999999999988888899999999999


Q ss_pred             hchHHhHHhhHHHHHHHHHhhccC
Q psy16860        116 VWNSFDVYFSQYRKHFNMYEVDFE  139 (139)
Q Consensus       116 ~~~~~~~~~~~~Si~~~m~~~~~~  139 (139)
                      +..|+......+|+ ++|+|+++|
T Consensus       110 f~nf~~itav~vSV-fTlvAiA~D  132 (423)
T KOG4219|consen  110 FVNFFPITAVFVSV-FTLVAIAID  132 (423)
T ss_pred             eccccchhhhhHhH-HHHHHHHHH
Confidence            99999999999999 788888875


No 2  
>PHA03234 DNA packaging protein UL33; Provisional
Probab=99.75  E-value=1.3e-17  Score=126.62  Aligned_cols=99  Identities=14%  Similarity=0.165  Sum_probs=82.6

Q ss_pred             hHHHHHHHHHHHHHHHHHHHHHHHhHHHh--hccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccch
Q psy16860         38 IFPFIIKGMLMCFIIITAILGNLLVIISV--IKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCD  115 (139)
Q Consensus        38 ~~~~~~~~~~~~~i~~~g~~gN~lvi~v~--~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~  115 (139)
                      +..+.....+|.+++++|++||+++++++  .+++++|+++|+|+.|||++|++.++. .|+.... ..+.|.+|+..||
T Consensus        29 ~~~~~~~~~~y~~vf~~gl~gN~lvl~v~~~~~~~~~rt~tn~fi~NLAvaDLL~~l~-lp~~~~~-~~~~w~fG~~lCk  106 (338)
T PHA03234         29 KKAQILESAINGIMLTLIIPMIIIVICTLIIYHKVAKHNATSFYLITLFASDFLHMLC-VFFLTLN-REALFNFNQAFCQ  106 (338)
T ss_pred             HHHHHHhhHHHHHHHHHHhhhHHHHHHHHHHHhccccccHHHHHHHHHHHHHHHHHHH-HHHHHHH-HhCCccCchhHHH
Confidence            34478889999999999999999999955  456677999999999999999999765 5555443 3457999999999


Q ss_pred             hchHHhHHhhHHHHHHHHHhhccC
Q psy16860        116 VWNSFDVYFSQYRKHFNMYEVDFE  139 (139)
Q Consensus       116 ~~~~~~~~~~~~Si~~~m~~~~~~  139 (139)
                      +..++.....++|+ +++.++|+|
T Consensus       107 ~~~~~~~~~~~~Si-~~L~~ISiD  129 (338)
T PHA03234        107 CVLFIYHASCSYSI-CMLAIIATI  129 (338)
T ss_pred             HHHHHHHHHHHHHH-HHHHHHHHH
Confidence            99999999999999 557777765


No 3  
>PHA02834 chemokine receptor-like protein; Provisional
Probab=99.65  E-value=1.3e-15  Score=114.85  Aligned_cols=95  Identities=23%  Similarity=0.426  Sum_probs=80.4

Q ss_pred             HHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHH
Q psy16860         41 FIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSF  120 (139)
Q Consensus        41 ~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~  120 (139)
                      +.....++.+++++|++||+++++++.++|+ +++.|+|+.|||++|++. .+..|+.+.... ++|.+|+..|++..+.
T Consensus        28 ~~~~~~~~~li~v~~~~gN~lVi~vi~~~~~-~~~~n~~i~nLAiaDll~-~~~lP~~i~~~~-~~w~~g~~~C~~~~~~  104 (323)
T PHA02834         28 NYFVIVFYILLFIFGLIGNVLVIAVLIVKRF-MFVVDVYLFNIAMSDLML-VFSFPFIIHNDL-NEWIFGEFMCKLVLGV  104 (323)
T ss_pred             hhhHHHHHHHHHHHHHhhHHHHHHHHHhccc-cchhhhhhHHHHHHHHHH-HHHHHHHHHHHc-CCcCCcchHHHhHHHH
Confidence            5577889999999999999999999887665 457899999999999987 667898765544 4799999999999998


Q ss_pred             hHHhhHHHHHHHHHhhccC
Q psy16860        121 DVYFSQYRKHFNMYEVDFE  139 (139)
Q Consensus       121 ~~~~~~~Si~~~m~~~~~~  139 (139)
                      ......+|+ +++.++|+|
T Consensus       105 ~~~~~~~Si-~tL~~Isid  122 (323)
T PHA02834        105 YFVGFFSNM-FFVTLISID  122 (323)
T ss_pred             HHHHHHHHH-HHHHHHHHH
Confidence            888888888 678888776


No 4  
>PHA02638 CC chemokine receptor-like protein; Provisional
Probab=99.57  E-value=6.6e-14  Score=109.10  Aligned_cols=96  Identities=23%  Similarity=0.389  Sum_probs=79.5

Q ss_pred             HHHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchH
Q psy16860         40 PFIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNS  119 (139)
Q Consensus        40 ~~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~  119 (139)
                      .......++.+++++|++||+++++++.+ |++|+++|++++|||++|++. ++..|+.+... .+.|.+|+..||+..+
T Consensus        97 ~~~~l~~~y~lvfvlgliGN~LVl~il~~-k~lrt~t~i~llnLAisDLl~-~l~lPf~i~~~-~~~W~fg~~~Ck~~~~  173 (417)
T PHA02638         97 ISEYIKIFYIIIFILGLFGNAAIIMILFC-KKIKTITDIYIFNLAISDLIF-VIDFPFIIYNE-FDQWIFGDFMCKVISA  173 (417)
T ss_pred             hhhHHHHHHHHHHHHHHHHHHHHHHHHHh-ccCCCHhHHHHHHHHHHHHHH-HHHHHHHHHHH-hccccccccchhhHHH
Confidence            35677888999999999999999987654 778999999999999999988 55788877654 4679999999999999


Q ss_pred             HhHHhhHHHHHHHHHhhccC
Q psy16860        120 FDVYFSQYRKHFNMYEVDFE  139 (139)
Q Consensus       120 ~~~~~~~~Si~~~m~~~~~~  139 (139)
                      +.....++|++ .+.++++|
T Consensus       174 l~~~~~~~Si~-~L~~isiD  192 (417)
T PHA02638        174 SYYIGFFSNMF-LITLMSID  192 (417)
T ss_pred             HHHHHHHHHHH-HHHHHHHH
Confidence            99888888875 55555543


No 5  
>PHA03235 DNA packaging protein UL33; Provisional
Probab=99.56  E-value=3.7e-14  Score=110.20  Aligned_cols=99  Identities=14%  Similarity=0.085  Sum_probs=76.3

Q ss_pred             hHHHHHHHHHHHHHHHHHHHHHHHhHHHhhcc-CC-CCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccch
Q psy16860         38 IFPFIIKGMLMCFIIITAILGNLLVIISVIKH-RK-LRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCD  115 (139)
Q Consensus        38 ~~~~~~~~~~~~~i~~~g~~gN~lvi~v~~~~-~~-~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~  115 (139)
                      +..+.+...++.+++++|++||+++++++.++ |+ .++..++|+.|||++|++. ++.+|+.+... ...|..|+..||
T Consensus        29 ~~~~~~~~~~~~li~vvGiigN~lVL~~~~~~~r~~~~~~~~~~I~NLAvsDLl~-l~~lP~~i~~~-~~~~~~g~~~Ck  106 (409)
T PHA03235         29 SAARTTETFINLLIISVGGPLNLIVLVTQLLANRVHGFSTPTLYMTNLYLANLLT-VFVLPFIMLSN-QGLLSGSVAGCK  106 (409)
T ss_pred             hhhHhHHHHHHHHHHHHHHHHHHHHHHHHHHhhhcccCCccHHHHHHHHHHHHHH-HHHHHHHHHhc-CccccCCCCeeh
Confidence            44578899999999999999999999986533 32 2356679999999999987 66788776432 122334578999


Q ss_pred             hchHHhHHhhHHHHHHHHHhhccC
Q psy16860        116 VWNSFDVYFSQYRKHFNMYEVDFE  139 (139)
Q Consensus       116 ~~~~~~~~~~~~Si~~~m~~~~~~  139 (139)
                      +..++....+.+|+ .++.++++|
T Consensus       107 ~~~~l~~~~~~~Si-~tL~~ISiD  129 (409)
T PHA03235        107 FASLLYYASCTVGF-ATVALIAAD  129 (409)
T ss_pred             hHHHHHHHHHHHHH-HHHHHHHHH
Confidence            99999999999998 567777765


No 6  
>KOG4220|consensus
Probab=99.54  E-value=6.2e-16  Score=117.61  Aligned_cols=98  Identities=30%  Similarity=0.569  Sum_probs=90.4

Q ss_pred             HHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHH
Q psy16860         41 FIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSF  120 (139)
Q Consensus        41 ~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~  120 (139)
                      -...+++..++.++.++||++|++.+...|++++..|||+++||++|++++.+.+|+...+.+-|+|.+|...|.+.-.+
T Consensus        30 ~v~i~~v~~~lsLVTv~GNlLVmiSfKvnrqLqTVnNYfLfSLAcADliIG~~SMnl~t~Y~lmg~W~LG~~~CdlWLal  109 (503)
T KOG4220|consen   30 VVFIVVVTGSLSLVTVVGNLLVMISFKVNRQLQTVNNYFLFSLACADLIIGAFSMNLYTTYTLMGYWPLGPLVCDLWLAL  109 (503)
T ss_pred             EEeeehhhhHHHHHhhhccEEEEEEEEecceeeeecceeehHHHHhhhhhheeechHHHHHHHHcccccchHHHHHHHHH
Confidence            34566677888999999999999999999999999999999999999999999999999999999999999999999999


Q ss_pred             hHHhhHHHHHHHHHhhccC
Q psy16860        121 DVYFSQYRKHFNMYEVDFE  139 (139)
Q Consensus       121 ~~~~~~~Si~~~m~~~~~~  139 (139)
                      .++...+|+ +|+-.++||
T Consensus       110 DYvaSNASV-mNLLiISFD  127 (503)
T KOG4220|consen  110 DYVASNASV-MNLLIISFD  127 (503)
T ss_pred             HHHhhhhhh-hhhheeeee
Confidence            999999999 677777775


No 7  
>PHA03087 G protein-coupled chemokine receptor-like protein; Provisional
Probab=99.54  E-value=4.9e-14  Score=106.43  Aligned_cols=98  Identities=21%  Similarity=0.403  Sum_probs=83.0

Q ss_pred             HHHHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhch
Q psy16860         39 FPFIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWN  118 (139)
Q Consensus        39 ~~~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~  118 (139)
                      ..+.+...++.+++++|++||+++++++.++ ++|++.|+++.|||++|++.++ ..|........++|.+|+..|+...
T Consensus        38 ~~~~~~~~~~~~i~~~gl~gN~lvl~~~~~~-~~~~~~~~ll~~laisDll~~~-~~~~~~~~~~~~~~~~~~~~C~~~~  115 (335)
T PHA03087         38 TNSTILIVVYSTIFFFGLVGNIIVIYVLTKT-KIKTPMDIYLLNLAVSDLLFVM-TLPFQIYYYILFQWSFGEFACKIVS  115 (335)
T ss_pred             chhhHHHHHHHHHHHHHHHhhHhEEeeehhc-cccCchHHHHHHHHHHHHHHHH-hHHHHHHHHhCCCCCCCcHHHHHHH
Confidence            3466788899999999999999999998887 8899999999999999998865 4676665666678999999999999


Q ss_pred             HHhHHhhHHHHHHHHHhhccC
Q psy16860        119 SFDVYFSQYRKHFNMYEVDFE  139 (139)
Q Consensus       119 ~~~~~~~~~Si~~~m~~~~~~  139 (139)
                      ++......+|+++ +.++++|
T Consensus       116 ~~~~~~~~~S~~~-l~~iaid  135 (335)
T PHA03087        116 GLYYIGFYNSMNF-ITVMSVD  135 (335)
T ss_pred             HHHHHHHHHHHHH-HHHHHHH
Confidence            9999999999854 6666654


No 8  
>PF00001 7tm_1:  7 transmembrane receptor (rhodopsin family) Rhodopsin-like GPCR superfamily signature 5-hydroxytryptamine 7 receptor signature bradykinin receptor signature gastrin receptor signature melatonin receptor signature olfactory receptor signature;  InterPro: IPR000276 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The rhodopsin-like GPCRs themselves represent a widespread protein family that includes hormone, neurotransmitter and light receptors, all of which transduce extracellular signals through interaction with guanine nucleotide-binding (G) proteins. Although their activating ligands vary widely in structure and character, the amino acid sequences of the receptors are very similar and are believed to adopt a common structural framework comprising 7 transmembrane (TM) helices [, , ].; GO: 0007186 G-protein coupled receptor protein signaling pathway, 0016021 integral to membrane; PDB: 2KI9_A 3QAK_A 2YDV_A 3VGA_A 3PWH_A 3RFM_A 3EML_A 3VG9_A 3REY_A 3UZA_A ....
Probab=99.35  E-value=2.4e-12  Score=91.84  Aligned_cols=81  Identities=31%  Similarity=0.634  Sum_probs=70.9

Q ss_pred             HHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHHhHHhhHHHHHHHHHhhc
Q psy16860         58 GNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSFDVYFSQYRKHFNMYEVD  137 (139)
Q Consensus        58 gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~~~~~~~~Si~~~m~~~~  137 (139)
                      ||+++++++.++|++|++.++++.|||++|++.++...|........++|..|+..|++..++.......|.+ ++.+++
T Consensus         1 GN~lvi~~~~~~~~~~~~~~~~l~~Lav~Dll~~~~~~~~~~~~~~~~~~~~~~~~C~~~~~~~~~~~~~s~~-~~~~is   79 (257)
T PF00001_consen    1 GNILVILVILRSKRLRTPSNILLLNLAVADLLVGLFCIPFYIYSLLFDDWIFSSFLCRIFGFLFYFSSFSSIF-SLVAIS   79 (257)
T ss_dssp             HHHHHHHHHHHSGGG-SHHHHHHHHHHHHHHHHHHTHHHHHHHHHHHSSCTSHHHHHHHHHHHHHHHHHHHHH-HHHHHH
T ss_pred             CchhehhhhhhhccCCChhHHHHHHHHHHHHhhcccccccccccccccccccccccccccccccccccccccc-cccccc
Confidence            8999999999999999999999999999999999999998887777788999999999999999988888884 555555


Q ss_pred             cC
Q psy16860        138 FE  139 (139)
Q Consensus       138 ~~  139 (139)
                      +|
T Consensus        80 ~d   81 (257)
T PF00001_consen   80 ID   81 (257)
T ss_dssp             HH
T ss_pred             cc
Confidence            44


No 9  
>PF10320 7TM_GPCR_Srsx:  Serpentine type 7TM GPCR chemoreceptor Srsx;  InterPro: IPR019424 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class sx (Srsx), which is a solo family amongst the superfamilies of chemoreceptors. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. 
Probab=98.53  E-value=7.1e-08  Score=70.86  Aligned_cols=75  Identities=28%  Similarity=0.467  Sum_probs=59.7

Q ss_pred             HHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHHhHHhhHHHH
Q psy16860         53 ITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSFDVYFSQYRK  129 (139)
Q Consensus        53 ~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~~~~~~~~Si  129 (139)
                      ++|++||..++.++.|+|++|+|.++++..+|++|++......|..... +.+. ......|-.+.+...++..+..
T Consensus         2 ~ig~~gN~~~i~~~~~~~~Lrs~~~~li~~~~~~d~~~~~~~~~~~~~~-~~~~-~i~~~~Cf~~~~~~~f~~~~qs   76 (257)
T PF10320_consen    2 IIGLFGNLLLIILIFRNKSLRSPCYILICILCFADLICLLGTLPFMLFL-FRDH-QITRSECFWQIFFYIFFQCAQS   76 (257)
T ss_pred             EEEEEccHHHHHHHHhccccccchHHHHHHHHHHHHHHHhhHHHHHHHH-Hhhe-eccHHHHHHHHHHHHHHHHHHH
Confidence            4688999999999999999999999999999999999988888866633 3332 3567789877776665555443


No 10 
>KOG2087|consensus
Probab=98.20  E-value=1.1e-06  Score=66.62  Aligned_cols=92  Identities=25%  Similarity=0.369  Sum_probs=71.2

Q ss_pred             HHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhc-C-------cccccccc
Q psy16860         42 IIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLT-D-------EWYFGYFM  113 (139)
Q Consensus        42 ~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~-~-------~~~~g~~~  113 (139)
                      .+.-....++..+++.||.+|++.+...|...+...+++.|||++|++.++-..-...+.... +       .| .+...
T Consensus        25 ~~lRi~vW~i~~lAi~gN~~Vl~~~~~~~~~~~~~~~li~~la~ad~~mGiYl~~ia~vD~~~~gey~~~ai~W-~tg~g  103 (363)
T KOG2087|consen   25 WILRISVWVIALLAIVGNLLVLLTRFTSRYELNSHRFLICNLAFADLLMGIYLGLIASVDAKTRGEYYKHAIDW-QTGLG  103 (363)
T ss_pred             ceeeehhhhhhhHHhccCeeeeeeeeehhhhccchHHHHHHHHHHHHHcchHHHHHHHhhHHHHHHHHHHHHhh-hhcCC
Confidence            344445667888899999999999888888788889999999999999988655544443332 1       35 35678


Q ss_pred             chhchHHhHHhhHHHHHHHHH
Q psy16860        114 CDVWNSFDVYFSQYRKHFNMY  134 (139)
Q Consensus       114 C~~~~~~~~~~~~~Si~~~m~  134 (139)
                      |++.+|+..+..-.|+++.+.
T Consensus       104 C~~aGflavFASElSv~~LT~  124 (363)
T KOG2087|consen  104 CPVAGFLAVFASELSVFLLTL  124 (363)
T ss_pred             CchHHHHHHHHHHHHHHHHHH
Confidence            999999999999999965543


No 11 
>PF05296 TAS2R:  Mammalian taste receptor protein (TAS2R);  InterPro: IPR007960 This family consists of several forms of mammalian taste receptor proteins (TAS2Rs). TAS2Rs are G protein-coupled receptors expressed in subsets of taste receptor cells of the tongue and palate epithelia and are organised in the genome in clusters. The proteins are genetically linked to loci that influence bitter perception in mice and humans [].; GO: 0004930 G-protein coupled receptor activity, 0007186 G-protein coupled receptor protein signaling pathway, 0050909 sensory perception of taste, 0016021 integral to membrane
Probab=97.86  E-value=0.00044  Score=52.10  Aligned_cols=95  Identities=15%  Similarity=0.200  Sum_probs=70.6

Q ss_pred             HHHHHHHHHHHHHHHHHHHHHHhHHHhh---ccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchh
Q psy16860         40 PFIIKGMLMCFIIITAILGNLLVIISVI---KHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDV  116 (139)
Q Consensus        40 ~~~~~~~~~~~i~~~g~~gN~lvi~v~~---~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~  116 (139)
                      .+.+...+..+.+++|++||+.++.+-+   +++|.-.|.+..+.+||++.++.-....-......+.......+..++.
T Consensus         5 ~~~i~~~i~~~~~~~Gi~~N~FI~~vn~~~w~k~~~l~~~d~IL~~La~sr~~l~~~~~~~~~~~~~~~~~~~~~~~~~~   84 (303)
T PF05296_consen    5 LEIIFLIILVVEFIIGILGNGFIVLVNCSDWVKSRKLSPSDQILTSLAISRILLQWVILLNSFLSFFFPNIYFSENVYKI   84 (303)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHhHHHHHcCCCCChHHHHHHHHHHHHHHHHHHHHHHHHHHHHcchhhhhhhHHHH
Confidence            3567788899999999999998886665   3344456899999999999999866544444444444444456678888


Q ss_pred             chHHhHHhhHHHHHHHHH
Q psy16860        117 WNSFDVYFSQYRKHFNMY  134 (139)
Q Consensus       117 ~~~~~~~~~~~Si~~~m~  134 (139)
                      ..+...+....|++++..
T Consensus        85 ~~~~~~f~~~~s~W~tt~  102 (303)
T PF05296_consen   85 IDFLWMFSNSSSLWFTTW  102 (303)
T ss_pred             HHHHHHHHhHHHHHHHHH
Confidence            888888888888887654


No 12 
>PF05462 Dicty_CAR:  Slime mold cyclic AMP receptor
Probab=97.78  E-value=0.0003  Score=53.00  Aligned_cols=87  Identities=18%  Similarity=0.235  Sum_probs=67.0

Q ss_pred             HHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHHhH
Q psy16860         43 IKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSFDV  122 (139)
Q Consensus        43 ~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~~~  122 (139)
                      ....+..+..+++++|-+.++..+++.|++|++.+.++.-++++|++..+......   .. +.-..+...|++++++..
T Consensus         8 ~~~~i~~~~s~lSllGclfiI~tf~~~k~~r~~~~rli~yl~~~~ll~~v~~~~~~---~~-~~~~~~s~lC~~Qafliq   83 (303)
T PF05462_consen    8 TLYAIELVASVLSLLGCLFIIITFCLFKRLRKPINRLIFYLSIANLLTNVASMIMT---LS-PSAGENSFLCQFQAFLIQ   83 (303)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHHHHhCccHHHHHHHHHHHHHHHHHHHHHHH---hc-ccCCCCCcchhhHhHHHH
Confidence            34455666678888999999999999999999999999999999999765433221   11 222345778999999999


Q ss_pred             HhhHHHHHHHH
Q psy16860        123 YFSQYRKHFNM  133 (139)
Q Consensus       123 ~~~~~Si~~~m  133 (139)
                      ++..++.+.++
T Consensus        84 ~f~~as~lWt~   94 (303)
T PF05462_consen   84 FFMLASFLWTL   94 (303)
T ss_pred             HhhHHHHHHHH
Confidence            99999987554


No 13 
>PF11710 Git3:  G protein-coupled glucose receptor regulating Gpa2;  InterPro: IPR023041 This entry contains a functionally uncharacterised region belonging to the Git3 G-protein coupled receptor. Git3 is one of six proteins required for glucose-triggered adenylate cyclase activation, and is a G protein-coupled receptor responsible for the activation of adenylate cyclase through Gpa2 - heterotrimeric G protein alpha subunit, part of the glucose-detection pathway. Git3 contains seven predicted transmembrane domains, a third cytoplasmic loop and a cytoplasmic tail []. This is the conserved N-terminal domain of the member proteins. 
Probab=97.71  E-value=0.00064  Score=48.33  Aligned_cols=60  Identities=13%  Similarity=0.062  Sum_probs=44.0

Q ss_pred             CCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHHhHHhhHHHH
Q psy16860         70 RKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSFDVYFSQYRK  129 (139)
Q Consensus        70 ~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~~~~~~~~Si  129 (139)
                      +++|.-.+.++.||.++|++.++........+...++-.-+...|..++++....-.++=
T Consensus        30 ~r~~~fR~~LIl~L~~aD~~qal~~~i~~~~~l~~~~i~~~s~~C~aqGf~~q~g~~~sd   89 (201)
T PF11710_consen   30 YRRRSFRHQLILNLLLADFIQALAFLISPIRWLARGGIIAPSPFCQAQGFFLQVGDEASD   89 (201)
T ss_pred             hhhhhHHHHHHHHHHHHHHHHHHHHHHHHHHHHhcCCeeCCCCchhhhHHHHHHHHHHHH
Confidence            445667778999999999999887665455555555444567999999998776654443


No 14 
>PF10328 7TM_GPCR_Srx:  Serpentine type 7TM GPCR chemoreceptor Srx;  InterPro: IPR019430 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class x (Srx) from the Srg superfamily [, ]. Srg receptors contain seven hydrophobic, putative transmembrane, regions and can be distinguished from other 7TM GPCR receptors by their own characteristic TM signatures. 
Probab=96.79  E-value=0.012  Score=43.44  Aligned_cols=45  Identities=31%  Similarity=0.428  Sum_probs=40.6

Q ss_pred             HHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhh
Q psy16860         51 IIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVM   95 (139)
Q Consensus        51 i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~   95 (139)
                      +.+.|++.|.++++.+.|.|++|++.+.+-.+.|++|.+.++...
T Consensus         3 ~s~~G~~~N~~v~~~~~~~~~~~~sF~~l~~~~a~~n~i~~~~~l   47 (274)
T PF10328_consen    3 ISIIGIILNWLVFIIIFKLKSLRNSFGILCASQAIANIIICLIFL   47 (274)
T ss_pred             eeHHHHHHHHHHHHHHHhcccccCCHHHHHHHHHHHHHHHHHHHH
Confidence            467899999999999999999999999999999999999977543


No 15 
>PF10324 7TM_GPCR_Srw:  Serpentine type 7TM GPCR chemoreceptor Srw;  InterPro: IPR019427 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class w (Srw), which is a solo family amongst the superfamilies of chemoreceptors. The genes encoding Srw do not appear to be under as strong an adaptive evolutionary pressure as those of Srz []. 
Probab=96.37  E-value=0.02  Score=42.94  Aligned_cols=51  Identities=20%  Similarity=0.399  Sum_probs=41.5

Q ss_pred             HHHHHHHHHHHhHHHhhccCCCCc-hHHHHHHHHHHHHHHHHHhhhHHHHHHH
Q psy16860         51 IIITAILGNLLVIISVIKHRKLRI-ITNYFVVSLAFADLLVALCVMPFNAIVS  102 (139)
Q Consensus        51 i~~~g~~gN~lvi~v~~~~~~~~~-~~~~~i~nLa~~Dl~~~l~~~p~~~~~~  102 (139)
                      +.++|+++|.+-+.++. +|++|+ +.|.++..+|++|++..+...+......
T Consensus         6 ~~~~g~~~N~~h~~VLt-rk~mR~~~in~~l~~Iai~Dl~~~~~~~~~~~~~~   57 (318)
T PF10324_consen    6 LSIFGLFINIFHLIVLT-RKSMRSSSINILLIGIAICDLLYMLSILIWELFFF   57 (318)
T ss_pred             EeHHHHHHHHHHhhhcC-ChhhhcCCHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            46789999999997764 566676 8999999999999999888777665443


No 16 
>PF10321 7TM_GPCR_Srt:  Serpentine type 7TM GPCR chemoreceptor Srt;  InterPro: IPR019425  Chemoreception is mediated in Caenorhabditis elegans by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs) of proteins which are of the serpentine type []. Srt is a member of the Srg superfamily of chemoreceptors. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. 
Probab=96.07  E-value=0.065  Score=40.71  Aligned_cols=75  Identities=17%  Similarity=0.208  Sum_probs=54.6

Q ss_pred             HHHHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCc-cccccccchhc
Q psy16860         39 FPFIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDE-WYFGYFMCDVW  117 (139)
Q Consensus        39 ~~~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~-~~~g~~~C~~~  117 (139)
                      ..+......+.+.+++-.+-...++.++.++|+.|.+.+....-|++.|++......-      .+|- -..|..+|+.-
T Consensus        30 ~~~p~~G~~~~~~g~~~~~lY~p~~~~i~~~~~~k~~~ykiM~~L~i~Di~~l~~~si------~tG~l~i~G~vfC~~P  103 (313)
T PF10321_consen   30 VKRPILGIYFLIFGIIIIILYIPCLIAIFKKKLFKMSCYKIMFFLAIFDIIQLFINSI------ITGILAIFGAVFCSYP  103 (313)
T ss_pred             CcccchhHHHHHHHHHHHHHHHHHHHHHHHhccccCcHHHHHHHHHHHHHHHHHhhhh------hhhHHHhcCccccCCc
Confidence            3355677778888888888889999999988888899999999999999998554322      2221 12466777654


Q ss_pred             hH
Q psy16860        118 NS  119 (139)
Q Consensus       118 ~~  119 (139)
                      .+
T Consensus       104 ~~  105 (313)
T PF10321_consen  104 RF  105 (313)
T ss_pred             hH
Confidence            44


No 17 
>PF03402 V1R:  Vomeronasal organ pheromone receptor family, V1R;  InterPro: IPR004072 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The rhodopsin-like GPCRs themselves represent a widespread protein family that includes hormone, neurotransmitter and light receptors, all of which transduce extracellular signals through interaction with guanine nucleotide-binding (G) proteins. Although their activating ligands vary widely in structure and character, the amino acid sequences of the receptors are very similar and are believed to adopt a common structural framework comprising 7 transmembrane (TM) helices [, , ]. Pheromones have evolved in all animal phyla, to signal sex and dominance status, and are responsible for stereotypical social and sexual behaviour among members of the same species. In mammals, these chemical signals are believed to be detected primarily by the vomeronasal organ (VNO), a chemosensory organ located at the base of the nasal septum []. The VNO is present in most amphibia, reptiles and non-primate mammals but is absent in birds, adult catarrhine monkeys and apes []. An active role for the human VNO in the detection of pheromones is disputed; the VNO is clearly present in the foetus but appears to be atrophied or absent in adults. Three distinct families of putative pheromone receptors have been identified in the vomeronasal organ (V1Rs, V2Rs and V3Rs). All are G protein-coupled receptors but are only distantly related to the receptors of the main olfactory system, highlighting their different role []. The V1 receptors share between 50 and 90% sequence identity but have little similarity to other families of G protein-coupled receptors. They appear to be distantly related to the mammalian T2R bitter taste receptors and the rhodopsin-like GPCRs []. In rat, the family comprises 30-40 genes. These are expressed in the apical regions of the VNO, in neurons expressing Gi2. Coupling of the receptors to this protein mediates inositol trisphosphate signalling []. A number of human V1 receptor homologues have also been found. The majority of these human sequences are pseudogenes [] but an apparently functional receptor has been identified that is expressed in the human olfactory system [].; GO: 0016503 pheromone receptor activity, 0007186 G-protein coupled receptor protein signaling pathway, 0016021 integral to membrane
Probab=96.00  E-value=0.018  Score=42.64  Aligned_cols=61  Identities=15%  Similarity=0.163  Sum_probs=44.7

Q ss_pred             CCCCchHHHHHHHHHHHHHHHHHh-hhHHHHHHHhcCccccccccchhchHHhHHhhHHHHHHH
Q psy16860         70 RKLRIITNYFVVSLAFADLLVALC-VMPFNAIVSLTDEWYFGYFMCDVWNSFDVYFSQYRKHFN  132 (139)
Q Consensus        70 ~~~~~~~~~~i~nLa~~Dl~~~l~-~~p~~~~~~~~~~~~~g~~~C~~~~~~~~~~~~~Si~~~  132 (139)
                      .++.+|++..+.|||+++.++.+. ++| ........+ .+++..||+..|+.-..-+.|+-++
T Consensus         5 ~~r~kp~dlIl~hLa~aN~lvLl~rGip-~~~~~~~~~-~~~d~gCK~v~Y~~RV~RglSictT   66 (265)
T PF03402_consen    5 GHRLKPIDLILIHLALANILVLLSRGIP-QTMAFFGWK-FFDDIGCKIVFYIYRVARGLSICTT   66 (265)
T ss_pred             CCCCCcHHHHHHHHHHHHHHHHHHhhHH-HHHHHhhcc-cCCCceeeeeeeehHHhchhhHHhh
Confidence            455788999999999999998654 344 333333323 3689999999999888888877543


No 18 
>PF10317 7TM_GPCR_Srd:  Serpentine type 7TM GPCR chemoreceptor Srd;  InterPro: IPR019421 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents the chemoreceptor Srd []. 
Probab=94.64  E-value=0.27  Score=36.72  Aligned_cols=50  Identities=18%  Similarity=0.224  Sum_probs=38.8

Q ss_pred             HHHHHHHHHHHHHHHhHHHhhccC-CCCchHHHHHHHHHHHHHHHHHhhhH
Q psy16860         47 LMCFIIITAILGNLLVIISVIKHR-KLRIITNYFVVSLAFADLLVALCVMP   96 (139)
Q Consensus        47 ~~~~i~~~g~~gN~lvi~v~~~~~-~~~~~~~~~i~nLa~~Dl~~~l~~~p   96 (139)
                      ++.+.+.+|++.|.+.++.+.++. +.-+...+++.|-|+.|++.+....-
T Consensus         4 ~~~~~~~~~~~~n~~Ll~~i~~~tp~~l~~~~~~l~~~~~~~~~~~~~~~~   54 (292)
T PF10317_consen    4 YHPIFFILGIILNILLLYLIIFKTPKSLRTYSILLLNTAIFDLISIISAFL   54 (292)
T ss_pred             eHHHHHHHHHHHHHHHHHHHHHhChHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence            466788999999999998777554 33445679999999999999765433


No 19 
>PF00002 7tm_2:  7 transmembrane receptor (Secretin family);  InterPro: IPR000832 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The secretin-like GPCRs include secretin [], calcitonin [], parathyroid hormone/parathyroid hormone-related peptides [] and vasoactive intestinal peptide [], all of which activate adenylyl cyclase and the phosphatidyl-inositol-calcium pathway. These receptors contain seven transmembrane regions, in a manner reminiscent of the rhodopsins and other receptors believed to interact with G-proteins (however there is no significant sequence identity between these families, the secretin-like receptors thus bear their own unique '7TM' signature). Their N terminus is probably located on the extracellular side of the membrane and potentially glycosylated. This N-terminal region contains a long conserved region which allow the binding of large peptidic ligand such as glucagon, secretin, VIP and PACAP; this region contains five conserved cysteines residues which could be involved in disulphide bond. The C-terminal region of these receptor is probably cytoplasmic. Every receptor gene in this family is encoded on multiple exons, and several of these genes are alternatively spliced to yield functionally distinct products. ; GO: 0004930 G-protein coupled receptor activity, 0007186 G-protein coupled receptor protein signaling pathway, 0016021 integral to membrane; PDB: 3L2J_A 1BL1_A.
Probab=87.94  E-value=0.28  Score=35.22  Aligned_cols=75  Identities=25%  Similarity=0.215  Sum_probs=0.4

Q ss_pred             HHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCcccc--ccccchhchHHhHHhhHHHHH
Q psy16860         53 ITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYF--GYFMCDVWNSFDVYFSQYRKH  130 (139)
Q Consensus        53 ~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~--g~~~C~~~~~~~~~~~~~Si~  130 (139)
                      .+++++-.+.+......|++|+..+....||++++++..+..+..   .....+...  .+..|+....+.+.+..++..
T Consensus        12 ~~Si~~ll~~i~~~~~~r~lr~~~~~i~~~l~~sll~~~~~~l~~---~~~~~~~~~~~~~~~C~~~a~~~hy~~la~f~   88 (242)
T PF00002_consen   12 SLSIICLLLTIITYLLFRKLRSFRNKIHLNLCLSLLLANLSFLIG---ISQTFSPISTTNHCLCRAIAILLHYFFLASFF   88 (242)
T ss_dssp             H-------------------------------------------------------------------------------
T ss_pred             HHHHHHHHHHHHHHHHHHhhcccchhhhhhhHHHHHHHHHHHhee---hhhccccccccccccchhhhhHhHHHHHHHHH
Confidence            333444444444445557777777888999999998876543221   111111111  223599998877776655543


No 20 
>PF10292 7TM_GPCR_Srab:  Serpentine type 7TM GPCR receptor class ab chemoreceptor;  InterPro: IPR019408 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. Srab is part of the Sra superfamily of chemoreceptors. The expression pattern of the srab genes is biologically intriguing. Of the six promoters successfully expressed in transgenic organisms, one was exclusively expressed in the tail phasmid neurons, two were exclusively expressed in a head amphid neuron, and two were expressed both in the head and tail neurons as well as a limited number of other cells []. 
Probab=84.12  E-value=16  Score=27.67  Aligned_cols=97  Identities=8%  Similarity=0.039  Sum_probs=59.0

Q ss_pred             HHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHH---HHHHhc---C-ccccccccc
Q psy16860         42 IIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFN---AIVSLT---D-EWYFGYFMC  114 (139)
Q Consensus        42 ~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~---~~~~~~---~-~~~~g~~~C  114 (139)
                      .....+-.++.++|++.++..++...+++..|....+.+....++.++-++...-..   +..+..   + +.......|
T Consensus        17 ~~~~~~~~~~s~~~~~~~~~~~~~~~~~~~~H~N~ril~~~~~~~~l~~~~~r~~~h~~~l~~~~~~~~~Cd~~~~~~~C   96 (324)
T PF10292_consen   17 RLSLIFNLLLSIIAFPVIIYALWKIRNSKLFHFNTRILFIVHCFSFLIHCTGRIILHTYDLYNYFFPDDPCDMIPSTYRC   96 (324)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHhhcchhchhHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHccCCCcccccchhHH
Confidence            344445677777888787777777777777888888888888888887765433322   222222   1 233455667


Q ss_pred             hhchHHhHHhhHHHHHHHHHhhccC
Q psy16860        115 DVWNSFDVYFSQYRKHFNMYEVDFE  139 (139)
Q Consensus       115 ~~~~~~~~~~~~~Si~~~m~~~~~~  139 (139)
                      -............+. .+..++.+|
T Consensus        97 ~~lR~~~~~~~~~~~-~t~v~l~IE  120 (324)
T PF10292_consen   97 FILRIPYNFGLFLVS-FTTVSLVIE  120 (324)
T ss_pred             HHHHHHHHHHHHHHH-HHHHHHHHH
Confidence            655555555544444 555555554


No 21 
>PF02101 Ocular_alb:  Ocular albinism type 1 protein;  InterPro: IPR001414 Ocular albinism type 1 (OA1) is an X-linked disorder characterised by severe impairment of visual acuity, retinal hypopigmentation and the presence of macromelanosomes. A novel transcript from the OA1 critical region is expressed in high levels in RNA samples from retina and from melanoma and encodes a potential integral membrane protein []. This protein is of unknown function but is known to bind heterotrimeric G proteins.; GO: 0016020 membrane
Probab=78.89  E-value=19  Score=28.33  Aligned_cols=95  Identities=19%  Similarity=0.144  Sum_probs=53.6

Q ss_pred             HHHHHHHHHHHHHHHHHHHhHHHhhccC----CCCc--hHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccc--------
Q psy16860         43 IKGMLMCFIIITAILGNLLVIISVIKHR----KLRI--ITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWY--------  108 (139)
Q Consensus        43 ~~~~~~~~i~~~g~~gN~lvi~v~~~~~----~~~~--~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~--------  108 (139)
                      ..-.+.+....+|+.|-++-++=-.|..    +.++  ...-.+..|+++|++-++-.+--..++....+..        
T Consensus        28 ~f~avCLgSs~l~l~gallQLlp~rr~~~~~~~~~sp~~~~rIl~~la~aDlLaclGVivRS~vWl~~p~~~~s~s~~~~  107 (405)
T PF02101_consen   28 AFNAVCLGSSVLSLLGALLQLLPRRRSAGPRAPARSPSSSRRILFWLAVADLLACLGVIVRSSVWLGFPNFIDSISDVNG  107 (405)
T ss_pred             hhhhhHHHHHHHHHHHHHHhhccccccccccccccCCcCCchhHHHHHHHHHHhhhhHHHHhhhhhcCCcccccccCCCC
Confidence            3444556666677777555553111110    1111  2346889999999999887666665555443221        


Q ss_pred             ---cccccchhchHHhHHhhHHHHH-HHHHhhc
Q psy16860        109 ---FGYFMCDVWNSFDVYFSQYRKH-FNMYEVD  137 (139)
Q Consensus       109 ---~g~~~C~~~~~~~~~~~~~Si~-~~m~~~~  137 (139)
                         .+..+|........++..++-+ +.-||++
T Consensus       108 ~d~wp~afCv~ss~WIq~fYsAtfwWtfcYAVD  140 (405)
T PF02101_consen  108 TDIWPAAFCVGSSMWIQLFYSATFWWTFCYAVD  140 (405)
T ss_pred             CccccHHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence               1347897766655555555544 5556654


No 22 
>KOG4564|consensus
Probab=75.16  E-value=43  Score=27.19  Aligned_cols=84  Identities=18%  Similarity=0.195  Sum_probs=55.0

Q ss_pred             HHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCc--------c----ccccccch
Q psy16860         48 MCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDE--------W----YFGYFMCD  115 (139)
Q Consensus        48 ~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~--------~----~~g~~~C~  115 (139)
                      +.+=.-++++.=++.+.++...|++|-..|+.-.||.++=++-++..+-........++        +    .-+...||
T Consensus       151 ytvGyslSl~sL~vAl~If~~FR~L~CtRn~IH~nLF~SfiLra~~~~i~~~~l~~~~~~~~~~~~~~~~~~~~~~~~Ck  230 (473)
T KOG4564|consen  151 YTVGYSLSLVSLLVALIIFLYFRSLHCTRNYIHMNLFASFILRAASVLIKDLVLVVNGEQDASSDTSLHCLISSNPVGCK  230 (473)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHhhhhcchHHHHHHHHHHHHHHHHHHHHHHHHHhhccccccccccccccccccccchhHH
Confidence            33333333443334445566778999999999999999999988876665554332222        1    13567899


Q ss_pred             hchHHhHHhhHHHHHH
Q psy16860        116 VWNSFDVYFSQYRKHF  131 (139)
Q Consensus       116 ~~~~~~~~~~~~Si~~  131 (139)
                      ....+...+..+.-+.
T Consensus       231 ~~~~~~~Yf~~aNf~W  246 (473)
T KOG4564|consen  231 LLFVFFQYFVLANFFW  246 (473)
T ss_pred             HHHHHHHHHHHHHHHH
Confidence            8888777776665543


No 23 
>PF10327 7TM_GPCR_Sri:  Serpentine type 7TM GPCR chemoreceptor Sri;  InterPro: IPR019429 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents Sri, which is part of the Str superfamily of chemoreceptors.
Probab=67.61  E-value=21  Score=26.95  Aligned_cols=64  Identities=19%  Similarity=0.173  Sum_probs=45.1

Q ss_pred             HHHHHHHHHHHHHHHHHHHHhHHHhh-ccCCCCchHHHHH-H--HHHHHHHHHHHhhhHHHHHHHhcC
Q psy16860         42 IIKGMLMCFIIITAILGNLLVIISVI-KHRKLRIITNYFV-V--SLAFADLLVALCVMPFNAIVSLTD  105 (139)
Q Consensus        42 ~~~~~~~~~i~~~g~~gN~lvi~v~~-~~~~~~~~~~~~i-~--nLa~~Dl~~~l~~~p~~~~~~~~~  105 (139)
                      .+....+-+++.++++-|.+.++.+. |.+|+.+-.++++ .  ...++|+-.+.+..|........|
T Consensus         9 ~~li~~~~~ig~iS~~~n~~~iyLi~fks~k~~~fry~ll~~Qi~~~l~di~~t~L~qpipLfP~~ag   76 (303)
T PF10327_consen    9 QWLINYYHIIGVISFILNSLGIYLIIFKSPKLDNFRYYLLYFQISCTLTDIHLTFLMQPIPLFPIPAG   76 (303)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHheeEEEecCCccchhhHHHHHHHHHHHhhhhhhhhccchhhcceeEE
Confidence            35566788899999999999997666 4555555443333 2  234679999999888777665555


No 24 
>PF09882 DUF2109:  Predicted membrane protein (DUF2109);  InterPro: IPR019214  This entry is found in various hypothetical archaeal proteins and has no known function. 
Probab=66.67  E-value=23  Score=21.22  Aligned_cols=47  Identities=23%  Similarity=0.353  Sum_probs=32.6

Q ss_pred             HHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHH
Q psy16860         52 IITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFN   98 (139)
Q Consensus        52 ~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~   98 (139)
                      .+.|+++=...+-++.-+.+.|+-.+.-.+|-+++-++....--|+.
T Consensus         4 ~i~g~Iai~~~iR~~~~~~r~~KL~yLnv~~F~iaalIaL~i~~P~g   50 (78)
T PF09882_consen    4 IIIGIIAILMAIRIFLTKSRARKLLYLNVINFAIAALIALYIKSPMG   50 (78)
T ss_pred             HHHHHHHHHHHHHHHHhHhHHHhhhHHHHHHHHHHHHHHHHhCCcHH
Confidence            44566665566666666677788788888888888887766555544


No 25 
>PF10316 7TM_GPCR_Srbc:  Serpentine type 7TM GPCR chemoreceptor Srbc ;  InterPro: IPR019420 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class b (Srb) from the Sra superfamily []. Srb receptors contain 6-8 hydrophobic, putative transmembrane, regions and can be distinguished from other 7TM GPCR receptors by their own characteristic TM signatures. Srbc is a solo family amongst the superfamilies of chemoreceptors.
Probab=66.57  E-value=32  Score=25.70  Aligned_cols=61  Identities=15%  Similarity=0.118  Sum_probs=41.7

Q ss_pred             HHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHH
Q psy16860         42 IIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVS  102 (139)
Q Consensus        42 ~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~  102 (139)
                      .+...+-++........|...++.+.++|+.|++--.++---...|.+.+....+......
T Consensus         6 ~iv~~i~i~~s~~~~~iN~~lL~~if~~Kk~kk~~l~LfY~Rf~~D~~~~~~~~~~~~~~~   66 (273)
T PF10316_consen    6 IIVSIIGIIFSIITCLINFYLLYSIFYSKKKKKPDLSLFYFRFAIDVFYGFSVFIYLIYYI   66 (273)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHhccccCCCCEEeeHHHHHHHHHHHHHHHHHHHHHH
Confidence            3445555666777788899888888866664555445555567889999988776544433


No 26 
>PF01102 Glycophorin_A:  Glycophorin A;  InterPro: IPR001195 Proteins in this group are responsible for the molecular basis of the blood group antigens, surface markers on the outside of the red blood cell membrane. Most of these markers are proteins, but some are carbohydrates attached to lipids or proteins [Reid M.E., Lomas-Francis C. The Blood Group Antigen FactsBook Academic Press, London / San Diego, (1997)]. Glycophorin A (PAS-2) and glycophorin B (PAS-3) belong to the MNS blood group system and are associated with antigens that include M/N, S/s, U, He, Mi(a), M(c), Vw, Mur, M(g), Vr, M(e), Mt(a), St(a), Ri(a), Cl(a), Ny(a), Hut, Hil, M(v), Far, Mit, Dantu, Hop, Nob, En(a), ENKT, amongst others. Glycophorin A is the major sialoglycoprotein of the erythrocyte membrane []. Structurally, glycophorin A consists of an N-terminal extracellular domain, heavily glycosylated on serine and threonine residues, followed by a transmembrane region and a C-terminal cytoplasmic domain. Other glycophorins in this entry such as Glycophorin B and Glycophorin E represent minor sialoglycoproteins in the erythrocyte membrane.; GO: 0016021 integral to membrane; PDB: 2KPF_B 1AFO_B 2KPE_A.
Probab=57.16  E-value=15  Score=24.09  Aligned_cols=11  Identities=27%  Similarity=0.428  Sum_probs=4.2

Q ss_pred             HhHHHhhccCC
Q psy16860         61 LVIISVIKHRK   71 (139)
Q Consensus        61 lvi~v~~~~~~   71 (139)
                      ++.|+++|.||
T Consensus        83 li~y~irR~~K   93 (122)
T PF01102_consen   83 LISYCIRRLRK   93 (122)
T ss_dssp             HHHHHHHHHS-
T ss_pred             HHHHHHHHHhc
Confidence            33344444443


No 27 
>PF10323 7TM_GPCR_Srv:  Serpentine type 7TM GPCR chemoreceptor Srv;  InterPro: IPR019426 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae.  This entry represents serpentine receptor class v (Srv) from the Srg superfamily [, ]. Srg receptors contain seven hydrophobic, putative transmembrane, regions and can be distinguished from other 7TM GPCR receptors by their own characteristic TM signatures. 
Probab=57.16  E-value=46  Score=24.71  Aligned_cols=41  Identities=17%  Similarity=0.292  Sum_probs=28.2

Q ss_pred             HHHHHHhHHHhhccCC----CCchHHHHHHHHHHHHHHHHHhhhH
Q psy16860         56 ILGNLLVIISVIKHRK----LRIITNYFVVSLAFADLLVALCVMP   96 (139)
Q Consensus        56 ~~gN~lvi~v~~~~~~----~~~~~~~~i~nLa~~Dl~~~l~~~p   96 (139)
                      ++--..+++++.|.|+    .+++.+.++.+-+++|++..+....
T Consensus         9 lply~~il~~l~~~r~~~~~~~~~Fy~l~~~~~iaDi~~~~~~~~   53 (283)
T PF10323_consen    9 LPLYIFILYCLLKLRKRSKTFKSTFYTLLIQHCIADILSMLFYFL   53 (283)
T ss_pred             HHHHHHHHHHHHHcccCccccCCHHHHHHHHHHHHHHHHHHHHHH
Confidence            3444555555554443    5688999999999999998665433


No 28 
>PF11446 DUF2897:  Protein of unknown function (DUF2897);  InterPro: IPR021550  This is a bacterial family of uncharacterised proteins. 
Probab=53.70  E-value=20  Score=20.05  Aligned_cols=22  Identities=23%  Similarity=0.445  Sum_probs=13.9

Q ss_pred             HHHHHHHHHHHHHHHHhHHHhh
Q psy16860         46 MLMCFIIITAILGNLLVIISVI   67 (139)
Q Consensus        46 ~~~~~i~~~g~~gN~lvi~v~~   67 (139)
                      .+.+++.+.-++||+.++---.
T Consensus         6 wlIIviVlgvIigNia~LK~sA   27 (55)
T PF11446_consen    6 WLIIVIVLGVIIGNIAALKYSA   27 (55)
T ss_pred             hHHHHHHHHHHHhHHHHHHHhc
Confidence            3445555556789998885433


No 29 
>TIGR01477 RIFIN variant surface antigen, rifin family. This model represents the rifin branch of the rifin/stevor family (pfam02009) of predicted variant surface antigens as found in Plasmodium falciparum. This model is based on a set of rifin sequences kindly provided by Matt Berriman from the Sanger Center. This is a global model and assesses a penalty for incomplete sequence. Additional fragmentary sequences may be found with the fragment model and a cutoff of 20 bits.
Probab=52.75  E-value=20  Score=27.84  Aligned_cols=31  Identities=16%  Similarity=0.314  Sum_probs=21.1

Q ss_pred             HHHHHHHHHHHHHHHHHHhHHHhhccCCCCc
Q psy16860         44 KGMLMCFIIITAILGNLLVIISVIKHRKLRI   74 (139)
Q Consensus        44 ~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~   74 (139)
                      .+....++.++.++.-.++|+.++|+||.++
T Consensus       310 t~IiaSiIAIvvIVLIMvIIYLILRYRRKKK  340 (353)
T TIGR01477       310 TPIIASIIAILIIVLIMVIIYLILRYRRKKK  340 (353)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHhhhcch
Confidence            3455666666666666678888888877554


No 30 
>PTZ00046 rifin; Provisional
Probab=49.96  E-value=24  Score=27.52  Aligned_cols=30  Identities=13%  Similarity=0.361  Sum_probs=20.2

Q ss_pred             HHHHHHHHHHHHHHHHHhHHHhhccCCCCc
Q psy16860         45 GMLMCFIIITAILGNLLVIISVIKHRKLRI   74 (139)
Q Consensus        45 ~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~   74 (139)
                      .....++.++-++.-.++|+.++|+||.++
T Consensus       316 aIiaSiiAIvVIVLIMvIIYLILRYRRKKK  345 (358)
T PTZ00046        316 AIIASIVAIVVIVLIMVIIYLILRYRRKKK  345 (358)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHhhhcch
Confidence            445566666666666677888888877554


No 31 
>PF06024 DUF912:  Nucleopolyhedrovirus protein of unknown function (DUF912);  InterPro: IPR009261 This entry is represented by Autographa californica nuclear polyhedrosis virus (AcMNPV), Orf78; it is a family of uncharacterised viral proteins.
Probab=46.02  E-value=29  Score=21.74  Aligned_cols=11  Identities=9%  Similarity=0.371  Sum_probs=5.7

Q ss_pred             HHHhhccCCCC
Q psy16860         63 IISVIKHRKLR   73 (139)
Q Consensus        63 i~v~~~~~~~~   73 (139)
                      -+++.|.|+.+
T Consensus        83 YFVILRer~~~   93 (101)
T PF06024_consen   83 YFVILRERQKS   93 (101)
T ss_pred             EEEEEeccccc
Confidence            35555665543


No 32 
>PF02009 Rifin_STEVOR:  Rifin/stevor family;  InterPro: IPR002858 Malaria is still a major cause of mortality in many areas of the world. Plasmodium falciparum causes the most severe human form of the disease and is responsible for most fatalities. Severe cases of malaria can occur when the parasite invades and then proliferates within red blood cell erythrocytes. The parasite produces many variant antigenic proteins, encoded by multigene families, which are present on the surface of the infected erythrocyte and play important roles in virulence. A crucial survival mechanism for the malaria parasite is its ability to evade the immune response by switching these variant surface antigens. The high virulence of P. falciparum relative to other malarial parasites is in large part due to the fact that in this organism many of these surface antigens mediate the binding of infected erythrocytes to the vascular endothelium (cytoadherence) and non-infected erythrocytes (rosetting). This can lead to the accumulation of infected cells in the vasculature of a variety of organs, blocking the blood flow and reducing the oxygen supply. Clinical symptoms of severe infection can include fever, progressive anaemia, multi-organ dysfunction and coma. For more information see []. Several multicopy gene families have been described in Plasmodium falciparum, including the stevor family of subtelomeric open reading frames and the rif interspersed repetitive elements. Both families contain three predicted transmembrane segments. It has been proposed that stevor and rif are members of a larger superfamily that code for variant surface antigens [].
Probab=45.57  E-value=23  Score=26.90  Aligned_cols=27  Identities=19%  Similarity=0.374  Sum_probs=15.3

Q ss_pred             HHHHHHHHHHHHHHHhHHHhhccCCCC
Q psy16860         47 LMCFIIITAILGNLLVIISVIKHRKLR   73 (139)
Q Consensus        47 ~~~~i~~~g~~gN~lvi~v~~~~~~~~   73 (139)
                      ...++.++-++.=.++|+.++|+||.|
T Consensus       259 ~aSiiaIliIVLIMvIIYLILRYRRKK  285 (299)
T PF02009_consen  259 IASIIAILIIVLIMVIIYLILRYRRKK  285 (299)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHHHHh
Confidence            444444444444456777777776643


No 33 
>PF02532 PsbI:  Photosystem II reaction centre I protein (PSII 4.8 kDa protein);  InterPro: IPR003686 Oxygenic photosynthesis uses two multi-subunit photosystems (I and II) located in the cell membranes of cyanobacteria and in the thylakoid membranes of chloroplasts in plants and algae. Photosystem II (PSII) has a P680 reaction centre containing chlorophyll 'a' that uses light energy to carry out the oxidation (splitting) of water molecules, and to produce ATP via a proton pump. Photosystem I (PSI) has a P700 reaction centre containing chlorophyll that takes the electron and associated hydrogen donated from PSII to reduce NADP+ to NADPH. Both ATP and NADPH are subsequently used in the light-independent reactions to convert carbon dioxide to glucose using the hydrogen atom extracted from water by PSII, releasing oxygen as a by-product. PSII is a multisubunit protein-pigment complex containing polypeptides both intrinsic and extrinsic to the photosynthetic membrane [, ]. Within the core of the complex, the chlorophyll and beta-carotene pigments are mainly bound to the antenna proteins CP43 (PsbC) and CP47 (PsbB), which pass the excitation energy on to the reaction centre proteins D1 (Qb, PsbA) and D2 (Qa, PsbD) that bind all the redox-active cofactors involved in the energy conversion process. The PSII oxygen-evolving complex (OEC) oxidises water to provide protons for use by PSI, and consists of OEE1 (PsbO), OEE2 (PsbP) and OEE3 (PsbQ). The remaining subunits in PSII are of low molecular weight (less than 10 kDa), and are involved in PSII assembly, stabilisation, dimerisation, and photo-protection [].  This family represents the low molecular weight transmembrane protein PsbI, which is tightly associated with the D1/D2 heterodimer in PSII. The function of PsbI is unknown, but it may be involved in the assembly, dimerisation or stabilisation of PSII dimers [].; GO: 0015979 photosynthesis, 0009523 photosystem II, 0009539 photosystem II reaction center, 0016020 membrane; PDB: 3A0H_i 3ARC_I 3A0B_i 3BZ2_I 3PRQ_I 3KZI_I 3PRR_I 2AXT_i 4FBY_I 1S5L_i ....
Probab=42.40  E-value=42  Score=16.93  Aligned_cols=18  Identities=17%  Similarity=0.229  Sum_probs=9.5

Q ss_pred             HHHHHHHHHHHHHHHHHH
Q psy16860         42 IIKGMLMCFIIITAILGN   59 (139)
Q Consensus        42 ~~~~~~~~~i~~~g~~gN   59 (139)
                      +..+.+++.+++.|.+.|
T Consensus         9 y~vV~ffv~LFifGflsn   26 (36)
T PF02532_consen    9 YTVVIFFVSLFIFGFLSN   26 (36)
T ss_dssp             HHHHHHHHHHHHHHHHTT
T ss_pred             hhhHHHHHHHHhccccCC
Confidence            344455555666655544


No 34 
>PF15330 SIT:  SHP2-interacting transmembrane adaptor protein, SIT
Probab=42.17  E-value=53  Score=20.96  Aligned_cols=28  Identities=11%  Similarity=0.216  Sum_probs=18.0

Q ss_pred             HHHHHHHHHHHHHHHHHHhHHHhhccCC
Q psy16860         44 KGMLMCFIIITAILGNLLVIISVIKHRK   71 (139)
Q Consensus        44 ~~~~~~~i~~~g~~gN~lvi~v~~~~~~   71 (139)
                      ...++.++.++.++.|++......|++|
T Consensus         3 Ll~il~llLll~l~asl~~wr~~~rq~k   30 (107)
T PF15330_consen    3 LLGILALLLLLSLAASLLAWRMKQRQKK   30 (107)
T ss_pred             HHHHHHHHHHHHHHHHHHHHHHHhhhcc
Confidence            4556777777778888777654444444


No 35 
>KOG4193|consensus
Probab=40.61  E-value=2e+02  Score=24.34  Aligned_cols=63  Identities=17%  Similarity=0.092  Sum_probs=34.4

Q ss_pred             HHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCcccccc-c-cchhchHHhHHhhHHHHH
Q psy16860         60 LLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGY-F-MCDVWNSFDVYFSQYRKH  130 (139)
Q Consensus        60 ~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~-~-~C~~~~~~~~~~~~~Si~  130 (139)
                      .+.+++.+..|++++..+....||+++-++. -+      . .+.+.|.-+. . .|+...++-+++..+...
T Consensus       338 ~lti~ty~~~~~l~~~~~~i~~~l~~~L~l~-~l------~-fL~~~~~~~~~~~~C~~~a~llhff~LaaF~  402 (610)
T KOG4193|consen  338 LLTIATYLLFRKLQNDRTKIHINLCLCLFLA-EL------L-FLLGIDRTSTSVVLCIAAAILLHFFFLAAFF  402 (610)
T ss_pred             HHHHHHHHHHHHHHhhcchhHHHHHHHHHHH-HH------H-HhcccccccCcccccHHHHHHHHHHHHHHHH
Confidence            3444444444445544478888998872222 11      1 2223344332 2 699988877777666543


No 36 
>KOG2927|consensus
Probab=39.87  E-value=95  Score=24.28  Aligned_cols=9  Identities=22%  Similarity=0.586  Sum_probs=6.0

Q ss_pred             Ccccccccc
Q psy16860        105 DEWYFGYFM  113 (139)
Q Consensus       105 ~~~~~g~~~  113 (139)
                      +-|.+++.+
T Consensus       255 g~W~FPNL~  263 (372)
T KOG2927|consen  255 GFWLFPNLT  263 (372)
T ss_pred             ceEeccchh
Confidence            358887654


No 37 
>PF10326 7TM_GPCR_Str:  Serpentine type 7TM GPCR chemoreceptor Str;  InterPro: IPR019428 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class r (Str) from the Str superfamily [, ]. Almost a quarter (22.5%) of str and srj family genes and pseudogenes in C. elegans appear to have been newly formed by gene duplications since the species split []. 
Probab=39.84  E-value=28  Score=25.92  Aligned_cols=48  Identities=8%  Similarity=0.238  Sum_probs=34.0

Q ss_pred             HHHHHHHHHHHHHhHHHhhccCCCCc-hHHHHHHHHHHHHHHHHHhhhH
Q psy16860         49 CFIIITAILGNLLVIISVIKHRKLRI-ITNYFVVSLAFADLLVALCVMP   96 (139)
Q Consensus        49 ~~i~~~g~~gN~lvi~v~~~~~~~~~-~~~~~i~nLa~~Dl~~~l~~~p   96 (139)
                      -+-++++++.|.+.++.+.++.+.+. ...+++.--|+.|+..+..-..
T Consensus         6 ~~~~~~s~~~N~~Li~Li~~~s~k~~G~Yk~Lm~~fs~~~i~fs~~~~~   54 (307)
T PF10326_consen    6 YIGFVLSLFLNSLLIYLILTKSPKSLGSYKYLMIYFSIFEIIFSILDFL   54 (307)
T ss_pred             HHHHHHHHHHHHHHHHHHHhccCCCCCCEEEEEehhHHHHHHHHHHHHH
Confidence            45678889999999988775544333 3456777788888888776543


No 38 
>COG1230 CzcD Co/Zn/Cd efflux system component [Inorganic ion transport and metabolism]
Probab=37.73  E-value=1.8e+02  Score=22.12  Aligned_cols=70  Identities=17%  Similarity=0.177  Sum_probs=43.2

Q ss_pred             HHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHH-HHHHHHHHHhhhHHHHHHHhcCccccccccchhchH
Q psy16860         47 LMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSL-AFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNS  119 (139)
Q Consensus        47 ~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nL-a~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~  119 (139)
                      -++.+.++|++-|.+..+.+.+.++ + ..|.=-..| +++|.+-++..+--.+.-.+.+ |..-|..+.+...
T Consensus       126 ~ml~va~~GL~vN~~~a~ll~~~~~-~-~lN~r~a~LHvl~D~Lgsv~vIia~i~i~~~~-w~~~Dpi~si~i~  196 (296)
T COG1230         126 GMLVVAIIGLVVNLVSALLLHKGHE-E-NLNMRGAYLHVLGDALGSVGVIIAAIVIRFTG-WSWLDPILSIVIA  196 (296)
T ss_pred             chHHHHHHHHHHHHHHHHHhhCCCc-c-cchHHHHHHHHHHHHHHHHHHHHHHHHHHHhC-CCccchHHHHHHH
Confidence            3678889999999999998877622 1 122222223 3579998777666555555555 4445665544333


No 39 
>PRK03557 zinc transporter ZitB; Provisional
Probab=36.75  E-value=1.9e+02  Score=21.91  Aligned_cols=73  Identities=16%  Similarity=0.145  Sum_probs=34.1

Q ss_pred             HHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHHhHH
Q psy16860         50 FIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSFDVY  123 (139)
Q Consensus        50 ~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~~~~  123 (139)
                      .+.+.+++.|.+..+..++.++.++..-.--.-=...|.+.++..+--.+...+.+ |..-+..+.+...+..+
T Consensus       126 ~v~~~~~~~~~~~~~~~~~~~~~~s~~l~a~~~h~~~D~l~s~~vlv~~~~~~~~g-~~~~Dpi~~ilis~~i~  198 (312)
T PRK03557        126 AIAVAGLLANILSFWLLHHGSEEKNLNVRAAALHVLGDLLGSVGAIIAALIIIWTG-WTPADPILSILVSVLVL  198 (312)
T ss_pred             HHHHHHHHHHHHHHHHHhcccccCCHHHHHHHHHHHHHHHHHHHHHHHHHHHHHcC-CcchhHHHHHHHHHHHH
Confidence            34556777787766655554443443211111112557777665333222222333 33345666554444333


No 40 
>PF08114 PMP1_2:  ATPase proteolipid family;  InterPro: IPR012589 This family consists of small proteolipids associated with the plasma membrane H+ ATPase. Two proteolipids (PMP1 and PMP2) are associated with the ATPase and both genes are similarly expressed in the wild-type strain of yeast. No modification of the level of transcription of one PMP gene is detected in a strain deleted of the other. Though both proteolipids show similarity with other small proteolipids associated with other cation -transporting ATPases, their functions remain unclear [].
Probab=35.60  E-value=13  Score=19.39  Aligned_cols=24  Identities=8%  Similarity=0.270  Sum_probs=14.0

Q ss_pred             HHHHHHHHHHHHHHhHHHhhccCC
Q psy16860         48 MCFIIITAILGNLLVIISVIKHRK   71 (139)
Q Consensus        48 ~~~i~~~g~~gN~lvi~v~~~~~~   71 (139)
                      ..+++++|+.|-+++...+.|+..
T Consensus        11 IlVF~lVglv~i~iva~~iYRKw~   34 (43)
T PF08114_consen   11 ILVFCLVGLVGIGIVALFIYRKWQ   34 (43)
T ss_pred             eeehHHHHHHHHHHHHHHHHHHHH
Confidence            344556666666666666665543


No 41 
>cd07912 Tweety_N N-terminal domain of the protein encoded by the Drosophila tweety gene and related proteins, a family of chloride ion channels. The protein product of the Drosophila tweety (tty) gene is thought to form a trans-membrane protein with five membrane-spanning regions and a cytoplasmic C-terminus. This N-terminal domain contains the putative transmembrane spanning regions. Tweety has been suggested as a candidate for a large conductance chloride channel, both in vertebrate and insect cells. Three human homologs have been identified and designated TTYH1-3. TTYH2 has been associated with the progression of cancer, and Drosophila melanogaster tweety has been assumed to play a role in development. TTYH2, and TTYH3 bind to and are ubiquinated by Nedd4-2, a HECT type E3 ubiquitin ligase, which most likely plays a role in controlling the cellular levels of tweety family proteins.
Probab=34.54  E-value=2.4e+02  Score=22.60  Aligned_cols=15  Identities=33%  Similarity=0.279  Sum_probs=8.3

Q ss_pred             HHHHHHHHHHHHHHH
Q psy16860         76 TNYFVVSLAFADLLV   90 (139)
Q Consensus        76 ~~~~i~nLa~~Dl~~   90 (139)
                      ...+...|.+.-++.
T Consensus        79 ~~c~~~sLiiltL~~   93 (418)
T cd07912          79 ICCLKWSLVIATLLC   93 (418)
T ss_pred             ccHHHHHHHHHHHHH
Confidence            445555666555554


No 42 
>PF12606 RELT:  Tumour necrosis factor receptor superfamily member 19;  InterPro: IPR022248 The members of tumor necrosis factor receptor (TNFR) superfamily have been designated as the "guardians of the immune system" due to their roles in immune cell proliferation, differentiation, activation, and death (apoptosis).  RELT (receptor expressed in lymphoid tissues) is a member of the TNFR superfamily. The messenger RNA of RELT is especially abundant in hematologic tissues such as spleen, lymph node, and peripheral blood leukocytes as well as in leukemias and lymphomas. RELT is able to activate the NF-kappaB pathway and selectively binds tumor necrosis factor receptor-associated factor 1 []. RELT like proteins 1 and 2 (RELL1 and RELL2) are two RELT homologues that bind to RELT. The expression of RELL1 at the mRNA level is ubiquitous, whereas expression of RELL2 mRNA is more restricted to particular tissues [].
Probab=33.18  E-value=51  Score=18.03  Aligned_cols=23  Identities=26%  Similarity=0.531  Sum_probs=11.1

Q ss_pred             HHHHHHHHHHHHHHHhHHHhhccCC
Q psy16860         47 LMCFIIITAILGNLLVIISVIKHRK   71 (139)
Q Consensus        47 ~~~~i~~~g~~gN~lvi~v~~~~~~   71 (139)
                      +..++++.|++|  +.++.+.|.+.
T Consensus         6 iV~i~iv~~lLg--~~I~~~~K~yg   28 (50)
T PF12606_consen    6 IVSIFIVMGLLG--LSICTTLKAYG   28 (50)
T ss_pred             HHHHHHHHHHHH--HHHHHHhhccc
Confidence            344555556655  33334444443


No 43 
>PF10873 DUF2668:  Protein of unknown function (DUF2668);  InterPro: IPR022640  Members in this family of proteins are annotated as cysteine and tyrosine-rich protein 1, however currently no function is known []. 
Probab=30.08  E-value=59  Score=22.04  Aligned_cols=32  Identities=9%  Similarity=0.322  Sum_probs=21.5

Q ss_pred             HHHHHHHHHHHHHHHHHHHhHHHhhccCCCCc
Q psy16860         43 IKGMLMCFIIITAILGNLLVIISVIKHRKLRI   74 (139)
Q Consensus        43 ~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~   74 (139)
                      +..+++.++++.|+++-+.+....+.++..++
T Consensus        63 IaGIVfgiVfimgvva~i~icvCmc~kn~rgs   94 (155)
T PF10873_consen   63 IAGIVFGIVFIMGVVAGIAICVCMCMKNSRGS   94 (155)
T ss_pred             eeeeehhhHHHHHHHHHHHHHHhhhhhcCCCc
Confidence            33446788888888887777766665555443


No 44 
>KOG1482|consensus
Probab=29.92  E-value=1.2e+02  Score=23.92  Aligned_cols=78  Identities=13%  Similarity=0.072  Sum_probs=51.6

Q ss_pred             HHHHHHHHHhHHHhhccCCCCch-----------------HH-HHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccc
Q psy16860         53 ITAILGNLLVIISVIKHRKLRII-----------------TN-YFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMC  114 (139)
Q Consensus        53 ~~g~~gN~lvi~v~~~~~~~~~~-----------------~~-~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C  114 (139)
                      -+|++-|++...++...-+-++.                 .| ---..=++.|++-+.-..-......+...|..-+..|
T Consensus       184 ~~gv~vNiim~~vL~~~~h~h~H~~~~s~g~~h~~~~~~n~nvraAyiHVlGDliQSvGV~iaa~Ii~f~P~~~i~DpIC  263 (379)
T KOG1482|consen  184 AVGVAVNIIMGFVLHQSGHGHSHGGSHSHGHSHDHGEELNLNVRAAFVHVLGDLIQSVGVLIAALIIYFKPEYKIADPIC  263 (379)
T ss_pred             ehhhhhhhhhhhhhcccCCCCCCCCCCCcCcccccccccchHHHHHHHHHHHHHHHHHHHHhhheeEEecccceecCchh
Confidence            45667788777776644122221                 11 1112245679999877666666667777899999999


Q ss_pred             hhchHHhHHhhHHHHH
Q psy16860        115 DVWNSFDVYFSQYRKH  130 (139)
Q Consensus       115 ~~~~~~~~~~~~~Si~  130 (139)
                      .+...........+++
T Consensus       264 T~~FSiivl~TT~~i~  279 (379)
T KOG1482|consen  264 TFVFSIIVLGTTITIL  279 (379)
T ss_pred             hhhHHHHHHHhHHHHH
Confidence            9988888888777775


No 45 
>PF10319 7TM_GPCR_Srj:  Serpentine type 7TM GPCR chemoreceptor Srj;  InterPro: IPR019423 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/).  The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae.  This entry represents serpentine receptor class j (Srj) from the Str superfamily [, ]. The Srj family is designated as the out-group based on its location in preliminary phylogenetic analyses of the entire superfamily []. 
Probab=28.62  E-value=1.2e+02  Score=23.36  Aligned_cols=48  Identities=15%  Similarity=0.231  Sum_probs=37.3

Q ss_pred             HHHHHHHHHHHHHHHhHHHhhccCCCCc-hHHHHHHHHHHHHHHHHHhh
Q psy16860         47 LMCFIIITAILGNLLVIISVIKHRKLRI-ITNYFVVSLAFADLLVALCV   94 (139)
Q Consensus        47 ~~~~i~~~g~~gN~lvi~v~~~~~~~~~-~~~~~i~nLa~~Dl~~~l~~   94 (139)
                      +.-+.++++.+-|-+.++.+..+|+.+- ...+++.--|+-|++.++.-
T Consensus        10 ~Pk~~~~lsf~~Np~fiyli~~~~~~~~G~Yr~LL~~Fa~fn~~~S~~~   58 (310)
T PF10319_consen   10 IPKIFGILSFIVNPIFIYLIFTEKKSQFGNYRYLLLFFAIFNLIYSVVD   58 (310)
T ss_pred             HHHHHHHHHHHHhhhhheeEEcccccccccHHHHHHHHHHHHHHHHHHH
Confidence            3456778888999999999888887655 34577888888899988753


No 46 
>KOG4349|consensus
Probab=26.11  E-value=2e+02  Score=18.99  Aligned_cols=21  Identities=5%  Similarity=0.159  Sum_probs=11.9

Q ss_pred             ccccccchhchHHhHHhhHHH
Q psy16860        108 YFGYFMCDVWNSFDVYFSQYR  128 (139)
Q Consensus       108 ~~g~~~C~~~~~~~~~~~~~S  128 (139)
                      ......|...+..+.+...+.
T Consensus       114 ~ms~~y~~i~G~~QT~~i~v~  134 (143)
T KOG4349|consen  114 AVSTWYCAIMGIIQTLLIFVT  134 (143)
T ss_pred             CccHHHHHHHHHHHHHHHHHH
Confidence            456667776665555544443


No 47 
>PF05545 FixQ:  Cbb3-type cytochrome oxidase component FixQ;  InterPro: IPR008621 This family consists of several Cbb3-type cytochrome oxidase components (FixQ/CcoQ). FixQ is found in nitrogen fixing bacteria. Since nitrogen fixation is an energy-consuming process, effective symbioses depend on operation of a respiratory chain with a high affinity for O2, closely coupled to ATP production. This requirement is fulfilled by a special three-subunit terminal oxidase (cytochrome terminal oxidase cbb3), which was first identified in Bradyrhizobium japonicum as the product of the fixNOQP operon [].
Probab=26.07  E-value=1.1e+02  Score=16.13  Aligned_cols=9  Identities=22%  Similarity=0.224  Sum_probs=4.8

Q ss_pred             HhHHHhhcc
Q psy16860         61 LVIISVIKH   69 (139)
Q Consensus        61 lvi~v~~~~   69 (139)
                      +++|+..++
T Consensus        25 i~~w~~~~~   33 (49)
T PF05545_consen   25 IVIWAYRPR   33 (49)
T ss_pred             HHHHHHccc
Confidence            556665443


No 48 
>PF05961 Chordopox_A13L:  Chordopoxvirus A13L protein;  InterPro: IPR009236 This family consists of A13L proteins from the Chordopoxviruses. A13L or p8 is one of the three most abundant membrane proteins of the intracellular mature Vaccinia virus [].
Probab=25.87  E-value=1.1e+02  Score=17.79  Aligned_cols=13  Identities=15%  Similarity=0.457  Sum_probs=8.3

Q ss_pred             HhHHHhhccCCCC
Q psy16860         61 LVIISVIKHRKLR   73 (139)
Q Consensus        61 lvi~v~~~~~~~~   73 (139)
                      ++++.+..+++-+
T Consensus        17 lIlY~iYnr~~~~   29 (68)
T PF05961_consen   17 LILYGIYNRKKTT   29 (68)
T ss_pred             HHHHHHHhccccc
Confidence            6777777665543


No 49 
>CHL00024 psbI photosystem II protein I
Probab=24.50  E-value=34  Score=17.25  Aligned_cols=17  Identities=18%  Similarity=0.257  Sum_probs=8.8

Q ss_pred             HHHHHHHHHHHHHHHHH
Q psy16860         43 IKGMLMCFIIITAILGN   59 (139)
Q Consensus        43 ~~~~~~~~i~~~g~~gN   59 (139)
                      ..+.+++.+++.|.+.|
T Consensus        10 ~vV~ffvsLFifGFlsn   26 (36)
T CHL00024         10 TVVIFFVSLFIFGFLSN   26 (36)
T ss_pred             hHHHHHHHHHHccccCC
Confidence            34455555565555443


No 50 
>PF06679 DUF1180:  Protein of unknown function (DUF1180);  InterPro: IPR009565 This entry consists of several hypothetical eukaryotic proteins thought to be membrane proteins. Their function is unknown.
Probab=24.11  E-value=2.5e+02  Score=19.39  Aligned_cols=26  Identities=19%  Similarity=0.235  Sum_probs=14.4

Q ss_pred             HHHHHHHHHHHHHHHHHHHHhHHHhh
Q psy16860         42 IIKGMLMCFIIITAILGNLLVIISVI   67 (139)
Q Consensus        42 ~~~~~~~~~i~~~g~~gN~lvi~v~~   67 (139)
                      .++-.+|+++.+.+++.=-+++-+++
T Consensus        93 ~l~R~~~Vl~g~s~l~i~yfvir~~R  118 (163)
T PF06679_consen   93 MLKRALYVLVGLSALAILYFVIRTFR  118 (163)
T ss_pred             chhhhHHHHHHHHHHHHHHHHHHHHh
Confidence            34555566666666655445555443


No 51 
>PRK13664 hypothetical protein; Provisional
Probab=23.98  E-value=1.1e+02  Score=17.29  Aligned_cols=17  Identities=12%  Similarity=0.405  Sum_probs=11.6

Q ss_pred             HHHHHHHHHHHHHHHhH
Q psy16860         47 LMCFIIITAILGNLLVI   63 (139)
Q Consensus        47 ~~~~i~~~g~~gN~lvi   63 (139)
                      +++++.++|++-|++=-
T Consensus        10 ilill~lvG~i~N~iK~   26 (62)
T PRK13664         10 ILVLVFLVGVLLNVIKD   26 (62)
T ss_pred             HHHHHHHHHHHHHHHHH
Confidence            34667788888886543


No 52 
>PRK02655 psbI photosystem II reaction center I protein I; Provisional
Probab=22.12  E-value=38  Score=17.23  Aligned_cols=16  Identities=13%  Similarity=0.202  Sum_probs=8.1

Q ss_pred             HHHHHHHHHHHHHHHH
Q psy16860         43 IKGMLMCFIIITAILG   58 (139)
Q Consensus        43 ~~~~~~~~i~~~g~~g   58 (139)
                      ..+.+++.+++.|.+.
T Consensus        10 ~vV~ffvsLFiFGfls   25 (38)
T PRK02655         10 IVVFFFVGLFVFGFLS   25 (38)
T ss_pred             hhHHHHHHHHHcccCC
Confidence            3444555555555443


No 53 
>PF10329 DUF2417:  Region of unknown function (DUF2417);  InterPro: IPR019431  This entry represents a family of fungal proteins with no known function. In some cases these proteins also contain an alpha/beta hydrolase fold (IPR000073 from INTERPRO). 
Probab=21.36  E-value=2.2e+02  Score=20.89  Aligned_cols=15  Identities=27%  Similarity=0.499  Sum_probs=7.3

Q ss_pred             HHHHHHHHHHHHHHH
Q psy16860         76 TNYFVVSLAFADLLV   90 (139)
Q Consensus        76 ~~~~i~nLa~~Dl~~   90 (139)
                      .++.+.-|-+.|++.
T Consensus       104 l~~vl~~Lllvdlil  118 (232)
T PF10329_consen  104 LNIVLAGLLLVDLIL  118 (232)
T ss_pred             HHHHHHHHHHHHHHH
Confidence            344444444555554


No 54 
>PHA03049 IMV membrane protein; Provisional
Probab=20.79  E-value=1.7e+02  Score=17.02  Aligned_cols=17  Identities=18%  Similarity=0.567  Sum_probs=9.7

Q ss_pred             HHHHHHHHHhHHHhhccCC
Q psy16860         53 ITAILGNLLVIISVIKHRK   71 (139)
Q Consensus        53 ~~g~~gN~lvi~v~~~~~~   71 (139)
                      +++++|  ++++.+..+++
T Consensus        11 CVaIi~--lIvYgiYnkk~   27 (68)
T PHA03049         11 CVVIIG--LIVYGIYNKKT   27 (68)
T ss_pred             HHHHHH--HHHHHHHhccc
Confidence            334444  66777776554


Done!