Query psy16860
Match_columns 139
No_of_seqs 123 out of 1420
Neff 9.1
Searched_HMMs 46136
Date Fri Aug 16 23:55:19 2013
Command hhsearch -i /work/01045/syshi/Psyhhblits/psy16860.a3m -d /work/01045/syshi/HHdatabase/Cdd.hhm -o /work/01045/syshi/hhsearch_cdd/16860hhsearch_cdd -cpu 12 -v 0
No Hit Prob E-value P-value Score SS Cols Query HMM Template HMM
1 KOG4219|consensus 99.8 4.2E-19 9.2E-24 134.2 6.4 103 36-139 30-132 (423)
2 PHA03234 DNA packaging protein 99.8 1.3E-17 2.8E-22 126.6 11.5 99 38-139 29-129 (338)
3 PHA02834 chemokine receptor-li 99.7 1.3E-15 2.9E-20 114.8 10.8 95 41-139 28-122 (323)
4 PHA02638 CC chemokine receptor 99.6 6.6E-14 1.4E-18 109.1 12.8 96 40-139 97-192 (417)
5 PHA03235 DNA packaging protein 99.6 3.7E-14 8.1E-19 110.2 11.2 99 38-139 29-129 (409)
6 KOG4220|consensus 99.5 6.2E-16 1.3E-20 117.6 -0.4 98 41-139 30-127 (503)
7 PHA03087 G protein-coupled che 99.5 4.9E-14 1.1E-18 106.4 9.6 98 39-139 38-135 (335)
8 PF00001 7tm_1: 7 transmembran 99.3 2.4E-12 5.2E-17 91.8 6.8 81 58-139 1-81 (257)
9 PF10320 7TM_GPCR_Srsx: Serpen 98.5 7.1E-08 1.5E-12 70.9 3.5 75 53-129 2-76 (257)
10 KOG2087|consensus 98.2 1.1E-06 2.4E-11 66.6 3.1 92 42-134 25-124 (363)
11 PF05296 TAS2R: Mammalian tast 97.9 0.00044 9.5E-09 52.1 11.5 95 40-134 5-102 (303)
12 PF05462 Dicty_CAR: Slime mold 97.8 0.0003 6.6E-09 53.0 9.5 87 43-133 8-94 (303)
13 PF11710 Git3: G protein-coupl 97.7 0.00064 1.4E-08 48.3 9.8 60 70-129 30-89 (201)
14 PF10328 7TM_GPCR_Srx: Serpent 96.8 0.012 2.6E-07 43.4 8.4 45 51-95 3-47 (274)
15 PF10324 7TM_GPCR_Srw: Serpent 96.4 0.02 4.4E-07 42.9 7.4 51 51-102 6-57 (318)
16 PF10321 7TM_GPCR_Srt: Serpent 96.1 0.065 1.4E-06 40.7 8.7 75 39-119 30-105 (313)
17 PF03402 V1R: Vomeronasal orga 96.0 0.018 4E-07 42.6 5.3 61 70-132 5-66 (265)
18 PF10317 7TM_GPCR_Srd: Serpent 94.6 0.27 5.8E-06 36.7 7.9 50 47-96 4-54 (292)
19 PF00002 7tm_2: 7 transmembran 87.9 0.28 6.2E-06 35.2 1.4 75 53-130 12-88 (242)
20 PF10292 7TM_GPCR_Srab: Serpen 84.1 16 0.00035 27.7 9.8 97 42-139 17-120 (324)
21 PF02101 Ocular_alb: Ocular al 78.9 19 0.00042 28.3 8.0 95 43-137 28-140 (405)
22 KOG4564|consensus 75.2 43 0.00093 27.2 12.2 84 48-131 151-246 (473)
23 PF10327 7TM_GPCR_Sri: Serpent 67.6 21 0.00045 26.9 5.8 64 42-105 9-76 (303)
24 PF09882 DUF2109: Predicted me 66.7 23 0.0005 21.2 4.6 47 52-98 4-50 (78)
25 PF10316 7TM_GPCR_Srbc: Serpen 66.6 32 0.00069 25.7 6.5 61 42-102 6-66 (273)
26 PF01102 Glycophorin_A: Glycop 57.2 15 0.00032 24.1 2.9 11 61-71 83-93 (122)
27 PF10323 7TM_GPCR_Srv: Serpent 57.2 46 0.001 24.7 6.0 41 56-96 9-53 (283)
28 PF11446 DUF2897: Protein of u 53.7 20 0.00042 20.0 2.7 22 46-67 6-27 (55)
29 TIGR01477 RIFIN variant surfac 52.8 20 0.00044 27.8 3.4 31 44-74 310-340 (353)
30 PTZ00046 rifin; Provisional 50.0 24 0.00052 27.5 3.5 30 45-74 316-345 (358)
31 PF06024 DUF912: Nucleopolyhed 46.0 29 0.00064 21.7 2.9 11 63-73 83-93 (101)
32 PF02009 Rifin_STEVOR: Rifin/s 45.6 23 0.0005 26.9 2.8 27 47-73 259-285 (299)
33 PF02532 PsbI: Photosystem II 42.4 42 0.00091 16.9 2.5 18 42-59 9-26 (36)
34 PF15330 SIT: SHP2-interacting 42.2 53 0.0011 21.0 3.7 28 44-71 3-30 (107)
35 KOG4193|consensus 40.6 2E+02 0.0043 24.3 7.6 63 60-130 338-402 (610)
36 KOG2927|consensus 39.9 95 0.0021 24.3 5.3 9 105-113 255-263 (372)
37 PF10326 7TM_GPCR_Str: Serpent 39.8 28 0.0006 25.9 2.5 48 49-96 6-54 (307)
38 COG1230 CzcD Co/Zn/Cd efflux s 37.7 1.8E+02 0.004 22.1 9.3 70 47-119 126-196 (296)
39 PRK03557 zinc transporter ZitB 36.8 1.9E+02 0.004 21.9 10.0 73 50-123 126-198 (312)
40 PF08114 PMP1_2: ATPase proteo 35.6 13 0.00029 19.4 0.1 24 48-71 11-34 (43)
41 cd07912 Tweety_N N-terminal do 34.5 2.4E+02 0.0053 22.6 8.4 15 76-90 79-93 (418)
42 PF12606 RELT: Tumour necrosis 33.2 51 0.0011 18.0 2.2 23 47-71 6-28 (50)
43 PF10873 DUF2668: Protein of u 30.1 59 0.0013 22.0 2.5 32 43-74 63-94 (155)
44 KOG1482|consensus 29.9 1.2E+02 0.0026 23.9 4.5 78 53-130 184-279 (379)
45 PF10319 7TM_GPCR_Srj: Serpent 28.6 1.2E+02 0.0025 23.4 4.2 48 47-94 10-58 (310)
46 KOG4349|consensus 26.1 2E+02 0.0043 19.0 7.8 21 108-128 114-134 (143)
47 PF05545 FixQ: Cbb3-type cytoc 26.1 1.1E+02 0.0025 16.1 3.8 9 61-69 25-33 (49)
48 PF05961 Chordopox_A13L: Chord 25.9 1.1E+02 0.0024 17.8 2.9 13 61-73 17-29 (68)
49 CHL00024 psbI photosystem II p 24.5 34 0.00074 17.3 0.5 17 43-59 10-26 (36)
50 PF06679 DUF1180: Protein of u 24.1 2.5E+02 0.0054 19.4 4.9 26 42-67 93-118 (163)
51 PRK13664 hypothetical protein; 24.0 1.1E+02 0.0023 17.3 2.5 17 47-63 10-26 (62)
52 PRK02655 psbI photosystem II r 22.1 38 0.00083 17.2 0.4 16 43-58 10-25 (38)
53 PF10329 DUF2417: Region of un 21.4 2.2E+02 0.0047 20.9 4.3 15 76-90 104-118 (232)
54 PHA03049 IMV membrane protein; 20.8 1.7E+02 0.0037 17.0 2.9 17 53-71 11-27 (68)
No 1
>KOG4219|consensus
Probab=99.77 E-value=4.2e-19 Score=134.25 Aligned_cols=103 Identities=31% Similarity=0.627 Sum_probs=96.3
Q ss_pred cchHHHHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccch
Q psy16860 36 IHIFPFIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCD 115 (139)
Q Consensus 36 ~~~~~~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~ 115 (139)
...+.+.+...++.++.+++++||.+|+|++..+|++|+.+|+|++|||++|++.+++..|+.........|.+|...|+
T Consensus 30 lp~~~~~~wai~yg~l~~vAv~GN~iVlwIil~hrrMRtvtnyfL~NLAfADl~~s~Fn~~f~f~yal~~~W~~G~f~C~ 109 (423)
T KOG4219|consen 30 LPAWQQALWAIAYGLLVFVAVVGNLIVLWIILAHRRMRTVTNYFLVNLAFADLSMSIFNTVFNFQYALHQEWYFGSFYCR 109 (423)
T ss_pred CCHHHHHHHHHHHHHHHHHHHhcCceEEEEEeehhehhhhHHHHHHHHHHHHHHHHHHhhHHHHHHHHHhccccccceee
Confidence 34555889999999999999999999999999999999999999999999999999999999988888899999999999
Q ss_pred hchHHhHHhhHHHHHHHHHhhccC
Q psy16860 116 VWNSFDVYFSQYRKHFNMYEVDFE 139 (139)
Q Consensus 116 ~~~~~~~~~~~~Si~~~m~~~~~~ 139 (139)
+..|+......+|+ ++|+|+++|
T Consensus 110 f~nf~~itav~vSV-fTlvAiA~D 132 (423)
T KOG4219|consen 110 FVNFFPITAVFVSV-FTLVAIAID 132 (423)
T ss_pred eccccchhhhhHhH-HHHHHHHHH
Confidence 99999999999999 788888875
No 2
>PHA03234 DNA packaging protein UL33; Provisional
Probab=99.75 E-value=1.3e-17 Score=126.62 Aligned_cols=99 Identities=14% Similarity=0.165 Sum_probs=82.6
Q ss_pred hHHHHHHHHHHHHHHHHHHHHHHHhHHHh--hccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccch
Q psy16860 38 IFPFIIKGMLMCFIIITAILGNLLVIISV--IKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCD 115 (139)
Q Consensus 38 ~~~~~~~~~~~~~i~~~g~~gN~lvi~v~--~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~ 115 (139)
+..+.....+|.+++++|++||+++++++ .+++++|+++|+|+.|||++|++.++. .|+.... ..+.|.+|+..||
T Consensus 29 ~~~~~~~~~~y~~vf~~gl~gN~lvl~v~~~~~~~~~rt~tn~fi~NLAvaDLL~~l~-lp~~~~~-~~~~w~fG~~lCk 106 (338)
T PHA03234 29 KKAQILESAINGIMLTLIIPMIIIVICTLIIYHKVAKHNATSFYLITLFASDFLHMLC-VFFLTLN-REALFNFNQAFCQ 106 (338)
T ss_pred HHHHHHhhHHHHHHHHHHhhhHHHHHHHHHHHhccccccHHHHHHHHHHHHHHHHHHH-HHHHHHH-HhCCccCchhHHH
Confidence 34478889999999999999999999955 456677999999999999999999765 5555443 3457999999999
Q ss_pred hchHHhHHhhHHHHHHHHHhhccC
Q psy16860 116 VWNSFDVYFSQYRKHFNMYEVDFE 139 (139)
Q Consensus 116 ~~~~~~~~~~~~Si~~~m~~~~~~ 139 (139)
+..++.....++|+ +++.++|+|
T Consensus 107 ~~~~~~~~~~~~Si-~~L~~ISiD 129 (338)
T PHA03234 107 CVLFIYHASCSYSI-CMLAIIATI 129 (338)
T ss_pred HHHHHHHHHHHHHH-HHHHHHHHH
Confidence 99999999999999 557777765
No 3
>PHA02834 chemokine receptor-like protein; Provisional
Probab=99.65 E-value=1.3e-15 Score=114.85 Aligned_cols=95 Identities=23% Similarity=0.426 Sum_probs=80.4
Q ss_pred HHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHH
Q psy16860 41 FIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSF 120 (139)
Q Consensus 41 ~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~ 120 (139)
+.....++.+++++|++||+++++++.++|+ +++.|+|+.|||++|++. .+..|+.+.... ++|.+|+..|++..+.
T Consensus 28 ~~~~~~~~~li~v~~~~gN~lVi~vi~~~~~-~~~~n~~i~nLAiaDll~-~~~lP~~i~~~~-~~w~~g~~~C~~~~~~ 104 (323)
T PHA02834 28 NYFVIVFYILLFIFGLIGNVLVIAVLIVKRF-MFVVDVYLFNIAMSDLML-VFSFPFIIHNDL-NEWIFGEFMCKLVLGV 104 (323)
T ss_pred hhhHHHHHHHHHHHHHhhHHHHHHHHHhccc-cchhhhhhHHHHHHHHHH-HHHHHHHHHHHc-CCcCCcchHHHhHHHH
Confidence 5577889999999999999999999887665 457899999999999987 667898765544 4799999999999998
Q ss_pred hHHhhHHHHHHHHHhhccC
Q psy16860 121 DVYFSQYRKHFNMYEVDFE 139 (139)
Q Consensus 121 ~~~~~~~Si~~~m~~~~~~ 139 (139)
......+|+ +++.++|+|
T Consensus 105 ~~~~~~~Si-~tL~~Isid 122 (323)
T PHA02834 105 YFVGFFSNM-FFVTLISID 122 (323)
T ss_pred HHHHHHHHH-HHHHHHHHH
Confidence 888888888 678888776
No 4
>PHA02638 CC chemokine receptor-like protein; Provisional
Probab=99.57 E-value=6.6e-14 Score=109.10 Aligned_cols=96 Identities=23% Similarity=0.389 Sum_probs=79.5
Q ss_pred HHHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchH
Q psy16860 40 PFIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNS 119 (139)
Q Consensus 40 ~~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~ 119 (139)
.......++.+++++|++||+++++++.+ |++|+++|++++|||++|++. ++..|+.+... .+.|.+|+..||+..+
T Consensus 97 ~~~~l~~~y~lvfvlgliGN~LVl~il~~-k~lrt~t~i~llnLAisDLl~-~l~lPf~i~~~-~~~W~fg~~~Ck~~~~ 173 (417)
T PHA02638 97 ISEYIKIFYIIIFILGLFGNAAIIMILFC-KKIKTITDIYIFNLAISDLIF-VIDFPFIIYNE-FDQWIFGDFMCKVISA 173 (417)
T ss_pred hhhHHHHHHHHHHHHHHHHHHHHHHHHHh-ccCCCHhHHHHHHHHHHHHHH-HHHHHHHHHHH-hccccccccchhhHHH
Confidence 35677888999999999999999987654 778999999999999999988 55788877654 4679999999999999
Q ss_pred HhHHhhHHHHHHHHHhhccC
Q psy16860 120 FDVYFSQYRKHFNMYEVDFE 139 (139)
Q Consensus 120 ~~~~~~~~Si~~~m~~~~~~ 139 (139)
+.....++|++ .+.++++|
T Consensus 174 l~~~~~~~Si~-~L~~isiD 192 (417)
T PHA02638 174 SYYIGFFSNMF-LITLMSID 192 (417)
T ss_pred HHHHHHHHHHH-HHHHHHHH
Confidence 99888888875 55555543
No 5
>PHA03235 DNA packaging protein UL33; Provisional
Probab=99.56 E-value=3.7e-14 Score=110.20 Aligned_cols=99 Identities=14% Similarity=0.085 Sum_probs=76.3
Q ss_pred hHHHHHHHHHHHHHHHHHHHHHHHhHHHhhcc-CC-CCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccch
Q psy16860 38 IFPFIIKGMLMCFIIITAILGNLLVIISVIKH-RK-LRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCD 115 (139)
Q Consensus 38 ~~~~~~~~~~~~~i~~~g~~gN~lvi~v~~~~-~~-~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~ 115 (139)
+..+.+...++.+++++|++||+++++++.++ |+ .++..++|+.|||++|++. ++.+|+.+... ...|..|+..||
T Consensus 29 ~~~~~~~~~~~~li~vvGiigN~lVL~~~~~~~r~~~~~~~~~~I~NLAvsDLl~-l~~lP~~i~~~-~~~~~~g~~~Ck 106 (409)
T PHA03235 29 SAARTTETFINLLIISVGGPLNLIVLVTQLLANRVHGFSTPTLYMTNLYLANLLT-VFVLPFIMLSN-QGLLSGSVAGCK 106 (409)
T ss_pred hhhHhHHHHHHHHHHHHHHHHHHHHHHHHHHhhhcccCCccHHHHHHHHHHHHHH-HHHHHHHHHhc-CccccCCCCeeh
Confidence 44578899999999999999999999986533 32 2356679999999999987 66788776432 122334578999
Q ss_pred hchHHhHHhhHHHHHHHHHhhccC
Q psy16860 116 VWNSFDVYFSQYRKHFNMYEVDFE 139 (139)
Q Consensus 116 ~~~~~~~~~~~~Si~~~m~~~~~~ 139 (139)
+..++....+.+|+ .++.++++|
T Consensus 107 ~~~~l~~~~~~~Si-~tL~~ISiD 129 (409)
T PHA03235 107 FASLLYYASCTVGF-ATVALIAAD 129 (409)
T ss_pred hHHHHHHHHHHHHH-HHHHHHHHH
Confidence 99999999999998 567777765
No 6
>KOG4220|consensus
Probab=99.54 E-value=6.2e-16 Score=117.61 Aligned_cols=98 Identities=30% Similarity=0.569 Sum_probs=90.4
Q ss_pred HHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHH
Q psy16860 41 FIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSF 120 (139)
Q Consensus 41 ~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~ 120 (139)
-...+++..++.++.++||++|++.+...|++++..|||+++||++|++++.+.+|+...+.+-|+|.+|...|.+.-.+
T Consensus 30 ~v~i~~v~~~lsLVTv~GNlLVmiSfKvnrqLqTVnNYfLfSLAcADliIG~~SMnl~t~Y~lmg~W~LG~~~CdlWLal 109 (503)
T KOG4220|consen 30 VVFIVVVTGSLSLVTVVGNLLVMISFKVNRQLQTVNNYFLFSLACADLIIGAFSMNLYTTYTLMGYWPLGPLVCDLWLAL 109 (503)
T ss_pred EEeeehhhhHHHHHhhhccEEEEEEEEecceeeeecceeehHHHHhhhhhheeechHHHHHHHHcccccchHHHHHHHHH
Confidence 34566677888999999999999999999999999999999999999999999999999999999999999999999999
Q ss_pred hHHhhHHHHHHHHHhhccC
Q psy16860 121 DVYFSQYRKHFNMYEVDFE 139 (139)
Q Consensus 121 ~~~~~~~Si~~~m~~~~~~ 139 (139)
.++...+|+ +|+-.++||
T Consensus 110 DYvaSNASV-mNLLiISFD 127 (503)
T KOG4220|consen 110 DYVASNASV-MNLLIISFD 127 (503)
T ss_pred HHHhhhhhh-hhhheeeee
Confidence 999999999 677777775
No 7
>PHA03087 G protein-coupled chemokine receptor-like protein; Provisional
Probab=99.54 E-value=4.9e-14 Score=106.43 Aligned_cols=98 Identities=21% Similarity=0.403 Sum_probs=83.0
Q ss_pred HHHHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhch
Q psy16860 39 FPFIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWN 118 (139)
Q Consensus 39 ~~~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~ 118 (139)
..+.+...++.+++++|++||+++++++.++ ++|++.|+++.|||++|++.++ ..|........++|.+|+..|+...
T Consensus 38 ~~~~~~~~~~~~i~~~gl~gN~lvl~~~~~~-~~~~~~~~ll~~laisDll~~~-~~~~~~~~~~~~~~~~~~~~C~~~~ 115 (335)
T PHA03087 38 TNSTILIVVYSTIFFFGLVGNIIVIYVLTKT-KIKTPMDIYLLNLAVSDLLFVM-TLPFQIYYYILFQWSFGEFACKIVS 115 (335)
T ss_pred chhhHHHHHHHHHHHHHHHhhHhEEeeehhc-cccCchHHHHHHHHHHHHHHHH-hHHHHHHHHhCCCCCCCcHHHHHHH
Confidence 3466788899999999999999999998887 8899999999999999998865 4676665666678999999999999
Q ss_pred HHhHHhhHHHHHHHHHhhccC
Q psy16860 119 SFDVYFSQYRKHFNMYEVDFE 139 (139)
Q Consensus 119 ~~~~~~~~~Si~~~m~~~~~~ 139 (139)
++......+|+++ +.++++|
T Consensus 116 ~~~~~~~~~S~~~-l~~iaid 135 (335)
T PHA03087 116 GLYYIGFYNSMNF-ITVMSVD 135 (335)
T ss_pred HHHHHHHHHHHHH-HHHHHHH
Confidence 9999999999854 6666654
No 8
>PF00001 7tm_1: 7 transmembrane receptor (rhodopsin family) Rhodopsin-like GPCR superfamily signature 5-hydroxytryptamine 7 receptor signature bradykinin receptor signature gastrin receptor signature melatonin receptor signature olfactory receptor signature; InterPro: IPR000276 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The rhodopsin-like GPCRs themselves represent a widespread protein family that includes hormone, neurotransmitter and light receptors, all of which transduce extracellular signals through interaction with guanine nucleotide-binding (G) proteins. Although their activating ligands vary widely in structure and character, the amino acid sequences of the receptors are very similar and are believed to adopt a common structural framework comprising 7 transmembrane (TM) helices [, , ].; GO: 0007186 G-protein coupled receptor protein signaling pathway, 0016021 integral to membrane; PDB: 2KI9_A 3QAK_A 2YDV_A 3VGA_A 3PWH_A 3RFM_A 3EML_A 3VG9_A 3REY_A 3UZA_A ....
Probab=99.35 E-value=2.4e-12 Score=91.84 Aligned_cols=81 Identities=31% Similarity=0.634 Sum_probs=70.9
Q ss_pred HHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHHhHHhhHHHHHHHHHhhc
Q psy16860 58 GNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSFDVYFSQYRKHFNMYEVD 137 (139)
Q Consensus 58 gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~~~~~~~~Si~~~m~~~~ 137 (139)
||+++++++.++|++|++.++++.|||++|++.++...|........++|..|+..|++..++.......|.+ ++.+++
T Consensus 1 GN~lvi~~~~~~~~~~~~~~~~l~~Lav~Dll~~~~~~~~~~~~~~~~~~~~~~~~C~~~~~~~~~~~~~s~~-~~~~is 79 (257)
T PF00001_consen 1 GNILVILVILRSKRLRTPSNILLLNLAVADLLVGLFCIPFYIYSLLFDDWIFSSFLCRIFGFLFYFSSFSSIF-SLVAIS 79 (257)
T ss_dssp HHHHHHHHHHHSGGG-SHHHHHHHHHHHHHHHHHHTHHHHHHHHHHHSSCTSHHHHHHHHHHHHHHHHHHHHH-HHHHHH
T ss_pred CchhehhhhhhhccCCChhHHHHHHHHHHHHhhcccccccccccccccccccccccccccccccccccccccc-cccccc
Confidence 8999999999999999999999999999999999999998887777788999999999999999988888884 555555
Q ss_pred cC
Q psy16860 138 FE 139 (139)
Q Consensus 138 ~~ 139 (139)
+|
T Consensus 80 ~d 81 (257)
T PF00001_consen 80 ID 81 (257)
T ss_dssp HH
T ss_pred cc
Confidence 44
No 9
>PF10320 7TM_GPCR_Srsx: Serpentine type 7TM GPCR chemoreceptor Srsx; InterPro: IPR019424 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class sx (Srsx), which is a solo family amongst the superfamilies of chemoreceptors. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' [].
Probab=98.53 E-value=7.1e-08 Score=70.86 Aligned_cols=75 Identities=28% Similarity=0.467 Sum_probs=59.7
Q ss_pred HHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHHhHHhhHHHH
Q psy16860 53 ITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSFDVYFSQYRK 129 (139)
Q Consensus 53 ~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~~~~~~~~Si 129 (139)
++|++||..++.++.|+|++|+|.++++..+|++|++......|..... +.+. ......|-.+.+...++..+..
T Consensus 2 ~ig~~gN~~~i~~~~~~~~Lrs~~~~li~~~~~~d~~~~~~~~~~~~~~-~~~~-~i~~~~Cf~~~~~~~f~~~~qs 76 (257)
T PF10320_consen 2 IIGLFGNLLLIILIFRNKSLRSPCYILICILCFADLICLLGTLPFMLFL-FRDH-QITRSECFWQIFFYIFFQCAQS 76 (257)
T ss_pred EEEEEccHHHHHHHHhccccccchHHHHHHHHHHHHHHHhhHHHHHHHH-Hhhe-eccHHHHHHHHHHHHHHHHHHH
Confidence 4688999999999999999999999999999999999988888866633 3332 3567789877776665555443
No 10
>KOG2087|consensus
Probab=98.20 E-value=1.1e-06 Score=66.62 Aligned_cols=92 Identities=25% Similarity=0.369 Sum_probs=71.2
Q ss_pred HHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhc-C-------cccccccc
Q psy16860 42 IIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLT-D-------EWYFGYFM 113 (139)
Q Consensus 42 ~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~-~-------~~~~g~~~ 113 (139)
.+.-....++..+++.||.+|++.+...|...+...+++.|||++|++.++-..-...+.... + .| .+...
T Consensus 25 ~~lRi~vW~i~~lAi~gN~~Vl~~~~~~~~~~~~~~~li~~la~ad~~mGiYl~~ia~vD~~~~gey~~~ai~W-~tg~g 103 (363)
T KOG2087|consen 25 WILRISVWVIALLAIVGNLLVLLTRFTSRYELNSHRFLICNLAFADLLMGIYLGLIASVDAKTRGEYYKHAIDW-QTGLG 103 (363)
T ss_pred ceeeehhhhhhhHHhccCeeeeeeeeehhhhccchHHHHHHHHHHHHHcchHHHHHHHhhHHHHHHHHHHHHhh-hhcCC
Confidence 344445667888899999999999888888788889999999999999988655544443332 1 35 35678
Q ss_pred chhchHHhHHhhHHHHHHHHH
Q psy16860 114 CDVWNSFDVYFSQYRKHFNMY 134 (139)
Q Consensus 114 C~~~~~~~~~~~~~Si~~~m~ 134 (139)
|++.+|+..+..-.|+++.+.
T Consensus 104 C~~aGflavFASElSv~~LT~ 124 (363)
T KOG2087|consen 104 CPVAGFLAVFASELSVFLLTL 124 (363)
T ss_pred CchHHHHHHHHHHHHHHHHHH
Confidence 999999999999999965543
No 11
>PF05296 TAS2R: Mammalian taste receptor protein (TAS2R); InterPro: IPR007960 This family consists of several forms of mammalian taste receptor proteins (TAS2Rs). TAS2Rs are G protein-coupled receptors expressed in subsets of taste receptor cells of the tongue and palate epithelia and are organised in the genome in clusters. The proteins are genetically linked to loci that influence bitter perception in mice and humans [].; GO: 0004930 G-protein coupled receptor activity, 0007186 G-protein coupled receptor protein signaling pathway, 0050909 sensory perception of taste, 0016021 integral to membrane
Probab=97.86 E-value=0.00044 Score=52.10 Aligned_cols=95 Identities=15% Similarity=0.200 Sum_probs=70.6
Q ss_pred HHHHHHHHHHHHHHHHHHHHHHhHHHhh---ccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchh
Q psy16860 40 PFIIKGMLMCFIIITAILGNLLVIISVI---KHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDV 116 (139)
Q Consensus 40 ~~~~~~~~~~~i~~~g~~gN~lvi~v~~---~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~ 116 (139)
.+.+...+..+.+++|++||+.++.+-+ +++|.-.|.+..+.+||++.++.-....-......+.......+..++.
T Consensus 5 ~~~i~~~i~~~~~~~Gi~~N~FI~~vn~~~w~k~~~l~~~d~IL~~La~sr~~l~~~~~~~~~~~~~~~~~~~~~~~~~~ 84 (303)
T PF05296_consen 5 LEIIFLIILVVEFIIGILGNGFIVLVNCSDWVKSRKLSPSDQILTSLAISRILLQWVILLNSFLSFFFPNIYFSENVYKI 84 (303)
T ss_pred HHHHHHHHHHHHHHHHHHHHHHHHHHhHHHHHcCCCCChHHHHHHHHHHHHHHHHHHHHHHHHHHHHcchhhhhhhHHHH
Confidence 3567788899999999999998886665 3344456899999999999999866544444444444444456678888
Q ss_pred chHHhHHhhHHHHHHHHH
Q psy16860 117 WNSFDVYFSQYRKHFNMY 134 (139)
Q Consensus 117 ~~~~~~~~~~~Si~~~m~ 134 (139)
..+...+....|++++..
T Consensus 85 ~~~~~~f~~~~s~W~tt~ 102 (303)
T PF05296_consen 85 IDFLWMFSNSSSLWFTTW 102 (303)
T ss_pred HHHHHHHHhHHHHHHHHH
Confidence 888888888888887654
No 12
>PF05462 Dicty_CAR: Slime mold cyclic AMP receptor
Probab=97.78 E-value=0.0003 Score=53.00 Aligned_cols=87 Identities=18% Similarity=0.235 Sum_probs=67.0
Q ss_pred HHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHHhH
Q psy16860 43 IKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSFDV 122 (139)
Q Consensus 43 ~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~~~ 122 (139)
....+..+..+++++|-+.++..+++.|++|++.+.++.-++++|++..+...... .. +.-..+...|++++++..
T Consensus 8 ~~~~i~~~~s~lSllGclfiI~tf~~~k~~r~~~~rli~yl~~~~ll~~v~~~~~~---~~-~~~~~~s~lC~~Qafliq 83 (303)
T PF05462_consen 8 TLYAIELVASVLSLLGCLFIIITFCLFKRLRKPINRLIFYLSIANLLTNVASMIMT---LS-PSAGENSFLCQFQAFLIQ 83 (303)
T ss_pred HHHHHHHHHHHHHHHHHHHHHHHHHHHHHhCccHHHHHHHHHHHHHHHHHHHHHHH---hc-ccCCCCCcchhhHhHHHH
Confidence 34455666678888999999999999999999999999999999999765433221 11 222345778999999999
Q ss_pred HhhHHHHHHHH
Q psy16860 123 YFSQYRKHFNM 133 (139)
Q Consensus 123 ~~~~~Si~~~m 133 (139)
++..++.+.++
T Consensus 84 ~f~~as~lWt~ 94 (303)
T PF05462_consen 84 FFMLASFLWTL 94 (303)
T ss_pred HhhHHHHHHHH
Confidence 99999987554
No 13
>PF11710 Git3: G protein-coupled glucose receptor regulating Gpa2; InterPro: IPR023041 This entry contains a functionally uncharacterised region belonging to the Git3 G-protein coupled receptor. Git3 is one of six proteins required for glucose-triggered adenylate cyclase activation, and is a G protein-coupled receptor responsible for the activation of adenylate cyclase through Gpa2 - heterotrimeric G protein alpha subunit, part of the glucose-detection pathway. Git3 contains seven predicted transmembrane domains, a third cytoplasmic loop and a cytoplasmic tail []. This is the conserved N-terminal domain of the member proteins.
Probab=97.71 E-value=0.00064 Score=48.33 Aligned_cols=60 Identities=13% Similarity=0.062 Sum_probs=44.0
Q ss_pred CCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHHhHHhhHHHH
Q psy16860 70 RKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSFDVYFSQYRK 129 (139)
Q Consensus 70 ~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~~~~~~~~Si 129 (139)
+++|.-.+.++.||.++|++.++........+...++-.-+...|..++++....-.++=
T Consensus 30 ~r~~~fR~~LIl~L~~aD~~qal~~~i~~~~~l~~~~i~~~s~~C~aqGf~~q~g~~~sd 89 (201)
T PF11710_consen 30 YRRRSFRHQLILNLLLADFIQALAFLISPIRWLARGGIIAPSPFCQAQGFFLQVGDEASD 89 (201)
T ss_pred hhhhhHHHHHHHHHHHHHHHHHHHHHHHHHHHHhcCCeeCCCCchhhhHHHHHHHHHHHH
Confidence 445667778999999999999887665455555555444567999999998776654443
No 14
>PF10328 7TM_GPCR_Srx: Serpentine type 7TM GPCR chemoreceptor Srx; InterPro: IPR019430 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class x (Srx) from the Srg superfamily [, ]. Srg receptors contain seven hydrophobic, putative transmembrane, regions and can be distinguished from other 7TM GPCR receptors by their own characteristic TM signatures.
Probab=96.79 E-value=0.012 Score=43.44 Aligned_cols=45 Identities=31% Similarity=0.428 Sum_probs=40.6
Q ss_pred HHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhh
Q psy16860 51 IIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVM 95 (139)
Q Consensus 51 i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~ 95 (139)
+.+.|++.|.++++.+.|.|++|++.+.+-.+.|++|.+.++...
T Consensus 3 ~s~~G~~~N~~v~~~~~~~~~~~~sF~~l~~~~a~~n~i~~~~~l 47 (274)
T PF10328_consen 3 ISIIGIILNWLVFIIIFKLKSLRNSFGILCASQAIANIIICLIFL 47 (274)
T ss_pred eeHHHHHHHHHHHHHHHhcccccCCHHHHHHHHHHHHHHHHHHHH
Confidence 467899999999999999999999999999999999999977543
No 15
>PF10324 7TM_GPCR_Srw: Serpentine type 7TM GPCR chemoreceptor Srw; InterPro: IPR019427 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class w (Srw), which is a solo family amongst the superfamilies of chemoreceptors. The genes encoding Srw do not appear to be under as strong an adaptive evolutionary pressure as those of Srz [].
Probab=96.37 E-value=0.02 Score=42.94 Aligned_cols=51 Identities=20% Similarity=0.399 Sum_probs=41.5
Q ss_pred HHHHHHHHHHHhHHHhhccCCCCc-hHHHHHHHHHHHHHHHHHhhhHHHHHHH
Q psy16860 51 IIITAILGNLLVIISVIKHRKLRI-ITNYFVVSLAFADLLVALCVMPFNAIVS 102 (139)
Q Consensus 51 i~~~g~~gN~lvi~v~~~~~~~~~-~~~~~i~nLa~~Dl~~~l~~~p~~~~~~ 102 (139)
+.++|+++|.+-+.++. +|++|+ +.|.++..+|++|++..+...+......
T Consensus 6 ~~~~g~~~N~~h~~VLt-rk~mR~~~in~~l~~Iai~Dl~~~~~~~~~~~~~~ 57 (318)
T PF10324_consen 6 LSIFGLFINIFHLIVLT-RKSMRSSSINILLIGIAICDLLYMLSILIWELFFF 57 (318)
T ss_pred EeHHHHHHHHHHhhhcC-ChhhhcCCHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence 46789999999997764 566676 8999999999999999888777665443
No 16
>PF10321 7TM_GPCR_Srt: Serpentine type 7TM GPCR chemoreceptor Srt; InterPro: IPR019425 Chemoreception is mediated in Caenorhabditis elegans by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs) of proteins which are of the serpentine type []. Srt is a member of the Srg superfamily of chemoreceptors. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' [].
Probab=96.07 E-value=0.065 Score=40.71 Aligned_cols=75 Identities=17% Similarity=0.208 Sum_probs=54.6
Q ss_pred HHHHHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCc-cccccccchhc
Q psy16860 39 FPFIIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDE-WYFGYFMCDVW 117 (139)
Q Consensus 39 ~~~~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~-~~~g~~~C~~~ 117 (139)
..+......+.+.+++-.+-...++.++.++|+.|.+.+....-|++.|++......- .+|- -..|..+|+.-
T Consensus 30 ~~~p~~G~~~~~~g~~~~~lY~p~~~~i~~~~~~k~~~ykiM~~L~i~Di~~l~~~si------~tG~l~i~G~vfC~~P 103 (313)
T PF10321_consen 30 VKRPILGIYFLIFGIIIIILYIPCLIAIFKKKLFKMSCYKIMFFLAIFDIIQLFINSI------ITGILAIFGAVFCSYP 103 (313)
T ss_pred CcccchhHHHHHHHHHHHHHHHHHHHHHHHhccccCcHHHHHHHHHHHHHHHHHhhhh------hhhHHHhcCccccCCc
Confidence 3355677778888888888889999999988888899999999999999998554322 2221 12466777654
Q ss_pred hH
Q psy16860 118 NS 119 (139)
Q Consensus 118 ~~ 119 (139)
.+
T Consensus 104 ~~ 105 (313)
T PF10321_consen 104 RF 105 (313)
T ss_pred hH
Confidence 44
No 17
>PF03402 V1R: Vomeronasal organ pheromone receptor family, V1R; InterPro: IPR004072 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The rhodopsin-like GPCRs themselves represent a widespread protein family that includes hormone, neurotransmitter and light receptors, all of which transduce extracellular signals through interaction with guanine nucleotide-binding (G) proteins. Although their activating ligands vary widely in structure and character, the amino acid sequences of the receptors are very similar and are believed to adopt a common structural framework comprising 7 transmembrane (TM) helices [, , ]. Pheromones have evolved in all animal phyla, to signal sex and dominance status, and are responsible for stereotypical social and sexual behaviour among members of the same species. In mammals, these chemical signals are believed to be detected primarily by the vomeronasal organ (VNO), a chemosensory organ located at the base of the nasal septum []. The VNO is present in most amphibia, reptiles and non-primate mammals but is absent in birds, adult catarrhine monkeys and apes []. An active role for the human VNO in the detection of pheromones is disputed; the VNO is clearly present in the foetus but appears to be atrophied or absent in adults. Three distinct families of putative pheromone receptors have been identified in the vomeronasal organ (V1Rs, V2Rs and V3Rs). All are G protein-coupled receptors but are only distantly related to the receptors of the main olfactory system, highlighting their different role []. The V1 receptors share between 50 and 90% sequence identity but have little similarity to other families of G protein-coupled receptors. They appear to be distantly related to the mammalian T2R bitter taste receptors and the rhodopsin-like GPCRs []. In rat, the family comprises 30-40 genes. These are expressed in the apical regions of the VNO, in neurons expressing Gi2. Coupling of the receptors to this protein mediates inositol trisphosphate signalling []. A number of human V1 receptor homologues have also been found. The majority of these human sequences are pseudogenes [] but an apparently functional receptor has been identified that is expressed in the human olfactory system [].; GO: 0016503 pheromone receptor activity, 0007186 G-protein coupled receptor protein signaling pathway, 0016021 integral to membrane
Probab=96.00 E-value=0.018 Score=42.64 Aligned_cols=61 Identities=15% Similarity=0.163 Sum_probs=44.7
Q ss_pred CCCCchHHHHHHHHHHHHHHHHHh-hhHHHHHHHhcCccccccccchhchHHhHHhhHHHHHHH
Q psy16860 70 RKLRIITNYFVVSLAFADLLVALC-VMPFNAIVSLTDEWYFGYFMCDVWNSFDVYFSQYRKHFN 132 (139)
Q Consensus 70 ~~~~~~~~~~i~nLa~~Dl~~~l~-~~p~~~~~~~~~~~~~g~~~C~~~~~~~~~~~~~Si~~~ 132 (139)
.++.+|++..+.|||+++.++.+. ++| ........+ .+++..||+..|+.-..-+.|+-++
T Consensus 5 ~~r~kp~dlIl~hLa~aN~lvLl~rGip-~~~~~~~~~-~~~d~gCK~v~Y~~RV~RglSictT 66 (265)
T PF03402_consen 5 GHRLKPIDLILIHLALANILVLLSRGIP-QTMAFFGWK-FFDDIGCKIVFYIYRVARGLSICTT 66 (265)
T ss_pred CCCCCcHHHHHHHHHHHHHHHHHHhhHH-HHHHHhhcc-cCCCceeeeeeeehHHhchhhHHhh
Confidence 455788999999999999998654 344 333333323 3689999999999888888877543
No 18
>PF10317 7TM_GPCR_Srd: Serpentine type 7TM GPCR chemoreceptor Srd; InterPro: IPR019421 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents the chemoreceptor Srd [].
Probab=94.64 E-value=0.27 Score=36.72 Aligned_cols=50 Identities=18% Similarity=0.224 Sum_probs=38.8
Q ss_pred HHHHHHHHHHHHHHHhHHHhhccC-CCCchHHHHHHHHHHHHHHHHHhhhH
Q psy16860 47 LMCFIIITAILGNLLVIISVIKHR-KLRIITNYFVVSLAFADLLVALCVMP 96 (139)
Q Consensus 47 ~~~~i~~~g~~gN~lvi~v~~~~~-~~~~~~~~~i~nLa~~Dl~~~l~~~p 96 (139)
++.+.+.+|++.|.+.++.+.++. +.-+...+++.|-|+.|++.+....-
T Consensus 4 ~~~~~~~~~~~~n~~Ll~~i~~~tp~~l~~~~~~l~~~~~~~~~~~~~~~~ 54 (292)
T PF10317_consen 4 YHPIFFILGIILNILLLYLIIFKTPKSLRTYSILLLNTAIFDLISIISAFL 54 (292)
T ss_pred eHHHHHHHHHHHHHHHHHHHHHhChHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence 466788999999999998777554 33445679999999999999765433
No 19
>PF00002 7tm_2: 7 transmembrane receptor (Secretin family); InterPro: IPR000832 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The secretin-like GPCRs include secretin [], calcitonin [], parathyroid hormone/parathyroid hormone-related peptides [] and vasoactive intestinal peptide [], all of which activate adenylyl cyclase and the phosphatidyl-inositol-calcium pathway. These receptors contain seven transmembrane regions, in a manner reminiscent of the rhodopsins and other receptors believed to interact with G-proteins (however there is no significant sequence identity between these families, the secretin-like receptors thus bear their own unique '7TM' signature). Their N terminus is probably located on the extracellular side of the membrane and potentially glycosylated. This N-terminal region contains a long conserved region which allow the binding of large peptidic ligand such as glucagon, secretin, VIP and PACAP; this region contains five conserved cysteines residues which could be involved in disulphide bond. The C-terminal region of these receptor is probably cytoplasmic. Every receptor gene in this family is encoded on multiple exons, and several of these genes are alternatively spliced to yield functionally distinct products. ; GO: 0004930 G-protein coupled receptor activity, 0007186 G-protein coupled receptor protein signaling pathway, 0016021 integral to membrane; PDB: 3L2J_A 1BL1_A.
Probab=87.94 E-value=0.28 Score=35.22 Aligned_cols=75 Identities=25% Similarity=0.215 Sum_probs=0.4
Q ss_pred HHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCcccc--ccccchhchHHhHHhhHHHHH
Q psy16860 53 ITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYF--GYFMCDVWNSFDVYFSQYRKH 130 (139)
Q Consensus 53 ~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~--g~~~C~~~~~~~~~~~~~Si~ 130 (139)
.+++++-.+.+......|++|+..+....||++++++..+..+.. .....+... .+..|+....+.+.+..++..
T Consensus 12 ~~Si~~ll~~i~~~~~~r~lr~~~~~i~~~l~~sll~~~~~~l~~---~~~~~~~~~~~~~~~C~~~a~~~hy~~la~f~ 88 (242)
T PF00002_consen 12 SLSIICLLLTIITYLLFRKLRSFRNKIHLNLCLSLLLANLSFLIG---ISQTFSPISTTNHCLCRAIAILLHYFFLASFF 88 (242)
T ss_dssp H-------------------------------------------------------------------------------
T ss_pred HHHHHHHHHHHHHHHHHHhhcccchhhhhhhHHHHHHHHHHHhee---hhhccccccccccccchhhhhHhHHHHHHHHH
Confidence 333444444444445557777777888999999998876543221 111111111 223599998877776655543
No 20
>PF10292 7TM_GPCR_Srab: Serpentine type 7TM GPCR receptor class ab chemoreceptor; InterPro: IPR019408 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. Srab is part of the Sra superfamily of chemoreceptors. The expression pattern of the srab genes is biologically intriguing. Of the six promoters successfully expressed in transgenic organisms, one was exclusively expressed in the tail phasmid neurons, two were exclusively expressed in a head amphid neuron, and two were expressed both in the head and tail neurons as well as a limited number of other cells [].
Probab=84.12 E-value=16 Score=27.67 Aligned_cols=97 Identities=8% Similarity=0.039 Sum_probs=59.0
Q ss_pred HHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHH---HHHHhc---C-ccccccccc
Q psy16860 42 IIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFN---AIVSLT---D-EWYFGYFMC 114 (139)
Q Consensus 42 ~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~---~~~~~~---~-~~~~g~~~C 114 (139)
.....+-.++.++|++.++..++...+++..|....+.+....++.++-++...-.. +..+.. + +.......|
T Consensus 17 ~~~~~~~~~~s~~~~~~~~~~~~~~~~~~~~H~N~ril~~~~~~~~l~~~~~r~~~h~~~l~~~~~~~~~Cd~~~~~~~C 96 (324)
T PF10292_consen 17 RLSLIFNLLLSIIAFPVIIYALWKIRNSKLFHFNTRILFIVHCFSFLIHCTGRIILHTYDLYNYFFPDDPCDMIPSTYRC 96 (324)
T ss_pred HHHHHHHHHHHHHHHHHHHHHHHHHhhcchhchhHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHHccCCCcccccchhHH
Confidence 344445677777888787777777777777888888888888888887765433322 222222 1 233455667
Q ss_pred hhchHHhHHhhHHHHHHHHHhhccC
Q psy16860 115 DVWNSFDVYFSQYRKHFNMYEVDFE 139 (139)
Q Consensus 115 ~~~~~~~~~~~~~Si~~~m~~~~~~ 139 (139)
-............+. .+..++.+|
T Consensus 97 ~~lR~~~~~~~~~~~-~t~v~l~IE 120 (324)
T PF10292_consen 97 FILRIPYNFGLFLVS-FTTVSLVIE 120 (324)
T ss_pred HHHHHHHHHHHHHHH-HHHHHHHHH
Confidence 655555555544444 555555554
No 21
>PF02101 Ocular_alb: Ocular albinism type 1 protein; InterPro: IPR001414 Ocular albinism type 1 (OA1) is an X-linked disorder characterised by severe impairment of visual acuity, retinal hypopigmentation and the presence of macromelanosomes. A novel transcript from the OA1 critical region is expressed in high levels in RNA samples from retina and from melanoma and encodes a potential integral membrane protein []. This protein is of unknown function but is known to bind heterotrimeric G proteins.; GO: 0016020 membrane
Probab=78.89 E-value=19 Score=28.33 Aligned_cols=95 Identities=19% Similarity=0.144 Sum_probs=53.6
Q ss_pred HHHHHHHHHHHHHHHHHHHhHHHhhccC----CCCc--hHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccc--------
Q psy16860 43 IKGMLMCFIIITAILGNLLVIISVIKHR----KLRI--ITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWY-------- 108 (139)
Q Consensus 43 ~~~~~~~~i~~~g~~gN~lvi~v~~~~~----~~~~--~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~-------- 108 (139)
..-.+.+....+|+.|-++-++=-.|.. +.++ ...-.+..|+++|++-++-.+--..++....+..
T Consensus 28 ~f~avCLgSs~l~l~gallQLlp~rr~~~~~~~~~sp~~~~rIl~~la~aDlLaclGVivRS~vWl~~p~~~~s~s~~~~ 107 (405)
T PF02101_consen 28 AFNAVCLGSSVLSLLGALLQLLPRRRSAGPRAPARSPSSSRRILFWLAVADLLACLGVIVRSSVWLGFPNFIDSISDVNG 107 (405)
T ss_pred hhhhhHHHHHHHHHHHHHHhhccccccccccccccCCcCCchhHHHHHHHHHHhhhhHHHHhhhhhcCCcccccccCCCC
Confidence 3444556666677777555553111110 1111 2346889999999999887666665555443221
Q ss_pred ---cccccchhchHHhHHhhHHHHH-HHHHhhc
Q psy16860 109 ---FGYFMCDVWNSFDVYFSQYRKH-FNMYEVD 137 (139)
Q Consensus 109 ---~g~~~C~~~~~~~~~~~~~Si~-~~m~~~~ 137 (139)
.+..+|........++..++-+ +.-||++
T Consensus 108 ~d~wp~afCv~ss~WIq~fYsAtfwWtfcYAVD 140 (405)
T PF02101_consen 108 TDIWPAAFCVGSSMWIQLFYSATFWWTFCYAVD 140 (405)
T ss_pred CccccHHHHHHHHHHHHHHHHHHHHHHHHHHHH
Confidence 1347897766655555555544 5556654
No 22
>KOG4564|consensus
Probab=75.16 E-value=43 Score=27.19 Aligned_cols=84 Identities=18% Similarity=0.195 Sum_probs=55.0
Q ss_pred HHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCc--------c----ccccccch
Q psy16860 48 MCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDE--------W----YFGYFMCD 115 (139)
Q Consensus 48 ~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~--------~----~~g~~~C~ 115 (139)
+.+=.-++++.=++.+.++...|++|-..|+.-.||.++=++-++..+-........++ + .-+...||
T Consensus 151 ytvGyslSl~sL~vAl~If~~FR~L~CtRn~IH~nLF~SfiLra~~~~i~~~~l~~~~~~~~~~~~~~~~~~~~~~~~Ck 230 (473)
T KOG4564|consen 151 YTVGYSLSLVSLLVALIIFLYFRSLHCTRNYIHMNLFASFILRAASVLIKDLVLVVNGEQDASSDTSLHCLISSNPVGCK 230 (473)
T ss_pred HHHHHHHHHHHHHHHHHHHHHhhhhcchHHHHHHHHHHHHHHHHHHHHHHHHHhhccccccccccccccccccccchhHH
Confidence 33333333443334445566778999999999999999999988876665554332222 1 13567899
Q ss_pred hchHHhHHhhHHHHHH
Q psy16860 116 VWNSFDVYFSQYRKHF 131 (139)
Q Consensus 116 ~~~~~~~~~~~~Si~~ 131 (139)
....+...+..+.-+.
T Consensus 231 ~~~~~~~Yf~~aNf~W 246 (473)
T KOG4564|consen 231 LLFVFFQYFVLANFFW 246 (473)
T ss_pred HHHHHHHHHHHHHHHH
Confidence 8888777776665543
No 23
>PF10327 7TM_GPCR_Sri: Serpentine type 7TM GPCR chemoreceptor Sri; InterPro: IPR019429 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents Sri, which is part of the Str superfamily of chemoreceptors.
Probab=67.61 E-value=21 Score=26.95 Aligned_cols=64 Identities=19% Similarity=0.173 Sum_probs=45.1
Q ss_pred HHHHHHHHHHHHHHHHHHHHhHHHhh-ccCCCCchHHHHH-H--HHHHHHHHHHHhhhHHHHHHHhcC
Q psy16860 42 IIKGMLMCFIIITAILGNLLVIISVI-KHRKLRIITNYFV-V--SLAFADLLVALCVMPFNAIVSLTD 105 (139)
Q Consensus 42 ~~~~~~~~~i~~~g~~gN~lvi~v~~-~~~~~~~~~~~~i-~--nLa~~Dl~~~l~~~p~~~~~~~~~ 105 (139)
.+....+-+++.++++-|.+.++.+. |.+|+.+-.++++ . ...++|+-.+.+..|........|
T Consensus 9 ~~li~~~~~ig~iS~~~n~~~iyLi~fks~k~~~fry~ll~~Qi~~~l~di~~t~L~qpipLfP~~ag 76 (303)
T PF10327_consen 9 QWLINYYHIIGVISFILNSLGIYLIIFKSPKLDNFRYYLLYFQISCTLTDIHLTFLMQPIPLFPIPAG 76 (303)
T ss_pred HHHHHHHHHHHHHHHHHHHHHheeEEEecCCccchhhHHHHHHHHHHHhhhhhhhhccchhhcceeEE
Confidence 35566788899999999999997666 4555555443333 2 234679999999888777665555
No 24
>PF09882 DUF2109: Predicted membrane protein (DUF2109); InterPro: IPR019214 This entry is found in various hypothetical archaeal proteins and has no known function.
Probab=66.67 E-value=23 Score=21.22 Aligned_cols=47 Identities=23% Similarity=0.353 Sum_probs=32.6
Q ss_pred HHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHH
Q psy16860 52 IITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFN 98 (139)
Q Consensus 52 ~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~ 98 (139)
.+.|+++=...+-++.-+.+.|+-.+.-.+|-+++-++....--|+.
T Consensus 4 ~i~g~Iai~~~iR~~~~~~r~~KL~yLnv~~F~iaalIaL~i~~P~g 50 (78)
T PF09882_consen 4 IIIGIIAILMAIRIFLTKSRARKLLYLNVINFAIAALIALYIKSPMG 50 (78)
T ss_pred HHHHHHHHHHHHHHHHhHhHHHhhhHHHHHHHHHHHHHHHHhCCcHH
Confidence 44566665566666666677788788888888888887766555544
No 25
>PF10316 7TM_GPCR_Srbc: Serpentine type 7TM GPCR chemoreceptor Srbc ; InterPro: IPR019420 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class b (Srb) from the Sra superfamily []. Srb receptors contain 6-8 hydrophobic, putative transmembrane, regions and can be distinguished from other 7TM GPCR receptors by their own characteristic TM signatures. Srbc is a solo family amongst the superfamilies of chemoreceptors.
Probab=66.57 E-value=32 Score=25.70 Aligned_cols=61 Identities=15% Similarity=0.118 Sum_probs=41.7
Q ss_pred HHHHHHHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHH
Q psy16860 42 IIKGMLMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVS 102 (139)
Q Consensus 42 ~~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~ 102 (139)
.+...+-++........|...++.+.++|+.|++--.++---...|.+.+....+......
T Consensus 6 ~iv~~i~i~~s~~~~~iN~~lL~~if~~Kk~kk~~l~LfY~Rf~~D~~~~~~~~~~~~~~~ 66 (273)
T PF10316_consen 6 IIVSIIGIIFSIITCLINFYLLYSIFYSKKKKKPDLSLFYFRFAIDVFYGFSVFIYLIYYI 66 (273)
T ss_pred HHHHHHHHHHHHHHHHHHHHHHHHHHhccccCCCCEEeeHHHHHHHHHHHHHHHHHHHHHH
Confidence 3445555666777788899888888866664555445555567889999988776544433
No 26
>PF01102 Glycophorin_A: Glycophorin A; InterPro: IPR001195 Proteins in this group are responsible for the molecular basis of the blood group antigens, surface markers on the outside of the red blood cell membrane. Most of these markers are proteins, but some are carbohydrates attached to lipids or proteins [Reid M.E., Lomas-Francis C. The Blood Group Antigen FactsBook Academic Press, London / San Diego, (1997)]. Glycophorin A (PAS-2) and glycophorin B (PAS-3) belong to the MNS blood group system and are associated with antigens that include M/N, S/s, U, He, Mi(a), M(c), Vw, Mur, M(g), Vr, M(e), Mt(a), St(a), Ri(a), Cl(a), Ny(a), Hut, Hil, M(v), Far, Mit, Dantu, Hop, Nob, En(a), ENKT, amongst others. Glycophorin A is the major sialoglycoprotein of the erythrocyte membrane []. Structurally, glycophorin A consists of an N-terminal extracellular domain, heavily glycosylated on serine and threonine residues, followed by a transmembrane region and a C-terminal cytoplasmic domain. Other glycophorins in this entry such as Glycophorin B and Glycophorin E represent minor sialoglycoproteins in the erythrocyte membrane.; GO: 0016021 integral to membrane; PDB: 2KPF_B 1AFO_B 2KPE_A.
Probab=57.16 E-value=15 Score=24.09 Aligned_cols=11 Identities=27% Similarity=0.428 Sum_probs=4.2
Q ss_pred HhHHHhhccCC
Q psy16860 61 LVIISVIKHRK 71 (139)
Q Consensus 61 lvi~v~~~~~~ 71 (139)
++.|+++|.||
T Consensus 83 li~y~irR~~K 93 (122)
T PF01102_consen 83 LISYCIRRLRK 93 (122)
T ss_dssp HHHHHHHHHS-
T ss_pred HHHHHHHHHhc
Confidence 33344444443
No 27
>PF10323 7TM_GPCR_Srv: Serpentine type 7TM GPCR chemoreceptor Srv; InterPro: IPR019426 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class v (Srv) from the Srg superfamily [, ]. Srg receptors contain seven hydrophobic, putative transmembrane, regions and can be distinguished from other 7TM GPCR receptors by their own characteristic TM signatures.
Probab=57.16 E-value=46 Score=24.71 Aligned_cols=41 Identities=17% Similarity=0.292 Sum_probs=28.2
Q ss_pred HHHHHHhHHHhhccCC----CCchHHHHHHHHHHHHHHHHHhhhH
Q psy16860 56 ILGNLLVIISVIKHRK----LRIITNYFVVSLAFADLLVALCVMP 96 (139)
Q Consensus 56 ~~gN~lvi~v~~~~~~----~~~~~~~~i~nLa~~Dl~~~l~~~p 96 (139)
++--..+++++.|.|+ .+++.+.++.+-+++|++..+....
T Consensus 9 lply~~il~~l~~~r~~~~~~~~~Fy~l~~~~~iaDi~~~~~~~~ 53 (283)
T PF10323_consen 9 LPLYIFILYCLLKLRKRSKTFKSTFYTLLIQHCIADILSMLFYFL 53 (283)
T ss_pred HHHHHHHHHHHHHcccCccccCCHHHHHHHHHHHHHHHHHHHHHH
Confidence 3444555555554443 5688999999999999998665433
No 28
>PF11446 DUF2897: Protein of unknown function (DUF2897); InterPro: IPR021550 This is a bacterial family of uncharacterised proteins.
Probab=53.70 E-value=20 Score=20.05 Aligned_cols=22 Identities=23% Similarity=0.445 Sum_probs=13.9
Q ss_pred HHHHHHHHHHHHHHHHhHHHhh
Q psy16860 46 MLMCFIIITAILGNLLVIISVI 67 (139)
Q Consensus 46 ~~~~~i~~~g~~gN~lvi~v~~ 67 (139)
.+.+++.+.-++||+.++---.
T Consensus 6 wlIIviVlgvIigNia~LK~sA 27 (55)
T PF11446_consen 6 WLIIVIVLGVIIGNIAALKYSA 27 (55)
T ss_pred hHHHHHHHHHHHhHHHHHHHhc
Confidence 3445555556789998885433
No 29
>TIGR01477 RIFIN variant surface antigen, rifin family. This model represents the rifin branch of the rifin/stevor family (pfam02009) of predicted variant surface antigens as found in Plasmodium falciparum. This model is based on a set of rifin sequences kindly provided by Matt Berriman from the Sanger Center. This is a global model and assesses a penalty for incomplete sequence. Additional fragmentary sequences may be found with the fragment model and a cutoff of 20 bits.
Probab=52.75 E-value=20 Score=27.84 Aligned_cols=31 Identities=16% Similarity=0.314 Sum_probs=21.1
Q ss_pred HHHHHHHHHHHHHHHHHHhHHHhhccCCCCc
Q psy16860 44 KGMLMCFIIITAILGNLLVIISVIKHRKLRI 74 (139)
Q Consensus 44 ~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~ 74 (139)
.+....++.++.++.-.++|+.++|+||.++
T Consensus 310 t~IiaSiIAIvvIVLIMvIIYLILRYRRKKK 340 (353)
T TIGR01477 310 TPIIASIIAILIIVLIMVIIYLILRYRRKKK 340 (353)
T ss_pred HHHHHHHHHHHHHHHHHHHHHHHHHhhhcch
Confidence 3455666666666666678888888877554
No 30
>PTZ00046 rifin; Provisional
Probab=49.96 E-value=24 Score=27.52 Aligned_cols=30 Identities=13% Similarity=0.361 Sum_probs=20.2
Q ss_pred HHHHHHHHHHHHHHHHHhHHHhhccCCCCc
Q psy16860 45 GMLMCFIIITAILGNLLVIISVIKHRKLRI 74 (139)
Q Consensus 45 ~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~ 74 (139)
.....++.++-++.-.++|+.++|+||.++
T Consensus 316 aIiaSiiAIvVIVLIMvIIYLILRYRRKKK 345 (358)
T PTZ00046 316 AIIASIVAIVVIVLIMVIIYLILRYRRKKK 345 (358)
T ss_pred HHHHHHHHHHHHHHHHHHHHHHHHhhhcch
Confidence 445566666666666677888888877554
No 31
>PF06024 DUF912: Nucleopolyhedrovirus protein of unknown function (DUF912); InterPro: IPR009261 This entry is represented by Autographa californica nuclear polyhedrosis virus (AcMNPV), Orf78; it is a family of uncharacterised viral proteins.
Probab=46.02 E-value=29 Score=21.74 Aligned_cols=11 Identities=9% Similarity=0.371 Sum_probs=5.7
Q ss_pred HHHhhccCCCC
Q psy16860 63 IISVIKHRKLR 73 (139)
Q Consensus 63 i~v~~~~~~~~ 73 (139)
-+++.|.|+.+
T Consensus 83 YFVILRer~~~ 93 (101)
T PF06024_consen 83 YFVILRERQKS 93 (101)
T ss_pred EEEEEeccccc
Confidence 35555665543
No 32
>PF02009 Rifin_STEVOR: Rifin/stevor family; InterPro: IPR002858 Malaria is still a major cause of mortality in many areas of the world. Plasmodium falciparum causes the most severe human form of the disease and is responsible for most fatalities. Severe cases of malaria can occur when the parasite invades and then proliferates within red blood cell erythrocytes. The parasite produces many variant antigenic proteins, encoded by multigene families, which are present on the surface of the infected erythrocyte and play important roles in virulence. A crucial survival mechanism for the malaria parasite is its ability to evade the immune response by switching these variant surface antigens. The high virulence of P. falciparum relative to other malarial parasites is in large part due to the fact that in this organism many of these surface antigens mediate the binding of infected erythrocytes to the vascular endothelium (cytoadherence) and non-infected erythrocytes (rosetting). This can lead to the accumulation of infected cells in the vasculature of a variety of organs, blocking the blood flow and reducing the oxygen supply. Clinical symptoms of severe infection can include fever, progressive anaemia, multi-organ dysfunction and coma. For more information see []. Several multicopy gene families have been described in Plasmodium falciparum, including the stevor family of subtelomeric open reading frames and the rif interspersed repetitive elements. Both families contain three predicted transmembrane segments. It has been proposed that stevor and rif are members of a larger superfamily that code for variant surface antigens [].
Probab=45.57 E-value=23 Score=26.90 Aligned_cols=27 Identities=19% Similarity=0.374 Sum_probs=15.3
Q ss_pred HHHHHHHHHHHHHHHhHHHhhccCCCC
Q psy16860 47 LMCFIIITAILGNLLVIISVIKHRKLR 73 (139)
Q Consensus 47 ~~~~i~~~g~~gN~lvi~v~~~~~~~~ 73 (139)
...++.++-++.=.++|+.++|+||.|
T Consensus 259 ~aSiiaIliIVLIMvIIYLILRYRRKK 285 (299)
T PF02009_consen 259 IASIIAILIIVLIMVIIYLILRYRRKK 285 (299)
T ss_pred HHHHHHHHHHHHHHHHHHHHHHHHHHh
Confidence 444444444444456777777776643
No 33
>PF02532 PsbI: Photosystem II reaction centre I protein (PSII 4.8 kDa protein); InterPro: IPR003686 Oxygenic photosynthesis uses two multi-subunit photosystems (I and II) located in the cell membranes of cyanobacteria and in the thylakoid membranes of chloroplasts in plants and algae. Photosystem II (PSII) has a P680 reaction centre containing chlorophyll 'a' that uses light energy to carry out the oxidation (splitting) of water molecules, and to produce ATP via a proton pump. Photosystem I (PSI) has a P700 reaction centre containing chlorophyll that takes the electron and associated hydrogen donated from PSII to reduce NADP+ to NADPH. Both ATP and NADPH are subsequently used in the light-independent reactions to convert carbon dioxide to glucose using the hydrogen atom extracted from water by PSII, releasing oxygen as a by-product. PSII is a multisubunit protein-pigment complex containing polypeptides both intrinsic and extrinsic to the photosynthetic membrane [, ]. Within the core of the complex, the chlorophyll and beta-carotene pigments are mainly bound to the antenna proteins CP43 (PsbC) and CP47 (PsbB), which pass the excitation energy on to the reaction centre proteins D1 (Qb, PsbA) and D2 (Qa, PsbD) that bind all the redox-active cofactors involved in the energy conversion process. The PSII oxygen-evolving complex (OEC) oxidises water to provide protons for use by PSI, and consists of OEE1 (PsbO), OEE2 (PsbP) and OEE3 (PsbQ). The remaining subunits in PSII are of low molecular weight (less than 10 kDa), and are involved in PSII assembly, stabilisation, dimerisation, and photo-protection []. This family represents the low molecular weight transmembrane protein PsbI, which is tightly associated with the D1/D2 heterodimer in PSII. The function of PsbI is unknown, but it may be involved in the assembly, dimerisation or stabilisation of PSII dimers [].; GO: 0015979 photosynthesis, 0009523 photosystem II, 0009539 photosystem II reaction center, 0016020 membrane; PDB: 3A0H_i 3ARC_I 3A0B_i 3BZ2_I 3PRQ_I 3KZI_I 3PRR_I 2AXT_i 4FBY_I 1S5L_i ....
Probab=42.40 E-value=42 Score=16.93 Aligned_cols=18 Identities=17% Similarity=0.229 Sum_probs=9.5
Q ss_pred HHHHHHHHHHHHHHHHHH
Q psy16860 42 IIKGMLMCFIIITAILGN 59 (139)
Q Consensus 42 ~~~~~~~~~i~~~g~~gN 59 (139)
+..+.+++.+++.|.+.|
T Consensus 9 y~vV~ffv~LFifGflsn 26 (36)
T PF02532_consen 9 YTVVIFFVSLFIFGFLSN 26 (36)
T ss_dssp HHHHHHHHHHHHHHHHTT
T ss_pred hhhHHHHHHHHhccccCC
Confidence 344455555666655544
No 34
>PF15330 SIT: SHP2-interacting transmembrane adaptor protein, SIT
Probab=42.17 E-value=53 Score=20.96 Aligned_cols=28 Identities=11% Similarity=0.216 Sum_probs=18.0
Q ss_pred HHHHHHHHHHHHHHHHHHhHHHhhccCC
Q psy16860 44 KGMLMCFIIITAILGNLLVIISVIKHRK 71 (139)
Q Consensus 44 ~~~~~~~i~~~g~~gN~lvi~v~~~~~~ 71 (139)
...++.++.++.++.|++......|++|
T Consensus 3 Ll~il~llLll~l~asl~~wr~~~rq~k 30 (107)
T PF15330_consen 3 LLGILALLLLLSLAASLLAWRMKQRQKK 30 (107)
T ss_pred HHHHHHHHHHHHHHHHHHHHHHHhhhcc
Confidence 4556777777778888777654444444
No 35
>KOG4193|consensus
Probab=40.61 E-value=2e+02 Score=24.34 Aligned_cols=63 Identities=17% Similarity=0.092 Sum_probs=34.4
Q ss_pred HHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCcccccc-c-cchhchHHhHHhhHHHHH
Q psy16860 60 LLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGY-F-MCDVWNSFDVYFSQYRKH 130 (139)
Q Consensus 60 ~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~-~-~C~~~~~~~~~~~~~Si~ 130 (139)
.+.+++.+..|++++..+....||+++-++. -+ . .+.+.|.-+. . .|+...++-+++..+...
T Consensus 338 ~lti~ty~~~~~l~~~~~~i~~~l~~~L~l~-~l------~-fL~~~~~~~~~~~~C~~~a~llhff~LaaF~ 402 (610)
T KOG4193|consen 338 LLTIATYLLFRKLQNDRTKIHINLCLCLFLA-EL------L-FLLGIDRTSTSVVLCIAAAILLHFFFLAAFF 402 (610)
T ss_pred HHHHHHHHHHHHHHhhcchhHHHHHHHHHHH-HH------H-HhcccccccCcccccHHHHHHHHHHHHHHHH
Confidence 3444444444445544478888998872222 11 1 2223344332 2 699988877777666543
No 36
>KOG2927|consensus
Probab=39.87 E-value=95 Score=24.28 Aligned_cols=9 Identities=22% Similarity=0.586 Sum_probs=6.0
Q ss_pred Ccccccccc
Q psy16860 105 DEWYFGYFM 113 (139)
Q Consensus 105 ~~~~~g~~~ 113 (139)
+-|.+++.+
T Consensus 255 g~W~FPNL~ 263 (372)
T KOG2927|consen 255 GFWLFPNLT 263 (372)
T ss_pred ceEeccchh
Confidence 358887654
No 37
>PF10326 7TM_GPCR_Str: Serpentine type 7TM GPCR chemoreceptor Str; InterPro: IPR019428 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class r (Str) from the Str superfamily [, ]. Almost a quarter (22.5%) of str and srj family genes and pseudogenes in C. elegans appear to have been newly formed by gene duplications since the species split [].
Probab=39.84 E-value=28 Score=25.92 Aligned_cols=48 Identities=8% Similarity=0.238 Sum_probs=34.0
Q ss_pred HHHHHHHHHHHHHhHHHhhccCCCCc-hHHHHHHHHHHHHHHHHHhhhH
Q psy16860 49 CFIIITAILGNLLVIISVIKHRKLRI-ITNYFVVSLAFADLLVALCVMP 96 (139)
Q Consensus 49 ~~i~~~g~~gN~lvi~v~~~~~~~~~-~~~~~i~nLa~~Dl~~~l~~~p 96 (139)
-+-++++++.|.+.++.+.++.+.+. ...+++.--|+.|+..+..-..
T Consensus 6 ~~~~~~s~~~N~~Li~Li~~~s~k~~G~Yk~Lm~~fs~~~i~fs~~~~~ 54 (307)
T PF10326_consen 6 YIGFVLSLFLNSLLIYLILTKSPKSLGSYKYLMIYFSIFEIIFSILDFL 54 (307)
T ss_pred HHHHHHHHHHHHHHHHHHHhccCCCCCCEEEEEehhHHHHHHHHHHHHH
Confidence 45678889999999988775544333 3456777788888888776543
No 38
>COG1230 CzcD Co/Zn/Cd efflux system component [Inorganic ion transport and metabolism]
Probab=37.73 E-value=1.8e+02 Score=22.12 Aligned_cols=70 Identities=17% Similarity=0.177 Sum_probs=43.2
Q ss_pred HHHHHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHH-HHHHHHHHHhhhHHHHHHHhcCccccccccchhchH
Q psy16860 47 LMCFIIITAILGNLLVIISVIKHRKLRIITNYFVVSL-AFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNS 119 (139)
Q Consensus 47 ~~~~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nL-a~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~ 119 (139)
-++.+.++|++-|.+..+.+.+.++ + ..|.=-..| +++|.+-++..+--.+.-.+.+ |..-|..+.+...
T Consensus 126 ~ml~va~~GL~vN~~~a~ll~~~~~-~-~lN~r~a~LHvl~D~Lgsv~vIia~i~i~~~~-w~~~Dpi~si~i~ 196 (296)
T COG1230 126 GMLVVAIIGLVVNLVSALLLHKGHE-E-NLNMRGAYLHVLGDALGSVGVIIAAIVIRFTG-WSWLDPILSIVIA 196 (296)
T ss_pred chHHHHHHHHHHHHHHHHHhhCCCc-c-cchHHHHHHHHHHHHHHHHHHHHHHHHHHHhC-CCccchHHHHHHH
Confidence 3678889999999999998877622 1 122222223 3579998777666555555555 4445665544333
No 39
>PRK03557 zinc transporter ZitB; Provisional
Probab=36.75 E-value=1.9e+02 Score=21.91 Aligned_cols=73 Identities=16% Similarity=0.145 Sum_probs=34.1
Q ss_pred HHHHHHHHHHHHhHHHhhccCCCCchHHHHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccchhchHHhHH
Q psy16860 50 FIIITAILGNLLVIISVIKHRKLRIITNYFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMCDVWNSFDVY 123 (139)
Q Consensus 50 ~i~~~g~~gN~lvi~v~~~~~~~~~~~~~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C~~~~~~~~~ 123 (139)
.+.+.+++.|.+..+..++.++.++..-.--.-=...|.+.++..+--.+...+.+ |..-+..+.+...+..+
T Consensus 126 ~v~~~~~~~~~~~~~~~~~~~~~~s~~l~a~~~h~~~D~l~s~~vlv~~~~~~~~g-~~~~Dpi~~ilis~~i~ 198 (312)
T PRK03557 126 AIAVAGLLANILSFWLLHHGSEEKNLNVRAAALHVLGDLLGSVGAIIAALIIIWTG-WTPADPILSILVSVLVL 198 (312)
T ss_pred HHHHHHHHHHHHHHHHHhcccccCCHHHHHHHHHHHHHHHHHHHHHHHHHHHHHcC-CcchhHHHHHHHHHHHH
Confidence 34556777787766655554443443211111112557777665333222222333 33345666554444333
No 40
>PF08114 PMP1_2: ATPase proteolipid family; InterPro: IPR012589 This family consists of small proteolipids associated with the plasma membrane H+ ATPase. Two proteolipids (PMP1 and PMP2) are associated with the ATPase and both genes are similarly expressed in the wild-type strain of yeast. No modification of the level of transcription of one PMP gene is detected in a strain deleted of the other. Though both proteolipids show similarity with other small proteolipids associated with other cation -transporting ATPases, their functions remain unclear [].
Probab=35.60 E-value=13 Score=19.39 Aligned_cols=24 Identities=8% Similarity=0.270 Sum_probs=14.0
Q ss_pred HHHHHHHHHHHHHHhHHHhhccCC
Q psy16860 48 MCFIIITAILGNLLVIISVIKHRK 71 (139)
Q Consensus 48 ~~~i~~~g~~gN~lvi~v~~~~~~ 71 (139)
..+++++|+.|-+++...+.|+..
T Consensus 11 IlVF~lVglv~i~iva~~iYRKw~ 34 (43)
T PF08114_consen 11 ILVFCLVGLVGIGIVALFIYRKWQ 34 (43)
T ss_pred eeehHHHHHHHHHHHHHHHHHHHH
Confidence 344556666666666666665543
No 41
>cd07912 Tweety_N N-terminal domain of the protein encoded by the Drosophila tweety gene and related proteins, a family of chloride ion channels. The protein product of the Drosophila tweety (tty) gene is thought to form a trans-membrane protein with five membrane-spanning regions and a cytoplasmic C-terminus. This N-terminal domain contains the putative transmembrane spanning regions. Tweety has been suggested as a candidate for a large conductance chloride channel, both in vertebrate and insect cells. Three human homologs have been identified and designated TTYH1-3. TTYH2 has been associated with the progression of cancer, and Drosophila melanogaster tweety has been assumed to play a role in development. TTYH2, and TTYH3 bind to and are ubiquinated by Nedd4-2, a HECT type E3 ubiquitin ligase, which most likely plays a role in controlling the cellular levels of tweety family proteins.
Probab=34.54 E-value=2.4e+02 Score=22.60 Aligned_cols=15 Identities=33% Similarity=0.279 Sum_probs=8.3
Q ss_pred HHHHHHHHHHHHHHH
Q psy16860 76 TNYFVVSLAFADLLV 90 (139)
Q Consensus 76 ~~~~i~nLa~~Dl~~ 90 (139)
...+...|.+.-++.
T Consensus 79 ~~c~~~sLiiltL~~ 93 (418)
T cd07912 79 ICCLKWSLVIATLLC 93 (418)
T ss_pred ccHHHHHHHHHHHHH
Confidence 445555666555554
No 42
>PF12606 RELT: Tumour necrosis factor receptor superfamily member 19; InterPro: IPR022248 The members of tumor necrosis factor receptor (TNFR) superfamily have been designated as the "guardians of the immune system" due to their roles in immune cell proliferation, differentiation, activation, and death (apoptosis). RELT (receptor expressed in lymphoid tissues) is a member of the TNFR superfamily. The messenger RNA of RELT is especially abundant in hematologic tissues such as spleen, lymph node, and peripheral blood leukocytes as well as in leukemias and lymphomas. RELT is able to activate the NF-kappaB pathway and selectively binds tumor necrosis factor receptor-associated factor 1 []. RELT like proteins 1 and 2 (RELL1 and RELL2) are two RELT homologues that bind to RELT. The expression of RELL1 at the mRNA level is ubiquitous, whereas expression of RELL2 mRNA is more restricted to particular tissues [].
Probab=33.18 E-value=51 Score=18.03 Aligned_cols=23 Identities=26% Similarity=0.531 Sum_probs=11.1
Q ss_pred HHHHHHHHHHHHHHHhHHHhhccCC
Q psy16860 47 LMCFIIITAILGNLLVIISVIKHRK 71 (139)
Q Consensus 47 ~~~~i~~~g~~gN~lvi~v~~~~~~ 71 (139)
+..++++.|++| +.++.+.|.+.
T Consensus 6 iV~i~iv~~lLg--~~I~~~~K~yg 28 (50)
T PF12606_consen 6 IVSIFIVMGLLG--LSICTTLKAYG 28 (50)
T ss_pred HHHHHHHHHHHH--HHHHHHhhccc
Confidence 344555556655 33334444443
No 43
>PF10873 DUF2668: Protein of unknown function (DUF2668); InterPro: IPR022640 Members in this family of proteins are annotated as cysteine and tyrosine-rich protein 1, however currently no function is known [].
Probab=30.08 E-value=59 Score=22.04 Aligned_cols=32 Identities=9% Similarity=0.322 Sum_probs=21.5
Q ss_pred HHHHHHHHHHHHHHHHHHHhHHHhhccCCCCc
Q psy16860 43 IKGMLMCFIIITAILGNLLVIISVIKHRKLRI 74 (139)
Q Consensus 43 ~~~~~~~~i~~~g~~gN~lvi~v~~~~~~~~~ 74 (139)
+..+++.++++.|+++-+.+....+.++..++
T Consensus 63 IaGIVfgiVfimgvva~i~icvCmc~kn~rgs 94 (155)
T PF10873_consen 63 IAGIVFGIVFIMGVVAGIAICVCMCMKNSRGS 94 (155)
T ss_pred eeeeehhhHHHHHHHHHHHHHHhhhhhcCCCc
Confidence 33446788888888887777766665555443
No 44
>KOG1482|consensus
Probab=29.92 E-value=1.2e+02 Score=23.92 Aligned_cols=78 Identities=13% Similarity=0.072 Sum_probs=51.6
Q ss_pred HHHHHHHHHhHHHhhccCCCCch-----------------HH-HHHHHHHHHHHHHHHhhhHHHHHHHhcCccccccccc
Q psy16860 53 ITAILGNLLVIISVIKHRKLRII-----------------TN-YFVVSLAFADLLVALCVMPFNAIVSLTDEWYFGYFMC 114 (139)
Q Consensus 53 ~~g~~gN~lvi~v~~~~~~~~~~-----------------~~-~~i~nLa~~Dl~~~l~~~p~~~~~~~~~~~~~g~~~C 114 (139)
-+|++-|++...++...-+-++. .| ---..=++.|++-+.-..-......+...|..-+..|
T Consensus 184 ~~gv~vNiim~~vL~~~~h~h~H~~~~s~g~~h~~~~~~n~nvraAyiHVlGDliQSvGV~iaa~Ii~f~P~~~i~DpIC 263 (379)
T KOG1482|consen 184 AVGVAVNIIMGFVLHQSGHGHSHGGSHSHGHSHDHGEELNLNVRAAFVHVLGDLIQSVGVLIAALIIYFKPEYKIADPIC 263 (379)
T ss_pred ehhhhhhhhhhhhhcccCCCCCCCCCCCcCcccccccccchHHHHHHHHHHHHHHHHHHHHhhheeEEecccceecCchh
Confidence 45667788777776644122221 11 1112245679999877666666667777899999999
Q ss_pred hhchHHhHHhhHHHHH
Q psy16860 115 DVWNSFDVYFSQYRKH 130 (139)
Q Consensus 115 ~~~~~~~~~~~~~Si~ 130 (139)
.+...........+++
T Consensus 264 T~~FSiivl~TT~~i~ 279 (379)
T KOG1482|consen 264 TFVFSIIVLGTTITIL 279 (379)
T ss_pred hhhHHHHHHHhHHHHH
Confidence 9988888888777775
No 45
>PF10319 7TM_GPCR_Srj: Serpentine type 7TM GPCR chemoreceptor Srj; InterPro: IPR019423 G-protein-coupled receptors, GPCRs, constitute a vast protein family that encompasses a wide range of functions (including various autocrine, paracrine and endocrine processes). They show considerable diversity at the sequence level, on the basis of which they can be separated into distinct groups. We use the term clan to describe the GPCRs, as they embrace a group of families for which there are indications of evolutionary relationship, but between which there is no statistically significant similarity in sequence []. The currently known clan members include the rhodopsin-like GPCRs, the secretin-like GPCRs, the cAMP receptors, the fungal mating pheromone receptors, and the metabotropic glutamate receptor family. There is a specialised database for GPCRs (http://www.gpcr.org/7tm/). The nematode Caenorhabditis elegans has only 14 types of chemosensory neuron, yet is able to sense and respond to several hundred different chemicals because each neuron detects several stimuli []. Chemoperception is one of the central senses of soil nematodes like C. elegans which are otherwise 'blind' and 'deaf' []. Chemoreception in C. elegans is mediated by members of the seven-transmembrane G-protein-coupled receptor class (7TM GPCRs). More than 1300 potential chemoreceptor genes have been identified in C. elegans, which are generally prefixed sr for serpentine receptor. The receptor superfamilies include Sra (Sra, Srb, Srab, Sre), Str (Srh, Str, Sri, Srd, Srj, Srm, Srn) and Srg (Srx, Srt, Srg, Sru, Srv, Srxa), as well as the families Srw, Srz, Srbc, Srsx and Srr [, , ]. Many of these proteins have homologues in Caenorhabditis briggsae. This entry represents serpentine receptor class j (Srj) from the Str superfamily [, ]. The Srj family is designated as the out-group based on its location in preliminary phylogenetic analyses of the entire superfamily [].
Probab=28.62 E-value=1.2e+02 Score=23.36 Aligned_cols=48 Identities=15% Similarity=0.231 Sum_probs=37.3
Q ss_pred HHHHHHHHHHHHHHHhHHHhhccCCCCc-hHHHHHHHHHHHHHHHHHhh
Q psy16860 47 LMCFIIITAILGNLLVIISVIKHRKLRI-ITNYFVVSLAFADLLVALCV 94 (139)
Q Consensus 47 ~~~~i~~~g~~gN~lvi~v~~~~~~~~~-~~~~~i~nLa~~Dl~~~l~~ 94 (139)
+.-+.++++.+-|-+.++.+..+|+.+- ...+++.--|+-|++.++.-
T Consensus 10 ~Pk~~~~lsf~~Np~fiyli~~~~~~~~G~Yr~LL~~Fa~fn~~~S~~~ 58 (310)
T PF10319_consen 10 IPKIFGILSFIVNPIFIYLIFTEKKSQFGNYRYLLLFFAIFNLIYSVVD 58 (310)
T ss_pred HHHHHHHHHHHHhhhhheeEEcccccccccHHHHHHHHHHHHHHHHHHH
Confidence 3456778888999999999888887655 34577888888899988753
No 46
>KOG4349|consensus
Probab=26.11 E-value=2e+02 Score=18.99 Aligned_cols=21 Identities=5% Similarity=0.159 Sum_probs=11.9
Q ss_pred ccccccchhchHHhHHhhHHH
Q psy16860 108 YFGYFMCDVWNSFDVYFSQYR 128 (139)
Q Consensus 108 ~~g~~~C~~~~~~~~~~~~~S 128 (139)
......|...+..+.+...+.
T Consensus 114 ~ms~~y~~i~G~~QT~~i~v~ 134 (143)
T KOG4349|consen 114 AVSTWYCAIMGIIQTLLIFVT 134 (143)
T ss_pred CccHHHHHHHHHHHHHHHHHH
Confidence 456667776665555544443
No 47
>PF05545 FixQ: Cbb3-type cytochrome oxidase component FixQ; InterPro: IPR008621 This family consists of several Cbb3-type cytochrome oxidase components (FixQ/CcoQ). FixQ is found in nitrogen fixing bacteria. Since nitrogen fixation is an energy-consuming process, effective symbioses depend on operation of a respiratory chain with a high affinity for O2, closely coupled to ATP production. This requirement is fulfilled by a special three-subunit terminal oxidase (cytochrome terminal oxidase cbb3), which was first identified in Bradyrhizobium japonicum as the product of the fixNOQP operon [].
Probab=26.07 E-value=1.1e+02 Score=16.13 Aligned_cols=9 Identities=22% Similarity=0.224 Sum_probs=4.8
Q ss_pred HhHHHhhcc
Q psy16860 61 LVIISVIKH 69 (139)
Q Consensus 61 lvi~v~~~~ 69 (139)
+++|+..++
T Consensus 25 i~~w~~~~~ 33 (49)
T PF05545_consen 25 IVIWAYRPR 33 (49)
T ss_pred HHHHHHccc
Confidence 556665443
No 48
>PF05961 Chordopox_A13L: Chordopoxvirus A13L protein; InterPro: IPR009236 This family consists of A13L proteins from the Chordopoxviruses. A13L or p8 is one of the three most abundant membrane proteins of the intracellular mature Vaccinia virus [].
Probab=25.87 E-value=1.1e+02 Score=17.79 Aligned_cols=13 Identities=15% Similarity=0.457 Sum_probs=8.3
Q ss_pred HhHHHhhccCCCC
Q psy16860 61 LVIISVIKHRKLR 73 (139)
Q Consensus 61 lvi~v~~~~~~~~ 73 (139)
++++.+..+++-+
T Consensus 17 lIlY~iYnr~~~~ 29 (68)
T PF05961_consen 17 LILYGIYNRKKTT 29 (68)
T ss_pred HHHHHHHhccccc
Confidence 6777777665543
No 49
>CHL00024 psbI photosystem II protein I
Probab=24.50 E-value=34 Score=17.25 Aligned_cols=17 Identities=18% Similarity=0.257 Sum_probs=8.8
Q ss_pred HHHHHHHHHHHHHHHHH
Q psy16860 43 IKGMLMCFIIITAILGN 59 (139)
Q Consensus 43 ~~~~~~~~i~~~g~~gN 59 (139)
..+.+++.+++.|.+.|
T Consensus 10 ~vV~ffvsLFifGFlsn 26 (36)
T CHL00024 10 TVVIFFVSLFIFGFLSN 26 (36)
T ss_pred hHHHHHHHHHHccccCC
Confidence 34455555565555443
No 50
>PF06679 DUF1180: Protein of unknown function (DUF1180); InterPro: IPR009565 This entry consists of several hypothetical eukaryotic proteins thought to be membrane proteins. Their function is unknown.
Probab=24.11 E-value=2.5e+02 Score=19.39 Aligned_cols=26 Identities=19% Similarity=0.235 Sum_probs=14.4
Q ss_pred HHHHHHHHHHHHHHHHHHHHhHHHhh
Q psy16860 42 IIKGMLMCFIIITAILGNLLVIISVI 67 (139)
Q Consensus 42 ~~~~~~~~~i~~~g~~gN~lvi~v~~ 67 (139)
.++-.+|+++.+.+++.=-+++-+++
T Consensus 93 ~l~R~~~Vl~g~s~l~i~yfvir~~R 118 (163)
T PF06679_consen 93 MLKRALYVLVGLSALAILYFVIRTFR 118 (163)
T ss_pred chhhhHHHHHHHHHHHHHHHHHHHHh
Confidence 34555566666666655445555443
No 51
>PRK13664 hypothetical protein; Provisional
Probab=23.98 E-value=1.1e+02 Score=17.29 Aligned_cols=17 Identities=12% Similarity=0.405 Sum_probs=11.6
Q ss_pred HHHHHHHHHHHHHHHhH
Q psy16860 47 LMCFIIITAILGNLLVI 63 (139)
Q Consensus 47 ~~~~i~~~g~~gN~lvi 63 (139)
+++++.++|++-|++=-
T Consensus 10 ilill~lvG~i~N~iK~ 26 (62)
T PRK13664 10 ILVLVFLVGVLLNVIKD 26 (62)
T ss_pred HHHHHHHHHHHHHHHHH
Confidence 34667788888886543
No 52
>PRK02655 psbI photosystem II reaction center I protein I; Provisional
Probab=22.12 E-value=38 Score=17.23 Aligned_cols=16 Identities=13% Similarity=0.202 Sum_probs=8.1
Q ss_pred HHHHHHHHHHHHHHHH
Q psy16860 43 IKGMLMCFIIITAILG 58 (139)
Q Consensus 43 ~~~~~~~~i~~~g~~g 58 (139)
..+.+++.+++.|.+.
T Consensus 10 ~vV~ffvsLFiFGfls 25 (38)
T PRK02655 10 IVVFFFVGLFVFGFLS 25 (38)
T ss_pred hhHHHHHHHHHcccCC
Confidence 3444555555555443
No 53
>PF10329 DUF2417: Region of unknown function (DUF2417); InterPro: IPR019431 This entry represents a family of fungal proteins with no known function. In some cases these proteins also contain an alpha/beta hydrolase fold (IPR000073 from INTERPRO).
Probab=21.36 E-value=2.2e+02 Score=20.89 Aligned_cols=15 Identities=27% Similarity=0.499 Sum_probs=7.3
Q ss_pred HHHHHHHHHHHHHHH
Q psy16860 76 TNYFVVSLAFADLLV 90 (139)
Q Consensus 76 ~~~~i~nLa~~Dl~~ 90 (139)
.++.+.-|-+.|++.
T Consensus 104 l~~vl~~Lllvdlil 118 (232)
T PF10329_consen 104 LNIVLAGLLLVDLIL 118 (232)
T ss_pred HHHHHHHHHHHHHHH
Confidence 344444444555554
No 54
>PHA03049 IMV membrane protein; Provisional
Probab=20.79 E-value=1.7e+02 Score=17.02 Aligned_cols=17 Identities=18% Similarity=0.567 Sum_probs=9.7
Q ss_pred HHHHHHHHHhHHHhhccCC
Q psy16860 53 ITAILGNLLVIISVIKHRK 71 (139)
Q Consensus 53 ~~g~~gN~lvi~v~~~~~~ 71 (139)
+++++| ++++.+..+++
T Consensus 11 CVaIi~--lIvYgiYnkk~ 27 (68)
T PHA03049 11 CVVIIG--LIVYGIYNKKT 27 (68)
T ss_pred HHHHHH--HHHHHHHhccc
Confidence 334444 66777776554
Done!