Query 047041
Match_columns 91
No_of_seqs 127 out of 716
Neff 6.2
Searched_HMMs 46136
Date Fri Mar 29 07:04:38 2013
Command hhsearch -i /work/01045/syshi/csienesis_hhblits_a3m/047041.a3m -d /work/01045/syshi/HHdatabase/Cdd.hhm -o /work/01045/syshi/hhsearch_cdd/047041hhsearch_cdd -cpu 12 -v 0
No Hit Prob E-value P-value Score SS Cols Query HMM Template HMM
1 smart00768 X8 Possibly involve 100.0 1.8E-38 3.9E-43 201.6 8.6 85 2-87 1-85 (85)
2 PF07983 X8: X8 domain; Inter 100.0 8.7E-31 1.9E-35 164.2 7.0 73 2-74 1-78 (78)
3 PF09628 YvfG: YvfG protein; 40.0 19 0.00041 21.7 1.3 9 51-59 27-35 (68)
4 KOG3679 Predicted coiled-coil 25.9 42 0.00091 27.6 1.5 27 48-74 530-556 (802)
5 PF04255 DUF433: Protein of un 19.8 1.1E+02 0.0023 17.3 2.0 15 8-22 42-56 (56)
6 cd04366 IlGF_insulin_bombyxin_ 16.0 2.2E+02 0.0048 15.5 3.5 34 13-51 4-41 (42)
7 PF01054 MMTV_SAg: Mouse mamma 15.6 79 0.0017 24.0 1.0 22 2-23 261-282 (313)
8 cd00101 IlGF_like Insulin/insu 15.3 2.2E+02 0.0047 15.1 3.5 37 13-51 4-40 (41)
9 KOG0044 Ca2+ sensor (EF-Hand s 15.2 67 0.0015 23.1 0.5 13 44-56 97-109 (193)
10 COG0066 LeuD 3-isopropylmalate 13.1 1.3E+02 0.0029 21.8 1.6 17 40-56 72-88 (191)
No 1
>smart00768 X8 Possibly involved in carbohydrate binding. The X8 domain, which may be involved in carbohydrate binding, is found in an Olive pollen antigen as well as at the C terminus of family 17 glycosyl hydrolases. It contains 6 conserved cysteine residues which presumably form three disulfide bridges.
Probab=100.00 E-value=1.8e-38 Score=201.61 Aligned_cols=85 Identities=52% Similarity=0.980 Sum_probs=82.8
Q ss_pred eeeeeCCCCCHHHHHHHHHHHhccCcccccccccCcCcCCCCChhhhHhHHHHHHHHHhcCCCccccCCcceEEeecCCC
Q 047041 2 EWCIADGQTPDDELQMALDWACGKGGADCNKLQAKQPCYFPNTLRDHASYAFNDYYQKFKHQGATCYFHAAAMITDLDPS 81 (91)
Q Consensus 2 ~wCV~~~~~~~~~l~~~l~~aCg~~~~dC~~I~~~g~c~~p~t~~~~~Sya~N~YYq~~~~~~~aCdF~G~a~i~~~~ps 81 (91)
+|||+|+++++++|+++|+||||++ +||++|++||+||+|+++++|||||||+|||++++.+++|||+|+|+|++.+|+
T Consensus 1 ~wCv~~~~~~~~~l~~~~~yaCg~~-~dC~~I~~~g~c~~~~~~~~~aS~a~N~YYq~~~~~~~aC~F~G~a~~~~~~ps 79 (85)
T smart00768 1 LWCVAKPDADEAALQAALDYACGQG-ADCTAIQPGGSCYSPNTVKAHASYAFNSYYQKQGQSSGACDFGGTATITTTDPS 79 (85)
T ss_pred CccccCCCCCHHHHHHHHHHHhcCC-CCccccCCCCcccCCCCHHHHHHHHHHHHHHHcCCCCCcCCCCCceEEEecCCC
Confidence 5999999999999999999999987 999999999999999999999999999999999999999999999999999999
Q ss_pred CCceee
Q 047041 82 HRSCKF 87 (91)
Q Consensus 82 ~~~C~~ 87 (91)
.++|+|
T Consensus 80 ~~~C~~ 85 (85)
T smart00768 80 TGSCKF 85 (85)
T ss_pred CCccCC
Confidence 999986
No 2
>PF07983 X8: X8 domain; InterPro: IPR012946 The X8 domain [] contains 6 conserved cysteine residues that presumably form three disulphide bridges. The domain is found in an Olive pollen allergen [] as well as at the C terminus of family 17 glycosyl hydrolases []. This domain may be involved in carbohydrate binding.; PDB: 2JON_A 2W61_A 2W62_A 2W63_A.
Probab=99.97 E-value=8.7e-31 Score=164.23 Aligned_cols=73 Identities=44% Similarity=0.837 Sum_probs=62.5
Q ss_pred eeeeeCCCCCHHHHHHHHHHHhccCcccccccccCcC-----cCCCCChhhhHhHHHHHHHHHhcCCCccccCCcceE
Q 047041 2 EWCIADGQTPDDELQMALDWACGKGGADCNKLQAKQP-----CYFPNTLRDHASYAFNDYYQKFKHQGATCYFHAAAM 74 (91)
Q Consensus 2 ~wCV~~~~~~~~~l~~~l~~aCg~~~~dC~~I~~~g~-----c~~p~t~~~~~Sya~N~YYq~~~~~~~aCdF~G~a~ 74 (91)
+|||+++++++++|+++|||||+++++||++|++||+ .|++|++++|||||||+|||++++.+.+|||+|+||
T Consensus 1 l~Cv~~~~~~~~~l~~~l~~aC~~~~~dC~~I~~~g~~G~YG~~S~C~~~~~lSya~N~YY~~~~~~~~~C~F~G~at 78 (78)
T PF07983_consen 1 LWCVAKPDADDKELQDLLDYACGQGGVDCSPIQPNGTTGVYGAYSMCSPRQHLSYAFNQYYQKQGRNSSACDFSGNAT 78 (78)
T ss_dssp -EEEE-TTS-HHHHHHHHHHHTTT-SSSCCCC-EETTTTEE-TTTTS-CCHHHHHHHHHHHHHHTSSCCG-SS-STEE
T ss_pred CcceeCCCCCHHHHHHHHHHHHcCCCCChhhhCCCCcccccccccCCCHHHHHHHHHHHHHHHcCCCCCcCCCCCCCC
Confidence 6999999999999999999999998899999999998 577889999999999999999999999999999996
No 3
>PF09628 YvfG: YvfG protein; InterPro: IPR018590 Yvfg is a hypothetical protein of 71 residues expressed in some bacteria. The monomer consists of two parallel alpha helices, and the protein crystallises as a homo-dimer. ; PDB: 2GSV_A 2JS1_B.
Probab=40.04 E-value=19 Score=21.73 Aligned_cols=9 Identities=44% Similarity=0.833 Sum_probs=7.7
Q ss_pred HHHHHHHHH
Q 047041 51 YAFNDYYQK 59 (91)
Q Consensus 51 ya~N~YYq~ 59 (91)
-|||+||..
T Consensus 27 ~AmNaYYr~ 35 (68)
T PF09628_consen 27 HAMNAYYRS 35 (68)
T ss_dssp HHHHHHHHH
T ss_pred HHHHHHHHH
Confidence 589999986
No 4
>KOG3679 consensus Predicted coiled-coil protein [General function prediction only]
Probab=25.86 E-value=42 Score=27.62 Aligned_cols=27 Identities=15% Similarity=0.359 Sum_probs=22.4
Q ss_pred hHhHHHHHHHHHhcCCCccccCCcceE
Q 047041 48 HASYAFNDYYQKFKHQGATCYFHAAAM 74 (91)
Q Consensus 48 ~~Sya~N~YYq~~~~~~~aCdF~G~a~ 74 (91)
.+-.-+..||++|+-..-+|.|+|.--
T Consensus 530 elilrlqeyfekqgvkdfacsfsgsip 556 (802)
T KOG3679|consen 530 ELILRLQEYFEKQGVKDFACSFSGSIP 556 (802)
T ss_pred HHHHHHHHHHHHcCcceeeeeccCCcc
Confidence 355567889999999999999999754
No 5
>PF04255 DUF433: Protein of unknown function (DUF433); InterPro: IPR007367 This is a family of uncharacterised proteins.; PDB: 2GA1_B.
Probab=19.83 E-value=1.1e+02 Score=17.26 Aligned_cols=15 Identities=20% Similarity=0.335 Sum_probs=9.8
Q ss_pred CCCCHHHHHHHHHHH
Q 047041 8 GQTPDDELQMALDWA 22 (91)
Q Consensus 8 ~~~~~~~l~~~l~~a 22 (91)
|.++.+++.++|.|+
T Consensus 42 p~Lt~~~i~aAl~ya 56 (56)
T PF04255_consen 42 PSLTLEDIRAALAYA 56 (56)
T ss_dssp TT--HHHHHHHHHHH
T ss_pred CCCCHHHHHHHHHhC
Confidence 456778888888875
No 6
>cd04366 IlGF_insulin_bombyxin_like IlGF_like family, insulin_bombyxin_like subgroup. Members include a number of peptides including insulin, insulin-like growth factors I and II, insect prothoracicotropic hormone (bombyxin), locust insulin-related peptide (LIRP), molluscan insulin-related peptides 1 to 5 (MIP), and C. elegans insulin-like peptides. With the exception of insulin-like growth factors, the active forms of these peptide hormones are composed of two chains (A and B) linked by two disulfide bonds; the arrangement of four cysteines is conserved in the "A" chain: Cys1 is linked by a disulfide bond to Cys3, Cys2 and Cys4 are linked by interchain disulfide bonds to cysteines in the "B" chain. This alignment contains both chains, plus the intervening linker region, arranged as found in the propeptide form. Propeptides are cleaved to yield two separate chains linked covalently by the two disulfide bonds.
Probab=16.02 E-value=2.2e+02 Score=15.50 Aligned_cols=34 Identities=24% Similarity=0.315 Sum_probs=23.3
Q ss_pred HHHHHHHHHHhccCcccccccccCc----CcCCCCChhhhHhH
Q 047041 13 DELQMALDWACGKGGADCNKLQAKQ----PCYFPNTLRDHASY 51 (91)
Q Consensus 13 ~~l~~~l~~aCg~~~~dC~~I~~~g----~c~~p~t~~~~~Sy 51 (91)
+.|.+.|.++|+..+.. ..| =|+.+||+.+=.+|
T Consensus 4 ~~L~~~L~~vC~~~~~~-----~~gIvdeCC~~~Ct~~~L~~Y 41 (42)
T cd04366 4 RHLADTLALLCSEYNSP-----RRGIVDECCRKSCTLDELLSY 41 (42)
T ss_pred HHHHHHHHHHhCCCCCC-----CCChhhccCCCcCCHHHHHhh
Confidence 57889999999874431 123 26888888765554
No 7
>PF01054 MMTV_SAg: Mouse mammary tumour virus superantigen; InterPro: IPR001213 The Mouse mammary tumor virus (MMTV) is a milk-transmitted type B retrovirus. The superantigen (SAg) is encoded in the long terminal repeat [].
Probab=15.58 E-value=79 Score=24.02 Aligned_cols=22 Identities=27% Similarity=0.601 Sum_probs=20.1
Q ss_pred eeeeeCCCCCHHHHHHHHHHHh
Q 047041 2 EWCIADGQTPDDELQMALDWAC 23 (91)
Q Consensus 2 ~wCV~~~~~~~~~l~~~l~~aC 23 (91)
-|||..+.-.++.++..-||+-
T Consensus 261 PWCv~Tq~EKddm~qQvhdyiy 282 (313)
T PF01054_consen 261 PWCVLTQKEKDDMKQQVHDYIY 282 (313)
T ss_pred CcEeecHHHHHHHHHHHhhhee
Confidence 4999999999999999999987
No 8
>cd00101 IlGF_like Insulin/insulin-like growth factor/relaxin family; insulin family of proteins. Members include a number of active peptides which are evolutionary related including insulin, relaxin, prorelaxin, insulin-like growth factors I and II, mammalian Leydig cell-specific insulin-like peptide (gene INSL3), early placenta insulin-like peptide (ELIP; gene INSL4), insect prothoracicotropic hormone (bombyxin), locust insulin-related peptide (LIRP), molluscan insulin-related peptides 1 to 5 (MIP), and C. elegans insulin-like peptides. Typically, the active forms of these peptide hormones are composed of two chains (A and B) linked by two disulfide bonds; the arrangement of four cysteines is conserved in the "A" chain: Cys1 is linked by a disulfide bond to Cys3, Cys2 and Cys4 are linked by interchain disulfide bonds to cysteines in the "B" chain. This alignment contains both chains, plus the intervening linker region, arranged as found in the propeptide form. Propeptides are cleaved
Probab=15.29 E-value=2.2e+02 Score=15.13 Aligned_cols=37 Identities=38% Similarity=0.565 Sum_probs=21.8
Q ss_pred HHHHHHHHHHhccCcccccccccCcCcCCCCChhhhHhH
Q 047041 13 DELQMALDWACGKGGADCNKLQAKQPCYFPNTLRDHASY 51 (91)
Q Consensus 13 ~~l~~~l~~aCg~~~~dC~~I~~~g~c~~p~t~~~~~Sy 51 (91)
.+|.+++.++|+..+.. .+|.. -=|+.+||..+=++|
T Consensus 4 ~~Lv~~l~~vC~~~~~~-~giv~-eCC~~~Ct~~~L~~Y 40 (41)
T cd00101 4 RELVRALIFVCGDRGFY-RGIVD-ECCFRGCTLRELASY 40 (41)
T ss_pred HHHHHHHHHhcCCCCCc-CCccc-ccCCCCCChHHHHhh
Confidence 56888999999874433 11111 115777777654443
No 9
>KOG0044 consensus Ca2+ sensor (EF-Hand superfamily) [Signal transduction mechanisms]
Probab=15.18 E-value=67 Score=23.11 Aligned_cols=13 Identities=38% Similarity=0.685 Sum_probs=12.0
Q ss_pred ChhhhHhHHHHHH
Q 047041 44 TLRDHASYAFNDY 56 (91)
Q Consensus 44 t~~~~~Sya~N~Y 56 (91)
++.++|.|+|..|
T Consensus 97 t~eekl~w~F~ly 109 (193)
T KOG0044|consen 97 TLEEKLKWAFRLY 109 (193)
T ss_pred cHHHHhhhhheee
Confidence 8899999999988
No 10
>COG0066 LeuD 3-isopropylmalate dehydratase small subunit [Amino acid transport and metabolism]
Probab=13.05 E-value=1.3e+02 Score=21.80 Aligned_cols=17 Identities=35% Similarity=0.679 Sum_probs=13.6
Q ss_pred CCCCChhhhHhHHHHHH
Q 047041 40 YFPNTLRDHASYAFNDY 56 (91)
Q Consensus 40 ~~p~t~~~~~Sya~N~Y 56 (91)
|+--+.++||.||+..|
T Consensus 72 FGcGSSREHApwALk~~ 88 (191)
T COG0066 72 FGCGSSREHAPWALKDY 88 (191)
T ss_pred CCCCccHHHHHHHHHHc
Confidence 33357899999999887
Done!