Query         047041
Match_columns 91
No_of_seqs    127 out of 716
Neff          6.2 
Searched_HMMs 46136
Date          Fri Mar 29 07:04:38 2013
Command       hhsearch -i /work/01045/syshi/csienesis_hhblits_a3m/047041.a3m -d /work/01045/syshi/HHdatabase/Cdd.hhm -o /work/01045/syshi/hhsearch_cdd/047041hhsearch_cdd -cpu 12 -v 0 

 No Hit                             Prob E-value P-value  Score    SS Cols Query HMM  Template HMM
  1 smart00768 X8 Possibly involve 100.0 1.8E-38 3.9E-43  201.6   8.6   85    2-87      1-85  (85)
  2 PF07983 X8:  X8 domain;  Inter 100.0 8.7E-31 1.9E-35  164.2   7.0   73    2-74      1-78  (78)
  3 PF09628 YvfG:  YvfG protein;    40.0      19 0.00041   21.7   1.3    9   51-59     27-35  (68)
  4 KOG3679 Predicted coiled-coil   25.9      42 0.00091   27.6   1.5   27   48-74    530-556 (802)
  5 PF04255 DUF433:  Protein of un  19.8 1.1E+02  0.0023   17.3   2.0   15    8-22     42-56  (56)
  6 cd04366 IlGF_insulin_bombyxin_  16.0 2.2E+02  0.0048   15.5   3.5   34   13-51      4-41  (42)
  7 PF01054 MMTV_SAg:  Mouse mamma  15.6      79  0.0017   24.0   1.0   22    2-23    261-282 (313)
  8 cd00101 IlGF_like Insulin/insu  15.3 2.2E+02  0.0047   15.1   3.5   37   13-51      4-40  (41)
  9 KOG0044 Ca2+ sensor (EF-Hand s  15.2      67  0.0015   23.1   0.5   13   44-56     97-109 (193)
 10 COG0066 LeuD 3-isopropylmalate  13.1 1.3E+02  0.0029   21.8   1.6   17   40-56     72-88  (191)

No 1  
>smart00768 X8 Possibly involved in carbohydrate binding. The X8 domain, which may be involved in carbohydrate binding, is found in an Olive pollen antigen as well as at the C terminus of family 17 glycosyl hydrolases. It contains 6 conserved cysteine residues which presumably form three disulfide bridges.
Probab=100.00  E-value=1.8e-38  Score=201.61  Aligned_cols=85  Identities=52%  Similarity=0.980  Sum_probs=82.8

Q ss_pred             eeeeeCCCCCHHHHHHHHHHHhccCcccccccccCcCcCCCCChhhhHhHHHHHHHHHhcCCCccccCCcceEEeecCCC
Q 047041            2 EWCIADGQTPDDELQMALDWACGKGGADCNKLQAKQPCYFPNTLRDHASYAFNDYYQKFKHQGATCYFHAAAMITDLDPS   81 (91)
Q Consensus         2 ~wCV~~~~~~~~~l~~~l~~aCg~~~~dC~~I~~~g~c~~p~t~~~~~Sya~N~YYq~~~~~~~aCdF~G~a~i~~~~ps   81 (91)
                      +|||+|+++++++|+++|+||||++ +||++|++||+||+|+++++|||||||+|||++++.+++|||+|+|+|++.+|+
T Consensus         1 ~wCv~~~~~~~~~l~~~~~yaCg~~-~dC~~I~~~g~c~~~~~~~~~aS~a~N~YYq~~~~~~~aC~F~G~a~~~~~~ps   79 (85)
T smart00768        1 LWCVAKPDADEAALQAALDYACGQG-ADCTAIQPGGSCYSPNTVKAHASYAFNSYYQKQGQSSGACDFGGTATITTTDPS   79 (85)
T ss_pred             CccccCCCCCHHHHHHHHHHHhcCC-CCccccCCCCcccCCCCHHHHHHHHHHHHHHHcCCCCCcCCCCCceEEEecCCC
Confidence            5999999999999999999999987 999999999999999999999999999999999999999999999999999999


Q ss_pred             CCceee
Q 047041           82 HRSCKF   87 (91)
Q Consensus        82 ~~~C~~   87 (91)
                      .++|+|
T Consensus        80 ~~~C~~   85 (85)
T smart00768       80 TGSCKF   85 (85)
T ss_pred             CCccCC
Confidence            999986


No 2  
>PF07983 X8:  X8 domain;  InterPro: IPR012946 The X8 domain [] contains 6 conserved cysteine residues that presumably form three disulphide bridges. The domain is found in an Olive pollen allergen [] as well as at the C terminus of family 17 glycosyl hydrolases []. This domain may be involved in carbohydrate binding.; PDB: 2JON_A 2W61_A 2W62_A 2W63_A.
Probab=99.97  E-value=8.7e-31  Score=164.23  Aligned_cols=73  Identities=44%  Similarity=0.837  Sum_probs=62.5

Q ss_pred             eeeeeCCCCCHHHHHHHHHHHhccCcccccccccCcC-----cCCCCChhhhHhHHHHHHHHHhcCCCccccCCcceE
Q 047041            2 EWCIADGQTPDDELQMALDWACGKGGADCNKLQAKQP-----CYFPNTLRDHASYAFNDYYQKFKHQGATCYFHAAAM   74 (91)
Q Consensus         2 ~wCV~~~~~~~~~l~~~l~~aCg~~~~dC~~I~~~g~-----c~~p~t~~~~~Sya~N~YYq~~~~~~~aCdF~G~a~   74 (91)
                      +|||+++++++++|+++|||||+++++||++|++||+     .|++|++++|||||||+|||++++.+.+|||+|+||
T Consensus         1 l~Cv~~~~~~~~~l~~~l~~aC~~~~~dC~~I~~~g~~G~YG~~S~C~~~~~lSya~N~YY~~~~~~~~~C~F~G~at   78 (78)
T PF07983_consen    1 LWCVAKPDADDKELQDLLDYACGQGGVDCSPIQPNGTTGVYGAYSMCSPRQHLSYAFNQYYQKQGRNSSACDFSGNAT   78 (78)
T ss_dssp             -EEEE-TTS-HHHHHHHHHHHTTT-SSSCCCC-EETTTTEE-TTTTS-CCHHHHHHHHHHHHHHTSSCCG-SS-STEE
T ss_pred             CcceeCCCCCHHHHHHHHHHHHcCCCCChhhhCCCCcccccccccCCCHHHHHHHHHHHHHHHcCCCCCcCCCCCCCC
Confidence            6999999999999999999999998899999999998     577889999999999999999999999999999996


No 3  
>PF09628 YvfG:  YvfG protein;  InterPro: IPR018590  Yvfg is a hypothetical protein of 71 residues expressed in some bacteria. The monomer consists of two parallel alpha helices, and the protein crystallises as a homo-dimer. ; PDB: 2GSV_A 2JS1_B.
Probab=40.04  E-value=19  Score=21.73  Aligned_cols=9  Identities=44%  Similarity=0.833  Sum_probs=7.7

Q ss_pred             HHHHHHHHH
Q 047041           51 YAFNDYYQK   59 (91)
Q Consensus        51 ya~N~YYq~   59 (91)
                      -|||+||..
T Consensus        27 ~AmNaYYr~   35 (68)
T PF09628_consen   27 HAMNAYYRS   35 (68)
T ss_dssp             HHHHHHHHH
T ss_pred             HHHHHHHHH
Confidence            589999986


No 4  
>KOG3679 consensus Predicted coiled-coil protein [General function prediction only]
Probab=25.86  E-value=42  Score=27.62  Aligned_cols=27  Identities=15%  Similarity=0.359  Sum_probs=22.4

Q ss_pred             hHhHHHHHHHHHhcCCCccccCCcceE
Q 047041           48 HASYAFNDYYQKFKHQGATCYFHAAAM   74 (91)
Q Consensus        48 ~~Sya~N~YYq~~~~~~~aCdF~G~a~   74 (91)
                      .+-.-+..||++|+-..-+|.|+|.--
T Consensus       530 elilrlqeyfekqgvkdfacsfsgsip  556 (802)
T KOG3679|consen  530 ELILRLQEYFEKQGVKDFACSFSGSIP  556 (802)
T ss_pred             HHHHHHHHHHHHcCcceeeeeccCCcc
Confidence            355567889999999999999999754


No 5  
>PF04255 DUF433:  Protein of unknown function (DUF433);  InterPro: IPR007367 This is a family of uncharacterised proteins.; PDB: 2GA1_B.
Probab=19.83  E-value=1.1e+02  Score=17.26  Aligned_cols=15  Identities=20%  Similarity=0.335  Sum_probs=9.8

Q ss_pred             CCCCHHHHHHHHHHH
Q 047041            8 GQTPDDELQMALDWA   22 (91)
Q Consensus         8 ~~~~~~~l~~~l~~a   22 (91)
                      |.++.+++.++|.|+
T Consensus        42 p~Lt~~~i~aAl~ya   56 (56)
T PF04255_consen   42 PSLTLEDIRAALAYA   56 (56)
T ss_dssp             TT--HHHHHHHHHHH
T ss_pred             CCCCHHHHHHHHHhC
Confidence            456778888888875


No 6  
>cd04366 IlGF_insulin_bombyxin_like IlGF_like family, insulin_bombyxin_like subgroup. Members include a number of peptides including insulin, insulin-like growth factors I and II, insect prothoracicotropic hormone (bombyxin), locust insulin-related peptide (LIRP), molluscan insulin-related peptides 1 to 5 (MIP), and C. elegans insulin-like peptides. With the exception of insulin-like growth factors, the active forms of these peptide hormones are composed of two chains (A and B) linked by two disulfide bonds; the arrangement of four cysteines is conserved in the "A" chain:  Cys1 is linked by a disulfide bond to Cys3, Cys2 and Cys4 are linked by interchain disulfide bonds to cysteines in the "B" chain. This alignment contains both chains, plus the intervening linker region, arranged as found in the propeptide form. Propeptides are cleaved to yield two separate chains linked covalently by the two disulfide bonds.
Probab=16.02  E-value=2.2e+02  Score=15.50  Aligned_cols=34  Identities=24%  Similarity=0.315  Sum_probs=23.3

Q ss_pred             HHHHHHHHHHhccCcccccccccCc----CcCCCCChhhhHhH
Q 047041           13 DELQMALDWACGKGGADCNKLQAKQ----PCYFPNTLRDHASY   51 (91)
Q Consensus        13 ~~l~~~l~~aCg~~~~dC~~I~~~g----~c~~p~t~~~~~Sy   51 (91)
                      +.|.+.|.++|+..+..     ..|    =|+.+||+.+=.+|
T Consensus         4 ~~L~~~L~~vC~~~~~~-----~~gIvdeCC~~~Ct~~~L~~Y   41 (42)
T cd04366           4 RHLADTLALLCSEYNSP-----RRGIVDECCRKSCTLDELLSY   41 (42)
T ss_pred             HHHHHHHHHHhCCCCCC-----CCChhhccCCCcCCHHHHHhh
Confidence            57889999999874431     123    26888888765554


No 7  
>PF01054 MMTV_SAg:  Mouse mammary tumour virus superantigen;  InterPro: IPR001213 The Mouse mammary tumor virus (MMTV) is a milk-transmitted type B retrovirus. The superantigen (SAg) is encoded in the long terminal repeat [].
Probab=15.58  E-value=79  Score=24.02  Aligned_cols=22  Identities=27%  Similarity=0.601  Sum_probs=20.1

Q ss_pred             eeeeeCCCCCHHHHHHHHHHHh
Q 047041            2 EWCIADGQTPDDELQMALDWAC   23 (91)
Q Consensus         2 ~wCV~~~~~~~~~l~~~l~~aC   23 (91)
                      -|||..+.-.++.++..-||+-
T Consensus       261 PWCv~Tq~EKddm~qQvhdyiy  282 (313)
T PF01054_consen  261 PWCVLTQKEKDDMKQQVHDYIY  282 (313)
T ss_pred             CcEeecHHHHHHHHHHHhhhee
Confidence            4999999999999999999987


No 8  
>cd00101 IlGF_like Insulin/insulin-like growth factor/relaxin family; insulin family of proteins. Members include a number of active peptides which are evolutionary related including insulin, relaxin, prorelaxin, insulin-like growth factors I and II, mammalian Leydig cell-specific insulin-like peptide (gene INSL3), early placenta insulin-like peptide (ELIP; gene INSL4), insect prothoracicotropic hormone (bombyxin), locust insulin-related peptide (LIRP), molluscan insulin-related peptides 1 to 5 (MIP), and C. elegans insulin-like peptides. Typically, the active forms of these peptide hormones are composed of two chains (A and B) linked by two disulfide bonds; the arrangement of four cysteines is conserved in the "A" chain: Cys1 is linked by a disulfide bond to Cys3, Cys2 and Cys4 are linked by interchain disulfide bonds to cysteines in the "B" chain. This alignment contains both chains, plus the intervening linker region, arranged as found in the propeptide form. Propeptides are cleaved 
Probab=15.29  E-value=2.2e+02  Score=15.13  Aligned_cols=37  Identities=38%  Similarity=0.565  Sum_probs=21.8

Q ss_pred             HHHHHHHHHHhccCcccccccccCcCcCCCCChhhhHhH
Q 047041           13 DELQMALDWACGKGGADCNKLQAKQPCYFPNTLRDHASY   51 (91)
Q Consensus        13 ~~l~~~l~~aCg~~~~dC~~I~~~g~c~~p~t~~~~~Sy   51 (91)
                      .+|.+++.++|+..+.. .+|.. -=|+.+||..+=++|
T Consensus         4 ~~Lv~~l~~vC~~~~~~-~giv~-eCC~~~Ct~~~L~~Y   40 (41)
T cd00101           4 RELVRALIFVCGDRGFY-RGIVD-ECCFRGCTLRELASY   40 (41)
T ss_pred             HHHHHHHHHhcCCCCCc-CCccc-ccCCCCCChHHHHhh
Confidence            56888999999874433 11111 115777777654443


No 9  
>KOG0044 consensus Ca2+ sensor (EF-Hand superfamily) [Signal transduction mechanisms]
Probab=15.18  E-value=67  Score=23.11  Aligned_cols=13  Identities=38%  Similarity=0.685  Sum_probs=12.0

Q ss_pred             ChhhhHhHHHHHH
Q 047041           44 TLRDHASYAFNDY   56 (91)
Q Consensus        44 t~~~~~Sya~N~Y   56 (91)
                      ++.++|.|+|..|
T Consensus        97 t~eekl~w~F~ly  109 (193)
T KOG0044|consen   97 TLEEKLKWAFRLY  109 (193)
T ss_pred             cHHHHhhhhheee
Confidence            8899999999988


No 10 
>COG0066 LeuD 3-isopropylmalate dehydratase small subunit [Amino acid transport and metabolism]
Probab=13.05  E-value=1.3e+02  Score=21.80  Aligned_cols=17  Identities=35%  Similarity=0.679  Sum_probs=13.6

Q ss_pred             CCCCChhhhHhHHHHHH
Q 047041           40 YFPNTLRDHASYAFNDY   56 (91)
Q Consensus        40 ~~p~t~~~~~Sya~N~Y   56 (91)
                      |+--+.++||.||+..|
T Consensus        72 FGcGSSREHApwALk~~   88 (191)
T COG0066          72 FGCGSSREHAPWALKDY   88 (191)
T ss_pred             CCCCccHHHHHHHHHHc
Confidence            33357899999999887


Done!