Query 018402
Match_columns 356
No_of_seqs 143 out of 1224
Neff 7.6
Searched_HMMs 46136
Date Fri Mar 29 08:43:13 2013
Command hhsearch -i /work/01045/syshi/csienesis_hhblits_a3m/018402.a3m -d /work/01045/syshi/HHdatabase/Cdd.hhm -o /work/01045/syshi/hhsearch_cdd/018402hhsearch_cdd -cpu 12 -v 0
No Hit Prob E-value P-value Score SS Cols Query HMM Template HMM
1 PF00954 S_locus_glycop: S-loc 99.8 4.9E-20 1.1E-24 150.3 10.7 104 139-261 1-108 (110)
2 PF01453 B_lectin: D-mannose b 99.8 1.1E-20 2.5E-25 154.9 3.1 75 36-110 2-114 (114)
3 PF08276 PAN_2: PAN-like domai 99.6 1.2E-15 2.5E-20 112.9 5.7 61 277-339 3-66 (66)
4 cd01098 PAN_AP_plant Plant PAN 99.6 4.7E-15 1E-19 114.3 8.2 80 271-354 2-84 (84)
5 cd00028 B_lectin Bulb-type man 99.6 1E-14 2.2E-19 120.0 8.8 61 5-69 13-74 (116)
6 smart00108 B_lectin Bulb-type 99.5 5.4E-14 1.2E-18 115.3 8.2 67 5-78 13-79 (114)
7 cd00129 PAN_APPLE PAN/APPLE-li 99.4 1.9E-13 4.2E-18 104.6 5.8 67 279-353 9-80 (80)
8 smart00473 PAN_AP divergent su 98.6 1.9E-07 4E-12 70.2 7.4 72 279-353 4-78 (78)
9 cd01100 APPLE_Factor_XI_like S 97.8 2.6E-05 5.7E-10 58.5 4.5 51 283-336 8-58 (73)
10 PF00024 PAN_1: PAN domain Thi 94.1 0.061 1.3E-06 39.9 3.6 52 280-334 3-55 (79)
11 smart00223 APPLE APPLE domain. 92.8 0.16 3.4E-06 38.7 3.9 49 284-335 6-57 (79)
12 PF14295 PAN_4: PAN domain; PD 91.9 0.14 3.1E-06 34.7 2.5 32 302-333 15-51 (51)
13 PF01453 B_lectin: D-mannose b 91.5 0.39 8.4E-06 39.1 5.1 35 37-71 39-74 (114)
14 smart00108 B_lectin Bulb-type 91.0 0.36 7.8E-06 39.0 4.4 41 37-81 74-114 (114)
15 smart00605 CW CW domain. 90.4 1.7 3.7E-05 33.9 7.6 52 302-354 21-75 (94)
16 cd00028 B_lectin Bulb-type man 90.3 0.47 1E-05 38.5 4.5 42 37-82 75-116 (116)
17 PF08277 PAN_3: PAN-like domai 90.2 0.9 1.9E-05 33.2 5.5 51 301-353 18-71 (71)
18 cd01099 PAN_AP_HGF Subfamily o 85.4 1.3 2.7E-05 33.6 3.8 33 302-334 24-58 (80)
19 cd00053 EGF Epidermal growth f 66.5 6.3 0.00014 23.7 2.5 27 236-262 5-32 (36)
20 cd05845 Ig2_L1-CAM_like Second 62.2 12 0.00027 29.4 3.9 34 34-67 32-65 (95)
21 PF01683 EB: EB module; Inter 57.2 12 0.00027 25.5 2.8 31 228-261 17-47 (52)
22 PF07354 Sp38: Zona-pellucida- 51.4 17 0.00037 34.0 3.5 30 37-66 14-43 (271)
23 PF07645 EGF_CA: Calcium-bindi 50.3 4.8 0.0001 26.4 -0.2 30 231-260 3-34 (42)
24 smart00179 EGF_CA Calcium-bind 48.7 15 0.00033 22.7 2.0 30 231-260 3-33 (39)
25 cd05852 Ig5_Contactin-1 Fifth 42.9 28 0.0006 25.5 2.9 33 34-67 14-46 (73)
26 cd00054 EGF_CA Calcium-binding 42.4 22 0.00048 21.5 2.0 30 231-260 3-33 (38)
27 PF10681 Rot1: Chaperone for p 39.4 1.3E+02 0.0027 27.3 6.9 102 11-119 16-132 (212)
28 PF07974 EGF_2: EGF-like domai 37.4 26 0.00057 21.6 1.7 23 237-260 6-28 (32)
29 PF09064 Tme5_EGF_like: Thromb 32.6 26 0.00057 22.1 1.1 16 245-260 12-27 (34)
30 smart00564 PQQ beta-propeller 27.7 62 0.0013 19.1 2.3 19 57-77 13-31 (33)
31 KOG4649 PQQ (pyrrolo-quinoline 27.5 2.1E+02 0.0044 27.2 6.5 41 34-77 167-213 (354)
32 PF01436 NHL: NHL repeat; Int 24.1 85 0.0018 18.4 2.3 18 52-69 5-22 (28)
33 cd05764 Ig_2 Subgroup of the i 23.3 1.3E+02 0.0027 21.4 3.7 32 34-65 14-45 (74)
No 1
>PF00954 S_locus_glycop: S-locus glycoprotein family; InterPro: IPR000858 In Brassicaceae, self-incompatible plants have a self/non-self recognition system, which involves the inability of flowering plants to achieve self-fertilisation. This is sporophytically controlled by multiple alleles at a single locus (S). There are a total of 50 different S alleles in Brassica oleracea. S-locus glycoproteins, as well as S-receptor kinases, are in linkage with the S-alleles []. Most of the proteins within this family contain apple-like domain (IPR003609 from INTERPRO), which is predicted to possess protein- and/or carbohydrate-binding functions.; GO: 0048544 recognition of pollen
Probab=99.82 E-value=4.9e-20 Score=150.31 Aligned_cols=104 Identities=16% Similarity=0.267 Sum_probs=82.3
Q ss_pred EeCCCCCCCcceEEEecCCCCc--ceEEEEEeCCCC--ceEEecCCCCceEEEEEEccCCcEEEEEeecCCCCCCceEEE
Q 018402 139 YTFPISYNGLKSLTLKSSPGKM--HELTLVDSSGDN--GFIFDRPKYDSTISFLRLGIDGNLRVFTYSQEVDWLPEEERF 214 (356)
Q Consensus 139 W~~~~~~~~~~~~~~~~~~~~~--~~~~~~~~~~~~--~~~~~~~~~~~~~~~l~Ld~dG~l~~~~~~~~~~~~~W~~~~ 214 (356)
||+| +|++..|++.|++. ..+.+.|+.++. ++.+ ...+.+.++|++||++|+|+++.|.+. .+.|.+.|
T Consensus 1 wrsG----~WnG~~f~g~p~~~~~~~~~~~fv~~~~e~~~t~-~~~~~s~~~r~~ld~~G~l~~~~w~~~--~~~W~~~~ 73 (110)
T PF00954_consen 1 WRSG----PWNGQRFSGIPEMSSNSLYNYSFVSNNEEVYYTY-SLSNSSVLSRLVLDSDGQLQRYIWNES--TQSWSVFW 73 (110)
T ss_pred CCcc----ccCCeEECCcccccccceeEEEEEECCCeEEEEE-ecCCCceEEEEEEeeeeEEEEEEEecC--CCcEEEEE
Confidence 6664 89999999999872 334444444333 3344 344567789999999999999999874 68899999
Q ss_pred EecccccCCCCCCCCCCCCCCCCCCCCCcccCCCCCccCCCCCCCCC
Q 018402 215 TLFGKDFRGSNARNWDSECQMPERCGKLGVCEDNQCVACPTEKGLIG 261 (356)
Q Consensus 215 ~~~~~~~~~~~~~~p~~~C~~~~~CG~~g~C~~~~~~~C~C~~g~~~ 261 (356)
. +|.|+||+|+.||+||+|+.++.+.|.|++||..
T Consensus 74 ~------------~p~d~Cd~y~~CG~~g~C~~~~~~~C~Cl~GF~P 108 (110)
T PF00954_consen 74 S------------APKDQCDVYGFCGPNGICNSNNSPKCSCLPGFEP 108 (110)
T ss_pred E------------ecccCCCCccccCCccEeCCCCCCceECCCCcCC
Confidence 6 4789999999999999999888889999999963
No 2
>PF01453 B_lectin: D-mannose binding lectin; InterPro: IPR001480 A bulb lectin super-family (Amaryllidaceae, Orchidaceae and Aliaceae) contains a ~115-residue-long domain whose overall three dimensional fold is very similar to that of [, ]: Dictyostelium discoideum comitin, an actin binding protein Curculigo latifolia curculin, a sweet tasting and taste-modifying protein This domain generally binds mannose, but in at least one protein, curculin, it is apparently devoid of mannose-binding activity. Each bulb-type lectin domain consists of three sequential beta-sheet subdomains (I, II, III) that are inter-related by pseudo three-fold symmetry. The three subdomains are flat four-stranded, antiparrallel beta-sheets. Together they form a 12-stranded beta-barrel in which the barrel axis coincides with the pseudo 3-fold axis.; GO: 0005529 sugar binding; PDB: 3M7H_A 3M7J_B 3MEZ_D 1DLP_A 1BWU_D 1KJ1_A 1B2P_A 1XD6_A 2DPF_C 2D04_B ....
Probab=99.80 E-value=1.1e-20 Score=154.93 Aligned_cols=75 Identities=51% Similarity=0.782 Sum_probs=48.8
Q ss_pred eeEEEEecCCCCcC---CCceEEEccCccEEEEcCCCCc-------------------------------ceEEeeeccC
Q 018402 36 FRWVWEANRGKPVR---ENAVFSLGADGNLVLAEADGTV-------------------------------GNFIWQSFDY 81 (356)
Q Consensus 36 ~t~VWvANr~~Pv~---~~~~l~l~~~GnLvL~d~~~~~-------------------------------~~~lWQSFD~ 81 (356)
+||||+|||+.|+. ...+|.|..||+|||.+..+++ +++|||||||
T Consensus 2 ~tvvW~an~~~p~~~~s~~~~L~l~~dGnLvl~~~~~~~iWss~~t~~~~~~~~~~~L~~~GNlvl~d~~~~~lW~Sf~~ 81 (114)
T PF01453_consen 2 RTVVWVANRNSPLTSSSGNYTLILQSDGNLVLYDSNGSVIWSSNNTSGRGNSGCYLVLQDDGNLVLYDSSGNVLWQSFDY 81 (114)
T ss_dssp --------TTEEEEECETTEEEEEETTSEEEEEETTTEEEEE--S-TTSS-SSEEEEEETTSEEEEEETTSEEEEESTTS
T ss_pred cccccccccccccccccccccceECCCCeEEEEcCCCCEEEEecccCCccccCeEEEEeCCCCEEEEeecceEEEeecCC
Confidence 48899999999984 2478999999999998876432 5689999999
Q ss_pred CCCccccCceeecCCe----EEEEecCCCCCCC
Q 018402 82 PTDTLLVGQSLRVGRV----TKLVSRLSVKENV 110 (356)
Q Consensus 82 PTDTlLPgqkL~~g~~----~~L~Sw~s~~dps 110 (356)
||||+||+|+|+.+.. ..|+||++.+|||
T Consensus 82 ptdt~L~~q~l~~~~~~~~~~~~~sw~s~~dps 114 (114)
T PF01453_consen 82 PTDTLLPGQKLGDGNVTGKNDSLTSWSSNTDPS 114 (114)
T ss_dssp SS-EEEEEET--TSEEEEESTSSEEEESS----
T ss_pred CccEEEeccCcccCCCccccceEEeECCCCCCC
Confidence 9999999999987532 2489999999996
No 3
>PF08276 PAN_2: PAN-like domain; InterPro: IPR013227 PAN domains have significant functional versatility fulfilling diverse biological functions by mediating protein-protein or protein-carbohydrate interactions []. These domains contain a hair-pin loop like structure, similar to knottins, but the pattern of disulphide bonds differs
Probab=99.60 E-value=1.2e-15 Score=112.93 Aligned_cols=61 Identities=26% Similarity=0.529 Sum_probs=50.2
Q ss_pred CCceEEEeccccccccccccccCCCCCCHHHHHHHhhccCCeEEEEecC--CCceeEEcc-cccce
Q 018402 277 KHFHYYKLESVRHYMCTYNFYDGIGDITIEDCGKRCSSNCRCVGYFYDT--SVSRCWIAF-DLKTL 339 (356)
Q Consensus 277 ~~~~f~~l~~~~~~~~~~~~~~~~~~~s~~~C~~~Cl~nCsC~Ay~y~~--~~~~C~~w~-~l~~~ 339 (356)
.+++|++|++|++|+..... ...++++++|+++||+||||+||+|.+ ++++|++|. +|+|+
T Consensus 3 ~~d~F~~l~~~~~p~~~~~~--~~~~~s~~~C~~~Cl~nCsC~Ayay~~~~~~~~C~lW~~~L~d~ 66 (66)
T PF08276_consen 3 SGDGFLKLPNMKLPDFDNAI--VDSSVSLEECEKACLSNCSCTAYAYSNLSGGGGCLLWYGDLVDL 66 (66)
T ss_pred CCCEEEEECCeeCCCCccee--eecCCCHHHHHhhcCCCCCEeeEEeeccCCCCEEEEEcCEeecC
Confidence 35799999999998765442 224589999999999999999999975 678999995 78774
No 4
>cd01098 PAN_AP_plant Plant PAN/APPLE-like domain; present in plant S-receptor protein kinases and secreted glycoproteins. PAN/APPLE domains fulfill diverse biological functions by mediating protein-protein or protein-carbohydrate interactions. S-receptor protein kinases and S-locus glycoproteins are involved in sporophytic self-incompatibility response in Brassica, one of probably many molecular mechanisms, by which hermaphrodite flowering plants avoid self-fertilization.
Probab=99.59 E-value=4.7e-15 Score=114.27 Aligned_cols=80 Identities=23% Similarity=0.415 Sum_probs=63.6
Q ss_pred ccCCCCC--CceEEEeccccccccccccccCCCCCCHHHHHHHhhccCCeEEEEecCCCceeEEcc-cccceeecCCCCc
Q 018402 271 ANFCGTK--HFHYYKLESVRHYMCTYNFYDGIGDITIEDCGKRCSSNCRCVGYFYDTSVSRCWIAF-DLKTLTKEPDSPI 347 (356)
Q Consensus 271 ~~~C~~~--~~~f~~l~~~~~~~~~~~~~~~~~~~s~~~C~~~Cl~nCsC~Ay~y~~~~~~C~~w~-~l~~~~~~~~~~~ 347 (356)
+++|+.. .+.|++++++++++..+. . ...++++|+++||+||+|+||+|.+++++|++|. .+.+.+.....+.
T Consensus 2 ~~~C~~~~~~~~f~~~~~~~~~~~~~~---~-~~~s~~~C~~~Cl~nCsC~a~~~~~~~~~C~~~~~~~~~~~~~~~~~~ 77 (84)
T cd01098 2 PLNCGGDGSTDGFLKLPDVKLPDNASA---I-TAISLEECREACLSNCSCTAYAYNNGSGGCLLWNGLLNNLRSLSSGGG 77 (84)
T ss_pred CcccCCCCCCCEEEEeCCeeCCCchhh---h-ccCCHHHHHHHHhcCCCcceeeecCCCCeEEEEeceecceEeecCCCc
Confidence 4578654 368999999998865443 1 4579999999999999999999987778999995 6777766444468
Q ss_pred eEEEEec
Q 018402 348 VGFIKVS 354 (356)
Q Consensus 348 ~~yikv~ 354 (356)
++||||+
T Consensus 78 ~~yiKv~ 84 (84)
T cd01098 78 TLYLRLA 84 (84)
T ss_pred EEEEEeC
Confidence 9999996
No 5
>cd00028 B_lectin Bulb-type mannose-specific lectin. The domain contains a three-fold internal repeat (beta-prism architecture). The consensus sequence motif QXDXNXVXY is involved in alpha-D-mannose recognition. Lectins are carbohydrate-binding proteins which specifically recognize diverse carbohydrates and mediate a wide variety of biological processes, such as cell-cell and host-pathogen interactions, serum glycoprotein turnover, and innate immune responses.
Probab=99.57 E-value=1e-14 Score=120.00 Aligned_cols=61 Identities=28% Similarity=0.442 Sum_probs=51.9
Q ss_pred CCCcEEEEeecCCCCC-eEEEEEEceeCCCCceeEEEEecCCCCcCCCceEEEccCccEEEEcCCC
Q 018402 5 YNDPFQLGFYNTTPNA-FTLALRLGIKKQEPVFRWVWEANRGKPVRENAVFSLGADGNLVLAEADG 69 (356)
Q Consensus 5 ~~g~F~lGFf~~~~~~-~~l~Iw~~~~~~~~~~t~VWvANr~~Pv~~~~~l~l~~~GnLvL~d~~~ 69 (356)
.++.|++|||.+.... .+++|||.... .++||+||++.|....++|.|..+|+|||.+.++
T Consensus 13 ~~~~f~~G~~~~~~q~~dgnlv~~~~~~----~~~vW~snt~~~~~~~~~l~l~~dGnLvl~~~~g 74 (116)
T cd00028 13 SGSLFELGFFKLIMQSRDYNLILYKGSS----RTVVWVANRDNPSGSSCTLTLQSDGNLVIYDGSG 74 (116)
T ss_pred CCCcEEEecccCCCCCCeEEEEEEeCCC----CeEEEECCCCCCCCCCEEEEEecCCCeEEEcCCC
Confidence 4688999999987665 89999998642 4789999999996667899999999999998765
No 6
>smart00108 B_lectin Bulb-type mannose-specific lectin.
Probab=99.50 E-value=5.4e-14 Score=115.30 Aligned_cols=67 Identities=34% Similarity=0.579 Sum_probs=57.2
Q ss_pred CCCcEEEEeecCCCCCeEEEEEEceeCCCCceeEEEEecCCCCcCCCceEEEccCccEEEEcCCCCcceEEeee
Q 018402 5 YNDPFQLGFYNTTPNAFTLALRLGIKKQEPVFRWVWEANRGKPVRENAVFSLGADGNLVLAEADGTVGNFIWQS 78 (356)
Q Consensus 5 ~~g~F~lGFf~~~~~~~~l~Iw~~~~~~~~~~t~VWvANr~~Pv~~~~~l~l~~~GnLvL~d~~~~~~~~lWQS 78 (356)
.++.|++|||.+....++++|||+..+ .++||+|||+.|+..++.|.|.++|+|||.+.++ .++|+|
T Consensus 13 ~~~~f~~G~~~~~~q~dgnlV~~~~~~----~~~vW~snt~~~~~~~~~l~l~~dGnLvl~~~~g---~~vW~S 79 (114)
T smart00108 13 GNSLFELGFFTLIMQNDYNLILYKSSS----RTVVWVANRDNPVSDSCTLTLQSDGNLVLYDGDG---RVVWSS 79 (114)
T ss_pred CCCcEeeeccccCCCCCEEEEEEECCC----CcEEEECCCCCCCCCCEEEEEeCCCCEEEEeCCC---CEEEEe
Confidence 478899999998767789999998642 4789999999998777899999999999998874 568887
No 7
>cd00129 PAN_APPLE PAN/APPLE-like domain; present in N-terminal (N) domains of plasminogen/ hepatocyte growth factor proteins, plasma prekallikrein/coagulation factor XI and microneme antigen proteins, plant receptor-like protein kinases, and various nematode and leech anti-platelet proteins. Common structural features include two disulfide bonds that link the alpha-helix to the central region of the protein. PAN domains have significant functional versatility, fulfilling diverse biological functions by mediating protein-protein or protein-carbohydrate interactions.
Probab=99.43 E-value=1.9e-13 Score=104.59 Aligned_cols=67 Identities=12% Similarity=0.196 Sum_probs=56.3
Q ss_pred ceEEEeccccccccccccccCCCCCCHHHHHHHhhc---cCCeEEEEecCCCceeEEcc-cc-cceeecCCCCceEEEEe
Q 018402 279 FHYYKLESVRHYMCTYNFYDGIGDITIEDCGKRCSS---NCRCVGYFYDTSVSRCWIAF-DL-KTLTKEPDSPIVGFIKV 353 (356)
Q Consensus 279 ~~f~~l~~~~~~~~~~~~~~~~~~~s~~~C~~~Cl~---nCsC~Ay~y~~~~~~C~~w~-~l-~~~~~~~~~~~~~yikv 353 (356)
..|+++.+|+.|+.. ..+++||+++|++ ||||+||+|.+.+++|++|. +| .+++...+.+.++|||.
T Consensus 9 g~fl~~~~~klpd~~--------~~s~~eC~~~Cl~~~~nCsC~Aya~~~~~~gC~~W~~~l~~d~~~~~~~g~~Ly~r~ 80 (80)
T cd00129 9 GTTLIKIALKIKTTK--------ANTADECANRCEKNGLPFSCKAFVFAKARKQCLWFPFNSMSGVRKEFSHGFDLYENK 80 (80)
T ss_pred CeEEEeecccCCccc--------ccCHHHHHHHHhcCCCCCCceeeeccCCCCCeEEecCcchhhHHhccCCCceeEeEC
Confidence 478999888876432 2589999999999 99999999976567899996 78 89988777789999984
No 8
>smart00473 PAN_AP divergent subfamily of APPLE domains. Apple-like domains present in Plasminogen, C. elegans hypothetical ORFs and the extracellular portion of plant receptor-like protein kinases. Predicted to possess protein- and/or carbohydrate-binding functions.
Probab=98.59 E-value=1.9e-07 Score=70.18 Aligned_cols=72 Identities=18% Similarity=0.377 Sum_probs=53.9
Q ss_pred ceEEEeccccccccccccccCCCCCCHHHHHHHhhc-cCCeEEEEecCCCceeEEcc--cccceeecCCCCceEEEEe
Q 018402 279 FHYYKLESVRHYMCTYNFYDGIGDITIEDCGKRCSS-NCRCVGYFYDTSVSRCWIAF--DLKTLTKEPDSPIVGFIKV 353 (356)
Q Consensus 279 ~~f~~l~~~~~~~~~~~~~~~~~~~s~~~C~~~Cl~-nCsC~Ay~y~~~~~~C~~w~--~l~~~~~~~~~~~~~yikv 353 (356)
..|++++++.++..... .....++++|++.|++ +|+|.||.|..+++.|.+|. .+.+.......+.++|.|.
T Consensus 4 ~~f~~~~~~~l~~~~~~---~~~~~s~~~C~~~C~~~~~~C~s~~y~~~~~~C~l~~~~~~~~~~~~~~~~~~~y~~~ 78 (78)
T smart00473 4 DCFVRLPNTKLPGFSRI---VISVASLEECASKCLNSNCSCRSFTYNNGTKGCLLWSESSLGDARLFPSGGVDLYEKI 78 (78)
T ss_pred ceeEEecCccCCCCcce---eEcCCCHHHHHHHhCCCCCceEEEEEcCCCCEEEEeeCCccccceecccCCceeEEeC
Confidence 46889999887633221 1235799999999999 99999999976568999997 4666664455567888763
No 9
>cd01100 APPLE_Factor_XI_like Subfamily of PAN/APPLE-like domains; present in plasma prekallikrein/coagulation factor XI, microneme antigen proteins, and a few prokaryotic proteins. PAN/APPLE domains fulfill diverse biological functions by mediating protein-protein or protein-carbohydrate interactions.
Probab=97.82 E-value=2.6e-05 Score=58.48 Aligned_cols=51 Identities=16% Similarity=0.358 Sum_probs=37.9
Q ss_pred EeccccccccccccccCCCCCCHHHHHHHhhccCCeEEEEecCCCceeEEcccc
Q 018402 283 KLESVRHYMCTYNFYDGIGDITIEDCGKRCSSNCRCVGYFYDTSVSRCWIAFDL 336 (356)
Q Consensus 283 ~l~~~~~~~~~~~~~~~~~~~s~~~C~~~Cl~nCsC~Ay~y~~~~~~C~~w~~l 336 (356)
.+++++++..+.. .. ...+.++|+++|+.+|+|.||.|....+.|+++...
T Consensus 8 ~~~~~~~~g~d~~--~~-~~~s~~~Cq~~C~~~~~C~afT~~~~~~~C~lk~~~ 58 (73)
T cd01100 8 QGSNVDFRGGDLS--TV-FASSAEQCQAACTADPGCLAFTYNTKSKKCFLKSSE 58 (73)
T ss_pred ccCCCccccCCcc--ee-ecCCHHHHHHHcCCCCCceEEEEECCCCeEEcccCC
Confidence 3356666544433 11 246899999999999999999998777899999643
No 10
>PF00024 PAN_1: PAN domain This Prosite entry concerns apple domains, a subset of PAN domains; InterPro: IPR003014 PAN domains have significant functional versatility fulfilling diverse biological functions by mediating protein-protein or protein-carbohydrate interactions []. These domains contain a hair-pin loop like structure, similar to knottins, but the pattern of disulphide bonds differs It has been shown that, the N-terminal N domains of members of the plasminogen/hepatocyte growth factor family, the apple domains of the plasma prekallikrein/coagulation factor XI family, and domains of various nematode proteins belong to the same module superfamily, the PAN module []. PAN contains a conserved core of three disulphide bridges. In some members of the family there is an additional fourth disulphide bridge that links the N and C termini of the domain.; PDB: 1GP9_C 2QJ2_B 1GMO_H 1NK1_B 3MKP_B 1BHT_B 3HN4_A 1GMN_A 3HMS_A 3HMT_B ....
Probab=94.14 E-value=0.061 Score=39.90 Aligned_cols=52 Identities=21% Similarity=0.507 Sum_probs=38.2
Q ss_pred eEEEeccccccccccccccCCCCCCHHHHHHHhhccCC-eEEEEecCCCceeEEcc
Q 018402 280 HYYKLESVRHYMCTYNFYDGIGDITIEDCGKRCSSNCR-CVGYFYDTSVSRCWIAF 334 (356)
Q Consensus 280 ~f~~l~~~~~~~~~~~~~~~~~~~s~~~C~~~Cl~nCs-C~Ay~y~~~~~~C~~w~ 334 (356)
.|.++++..+...... .....++++|.+.|+.+=. |.+|.|......|.+..
T Consensus 3 ~f~~~~~~~l~~~~~~---~~~v~s~~~C~~~C~~~~~~C~s~~y~~~~~~C~L~~ 55 (79)
T PF00024_consen 3 AFERIPGYRLSGHSIK---EINVPSLEECAQLCLNEPRRCKSFNYDPSSKTCYLSS 55 (79)
T ss_dssp TEEEEEEEEEESCEEE---EEEESSHHHHHHHHHHSTT-ESEEEEETTTTEEEEEC
T ss_pred CeEEECCEEEeCCcce---EEcCCCHHHHHhhcCcCcccCCeEEEECCCCEEEEcC
Confidence 3666666655432222 1133589999999999999 99999987778999974
No 11
>smart00223 APPLE APPLE domain. Four-fold repeat in plasma kallikrein and coagulation factor XI. Factor XI apple 3 mediates binding to platelets. Factor XI apple 1 binds high-molecular-mass kininogen. Apple 4 in factor XI mediates dimer formation and binds to factor XIIa. Mutations in apple 4 cause factor XI deficiency, an inherited bleeding disorder.
Probab=92.83 E-value=0.16 Score=38.69 Aligned_cols=49 Identities=22% Similarity=0.421 Sum_probs=36.7
Q ss_pred eccccccccccccccCCCCCCHHHHHHHhhccCCeEEEEecCCCc---eeEEccc
Q 018402 284 LESVRHYMCTYNFYDGIGDITIEDCGKRCSSNCRCVGYFYDTSVS---RCWIAFD 335 (356)
Q Consensus 284 l~~~~~~~~~~~~~~~~~~~s~~~C~~~Cl~nCsC~Ay~y~~~~~---~C~~w~~ 335 (356)
.+++++...+.. .....+.++|++.|..+=.|.||.|..... .|+++..
T Consensus 6 ~~~~df~G~Dl~---~~~~~~~~~Cq~~Ct~~~~C~~FTf~~~~~~~~~C~LK~s 57 (79)
T smart00223 6 YKNVDFRGSDIN---TVYVPSAQVCQKRCTSHPRCLFFTFSTNEPPEEKCLLKDS 57 (79)
T ss_pred ccCccccCceee---eeecCCHHHHHHhhcCCCCccEEEeeCCCCCCCEeEeCcC
Confidence 356666655443 223468999999999999999999975555 8999864
No 12
>PF14295 PAN_4: PAN domain; PDB: 2YIL_E 2YIP_C 2YIO_A.
Probab=91.88 E-value=0.14 Score=34.66 Aligned_cols=32 Identities=22% Similarity=0.724 Sum_probs=17.8
Q ss_pred CCCHHHHHHHhhccCCeEEEEecC-----CCceeEEc
Q 018402 302 DITIEDCGKRCSSNCRCVGYFYDT-----SVSRCWIA 333 (356)
Q Consensus 302 ~~s~~~C~~~Cl~nCsC~Ay~y~~-----~~~~C~~w 333 (356)
..+.++|.++|..+=.|.+|.|.. ..+.|+++
T Consensus 15 ~~s~~~C~~~C~~~~~C~~~~~~~~~~~~~~~~C~LK 51 (51)
T PF14295_consen 15 ASSPEECQAACAADPGCQAFTFNPPGCPSSSGRCYLK 51 (51)
T ss_dssp ---HHHHHHHHHTSTT--EEEEETTEE----------
T ss_pred CCCHHHHHHHccCCCCCCEEEEECCCcccccccccCC
Confidence 468999999999999999999975 45678764
No 13
>PF01453 B_lectin: D-mannose binding lectin; InterPro: IPR001480 A bulb lectin super-family (Amaryllidaceae, Orchidaceae and Aliaceae) contains a ~115-residue-long domain whose overall three dimensional fold is very similar to that of [, ]: Dictyostelium discoideum comitin, an actin binding protein Curculigo latifolia curculin, a sweet tasting and taste-modifying protein This domain generally binds mannose, but in at least one protein, curculin, it is apparently devoid of mannose-binding activity. Each bulb-type lectin domain consists of three sequential beta-sheet subdomains (I, II, III) that are inter-related by pseudo three-fold symmetry. The three subdomains are flat four-stranded, antiparrallel beta-sheets. Together they form a 12-stranded beta-barrel in which the barrel axis coincides with the pseudo 3-fold axis.; GO: 0005529 sugar binding; PDB: 3M7H_A 3M7J_B 3MEZ_D 1DLP_A 1BWU_D 1KJ1_A 1B2P_A 1XD6_A 2DPF_C 2D04_B ....
Probab=91.53 E-value=0.39 Score=39.10 Aligned_cols=35 Identities=31% Similarity=0.549 Sum_probs=24.8
Q ss_pred eEEEEe-cCCCCcCCCceEEEccCccEEEEcCCCCc
Q 018402 37 RWVWEA-NRGKPVRENAVFSLGADGNLVLAEADGTV 71 (356)
Q Consensus 37 t~VWvA-Nr~~Pv~~~~~l~l~~~GnLvL~d~~~~~ 71 (356)
++||.. +..........+.|.++|||||.+..+.+
T Consensus 39 ~~iWss~~t~~~~~~~~~~~L~~~GNlvl~d~~~~~ 74 (114)
T PF01453_consen 39 SVIWSSNNTSGRGNSGCYLVLQDDGNLVLYDSSGNV 74 (114)
T ss_dssp EEEEE--S-TTSS-SSEEEEEETTSEEEEEETTSEE
T ss_pred CEEEEecccCCccccCeEEEEeCCCCEEEEeecceE
Confidence 569999 43333335688999999999999977654
No 14
>smart00108 B_lectin Bulb-type mannose-specific lectin.
Probab=91.03 E-value=0.36 Score=39.05 Aligned_cols=41 Identities=54% Similarity=0.952 Sum_probs=33.2
Q ss_pred eEEEEecCCCCcCCCceEEEccCccEEEEcCCCCcceEEeeeccC
Q 018402 37 RWVWEANRGKPVRENAVFSLGADGNLVLAEADGTVGNFIWQSFDY 81 (356)
Q Consensus 37 t~VWvANr~~Pv~~~~~l~l~~~GnLvL~d~~~~~~~~lWQSFD~ 81 (356)
.+||.++.. .-.....++|.+||||||.++. +++|||||||
T Consensus 74 ~~vW~S~t~-~~~~~~~~~L~ddGnlvl~~~~---~~~~W~Sf~~ 114 (114)
T smart00108 74 RVVWSSNTT-GANGNYVLVLLDDGNLVIYDSD---GNFLWQSFDY 114 (114)
T ss_pred CEEEEeccc-CCCCceEEEEeCCCCEEEECCC---CCEEeCCCCC
Confidence 578998886 2223468999999999999987 4689999997
No 15
>smart00605 CW CW domain.
Probab=90.43 E-value=1.7 Score=33.88 Aligned_cols=52 Identities=21% Similarity=0.407 Sum_probs=39.0
Q ss_pred CCCHHHHHHHhhccCCeEEEEecCCCceeEEcc--cccceeecCC-CCceEEEEec
Q 018402 302 DITIEDCGKRCSSNCRCVGYFYDTSVSRCWIAF--DLKTLTKEPD-SPIVGFIKVS 354 (356)
Q Consensus 302 ~~s~~~C~~~Cl~nCsC~Ay~y~~~~~~C~~w~--~l~~~~~~~~-~~~~~yikv~ 354 (356)
..+.++|...|..+..|+.+.... ...|.+.. .+..+++... .+..+=||+.
T Consensus 21 ~~sw~~Ci~~C~~~~~Cvlay~~~-~~~C~~f~~~~~~~v~~~~~~~~~~VAfK~~ 75 (94)
T smart00605 21 TLSWDECIQKCYEDSNCVLAYGNS-SETCYLFSYGTVLTVKKLSSSSGKKVAFKVS 75 (94)
T ss_pred CCCHHHHHHHHhCCCceEEEecCC-CCceEEEEcCCeEEEEEccCCCCcEEEEEEe
Confidence 468899999999999999876543 46898863 5677776553 4566778874
No 16
>cd00028 B_lectin Bulb-type mannose-specific lectin. The domain contains a three-fold internal repeat (beta-prism architecture). The consensus sequence motif QXDXNXVXY is involved in alpha-D-mannose recognition. Lectins are carbohydrate-binding proteins which specifically recognize diverse carbohydrates and mediate a wide variety of biological processes, such as cell-cell and host-pathogen interactions, serum glycoprotein turnover, and innate immune responses.
Probab=90.30 E-value=0.47 Score=38.53 Aligned_cols=42 Identities=57% Similarity=1.021 Sum_probs=35.1
Q ss_pred eEEEEecCCCCcCCCceEEEccCccEEEEcCCCCcceEEeeeccCC
Q 018402 37 RWVWEANRGKPVRENAVFSLGADGNLVLAEADGTVGNFIWQSFDYP 82 (356)
Q Consensus 37 t~VWvANr~~Pv~~~~~l~l~~~GnLvL~d~~~~~~~~lWQSFD~P 82 (356)
.+||..+... ......++|.+||||||.+.+ +++||||||||
T Consensus 75 ~~vW~S~~~~-~~~~~~~~L~ddGnlvl~~~~---~~~~W~Sf~~P 116 (116)
T cd00028 75 TVVWSSNTTR-VNGNYVLVLLDDGNLVLYDSD---GNFLWQSFDYP 116 (116)
T ss_pred cEEEEecccC-CCCceEEEEeCCCCEEEECCC---CCEEEcCCCCC
Confidence 5789988875 234578999999999999987 57899999999
No 17
>PF08277 PAN_3: PAN-like domain; InterPro: IPR006583 PAN domains have significant functional versatility fulfilling diverse biological functions by mediating protein-protein or protein-carbohydrate interactions []. These domains contain a hair-pin loop like structure, similar to knottins, but the pattern of disulphide bonds differs The PAN-3 or CW is a domain associated with a number of Caenorhabditis elegans hypothetical proteins.
Probab=90.18 E-value=0.9 Score=33.22 Aligned_cols=51 Identities=22% Similarity=0.503 Sum_probs=38.2
Q ss_pred CCCCHHHHHHHhhccCCeEEEEecCCCceeEEcc--cccceeecC-CCCceEEEEe
Q 018402 301 GDITIEDCGKRCSSNCRCVGYFYDTSVSRCWIAF--DLKTLTKEP-DSPIVGFIKV 353 (356)
Q Consensus 301 ~~~s~~~C~~~Cl~nCsC~Ay~y~~~~~~C~~w~--~l~~~~~~~-~~~~~~yikv 353 (356)
...+.++|-+.|+.+=+|.++.++ ...|.+.. .+..+++.. ..+..+-||+
T Consensus 18 ~~~sw~~Cv~~C~~~~~C~la~~~--~~~C~~y~~~~i~~v~~~~~~~~~~VA~K~ 71 (71)
T PF08277_consen 18 TNTSWDDCVQKCYNDENCVLAYFD--SGKCYLYNYGSISTVQKTDSSSGNKVAFKI 71 (71)
T ss_pred cCCCHHHHhHHhCCCCEEEEEEeC--CCCEEEEEcCCEEEEEEeecCCCeEEEEEC
Confidence 346889999999999999999887 56899973 566666543 3455556664
No 18
>cd01099 PAN_AP_HGF Subfamily of PAN/APPLE-like domains; present in N-terminal (N) domains of plasminogen/hepatocyte growth factor proteins, and various proteins found in Bilateria, such as leech anti-platelet proteins. PAN/APPLE domains fulfill diverse biological functions by mediating protein-protein or protein-carbohydrate interactions.
Probab=85.37 E-value=1.3 Score=33.60 Aligned_cols=33 Identities=18% Similarity=0.582 Sum_probs=28.7
Q ss_pred CCCHHHHHHHhhc--cCCeEEEEecCCCceeEEcc
Q 018402 302 DITIEDCGKRCSS--NCRCVGYFYDTSVSRCWIAF 334 (356)
Q Consensus 302 ~~s~~~C~~~Cl~--nCsC~Ay~y~~~~~~C~~w~ 334 (356)
..++++|.++|++ +=.|.++.|......|.+-.
T Consensus 24 ~~s~~~C~~~C~~~~~f~CrSf~y~~~~~~C~L~~ 58 (80)
T cd01099 24 VASLEECLRKCLEETEFTCRSFNYNYKSKECILSD 58 (80)
T ss_pred cCCHHHHHHHhCCCCCceEeEEEEEcCCCEEEEeC
Confidence 4789999999999 88999999976678999864
No 19
>cd00053 EGF Epidermal growth factor domain, found in epidermal growth factor (EGF) presents in a large number of proteins, mostly animal; the list of proteins currently known to contain one or more copies of an EGF-like pattern is large and varied; the functional significance of EGF-like domains in what appear to be unrelated proteins is not yet clear; a common feature is that these repeats are found in the extracellular domain of membrane-bound proteins or in proteins known to be secreted (exception: prostaglandin G/H synthase); the domain includes six cysteine residues which have been shown to be involved in disulfide bonds; the main structure is a two-stranded beta-sheet followed by a loop to a C-terminal short two-stranded sheet; Subdomains between the conserved cysteines vary in length; the region between the 5th and 6th cysteine contains two conserved glycines of which at least one is present in most EGF-like domains; a subset of these bind calcium.
Probab=66.53 E-value=6.3 Score=23.66 Aligned_cols=27 Identities=22% Similarity=0.501 Sum_probs=19.4
Q ss_pred CCCCCCCcccCCC-CCccCCCCCCCCCC
Q 018402 236 PERCGKLGVCEDN-QCVACPTEKGLIGW 262 (356)
Q Consensus 236 ~~~CG~~g~C~~~-~~~~C~C~~g~~~~ 262 (356)
...|...+.|... ....|.|+.||...
T Consensus 5 ~~~C~~~~~C~~~~~~~~C~C~~g~~g~ 32 (36)
T cd00053 5 SNPCSNGGTCVNTPGSYRCVCPPGYTGD 32 (36)
T ss_pred CCCCCCCCEEecCCCCeEeECCCCCccc
Confidence 4568888899753 34569999998643
No 20
>cd05845 Ig2_L1-CAM_like Second immunoglobulin (Ig)-like domain of the L1 cell adhesion molecule (CAM) and similar proteins. Ig2_L1-CAM_like: domain similar to the second immunoglobulin (Ig)-like domain of the L1 cell adhesion molecule (CAM). L1 belongs to the L1 subfamily of cell adhesion molecules (CAMs) and is comprised of an extracellular region having six Ig-like domains, five fibronectin type III domains, a transmembrane region and an intracellular domain. L1 is primarily expressed in the nervous system and is involved in its development and function. L1 is associated with an X-linked recessive disorder, X-linked hydrocephalus, MASA syndrome, or spastic paraplegia type 1, that involves abnormalities of axonal growth.
Probab=62.21 E-value=12 Score=29.37 Aligned_cols=34 Identities=24% Similarity=0.416 Sum_probs=25.6
Q ss_pred CceeEEEEecCCCCcCCCceEEEccCccEEEEcC
Q 018402 34 PVFRWVWEANRGKPVRENAVFSLGADGNLVLAEA 67 (356)
Q Consensus 34 ~~~t~VWvANr~~Pv~~~~~l~l~~~GnLvL~d~ 67 (356)
+..++.|+-+....+..+..+.++.+|||.+.+-
T Consensus 32 P~P~i~W~~~~~~~i~~~~Ri~~~~~GnL~fs~v 65 (95)
T cd05845 32 VPLRIYWMNSDLLHITQDERVSMGQNGNLYFANV 65 (95)
T ss_pred CCCEEEEECCCCccccccccEEECCCceEEEEEE
Confidence 4558899955545566677889988999999763
No 21
>PF01683 EB: EB module; InterPro: IPR006149 The EB domain has no known function. It is found in several Caenorhabditis sp. and Drosophila sp. proteins. The domain contains 8 conserved cysteines that probably form four disulphide bridges and is found associated with kunitz domains IPR002223 from INTERPRO
Probab=57.16 E-value=12 Score=25.46 Aligned_cols=31 Identities=23% Similarity=0.516 Sum_probs=24.7
Q ss_pred CCCCCCCCCCCCCCCcccCCCCCccCCCCCCCCC
Q 018402 228 NWDSECQMPERCGKLGVCEDNQCVACPTEKGLIG 261 (356)
Q Consensus 228 ~p~~~C~~~~~CG~~g~C~~~~~~~C~C~~g~~~ 261 (356)
.|.+.|....-|-.+++|..+ .|.|++||..
T Consensus 17 ~~g~~C~~~~qC~~~s~C~~g---~C~C~~g~~~ 47 (52)
T PF01683_consen 17 QPGESCESDEQCIGGSVCVNG---RCQCPPGYVE 47 (52)
T ss_pred CCCCCCCCcCCCCCcCEEcCC---EeECCCCCEe
Confidence 356789988899999999654 5999999854
No 22
>PF07354 Sp38: Zona-pellucida-binding protein (Sp38); InterPro: IPR010857 This family contains a number of zona-pellucida-binding proteins that seem to be restricted to mammals. These are sperm proteins that bind to the 90 kDa family of zona pellucida glycoproteins in a calcium-dependent manner []. These represent some of the specific molecules that mediate the first steps of gamete interaction, allowing fertilisation to occur [].; GO: 0007339 binding of sperm to zona pellucida, 0005576 extracellular region
Probab=51.39 E-value=17 Score=33.99 Aligned_cols=30 Identities=23% Similarity=0.666 Sum_probs=27.4
Q ss_pred eEEEEecCCCCcCCCceEEEccCccEEEEc
Q 018402 37 RWVWEANRGKPVRENAVFSLGADGNLVLAE 66 (356)
Q Consensus 37 t~VWvANr~~Pv~~~~~l~l~~~GnLvL~d 66 (356)
+..|+--.+.++.+++.++|++.|.|++.+
T Consensus 14 ~y~W~GP~g~~l~gn~~~nIT~TG~L~~~~ 43 (271)
T PF07354_consen 14 TYLWTGPNGKPLSGNSYVNITETGKLMFKN 43 (271)
T ss_pred ceEEECCCCcccCCCCeEEEccCceEEeec
Confidence 678999999999999999999999999854
No 23
>PF07645 EGF_CA: Calcium-binding EGF domain; InterPro: IPR001881 A sequence of about forty amino-acid residues found in epidermal growth factor (EGF) has been shown [, , , , , ] to be present in a large number of membrane-bound and extracellular, mostly animal, proteins. Many of these proteins require calcium for their biological function and a calcium-binding site has been found at the N terminus of some EGF-like domains []. Calcium-binding may be crucial for numerous protein-protein interactions. For human coagulation factor IX it has been shown [] that the calcium-ligands form a pentagonal bipyramid. The first, third and fourth conserved negatively charged or polar residues are side chain ligands. The latter is possibly hydroxylated (see aspartic acid and asparagine hydroxylation site) []. A conserved aromatic residue, as well as the second conserved negative residue, are thought to be involved in stabilising the calcium-binding site. As in non-calcium binding EGF-like domains, there are six conserved cysteines and the structure of both types is very similar as calcium-binding induces only strictly local structural changes []. +------------------+ +---------+ | | | | nxnnC-x(3,14)-C-x(3,7)-CxxbxxxxaxC-x(1,6)-C-x(8,13)-Cx | | +------------------+ 'n': negatively charged or polar residue [DEQN] 'b': possibly beta-hydroxylated residue [DN] 'a': aromatic amino acid 'C': cysteine, involved in disulphide bond 'x': any amino acid. ; GO: 0005509 calcium ion binding; PDB: 2VJ3_A 1TOZ_A 1LMJ_A 1UZQ_A 1UZK_A 1UZJ_B 1UZP_A 1EMO_A 1EMN_A 2RR0_A ....
Probab=50.28 E-value=4.8 Score=26.35 Aligned_cols=30 Identities=23% Similarity=0.577 Sum_probs=23.2
Q ss_pred CCCCCCC-CCCCCcccCCC-CCccCCCCCCCC
Q 018402 231 SECQMPE-RCGKLGVCEDN-QCVACPTEKGLI 260 (356)
Q Consensus 231 ~~C~~~~-~CG~~g~C~~~-~~~~C~C~~g~~ 260 (356)
|+|.... .|..++.|... ..-.|.|++||.
T Consensus 3 dEC~~~~~~C~~~~~C~N~~Gsy~C~C~~Gy~ 34 (42)
T PF07645_consen 3 DECAEGPHNCPENGTCVNTEGSYSCSCPPGYE 34 (42)
T ss_dssp STTTTTSSSSSTTSEEEEETTEEEEEESTTEE
T ss_pred cccCCCCCcCCCCCEEEcCCCCEEeeCCCCcE
Confidence 6788754 79999999753 445699999986
No 24
>smart00179 EGF_CA Calcium-binding EGF-like domain.
Probab=48.66 E-value=15 Score=22.66 Aligned_cols=30 Identities=23% Similarity=0.567 Sum_probs=21.2
Q ss_pred CCCCCCCCCCCCcccCCC-CCccCCCCCCCC
Q 018402 231 SECQMPERCGKLGVCEDN-QCVACPTEKGLI 260 (356)
Q Consensus 231 ~~C~~~~~CG~~g~C~~~-~~~~C~C~~g~~ 260 (356)
++|.....|...+.|... ....|.|++||.
T Consensus 3 ~~C~~~~~C~~~~~C~~~~g~~~C~C~~g~~ 33 (39)
T smart00179 3 DECASGNPCQNGGTCVNTVGSYRCECPPGYT 33 (39)
T ss_pred ccCcCCCCcCCCCEeECCCCCeEeECCCCCc
Confidence 567654568888899753 334599999885
No 25
>cd05852 Ig5_Contactin-1 Fifth Ig domain of contactin-1. Ig5_Contactin-1: fifth Ig domain of the neural cell adhesion molecule contactin-1. Contactins are comprised of six Ig domains followed by four fibronectin type III (FnIII) domains anchored to the membrane by glycosylphosphatidylinositol. Contactin-1 is differentially expressed in tumor tissues and may through a RhoA mechanism, facilitate invasion and metastasis of human lung adenocarcinoma.
Probab=42.86 E-value=28 Score=25.46 Aligned_cols=33 Identities=21% Similarity=0.423 Sum_probs=22.9
Q ss_pred CceeEEEEecCCCCcCCCceEEEccCccEEEEcC
Q 018402 34 PVFRWVWEANRGKPVRENAVFSLGADGNLVLAEA 67 (356)
Q Consensus 34 ~~~t~VWvANr~~Pv~~~~~l~l~~~GnLvL~d~ 67 (356)
|..++.|.-+. .++..+..+.+..+|.|+|.+-
T Consensus 14 P~p~v~W~k~~-~~l~~~~r~~~~~~g~L~I~~v 46 (73)
T cd05852 14 PKPKFSWSKGT-ELLVNNSRISIWDDGSLEILNI 46 (73)
T ss_pred CCCEEEEEeCC-EecccCCCEEEcCCCEEEECcC
Confidence 44578998654 3555556777777888888654
No 26
>cd00054 EGF_CA Calcium-binding EGF-like domain, present in a large number of membrane-bound and extracellular (mostly animal) proteins. Many of these proteins require calcium for their biological function and calcium-binding sites have been found to be located at the N-terminus of particular EGF-like domains; calcium-binding may be crucial for numerous protein-protein interactions. Six conserved core cysteines form three disulfide bridges as in non calcium-binding EGF domains, whose structures are very similar. EGF_CA can be found in tandem repeat arrangements.
Probab=42.40 E-value=22 Score=21.50 Aligned_cols=30 Identities=23% Similarity=0.577 Sum_probs=20.5
Q ss_pred CCCCCCCCCCCCcccCCC-CCccCCCCCCCC
Q 018402 231 SECQMPERCGKLGVCEDN-QCVACPTEKGLI 260 (356)
Q Consensus 231 ~~C~~~~~CG~~g~C~~~-~~~~C~C~~g~~ 260 (356)
++|.....|...+.|... ....|.|++||.
T Consensus 3 ~~C~~~~~C~~~~~C~~~~~~~~C~C~~g~~ 33 (38)
T cd00054 3 DECASGNPCQNGGTCVNTVGSYRCSCPPGYT 33 (38)
T ss_pred ccCCCCCCcCCCCEeECCCCCeEeECCCCCc
Confidence 567654568777889653 334699998875
No 27
>PF10681 Rot1: Chaperone for protein-folding within the ER, fungal; InterPro: IPR019623 This conserved fungal family is an essential molecular chaperone in the endoplasmic reticulum. Molecular chaperones transiently interact with unfolded proteins to inhibit their self-aggregation and to support their folding and/or assembly. Rot1 is a general chaperone with some substrate specificity, its substrates being the structurally unrelated Kre5 Kre6 Big1 Atg22, which are type I, type II, and polytopic membrane proteins. The dependencies of each for Rot1 do not share similarities. However, their folding does require BiP, and one of these proteins was simultaneously associated with both Rot1 and BiP. In addition, Rot1 may cooperate with BiP/Kar2 in the folding of Kre6 [].
Probab=39.43 E-value=1.3e+02 Score=27.27 Aligned_cols=102 Identities=17% Similarity=0.228 Sum_probs=57.8
Q ss_pred EEeecCCCCCe----EEEEEEceeCC--CCceeEEEEecCCCCcCC-------CceEEEccCccEEEEc--CCCCcceEE
Q 018402 11 LGFYNTTPNAF----TLALRLGIKKQ--EPVFRWVWEANRGKPVRE-------NAVFSLGADGNLVLAE--ADGTVGNFI 75 (356)
Q Consensus 11 lGFf~~~~~~~----~l~Iw~~~~~~--~~~~t~VWvANr~~Pv~~-------~~~l~l~~~GnLvL~d--~~~~~~~~l 75 (356)
-|||+|-...| .=||-|.-..+ -.+-....++|..+|--- .++-+|.++|.|+|.- .+|+
T Consensus 16 pgFydPv~~~f~eP~~~GISYSFT~DG~~EeA~Yr~~~Np~~p~C~~a~l~wQHGtY~l~~nGsl~L~P~~~DGr----- 90 (212)
T PF10681_consen 16 PGFYDPVDELFIEPSLPGISYSFTEDGYFEEAQYRVTSNPTNPSCPTAVLIWQHGTYELNSNGSLTLTPFAVDGR----- 90 (212)
T ss_pred CcccChhHhhccCCCCCceeEEEcCCCeeeEEEEEEccCCCCCCCCceEEEEecceEEECCCCcEEEeecCCCCc-----
Confidence 48998732111 23455532211 012245678888777321 4678888899999853 3432
Q ss_pred eeeccCCCCccccCceeecCCeEEEEecCCCCCCCCcceEEEEe
Q 018402 76 WQSFDYPTDTLLVGQSLRVGRVTKLVSRLSVKENVDGPHSFVME 119 (356)
Q Consensus 76 WQSFD~PTDTlLPgqkL~~g~~~~L~Sw~s~~dps~G~fsl~l~ 119 (356)
|-+-.|...= -..--+.+..-.+.+|.-..|+-.|.|.|+|-
T Consensus 91 -Ql~sdPC~~~-~s~y~rYnq~e~f~~~~v~~D~y~~~~~L~L~ 132 (212)
T PF10681_consen 91 -QLVSDPCADD-SSTYTRYNQTELFKSFDVYVDPYHGRYRLQLY 132 (212)
T ss_pred -eeccCCCCCC-cccEEEEcceEEEEEEEEEEeCCCCeeEEEEE
Confidence 4444555443 11111233333567788888998999999885
No 28
>PF07974 EGF_2: EGF-like domain; InterPro: IPR013111 A sequence of about thirty to forty amino-acid residues long found in the sequence of epidermal growth factor (EGF) has been shown [, , , , ] to be present, in a more or less conserved form, in a large number of other, mostly animal proteins. The list of proteins currently known to contain one or more copies of an EGF-like pattern is large and varied. The functional significance of EGF domains in what appear to be unrelated proteins is not yet clear. However, a common feature is that these repeats are found in the extracellular domain of membrane-bound proteins or in proteins known to be secreted (exception: prostaglandin G/H synthase). The EGF domain includes six cysteine residues which have been shown (in EGF) to be involved in disulphide bonds. The main structure is a two-stranded beta-sheet followed by a loop to a C-terminal short two-stranded sheet. Subdomains between the conserved cysteines vary in length. This entry contains EGF domains found in a variety of extracellular and membrane proteins
Probab=37.42 E-value=26 Score=21.64 Aligned_cols=23 Identities=26% Similarity=0.707 Sum_probs=17.1
Q ss_pred CCCCCCcccCCCCCccCCCCCCCC
Q 018402 237 ERCGKLGVCEDNQCVACPTEKGLI 260 (356)
Q Consensus 237 ~~CG~~g~C~~~~~~~C~C~~g~~ 260 (356)
..|...|+|+.. ...|.|.+||.
T Consensus 6 ~~C~~~G~C~~~-~g~C~C~~g~~ 28 (32)
T PF07974_consen 6 NICSGHGTCVSP-CGRCVCDSGYT 28 (32)
T ss_pred CccCCCCEEeCC-CCEEECCCCCc
Confidence 468889999754 23599998874
No 29
>PF09064 Tme5_EGF_like: Thrombomodulin like fifth domain, EGF-like; InterPro: IPR015149 This domain adopts a fold similar to other EGF domains, with a flat major and a twisted minor beta sheet. Disulphide pairing, however, is not of the usual 1-3, 2-4, 5-6 type; rather 1-2, 3-4, 5-6 pairing is found. Its extended major sheet (strands beta-2 and beta-3 and the connecting loop) projects into thrombin's active site groove. This domain is required for interaction of thrombomodulin with thrombin, and subsequent activation of protein-C []. ; GO: 0004888 transmembrane signaling receptor activity, 0016021 integral to membrane
Probab=32.62 E-value=26 Score=22.11 Aligned_cols=16 Identities=31% Similarity=0.613 Sum_probs=10.7
Q ss_pred cCCCCCccCCCCCCCC
Q 018402 245 CEDNQCVACPTEKGLI 260 (356)
Q Consensus 245 C~~~~~~~C~C~~g~~ 260 (356)
|+.+..-.|.|+.||.
T Consensus 12 CDpn~~~~C~CPeGyI 27 (34)
T PF09064_consen 12 CDPNSPGQCFCPEGYI 27 (34)
T ss_pred cCCCCCCceeCCCceE
Confidence 4444344699999984
No 30
>smart00564 PQQ beta-propeller repeat. Beta-propeller repeat occurring in enzymes with pyrrolo-quinoline quinone (PQQ) as cofactor, in Ire1p-like Ser/Thr kinases, and in prokaryotic dehydrogenases.
Probab=27.72 E-value=62 Score=19.12 Aligned_cols=19 Identities=32% Similarity=0.760 Sum_probs=12.7
Q ss_pred ccCccEEEEcCCCCcceEEee
Q 018402 57 GADGNLVLAEADGTVGNFIWQ 77 (356)
Q Consensus 57 ~~~GnLvL~d~~~~~~~~lWQ 77 (356)
+.+|.|+-.|... |+.+|+
T Consensus 13 ~~~g~l~a~d~~~--G~~~W~ 31 (33)
T smart00564 13 STDGTLYALDAKT--GEILWT 31 (33)
T ss_pred cCCCEEEEEEccc--CcEEEE
Confidence 4567777666533 678887
No 31
>KOG4649 consensus PQQ (pyrrolo-quinoline quinone) repeat protein [Secondary metabolites biosynthesis, transport and catabolism]
Probab=27.50 E-value=2.1e+02 Score=27.22 Aligned_cols=41 Identities=22% Similarity=0.390 Sum_probs=29.5
Q ss_pred CceeEEEEecCCCCcCCC-----ceEEE-ccCccEEEEcCCCCcceEEee
Q 018402 34 PVFRWVWEANRGKPVREN-----AVFSL-GADGNLVLAEADGTVGNFIWQ 77 (356)
Q Consensus 34 ~~~t~VWvANr~~Pv~~~-----~~l~l-~~~GnLvL~d~~~~~~~~lWQ 77 (356)
.+.+..|-|.|..|+-.+ +.+.+ +-||+|.-.++.| +.+||
T Consensus 167 ~~~~~~w~~~~~~PiF~splcv~~sv~i~~VdG~l~~f~~sG---~qvwr 213 (354)
T KOG4649|consen 167 YSSTEFWAATRFGPIFASPLCVGSSVIITTVDGVLTSFDESG---RQVWR 213 (354)
T ss_pred CCcceehhhhcCCccccCceeccceEEEEEeccEEEEEcCCC---cEEEe
Confidence 345889999999998653 33444 5689998888775 45674
No 32
>PF01436 NHL: NHL repeat; InterPro: IPR001258 The NHL repeat, named after NCL-1, HT2A and Lin-41, is found largely in a large number of eukaryotic and prokaryotic proteins. For example, the repeat is found in a variety of enzymes of the copper type II, ascorbate-dependent monooxygenase family which catalyse the C terminus alpha-amidation of biological peptides []. In many it occurs in tandem arrays, for example in the ringfinger beta-box, coiled-coil (RBCC) eukaryotic growth regulators []. The 'Brain Tumor' protein (Brat) is one such growth regulator that contains a 6-bladed NHL-repeat beta-propeller [, ]. The NHL repeats are also found in serine/threonine protein kinase (STPK) in diverse range of pathogenic bacteria. These STPK are transmembrane receptors with a intracellular N-terminal kinase domain and extracellular C-terminal sensor domain. In the STPK, PknD, from Mycobacterium tuberculosis, the sensor domain forms a rigid, six-bladed b-propeller composed of NHL repeats with a flexible tether to the transmembrane domain.; GO: 0005515 protein binding; PDB: 3FVZ_A 3FW0_A 1RWL_A 1RWI_A 1Q7F_A.
Probab=24.14 E-value=85 Score=18.37 Aligned_cols=18 Identities=22% Similarity=0.471 Sum_probs=14.1
Q ss_pred ceEEEccCccEEEEcCCC
Q 018402 52 AVFSLGADGNLVLAEADG 69 (356)
Q Consensus 52 ~~l~l~~~GnLvL~d~~~ 69 (356)
.-+.++++|+|++.|..+
T Consensus 5 ~gvav~~~g~i~VaD~~n 22 (28)
T PF01436_consen 5 HGVAVDSDGNIYVADSGN 22 (28)
T ss_dssp EEEEEETTSEEEEEECCC
T ss_pred cEEEEeCCCCEEEEECCC
Confidence 457788999999999653
No 33
>cd05764 Ig_2 Subgroup of the immunoglobulin (Ig) superfamily. Ig_2: subgroup of the immunoglobulin (Ig) domain found in the Ig superfamily. The Ig superfamily is a heterogenous group of proteins, built on a common fold comprised of a sandwich of two beta sheets. Members of the Ig superfamily are components of immunoglobulin, neuroglia, cell surface glycoproteins, such as T-cell receptors, CD2, CD4, CD8, and membrane glycoproteins, such as butyrophilin and chondroitin sulfate proteoglycan core protein. A predominant feature of most Ig domains is a disulfide bridge connecting the two beta-sheets with a tryptophan residue packed against the disulfide bond.
Probab=23.35 E-value=1.3e+02 Score=21.38 Aligned_cols=32 Identities=19% Similarity=0.266 Sum_probs=18.4
Q ss_pred CceeEEEEecCCCCcCCCceEEEccCccEEEE
Q 018402 34 PVFRWVWEANRGKPVRENAVFSLGADGNLVLA 65 (356)
Q Consensus 34 ~~~t~VWvANr~~Pv~~~~~l~l~~~GnLvL~ 65 (356)
|..++.|.-+.+.++.......+..+|.|.|.
T Consensus 14 P~p~v~W~~~~~~~~~~~~~~~~~~~~~L~i~ 45 (74)
T cd05764 14 PEPAIHWISPDGKLISNSSRTLVYDNGTLDIL 45 (74)
T ss_pred CCCEEEEEeCCCEEecCCCeEEEecCCEEEEE
Confidence 44578898655555544444444455666654
Done!