Query psy17777
Match_columns 1375
No_of_seqs 829 out of 4393
Neff 6.2
Searched_HMMs 46136
Date Fri Aug 16 22:43:34 2013
Command hhsearch -i /work/01045/syshi/Psyhhblits/psy17777.a3m -d /work/01045/syshi/HHdatabase/Cdd.hhm -o /work/01045/syshi/hhsearch_cdd/17777hhsearch_cdd -cpu 12 -v 0
No Hit Prob E-value P-value Score SS Cols Query HMM Template HMM
1 KOG3546|consensus 100.0 1.5E-38 3.3E-43 366.5 57.5 12 1172-1183 1101-1112(1167)
2 KOG3546|consensus 100.0 9.3E-39 2E-43 368.3 55.2 83 326-411 455-537 (1167)
3 PF01413 C4: C-terminal tandem 100.0 4.6E-41 9.9E-46 326.7 6.1 112 1261-1374 2-114 (114)
4 smart00111 C4 C-terminal tande 100.0 8.6E-40 1.9E-44 317.3 8.5 112 1261-1374 2-114 (114)
5 smart00111 C4 C-terminal tande 100.0 3.7E-32 8.1E-37 264.1 6.7 110 909-1020 2-112 (114)
6 PF01413 C4: C-terminal tandem 100.0 2E-31 4.2E-36 259.5 3.2 106 1147-1258 1-113 (114)
7 PF01410 COLFI: Fibrillar coll 99.3 7.4E-13 1.6E-17 146.1 4.2 117 1204-1343 7-132 (214)
8 smart00038 COLFI Fibrillar col 99.0 1.5E-10 3.3E-15 129.1 4.5 129 1192-1344 13-150 (232)
9 TIGR03032 conserved hypothetic 24.1 1.6E+02 0.0035 35.1 6.4 28 1255-1284 189-217 (335)
10 KOG0808|consensus 21.2 36 0.00078 38.7 0.4 32 1310-1342 332-363 (387)
11 KOG1446|consensus 21.1 1.1E+02 0.0023 36.2 4.1 34 1339-1374 106-140 (311)
12 PHA03112 IL-18 binding protein 20.4 80 0.0017 33.3 2.7 20 1317-1340 41-61 (141)
13 KOG1388|consensus 20.1 1.2E+02 0.0025 34.2 4.0 50 1282-1344 147-196 (217)
No 1
>KOG3546|consensus
Probab=100.00 E-value=1.5e-38 Score=366.47 Aligned_cols=12 Identities=33% Similarity=0.614 Sum_probs=7.7
Q ss_pred eeeeeecccccc
Q psy17777 1172 TKLWEGYSLLYV 1183 (1375)
Q Consensus 1172 ~~l~~Gysll~~ 1183 (1375)
..+|.+.|.|||
T Consensus 1101 klvwhgsSplgi 1112 (1167)
T KOG3546|consen 1101 KLVWHGSSPLGI 1112 (1167)
T ss_pred hheecCCCccch
Confidence 356777776665
No 2
>KOG3546|consensus
Probab=100.00 E-value=9.3e-39 Score=368.25 Aligned_cols=83 Identities=42% Similarity=0.768 Sum_probs=38.3
Q ss_pred CCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCCC
Q psy17777 326 KGEPGINGTHGVKGEKGQPGGIGYKGEPGQDGAKGDSGGRCIGCLPGARGEKGDRGKDGLPGIPGPPGAPGMRGRDGDAG 405 (1375)
Q Consensus 326 ~G~~G~~G~~G~~G~~G~~G~~G~~G~~G~~G~~G~~G~~g~~G~~G~~G~~G~~G~~G~~G~~G~~G~~G~~G~~G~~G 405 (1375)
.|.+|.+|..|.+|++|.+|..|.+|..|.+|.+|++|.+ |..|..|+||.+|.+|.+|.+|+.|.+|..|.+|++|
T Consensus 455 ~g~pg~pg~~gp~g~pg~pgp~g~pg~~g~pg~~g~kg~~---g~~g~~g~kg~~g~pgp~g~~g~ag~pg~~gppgppg 531 (1167)
T KOG3546|consen 455 AGLPGVPGREGPPGFPGLPGPPGPPGREGPPGRTGQKGSL---GEAGAPGHKGSKGAPGPAGARGEAGAPGPAGPPGPPG 531 (1167)
T ss_pred CCCCCCCCCcCCCCCCCCCCCCCCCCCCCCCCcccccccc---CCCCCCcccCCCCCCCCCCcccCCCCCCCCCCCCCCC
Confidence 3445555555555544444444444444444444444333 3334444444444444444444444444444444444
Q ss_pred CCCCCC
Q psy17777 406 RDGPPG 411 (1375)
Q Consensus 406 ~~G~~G 411 (1375)
.+|.+|
T Consensus 532 ppgppg 537 (1167)
T KOG3546|consen 532 PPGPPG 537 (1167)
T ss_pred CCCCCC
Confidence 444443
No 3
>PF01413 C4: C-terminal tandem repeated domain in type 4 procollagen; InterPro: IPR001442 Collagens are major components of the extracellular matrices of all metazoan life and play crucial roles in developmental processes and tissue homeostasis. Collagens are composed of three polypeptide chains (alpha chains) that fold together to form the characteristic triple helical collagenous domain. Some types of triple helical protomers contain genetically identical alpha chains forming homotrimers, whereas others contain two or three different alpha chains forming heterotrimers. The sequences required to form a collagenous domain are Gly-X-Y repeats in which the X and Y positions are frequently proline and hydroxyproline. Glycine is required every third residue as it is the only amino acid small enough to pack into the central core of the triple helix. The triple helix-forming parts are surrounded by non-collagenous (NC) domains of variable sequence, size, and shape. Even if the triple helical parts represent the most striking feature of collagens, tissue specificity as well as defined binding of non-collagens seem to be encoded in the NC domains. The terminal NC domains are excised, modified, or incorporated directly into the final suprastructure, depending on protomer type and function [, ]. Type IV collagen is one of the major constituents of basement membranes, a specialised form of extracellular matrix underlying epithelia that compartmentalises tissues and provides molecular signals for influencing cell behaviour. Each type IV chain contains a long triple-helical collagenous domain flanked by a short 7S domain of 25 residues and a globular non-collagenous NC1 domain of ~230 residues at the N and C terminus, respectively. In protomer assembly, the NC1 domains (monomers) of three chains interact, forming an NC1 trimer, to select and register chains for triple helix formation. In network assembly, the NC1 trimers of two protomers interact, forming a NC1 hexamer structure, to select and connect protomers [, , ]. The collagen IV NC1 domain contains 12 cysteines, and all of them are involved in disulphide bonds. It folds into a tertiary structure with predominantly beta-strands. The collagen IV NC1 domain is composed of two similarly folded subdomains stabilised by 3 intrachain dissulphide bonds involving the following pairs: C1-C6, C2-C5, and C3-C4. Each subdomain represents a compact disulphide-stabilised triangular structure, from which a finger-like hairpin loop projects into an incompletely formed six-stranded beta-sheet of an adjacent subdomain of the same or of an adjacent chain clamping the subdomains tightly together [, , ]. This duplicated domain is present at the C-terminal of type 4 collagen, the major structural component of glomerular basement membranes (GMB) forming a 'chicken-wire' meshwork together with laminins, proteoglycans and entactin/nidogen. Mutations in alpha-5 collagen IV are associated with X-linked Alport syndrome.; GO: 0005201 extracellular matrix structural constituent, 0005581 collagen; PDB: 1T61_F 1M3D_L 1LI1_D.
Probab=100.00 E-value=4.6e-41 Score=326.65 Aligned_cols=112 Identities=61% Similarity=1.197 Sum_probs=95.0
Q ss_pred ceeEeccCCcccCCCCCCCccceeeeEEEEEcCCCCCCCCCCCCCCCchhhhcccCCceeeecCCCceeeccCCcceeee
Q psy17777 1261 NVIAVHSQSVVIPECPVGWNSLWIGYSFVMHTAAGGDGGGQSLASPGSCLEEFRATPFIECNGEHGSCHYFANKLSFWLA 1340 (1375)
Q Consensus 1261 ~~iavhs~~~~ip~Cp~gW~~~w~Gysf~~~t~~g~~g~gq~l~spgscl~~~r~~~~iec~~~~g~c~~~~~~~s~w~~ 1340 (1375)
.+||||||+..||+||.+|++||+||||+|||+ +++++||.|.||||||++||.+|||||+. +++|||++|.+||||+
T Consensus 2 ~~ia~HSQt~~iP~CP~G~~~LW~GYSfl~~tg-~~~~~gQdLgspGSCL~~F~~~Pfi~C~~-~g~C~~~~n~~S~WLs 79 (114)
T PF01413_consen 2 FVIAVHSQTDQIPDCPPGWTSLWTGYSFLMHTG-GGEGHGQDLGSPGSCLERFRTMPFIECNG-RGTCNYASNDYSYWLS 79 (114)
T ss_dssp EEEEEE-SSSS-----TTEEEEEEEEEEEEEEE-TTEEEE--TTSGGGEESSEESS-EEEEET-TSEEEESTT-EEEEEB
T ss_pred eEEEEecCCCCCCcCCCchhHHhhhhhhhheec-CCccccccccccchhHHHHhcCcceecCC-CCcCCccCCCcceeEE
Confidence 479999999999999999999999999999999 89999999999999999999999999999 9999999999999999
Q ss_pred ccCCCCCcCCCccccccCC-cCCCceeeeEecccC
Q psy17777 1341 TIDPQDQWSRPQQQTLKSG-NLRQRISRCQVCIKR 1374 (1375)
Q Consensus 1341 ~~~~~~~~~~p~~~t~~~~-~~~~~~src~vc~~~ 1374 (1375)
|++.++||++|+++|+|++ +++++||||+|||||
T Consensus 80 t~~~~~~f~~p~~~~~~~~~~~~~~ISRC~VC~~~ 114 (114)
T PF01413_consen 80 TIEPNDQFQMPMPMTPKAGEEIRSYISRCSVCEAN 114 (114)
T ss_dssp -S-CCCTTSS-TTEEEECCGCGGGGB-EEEEEEE-
T ss_pred ccCchhhccCCCCccccCccccccccccceEeecC
Confidence 9999999999999999999 889999999999997
No 4
>smart00111 C4 C-terminal tandem repeated domain in type 4 procollagens. Duplicated domain in C-terminus of type 4 collagens. Mutations in alpha-5 collagen IV are associated with X-linked Alport syndrome.
Probab=100.00 E-value=8.6e-40 Score=317.28 Aligned_cols=112 Identities=62% Similarity=1.203 Sum_probs=109.4
Q ss_pred ceeEeccCCcccCCCCCCCccceeeeEEEEEcCCCCCCCCCCCCCCCchhhhcccCCceeeecCCCceeecc-CCcceee
Q psy17777 1261 NVIAVHSQSVVIPECPVGWNSLWIGYSFVMHTAAGGDGGGQSLASPGSCLEEFRATPFIECNGEHGSCHYFA-NKLSFWL 1339 (1375)
Q Consensus 1261 ~~iavhs~~~~ip~Cp~gW~~~w~Gysf~~~t~~g~~g~gq~l~spgscl~~~r~~~~iec~~~~g~c~~~~-~~~s~w~ 1339 (1375)
.+||||||+..||+||.+|+++|+||||+|||+ +++++||+|.||||||++||++|||||+. +++|||++ |.+||||
T Consensus 2 ~~~a~HSQt~~iP~CP~g~~~LW~GyS~l~~tg-~~~g~gQdL~spGSCL~~F~~~Pfi~C~~-~~~C~y~~rn~~SfWL 79 (114)
T smart00111 2 FVIAVHSQTTNVPQCPAGWVELWTGYSFLMHTG-NGEGHGQDLGSPGSCLERFRTMPFIECNG-RGVCNYASRNDYSFWL 79 (114)
T ss_pred ceEEEecCCCCCCCCCCCCcccccceeEEEecC-CCCccCCCCCCCccchhhhccCCeEEECC-CCeecccccCCcceEE
Confidence 579999999999999999999999999999998 89999999999999999999999999999 99999999 9999999
Q ss_pred eccCCCCCcCCCccccccCCcCCCceeeeEecccC
Q psy17777 1340 ATIDPQDQWSRPQQQTLKSGNLRQRISRCQVCIKR 1374 (1375)
Q Consensus 1340 ~~~~~~~~~~~p~~~t~~~~~~~~~~src~vc~~~ 1374 (1375)
+++++++||++|+++|+++++++++||||+||||+
T Consensus 80 st~~~~~~f~~p~~~~~~~~~~~~~ISRC~VC~~~ 114 (114)
T smart00111 80 STIEPSDQFTAPRPMTPKAGDLRPYISRCQVCEKP 114 (114)
T ss_pred eccCccccccccccccccchhhhhhccccceeecc
Confidence 99999999999999999999999999999999985
No 5
>smart00111 C4 C-terminal tandem repeated domain in type 4 procollagens. Duplicated domain in C-terminus of type 4 collagens. Mutations in alpha-5 collagen IV are associated with X-linked Alport syndrome.
Probab=99.97 E-value=3.7e-32 Score=264.13 Aligned_cols=110 Identities=60% Similarity=1.176 Sum_probs=107.2
Q ss_pred ceEEEeccccccCCCCCCcccccccCCcccccCCCCCCCCCCCCCCCCcccCcccCcccccCCCCCCCCccc-cccceeE
Q psy17777 909 NVIAVHSQSVVIPECPVGWNSLWIGYSFVMHTAAGGDGGGQSLASPGSCLEEFRATPFIECNGEHGSCHYFA-NKLSFWL 987 (1375)
Q Consensus 909 ~~~a~hsq~~~ip~cp~gw~~l~~~~~~~~~~g~~g~~g~~~~~spGscl~~~~~~pfie~~G~~G~~~~~g-~~~~fwl 987 (1375)
.+||+|||+..||+||.+|+.||.+|||+|||| .++.+.|.|.||||||+.|+..|||||++ +++|+|+. +.++|||
T Consensus 2 ~~~a~HSQt~~iP~CP~g~~~LW~GyS~l~~tg-~~~g~gQdL~spGSCL~~F~~~Pfi~C~~-~~~C~y~~rn~~SfWL 79 (114)
T smart00111 2 FVIAVHSQTTNVPQCPAGWVELWTGYSFLMHTG-NGEGHGQDLGSPGSCLERFRTMPFIECNG-RGVCNYASRNDYSFWL 79 (114)
T ss_pred ceEEEecCCCCCCCCCCCCcccccceeEEEecC-CCCccCCCCCCCccchhhhccCCeEEECC-CCeecccccCCcceEE
Confidence 579999999999999999999999999999999 89999999999999999999999999999 79999999 9999999
Q ss_pred eeeCCCccccccccccccccccccccccccccC
Q psy17777 988 ATIDPQDQWSRPQQQTLKSGNLRQRISRCQLFL 1020 (1375)
Q Consensus 988 ~~i~~~~~~~~p~~~tlk~g~~~~risrC~g~~ 1020 (1375)
+++++.++|.+|.++++++++++++||||+||+
T Consensus 80 st~~~~~~f~~p~~~~~~~~~~~~~ISRC~VC~ 112 (114)
T smart00111 80 STIEPSDQFTAPRPMTPKAGDLRPYISRCQVCE 112 (114)
T ss_pred eccCccccccccccccccchhhhhhccccceee
Confidence 999999999999999999999999999999997
No 6
>PF01413 C4: C-terminal tandem repeated domain in type 4 procollagen; InterPro: IPR001442 Collagens are major components of the extracellular matrices of all metazoan life and play crucial roles in developmental processes and tissue homeostasis. Collagens are composed of three polypeptide chains (alpha chains) that fold together to form the characteristic triple helical collagenous domain. Some types of triple helical protomers contain genetically identical alpha chains forming homotrimers, whereas others contain two or three different alpha chains forming heterotrimers. The sequences required to form a collagenous domain are Gly-X-Y repeats in which the X and Y positions are frequently proline and hydroxyproline. Glycine is required every third residue as it is the only amino acid small enough to pack into the central core of the triple helix. The triple helix-forming parts are surrounded by non-collagenous (NC) domains of variable sequence, size, and shape. Even if the triple helical parts represent the most striking feature of collagens, tissue specificity as well as defined binding of non-collagens seem to be encoded in the NC domains. The terminal NC domains are excised, modified, or incorporated directly into the final suprastructure, depending on protomer type and function [, ]. Type IV collagen is one of the major constituents of basement membranes, a specialised form of extracellular matrix underlying epithelia that compartmentalises tissues and provides molecular signals for influencing cell behaviour. Each type IV chain contains a long triple-helical collagenous domain flanked by a short 7S domain of 25 residues and a globular non-collagenous NC1 domain of ~230 residues at the N and C terminus, respectively. In protomer assembly, the NC1 domains (monomers) of three chains interact, forming an NC1 trimer, to select and register chains for triple helix formation. In network assembly, the NC1 trimers of two protomers interact, forming a NC1 hexamer structure, to select and connect protomers [, , ]. The collagen IV NC1 domain contains 12 cysteines, and all of them are involved in disulphide bonds. It folds into a tertiary structure with predominantly beta-strands. The collagen IV NC1 domain is composed of two similarly folded subdomains stabilised by 3 intrachain dissulphide bonds involving the following pairs: C1-C6, C2-C5, and C3-C4. Each subdomain represents a compact disulphide-stabilised triangular structure, from which a finger-like hairpin loop projects into an incompletely formed six-stranded beta-sheet of an adjacent subdomain of the same or of an adjacent chain clamping the subdomains tightly together [, , ]. This duplicated domain is present at the C-terminal of type 4 collagen, the major structural component of glomerular basement membranes (GMB) forming a 'chicken-wire' meshwork together with laminins, proteoglycans and entactin/nidogen. Mutations in alpha-5 collagen IV are associated with X-linked Alport syndrome.; GO: 0005201 extracellular matrix structural constituent, 0005581 collagen; PDB: 1T61_F 1M3D_L 1LI1_D.
Probab=99.96 E-value=2e-31 Score=259.51 Aligned_cols=106 Identities=52% Similarity=1.042 Sum_probs=89.0
Q ss_pred cceeecccCCCCCCccCCCCCccceeeeeeeccccccccCccccccCcCcccccccccCCCCccccCCCCccccCCCCCc
Q psy17777 1147 GILMVKHSQTAEVPSCTDDKTNTQFTKLWEGYSLLYVEGNEQAHNQDLGFAGSCVRKFSTMPFLFCDPNNVCNYASRNDR 1226 (1375)
Q Consensus 1147 g~~~~rhsqs~~~p~cp~g~~~~~~~~l~~Gysll~~~g~~~a~~qdlG~~gScl~~fstmPf~~C~~~~~C~~a~~~d~ 1226 (1375)
+|++++|||+++||.||.+ +++||+|||||++.+++++|+|||++++|||++|++|||++|+.+++|||++ ||+
T Consensus 1 g~~ia~HSQt~~iP~CP~G-----~~~LW~GYSfl~~tg~~~~~gQdLgspGSCL~~F~~~Pfi~C~~~g~C~~~~-n~~ 74 (114)
T PF01413_consen 1 GFVIAVHSQTDQIPDCPPG-----WTSLWTGYSFLMHTGGGEGHGQDLGSPGSCLERFRTMPFIECNGRGTCNYAS-NDY 74 (114)
T ss_dssp SEEEEEE-SSSS-----TT-----EEEEEEEEEEEEEEETTEEEE--TTSGGGEESSEESS-EEEEETTSEEEEST-T-E
T ss_pred CeEEEEecCCCCCCcCCCc-----hhHHhhhhhhhheecCCccccccccccchhHHHHhcCcceecCCCCcCCccC-CCc
Confidence 5899999999999999999 9999999999999999999999999999999999999999999999999987 999
Q ss_pred cccccCC-------CCCCCCCCCccccceecccceeecc
Q psy17777 1227 SYWLSTE-------EPMPMMPVESTQIQKFISRCVVCEV 1258 (1375)
Q Consensus 1227 sYWld~n-------~~~p~~p~~~d~i~~~isrC~vCe~ 1258 (1375)
||||++. .+++|+++.+++|+.|||||+||++
T Consensus 75 S~WLst~~~~~~f~~p~~~~~~~~~~~~~~ISRC~VC~~ 113 (114)
T PF01413_consen 75 SYWLSTIEPNDQFQMPMPMTPKAGEEIRSYISRCSVCEA 113 (114)
T ss_dssp EEEEB-S-CCCTTSS-TTEEEECCGCGGGGB-EEEEEEE
T ss_pred ceeEEccCchhhccCCCCccccCccccccccccceEeec
Confidence 9999999 4566789999999999999999986
No 7
>PF01410 COLFI: Fibrillar collagen C-terminal domain; InterPro: IPR000885 Collagens contain a large number of globular domains in between the regions of triple helical repeats IPR008160 from INTERPRO. These domains are involved in binding diverse substrates. One of these domains is found at the C terminus of fibrillar collagens. The exact function of this domain is unknown.; GO: 0005201 extracellular matrix structural constituent, 0005581 collagen
Probab=99.32 E-value=7.4e-13 Score=146.10 Aligned_cols=117 Identities=16% Similarity=0.366 Sum_probs=90.3
Q ss_pred cCCCCccccCCCCccccCCCCCccccccCCCCCCCCCCCccccceecccceeeccCcceeEeccCCcccCCC------C-
Q psy17777 1204 FSTMPFLFCDPNNVCNYASRNDRSYWLSTEEPMPMMPVESTQIQKFISRCVVCEVPANVIAVHSQSVVIPEC------P- 1276 (1375)
Q Consensus 1204 fstmPf~~C~~~~~C~~a~~~d~sYWld~n~~~p~~p~~~d~i~~~isrC~vCe~pt~~iavhs~~~~ip~C------p- 1276 (1375)
.+.+|+++|.++.+||. ...|++||||+|+.+ ..|+|++| |++.+..+||+.....++.- .
T Consensus 7 t~~~PArtC~dL~~~~p-~~~dG~YwIDPN~G~-----~~Dai~v~------C~~~~g~TCi~p~~~~~~~~~~~~~~~~ 74 (214)
T PF01410_consen 7 TKENPARTCRDLKLCHP-ELPDGEYWIDPNGGS-----PRDAIEVF------CNFTTGETCISPDQSSVPSKSWYKGSPK 74 (214)
T ss_pred CccchhhhHHHHHhhCc-ccCCCcEeECCCCCC-----CCCcEEEE------EeeeeceEEEeccCcccccccccCCCCc
Confidence 46789999999999998 799999999999998 88999999 88877888888876555422 1
Q ss_pred CCCccc-ee-eeEEEEEcCCCCCCCCCCCCCCCchhhhcccCCceeeecCCCceeeccCCcceeeeccC
Q psy17777 1277 VGWNSL-WI-GYSFVMHTAAGGDGGGQSLASPGSCLEEFRATPFIECNGEHGSCHYFANKLSFWLATID 1343 (1375)
Q Consensus 1277 ~gW~~~-w~-Gysf~~~t~~g~~g~gq~l~spgscl~~~r~~~~iec~~~~g~c~~~~~~~s~w~~~~~ 1343 (1375)
..|.++ .. ++.|.|.... .++.+ ..||+||+|||-+| +|+.+|+|.++..|.+...
T Consensus 75 ~~w~~~~~~g~~~f~Y~~~~---------~~~~~-~vQL~FLrLlS~~A-~Q~iTy~C~ns~~~~d~~~ 132 (214)
T PF01410_consen 75 GHWFSESMNGGFQFSYSISD---------GSPVG-VVQLNFLRLLSSEA-RQNITYHCKNSVAWYDQST 132 (214)
T ss_pred cceeeeecccccceeecCCc---------CCccc-hHHHHHHHHHhhhh-eeeEEEEcCCCcccccccc
Confidence 123322 22 4457775542 12222 78999999999999 9999999999988876543
No 8
>smart00038 COLFI Fibrillar collagens C-terminal domain. Found at C-termini of fibrillar collagens: Ephydatia muelleri procollagen EMF1alpha, vertebrate collagens alpha(1)III, alpha(1)II, alpha(2)V etc.
Probab=99.03 E-value=1.5e-10 Score=129.10 Aligned_cols=129 Identities=15% Similarity=0.294 Sum_probs=97.6
Q ss_pred cCcCcccccccccCCCCccccCCCCccccCCCCCccccccCCCCCCCCCCCccccceecccceeeccCcceeEeccCCcc
Q psy17777 1192 QDLGFAGSCVRKFSTMPFLFCDPNNVCNYASRNDRSYWLSTEEPMPMMPVESTQIQKFISRCVVCEVPANVIAVHSQSVV 1271 (1375)
Q Consensus 1192 qdlG~~gScl~~fstmPf~~C~~~~~C~~a~~~d~sYWld~n~~~p~~p~~~d~i~~~isrC~vCe~pt~~iavhs~~~~ 1271 (1375)
+.+.+..+.+ .....|+++|.++..||. ..+|++||||+|+.+ ..|+|++| |.+.+..+||+....+
T Consensus 13 ~~i~~i~~P~-Gt~~~PartC~dl~~~~p-~~~~G~YwIdPn~G~-----~~Da~~V~------C~~~~g~Tcv~p~~~~ 79 (232)
T smart00038 13 NQIEQLKSPT-GTKDNPARTCKDLKLCHP-EFKDGEYWVDPNQGC-----IRDAIKVF------CNFETGETCVSPSPSS 79 (232)
T ss_pred hHHhcccCCC-CCCCCCcchhhHHHhcCC-CCCCCCEEECCCCCC-----CCCCeEEE------EEeecCCeEEcccccc
Confidence 3344333333 345689999999999998 689999999999998 88999999 8888889999887666
Q ss_pred cCCCC-------CCCcccee--eeEEEEEcCCCCCCCCCCCCCCCchhhhcccCCceeeecCCCceeeccCCcceeeecc
Q psy17777 1272 IPECP-------VGWNSLWI--GYSFVMHTAAGGDGGGQSLASPGSCLEEFRATPFIECNGEHGSCHYFANKLSFWLATI 1342 (1375)
Q Consensus 1272 ip~Cp-------~gW~~~w~--Gysf~~~t~~g~~g~gq~l~spgscl~~~r~~~~iec~~~~g~c~~~~~~~s~w~~~~ 1342 (1375)
|+.-. ..|+++.. ++.|.|.... .++ .-..|++||+++|..| +++.+|+|.++..|++..
T Consensus 80 ~~~~~~~~~~~~~~wf~e~~~~g~~f~Y~~~~---------~~~-~~~vQltfLrllS~~A-~Q~iTy~C~ns~a~~d~~ 148 (232)
T smart00038 80 IPRKTWYSGKSKHVWFGETMNGGFQFSYGDSE---------GPP-VGVVQLTFLRLLSTSA-HQNITYHCKNSVAYMDEA 148 (232)
T ss_pred ccccccccCCCCCceeeeeccCCeEEEEecCC---------CCc-cchHHHHHHHHhcccc-eeEEEEEEEcceeEEecc
Confidence 66443 12334333 7788885432 123 2257999999999999 999999999999999876
Q ss_pred CC
Q psy17777 1343 DP 1344 (1375)
Q Consensus 1343 ~~ 1344 (1375)
..
T Consensus 149 ~~ 150 (232)
T smart00038 149 TG 150 (232)
T ss_pred cc
Confidence 53
No 9
>TIGR03032 conserved hypothetical protein TIGR03032. This protein family is uncharacterized. A number of motifs are conserved perfectly among all member sequences. The function of this protein is unknown.
Probab=24.12 E-value=1.6e+02 Score=35.10 Aligned_cols=28 Identities=21% Similarity=0.688 Sum_probs=17.6
Q ss_pred eeccCcceeEeccCCcccCCCCCCC-cccee
Q psy17777 1255 VCEVPANVIAVHSQSVVIPECPVGW-NSLWI 1284 (1375)
Q Consensus 1255 vCe~pt~~iavhs~~~~ip~Cp~gW-~~~w~ 1284 (1375)
|.++++++|.+ .-+..|++|.|. .++|.
T Consensus 189 vidv~s~evl~--~GLsmPhSPRWhdgrLwv 217 (335)
T TIGR03032 189 VIDIPSGEVVA--SGLSMPHSPRWYQGKLWL 217 (335)
T ss_pred EEEeCCCCEEE--cCccCCcCCcEeCCeEEE
Confidence 46677776665 345678999444 33666
No 10
>KOG0808|consensus
Probab=21.17 E-value=36 Score=38.69 Aligned_cols=32 Identities=19% Similarity=0.358 Sum_probs=22.5
Q ss_pred hhhcccCCceeeecCCCceeeccCCcceeeecc
Q psy17777 1310 LEEFRATPFIECNGEHGSCHYFANKLSFWLATI 1342 (1375)
Q Consensus 1310 l~~~r~~~~iec~~~~g~c~~~~~~~s~w~~~~ 1342 (1375)
|+.+|.--||.--. -+.|..|.+...|-|+..
T Consensus 332 lsr~rdgllia~ld-lnlcrq~kd~wgfrmt~r 363 (387)
T KOG0808|consen 332 LSRYRDGLLIADLD-LNLCRQYKDKWGFRMTAR 363 (387)
T ss_pred ccccccceEEeecc-hHHHHHhhhhhcceehhh
Confidence 45566666777666 678888888777777654
No 11
>KOG1446|consensus
Probab=21.14 E-value=1.1e+02 Score=36.19 Aligned_cols=34 Identities=21% Similarity=0.367 Sum_probs=22.3
Q ss_pred eeccCCCCCcCCCccc-cccCCcCCCceeeeEecccC
Q psy17777 1339 LATIDPQDQWSRPQQQ-TLKSGNLRQRISRCQVCIKR 1374 (1375)
Q Consensus 1339 ~~~~~~~~~~~~p~~~-t~~~~~~~~~~src~vc~~~ 1374 (1375)
|+.-.-+++|-+-.-+ |+.=..| ||++||+||+.
T Consensus 106 L~~sP~~d~FlS~S~D~tvrLWDl--R~~~cqg~l~~ 140 (311)
T KOG1446|consen 106 LSVSPKDDTFLSSSLDKTVRLWDL--RVKKCQGLLNL 140 (311)
T ss_pred EEecCCCCeEEecccCCeEEeeEe--cCCCCceEEec
Confidence 3334445667664444 6666654 59999999986
No 12
>PHA03112 IL-18 binding protein; Provisional
Probab=20.43 E-value=80 Score=33.26 Aligned_cols=20 Identities=40% Similarity=1.197 Sum_probs=16.8
Q ss_pred CceeeecCCCceeeccC-Ccceeee
Q psy17777 1317 PFIECNGEHGSCHYFAN-KLSFWLA 1340 (1375)
Q Consensus 1317 ~~iec~~~~g~c~~~~~-~~s~w~~ 1340 (1375)
=+|+|.| |.||.+ ++-|||.
T Consensus 41 vvL~C~g----cs~fp~fS~lYWl~ 61 (141)
T PHA03112 41 VVLECRG----CSYFPKFSYVYWLI 61 (141)
T ss_pred EEEEEEC----ccCCCCceEEEEEc
Confidence 3689999 888775 9999994
No 13
>KOG1388|consensus
Probab=20.13 E-value=1.2e+02 Score=34.19 Aligned_cols=50 Identities=20% Similarity=0.203 Sum_probs=34.2
Q ss_pred ceeeeEEEEEcCCCCCCCCCCCCCCCchhhhcccCCceeeecCCCceeeccCCcceeeeccCC
Q psy17777 1282 LWIGYSFVMHTAAGGDGGGQSLASPGSCLEEFRATPFIECNGEHGSCHYFANKLSFWLATIDP 1344 (1375)
Q Consensus 1282 ~w~Gysf~~~t~~g~~g~gq~l~spgscl~~~r~~~~iec~~~~g~c~~~~~~~s~w~~~~~~ 1344 (1375)
+-+-|.|+.|..+-+.+ -+...-||+|+....-+||+...++|-+..+.+
T Consensus 147 l~id~~ftf~l~~~d~~-------------fv~sd~~i~~~d~fd~~~~~~~g~~~ic~~~~~ 196 (217)
T KOG1388|consen 147 LLIDGQFTFHLLQEDDG-------------FVTSDNFISTHDMFDYLHLTNAGNTFICNPLWE 196 (217)
T ss_pred eecccccccceeecCCC-------------ceeeccccccCCcccchhhccCCCceecchHHH
Confidence 34467777776643322 267778899988668888888888887766543
Done!