NCBI Home Page NCBI Site Search page NCBI Guide that lists and describes the NCBI resources
Conserved domains on  [gi|303284100|ref|XP_003061341|]
View 

uncharacterized protein MICPUCDRAFT_60988 [Micromonas pusilla CCMP1545]

Protein Classification

WD repeat NOL10/ENP2 family protein( domain architecture ID 11456824)

WD repeat NOL10/ENP2 family protein contains WD40 repeats that fold into a beta-propeller structure and functions as a scaffold, such as Schizosaccharomyces pombe ribosome biogenesis protein enp2 homolog that may be involved in rRNA-processing and ribosome biosynthesis

CATH:  2.130.10.10
Gene Ontology:  GO:0005515
SCOP:  4002744

Graphical summary

 Zoom to residue level

show extra options »

Show site features     Horizontal zoom: ×

List of domain hits

Name Accession Description Interval E-value
WD40 COG2319
WD40 repeat [General function prediction only];
80-381 1.37e-14

WD40 repeat [General function prediction only];


:

Pssm-ID: 441893 [Multi-domain]  Cd Length: 403  Bit Score: 76.49  E-value: 1.37e-14
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100  80 SADGQFLVvTGLHPPQVKVYDLS--QLSMKFERHLDA-EVVDFAILGedySKLAFLCADRSIN-FHAKFGKYYSTRTPRQ 155
Cdd:COG2319   87 SPDGRLLA-SASADGTVRLWDLAtgLLLRTLTGHTGAvRSVAFSPDG---KTLASGSADGTVRlWDLATGKLLRTLTGHS 162
                         90       100       110       120       130       140       150       160
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 156 G--RDMAYvRHTGDLVVVGGAHEMYRL-NLERGRFLTPMEARGGSLNVIGSCPTTGLLAVGSEDGTVECFDPRTRAAIGE 232
Cdd:COG2319  163 GavTSVAF-SPDGKLLASGSDDGTVRLwDLATGKLLRTLTGHTGAVRSVAFSPDGKLLASGSADGTVRLWDLATGKLLRT 241
                        170       180       190       200       210       220       230       240
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 233 LnvTRANGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSRPVLTKDHYNGfPIKDLKYHTgvDEKRRVISADTRVVKIWD 312
Cdd:COG2319  242 L--TGHSGSVRSVAFSPDGRLLASGSADGTVRLWDLATGELLRTLTGHSG-GVNSVAFSP--DGKLLASGSDDGTVRLWD 316
                        250       260       270       280       290       300       310
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|..
gi 303284100 313 PsADGKArsvsrrvprcfqsRHTSTPFNSASDAFELHPD---IASygpstqpfaaiepGSEINDVCVWDRTG 381
Cdd:COG2319  317 L-ATGKL-------------LRTLTGHTGAVRSVAFSPDgktLAS-------------GSDDGTVRLWDLAT 361
NUC153 pfam08159
NUC153 domain; This small domain is found in a a novel nucleolar family.
550-578 4.98e-08

NUC153 domain; This small domain is found in a a novel nucleolar family.


:

Pssm-ID: 462385 [Multi-domain]  Cd Length: 29  Bit Score: 49.25  E-value: 4.98e-08
                          10        20
                  ....*....|....*....|....*....
gi 303284100  550 DDRFAAMFKDARFEVDEQSEAYKTLHPNA 578
Cdd:pfam08159   1 DPRFKALFEDHDFAIDPTSPEFKKTNPMK 29
 
Name Accession Description Interval E-value
WD40 COG2319
WD40 repeat [General function prediction only];
80-381 1.37e-14

WD40 repeat [General function prediction only];


Pssm-ID: 441893 [Multi-domain]  Cd Length: 403  Bit Score: 76.49  E-value: 1.37e-14
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100  80 SADGQFLVvTGLHPPQVKVYDLS--QLSMKFERHLDA-EVVDFAILGedySKLAFLCADRSIN-FHAKFGKYYSTRTPRQ 155
Cdd:COG2319   87 SPDGRLLA-SASADGTVRLWDLAtgLLLRTLTGHTGAvRSVAFSPDG---KTLASGSADGTVRlWDLATGKLLRTLTGHS 162
                         90       100       110       120       130       140       150       160
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 156 G--RDMAYvRHTGDLVVVGGAHEMYRL-NLERGRFLTPMEARGGSLNVIGSCPTTGLLAVGSEDGTVECFDPRTRAAIGE 232
Cdd:COG2319  163 GavTSVAF-SPDGKLLASGSDDGTVRLwDLATGKLLRTLTGHTGAVRSVAFSPDGKLLASGSADGTVRLWDLATGKLLRT 241
                        170       180       190       200       210       220       230       240
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 233 LnvTRANGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSRPVLTKDHYNGfPIKDLKYHTgvDEKRRVISADTRVVKIWD 312
Cdd:COG2319  242 L--TGHSGSVRSVAFSPDGRLLASGSADGTVRLWDLATGELLRTLTGHSG-GVNSVAFSP--DGKLLASGSDDGTVRLWD 316
                        250       260       270       280       290       300       310
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|..
gi 303284100 313 PsADGKArsvsrrvprcfqsRHTSTPFNSASDAFELHPD---IASygpstqpfaaiepGSEINDVCVWDRTG 381
Cdd:COG2319  317 L-ATGKL-------------LRTLTGHTGAVRSVAFSPDgktLAS-------------GSDDGTVRLWDLAT 361
WD40 cd00200
WD40 domain, found in a number of eukaryotic proteins that cover a wide variety of functions ...
166-316 1.74e-11

WD40 domain, found in a number of eukaryotic proteins that cover a wide variety of functions including adaptor/regulatory modules in signal transduction, pre-mRNA processing and cytoskeleton assembly; typically contains a GH dipeptide 11-24 residues from its N-terminus and the WD dipeptide at its C-terminus and is 40 residues long, hence the name WD40; between GH and WD lies a conserved core; serves as a stable propeller-like platform to which proteins can bind either stably or reversibly; forms a propeller-like structure with several blades where each blade is composed of a four-stranded anti-parallel b-sheet; instances with few detectable copies are hypothesized to form larger structures by dimerization; each WD40 sequence repeat forms the first three strands of one blade and the last strand in the next blade; the last C-terminal WD40 repeat completes the blade structure of the first WD40 repeat to create the closed ring propeller-structure; residues on the top and bottom surface of the propeller are proposed to coordinate interactions with other proteins and/or small ligands; 7 copies of the repeat are present in this alignment.


Pssm-ID: 238121 [Multi-domain]  Cd Length: 289  Bit Score: 65.82  E-value: 1.74e-11
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 166 GDLVVVGGAHEMYRL-NLERGRFLTPMEARGGSLNVIGSCPTTGLLAVGSEDGTVECFDPRTRAAIGEL----------- 233
Cdd:cd00200   21 GKLLATGSGDGTIKVwDLETGELLRTLKGHTGPVRDVAASADGTYLASGSSDKTIRLWDLETGECVRTLtghtsyvssva 100
                         90       100       110       120       130       140       150       160
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 234 --------------------NVTRA---------NGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSRPVLT-KDHYNgf 283
Cdd:cd00200  101 fspdgrilssssrdktikvwDVETGkclttlrghTDWVNSVAFSPDGTFVASSSQDGTIKLWDLRTGKCVATlTGHTG-- 178
                        170       180       190
                 ....*....|....*....|....*....|...
gi 303284100 284 PIKDLKYHTgvDEKRRVISADTRVVKIWDPSAD 316
Cdd:cd00200  179 EVNSVAFSP--DGEKLLSSSSDGTIKLWDLSTG 209
NUC153 pfam08159
NUC153 domain; This small domain is found in a a novel nucleolar family.
550-578 4.98e-08

NUC153 domain; This small domain is found in a a novel nucleolar family.


Pssm-ID: 462385 [Multi-domain]  Cd Length: 29  Bit Score: 49.25  E-value: 4.98e-08
                          10        20
                  ....*....|....*....|....*....
gi 303284100  550 DDRFAAMFKDARFEVDEQSEAYKTLHPNA 578
Cdd:pfam08159   1 DPRFKALFEDHDFAIDPTSPEFKKTNPMK 29
PLN00181 PLN00181
protein SPA1-RELATED; Provisional
181-379 1.30e-07

protein SPA1-RELATED; Provisional


Pssm-ID: 177776 [Multi-domain]  Cd Length: 793  Bit Score: 55.09  E-value: 1.30e-07
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 181 NLERGRFLTPM---EARGGSLNVIGSCPTtgLLAVGSEDGTVECFDPRTRAAIGELNvTRANggVTAVRF-DPNGMSVAA 256
Cdd:PLN00181 561 DVARSQLVTEMkehEKRVWSIDYSSADPT--LLASGSDDGSVKLWSINQGVSIGTIK-TKAN--ICCVQFpSESGRSLAF 635
                         90       100       110       120       130       140       150       160
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 257 GDGEGLIRVYDLRSSR-PVLTKDHYNgfpiKDLKYHTGVDEKRRVISADTRVVKIWDPSADGkarsvsrrvprcfqSRHT 335
Cdd:PLN00181 636 GSADHKVYYYDLRNPKlPLCTMIGHS----KTVSYVRFVDSSTLVSSSTDNTLKLWDLSMSI--------------SGIN 697
                        170       180       190       200
                 ....*....|....*....|....*....|....*....|....
gi 303284100 336 STPFNSasdaFELHPDIASYGPSTQPFAAIEPGSEINDVCVWDR 379
Cdd:PLN00181 698 ETPLHS----FMGHTNVKNFVGLSVSDGYIATGSETNEVFVYHK 737
ANAPC4_WD40 pfam12894
Anaphase-promoting complex subunit 4 WD40 domain; Apc4 contains an N-terminal propeller-shaped ...
204-274 3.44e-04

Anaphase-promoting complex subunit 4 WD40 domain; Apc4 contains an N-terminal propeller-shaped WD40 domain.The N-terminus of Afi1 serves to stabilize the union between Apc4 and Apc5, both of which lie towards the bottom-front of the APC,


Pssm-ID: 403945 [Multi-domain]  Cd Length: 91  Bit Score: 40.34  E-value: 3.44e-04
                          10        20        30        40        50        60        70
                  ....*....|....*....|....*....|....*....|....*....|....*....|....*....|..
gi 303284100  204 CPTTGLLAVGSEDGTVECFdpRT-RAAIGELNVTRANGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSRPV 274
Cdd:pfam12894   4 CPTMDLIALATEDGELLLH--RLnWQRVWTLSPDKEDLEVTSLAWRPDGKLLAVGYSDGTVRLLDAENGKIV 73
 
Name Accession Description Interval E-value
WD40 COG2319
WD40 repeat [General function prediction only];
80-381 1.37e-14

WD40 repeat [General function prediction only];


Pssm-ID: 441893 [Multi-domain]  Cd Length: 403  Bit Score: 76.49  E-value: 1.37e-14
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100  80 SADGQFLVvTGLHPPQVKVYDLS--QLSMKFERHLDA-EVVDFAILGedySKLAFLCADRSIN-FHAKFGKYYSTRTPRQ 155
Cdd:COG2319   87 SPDGRLLA-SASADGTVRLWDLAtgLLLRTLTGHTGAvRSVAFSPDG---KTLASGSADGTVRlWDLATGKLLRTLTGHS 162
                         90       100       110       120       130       140       150       160
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 156 G--RDMAYvRHTGDLVVVGGAHEMYRL-NLERGRFLTPMEARGGSLNVIGSCPTTGLLAVGSEDGTVECFDPRTRAAIGE 232
Cdd:COG2319  163 GavTSVAF-SPDGKLLASGSDDGTVRLwDLATGKLLRTLTGHTGAVRSVAFSPDGKLLASGSADGTVRLWDLATGKLLRT 241
                        170       180       190       200       210       220       230       240
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 233 LnvTRANGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSRPVLTKDHYNGfPIKDLKYHTgvDEKRRVISADTRVVKIWD 312
Cdd:COG2319  242 L--TGHSGSVRSVAFSPDGRLLASGSADGTVRLWDLATGELLRTLTGHSG-GVNSVAFSP--DGKLLASGSDDGTVRLWD 316
                        250       260       270       280       290       300       310
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|..
gi 303284100 313 PsADGKArsvsrrvprcfqsRHTSTPFNSASDAFELHPD---IASygpstqpfaaiepGSEINDVCVWDRTG 381
Cdd:COG2319  317 L-ATGKL-------------LRTLTGHTGAVRSVAFSPDgktLAS-------------GSDDGTVRLWDLAT 361
WD40 COG2319
WD40 repeat [General function prediction only];
80-314 4.37e-14

WD40 repeat [General function prediction only];


Pssm-ID: 441893 [Multi-domain]  Cd Length: 403  Bit Score: 74.95  E-value: 4.37e-14
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100  80 SADGQFLVvTGLHPPQVKVYDLS--QLSMKFERHlDAEVVDFAilgedYS----KLAFLCADRSInfhakfgKYYSTRTP 153
Cdd:COG2319  171 SPDGKLLA-SGSDDGTVRLWDLAtgKLLRTLTGH-TGAVRSVA-----FSpdgkLLASGSADGTV-------RLWDLATG 236
                         90       100       110       120       130       140       150       160
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 154 RQGRDM----AYVR-----HTGDLVVVGGAHEMYRL-NLERGRFLTPMEARGGSLNVIGSCPTTGLLAVGSEDGTVECFD 223
Cdd:COG2319  237 KLLRTLtghsGSVRsvafsPDGRLLASGSADGTVRLwDLATGELLRTLTGHSGGVNSVAFSPDGKLLASGSDDGTVRLWD 316
                        170       180       190       200       210       220       230       240
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 224 PRTRAAIGELnvTRANGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSRPVLT-KDHYNgfPIKDLKYHTgvDEKRRVIS 302
Cdd:COG2319  317 LATGKLLRTL--TGHTGAVRSVAFSPDGKTLASGSDDGTVRLWDLATGELLRTlTGHTG--AVTSVAFSP--DGRTLASG 390
                        250
                 ....*....|..
gi 303284100 303 ADTRVVKIWDPS 314
Cdd:COG2319  391 SADGTVRLWDLA 402
WD40 cd00200
WD40 domain, found in a number of eukaryotic proteins that cover a wide variety of functions ...
166-316 1.74e-11

WD40 domain, found in a number of eukaryotic proteins that cover a wide variety of functions including adaptor/regulatory modules in signal transduction, pre-mRNA processing and cytoskeleton assembly; typically contains a GH dipeptide 11-24 residues from its N-terminus and the WD dipeptide at its C-terminus and is 40 residues long, hence the name WD40; between GH and WD lies a conserved core; serves as a stable propeller-like platform to which proteins can bind either stably or reversibly; forms a propeller-like structure with several blades where each blade is composed of a four-stranded anti-parallel b-sheet; instances with few detectable copies are hypothesized to form larger structures by dimerization; each WD40 sequence repeat forms the first three strands of one blade and the last strand in the next blade; the last C-terminal WD40 repeat completes the blade structure of the first WD40 repeat to create the closed ring propeller-structure; residues on the top and bottom surface of the propeller are proposed to coordinate interactions with other proteins and/or small ligands; 7 copies of the repeat are present in this alignment.


Pssm-ID: 238121 [Multi-domain]  Cd Length: 289  Bit Score: 65.82  E-value: 1.74e-11
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 166 GDLVVVGGAHEMYRL-NLERGRFLTPMEARGGSLNVIGSCPTTGLLAVGSEDGTVECFDPRTRAAIGEL----------- 233
Cdd:cd00200   21 GKLLATGSGDGTIKVwDLETGELLRTLKGHTGPVRDVAASADGTYLASGSSDKTIRLWDLETGECVRTLtghtsyvssva 100
                         90       100       110       120       130       140       150       160
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 234 --------------------NVTRA---------NGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSRPVLT-KDHYNgf 283
Cdd:cd00200  101 fspdgrilssssrdktikvwDVETGkclttlrghTDWVNSVAFSPDGTFVASSSQDGTIKLWDLRTGKCVATlTGHTG-- 178
                        170       180       190
                 ....*....|....*....|....*....|...
gi 303284100 284 PIKDLKYHTgvDEKRRVISADTRVVKIWDPSAD 316
Cdd:cd00200  179 EVNSVAFSP--DGEKLLSSSSDGTIKLWDLSTG 209
WD40 COG2319
WD40 repeat [General function prediction only];
80-270 1.88e-11

WD40 repeat [General function prediction only];


Pssm-ID: 441893 [Multi-domain]  Cd Length: 403  Bit Score: 66.86  E-value: 1.88e-11
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100  80 SADGQFLVvTGLHPPQVKVYDLS--QLSMKFERHlDAEVVDFAILGeDYSKLAFLCADRSIN-FHAKFGKYYSTRTPRQG 156
Cdd:COG2319  213 SPDGKLLA-SGSADGTVRLWDLAtgKLLRTLTGH-SGSVRSVAFSP-DGRLLASGSADGTVRlWDLATGELLRTLTGHSG 289
                         90       100       110       120       130       140       150       160
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 157 R--DMAYvRHTGDLVVVGGA-HEMYRLNLERGRFLTPMEARGGSLNVIGSCPTTGLLAVGSEDGTVECFDPRTRAAIGEL 233
Cdd:COG2319  290 GvnSVAF-SPDGKLLASGSDdGTVRLWDLATGKLLRTLTGHTGAVRSVAFSPDGKTLASGSDDGTVRLWDLATGELLRTL 368
                        170       180       190
                 ....*....|....*....|....*....|....*..
gi 303284100 234 nvTRANGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRS 270
Cdd:COG2319  369 --TGHTGAVTSVAFSPDGRTLASGSADGTVRLWDLAT 403
WD40 cd00200
WD40 domain, found in a number of eukaryotic proteins that cover a wide variety of functions ...
75-312 2.86e-11

WD40 domain, found in a number of eukaryotic proteins that cover a wide variety of functions including adaptor/regulatory modules in signal transduction, pre-mRNA processing and cytoskeleton assembly; typically contains a GH dipeptide 11-24 residues from its N-terminus and the WD dipeptide at its C-terminus and is 40 residues long, hence the name WD40; between GH and WD lies a conserved core; serves as a stable propeller-like platform to which proteins can bind either stably or reversibly; forms a propeller-like structure with several blades where each blade is composed of a four-stranded anti-parallel b-sheet; instances with few detectable copies are hypothesized to form larger structures by dimerization; each WD40 sequence repeat forms the first three strands of one blade and the last strand in the next blade; the last C-terminal WD40 repeat completes the blade structure of the first WD40 repeat to create the closed ring propeller-structure; residues on the top and bottom surface of the propeller are proposed to coordinate interactions with other proteins and/or small ligands; 7 copies of the repeat are present in this alignment.


Pssm-ID: 238121 [Multi-domain]  Cd Length: 289  Bit Score: 65.05  E-value: 2.86e-11
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100  75 SRCKMSADGQFLVVTGLhPPQVKVYDLS--QLSMKFERHLDaEVVDFAILgeDYSKLAFLC-ADRSINFhakfgkyYSTR 151
Cdd:cd00200   55 RDVAASADGTYLASGSS-DKTIRLWDLEtgECVRTLTGHTS-YVSSVAFS--PDGRILSSSsRDKTIKV-------WDVE 123
                         90       100       110       120       130       140       150       160
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 152 TPRQGRDM----AYVR----HTGDLVVVGGAHEMY-RL-NLERGRFLTPMEARGGSLNVIGSCPTTGLLAVGSEDGTVEC 221
Cdd:cd00200  124 TGKCLTTLrghtDWVNsvafSPDGTFVASSSQDGTiKLwDLRTGKCVATLTGHTGEVNSVAFSPDGEKLLSSSSDGTIKL 203
                        170       180       190       200       210       220       230       240
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 222 FDPRTRAAIGELnvTRANGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSRPVLTKDHYNGfPIKDLKYHtgvDEKRRVI 301
Cdd:cd00200  204 WDLSTGKCLGTL--RGHENGVNSVAFSPDGYLLASGSEDGTIRVWDLRTGECVQTLSGHTN-SVTSLAWS---PDGKRLA 277
                        250
                 ....*....|...
gi 303284100 302 SA--DTRvVKIWD 312
Cdd:cd00200  278 SGsaDGT-IRIWD 289
NUC153 pfam08159
NUC153 domain; This small domain is found in a a novel nucleolar family.
550-578 4.98e-08

NUC153 domain; This small domain is found in a a novel nucleolar family.


Pssm-ID: 462385 [Multi-domain]  Cd Length: 29  Bit Score: 49.25  E-value: 4.98e-08
                          10        20
                  ....*....|....*....|....*....
gi 303284100  550 DDRFAAMFKDARFEVDEQSEAYKTLHPNA 578
Cdd:pfam08159   1 DPRFKALFEDHDFAIDPTSPEFKKTNPMK 29
PLN00181 PLN00181
protein SPA1-RELATED; Provisional
181-379 1.30e-07

protein SPA1-RELATED; Provisional


Pssm-ID: 177776 [Multi-domain]  Cd Length: 793  Bit Score: 55.09  E-value: 1.30e-07
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 181 NLERGRFLTPM---EARGGSLNVIGSCPTtgLLAVGSEDGTVECFDPRTRAAIGELNvTRANggVTAVRF-DPNGMSVAA 256
Cdd:PLN00181 561 DVARSQLVTEMkehEKRVWSIDYSSADPT--LLASGSDDGSVKLWSINQGVSIGTIK-TKAN--ICCVQFpSESGRSLAF 635
                         90       100       110       120       130       140       150       160
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 257 GDGEGLIRVYDLRSSR-PVLTKDHYNgfpiKDLKYHTGVDEKRRVISADTRVVKIWDPSADGkarsvsrrvprcfqSRHT 335
Cdd:PLN00181 636 GSADHKVYYYDLRNPKlPLCTMIGHS----KTVSYVRFVDSSTLVSSSTDNTLKLWDLSMSI--------------SGIN 697
                        170       180       190       200
                 ....*....|....*....|....*....|....*....|....
gi 303284100 336 STPFNSasdaFELHPDIASYGPSTQPFAAIEPGSEINDVCVWDR 379
Cdd:PLN00181 698 ETPLHS----FMGHTNVKNFVGLSVSDGYIATGSETNEVFVYHK 737
WD40 cd00200
WD40 domain, found in a number of eukaryotic proteins that cover a wide variety of functions ...
196-312 5.58e-05

WD40 domain, found in a number of eukaryotic proteins that cover a wide variety of functions including adaptor/regulatory modules in signal transduction, pre-mRNA processing and cytoskeleton assembly; typically contains a GH dipeptide 11-24 residues from its N-terminus and the WD dipeptide at its C-terminus and is 40 residues long, hence the name WD40; between GH and WD lies a conserved core; serves as a stable propeller-like platform to which proteins can bind either stably or reversibly; forms a propeller-like structure with several blades where each blade is composed of a four-stranded anti-parallel b-sheet; instances with few detectable copies are hypothesized to form larger structures by dimerization; each WD40 sequence repeat forms the first three strands of one blade and the last strand in the next blade; the last C-terminal WD40 repeat completes the blade structure of the first WD40 repeat to create the closed ring propeller-structure; residues on the top and bottom surface of the propeller are proposed to coordinate interactions with other proteins and/or small ligands; 7 copies of the repeat are present in this alignment.


Pssm-ID: 238121 [Multi-domain]  Cd Length: 289  Bit Score: 45.79  E-value: 5.58e-05
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 196 GSLNVIGSC---PTTGLLAVGSEDGTVECFDPRTRAAIGELNVtrANGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSR 272
Cdd:cd00200    7 GHTGGVTCVafsPDGKLLATGSGDGTIKVWDLETGELLRTLKG--HTGPVRDVAASADGTYLASGSSDKTIRLWDLETGE 84
                         90       100       110       120
                 ....*....|....*....|....*....|....*....|....
gi 303284100 273 PVLTkdhYNGF--PIKDLKYHTgvdeKRRVISADTR--VVKIWD 312
Cdd:cd00200   85 CVRT---LTGHtsYVSSVAFSP----DGRILSSSSRdkTIKVWD 121
WD40 cd00200
WD40 domain, found in a number of eukaryotic proteins that cover a wide variety of functions ...
239-321 1.38e-04

WD40 domain, found in a number of eukaryotic proteins that cover a wide variety of functions including adaptor/regulatory modules in signal transduction, pre-mRNA processing and cytoskeleton assembly; typically contains a GH dipeptide 11-24 residues from its N-terminus and the WD dipeptide at its C-terminus and is 40 residues long, hence the name WD40; between GH and WD lies a conserved core; serves as a stable propeller-like platform to which proteins can bind either stably or reversibly; forms a propeller-like structure with several blades where each blade is composed of a four-stranded anti-parallel b-sheet; instances with few detectable copies are hypothesized to form larger structures by dimerization; each WD40 sequence repeat forms the first three strands of one blade and the last strand in the next blade; the last C-terminal WD40 repeat completes the blade structure of the first WD40 repeat to create the closed ring propeller-structure; residues on the top and bottom surface of the propeller are proposed to coordinate interactions with other proteins and/or small ligands; 7 copies of the repeat are present in this alignment.


Pssm-ID: 238121 [Multi-domain]  Cd Length: 289  Bit Score: 44.63  E-value: 1.38e-04
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 239 NGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSRPVLT-KDHYngFPIKDLKYHtgVDEKRRVISADTRVVKIWDPSADG 317
Cdd:cd00200    9 TGGVTCVAFSPDGKLLATGSGDGTIKVWDLETGELLRTlKGHT--GPVRDVAAS--ADGTYLASGSSDKTIRLWDLETGE 84

                 ....
gi 303284100 318 KARS 321
Cdd:cd00200   85 CVRT 88
ANAPC4_WD40 pfam12894
Anaphase-promoting complex subunit 4 WD40 domain; Apc4 contains an N-terminal propeller-shaped ...
204-274 3.44e-04

Anaphase-promoting complex subunit 4 WD40 domain; Apc4 contains an N-terminal propeller-shaped WD40 domain.The N-terminus of Afi1 serves to stabilize the union between Apc4 and Apc5, both of which lie towards the bottom-front of the APC,


Pssm-ID: 403945 [Multi-domain]  Cd Length: 91  Bit Score: 40.34  E-value: 3.44e-04
                          10        20        30        40        50        60        70
                  ....*....|....*....|....*....|....*....|....*....|....*....|....*....|..
gi 303284100  204 CPTTGLLAVGSEDGTVECFdpRT-RAAIGELNVTRANGGVTAVRFDPNGMSVAAGDGEGLIRVYDLRSSRPV 274
Cdd:pfam12894   4 CPTMDLIALATEDGELLLH--RLnWQRVWTLSPDKEDLEVTSLAWRPDGKLLAVGYSDGTVRLLDAENGKIV 73
WD40 COG2319
WD40 repeat [General function prediction only];
179-318 6.75e-03

WD40 repeat [General function prediction only];


Pssm-ID: 441893 [Multi-domain]  Cd Length: 403  Bit Score: 39.51  E-value: 6.75e-03
                         10        20        30        40        50        60        70        80
                 ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 179 RLNLERGRFLTPMEARGGSLNVIGSCPTTGLLAVGSEDGTVECFDPRTRAAIGELnvTRANGGVTAVRFDPNGMSVAAGD 258
Cdd:COG2319   20 LLAAALGALLLLLLGLAAAVASLAASPDGARLAAGAGDLTLLLLDAAAGALLATL--LGHTAAVLSVAFSPDGRLLASAS 97
                         90       100       110       120       130       140
                 ....*....|....*....|....*....|....*....|....*....|....*....|
gi 303284100 259 GEGLIRVYDLRSSRPVLTKDHYNGfPIKDLKYHTgvDEKRRVISADTRVVKIWDPsADGK 318
Cdd:COG2319   98 ADGTVRLWDLATGLLLRTLTGHTG-AVRSVAFSP--DGKTLASGSADGTVRLWDL-ATGK 153
 
Blast search parameters
Data Source: Precalculated data, version = cdd.v.3.21
Preset Options:Database: CDSEARCH/cdd   Low complexity filter: no  Composition Based Adjustment: yes   E-value threshold: 0.01

References:

  • Wang J et al. (2023), "The conserved domain database in 2023", Nucleic Acids Res.51(D)384-8.
  • Lu S et al. (2020), "The conserved domain database in 2020", Nucleic Acids Res.48(D)265-8.
  • Marchler-Bauer A et al. (2017), "CDD/SPARCLE: functional classification of proteins via subfamily domain architectures.", Nucleic Acids Res.45(D)200-3.
Help | Disclaimer | Write to the Help Desk
NCBI | NLM | NIH