NCBI Home Page NCBI Site Search page NCBI Guide that lists and describes the NCBI resources
Conserved domains on  [gi|1057838381|dbj|BAV54118|]
View 

dCas9-5xPlat2AflD [Cloning vector pCAG-dCas9-5xPlat2AflD]

Protein Classification

type II CRISPR RNA-guided endonuclease Cas9( domain architecture ID 11493220)

type II CRISPR RNA-guided endonuclease Cas9 is an RNA-guided endonuclease that cleaves double-stranded DNA

Graphical summary

 Zoom to residue level

show extra options »

Show site features     Horizontal zoom: ×

List of domain hits

Name Accession Description Interval E-value
cas_Csn1 TIGR01865
CRISPR subtype II/NMENI RNA-guided endonuclease Cas9/Csn1; CRISPR loci appear to be mobile ...
30-1076 0e+00

CRISPR subtype II/NMENI RNA-guided endonuclease Cas9/Csn1; CRISPR loci appear to be mobile elements with a wide host range. This model represents a protein found only in CRISPR-containing species, near other CRISPR-associated proteins (cas), as part of the NMENI subtype of CRISPR/Cas locus. The species range so far for this protein is animal pathogens and commensals only.


:

Pssm-ID: 273840  Cd Length: 805  Bit Score: 852.10  E-value: 0e+00
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381   30 KYSIGLAIGTNSVGWAVITDEYKVPSKKFKVLGNtdrhsiKKNLIGALLFDSGETAE-ATRLKRTARRRYTRRKNRICYL 108
Cdd:TIGR01865    1 EYILGLDIGIASVGWAIVEDDYKVPAAKRLIDGG------VRNFTGAELPKTGETAAlDRRLARGARRRIRRRKHRLLRL 74
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  109 QEIFSNEMAKVDDSFFHRLEESFLVEEDKKHerhpifgnivdevayhekypTIYHLRKKLVDSTDKADLrlIYLALAHMI 188
Cdd:TIGR01865   75 QELFSREGSLTDFDFFSRLENSFLVEEDKRN--------------------TIYHLRKAALENKLKPDE--LYLALLHII 132
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  189 KFRGHFLIEgdlnpdnsdvdklfiqlvqtynqlfeenpinasgvdakailsarlsksrrlenliaqlpgekknglfgnli 268
Cdd:TIGR01865  133 KHRGHFLIE----------------------------------------------------------------------- 141
                          250       260       270       280       290       300       310       320
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  269 alslgltpnfksnfdlaedaklqlskdtyDDDLDnllaqigdqyadlflaaknlsdaillsdilrVNTEITKAPLSASMI 348
Cdd:TIGR01865  142 -----------------------------GNDFD-------------------------------TANKETGALLSAVMI 161
                          330       340       350       360       370       380       390       400
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  349 KRYDEHHQDLTLLKALVRQQLPEKYKEIFFDqskngyagyidggasqeefykfikpilekmdgteellvklnreDLLRKQ 428
Cdd:TIGR01865  162 NRYLEHEADLRTLKELILKKFPKKYKEIFSE-------------------------------------------TFLRNQ 198
                          410       420       430       440       450       460       470       480
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  429 RTFDNGSIPHQIHLGELHAILRRQEDFYPFlkdnrekiEKILTFRIPYYVGPLARGNSRFAwmtrkseetitpwnfeeVV 508
Cdd:TIGR01865  199 RGFYNGSIPRQLLLEELEAIFRKQREYYPF--------IKLLTFRIPYYIGPLAEGKSEFA-----------------FV 253
                          490       500       510       520       530       540       550       560
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  509 DKGASAQSFIERMTNFDKNLPNEKVLPKHSLLYEYFTVYNELTKVKYVTEGMRKPAFLSGEQKKAIVDLLFKTNRKVTVK 588
Cdd:TIGR01865  254 DKPASAENFIEKMTGKCTYLPEEKRAPKHSLLAEKFTVLNELNNVRIIILEQGETKILSKEEKQELLDLLFKKKKLTYKK 333
                          570       580       590       600       610       620       630       640
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  589 QLKEDYFKKIECFDSVEISGVEDR---FNASLGTYHDLLKIIKDKDFLDNEENEDILEDIVLTLTLFEDREMIEERLKTY 665
Cdd:TIGR01865  334 LRKLLGLSEDAIFKGLRYEGLDNAekaFNISLKTYHKLRKALGDKDLLDNPKNPKDLDEIVKILTLYKDREMIKKRLELY 413
                          650       660       670       680       690       700       710       720
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  666 AHLFDDKVMKQLKRRRYTGWGRLSRKLINGIRDKQSGKTILDFLKSDGFANRNFMQLIHDDSLTFKEDIQKAqvsgqgds 745
Cdd:TIGR01865  414 KDVLNEEQVKKLVRLHFTGWGRLSLKALRGIRPLMEQGKRYDEAILELGGNRNFMQNINDSQLLPKINITKA-------- 485
                          730       740       750       760       770       780       790       800
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  746 lhehiANLAGSPAIKKGILQTVKVVDELVKVMGrhKPENIVIEMARENQTTQKGQKNSRERMKRIEEGIKELGS----QI 821
Cdd:TIGR01865  486 -----KDEILNPVVKRALLQARKVVNELVKKYG--PPDKIVIEMAREEQGTNFGKRNSKERYKKNEDKIKEFASalgkEI 558
                          810       820       830       840       850       860       870       880
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  822 LKEHPVENTQLQNEKLYLYYLQNGRDMYVDQELDINRL---SDYDVDAIVPQSFLKDDSIDNKVLTRSDKNRGKSDNVPS 898
Cdd:TIGR01865  559 LKEEPTENSSKNILKLRLYYQQNGKCMYTGKEIDIDDLfdlSYYEIDHILPQSRSFDDSISNKVLVLASENQEKGDQTPY 638
                          890       900       910       920       930       940       950       960
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  899 E-EVVKKMKNYWRQLLNAKLITQRKFDNLTKAERGGLSELDKAGFIKRQLVETRQITKHVAQILDSRMNTKYDendklIR 977
Cdd:TIGR01865  639 EaEIVKKDSAFWNKFEAYVLISKRKSDKLTRAERGGLSDDDKAGFIDRNLNDTRYITRVVANYLKDRFNFHLK-----KR 713
                          970       980       990      1000      1010      1020      1030      1040
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  978 EVKVITLKSKLVSDFRKDFQFYKVREINNYHHAHDAYLNAVVGTALIKKYPKLESEFVYGDYKVYDVRKMIAkseqeigK 1057
Cdd:TIGR01865  714 KVKVVTLKGQLTSQLRKKWGLYKKREINNYHHAHDAYINAVSTNALVKKFSQLEPEFRYKEYHNFDGRKKKK-------S 786
                         1050
                   ....*....|....*....
gi 1057838381 1058 ATAKYFFYSNIMNFFKTEI 1076
Cdd:TIGR01865  787 ATDKKVKFSNPMEFFKQKV 805
Cas9_PI super family cl24973
PAM-interacting domain of CRISPR-associated endonuclease Cas9; Cas9_PI is a family found at ...
1128-1384 1.46e-43

PAM-interacting domain of CRISPR-associated endonuclease Cas9; Cas9_PI is a family found at the C-terminal of bacterial type II CRISPR system Cas9 endonuclease. This domain adopts a novel protein fold that is unique to the Cas9 family. It is positioned in the structure-DNA-complex to recognize the PAM sequence on the non-complementary DNA strand of the crRNA. PAM sequence is protospacer-adjacent motifs on DNA. See family CRISPR-DR2, Rfam:RF01315. Cas9 carries two nuclease domains, HNH and RuvC, which cleave the DNA strands that are complementary and non-complementary to the 20 nucleotide guide sequence in crRNAs, respectively.


The actual alignment was detected with superfamily member pfam16595:

Pssm-ID: 435449  Cd Length: 264  Bit Score: 159.79  E-value: 1.46e-43
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381 1128 TGGFSKESILP--KRNSDKLIARKKD---WDPKKYGGFDSPTVAYSVLVVAKVEKGKSKKLksvkeLLGITIMERSSFEK 1202
Cdd:pfam16595    1 KGGLFNQTILPahKKKGKGLIPLKKDergLDVEKYGGYSSLTAAYFSLVEYTGKKGKRKRT-----IEGVPLYLAAKIEE 75
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381 1203 NPI--DFLEAKGYKEVKKDLIIKLPKYSLFElENGRKRMLASAGE---LQKGNELALPSKYVNFLYLASHYEKLKGSPED 1277
Cdd:pfam16595   76 NKDllEYLEEKLGLKEPKIILPKIKKNSLIK-IDGFRMLLTGKTEnrlLKNAVQLVLSNDDEKYIKKIEKFVKKNKDDII 154
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381 1278 NEQKQLFVEQHKHYLDEIIEQISEFSKrVILADANLDKVLSAYNKHRDKPIREQAENIIHLFTLTNLGA-PAAFKYFDTT 1356
Cdd:pfam16595  155 EEKDGLTEEKNIKLYDELLDKMKNTIY-YKRPSNQGEKLEKLKEKFIKLSLEEKCKVLIEILKLTHANPtSADLKLIGGS 233
                          250       260       270
                   ....*....|....*....|....*....|.
gi 1057838381 1357 IDRKRYTSTKEVLDA---TLIHQSITGLYET 1384
Cdd:pfam16595  234 KHAGRIKISNNISKAsniKLINQSVTGLYEK 264
 
Name Accession Description Interval E-value
cas_Csn1 TIGR01865
CRISPR subtype II/NMENI RNA-guided endonuclease Cas9/Csn1; CRISPR loci appear to be mobile ...
30-1076 0e+00

CRISPR subtype II/NMENI RNA-guided endonuclease Cas9/Csn1; CRISPR loci appear to be mobile elements with a wide host range. This model represents a protein found only in CRISPR-containing species, near other CRISPR-associated proteins (cas), as part of the NMENI subtype of CRISPR/Cas locus. The species range so far for this protein is animal pathogens and commensals only.


Pssm-ID: 273840  Cd Length: 805  Bit Score: 852.10  E-value: 0e+00
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381   30 KYSIGLAIGTNSVGWAVITDEYKVPSKKFKVLGNtdrhsiKKNLIGALLFDSGETAE-ATRLKRTARRRYTRRKNRICYL 108
Cdd:TIGR01865    1 EYILGLDIGIASVGWAIVEDDYKVPAAKRLIDGG------VRNFTGAELPKTGETAAlDRRLARGARRRIRRRKHRLLRL 74
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  109 QEIFSNEMAKVDDSFFHRLEESFLVEEDKKHerhpifgnivdevayhekypTIYHLRKKLVDSTDKADLrlIYLALAHMI 188
Cdd:TIGR01865   75 QELFSREGSLTDFDFFSRLENSFLVEEDKRN--------------------TIYHLRKAALENKLKPDE--LYLALLHII 132
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  189 KFRGHFLIEgdlnpdnsdvdklfiqlvqtynqlfeenpinasgvdakailsarlsksrrlenliaqlpgekknglfgnli 268
Cdd:TIGR01865  133 KHRGHFLIE----------------------------------------------------------------------- 141
                          250       260       270       280       290       300       310       320
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  269 alslgltpnfksnfdlaedaklqlskdtyDDDLDnllaqigdqyadlflaaknlsdaillsdilrVNTEITKAPLSASMI 348
Cdd:TIGR01865  142 -----------------------------GNDFD-------------------------------TANKETGALLSAVMI 161
                          330       340       350       360       370       380       390       400
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  349 KRYDEHHQDLTLLKALVRQQLPEKYKEIFFDqskngyagyidggasqeefykfikpilekmdgteellvklnreDLLRKQ 428
Cdd:TIGR01865  162 NRYLEHEADLRTLKELILKKFPKKYKEIFSE-------------------------------------------TFLRNQ 198
                          410       420       430       440       450       460       470       480
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  429 RTFDNGSIPHQIHLGELHAILRRQEDFYPFlkdnrekiEKILTFRIPYYVGPLARGNSRFAwmtrkseetitpwnfeeVV 508
Cdd:TIGR01865  199 RGFYNGSIPRQLLLEELEAIFRKQREYYPF--------IKLLTFRIPYYIGPLAEGKSEFA-----------------FV 253
                          490       500       510       520       530       540       550       560
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  509 DKGASAQSFIERMTNFDKNLPNEKVLPKHSLLYEYFTVYNELTKVKYVTEGMRKPAFLSGEQKKAIVDLLFKTNRKVTVK 588
Cdd:TIGR01865  254 DKPASAENFIEKMTGKCTYLPEEKRAPKHSLLAEKFTVLNELNNVRIIILEQGETKILSKEEKQELLDLLFKKKKLTYKK 333
                          570       580       590       600       610       620       630       640
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  589 QLKEDYFKKIECFDSVEISGVEDR---FNASLGTYHDLLKIIKDKDFLDNEENEDILEDIVLTLTLFEDREMIEERLKTY 665
Cdd:TIGR01865  334 LRKLLGLSEDAIFKGLRYEGLDNAekaFNISLKTYHKLRKALGDKDLLDNPKNPKDLDEIVKILTLYKDREMIKKRLELY 413
                          650       660       670       680       690       700       710       720
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  666 AHLFDDKVMKQLKRRRYTGWGRLSRKLINGIRDKQSGKTILDFLKSDGFANRNFMQLIHDDSLTFKEDIQKAqvsgqgds 745
Cdd:TIGR01865  414 KDVLNEEQVKKLVRLHFTGWGRLSLKALRGIRPLMEQGKRYDEAILELGGNRNFMQNINDSQLLPKINITKA-------- 485
                          730       740       750       760       770       780       790       800
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  746 lhehiANLAGSPAIKKGILQTVKVVDELVKVMGrhKPENIVIEMARENQTTQKGQKNSRERMKRIEEGIKELGS----QI 821
Cdd:TIGR01865  486 -----KDEILNPVVKRALLQARKVVNELVKKYG--PPDKIVIEMAREEQGTNFGKRNSKERYKKNEDKIKEFASalgkEI 558
                          810       820       830       840       850       860       870       880
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  822 LKEHPVENTQLQNEKLYLYYLQNGRDMYVDQELDINRL---SDYDVDAIVPQSFLKDDSIDNKVLTRSDKNRGKSDNVPS 898
Cdd:TIGR01865  559 LKEEPTENSSKNILKLRLYYQQNGKCMYTGKEIDIDDLfdlSYYEIDHILPQSRSFDDSISNKVLVLASENQEKGDQTPY 638
                          890       900       910       920       930       940       950       960
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  899 E-EVVKKMKNYWRQLLNAKLITQRKFDNLTKAERGGLSELDKAGFIKRQLVETRQITKHVAQILDSRMNTKYDendklIR 977
Cdd:TIGR01865  639 EaEIVKKDSAFWNKFEAYVLISKRKSDKLTRAERGGLSDDDKAGFIDRNLNDTRYITRVVANYLKDRFNFHLK-----KR 713
                          970       980       990      1000      1010      1020      1030      1040
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  978 EVKVITLKSKLVSDFRKDFQFYKVREINNYHHAHDAYLNAVVGTALIKKYPKLESEFVYGDYKVYDVRKMIAkseqeigK 1057
Cdd:TIGR01865  714 KVKVVTLKGQLTSQLRKKWGLYKKREINNYHHAHDAYINAVSTNALVKKFSQLEPEFRYKEYHNFDGRKKKK-------S 786
                         1050
                   ....*....|....*....
gi 1057838381 1058 ATAKYFFYSNIMNFFKTEI 1076
Cdd:TIGR01865  787 ATDKKVKFSNPMEFFKQKV 805
Csn1 cd09643
CRISPR/Cas system-associated protein Cas9; CRISPR (Clustered Regularly Interspaced Short ...
30-1075 0e+00

CRISPR/Cas system-associated protein Cas9; CRISPR (Clustered Regularly Interspaced Short Palindromic Repeats) and associated Cas proteins comprise a system for heritable host defense by prokaryotic cells against phage and other foreign DNA; Very large protein containing McrA/HNH-nuclease related domain and a RuvC-like nuclease domain; signature gene for type II


Pssm-ID: 187774 [Multi-domain]  Cd Length: 799  Bit Score: 838.23  E-value: 0e+00
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381   30 KYSIGLAIGTNSVGWAVITDEYKVPSKKFKvlgntdrHSIKKNLIGALLFDSGETAE-ATRLKRTARRRYTRRKNRICYL 108
Cdd:cd09643      1 EYILGLDIGIASVGWAIVEDDYKVPAKKMI-------DCGVKIFTGAELFKTGETAAlDRRLARGARRRIRRRKHRLLRL 73
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  109 QEIFSNEMAKVDDSFFHRLEESFLveedkkherhpifgnivdevAYHEKYPTIYHLRKKLVDSTDKADLrlIYLALAHMI 188
Cdd:cd09643     74 QELFAREGSLTDFDFFSRLEDSFL--------------------EYHKNYPTIYHLRKAALENKLKPDE--LYLALLHII 131
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  189 KFRGHFLIEGDLNPDNsdvdklfiqlvqtynqlfeenpinasgvdakailsarlsksrrlenliaqlpgekknglfgnli 268
Cdd:cd09643    132 KHRGHFLIEGDEDTTA---------------------------------------------------------------- 147
                          250       260       270       280       290       300       310       320
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  269 alslgltpnfksnfdlaedaklqlskdtydddldnllaqigdqyadlflaaknlsdaillsdilrvnTEITKAPLSASMI 348
Cdd:cd09643    148 -------------------------------------------------------------------DKETGALLSASMI 160
                          330       340       350       360       370       380       390       400
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  349 KRYDEHHQDLTLLKALVRQQLPEKYKEIFFDqskngyagyidggasqeefykfikpilekmdgteellvklnrEDLLRKQ 428
Cdd:cd09643    161 KRYDEHKADLRKLKELIKKEFFKKYKEIFGD------------------------------------------ETFLRNQ 198
                          410       420       430       440       450       460       470       480
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  429 RTFDNGSIPHQIHLGELHAILRRQEDFYPFlkdnrekiEKILTFRIPYYVGPLARGNSRFAWMTRKSEEtitpwnfeevv 508
Cdd:cd09643    199 RGFYNGSIPRQLLLEELEAIFRKQREYYPF--------EKILTFRIPYYIGPLAEGKSEFAWLTRPALS----------- 259
                          490       500       510       520       530       540       550       560
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  509 dkgasaQSFIERMTNFDKNLPNEKVLPKHSLLYEYFTVYNELTKVKYVTEgMRKPAFLSGEQKKAIVDLLFKTNRKVTVK 588
Cdd:cd09643    260 ------EAFIEKMTGKCTYLPEEKRAPKHSLLAEKFTVLNELNNLRIIEE-QGETKILSKEEKQELLDLLFKKNKLTYKQ 332
                          570       580       590       600       610       620       630       640
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  589 QLKEDYFKKIECFDSVEISG--VEDRFNASLGTYHDLLKIIKDKDFLDNEENEDILEDIVLTLTLFEDREMIEERLKTYA 666
Cdd:cd09643    333 KRKLLGLKEEEIFKGLRYEGlkAEKNFNISLKTYHDLRKALGKEFLKDLELNEKILDEIVKILTLYKDREMIEKILELYK 412
                          650       660       670       680       690       700       710       720
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  667 HLFDDKVMKQLKRRRYTGWGRLSRKLINGIRDKQSGKTILDFLKSDGFANRNfmQLIHDDSLTFKEDIQKAQVsgqgdsl 746
Cdd:cd09643    413 DLLNEEQLKKLLKRHFTGWGRLSLKALRGIRPLMEQGKRYDEAILELGGNHN--QKINSDELKFLPIIKKAQV------- 483
                          730       740       750       760       770       780       790       800
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  747 hehiANLAGSPAIKKGILQTVKVVDELVKVMGrhKPENIVIEMARENQtTQKGQKNSRERMKRIEEGIKELGS---QILK 823
Cdd:cd09643    484 ----KDEILNPVVKRALLQARKVVNELVKKYG--PPDKIVIEMARENG-TNKGTKNRKKRQKKNEDNIKEAASaleQKLK 556
                          810       820       830       840       850       860       870       880
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  824 EHPVENTQLQNEKLYLYYLQNGRDMYVDQELDINRL---SDYDVDAIVPQSFLKDDSIDNKVLTRSDKNRGKSDNVPSEE 900
Cdd:cd09643    557 ELPLDIKSKNILKLRLYYQQNGKCMYTGKEIDIDDLfdlSYYEIDHILPQSRSFDDSISNKVLVLASENQEKGDQTPYEE 636
                          890       900       910       920       930       940       950       960
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  901 VVKKMKNYWRQLLNAKLITQR---KFDNLtKAERgGLSELDKAGFIKRQLVETRQITKHVAQILDSRMNTKYDendklIR 977
Cdd:cd09643    637 IVSKMSAFWNKLEAAKLISQRgdsKKDRL-LLEK-GISDDEKAGFIDRNLNDTRYITRVVANYLKDRFNFHLK-----KR 709
                          970       980       990      1000      1010      1020      1030      1040
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  978 EVKVITLKSKLVSDFRKDFQFYKVREINNYHHAHDAYLNAVVGTALIKKYPKLEsefVYGDYKVYDVRKMIAKSEQEIgk 1057
Cdd:cd09643    710 KVKVVTLKGQLTSQLRKKWGLYKKREINNYHHAHDAYINAVVTNALVKKFSQLE---RYKEYKRFDSEKGNKKTLDEN-- 784
                         1050
                   ....*....|....*...
gi 1057838381 1058 ataKYFFYSNIMNFFKTE 1075
Cdd:cd09643    785 ---KKFFFANPMNFFKQE 799
Cas9_REC pfam16592
REC lobe of CRISPR-associated endonuclease Cas9; The REC lobe of Cas9 - the CRISPR-associated ...
207-736 7.56e-180

REC lobe of CRISPR-associated endonuclease Cas9; The REC lobe of Cas9 - the CRISPR-associated endonuclease Cas9 - includes the REC1 and REC2 domains. REC1 forms an elongated, alpha-helical structure consisting of 25 alpha helices and two beta-sheets, whereas REC2 inserted within REC1 adopts a six-helix bundle structure. The REC lobe and the NUC lobe of Cas9 fold to present a positively charged groove at their interface which accommodates the negatively charged sgRNA:target DNA heteroduplex. CRISPR (clustered regularly interspaced short palindromic repeat)-Cas system occurs naturally in bacteria as a defence against invasion by phages or other mobile genetic elements. Cas9 is targeted to specific genomic locations by sgRNAs or single guide RNAs, in order to complex with invading DNA in order to cleave it and render it inactive.


Pssm-ID: 435447  Cd Length: 539  Bit Score: 550.13  E-value: 7.56e-180
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  207 VDKLFIQLVQTYNQLFEENPINASGVDAKAILSA-RLSKSRRLENLIAQLPGEK-KNGLFGNLIALSLGLTPNFKSNFDL 284
Cdd:pfam16592    1 VEESFQDLLNILYEQLENLELETQNVEIEKILKKtKISKKAKLDELLALPPNEKnSKKIFAEILKLILGNKADFTKIFEL 80
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  285 ------AEDAKLQLSKDTYDDDLDNLLAQIGDQYADLFLAAKNLSDAILLSDILRVNTEITKAPLSASMIKRYDEHHQDL 358
Cdd:pfam16592   81 ekfveePKKIKLSFSDSNYDEKIEELENQLGDEKAEIILILKKIYDWVVLSDILTVSTDNGKAYLSEAMVNRYDKHKEDL 160
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  359 TLLKALVRQQLPEKYKEIFFDQSKNGYAGYID----GGASQEEFYKFIKPILEKMDGTEE--LLVKLNREDLLRKQRTFD 432
Cdd:pfam16592  161 AQLKKVIKQNLSEKYNDMFRKEKKKGYSAYINgknnGKTSKEDFYKYIKKLINKVETSEAqyILSKIDNENFLPKQRTKS 240
                          250       260       270       280       290       300       310       320
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  433 NGSIPHQIHLGELHAILRRQEDFYPFLKDNREKIEKILTFRIPYYVGPLARGNSRFAWMTRKSEETITPWNFEEVVDKGA 512
Cdd:pfam16592  241 NGSIPYQVHLQELKKIIKNQAEYYPFLKENQEKILKLLTFRIPYYVGPLAEKKSKFAWMKRKEQGKIYPWNFEQKVDIDK 320
                          330       340       350       360       370       380       390       400
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  513 SAQSFIERMTNFDKNLPNEKVLPKHSLLYEYFTVYNELTKVKYVTEgmrkpaFLSGEQKKAIVDLLFKTNRKVTVKQLKE 592
Cdd:pfam16592  321 TAEAFITRMTNYCTYLPDEKVLPKNSLLYSKFTVLNELNKIKINGE------KISVELKQDIFNGLFKKNKKVTKKKLKD 394
                          410       420       430       440       450       460       470       480
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  593 DYFKKIECFDSVEISGV--EDRFNASLGTYHDLLKIIkdKDFLDNEENEDILEDIVLTLTLFEDREMIEERL-KTYAHLF 669
Cdd:pfam16592  395 WLVKEGYNFKAVEIKGFdkENNFNNSLTTYIDLAKIF--GDFLDNPDNEDIIEDIIYWLTLFEDRKILKRRLqKKYSNLL 472
                          490       500       510       520       530       540       550
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  670 DDKVMKQLKRRRYTGWGRLSRKLINGIRDKQS---GKTILDFLKSDgfaNRNFMQLIHDDSLTFKEDIQK 736
Cdd:pfam16592  473 TEKQIKQILKLKYKGWGRLSKELLNGIRGADRqgeIKTIIDLLWND---NRNLMQLINDERLSFKEEIEK 539
Cas9 COG3513
CRISPR-Cas system type-II protein Cas9 [Defense mechanisms]; CRISPR-Cas system type-II protein ...
29-1044 2.66e-119

CRISPR-Cas system type-II protein Cas9 [Defense mechanisms]; CRISPR-Cas system type-II protein Cas9 is part of the Pathway/BioSystem: CRISPR-Cas system


Pssm-ID: 442735 [Multi-domain]  Cd Length: 812  Bit Score: 396.26  E-value: 2.66e-119
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381   29 KKYSIGLAIGTNSVGWAVITDEYKVpskkfkvlgntdrHSIKKNLIGALLFDSGET-------AEATRLKRTARRRYTRR 101
Cdd:COG3513      2 DKYILGLDLGINSVGWAVLELDEDG-------------EPGEIIDAGVRIFDDGEDpksgeslAAARREARGARRRRRRR 68
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  102 KNRICYLQEIFSNEMakvddsffhrleesFLVEEDKKHERHPifgnivdevayhekYPTIYHLRKKLVDstDKADLRLIY 181
Cdd:COG3513     69 KHRLRRLKRLLVEEG--------------LLPADDAERKALL--------------PLNPYELRAKALD--EKLSPEELG 118
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  182 LALAHMIKFRGHfliegdLNPDNSDVDKLfiqlvqtynqlfeenpinasgvDAKAILSARLSKSRRLENLIAQLPGEkkn 261
Cdd:COG3513    119 RALFHLAQRRGF------KSNRKTDSKDN----------------------ESGKVKDAIKELRERLEAKGARTVGE--- 167
                          250       260       270       280       290       300       310       320
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  262 glfgnlialslgltpnfksnfdlaedaklqlskdtydddldnllaqigdqyadlFLAaknlsdaillsdilrvnteitka 341
Cdd:COG3513    168 ------------------------------------------------------YLY----------------------- 170
                          330       340       350       360       370       380       390       400
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  342 plsasmiKRYDEHHQdltllkalvrqqlpekykeiffdqskngyagyidggasqeefykfikpilekmdgteellvklnr 421
Cdd:COG3513    171 -------RRLQENGK----------------------------------------------------------------- 178
                          410       420       430       440       450       460       470       480
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  422 edlLRKQRTFDNGSIPHQIHLGELHAILRRQEDFYPFLKDN--REKIEKILTFRIPYYVGplargnsrfawmtrkseeti 499
Cdd:COG3513    179 ---VRNRKGDYDFYIPREDLEDEFEAIWAAQAEFGPALLTEelRDELLEIIFFQRPLKSG-------------------- 235
                          490       500       510       520       530       540       550       560
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  500 tpwnfeevvdkgasaqsfiERMTNFDKNLPNEKVLPKHSLLYEYFTVYNELTKVKYVTEGmRKPAFLSGEQKKAIVDLLF 579
Cdd:COG3513    236 -------------------KKLVGKCTFEPDEKRAPKASPLFQRFRILQKLNNLRIVDDG-GEERPLTLEERQKIIDLLE 295
                          570       580       590       600       610       620       630       640
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  580 KtNRKVTVKQLKEDYfkKIEcfDSVEISGVEDRFN-----ASLGTYHDLLKIIKDKDFldNEENEDILEDIVLTLTLFED 654
Cdd:COG3513    296 N-KKKLTFKKLRKLL--GLP--DGVIFKGFNYEDDdraklKGDKTYAKLAKIFGKAWL--NEFDPEILDDIVEALTLFKD 368
                          650       660       670       680       690       700       710       720
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  655 REMIEERLKTYAHLfDDKVMKQLKRRR-YTGWGRLSRKLINGIrdkqsgktiLDFLKSDgfanrnfmqlihddsLTFKED 733
Cdd:COG3513    369 DEELKEWLKKLYGL-DEEQAEALANLPlPDGYGNLSLKALRKI---------LPLLEEG---------------LDYDEA 423
                          730       740       750       760       770       780       790       800
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  734 IQKAQVSGQGDSLH--------EHIANLAGSPAIKKGILQTVKVVDELVKVMGrhKPENIVIEMARENQTTQKGQKNSRE 805
Cdd:COG3513    424 VKAAGYDHSSLEILdrlppigeEKRKGSIRNPVVHRALNQLRKVVNALIRKYG--KPDEIHIELARDLKKSKKERKEIQK 501
                          810       820       830       840       850       860       870       880
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  806 RMKRIEEGIKELGSQILKEHPVENTQLQNEKLYLYYLQNGRDMYVDQELDINRLSD--YDVDAIVPQSFLKDDSIDNKVL 883
Cdd:COG3513    502 RQRENEKAREKAREEIAEEGGGEPSRRDILKYRLWEEQNGRCPYTGKPISISDLLDgsVEIDHILPRSRTLDDSFNNKVL 581
                          890       900       910       920       930       940       950       960
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  884 TRSDKNRGKSDNVPSEEVVK----KMKNYWRQLLNAKLITQRKFDNLTKAERGglsELDKAGFIKRQLVETRQITKHVAQ 959
Cdd:COG3513    582 CLADANREKGNRTPYEALGGdeaeKWEEILARVENLKLIPQKKKKRFLKKELD---RDDDEGFIARQLNDTRYISRLAAE 658
                          970       980       990      1000      1010      1020      1030      1040
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  960 ILDSRMNTKYdendklIREVKVITLKSKLVSDFRKDFQFYKV-------REINNYHHAHDAYLNAVVGTALIKKYPKLES 1032
Cdd:COG3513    659 YLKSLYPFED------KGKRKVRVVPGQLTAMLRRAWGLNKIlsddgekNRDDHRHHAIDALVIACTTQGLLQRLAKASR 732
                         1050
                   ....*....|..
gi 1057838381 1033 EFVYGDYKVYDV 1044
Cdd:COG3513    733 EREDAEKAEEHF 744
Cas9_PI pfam16595
PAM-interacting domain of CRISPR-associated endonuclease Cas9; Cas9_PI is a family found at ...
1128-1384 1.46e-43

PAM-interacting domain of CRISPR-associated endonuclease Cas9; Cas9_PI is a family found at the C-terminal of bacterial type II CRISPR system Cas9 endonuclease. This domain adopts a novel protein fold that is unique to the Cas9 family. It is positioned in the structure-DNA-complex to recognize the PAM sequence on the non-complementary DNA strand of the crRNA. PAM sequence is protospacer-adjacent motifs on DNA. See family CRISPR-DR2, Rfam:RF01315. Cas9 carries two nuclease domains, HNH and RuvC, which cleave the DNA strands that are complementary and non-complementary to the 20 nucleotide guide sequence in crRNAs, respectively.


Pssm-ID: 435449  Cd Length: 264  Bit Score: 159.79  E-value: 1.46e-43
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381 1128 TGGFSKESILP--KRNSDKLIARKKD---WDPKKYGGFDSPTVAYSVLVVAKVEKGKSKKLksvkeLLGITIMERSSFEK 1202
Cdd:pfam16595    1 KGGLFNQTILPahKKKGKGLIPLKKDergLDVEKYGGYSSLTAAYFSLVEYTGKKGKRKRT-----IEGVPLYLAAKIEE 75
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381 1203 NPI--DFLEAKGYKEVKKDLIIKLPKYSLFElENGRKRMLASAGE---LQKGNELALPSKYVNFLYLASHYEKLKGSPED 1277
Cdd:pfam16595   76 NKDllEYLEEKLGLKEPKIILPKIKKNSLIK-IDGFRMLLTGKTEnrlLKNAVQLVLSNDDEKYIKKIEKFVKKNKDDII 154
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381 1278 NEQKQLFVEQHKHYLDEIIEQISEFSKrVILADANLDKVLSAYNKHRDKPIREQAENIIHLFTLTNLGA-PAAFKYFDTT 1356
Cdd:pfam16595  155 EEKDGLTEEKNIKLYDELLDKMKNTIY-YKRPSNQGEKLEKLKEKFIKLSLEEKCKVLIEILKLTHANPtSADLKLIGGS 233
                          250       260       270
                   ....*....|....*....|....*....|.
gi 1057838381 1357 IDRKRYTSTKEVLDA---TLIHQSITGLYET 1384
Cdd:pfam16595  234 KHAGRIKISNNISKAsniKLINQSVTGLYEK 264
 
Name Accession Description Interval E-value
cas_Csn1 TIGR01865
CRISPR subtype II/NMENI RNA-guided endonuclease Cas9/Csn1; CRISPR loci appear to be mobile ...
30-1076 0e+00

CRISPR subtype II/NMENI RNA-guided endonuclease Cas9/Csn1; CRISPR loci appear to be mobile elements with a wide host range. This model represents a protein found only in CRISPR-containing species, near other CRISPR-associated proteins (cas), as part of the NMENI subtype of CRISPR/Cas locus. The species range so far for this protein is animal pathogens and commensals only.


Pssm-ID: 273840  Cd Length: 805  Bit Score: 852.10  E-value: 0e+00
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381   30 KYSIGLAIGTNSVGWAVITDEYKVPSKKFKVLGNtdrhsiKKNLIGALLFDSGETAE-ATRLKRTARRRYTRRKNRICYL 108
Cdd:TIGR01865    1 EYILGLDIGIASVGWAIVEDDYKVPAAKRLIDGG------VRNFTGAELPKTGETAAlDRRLARGARRRIRRRKHRLLRL 74
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  109 QEIFSNEMAKVDDSFFHRLEESFLVEEDKKHerhpifgnivdevayhekypTIYHLRKKLVDSTDKADLrlIYLALAHMI 188
Cdd:TIGR01865   75 QELFSREGSLTDFDFFSRLENSFLVEEDKRN--------------------TIYHLRKAALENKLKPDE--LYLALLHII 132
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  189 KFRGHFLIEgdlnpdnsdvdklfiqlvqtynqlfeenpinasgvdakailsarlsksrrlenliaqlpgekknglfgnli 268
Cdd:TIGR01865  133 KHRGHFLIE----------------------------------------------------------------------- 141
                          250       260       270       280       290       300       310       320
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  269 alslgltpnfksnfdlaedaklqlskdtyDDDLDnllaqigdqyadlflaaknlsdaillsdilrVNTEITKAPLSASMI 348
Cdd:TIGR01865  142 -----------------------------GNDFD-------------------------------TANKETGALLSAVMI 161
                          330       340       350       360       370       380       390       400
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  349 KRYDEHHQDLTLLKALVRQQLPEKYKEIFFDqskngyagyidggasqeefykfikpilekmdgteellvklnreDLLRKQ 428
Cdd:TIGR01865  162 NRYLEHEADLRTLKELILKKFPKKYKEIFSE-------------------------------------------TFLRNQ 198
                          410       420       430       440       450       460       470       480
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  429 RTFDNGSIPHQIHLGELHAILRRQEDFYPFlkdnrekiEKILTFRIPYYVGPLARGNSRFAwmtrkseetitpwnfeeVV 508
Cdd:TIGR01865  199 RGFYNGSIPRQLLLEELEAIFRKQREYYPF--------IKLLTFRIPYYIGPLAEGKSEFA-----------------FV 253
                          490       500       510       520       530       540       550       560
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  509 DKGASAQSFIERMTNFDKNLPNEKVLPKHSLLYEYFTVYNELTKVKYVTEGMRKPAFLSGEQKKAIVDLLFKTNRKVTVK 588
Cdd:TIGR01865  254 DKPASAENFIEKMTGKCTYLPEEKRAPKHSLLAEKFTVLNELNNVRIIILEQGETKILSKEEKQELLDLLFKKKKLTYKK 333
                          570       580       590       600       610       620       630       640
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  589 QLKEDYFKKIECFDSVEISGVEDR---FNASLGTYHDLLKIIKDKDFLDNEENEDILEDIVLTLTLFEDREMIEERLKTY 665
Cdd:TIGR01865  334 LRKLLGLSEDAIFKGLRYEGLDNAekaFNISLKTYHKLRKALGDKDLLDNPKNPKDLDEIVKILTLYKDREMIKKRLELY 413
                          650       660       670       680       690       700       710       720
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  666 AHLFDDKVMKQLKRRRYTGWGRLSRKLINGIRDKQSGKTILDFLKSDGFANRNFMQLIHDDSLTFKEDIQKAqvsgqgds 745
Cdd:TIGR01865  414 KDVLNEEQVKKLVRLHFTGWGRLSLKALRGIRPLMEQGKRYDEAILELGGNRNFMQNINDSQLLPKINITKA-------- 485
                          730       740       750       760       770       780       790       800
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  746 lhehiANLAGSPAIKKGILQTVKVVDELVKVMGrhKPENIVIEMARENQTTQKGQKNSRERMKRIEEGIKELGS----QI 821
Cdd:TIGR01865  486 -----KDEILNPVVKRALLQARKVVNELVKKYG--PPDKIVIEMAREEQGTNFGKRNSKERYKKNEDKIKEFASalgkEI 558
                          810       820       830       840       850       860       870       880
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  822 LKEHPVENTQLQNEKLYLYYLQNGRDMYVDQELDINRL---SDYDVDAIVPQSFLKDDSIDNKVLTRSDKNRGKSDNVPS 898
Cdd:TIGR01865  559 LKEEPTENSSKNILKLRLYYQQNGKCMYTGKEIDIDDLfdlSYYEIDHILPQSRSFDDSISNKVLVLASENQEKGDQTPY 638
                          890       900       910       920       930       940       950       960
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  899 E-EVVKKMKNYWRQLLNAKLITQRKFDNLTKAERGGLSELDKAGFIKRQLVETRQITKHVAQILDSRMNTKYDendklIR 977
Cdd:TIGR01865  639 EaEIVKKDSAFWNKFEAYVLISKRKSDKLTRAERGGLSDDDKAGFIDRNLNDTRYITRVVANYLKDRFNFHLK-----KR 713
                          970       980       990      1000      1010      1020      1030      1040
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  978 EVKVITLKSKLVSDFRKDFQFYKVREINNYHHAHDAYLNAVVGTALIKKYPKLESEFVYGDYKVYDVRKMIAkseqeigK 1057
Cdd:TIGR01865  714 KVKVVTLKGQLTSQLRKKWGLYKKREINNYHHAHDAYINAVSTNALVKKFSQLEPEFRYKEYHNFDGRKKKK-------S 786
                         1050
                   ....*....|....*....
gi 1057838381 1058 ATAKYFFYSNIMNFFKTEI 1076
Cdd:TIGR01865  787 ATDKKVKFSNPMEFFKQKV 805
Csn1 cd09643
CRISPR/Cas system-associated protein Cas9; CRISPR (Clustered Regularly Interspaced Short ...
30-1075 0e+00

CRISPR/Cas system-associated protein Cas9; CRISPR (Clustered Regularly Interspaced Short Palindromic Repeats) and associated Cas proteins comprise a system for heritable host defense by prokaryotic cells against phage and other foreign DNA; Very large protein containing McrA/HNH-nuclease related domain and a RuvC-like nuclease domain; signature gene for type II


Pssm-ID: 187774 [Multi-domain]  Cd Length: 799  Bit Score: 838.23  E-value: 0e+00
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381   30 KYSIGLAIGTNSVGWAVITDEYKVPSKKFKvlgntdrHSIKKNLIGALLFDSGETAE-ATRLKRTARRRYTRRKNRICYL 108
Cdd:cd09643      1 EYILGLDIGIASVGWAIVEDDYKVPAKKMI-------DCGVKIFTGAELFKTGETAAlDRRLARGARRRIRRRKHRLLRL 73
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  109 QEIFSNEMAKVDDSFFHRLEESFLveedkkherhpifgnivdevAYHEKYPTIYHLRKKLVDSTDKADLrlIYLALAHMI 188
Cdd:cd09643     74 QELFAREGSLTDFDFFSRLEDSFL--------------------EYHKNYPTIYHLRKAALENKLKPDE--LYLALLHII 131
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  189 KFRGHFLIEGDLNPDNsdvdklfiqlvqtynqlfeenpinasgvdakailsarlsksrrlenliaqlpgekknglfgnli 268
Cdd:cd09643    132 KHRGHFLIEGDEDTTA---------------------------------------------------------------- 147
                          250       260       270       280       290       300       310       320
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  269 alslgltpnfksnfdlaedaklqlskdtydddldnllaqigdqyadlflaaknlsdaillsdilrvnTEITKAPLSASMI 348
Cdd:cd09643    148 -------------------------------------------------------------------DKETGALLSASMI 160
                          330       340       350       360       370       380       390       400
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  349 KRYDEHHQDLTLLKALVRQQLPEKYKEIFFDqskngyagyidggasqeefykfikpilekmdgteellvklnrEDLLRKQ 428
Cdd:cd09643    161 KRYDEHKADLRKLKELIKKEFFKKYKEIFGD------------------------------------------ETFLRNQ 198
                          410       420       430       440       450       460       470       480
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  429 RTFDNGSIPHQIHLGELHAILRRQEDFYPFlkdnrekiEKILTFRIPYYVGPLARGNSRFAWMTRKSEEtitpwnfeevv 508
Cdd:cd09643    199 RGFYNGSIPRQLLLEELEAIFRKQREYYPF--------EKILTFRIPYYIGPLAEGKSEFAWLTRPALS----------- 259
                          490       500       510       520       530       540       550       560
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  509 dkgasaQSFIERMTNFDKNLPNEKVLPKHSLLYEYFTVYNELTKVKYVTEgMRKPAFLSGEQKKAIVDLLFKTNRKVTVK 588
Cdd:cd09643    260 ------EAFIEKMTGKCTYLPEEKRAPKHSLLAEKFTVLNELNNLRIIEE-QGETKILSKEEKQELLDLLFKKNKLTYKQ 332
                          570       580       590       600       610       620       630       640
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  589 QLKEDYFKKIECFDSVEISG--VEDRFNASLGTYHDLLKIIKDKDFLDNEENEDILEDIVLTLTLFEDREMIEERLKTYA 666
Cdd:cd09643    333 KRKLLGLKEEEIFKGLRYEGlkAEKNFNISLKTYHDLRKALGKEFLKDLELNEKILDEIVKILTLYKDREMIEKILELYK 412
                          650       660       670       680       690       700       710       720
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  667 HLFDDKVMKQLKRRRYTGWGRLSRKLINGIRDKQSGKTILDFLKSDGFANRNfmQLIHDDSLTFKEDIQKAQVsgqgdsl 746
Cdd:cd09643    413 DLLNEEQLKKLLKRHFTGWGRLSLKALRGIRPLMEQGKRYDEAILELGGNHN--QKINSDELKFLPIIKKAQV------- 483
                          730       740       750       760       770       780       790       800
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  747 hehiANLAGSPAIKKGILQTVKVVDELVKVMGrhKPENIVIEMARENQtTQKGQKNSRERMKRIEEGIKELGS---QILK 823
Cdd:cd09643    484 ----KDEILNPVVKRALLQARKVVNELVKKYG--PPDKIVIEMARENG-TNKGTKNRKKRQKKNEDNIKEAASaleQKLK 556
                          810       820       830       840       850       860       870       880
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  824 EHPVENTQLQNEKLYLYYLQNGRDMYVDQELDINRL---SDYDVDAIVPQSFLKDDSIDNKVLTRSDKNRGKSDNVPSEE 900
Cdd:cd09643    557 ELPLDIKSKNILKLRLYYQQNGKCMYTGKEIDIDDLfdlSYYEIDHILPQSRSFDDSISNKVLVLASENQEKGDQTPYEE 636
                          890       900       910       920       930       940       950       960
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  901 VVKKMKNYWRQLLNAKLITQR---KFDNLtKAERgGLSELDKAGFIKRQLVETRQITKHVAQILDSRMNTKYDendklIR 977
Cdd:cd09643    637 IVSKMSAFWNKLEAAKLISQRgdsKKDRL-LLEK-GISDDEKAGFIDRNLNDTRYITRVVANYLKDRFNFHLK-----KR 709
                          970       980       990      1000      1010      1020      1030      1040
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  978 EVKVITLKSKLVSDFRKDFQFYKVREINNYHHAHDAYLNAVVGTALIKKYPKLEsefVYGDYKVYDVRKMIAKSEQEIgk 1057
Cdd:cd09643    710 KVKVVTLKGQLTSQLRKKWGLYKKREINNYHHAHDAYINAVVTNALVKKFSQLE---RYKEYKRFDSEKGNKKTLDEN-- 784
                         1050
                   ....*....|....*...
gi 1057838381 1058 ataKYFFYSNIMNFFKTE 1075
Cdd:cd09643    785 ---KKFFFANPMNFFKQE 799
Cas9_REC pfam16592
REC lobe of CRISPR-associated endonuclease Cas9; The REC lobe of Cas9 - the CRISPR-associated ...
207-736 7.56e-180

REC lobe of CRISPR-associated endonuclease Cas9; The REC lobe of Cas9 - the CRISPR-associated endonuclease Cas9 - includes the REC1 and REC2 domains. REC1 forms an elongated, alpha-helical structure consisting of 25 alpha helices and two beta-sheets, whereas REC2 inserted within REC1 adopts a six-helix bundle structure. The REC lobe and the NUC lobe of Cas9 fold to present a positively charged groove at their interface which accommodates the negatively charged sgRNA:target DNA heteroduplex. CRISPR (clustered regularly interspaced short palindromic repeat)-Cas system occurs naturally in bacteria as a defence against invasion by phages or other mobile genetic elements. Cas9 is targeted to specific genomic locations by sgRNAs or single guide RNAs, in order to complex with invading DNA in order to cleave it and render it inactive.


Pssm-ID: 435447  Cd Length: 539  Bit Score: 550.13  E-value: 7.56e-180
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  207 VDKLFIQLVQTYNQLFEENPINASGVDAKAILSA-RLSKSRRLENLIAQLPGEK-KNGLFGNLIALSLGLTPNFKSNFDL 284
Cdd:pfam16592    1 VEESFQDLLNILYEQLENLELETQNVEIEKILKKtKISKKAKLDELLALPPNEKnSKKIFAEILKLILGNKADFTKIFEL 80
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  285 ------AEDAKLQLSKDTYDDDLDNLLAQIGDQYADLFLAAKNLSDAILLSDILRVNTEITKAPLSASMIKRYDEHHQDL 358
Cdd:pfam16592   81 ekfveePKKIKLSFSDSNYDEKIEELENQLGDEKAEIILILKKIYDWVVLSDILTVSTDNGKAYLSEAMVNRYDKHKEDL 160
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  359 TLLKALVRQQLPEKYKEIFFDQSKNGYAGYID----GGASQEEFYKFIKPILEKMDGTEE--LLVKLNREDLLRKQRTFD 432
Cdd:pfam16592  161 AQLKKVIKQNLSEKYNDMFRKEKKKGYSAYINgknnGKTSKEDFYKYIKKLINKVETSEAqyILSKIDNENFLPKQRTKS 240
                          250       260       270       280       290       300       310       320
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  433 NGSIPHQIHLGELHAILRRQEDFYPFLKDNREKIEKILTFRIPYYVGPLARGNSRFAWMTRKSEETITPWNFEEVVDKGA 512
Cdd:pfam16592  241 NGSIPYQVHLQELKKIIKNQAEYYPFLKENQEKILKLLTFRIPYYVGPLAEKKSKFAWMKRKEQGKIYPWNFEQKVDIDK 320
                          330       340       350       360       370       380       390       400
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  513 SAQSFIERMTNFDKNLPNEKVLPKHSLLYEYFTVYNELTKVKYVTEgmrkpaFLSGEQKKAIVDLLFKTNRKVTVKQLKE 592
Cdd:pfam16592  321 TAEAFITRMTNYCTYLPDEKVLPKNSLLYSKFTVLNELNKIKINGE------KISVELKQDIFNGLFKKNKKVTKKKLKD 394
                          410       420       430       440       450       460       470       480
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  593 DYFKKIECFDSVEISGV--EDRFNASLGTYHDLLKIIkdKDFLDNEENEDILEDIVLTLTLFEDREMIEERL-KTYAHLF 669
Cdd:pfam16592  395 WLVKEGYNFKAVEIKGFdkENNFNNSLTTYIDLAKIF--GDFLDNPDNEDIIEDIIYWLTLFEDRKILKRRLqKKYSNLL 472
                          490       500       510       520       530       540       550
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  670 DDKVMKQLKRRRYTGWGRLSRKLINGIRDKQS---GKTILDFLKSDgfaNRNFMQLIHDDSLTFKEDIQK 736
Cdd:pfam16592  473 TEKQIKQILKLKYKGWGRLSKELLNGIRGADRqgeIKTIIDLLWND---NRNLMQLINDERLSFKEEIEK 539
Cas9 COG3513
CRISPR-Cas system type-II protein Cas9 [Defense mechanisms]; CRISPR-Cas system type-II protein ...
29-1044 2.66e-119

CRISPR-Cas system type-II protein Cas9 [Defense mechanisms]; CRISPR-Cas system type-II protein Cas9 is part of the Pathway/BioSystem: CRISPR-Cas system


Pssm-ID: 442735 [Multi-domain]  Cd Length: 812  Bit Score: 396.26  E-value: 2.66e-119
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381   29 KKYSIGLAIGTNSVGWAVITDEYKVpskkfkvlgntdrHSIKKNLIGALLFDSGET-------AEATRLKRTARRRYTRR 101
Cdd:COG3513      2 DKYILGLDLGINSVGWAVLELDEDG-------------EPGEIIDAGVRIFDDGEDpksgeslAAARREARGARRRRRRR 68
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  102 KNRICYLQEIFSNEMakvddsffhrleesFLVEEDKKHERHPifgnivdevayhekYPTIYHLRKKLVDstDKADLRLIY 181
Cdd:COG3513     69 KHRLRRLKRLLVEEG--------------LLPADDAERKALL--------------PLNPYELRAKALD--EKLSPEELG 118
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  182 LALAHMIKFRGHfliegdLNPDNSDVDKLfiqlvqtynqlfeenpinasgvDAKAILSARLSKSRRLENLIAQLPGEkkn 261
Cdd:COG3513    119 RALFHLAQRRGF------KSNRKTDSKDN----------------------ESGKVKDAIKELRERLEAKGARTVGE--- 167
                          250       260       270       280       290       300       310       320
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  262 glfgnlialslgltpnfksnfdlaedaklqlskdtydddldnllaqigdqyadlFLAaknlsdaillsdilrvnteitka 341
Cdd:COG3513    168 ------------------------------------------------------YLY----------------------- 170
                          330       340       350       360       370       380       390       400
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  342 plsasmiKRYDEHHQdltllkalvrqqlpekykeiffdqskngyagyidggasqeefykfikpilekmdgteellvklnr 421
Cdd:COG3513    171 -------RRLQENGK----------------------------------------------------------------- 178
                          410       420       430       440       450       460       470       480
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  422 edlLRKQRTFDNGSIPHQIHLGELHAILRRQEDFYPFLKDN--REKIEKILTFRIPYYVGplargnsrfawmtrkseeti 499
Cdd:COG3513    179 ---VRNRKGDYDFYIPREDLEDEFEAIWAAQAEFGPALLTEelRDELLEIIFFQRPLKSG-------------------- 235
                          490       500       510       520       530       540       550       560
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  500 tpwnfeevvdkgasaqsfiERMTNFDKNLPNEKVLPKHSLLYEYFTVYNELTKVKYVTEGmRKPAFLSGEQKKAIVDLLF 579
Cdd:COG3513    236 -------------------KKLVGKCTFEPDEKRAPKASPLFQRFRILQKLNNLRIVDDG-GEERPLTLEERQKIIDLLE 295
                          570       580       590       600       610       620       630       640
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  580 KtNRKVTVKQLKEDYfkKIEcfDSVEISGVEDRFN-----ASLGTYHDLLKIIKDKDFldNEENEDILEDIVLTLTLFED 654
Cdd:COG3513    296 N-KKKLTFKKLRKLL--GLP--DGVIFKGFNYEDDdraklKGDKTYAKLAKIFGKAWL--NEFDPEILDDIVEALTLFKD 368
                          650       660       670       680       690       700       710       720
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  655 REMIEERLKTYAHLfDDKVMKQLKRRR-YTGWGRLSRKLINGIrdkqsgktiLDFLKSDgfanrnfmqlihddsLTFKED 733
Cdd:COG3513    369 DEELKEWLKKLYGL-DEEQAEALANLPlPDGYGNLSLKALRKI---------LPLLEEG---------------LDYDEA 423
                          730       740       750       760       770       780       790       800
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  734 IQKAQVSGQGDSLH--------EHIANLAGSPAIKKGILQTVKVVDELVKVMGrhKPENIVIEMARENQTTQKGQKNSRE 805
Cdd:COG3513    424 VKAAGYDHSSLEILdrlppigeEKRKGSIRNPVVHRALNQLRKVVNALIRKYG--KPDEIHIELARDLKKSKKERKEIQK 501
                          810       820       830       840       850       860       870       880
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  806 RMKRIEEGIKELGSQILKEHPVENTQLQNEKLYLYYLQNGRDMYVDQELDINRLSD--YDVDAIVPQSFLKDDSIDNKVL 883
Cdd:COG3513    502 RQRENEKAREKAREEIAEEGGGEPSRRDILKYRLWEEQNGRCPYTGKPISISDLLDgsVEIDHILPRSRTLDDSFNNKVL 581
                          890       900       910       920       930       940       950       960
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  884 TRSDKNRGKSDNVPSEEVVK----KMKNYWRQLLNAKLITQRKFDNLTKAERGglsELDKAGFIKRQLVETRQITKHVAQ 959
Cdd:COG3513    582 CLADANREKGNRTPYEALGGdeaeKWEEILARVENLKLIPQKKKKRFLKKELD---RDDDEGFIARQLNDTRYISRLAAE 658
                          970       980       990      1000      1010      1020      1030      1040
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381  960 ILDSRMNTKYdendklIREVKVITLKSKLVSDFRKDFQFYKV-------REINNYHHAHDAYLNAVVGTALIKKYPKLES 1032
Cdd:COG3513    659 YLKSLYPFED------KGKRKVRVVPGQLTAMLRRAWGLNKIlsddgekNRDDHRHHAIDALVIACTTQGLLQRLAKASR 732
                         1050
                   ....*....|..
gi 1057838381 1033 EFVYGDYKVYDV 1044
Cdd:COG3513    733 EREDAEKAEEHF 744
Cas9_PI pfam16595
PAM-interacting domain of CRISPR-associated endonuclease Cas9; Cas9_PI is a family found at ...
1128-1384 1.46e-43

PAM-interacting domain of CRISPR-associated endonuclease Cas9; Cas9_PI is a family found at the C-terminal of bacterial type II CRISPR system Cas9 endonuclease. This domain adopts a novel protein fold that is unique to the Cas9 family. It is positioned in the structure-DNA-complex to recognize the PAM sequence on the non-complementary DNA strand of the crRNA. PAM sequence is protospacer-adjacent motifs on DNA. See family CRISPR-DR2, Rfam:RF01315. Cas9 carries two nuclease domains, HNH and RuvC, which cleave the DNA strands that are complementary and non-complementary to the 20 nucleotide guide sequence in crRNAs, respectively.


Pssm-ID: 435449  Cd Length: 264  Bit Score: 159.79  E-value: 1.46e-43
                           10        20        30        40        50        60        70        80
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381 1128 TGGFSKESILP--KRNSDKLIARKKD---WDPKKYGGFDSPTVAYSVLVVAKVEKGKSKKLksvkeLLGITIMERSSFEK 1202
Cdd:pfam16595    1 KGGLFNQTILPahKKKGKGLIPLKKDergLDVEKYGGYSSLTAAYFSLVEYTGKKGKRKRT-----IEGVPLYLAAKIEE 75
                           90       100       110       120       130       140       150       160
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381 1203 NPI--DFLEAKGYKEVKKDLIIKLPKYSLFElENGRKRMLASAGE---LQKGNELALPSKYVNFLYLASHYEKLKGSPED 1277
Cdd:pfam16595   76 NKDllEYLEEKLGLKEPKIILPKIKKNSLIK-IDGFRMLLTGKTEnrlLKNAVQLVLSNDDEKYIKKIEKFVKKNKDDII 154
                          170       180       190       200       210       220       230       240
                   ....*....|....*....|....*....|....*....|....*....|....*....|....*....|....*....|
gi 1057838381 1278 NEQKQLFVEQHKHYLDEIIEQISEFSKrVILADANLDKVLSAYNKHRDKPIREQAENIIHLFTLTNLGA-PAAFKYFDTT 1356
Cdd:pfam16595  155 EEKDGLTEEKNIKLYDELLDKMKNTIY-YKRPSNQGEKLEKLKEKFIKLSLEEKCKVLIEILKLTHANPtSADLKLIGGS 233
                          250       260       270
                   ....*....|....*....|....*....|.
gi 1057838381 1357 IDRKRYTSTKEVLDA---TLIHQSITGLYET 1384
Cdd:pfam16595  234 KHAGRIKISNNISKAsniKLINQSVTGLYEK 264
HNH_4 pfam13395
HNH endonuclease; This HNH nuclease domain is found in CRISPR-related proteins.
847-897 5.51e-12

HNH endonuclease; This HNH nuclease domain is found in CRISPR-related proteins.


Pssm-ID: 433172 [Multi-domain]  Cd Length: 55  Bit Score: 62.26  E-value: 5.51e-12
                           10        20        30        40        50
                   ....*....|....*....|....*....|....*....|....*....|....
gi 1057838381  847 DMYVDQELDINRLSD---YDVDAIVPQSFLKDDSIDNKVLTRSDKNRGKSDNVP 897
Cdd:pfam13395    1 CPYTGEQISIDDLFSeknYDIDHILPYSRSFDDSFSNKVLVLRSANQEKGNRTP 54
 
Blast search parameters
Data Source: Precalculated data, version = cdd.v.3.21
Preset Options:Database: CDSEARCH/cdd   Low complexity filter: no  Composition Based Adjustment: yes   E-value threshold: 0.01

References:

  • Wang J et al. (2023), "The conserved domain database in 2023", Nucleic Acids Res.51(D)384-8.
  • Lu S et al. (2020), "The conserved domain database in 2020", Nucleic Acids Res.48(D)265-8.
  • Marchler-Bauer A et al. (2017), "CDD/SPARCLE: functional classification of proteins via subfamily domain architectures.", Nucleic Acids Res.45(D)200-3.
Help | Disclaimer | Write to the Help Desk
NCBI | NLM | NIH