Show PDB file:   
         Plain Text   HTML   (compressed file size)
QuickSearch:   
by PDB,NDB,UniProt,PROSITE Code or Search Term(s)  
(-)Asymmetric Unit
(-)Asym. Unit - sites
(-)Biological Unit 1
(-)Biol. Unit 1 - sites
(-)Biological Unit 2
collapse expand < >
Image Asymmetric Unit
Asymmetric Unit  (Jmol Viewer)
Image Asym. Unit - sites
Asym. Unit - sites  (Jmol Viewer)
Image Biological Unit 1
Biological Unit 1  (Jmol Viewer)
Image Biol. Unit 1 - sites
Biol. Unit 1 - sites  (Jmol Viewer)
Image Biological Unit 2
Biological Unit 2  (Jmol Viewer)

(-) Description

Title :  CRYSTAL STRUCTURE OF WHY1 FROM ARABIDOPSIS THALIANA
 
Authors :  L. Cappadocia, J. S. Parent, N. Brisson, J. Sygusch
Date :  12 May 13  (Deposition) - 13 Nov 13  (Release) - 27 Nov 13  (Revision)
Method :  X-RAY DIFFRACTION
Resolution :  1.88
Chains :  Asym. Unit :  A,B,C,D
Biol. Unit 1:  A,B  (2x)
Biol. Unit 2:  C,D  (2x)
Keywords :  Plant, Whirly, Single-Stranded Dna Binding Protein, Dna Binding Protein (Keyword Search: [Gene Ontology, PubMed, Web (Google))
 
Reference :  L. Cappadocia, J. S. Parent, J. Sygusch, N. Brisson
A Family Portrait: Structural Comparison Of The Whirly Proteins From Arabidopsis Thaliana And Solanum Tuberosum.
Acta Crystallogr. , Sect. F V. 69 1207 2013
PubMed-ID: 24192350  |  Reference-DOI: 10.1107/S1744309113028698

(-) Compounds

Molecule 1 - SINGLE-STRANDED DNA-BINDING PROTEIN WHY1, CHLOROPLASTIC
    ChainsA, B, C, D
    EngineeredYES
    Expression SystemESCHERICHIA COLI
    Expression System PlasmidPET21A
    Expression System StrainBL21(DE3)
    Expression System Taxid469008
    Expression System Vector TypePLASMID
    FragmentUNP RESIDUES 74-241
    GeneAT1G14410, F14L17.18, PTAC1, WHY1
    Organism CommonMOUSE-EAR CRESS,THALE-CRESS
    Organism ScientificARABIDOPSIS THALIANA
    Organism Taxid3702
    StrainCOL-0
    SynonymPROTEIN PLASTID TRANSCRIPTIONALLY ACTIVE 1, PROTEIN WHIRLY 1, ATWHY1

 Structural Features

(-) Chains, Units

  1234
Asymmetric Unit ABCD
Biological Unit 1 (2x)AB  
Biological Unit 2 (2x)  CD

Summary Information (see also Sequences/Alignments below)

(-) Ligands, Modified Residues, Ions  (3, 8)

Asymmetric Unit (3, 8)
No.NameCountTypeFull Name
1MES6Ligand/Ion2-(N-MORPHOLINO)-ETHANESULFONIC ACID
2NI1Ligand/IonNICKEL (II) ION
3PO41Ligand/IonPHOSPHATE ION
Biological Unit 1 (2, 8)
No.NameCountTypeFull Name
1MES6Ligand/Ion2-(N-MORPHOLINO)-ETHANESULFONIC ACID
2NI-1Ligand/IonNICKEL (II) ION
3PO42Ligand/IonPHOSPHATE ION
Biological Unit 2 (1, 6)
No.NameCountTypeFull Name
1MES6Ligand/Ion2-(N-MORPHOLINO)-ETHANESULFONIC ACID
2NI-1Ligand/IonNICKEL (II) ION
3PO4-1Ligand/IonPHOSPHATE ION

(-) Sites  (8, 8)

Asymmetric Unit (8, 8)
No.NameEvidenceResiduesDescription
1AC1SOFTWAREARG A:130 , GLN A:131 , TYR A:132 , TRP A:134BINDING SITE FOR RESIDUE MES A 301
2AC2SOFTWAREPHE A:140 , SER A:141 , HIS A:163 , ASP A:164 , PRO A:165 , LYS A:177 , HOH A:432 , HOH A:443 , HOH A:505BINDING SITE FOR RESIDUE MES A 302
3AC3SOFTWAREHIS A:248 , HIS A:250 , HOH A:551 , HIS D:248 , HIS D:250BINDING SITE FOR RESIDUE NI A 303
4AC4SOFTWAREARG B:130 , GLN B:131 , TYR B:132 , TRP B:134BINDING SITE FOR RESIDUE MES B 301
5AC5SOFTWAREPHE B:140 , SER B:141 , HIS B:163 , PRO B:165 , LYS B:177 , HOH B:559BINDING SITE FOR RESIDUE PO4 B 302
6AC6SOFTWAREPHE C:140 , SER C:141 , HIS C:163 , PRO C:165 , LYS C:177BINDING SITE FOR RESIDUE MES C 301
7AC7SOFTWAREARG D:130 , TYR D:132 , TRP D:134 , HOH D:491BINDING SITE FOR RESIDUE MES D 301
8AC8SOFTWAREVAL D:139 , PHE D:140 , SER D:141 , HIS D:163 , PRO D:165 , LYS D:177 , HOH D:437 , HOH D:512BINDING SITE FOR RESIDUE MES D 302

(-) SS Bonds  (0, 0)

(no "SS Bond" information available for 4KOO)

(-) Cis Peptide Bonds  (0, 0)

(no "Cis Peptide Bond" information available for 4KOO)

 Sequence-Structure Mapping

(-) SAPs(SNPs)/Variants  (0, 0)

(no "SAP(SNP)/Variant" information available for 4KOO)

(-) PROSITE Motifs  (0, 0)

(no "PROSITE Motif" information available for 4KOO)

(-) Exons   (0, 0)

(no "Exon" information available for 4KOO)

(-) Sequences/Alignments

Asymmetric Unit
   Reformat: Number of residues per line =  ('0' or empty: single-line sequence representation)
  Number of residues per labelling interval =   
  UniProt sequence: complete  aligned part    
   Show mapping: SCOP domains CATH domains Pfam domains Secondary structure (by author)
SAPs(SNPs) PROSITE motifs Exons
(details for a mapped element are shown in a popup box when the mouse pointer rests over it)
Chain A from PDB  Type:PROTEIN  Length:176
                                                                                                                                                                                                                
               SCOP domains d4kooa_ A: automated matches                                                                                                                                                     SCOP domains
               CATH domains -------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- CATH domains
               Pfam domains -------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- Pfam domains
         Sec.struct. author ....ee..eeee...eeeeeeee..eeee.....eeeee..eeeeeeee.......hhhhheeeeehhhhhhhhhh......eeeee...........eeeeeeeee......eeeeeeeee....eeeeeeeeehhhhhhhhhhhhhhhhhhhhhhhhhhhhhhhhhh....... Sec.struct. author
                 SAPs(SNPs) -------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- SAPs(SNPs)
                    PROSITE -------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- PROSITE
                 Transcript -------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- Transcript
                 4koo A  77 LPARFYVGHSIYKGKAALTVDPRAPEFVALDSGAFKLSKDGFLLLQFAPSAGVRQYDWSKKQVFSLSVTEIGTLVSLGPRESCEFFHDPFKGKSDEGKVRKVLKVEPLPDGSGHFFNLSVQNKLVNVDESIYIPITRAEFAVLISAFNFVLPYLIGWHAFANSIKAAALEHHHHHH 252
                                    86        96       106       116       126       136       146       156       166       176       186       196       206       216       226       236       246      

Chain B from PDB  Type:PROTEIN  Length:162
                                                                                                                                                                                                  
               SCOP domains d4koob_ B: automated matches                                                                                                                                       SCOP domains
               CATH domains ------------------------------------------------------------------------------------------------------------------------------------------------------------------ CATH domains
               Pfam domains ------------------------------------------------------------------------------------------------------------------------------------------------------------------ Pfam domains
         Sec.struct. author .......eeee...eeeeeeee..eeee.....eeeee..eeeeeeee.......hhhhheeeeehhhhhhhhhhh.....eeeee...........eeeeeeeee......eeeeeeeee....eeeeeeeeehhhhhhhhhhhhhhhhhhhhhhhhh... Sec.struct. author
                 SAPs(SNPs) ------------------------------------------------------------------------------------------------------------------------------------------------------------------ SAPs(SNPs)
                    PROSITE ------------------------------------------------------------------------------------------------------------------------------------------------------------------ PROSITE
                 Transcript ------------------------------------------------------------------------------------------------------------------------------------------------------------------ Transcript
                 4koo B  78 PARFYVGHSIYKGKAALTVDPRAPEFVALDSGAFKLSKDGFLLLQFAPSAGVRQYDWSKKQVFSLSVTEIGTLVSLGPRESCEFFHDPFKGKSDEGKVRKVLKVEPLPDGSGHFFNLSVQNKLVNVDESIYIPITRAEFAVLISAFNFVLPYLIGWHAFANS 239
                                    87        97       107       117       127       137       147       157       167       177       187       197       207       217       227       237  

Chain C from PDB  Type:PROTEIN  Length:165
                                                                                                                                                                                                     
               SCOP domains d4kooc_ C: automated matches                                                                                                                                          SCOP domains
               CATH domains --------------------------------------------------------------------------------------------------------------------------------------------------------------------- CATH domains
               Pfam domains --------------------------------------------------------------------------------------------------------------------------------------------------------------------- Pfam domains
         Sec.struct. author ...ee..eeee...eeeeeeee..eeee.....eeeee..eeeeeeee.......hhhhheeeeehhhhhhhhhh......eeeee...........eeeeeeeee......eeeeeeeee....eeeeeeeeehhhhhhhhhhhhhhhhhhhhhhhhh...... Sec.struct. author
                 SAPs(SNPs) --------------------------------------------------------------------------------------------------------------------------------------------------------------------- SAPs(SNPs)
                    PROSITE --------------------------------------------------------------------------------------------------------------------------------------------------------------------- PROSITE
                 Transcript --------------------------------------------------------------------------------------------------------------------------------------------------------------------- Transcript
                 4koo C  78 PARFYVGHSIYKGKAALTVDPRAPEFVALDSGAFKLSKDGFLLLQFAPSAGVRQYDWSKKQVFSLSVTEIGTLVSLGPRESCEFFHDPFKGKSDEGKVRKVLKVEPLPDGSGHFFNLSVQNKLVNVDESIYIPITRAEFAVLISAFNFVLPYLIGWHAFANSIKA 242
                                    87        97       107       117       127       137       147       157       167       177       187       197       207       217       227       237     

Chain D from PDB  Type:PROTEIN  Length:175
                                                                                                                                                                                                               
               SCOP domains d4kood_ D: automated matches                                                                                                                                                    SCOP domains
               CATH domains ------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- CATH domains
               Pfam domains ------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- Pfam domains
         Sec.struct. author ........eeee...eeeeeeee..eeee.....eeeee..eeeeeeee.......hhhhheeeeehhhhhhhhhhh.....eeeee...........eeeeeeeee......eeeeeeeee....eeeeeeeeehhhhhhhhhhhhhhhhhhhhhhhhhhhhhhhhhhhhhee. Sec.struct. author
                 SAPs(SNPs) ------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- SAPs(SNPs)
                    PROSITE ------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- PROSITE
                 Transcript ------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- Transcript
                 4koo D  77 LPARFYVGHSIYKGKAALTVDPRAPEFVALDSGAFKLSKDGFLLLQFAPSAGVRQYDWSKKQVFSLSVTEIGTLVSLGPRESCEFFHDPFKGKSDEGKVRKVLKVEPLPDGSGHFFNLSVQNKLVNVDESIYIPITRAEFAVLISAFNFVLPYLIGWHAFANSIKAAALEHHHHH 251
                                    86        96       106       116       126       136       146       156       166       176       186       196       206       216       226       236       246     

   Legend:   → Mismatch (orange background)
  - → Gap (green background, '-', border residues have a numbering label)
    → Modified Residue (blue background, lower-case, 'x' indicates undefined single-letter code, labelled with number + name)
  x → Chemical Group (purple background, 'x', labelled with number + name, e.g. ACE or NH2)
  extra numbering lines below/above indicate numbering irregularities and modified residue names etc., number ends below/above '|'

 Classification and Annotation

(-) SCOP Domains  (1, 4)

Asymmetric Unit

(-) CATH Domains  (0, 0)

(no "CATH Domain" information available for 4KOO)

(-) Pfam Domains  (0, 0)

(no "Pfam Domain" information available for 4KOO)

(-) Gene Ontology  (14, 14)

Asymmetric Unit(hide GO term definitions)

 Visualization

(-) Interactive Views

Asymmetric Unit
  Complete Structure
    Jena3D(integrated viewing of ligand, site, SAP, PROSITE, SCOP information)
    WebMol | AstexViewer[tm]@PDBe
(Java Applets, require no local installation except for Java; loading may be slow)
    STRAP
(Java WebStart application, automatic local installation, requires Java; full application with system access!)
    RasMol
(require local installation)
    Molscript (VRML)
(requires installation of a VRML viewer; select preferred view via VRML and generate a mono or stereo PDF format file)
 
  Ligands, Modified Residues, Ions
    MES  [ RasMol | Jena3D ]  +environment [ RasMol | Jena3D ]
    NI  [ RasMol | Jena3D ]  +environment [ RasMol | Jena3D ]
    PO4  [ RasMol | Jena3D ]  +environment [ RasMol | Jena3D ]
 
  Sites
    AC1  [ RasMol ]  +environment [ RasMol ]
    AC2  [ RasMol ]  +environment [ RasMol ]
    AC3  [ RasMol ]  +environment [ RasMol ]
    AC4  [ RasMol ]  +environment [ RasMol ]
    AC5  [ RasMol ]  +environment [ RasMol ]
    AC6  [ RasMol ]  +environment [ RasMol ]
    AC7  [ RasMol ]  +environment [ RasMol ]
    AC8  [ RasMol ]  +environment [ RasMol ]
 
  Cis Peptide Bonds
(no "Cis Peptide Bonds" information available for 4koo)
 
Biological Units
  Complete Structure
    Biological Unit 1  [ Jena3D ]
    Biological Unit 2  [ Jena3D ]

(-) Still Images

Jmol
  protein: cartoon or spacefill or dots and stick; nucleic acid: cartoon and stick; ligands: spacefill; active site: stick
Molscript
  protein, nucleic acid: cartoon; ligands: spacefill; active site: ball and stick

 Databases and Analysis Tools

(-) Databases

Access by PDB/NDB ID
  4koo
    Family and Domain InformationProDom | SYSTERS
    General Structural InformationGlycoscienceDB | MMDB | NDB | OCA | PDB | PDBe | PDBj | PDBsum | PDBWiki | PQS | PROTEOPEDIA
    Orientation in MembranesOPM
    Protein SurfaceSURFACE
    Secondary StructureDSSP (structure derived) | HSSP (homology derived)
    Structural GenomicsGeneCensus
    Structural NeighboursCE | VAST
    Structure ClassificationCATH | Dali | SCOP
    Validation and Original DataBMRB Data View | BMRB Restraints Grid | EDS | PROCHECK | RECOORD | WHAT_CHECK
 
Access by UniProt ID/Accession number
  WHY1_ARATH | Q9M9S3
    Comparative Protein Structure ModelsModBase
    Genomic InformationEnsembl
    Protein-protein InteractionDIP
    Sequence, Family and Domain InformationInterPro | Pfam | SMART | UniProtKB/SwissProt
 
Access by Enzyme Classificator   (EC Number)
  (no 'Enzyme Classificator' available)
    General Enzyme InformationBRENDA | EC-PDB | Enzyme | IntEnz
    PathwayKEGG | MetaCyc
 
Access by Disease Identifier   (MIM ID)
  (no 'MIM ID' available)
    Disease InformationOMIM
 
Access by GenAge ID
  (no 'GenAge ID' available)
    Age Related InformationGenAge

(-) Analysis Tools

Access by PDB/NDB ID
    Domain InformationXDom
    Interatomic Contacts of Structural UnitsCSU
    Ligand-protein ContactsLPC
    Protein CavitiescastP
    Sequence and Secondary StructurePDBCartoon
    Structure AlignmentSTRAP(Java WebStart application, automatic local installation, requires Java; full application with system access!)
    Structure and Sequence BrowserSTING
 
Access by UniProt ID/Accession number
  WHY1_ARATH | Q9M9S3
    Protein Disorder PredictionDisEMBL | FoldIndex | GLOBPLOT (for more information see DisProt)

 Related Entries

(-) Entries Sharing at Least One Protein Chain (UniProt ID)

(no "Entries Sharing at Least One Protein Chain" available for 4KOO)

(-) Related Entries Specified in the PDB File

4kop 4koq