Complete the form by entering the relevant Dataset Details in each of the fields listed below. You cannot save the Dataset Details without completing all required fields. Note that when you first save a dataset it is only visible to you and local Private Data Enclave Administrators. However, you will have the option to edit Dataset Details or Dataset Access once you view the saved dataset. Dataset files can also be uploaded at that time.
Dataset Type (required): Choose between Generic, DICOM, and REDCap for the dataset type. The default type is Generic.
Dataset Title (required): Enter the name of your dataset.
Description (required): Type or paste in your dataset description.
Keywords (required): Enter a list of words or phrases that people can use to search for this dataset.
Storage Organization (required, not editable): The default value will be the organization associated with the parent project. Note that this is also the organization where the dataset details and dataset files will be stored. If you want to store data at another organization, someone at that organization must create a new project and the dataset can be added there.
Show Storage Organization Logo (required): This checkbox is selected as a default. However, you may choose to remove that selection if the project you are adding is multi-institutional, represents an external resource, or in any other circumstance where displaying your organization's logo on it is not appropriate or meaningful.
HIPAA Identifiers (required): Enter patient data types captured. If the data does not come from medical records, enter N/A. Note: This value cannot be changed for DICOM dataset objects after the initial save.
IRB Number (required if you indicate the dataset contains Highly Sensitive Data): Select the IRB # under which use of this data is approved. This field will present you with a list of active IRB protocols where you are listed as a member of the study team in your local IRB database.
Associated Data Security Plan (required if you indicate the dataset contains Highly Sensitive Data): Associating a DSP with a dataset provides the team with easy access to documentation that tells them where they are allowed to store the data and who it can be transferred to. This step is required for HSD because it is essential that team members with this type of data have access to this document so they can ensure compliance.
Data Granularity (required for data derived from medical records): Select the level of granularity of the dataset using either Row Level Data, Data Aggregated to Groups of 11 or more, or Other for unique cases such as imaging files, time series, etc.
Other Sensitive Data (required): Select which types of other sensitive data are contained in this dataset, such as HIV Status, Psychiatric, Financial, Employment, Student, Consumer, Criminal, Additional Status, Sexual Health, Gender Identity, Personally Identifiable Information, Social Security Number, Controlled Unclassified Information (CUI), and Other Sensitive.
Associated Contract (optional): If the project involves data sharing with another institution, then a contract should be uploaded to the project and associated with any datasets covered by that agreement.
License (optional): Specify a license that should be associated with this dataset (i.e. Creative Commons license).
Variables Measured (optional): This list of variables tells collaborators what data elements the dataset contains.
DOI (optional): A DOI is a Digital Object Identifier. If you are indexing a dataset that already has a DOI, you can add it here. If you publish a Dataset File to the Public Commons, a DOI will be created for you and this field will auto populate.
Data Attribution (optional): This is a free text field where you can describe appropriate attribution language for any future reuse of the dataset.
Link to External Dataset File (optional): Datasets can be indexed by entering Dataset Details even when the dataset file is stored elsewhere. Paste a link to an external website or file path where the file is stored if appropriate. Note that you can save files directly in the Private Data Commons by uploading a Dataset File after saving Dataset Details.
Select Cancel to discard any entries you have made in the form. Select Save to create and view the new dataset. Upon clicking save, the system will notify you if any of the entries are not valid. Once all elements are valid, the system will then prompt you to confirm that no Personally Identifiable Information (PII) has been included in any of the text fields.
You will need to select Edit Dataset Access from the View Dataset page if you need to add access for team members or if you want to make Dataset Details visible in the Public Data Commons.
Upon saving, the dataset record will be created in your local Private Data Enclave. The new dataset will be assigned a unique ID which you will see in the URL after you save the new dataset. You can share the URL with other people who have access to this dataset in the Private Data Enclave. NOTE: If you choose to make the Dataset Details publicly visible, a second URL for the dataset's view in the Public Data Commons will be available. This URL can be discovered by right-clicking on the "View in Public Data Commons" link when viewing the dataset in the Private Data Enclave. Use the Public Data Commons URL when sharing the dataset outside of your project team.