Applying Probability and Data with R to All of Us Datasets

 Applying Probability and Data with R to All of Us Datasets

Show more...

This badge represents completion of an HSLS Self-Paced Learning module, published by the Health Sciences Library System at the University of Pittsburgh. Participants who earn this badge should be able to: - Recall simple functions in R for exploring any dataset.  - Define the five main data domains in the All of Us database. - Classify AoU variables as categorical or numerical. - Calculate summary statistics for both categorical and numerical variables.  - Apply data cleaning procedures to datasets in R, including filtering, removing missing values (NA), and managing duplicates.  - Derive new variables from existing datasets.  - Generate a new dataframe by merging various datasets.  - Use random sampling techniques to accurately represent target distributions.  - Describe the data preparation processes involved in the 'All of Us' code templates developed by HSLS. 

Issuer

University of Pittsburgh

Issued On

1777896000

Expires

-
check_circle Badge assertion is valid
check_circle Badge has not been revoked
check_circle Badge has not expired

Check that the badge was awarded to the correct recipient by entering their email address:

 

If you have the original badge PDF file you can validate it here:


Validate another badge