How does each distribution look? Do you notice any potential issues with your data? Examples include lack of representation of groups of interest (e.g., gender, age), skew, etc. Based on this, what issues do you anticipate for testing hypotheses in part 3?