Work Motivation Under Lockdown Exploratory Data Analysis (EDA)

Two datasets...

  1. Post-workday evaluations of work motivation related stuff. Large spread across how many time points people collected; few decent-ish ones.

  2. Traditional questionnaires from the same participants.

    • Allocation indicates whether they took part in our motivation self-management training during the period or belonged to a waitlist control group.
    • Pre-, mid- and post-questionnaires --> TIME == 1, 2 and 3 respectively.
    • Everything from intrinsic1 to Psycap5 are raw answers. From Intrinsic to Optimism are sumscores.
    • This was a feasibility & acceptability study so it's not powered for or aiming to look at outcomes. But if there was one, it could be the "Relative Autonomy Index" (RAI), a sumscore conventionally built as:
      • Amotivation -3 + ExternalTotal -2 + Introjected -1 + Identified 2 + Intrinsic * 3
    • If I remember correctly, there were some people whose RAI was boosted a lot
    • In general:
      • Good things to go up during the intervention (or life in general):
        • Intrinsic, Identified
        • Competence, Relatedness, Autonomy
      • Bad things to go up during the intervention (or life in general):
        • Amotivation, External (material & social), Introjected
        • AutonomyThwarting, RelatednessThwarting, CompetenceThwarting
          • Unsure if these were named in the data, or just some of the items named autonomyx/relatednessx/competencex

This EDA only addresses the daily data, specifically data=="data/moti_feasibility_james_daily.csv"

Show the session information of the packages used in this analysis

Read in the raw daily data

Reshape the data frame and convert data types to numeric where possible

Count the amount of rows per user and plot

The aim of this is to assess whether there is enough data to do some of the more complext NLTSA methods

Plot the heatmaps of the top 5 users with the highest amount of rows

The aim of this is to assess the amount of missingness time wise (i.e. the user could have a lot of data, spread over a long period with large gaps in the middle).

Show the heatmap of the missing values for the last user plotted above

Based on the above plots it seems like there is some missingness that needs to be imputed

Take an example user and impute the missingness in their data

There are a few options, which we should think about and choose

Plot the recurrence plot of an example user and an example feature

Plot the imputations of an example feature per user

and calculate the fluctuation intensity, distribution uniformity, complexity resonance and cumulative complexity peaks data frames from the scaled imputed data

Plot the complexity resonance diagrams and cumulative complexity peaks plots per user