Skip to content

chore: 🏗️ add variable-resource map for prescreening form - #222

Open
lwjohnst86 wants to merge 1 commit into
chore/map-var-to-resource-bedqfrom
chore/var-resource-map-prescreening
Open

lwjohnst86 wants to merge 1 commit into
chore/map-var-to-resource-bedqfrom
chore/var-resource-map-prescreening

Conversation

@lwjohnst86

Copy link
Copy Markdown
Member

Description

Add variable to resource mapping for the prescreening form. This one is longer and more work to review.

Needs a thorough review.

@lwjohnst86
lwjohnst86 requested a review from a team as a code owner September 7, 2026 10:24
@lwjohnst86 lwjohnst86 moved this from Todo to In Review in Data development Sep 7, 2026
@lwjohnst86
lwjohnst86 force-pushed the chore/var-resource-map-prescreening branch from 2fe0006 to 174b85b Compare September 8, 2026 13:17
medications,,canagliflozin_dose_v1,,visit_01,besg_1_screening,
medications,,candesartan_dose_v1,,visit_01,besg_1_screening,
admin,cgm_info_v1_admin,cgm_info_v1,,visit_01,besg_1_screening,
qc,cgm_no_qc,cgm_no_v1,,visit_01,besg_1_screening,

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Suggested change
qc,cgm_no_qc,cgm_no_v1,,visit_01,besg_1_screening,
qc,cgm_no_v1_qc,cgm_no_v1,,visit_01,besg_1_screening,

There are a lot of CGMs in this study, we'll need to know which one it is.

Copy link
Copy Markdown
Member Author

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

We'll be rearranging these into long format (with visit number as a column), so this _v* tag isn't necessary. Plus, this is QC, which for now we are not including in the data package.

qc,cgm_no_qc,cgm_no_v1,,visit_01,besg_1_screening,
admin,cgm_v1_admin,cgm_v1,,visit_01,besg_1_screening,
admin,cgmreader_v1_admin,cgmreader_v1,,visit_01,besg_1_screening,
admin,comment_blood_sample_v1_admin,comment_blood_sample_v1,,visit_01,besg_1_screening,

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Not sure this should be admin, it is a full text description of why there may be outlier data regarding samples etc.

Copy link
Copy Markdown
Member Author

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Yea, there's a lot of variables I tagged with admin, but could also be qc. This one I can change to QC. But keep in mind, admin and QC resources won't go into this data package, as they aren't relevant from a research point of view. A researcher's job shouldn't be to figure out why data is missing or not, they should just know that if data is missing, there's a reason for it, it isn't random/a mistake. Our job as the engineers/builders of data packages is to make sure the data is high-quality enough that researchers can use it to answer their research questions effectively.

admin,cgm_v1_admin,cgm_v1,,visit_01,besg_1_screening,
admin,cgmreader_v1_admin,cgmreader_v1,,visit_01,besg_1_screening,
admin,comment_blood_sample_v1_admin,comment_blood_sample_v1,,visit_01,besg_1_screening,
admin,comments_end_v1_admin,comments_end_v1,,visit_01,besg_1_screening,

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

This is the field they use for describing anything that happened during or before the visit, I expect some of the qualitative analysis would like these for reference.

Comment on lines +55 to +56
admin,compliance_risk_expl_v1_admin,compliance_risk_expl_v1,,visit_01,besg_1_screening,
admin,compliance_risk_v1_admin,compliance_risk_v1,,visit_01,besg_1_screening,

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

These two talks about any concerns that the study team may have regarding compliance with the protocol, not sure this is admin or should be study data.

Comment on lines +159 to +160
admin,psychiatric_disease_barrier_reason_v1_admin,psychiatric_disease_barrier_reason_v1,,visit_01,besg_1_screening,
admin,psychiatric_disease_barrier_v1_admin,psychiatric_disease_barrier_v1,,visit_01,besg_1_screening,

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

This is where they detail if a participant fails screening on psychiatric reasons. Probably not strictly speaking research data on such a small population, but I'm not entirely sure it is admin either.

blood,,cpeptid_value_v1,,visit_01,besg_1_screening,
health_conditions,controlled_cvd,cvd_v1,,visit_01,besg_1_screening,
medications,,dapagliflozin_dose_v1,,visit_01,besg_1_screening,
*,,date_v1,,visit_01,besg_1_screening,Not sure how to include this in multiple resources...

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

if it is detailed in the overall ID table then if the fields have an identifier (like _v1) we could get the date from the ID table, and not from here?

@K-Beicher K-Beicher left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

A few comments but overall looks good.

@github-project-automation github-project-automation Bot moved this from In Review to In Progress in Data development Sep 9, 2026
blood,,cpeptid_value_v1,,visit_01,besg_1_screening,
health_conditions,controlled_cvd,cvd_v1,,visit_01,besg_1_screening,
medications,,dapagliflozin_dose_v1,,visit_01,besg_1_screening,
*,,date_v1,,visit_01,besg_1_screening,Not sure how to include this in multiple resources...

Copy link
Copy Markdown
Collaborator

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Independently of Kris' comment, just based on this table, what about having a row for each resource it should go in?

Copy link
Copy Markdown
Member Author

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Hmm, that's an idea 🤔

medications,,thiazid_type_v1,,visit_01,besg_1_screening,
medications,,tirzepatid_dose_v1,,visit_01,besg_1_screening,
admin,training_notes_v1_admin,training_notes_v1,,visit_01,besg_1_screening,
medications,,treatment_v1,,visit_01,besg_1_screening,

Copy link
Copy Markdown
Collaborator

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

treatment_v1 is a checkbox field. In the data this is given as one column for each checkbox option: treatment_v1___1, treatment_v1___2, treatment_v1___3.

How should we handle checkbox fields? Should we match the data and list each option in a separate row? This would be analogous to the checkbox expansion we have been doing so far. It would be more work upfront but it would make it possible to select these columns directly from the raw data when creating the resource.

This branch has not been deployed

No deployments
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

Status: In Progress

Development

Successfully merging this pull request may close these issues.

3 participants