The Comparative Validity of Interactive Multimedia Questionnaires to Paper-Administered Questionnaires for Beverage Intake and Physical Activity: Pilot Study
Bibliographic record
Abstract
BACKGROUND: Brief, valid, and reliable dietary and physical activity assessment tools are needed, and interactive computerized assessments (ie, those with visual cues, pictures, sounds, and voiceovers) can reduce administration and scoring burdens commonly encountered with paper-based assessments. OBJECTIVE: The purpose of this pilot investigation was to evaluate the comparative validity and reliability of interactive multimedia (IMM) versions (ie, IMM-1 and IMM-2) compared to validated paper-administered (PP) versions of the beverage intake questionnaire (BEVQ-15) and Stanford Leisure-Time Activity Categorical Item (L-Cat); a secondary purpose was to evaluate results across two education attainment levels. METHODS: Adults 21 years or older (n=60) were recruited to complete three laboratory sessions, separated by three to seven days in a randomly assigned sequence, with the following assessments-demographic information, two IMM and one paper-based (PP) version of the BEVQ-15 and L-Cat, health literacy, and an IMM usability survey. RESULTS: Responses across beverage categories from the IMM-1 and PP versions (validity; r=.34-.98) and the IMM-1 and IMM-2 administrations (reliability; r=.61-.94) (all P<.001) were significantly correlated. Paired t tests revealed significant differences in sugar-sweetened beverage (SSB) grams and kcal (P=.02 and P=.01, respectively) and total beverage kcal (P=.03), on IMM-1 and IMM-2; however, comparative validity was demonstrated between IMM-2 and the PP version suggesting familiarization with the IMM tool may influence participant responses (mean differences: SSB 63 grams, SEM 87; P=.52; SSB 21 kcal, SEM 33; P=.48; total beverage 65 kcal, SEM 49; P=.19). Overall mean scores between the PP and both IMM versions of the L-Cat were different (both P<.001); however, responses on all versions were correlated (P<.001). Differences between education categories were noted at each L-Cat administration (IMM-1: P=.008; IMM-2: P=.001; PP: P=.002). Major and minor themes from user feedback suggest that the IMM questionnaires were easy to complete, and relevant to participants' typical beverage choices and physical activity habits. CONCLUSIONS: In general, less educated participants consumed more total beverage and SSB energy, and reported less engagement in physical activity. The IMM BEVQ-15 appears to be a valid and reliable measure to assess habitual beverage intake, although software familiarization may increase response accuracy. The IMM-L-Cat can be considered reliable and may have permitted respondents to more freely disclose actual physical activity levels versus the paper-administered tool. Future larger-scale investigations are warranted to confirm these possibilities.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.005 | 0.004 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.002 | 0.000 |
| Scholarly communication | 0.000 | 0.002 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".