SYSTEM DYNAMICS AND GOAL-ORIENTED MEASUREMENT: A HYBRID APPROACH
Bibliographic record
Abstract
Goal-oriented measurement following the Goal/Question/Metric (GQM) approach is a well-defined and powerful tool in software management and decision-support. This chapter proposes the integration of GQM with a mature software process simulation approach, System Dynamics, in order to further enhance software managers' analytic, explorative, and decision-making capability. The proposed hybrid approach, which we denote “Dynamic GQM”, overcomes limitations that exist if applying GQM and system dynamics in isolation. It offers a new dimension of support to managers and decisionmakers by integrating traditional goal-oriented measurement and static modeling with the newly emerging paradigm of software process simulation and dynamic modeling. The hybrid approach is holistic by nature, i.e. it takes a global perspective on decision making in contrast to the local perspective advocated by traditional GQM. The proposed approach combines individual GQM plans into one consistent model and adds timedynamic behavior on top of it, thus offering a comprehensive view on what is actually happening in software projects.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".