Evaluation of ASSERT-PV V3R1 against the PSBT Benchmark
Bibliographic record
Abstract
Void fraction and DNB calculations conducted using ASSERT-PV V3R1 are evaluated against data from the NUPEC database as part of the OECD/NEA Pressurized Water Reactor Subchannel Benchmark Tests (PSBT). Void fraction measurements were well represented in the isolated single subchannel cases, with 77.0% of all predicted values falling within ±2σexp =0.06 of the experimental value. In the B5 type bundle, an average void fraction error of ϵ-α=-0.0540 was reported at the lower elevation, while this value was ϵ-α=-0.0405 at the upper measurement location. ASSERT was able to predict the steady state DNB power of the bundles to within ±10% of the measured value for a total of 344 times out of 432. Sensitivity studies conducted indicate that the Ahmad correlation with the Groeneveld 1995 CHF lookup table yielded the most accurate results, although some data points fell within the limiting quality region where the accuracy was reduced.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.005 | 0.015 |
| Meta-epidemiology (narrow) | 0.002 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.004 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.008 | 0.003 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".