{"id":"W4402442473","doi":"10.1145/3650212.3680332","title":"Semantic Constraint Inference for Web Form Test Generation","year":2024,"lang":"en","type":"article","venue":"","topic":"Software Testing and Debugging Techniques","field":"Computer Science","cited_by":1,"is_retracted":false,"has_abstract":true,"ca_institutions":"University of British Columbia","funders":"","keywords":"Computer science; Inference; Constraint (computer-aided design); Test (biology); Natural language processing; Artificial intelligence; Semantic Web; Information retrieval; Programming language; Mathematics; Geology","routes":{"ca_aff":true,"ca_fund":false,"ca_venue":false,"about_ca":false,"invisible_to_affiliation_only":false},"retraction":null,"screen":null,"direct_labels":[],"prediction":{"model_version":"metacan-v3-hybrid-931329e0061c","candidate_categories":[],"consensus_categories":[],"category_scores_codex":[0.002993048,0.001365778,0.0006829661,0.002259502,0.0006532948,0.001621489,0.002416485,0.001129668,0.006786807],"category_scores_gemma":[0.02355283,0.000601497,0.001861292,0.001296713,0.001964178,0.003178251,0.002443818,0.001896218,0.001315887],"about_ca_system_candidate":false,"about_ca_system_consensus":false,"about_ca_system_score_codex":0.001817009,"about_ca_system_score_gemma":0.003324518,"about_ca_topic_candidate":false,"about_ca_topic_consensus":false,"about_ca_topic_score_codex":0.008224616,"about_ca_topic_score_gemma":0.01768301,"domain_scores_codex":[0.9949326,0.002251221,0.0002746811,0.0006698532,0.001614862,0.000256786],"domain_scores_gemma":[0.9878163,0.008533084,0.0005852893,0.001478405,0.001409044,0.0001779458],"domain_codex":null,"domain_gemma":null,"domain_candidate":null,"domain_consensus":null,"study_design_codex":"design_other","study_design_gemma":"simulation_or_modeling","study_design_scores_codex":[0.0004951877,0.0003877885,0.008283139,0.001022848,0.0001755033,0.001051501,0.0005521163,0.380519,0.01517763,0.06970873,0.026736,0.4958905],"study_design_scores_gemma":[0.00005298916,0.00004217576,0.0003710771,0.00005760033,0.00002892399,0.0001742421,0.00008020637,0.9399783,0.0128715,0.038788,0.00753277,0.00002231924],"study_design_candidate":"simulation_or_modeling","study_design_consensus":null,"genre_codex":"methods","genre_gemma":"methods","genre_scores_codex":[0.01841316,0.0001573495,0.958282,0.0006430744,0.00005336124,0.0002757861,0.001415553,0.01716557,0.003594179],"genre_scores_gemma":[0.3029477,0.0001787123,0.6879247,0.0005234727,0.00004220715,0.0003001821,0.004342515,0.001970533,0.001769902],"genre_candidate":"methods","genre_consensus":"methods","teacher_disagreement_score":0.008224616,"threshold_uncertainty_score":0.02270406,"prediction_status":"machine_predicted_unvalidated"},"machine_scores":{"provisional":true,"baseline":true,"maturity_gate_passed":false,"score_opus":0.04611346186274883,"score_gpt":0.3088652585492415,"score_spread":0.2627517966864926,"validation_status":"score_only:v0-immature-baseline","note":"Baseline scores from an immature model (maturity gate not passed). Scores rank; they never assert a category."}}