{"id":"W4283323653","doi":"10.1145/3544790","title":"Toward More Efficient Statistical Debugging with Abstraction Refinement","year":2022,"lang":"en","type":"article","venue":"ACM Transactions on Software Engineering and Methodology","topic":"Software Testing and Debugging Techniques","field":"Computer Science","cited_by":2,"is_retracted":false,"has_abstract":true,"ca_institutions":"University of Waterloo","funders":"Office of Naval Research; Natural Science Foundation of Jiangsu Province; National Natural Science Foundation of China; National Science Foundation","keywords":"Debugging; Computer science; Algorithmic program debugging; Abstraction; Programming language; Pruning; Discriminative model; Software engineering; Machine learning","routes":{"ca_aff":true,"ca_fund":false,"ca_venue":false,"about_ca":false,"invisible_to_affiliation_only":false},"retraction":null,"screen":null,"direct_labels":[],"prediction":{"model_version":"metacan-v3-hybrid-931329e0061c","candidate_categories":[],"consensus_categories":[],"category_scores_codex":[0.0103191,0.001516928,0.001426284,0.002760945,0.0009093215,0.001784139,0.003148324,0.001030779,0.001131913],"category_scores_gemma":[0.04409249,0.001050193,0.001767394,0.00188672,0.001763011,0.004075697,0.003718386,0.003197089,0.0007225902],"about_ca_system_candidate":false,"about_ca_system_consensus":false,"about_ca_system_score_codex":0.0009712569,"about_ca_system_score_gemma":0.004262997,"about_ca_topic_candidate":false,"about_ca_topic_consensus":false,"about_ca_topic_score_codex":0.002662492,"about_ca_topic_score_gemma":0.004811287,"domain_scores_codex":[0.9844958,0.006229115,0.0009504119,0.001520658,0.00594931,0.000854819],"domain_scores_gemma":[0.9642211,0.01646041,0.00327577,0.01123442,0.00444098,0.0003672913],"domain_codex":null,"domain_gemma":null,"domain_candidate":null,"domain_consensus":null,"study_design_codex":"design_other","study_design_gemma":"simulation_or_modeling","study_design_scores_codex":[0.0004855957,0.0004074275,0.0191598,0.0006066352,0.0003169816,0.0006652528,0.001280406,0.2639868,0.07003285,0.09568154,0.006852462,0.5405242],"study_design_scores_gemma":[0.00008183279,0.0002103275,0.001687762,0.00009864586,0.0001150667,0.0003622162,0.0001322076,0.8957353,0.02301731,0.0720583,0.006444765,0.00005614976],"study_design_candidate":"simulation_or_modeling","study_design_consensus":null,"genre_codex":"methods","genre_gemma":"empirical","genre_scores_codex":[0.007685265,0.0001368255,0.9894302,0.000157604,0.00001411993,0.00006014936,0.00003695213,0.002163657,0.0003152491],"genre_scores_gemma":[0.2017989,0.0002408159,0.7957786,0.0002828134,0.00004325046,0.0001998899,0.0003258806,0.0007381718,0.00059171],"genre_candidate":"empirical","genre_consensus":null,"teacher_disagreement_score":0.0103191,"threshold_uncertainty_score":0.05457324,"prediction_status":"machine_predicted_unvalidated"},"machine_scores":{"provisional":true,"baseline":true,"maturity_gate_passed":false,"score_opus":0.07935506400377887,"score_gpt":0.3095148429020107,"score_spread":0.2301597788982318,"validation_status":"score_only:v0-immature-baseline","note":"Baseline scores from an immature model (maturity gate not passed). Scores rank; they never assert a category."}}