{"id":"W2016901333","doi":"10.1145/2479440.2482677","title":"Issues in big data testing and benchmarking","year":2013,"lang":"en","type":"article","venue":"","topic":"Data Quality and Management","field":"Decision Sciences","cited_by":29,"is_retracted":false,"has_abstract":true,"ca_institutions":"","funders":"European Institute of Innovation and Technology; University of Toronto; Deutsche Forschungsgemeinschaft; International Business Machines Corporation","keywords":"Benchmarking; Computer science; Big data; Scalability; Data science; Relational database; Data modeling; Data management; Database; Massively parallel; Volume (thermodynamics); Data mining; Parallel computing","routes":{"ca_aff":false,"ca_fund":true,"ca_venue":false,"about_ca":false,"invisible_to_affiliation_only":true},"retraction":null,"screen":null,"direct_labels":[],"prediction":{"model_version":"metacan-v3-hybrid-931329e0061c","candidate_categories":["metaresearch"],"consensus_categories":[],"category_scores_codex":[0.1545355,0.001507794,0.002321382,0.002632736,0.00252826,0.01052502,0.01062681,0.004396233,0.002507489],"category_scores_gemma":[0.3991296,0.00125992,0.001460282,0.007166184,0.007885386,0.01483966,0.005611114,0.005497478,0.001447589],"about_ca_system_candidate":false,"about_ca_system_consensus":false,"about_ca_system_score_codex":0.002911874,"about_ca_system_score_gemma":0.005957555,"about_ca_topic_candidate":false,"about_ca_topic_consensus":false,"about_ca_topic_score_codex":0.004214721,"about_ca_topic_score_gemma":0.002470536,"domain_scores_codex":[0.7748648,0.1553738,0.01341141,0.009121399,0.04362692,0.003601631],"domain_scores_gemma":[0.5498267,0.289373,0.01200561,0.09339369,0.04934018,0.006060833],"domain_codex":null,"domain_gemma":"methods","domain_candidate":"methods","domain_consensus":null,"study_design_codex":"design_other","study_design_gemma":"theoretical_or_conceptual","study_design_scores_codex":[0.001361823,0.001865567,0.04111477,0.00294474,0.000479689,0.001063458,0.003456547,0.1391547,0.006811608,0.2869593,0.09846684,0.4163209],"study_design_scores_gemma":[0.0003860627,0.000995066,0.008857262,0.001528628,0.0001069168,0.001057515,0.003971195,0.2811022,0.01364682,0.5916104,0.0965057,0.0002322459],"study_design_candidate":"theoretical_or_conceptual","study_design_consensus":null,"genre_codex":"methods","genre_gemma":"empirical","genre_scores_codex":[0.08604337,0.007536033,0.7755306,0.0795797,0.003497296,0.002378248,0.002321151,0.009115739,0.03399784],"genre_scores_gemma":[0.4714389,0.00170655,0.5062715,0.009312663,0.00155314,0.00258634,0.002649889,0.002521621,0.001959331],"genre_candidate":"empirical","genre_consensus":null,"teacher_disagreement_score":0.8454645,"threshold_uncertainty_score":0.8172714,"prediction_status":"machine_predicted_unvalidated"},"machine_scores":{"provisional":true,"baseline":true,"maturity_gate_passed":false,"score_opus":0.6107454406852676,"score_gpt":0.4649125596602122,"score_spread":0.1458328810250554,"validation_status":"score_only:v0-immature-baseline","note":"Baseline scores from an immature model (maturity gate not passed). Scores rank; they never assert a category."}}