{"id":"W2100247763","doi":"10.5555/1182635.1164182","title":"Similarity search: a matching based approach","year":2006,"lang":"en","type":"article","venue":"","topic":"Data Management and Algorithms","field":"Computer Science","cited_by":39,"is_retracted":false,"has_abstract":true,"ca_institutions":"University of Toronto","funders":"","keywords":"Nearest neighbor search; Similarity (geometry); Computer science; Matching (statistics); Object (grammar); Set (abstract data type); Curse of dimensionality; k-nearest neighbors algorithm; Data mining; Information retrieval; Mathematics; Artificial intelligence; Image (mathematics)","routes":{"ca_aff":true,"ca_fund":false,"ca_venue":false,"about_ca":false,"invisible_to_affiliation_only":false},"retraction":null,"screen":null,"direct_labels":[],"prediction":{"model_version":"metacan-v3-hybrid-931329e0061c","candidate_categories":[],"consensus_categories":[],"category_scores_codex":[0.002852505,0.0009557569,0.002428279,0.006248432,0.001674486,0.003912831,0.005267284,0.003796278,0.006787254],"category_scores_gemma":[0.01009417,0.0006751071,0.001781624,0.009226812,0.002296627,0.009454555,0.004152725,0.001982937,0.003385792],"about_ca_system_candidate":false,"about_ca_system_consensus":false,"about_ca_system_score_codex":0.001854427,"about_ca_system_score_gemma":0.001817591,"about_ca_topic_candidate":false,"about_ca_topic_consensus":false,"about_ca_topic_score_codex":0.003073362,"about_ca_topic_score_gemma":0.001624132,"domain_scores_codex":[0.9929507,0.00184552,0.0004914786,0.001430226,0.002992886,0.0002893124],"domain_scores_gemma":[0.9974908,0.001013841,0.0001645704,0.0005713793,0.0006264062,0.0001330789],"domain_codex":null,"domain_gemma":null,"domain_candidate":null,"domain_consensus":null,"study_design_codex":"design_other","study_design_gemma":"not_applicable","study_design_scores_codex":[0.0002623055,0.0004104379,0.001725393,0.0008626839,0.0002870971,0.0004027182,0.0004946709,0.04807879,0.006806426,0.3548755,0.01964732,0.5661466],"study_design_scores_gemma":[0.0001036414,0.0004029233,0.0009687631,0.0001850415,0.000173262,0.001893308,0.0004192079,0.4531498,0.006211148,0.4554287,0.08094673,0.0001174999],"study_design_candidate":"not_applicable","study_design_consensus":null,"genre_codex":"methods","genre_gemma":"methods","genre_scores_codex":[0.003582017,0.003382588,0.9820062,0.0009715826,0.0002091211,0.0002732844,0.0001599323,0.0005422798,0.008872991],"genre_scores_gemma":[0.1645725,0.005125431,0.8136652,0.0009206874,0.0007053725,0.0004142999,0.0009156212,0.0002298735,0.01345093],"genre_candidate":"methods","genre_consensus":"methods","teacher_disagreement_score":0.006787254,"threshold_uncertainty_score":0.02270555,"prediction_status":"machine_predicted_unvalidated"},"machine_scores":{"provisional":true,"baseline":true,"maturity_gate_passed":false,"score_opus":0.02780876414292226,"score_gpt":0.2324174349816486,"score_spread":0.2046086708387264,"validation_status":"score_only:v0-immature-baseline","note":"Baseline scores from an immature model (maturity gate not passed). Scores rank; they never assert a category."}}