Measuring the Effectiveness of Student Aid, 2005-2008 [Canada]
Notice bibliographique
Résumé
The Measuring the Effectiveness of Student Aid (MESA) dataset comprises a sample of low income students receiving student financial aid in 2006-07. Students were contacted first (Cycle I) in February-May of that academic year (the precise date varying by province), and were then followed up in 2007-08 (Cycle II), contacted in February-April of that year. Students will be contacted again in 2008-09 for the last time. The dataset represents a national sample, including all provinces-except for Prince Edward Island. In the spring and summer of 2005, the Canada Millennium Scholarship Foundation negotiated a series of agreements with provincial governments to deliver a set of bursaries (known as “Access Bursaries”) to first-time, first-year undergraduates from low-income families. These agreements are all broadly similar though eligibility criteria vary slightly by jurisdiction (section 1, below, describes the Access Bursaries as they exist in each province). Students do not need to apply for the award separately; instead, they are automatically considered for the award through their application for provincial student assistance. The sample represents a particular subset of the students who received student financial aid in their first year of post-secondary education in 2006-07. In the majority of the provinces, this subset consists of the students who received a Low Income Bursary from the Millennium Scholarship Foundation. In British Columbia and Nova Scotia, a control group made up of students who received financial aid but not the Millennium Bursary was surveyed as well. The Ontario sample is made up those Millennium Bursary recipients who also received a Canada Access Grant and those who did not, with sub-samples selected from each group (all appear together in the data but can be separately identified). The Bursary and the Grant are awarded in similar amounts, but the eligibility requirements are different. This dataset was freely received from the Canada Millennium Scholarship Foundation. Some work was required for the variable and value labels, and missing values. They were corrected as best as possible with the documentation received. Caution should be used with this dataset as some variables are lacking information.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,002 | 0,013 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,000 |
| Méta-épidémiologie (sens large) | 0,001 | 0,001 |
| Bibliométrie | 0,003 | 0,012 |
| Études des sciences et des technologies | 0,002 | 0,001 |
| Communication savante | 0,003 | 0,001 |
| Science ouverte | 0,002 | 0,001 |
| Intégrité de la recherche | 0,000 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,008 | 0,003 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».