Optimization and Evaluation of Spoken English CAF Based on Artificial Intelligence and Corpus
Notice bibliographique
Résumé
English is the most widely used language in the world, and the pronunciation of its spoken language is equally important. The traditional methods are not high in complexity, accuracy and fluency (CAF) for spoken English recognition. Therefore, it is very important to use AI and corpus to optimize and evaluate spoken English CAF. This paper aims to study the optimization and evaluation of spoken English CAF using AI and corpus, and proposes to use the Hidden Markov (HMM) model and convolutional neural network (CNN) model in the field of AI to optimize and evaluate spoken English CAF. By selecting a variety of English voices from the BNC corpus for model training and testing, and selecting the complexity, accuracy, fluency and harmonic average of the CNN model recognition as evaluation indicators, the HMM model's recognition spectrogram is added up and analyzed. In the experimental test, it was found that when the number of frames is 210, the indicators of the CNN model have been greatly improved, so the number of frames selected for the test in this paper is 210. The results show that the A value obtained by the HMM model test is about 85%, the CNN model is 67%, and the traditional SVM model is only 35%. The HMM model is tested with a C value of about 60%, the CNN model is 65%, and the traditional model is only 45%. The F-value obtained from the test of the HMM model is about 83%, the CNN model is 67%, and the traditional model is 46%. In contrast, the HMM model has higher recognition accuracy for spoken English, and the recognition results are more fluent. However, the CNN model can recognize spoken English with higher complexity, and both the CNN model and the HMM model can improve the CAF optimization effect of spoken English.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction distillée sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.
Scores Codex et Gemma par catégorie
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,005 | 0,008 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,000 |
| Méta-épidémiologie (sens large) | 0,000 | 0,000 |
| Bibliométrie | 0,001 | 0,001 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,000 | 0,001 |
| Science ouverte | 0,000 | 0,000 |
| Intégrité de la recherche | 0,000 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,000 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule tête enseignante, pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».