Academic Libraries Can Develop AI Chatbots for Virtual Reference Services with Minimal Technical Knowledge and Limited Resources
Notice bibliographique
Résumé
A Review of: Rodriguez, S., & Mune, C. (2022). Uncoding library chatbots: Deploying a new virtual reference tool at the San Jose State University Library. Reference Services Review, 50(3), 392-405. https://doi.org/10.1108/RSR-05-2022-0020 Objective – To describe the development of an artificial intelligence (AI) chatbot to support virtual reference services at an academic library. Design – Case study. Setting – A public university library in the United States. Subjects – 1,682 chatbot-user interactions. Methods – A university librarian and two graduate student interns researched and developed an AI chatbot to meet virtual reference needs. Developed using chatbot development software, Dialogflow, the chatbot was populated with questions, keywords, and other training phrases entered during user inquiries, text-based responses to inquiries, and intents (i.e., programmed mappings between user inquiries and chatbot responses). The chatbot utilized natural language processing and AI training for basic circulation and reference questions, and included interactive elements and embeddable widgets supported by Kommunicate (i.e., a bot support platform for chat widgets). The chatbot was enabled after live reference hours were over. User interactions with the chatbot were collected across 18 months since its launch. The authors used analytics from Kommunicate and Dialogflow to examine user interactions. Main Results – User interactions increased gradually since the launch of the chatbot. The chatbot logged approximately 44 monthly interactions during the spring 2021 term, which increased to approximately 137 monthly interactions during the spring 2022 term. The authors identified the most common reasons for users to engage the chatbot, using the chatbot’s triggered intents from user inquiries. These reasons included information about hours for the library building and live reference services, finding library resources (e.g., peer-reviewed articles, books), getting help from a librarian, locating databases and research guides, information about borrowing library items (e.g., laptops, books), and reporting issues with library resources. Conclusion – Libraries can successfully develop and train AI chatbots with minimal technical expertise and resources. The authors offered user experience considerations from their experience with the project, including editing library FAQs to be concise and easy to understand, testing and ensuring chatbot text and elements are accessible, and continuous maintenance of chatbot content. Kommunicate, Dialogflow, Google Analytics, and Crazy Egg (i.e., a web usage analytics tool) could not provide more in-depth user data (e.g., user clicks, scroll maps, heat maps), with plans to further explore other usage analysis software to collect the data. The authors noted that only 10% of users engaged the chatbot beyond the initial welcome prompt, requiring more research and user testing on how to facilitate user engagement.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,012 | 0,034 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,001 |
| Méta-épidémiologie (sens large) | 0,001 | 0,001 |
| Bibliométrie | 0,010 | 0,009 |
| Études des sciences et des technologies | 0,002 | 0,001 |
| Communication savante | 0,005 | 0,009 |
| Science ouverte | 0,002 | 0,004 |
| Intégrité de la recherche | 0,002 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,050 | 0,035 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».