Notice bibliographique
Résumé
For savvy Web users everywhere, there are now so many metadata initiatives that it is difficult to gain a clear understanding of what they all are and what their role is in the broader information science picture. The fact that these are usually represented by acronyms does not help, nor does the fact that although many become norms, it is not immediately clear which have this status and which do not. In addition, metadata initiatives involve many other communities than the information science community, and what the roles of each are is not always clear. Although the visual aid we are creating will necessarily be untidy and imperfect, we nevertheless hope to contribute to understanding of the relationships among metadata sets and other related information in this new and very necessary area of knowledge. The MetaMap Project is an attempt to sort out the very many efforts worldwide in working out norms and other information about metadata sets in information science. The goal is to produce a study aid that will help people interested in information science and technology to understand how the plethora of metadata initiatives spawned by the arrival of the World Wide Web are related to one another and to information studies. Sponsored by the Visual Information Research Group (GRIV) at the Université de Montréal, the project attempts to show relationships among the various norms, as well as to other pertinent information. It represents these metadata initiatives using the conventions of a subway map to help the user navigate this space, learning heavily on the conventions of the London Underground Map, noted for its clarity in helping users sort out complex reality. Each norm, metadata set, organization or other element is represented as a station on a line that has a theme. At present, lines representing themes include processes of information management: Creation, Organization, Dissemina-tion, Preservation; institutions with expertise in information management: Libraries, Archives, Museums; types of digital documentation: Text, Still Images, Moving Images, Sound. Organizations deeply involved in Web activity and metadata norms, such as the World Wide Web Consortium, OCLC, the IETF, the IEEE, and so on, are included on a separate “subway” line. Elements that have repercussions in a number of areas are represented as nodes in the “subway” network. For example, SMIL, the Synchronized Multimedia Integration Language, is common to the lines representing Text, Still Images, Moving Images, and Sound. Simpler nodes represent intersections of themes, for example where the lines representing Libraries, Archives and Museums cross the Organizations line, the nodes are respectively IFLA (International Federation of Library Associations and Institutions), ICA (International Council on Archives), and ICOM (International Council of Museums). Branch lines are created where norms such as XML or the Dublin Core spawn other norms, as a way of showing these relationships. Within a line, an attempt to order the stations to show relationships among them is made. Because of the complexity of the representation, no one criterion can be used for ordering all the lines, but those so far adopted have to do with various types of conceptual relationships among the norms such as the purpose for which they were created, their relatedness within the theme of the subway line, or chronological development (genesis) of them. Faced with the impossibility of ever being able to represent the information as clearly as we would like to within this structure, we have adopted the compromise philosophy of “well, it's much better than nothing.” To date, projected products of the initiative include a color poster which will be printed in English on one side and in French on the other. In French, the MetaMap is called MétroMéta. The other major product is a Web version of the map—actually two, one in English and one in French. SVG (Scalable Većtor Graphics) is being used to develop the Web versions, as it offers user help in navigating the information space such as zooming in and out, moving the visible area around the screen, and searching on individual acronyms. When the user passes the mouse over a station name, usually an acronym, a popup window displays the expanded name and gives other useful information such as the purpose of the metadata initiative and who is sponsoring it. Clicking on the name opens a new window and takes the user to the official Web site for the norm, metadata initiative, project, or organization. Where there is no official site, the click takes the user to the best or most complete available Web source of information. The nature of the project is such that it can never be completed. New initiatives spring up every day, it seems. Nor will the technological environment that creates the need for all this metadata activity probably ever completely settle. The inherent instability of the wonderful world of metadata means that keeping the content of our map up to date will be quite a challenge. For the moment, the best we can plan for is periodic updates, but these in turn will be dependent on the availability of human resources to carry them out. In order to be able to arrive at a product quickly, we are developing the visual representation of the MetaMap as an SVG graphic. However, what is needed in the long term is a database that will generate the map automatically from the records it contains. This will make use of exciting possibilities offered by XML and SVG as these norms themselves are developed and as we learn to use them creatively. Nevertheless, we feel that the MetaMap is a good start on organizing the information about the world we are trying to describe, and we hope it will help users in various fields of endeavor to gain an understanding of the many elements that make up that world. The Web version of the MetaMap can be found at the following site: http://mapageweb.umontreal.ca/turner/ Grateful acknowledgement is made to CoRIMedia, a consortium for research in image, video and multimedia indexing and navigation, which is financed by Valorisation-Recherche Québec, an initiative of the government of Québec, and industrial partners. CoRIMedia in turn is financing this project.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,013 | 0,026 |
| Méta-épidémiologie (sens strict) | 0,003 | 0,002 |
| Méta-épidémiologie (sens large) | 0,002 | 0,005 |
| Bibliométrie | 0,008 | 0,009 |
| Études des sciences et des technologies | 0,003 | 0,003 |
| Communication savante | 0,015 | 0,018 |
| Science ouverte | 0,008 | 0,018 |
| Intégrité de la recherche | 0,005 | 0,006 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,077 | 0,082 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».