MétaCan
Menu
Back to cohort
Record W4255233508 · doi:10.4242/balisagevol1.levy01

Beyond the Semantic Web: the Semantic Space

2009· article· en· W4255233508 on OpenAlexaff
Pierre Lévy

Bibliographic record

VenueBalisage series on markup technologies · 2009
Typearticle
Languageen
FieldComputer Science
TopicSemantic Web and Ontologies
Canadian institutionsUniversity of Ottawa
Fundersnot available
KeywordsComputer scienceParsingSemantics (computer science)Semantic WebSemantic Web Rule LanguageSpace (punctuation)Social Semantic WebVariety (cybernetics)Natural languageArtificial intelligenceWorld Wide WebNatural language processingSemantic analyticsProgramming language

Abstract

fetched live from OpenAlex

Today, the sharing of semantics remains a conundrum. Semantics can be shared within a universe of discourse, but individuals and communities cannot be relieved of the need to define their own universes of discourse. The emergence of collective intelligence is increasingly seen as necessary for human survival, but it is difficult for people who live in diverse universes of discourse to know when they are talking about the same things. Collective intelligence — the ability of a community to exhibit self-sustaining, rational behaviors — is inversely related to its participants' ability to understand each other. Diverse minds can create, recognize, and think in terms of diverse sets of distinct concepts and relationships between them. A conceptual addressing system can map such sets into a shared abstract "semantic space" that is structured by an algebraically definable group of transformations. Information Economy Meta Language (IEML) is such a "semantic space addressing system"; it defines a very large space of semantic addresses. A small number of the points in that space — more than 2,500 of them — are now listed in an "IEML Dictionary", along with interpretations of each of them in several natural languages. A language for compactly specifying sets of locations in the space exists, and a parser that translates expressions in this language into XML is available. A programming language for discovering and asserting relationships between sets of semantics is being developed, along with a variety of related software tools. The semantic space research program could provide a scientific (measurable, principled, experimentally repeatable) foundation on which technologies and professional disciplines can be created, including distributed collaborative semantic search engines, models and simulations of collective intelligences, tools and editorial practices for the automated production of multimedia documents, and many more.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.010
metaresearch head score (Gemma)0.015
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Theoretical or conceptual · Consensus signal: Theoretical or conceptual
GenreCandidate signal: Empirical · Consensus signal: none
Teacher disagreement score0.026
Threshold uncertainty score0.053

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0100.015
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0020.002
Bibliometrics0.0060.010
Science and technology studies0.0050.020
Scholarly communication0.0260.080
Open science0.0040.009
Research integrity0.0070.009
Insufficient payload (model declined to judge)0.0090.003

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.010
GPT teacher head0.225
Teacher spread0.215 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designTheoretical or conceptual
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations3
Published2009
Admission routes1
Has abstractyes

Explore more

Same venueBalisage series on markup technologiesSame topicSemantic Web and OntologiesFrench-language works237,207