MétaCan
Menu
Back to cohort
Record W2083622218 · doi:10.1017/s1360674309003025

The rise of<i>it</i>-clefting in English: areal-typological and contact-linguistic considerations

2009· article· en· W2083622218 on OpenAlexaboutno aff
Markku Filppula

Bibliographic record

VenueEnglish Language and Linguistics · 2009
Typearticle
Languageen
FieldSocial Sciences
TopicLinguistic Variation and Morphology
Canadian institutionsnot available
Fundersnot available
KeywordsGermanCeltic languagesLinguisticsHistoryNorth Germanic languagesGermanic languagesGrammarRomance languagesTRACE (psycholinguistics)Language contactDanishLiteratureArtPhilosophy

Abstract

fetched live from OpenAlex

Recent areal and typological research has brought to light several syntactic features which English shares with the Celtic languages as well as some of its neighbouring western European languages, but not with (all of) its Germanic sister languages, especially German. This study focuses on one of them, viz. the so-called it -cleft construction. What makes the it -cleft construction particularly interesting from an areal and typological point of view is the fact that, although it does not belong to the defining features of so-called Standard Average European (SAE), it has a strong presence in French, which is in the ‘nucleus’ of languages forming SAE alongside Dutch, German, and (northern dialects of) Italian. In German, however, clefting has remained a marginal option, not to mention most of the eastern European languages which hardly make use of clefting at all. This division in itself prompts the question of some kind of a historical-linguistic connection between the Celtic languages (both Insular and Continental), English, and French (or, more widely, Romance languages). Before tackling that question, one has to establish whether it -clefting is part of Old (and Middle) English grammar, and if so, to what extent it is used in these periods. In the first part of this article (sections 2 and 3), I trace the emergence of it -clefts on the basis of data from The York–Toronto–Helsinki Corpus of Old English Prose and The Penn–Helsinki Parsed Corpus of Middle English , second edition. Having established the gradually increasing use of it -clefts from OE to ME, I move on to discuss the areal distribution of clefting among European languages and its typological implications (section 4). This paves the way for a discussion of the possible role played by language contacts, and especially those with the Celtic languages, in the emergence of it -clefting in English (section 5). It is argued that contacts with the Celtic languages provide the most plausible explanation for the development of this feature of English. This conclusion is supported by the chronological precedence of the cleft construction in the Celtic languages, its prominence in modern-period ‘Celtic Englishes’, and close parallels between English and the Celtic languages with respect to several other syntactic features.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.001
metaresearch head score (Gemma)0.147
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Theoretical or conceptual · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: none
Teacher disagreement score0.734
Threshold uncertainty score0.861

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0010.147
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0000.000
Science and technology studies0.0000.000
Scholarly communication0.0000.000
Open science0.0000.000
Research integrity0.0000.000
Insufficient payload (model declined to judge)0.0000.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.018
GPT teacher head0.304
Teacher spread0.286 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

Study designTheoretical or conceptual
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations23
Published2009
Admission routes1
Has abstractyes

Explore more

Same venueEnglish Language and LinguisticsSame topicLinguistic Variation and MorphologyFrench-language works237,207