A corpus-based network analysis of onomastic references in 16th- and 17th century British grammar writing
The present study investigates who was considered authoritative in matters of language in the 16th and 17th centuries as well as how grammar authors position themselves with respect to these authorities. It evaluates whether a shift in referencing norms may already be observed from the 16th to the 17th century in these regards.
As it is claimed that in England the 16th century marks the beginning of English grammar writing (McCarthy 2020McCarthy, Michael 2020 Innovations and Challenges in Grammar. London: Routledge. McCarthy, Michael 2020 Innovations and Challenges in Grammar. London: Routledge. : 19–20) and the 17th century saw a shift in favour of English being recognized as a separate academic discipline (Beal 2004Beal, Joan C. 2004 English in Modern Times. London: Arnold.Beal, Joan C. 2004 English in Modern Times. London: Arnold.: 102), one can ask if and how onomastic — that is, name-based — references, appear in those grammars and how they can be categorized.
One of the findings of this study is that the onomastic references found in the 16th- and 17th-century grammars can be categorized along the six semantic categories suggested in previous work for 19th-century grammars (Busse et al. 2020 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. : 11–12). These include, for example, quotations, opinions, or mere mentions. To account for a possible shift in reference strategies over time, these semantic categories were reevaluated by means of inter-rater reliability (IRR) in the 16th- and 17th-century context.
Our main findings show that while some 16th-century authors put emphasis on Latinate authors, others embrace moving away from the Latinate approach by not referring to the established Latinate authorities at all or only rarely. A significant shift in onomastic referencing can be observed from the 16th to the 17th century, however the Latinate authorities still held significant ground in the 18th century grammar texts.
Publication history
1.Introduction
And not so much as but euen Quintilian that great writing, and speaking master wisheth sound to be obserued, as the surest teacher to write right, and not custom.(Mulcaster 1582Mulcaster, Richard 1582 The First Part of the Elementarie Which Entreateth Chefelie of the Right Writing of our English Tung. London: Thomas Vautroullier.Mulcaster, Richard 1582 The First Part of the Elementarie Which Entreateth Chefelie of the Right Writing of our English Tung. London: Thomas Vautroullier.: 94, emphasis by the author)
This reference to the ancient Roman rhetorician Quintilian is peculiar in a twofold way. The 16th-century grammar author, Richard Mulcaster, not only indirectly quotes Quintilian by stating his wishes, but also clarifies his own position towards the Roman scholar. By calling him “that great writing, and speaking master” he acknowledges Quintilian’s reputation and significance to the field. Onomastic references, that is, name-based references to specific persons (Busse et al. 2020 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. ), such as the one displayed in the example, are the central component of this study because of their frequency, their variation, and their function within the tradition of British grammar writing. The examination of onomastic references offers a unique perspective on language usage, sociolinguistics, and the cultural nuances embedded in linguistic artifacts of the Early modern English period.
This study aims to compare the influential figures referenced by grammar authors in the 16th and 17th centuries, highlighting the sources they used that contributed to shaping the scholarly tradition of English grammar writing. This shows to what extent authors of grammar books are aware of one another’s work, ideas, and shortcomings and which other figures they consider to be worthy of mentioning and whose work or musings they want to build upon.
This study is part of the larger HeidelGram project, which combines corpus-based historical linguistics (see Jucker & Taavitsainen 2014Jucker, Andreas H. & Irma Taavitsainen 2014 “Diachronic Corpus Pragmatics: Intersections and interactions”. Diachronic Corpus Pragmatics ed. by Irma Taavitsainen, Andreas H. Jucker & Jukka Tuominen, 3–26. Amsterdam & Philadelphia: John Benjamins. Jucker, Andreas H. & Irma Taavitsainen 2014 “Diachronic Corpus Pragmatics: Intersections and interactions”. Diachronic Corpus Pragmatics ed. by Irma Taavitsainen, Andreas H. Jucker & Jukka Tuominen, 3–26. Amsterdam & Philadelphia: John Benjamins. ) and network analytical approaches (see Freeman 2004Freeman, Linton C. 2004 The Development of Social Network Analysis: A study in the sociology of science. Vancouver: Empirical Press.Freeman, Linton C. 2004 The Development of Social Network Analysis: A study in the sociology of science. Vancouver: Empirical Press.; White 2011White, Howard D. 2011 “Scientific and Scholarly Networks”. In The SAGE Handbook of Social Network Analysis ed. by John Scott & Peter J. Carrington, 271–285. London: SAGE.White, Howard D. 2011 “Scientific and Scholarly Networks”. In The SAGE Handbook of Social Network Analysis ed. by John Scott & Peter J. Carrington, 271–285. London: SAGE.). The HeidelGram project investigates how and why grammarians refer to one another as well as other influential figures, how they employ evaluative strategies across linguistic topics and grammatical concepts over time, while also critically reassessing assumptions about the interplay between historical English norms, actual usage, and changing attitudes toward prescriptivism and descriptivism.
In previous studies, the practice of onomastic referencing has been investigated in a pilot corpus of 19th-century grammars from England (Busse et al. 2018Busse, Beatrix, Kirsten Gather & Ingo Kleiber 2018 “Assessing the Connections Between English Grammarians of the Nineteenth Century: A corpus-based network analysis”. Grammar and Corpora 2016 ed. by Eric Fuß, Marek Konopka, Beata Trawiński & Ulrich H. Waßner, 435–442. Heidelberg: Heidelberg University Publishing.Busse, Beatrix, Kirsten Gather & Ingo Kleiber 2018 “Assessing the Connections Between English Grammarians of the Nineteenth Century: A corpus-based network analysis”. Grammar and Corpora 2016 ed. by Eric Fuß, Marek Konopka, Beata Trawiński & Ulrich H. Waßner, 435–442. Heidelberg: Heidelberg University Publishing., 2019 2019 “Paradigm Shifts in 19th-Century British Grammar Writing: A network of texts and authors”. Norms and Conventions in the History of English ed. by Birte Bös & Claudia Claridge, 49–71. Amsterdam & Philadelphia: John Benjamins. 2019 “Paradigm Shifts in 19th-Century British Grammar Writing: A network of texts and authors”. Norms and Conventions in the History of English ed. by Birte Bös & Claudia Claridge, 49–71. Amsterdam & Philadelphia: John Benjamins. , 2020 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. ). The corpus and network analyses performed on the 19th-century data have revealed a turn away from traditional, prescriptivist writing in the first part of the 19th century towards describing language in the latter part of the century. (Kleiber et al. 2022Kleiber, Ingo, Beatrix Busse, Lyubomira Dimitrova, Sophie Du Bois & Julia Marcus 2022 SimpleCorpusNetwork [Computer software]. Cologne: University of Cologne. https://github.com/heidelgram/HGSimpleCorpusNetworkKleiber, Ingo, Beatrix Busse, Lyubomira Dimitrova, Sophie Du Bois & Julia Marcus 2022 SimpleCorpusNetwork [Computer software]. Cologne: University of Cologne. https://github.com/heidelgram/HGSimpleCorpusNetwork). This shift was observable in the amount and style of referencing chosen by the authors, as well as the choice of authors they referred to. In the Busse et al. (2020) 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. study, the focus was also on analysing further how reference is made to other names. For that purpose, six main categories of references were defined, namely: (1) quotation, (2) opinion, (3) comparison/contrast, (4) acknowledgement, (5) mention, and (6) application. An example taken from the 17th-century dataset of a reference of type ‘opinion’ is Howell’s praise of Ben Jonson:
Mr. Ben. Johnson a great Wit, who was as patient as he was elaborat in his re- sherches and compositions, as he was framing an English Syntaxis, confess’d the further he proceeded, the more he was puzzled.(Howell 1662Howell, James 1662 A New English Grammar. London: T. Williams, H. Brome, and H. Marsh.Howell, James 1662 A New English Grammar. London: T. Williams, H. Brome, and H. Marsh.: 80)
Here, Howell clearly outlines his own stance on the matter, with intentional subjectivity. It can be deduced that Howell not only appreciates Jonson’s meticulousness in his work but also admires Jonson’s honesty in expressing his perplexity when working on advanced English syntax. He also writes about Jonson’s character, referring to his patience and intelligence. The full list and definitions of the reference categories are outlined in more detail below.
The preliminary studies on onomastic references in the 19th-century grammars are focused on references to other grammar authors. One of the results is that only few references are made to grammars written before 1750, while many refer to authors of major works from the 1750s onward. Furthermore, Busse et al. (2019) 2019 “Paradigm Shifts in 19th-Century British Grammar Writing: A network of texts and authors”. Norms and Conventions in the History of English ed. by Birte Bös & Claudia Claridge, 49–71. Amsterdam & Philadelphia: John Benjamins. 2019 “Paradigm Shifts in 19th-Century British Grammar Writing: A network of texts and authors”. Norms and Conventions in the History of English ed. by Birte Bös & Claudia Claridge, 49–71. Amsterdam & Philadelphia: John Benjamins. found that the grammarians refer to one another not necessarily merely as authorities on language, but references are also used to put forth the author’s own thoughts and distance themselves from others. Since the 16th century marks the beginning of English grammar writing the present study investigates whether these practices of referencing were already established in these early days. The present study investigates the following research questions: (1) Which types of persons were considered authoritative in matters of language in the 16th and 17th centuries? (2) What linguistic means did the grammarians at the time employ to position themselves with regards to these authorities? (3) Can we observe a shift in these matters from the 16th to the 17th century?
For the grammar books under investigation in this study, we have the following hypotheses: (a) Fewer or no references are made to other authors of English grammars, as they were only beginning to emerge during the period in question. (b) Instead, other authorities may be referred to, such as influential figures from religion or politics due to Scripture and scholarly texts written by statesmen being more available and carrying authority in academic circles. (c) The previously established categories for references (see Busse et al. 2020 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. ) can be applied to the onomastic references extracted from the 16th- and 17th-century grammars.
The study approaches the above research questions adopting the mixed methods approach proposed in Busse et al. (2018Busse, Beatrix, Kirsten Gather & Ingo Kleiber 2018 “Assessing the Connections Between English Grammarians of the Nineteenth Century: A corpus-based network analysis”. Grammar and Corpora 2016 ed. by Eric Fuß, Marek Konopka, Beata Trawiński & Ulrich H. Waßner, 435–442. Heidelberg: Heidelberg University Publishing.Busse, Beatrix, Kirsten Gather & Ingo Kleiber 2018 “Assessing the Connections Between English Grammarians of the Nineteenth Century: A corpus-based network analysis”. Grammar and Corpora 2016 ed. by Eric Fuß, Marek Konopka, Beata Trawiński & Ulrich H. Waßner, 435–442. Heidelberg: Heidelberg University Publishing., 2019 2019 “Paradigm Shifts in 19th-Century British Grammar Writing: A network of texts and authors”. Norms and Conventions in the History of English ed. by Birte Bös & Claudia Claridge, 49–71. Amsterdam & Philadelphia: John Benjamins. 2019 “Paradigm Shifts in 19th-Century British Grammar Writing: A network of texts and authors”. Norms and Conventions in the History of English ed. by Birte Bös & Claudia Claridge, 49–71. Amsterdam & Philadelphia: John Benjamins. , 2020 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. ), combining methods from corpus linguistics and network analysis. Seven person categories are suggested based on grounded theory (see Bryant & Charmaz 2007 eds. 2007 The SAGE Handbook of Grounded Theory. London: SAGE. eds. 2007 The SAGE Handbook of Grounded Theory. London: SAGE. ), namely: (1) grammar author, (2) ancient scholar, (3) poet/literary author, (4) political figure, (5) religious figure, (6) contemporary scholar, and (7) none. The previously established reference categories (see Busse et al. 2020 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. ) are evaluated for their accuracy and applicability to the 16th- and 17th-century components of the HeidelGram corpus, spanning five and 17 grammar books, respectively using an inter-rater reliability approach. Once both sets of categories are evaluated and established, networks of onomastic references in the 16th- and 17th-century grammars are created using citation networks (White 2011White, Howard D. 2011 “Scientific and Scholarly Networks”. In The SAGE Handbook of Social Network Analysis ed. by John Scott & Peter J. Carrington, 271–285. London: SAGE.White, Howard D. 2011 “Scientific and Scholarly Networks”. In The SAGE Handbook of Social Network Analysis ed. by John Scott & Peter J. Carrington, 271–285. London: SAGE.). The resulting networks are analysed to suggest how the authors interact, make use of the practice of referencing to criticize or support each other’s works, and how ‘well-connected’ they are in a network-analytical sense. By extending our work to the 16th- and 17th-century grammar books, we are able to assess whether our established framework of categories is pertinent when adding grammar books from the 18th century to the historical corpus.
2.English grammar writing
2.1Defining grammar
As noticed by Mitchell, grammar books during the Early Modern English period covered a wide range of language-related material. The term grammar could mean anything from hard-word lists, spelling, pronunciation, synonyms, homonyms, etymology, and Latin- English dictionaries to poetry, logic, rhetoric, and Scripture lessons. For centuries grammar had been part of a classical education, forming the trivium with logic and rhetoric. Because of grammar’s traditional authority, it held the right to make language decisions (1994Mitchell, Linda C. 1994 “Inversion of Grammar Books and Dictionaries in the Seventeenth and Eighteenth Centuries”. In Euralex 94 Proceedings, 548–554.Mitchell, Linda C. 1994 “Inversion of Grammar Books and Dictionaries in the Seventeenth and Eighteenth Centuries”. In Euralex 94 Proceedings, 548–554.: 548–549). This view of grammar is applicable when considering the span of the Early Modern English period ranging approximately from the 16th-17th century, yet the definition of what constitutes a grammar for 16th and 17th century grammar authors differs.
Dons (2004Dons, Ute 2004 Descriptive Adequacy of Early Modern English Grammars. Berlin & New York: De Gruyter. Dons, Ute 2004 Descriptive Adequacy of Early Modern English Grammars. Berlin & New York: De Gruyter. : 4) claims that prior to the 16th century, a book was only referred to as a grammar when it dealt with the Latin language. Additionally, due to being overshadowed by Latin and French for centuries, writing about the grammar of English was considered uncommon even after the publication of so-called vernacular grammars. This is illustrated in Howell’s prologue addressing the reader: “Now, touching this new English Grammar, let not the Reader mistake, as if it were an English Grammar to learn another Language, as Lillie is for Latin” (1662: xxx).
Nonetheless, in the 16th century, authors attempted to define what grammar means for the English language: “For the first and chef pooint in Grammar for English iz too know what part of spech euery word in euery sentenc iz” (Bullokar 1586Bullokar, Wiliam 1586 Brief Grammar for English. London: Edmund Bollifant.Bullokar, Wiliam 1586 Brief Grammar for English. London: Edmund Bollifant.: 18–19).
By focusing on the individual sentence constituents in English, Bullokar is making an early attempt to create an English grammar framework which is distinct from the Latin model and create a bridge between the medieval tradition of Latin-based education to emphasis on the vernacular language. Bullokar emphasizing that grammar is not just concerned with the correctness of language but is a fundamental tool for understanding the structure and function of language using a practical and analytic approach, aligns with the Humanist approaches of the Renaissance period that aimed to study and revive classical antiquity (see Law 2015Law, Vivien 2015 The History of Linguistics in Europe: From Plato to 1600. Cambridge: Cambridge University Press.Law, Vivien 2015 The History of Linguistics in Europe: From Plato to 1600. Cambridge: Cambridge University Press.; Heath 1971Heath, Terrence 1971 “Logical Grammar, Grammatical Logic, and Humanism in Three German Universities”. Studies in the Renaissance 18. 9–64. Heath, Terrence 1971 “Logical Grammar, Grammatical Logic, and Humanism in Three German Universities”. Studies in the Renaissance 181. 9–64. ).
To understand the shift towards viewing grammar as a way of using language in a proper way, it is necessary to contextualise the process of standardisation unfolding throughout this period. In the late 16th century, the codification stage begins, which develops through the 17th century with Early Modern English grammars that are largely descriptive or normative rather than prescriptive, laying down rules in grammars and dictionaries as authoritative guides (Nevalainen & Tieken-Boon van Ostade 2006Nevalainen, Terttu & Ingrid Tieken-Boon van Ostade 2006 “Standardisation”. A History of the English Language ed. by Richard M. Hogg & David Denison, 271–311. Cambridge: Cambridge University Press. Nevalainen, Terttu & Ingrid Tieken-Boon van Ostade 2006 “Standardisation”. A History of the English Language ed. by Richard M. Hogg & David Denison, 271–311. Cambridge: Cambridge University Press. ). The prescriptive stage of standardisation does not fully emerge until the late 18th century and is often marked by Lowth’s Short Introduction to the English Language (1762Lowth, Robert 1762 A Short Introduction to English Grammar. London: J. Hughs.Lowth, Robert 1762 A Short Introduction to English Grammar. London: J. Hughs.). Lowth’s grammar is distinctive for explicitly identifying and criticising grammatical errors, even in the works of respected authors, drawing on earlier codification efforts to establish a norm of correct usage (ibid.). Vorlat (1998Vorlat, Emma 1998 “Criteria of Grammaticalness in 16th- and 17th-Century English Grammar”. A Reader in Early Modern English ed. by Mats Rydén, Ingrid Tieken-Boon van Ostade & Merja Kytö, 485–496. Frankfurt am Main: Lang.Vorlat, Emma 1998 “Criteria of Grammaticalness in 16th- and 17th-Century English Grammar”. A Reader in Early Modern English ed. by Mats Rydén, Ingrid Tieken-Boon van Ostade & Merja Kytö, 485–496. Frankfurt am Main: Lang.: 485–486) proposes a threefold distinction between: (1) descriptive grammars that record language without value judgments and including all language varieties, (2) normative grammars that still focus on language usage but favour how particular social or regional groups use language for pedagogical purposes, and (3) prescriptive grammars that impose rules based on logical or other criteria rather than actual usage. These different approaches shape how the grammar book is written due to the stance the author takes.
In the 17th century, the educational and prescriptive aims of the time began to broaden. This is reflected in Ben Jonson’s The English Grammar, where he states: Grammar is the art of true, and well speaking a Language: the writing is but an Accident (Jonson 1640Jonson, Ben 1640 The English Grammar.Jonson, Ben 1640 The English Grammar.: 35).
Jonson’s emphasis on rhetorical precision, clarity, and refinement in language use, rather than focusing on parts of speech and didactics, shows how grammar begins to be viewed less as a practical tool to systematically analyse linguistic features, but cultivating intellect.
As there has never been one consistent unified definition of the term grammar, the present study combines a conventional definition, that is “the rules and conventions of everyday language” (McCarthy 2020McCarthy, Michael 2020 Innovations and Challenges in Grammar. London: Routledge. McCarthy, Michael 2020 Innovations and Challenges in Grammar. London: Routledge. : 4) with the notion of a book which is intended to instruct the learner in the ways of using language according to time-period-specific norms. This view enables both a systematic approach to scrutinizing patterns in language and which linguistic conventions have been broadly accepted as a standardized form, whilst also acknowledging that discursive practices may be shaped by social, cultural, and even political factors at the time when the books were written.
2.2Previous work on 19th-century English grammars
In an effort to trace changing and stable language norms, attitudes towards language, and discursive practices in a diachronic way, the HeidelGram project aims to compile and analyse a corpus of English grammars spanning the 16th to 19th centuries. By combining corpus and network analyses, new perspectives on which figures and schools of thought influenced grammarians of the Early Modern English period are offered and qualitative assumptions on the history of grammar writing may be reassessed.
In a pilot study, a corpus of 19th-century grammars from England (40 texts, approximately 2.6 million words) was compiled and analysed. The selection of the texts was based on the criteria described in Busse et al. (2020 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. : 137–138), which are revisited in the methodology below. The focus of this study was grammarians’ references to each other in their grammar books and what purpose these references served. The addition of networks to the historical corpus analysis not only provides a new mode of visualization but also carries further its own suite of measures and scores, such as degree centrality. It also allows for investigation of individual authors by means of ego networks.
An example of an onomastic reference of the type ‘application’ is the following excerpt from Abbott’s 1871 grammar: The following table will be sufficient to illustrate Grimm’s law: […] (Abbott 1871, emphasis added).
Based on the frequency and variety of such name-based references, Busse et al. (2020) 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. introduced categories for the types of references that are being made. Six main categories were established using a grounded theory approach, which are listed and described in Table 1. For more detailed information see (Busse et al. 2020 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. : 143–144).
| Reference type | Description |
|---|---|
| Quotation | A citation of text passages from other grammar books (grammarians). |
| Opinion | An expression of positive or negative evaluation. |
| Comparison/Contrast | Comparison and contrast of grammarians’ approaches, terminologies, etc. |
| Acknowledgement | A general reference to grammarians or their works. |
| Mention | A simple mention of a grammarian without significant context or evaluation. |
| Application | The application of a rule or concept introduced by and named after a grammarian. |
For the present study, the reference types are understood more broadly. This means that not only references made to other grammarians will be considered, but also to other types of persons, which are to be categorized. For instance, for the reference type Quotation we decided to not only consider citations from other grammar books, but rather any other works in general, including philosophical texts, the Bible, etc.
Both corpus as well as network analyses allowed for initial findings on the types of references made as well as how the practices of referencing have changed, especially in the sense of evaluation. The analysis of the reference network and reference categories during the 19th century has revealed a marked change of both the quantity and the types of references made around 1850. This change indicates a turn away from an evaluative stance in the early decades to a more descriptive approach later in the century.
This was also reflected in who was being referenced. While most references are made by the grammarians who have published their works before 1850 and to the so-called prescriptivists such as Robert Lowth and Lindley Murray, the later grammarians make significantly fewer references and refer more commonly to contemporary descriptive grammarians in the later decades of the 19th century.
2.3The character of 16th- and 17th-century English grammars
Grammaticography, which refers to producing grammar books and dictionaries for specific languages (see Wolf 2015 2015 “English Grammaticography as Discourse Tradition: Comments on 18th-century developments”. In Anglistentag 2014 Hannover ed. by Rainer Emig & Jana Gohrisch, 19–33. Trier: Wissenschaftlicher Verlag. 2015 “English Grammaticography as Discourse Tradition: Comments on 18th-century developments”. In Anglistentag 2014 Hannover ed. by Rainer Emig & Jana Gohrisch, 19–33. Trier: Wissenschaftlicher Verlag.; Anderwald 2016Anderwald, Lieselotte 2016 Language between Description and Prescription: Verb and verb categories in nineteenth century grammars of English. Oxford: Oxford University Press. Anderwald, Lieselotte 2016 Language between Description and Prescription: Verb and verb categories in nineteenth century grammars of English. Oxford: Oxford University Press. ), developed in Western language sciences from the Graeco-Latin grammatical tradition (Auroux 1994Auroux, Sylvain 1994 La révolution technologique de la grammatisation. Liège: Mardaga.Auroux, Sylvain 1994 La révolution technologique de la grammatisation. Liège: Mardaga.). Raby and Andrieu (2018)Raby, Valérie & Wilfrid Andrieu 2018 “Norms and Rules in the History of Grammar: French and English handbooks in the seventeenth century”. Standardising English: Norms and margins in the history of the English language ed. by Linda Pillière, Wilfrid Andrieu, Valérie Kerfelec & Diana Lewis, 65–88. Cambridge: Cambridge University Press. Raby, Valérie & Wilfrid Andrieu 2018 “Norms and Rules in the History of Grammar: French and English handbooks in the seventeenth century”. Standardising English: Norms and margins in the history of the English language ed. by Linda Pillière, Wilfrid Andrieu, Valérie Kerfelec & Diana Lewis, 65–88. Cambridge: Cambridge University Press. outline that the Renaissance period was a turning point for the process of grammatication, as prior to the late Middle Ages few scholars were concerned with grammar writing. However, the widespread creation of grammars and dictionaries during this period became feasible due to the metalinguistic framework inherited from Graeco-Latin antiquity, although the use of categories established by the Latin model has often been viewed as “an artificial imposition, as a hindrance to valid descriptions of the vernaculars” (ibid.: 67).
The 16th century is said to mark the beginning of English grammar writing (Linn 2006Linn, Andrew 2006 “English Grammar Writing”. The Handbook of English Linguistics ed. by Bas Aarts & April McMahon, 72–92. Oxford: Blackwell. Linn, Andrew 2006 “English Grammar Writing”. The Handbook of English Linguistics ed. by Bas Aarts & April McMahon, 72–92. Oxford: Blackwell. ). It is therefore also a sign of a shifted interest in the varieties of English practised at the time which coincides with the onset of standardization processes of English during that century. As for grammar writing, we can witness a transition from the term grammar only referring to Latin grammatical phenomena towards a more inclusive terminology, allowing grammar to also display features and phenomena of the so-called vernacular languages — including the variety of English used. The reason for this development towards English grammar was due to the shift in the cultural climate of the Renaissance, Reformation, and Humanism (Dons 2004Dons, Ute 2004 Descriptive Adequacy of Early Modern English Grammars. Berlin & New York: De Gruyter. Dons, Ute 2004 Descriptive Adequacy of Early Modern English Grammars. Berlin & New York: De Gruyter. ).
The Renaissance period (late 15th to early 17th century) sparked an interest in classical antiquity and a desire for surpassing the achievements from that era by, for instance, refining the English language which was far less standardized in its pronunciation and spelling than Latin. The forerunners of the Reformation movement challenged perceived discrepancies and lack of transparency on behalf of the Catholic Church and yearned for the general public to understand clerical matters. The Bible is translated from Latin into English resulting in those who could read being able to read the Scriptures in their vernacular language. Humanism encouraged use of English for secular purposes. Another contributing factor to the emergence of English grammars was the invention of the printing press with books becoming more broadly accessible to more people than only those capable of reading and writing in Latin and it also led to a slow process of the standardization of orthography (see Blair 2010Blair, Ann M. 2010 Too Much to Know: Managing scholarly information before the Modern Age. New Haven, CT: Yale University Press.Blair, Ann M. 2010 Too Much to Know: Managing scholarly information before the Modern Age. New Haven, CT: Yale University Press.).
A widely held belief among scholars (e.g., Linn 2006Linn, Andrew 2006 “English Grammar Writing”. The Handbook of English Linguistics ed. by Bas Aarts & April McMahon, 72–92. Oxford: Blackwell. Linn, Andrew 2006 “English Grammar Writing”. The Handbook of English Linguistics ed. by Bas Aarts & April McMahon, 72–92. Oxford: Blackwell. ; McCarthy 2020McCarthy, Michael 2020 Innovations and Challenges in Grammar. London: Routledge. McCarthy, Michael 2020 Innovations and Challenges in Grammar. London: Routledge. ; Nevalainen 2006Nevalainen, Terttu 2006 An Introduction to Early Modern English. Edinburgh: Edinburgh University Press. Nevalainen, Terttu 2006 An Introduction to Early Modern English. Edinburgh: Edinburgh University Press. ; Wolf 2011Wolf, Göran 2011 Englische Grammatikschreibung 1600–1900: Der Wandel einer Diskurstradition. (= Arbeiten zur Sprachanalyse, 54). Frankfurt am Main: Lang.Wolf, Göran 2011 Englische Grammatikschreibung 1600–1900: Der Wandel einer Diskurstradition. (=Arbeiten zur Sprachanalyse, 54). Frankfurt am Main: Lang.) is that the first grammar book of English written in English emerged in 1586. It is William Bullokar’s Bref Grammar for English. However, we found that in our conceptualisation of what constitutes a grammar book, namely both a set of rules and conventions governing everyday language and a guidebook instructing learners based on time-period-specific norms, Sherry’s A Treatise of the Figures of Grammer and Rhetorike (1577Sherry, Richard 1577 A Treatise of the Figures of Grammer and Rhetorike. London: Ricardi Totteli.Sherry, Richard 1577 A Treatise of the Figures of Grammer and Rhetorike. London: Ricardi Totteli.) and Mulcaster’s The First Part of the Elementarie Which Entreateth Chefelie of the Right Writing of our English Tung (1582), both of which were published before, could be considered early examples of grammar books. This is because they emphasized the importance of observing and documenting a uniform grammar system, contributing to the standardization efforts of grammar in the Early Modern period. Algeo (1985)Algeo, John 1985 “The Earliest English Grammars”. Historical & Editorial Studies in Medieval & Early Modern English: For Johan Gerritsen ed. by Mary-Jo Arn, Hanneke Wirtjes & Hans Jansen, 191–207. Groningen: Wolters-Noordhoff.Algeo, John 1985 “The Earliest English Grammars”. Historical & Editorial Studies in Medieval & Early Modern English: For Johan Gerritsen ed. by Mary-Jo Arn, Hanneke Wirtjes & Hans Jansen, 191–207. Groningen: Wolters-Noordhoff. states that the Pamphlet and the Bref Grammar for English have mistakenly been treated as two works due to a binding error of the Bodleian copy of the book. The Pamphlet was modelled after William Lily’s (1534) work on Latin grammar, Rudimenta Grammatices (Usmonovna 2021Usmonovna, Latipova Umida 2021 “The History of English Grammar”. International Journal of Engineering and Information Systems 5:2. 194–198.Usmonovna, Latipova Umida 2021 “The History of English Grammar”. International Journal of Engineering and Information Systems 5:2. 194–198.). According to Le Prieult (2015Le Prieult, Henri 2015 “The ‘Stranger’ and the Grammarian: When Early English Grammarians Reached Out”. Caliban 54.309–326. Le Prieult, Henri 2015 “The ‘Stranger’ and the Grammarian: When Early English Grammarians Reached Out”. Caliban 541.309–326. : 321), Bullokar not only departed from a long-standing tradition, which bestowed on Latin the privilege of giving access to knowledge in all fields, he considered the grammar of English as a better passport for Europe. Padley (1988Padley, G. A. 1988 Grammatical Theory in Western Europe 1500–1700. Cambridge: Cambridge University Press.Padley, G. A. 1988 Grammatical Theory in Western Europe 1500–1700. Cambridge: Cambridge University Press.: 230) criticises Bullokar’s Pamphlet for Grammar, stating that it is: “the first in a long line of works assuming that what is appropriate to the description of Latin will be equally appropriate to the description of the mother tongue”.
However, it is crucial to keep in mind that authors such as Bullokar were among the first to write grammars of English in English. Just because they based their grammars on existing Latin grammar books does not mean they assumed the Latin model would be just as applicable to the vernacular. In fact, Algeo (1985Algeo, John 1985 “The Earliest English Grammars”. Historical & Editorial Studies in Medieval & Early Modern English: For Johan Gerritsen ed. by Mary-Jo Arn, Hanneke Wirtjes & Hans Jansen, 191–207. Groningen: Wolters-Noordhoff.Algeo, John 1985 “The Earliest English Grammars”. Historical & Editorial Studies in Medieval & Early Modern English: For Johan Gerritsen ed. by Mary-Jo Arn, Hanneke Wirtjes & Hans Jansen, 191–207. Groningen: Wolters-Noordhoff.: 206) argues that the grammarians at the time were “fully cognizant of the structural differences between Latin as an inflected language and English as an analytical one”. Algeo further claims that the Latinate approach in these early grammars fulfilled a particular purpose. Students learning English grammar through these books were more readily prepared to learn Latin grammar, as a comparable basis was established (Algeo 1985Algeo, John 1985 “The Earliest English Grammars”. Historical & Editorial Studies in Medieval & Early Modern English: For Johan Gerritsen ed. by Mary-Jo Arn, Hanneke Wirtjes & Hans Jansen, 191–207. Groningen: Wolters-Noordhoff.Algeo, John 1985 “The Earliest English Grammars”. Historical & Editorial Studies in Medieval & Early Modern English: For Johan Gerritsen ed. by Mary-Jo Arn, Hanneke Wirtjes & Hans Jansen, 191–207. Groningen: Wolters-Noordhoff.: 203). Moreover, foreigners who already knew Latin were more gently introduced into the structure of English (Algeo 1985Algeo, John 1985 “The Earliest English Grammars”. Historical & Editorial Studies in Medieval & Early Modern English: For Johan Gerritsen ed. by Mary-Jo Arn, Hanneke Wirtjes & Hans Jansen, 191–207. Groningen: Wolters-Noordhoff.Algeo, John 1985 “The Earliest English Grammars”. Historical & Editorial Studies in Medieval & Early Modern English: For Johan Gerritsen ed. by Mary-Jo Arn, Hanneke Wirtjes & Hans Jansen, 191–207. Groningen: Wolters-Noordhoff.: 191).
The purpose of Bullokar’s book was to dismantle the prevalent assumption that English was not rule-governed like Latin. While he used his own reformed spelling system, other traditional grammars at the time were still written in Latin (Linn 2006Linn, Andrew 2006 “English Grammar Writing”. The Handbook of English Linguistics ed. by Bas Aarts & April McMahon, 72–92. Oxford: Blackwell. Linn, Andrew 2006 “English Grammar Writing”. The Handbook of English Linguistics ed. by Bas Aarts & April McMahon, 72–92. Oxford: Blackwell. ). This is seen by Bullokar identifying the differences between the Latin system and the vernacular and adjusting his grammatical descriptions accordingly, for example by listing five cases, as opposed to six which exist in Latin (Raby & Andreu 2018Raby, Valérie & Wilfrid Andrieu 2018 “Norms and Rules in the History of Grammar: French and English handbooks in the seventeenth century”. Standardising English: Norms and margins in the history of the English language ed. by Linda Pillière, Wilfrid Andrieu, Valérie Kerfelec & Diana Lewis, 65–88. Cambridge: Cambridge University Press. Raby, Valérie & Wilfrid Andrieu 2018 “Norms and Rules in the History of Grammar: French and English handbooks in the seventeenth century”. Standardising English: Norms and margins in the history of the English language ed. by Linda Pillière, Wilfrid Andrieu, Valérie Kerfelec & Diana Lewis, 65–88. Cambridge: Cambridge University Press. ; Michael 1970Michael, Ian 1970 English Grammatical Categories and the Tradition to 1800. Cambridge: Cambridge University Press.Michael, Ian 1970 English Grammatical Categories and the Tradition to 1800. Cambridge: Cambridge University Press., 1987 1987 The Teaching of English from the Sixteenth Century to 1870. Cambridge: Cambridge University Press. 1987 The Teaching of English from the Sixteenth Century to 1870. Cambridge: Cambridge University Press. ).
Additionally, with the beginning of books being printed in London at the end of the 15th century, there was a rising economic interest in spelling reform and standardization (Robins 1986Robins, Robert H. 1986 “The Evolution of English Grammar Books Since the Renaissance”. The English Reference Grammar: Language and linguistics, writers and readers ed. by Gerhard Leitner, 292–306. Tübingen: Niemeyer.Robins, Robert H. 1986 “The Evolution of English Grammar Books Since the Renaissance”. The English Reference Grammar: Language and linguistics, writers and readers ed. by Gerhard Leitner, 292–306. Tübingen: Niemeyer.), further incentivizing English grammarians to outline language-specific features of English, particularly for the use in pedagogy (McCarthy 2020McCarthy, Michael 2020 Innovations and Challenges in Grammar. London: Routledge. McCarthy, Michael 2020 Innovations and Challenges in Grammar. London: Routledge. : 23; Michael 1970Michael, Ian 1970 English Grammatical Categories and the Tradition to 1800. Cambridge: Cambridge University Press.Michael, Ian 1970 English Grammatical Categories and the Tradition to 1800. Cambridge: Cambridge University Press., 1987 1987 The Teaching of English from the Sixteenth Century to 1870. Cambridge: Cambridge University Press. 1987 The Teaching of English from the Sixteenth Century to 1870. Cambridge: Cambridge University Press. ).
The 17th century represents an eventful time for the development of the English language and its instruction. The English language expanded to areas of language use where the classical languages had previously been dominating, especially in the late 17th century. This led to increased standardization in language usage (Nevalainen 2006Nevalainen, Terttu 2006 An Introduction to Early Modern English. Edinburgh: Edinburgh University Press. Nevalainen, Terttu 2006 An Introduction to Early Modern English. Edinburgh: Edinburgh University Press. : 42; Milroy & Milroy 2012Milroy, James & Lesley Milroy 2012 [1985] Authority in Language: Investigating Language Prescription and Standardisation. 4th ed. London & New York: Routledge.Milroy, James & Lesley Milroy 2012 [1985] Authority in Language: Investigating Language Prescription and Standardisation. 4th ed. London & New York: Routledge.; Nevalainen & Tieken-Boon van Ostade 2006Nevalainen, Terttu & Ingrid Tieken-Boon van Ostade 2006 “Standardisation”. A History of the English Language ed. by Richard M. Hogg & David Denison, 271–311. Cambridge: Cambridge University Press. Nevalainen, Terttu & Ingrid Tieken-Boon van Ostade 2006 “Standardisation”. A History of the English Language ed. by Richard M. Hogg & David Denison, 271–311. Cambridge: Cambridge University Press. ), which, together with other sociopolitical developments, sparked a shift in favour of English being recognized as a separate academic discipline (Beal 2004Beal, Joan C. 2004 English in Modern Times. London: Arnold.Beal, Joan C. 2004 English in Modern Times. London: Arnold.: 102). Authors were addressing a broader audience of intellectuals and scholars curious about language. There is also a gradual shift towards standardization in comparison to the 16th century, since 17th-century writers such as Jonson and Butler focus more on correctness of forms and codifying norms in English. Due to more books being published, there is also a broader range of topics that are being addressed in the grammars published in the 17th century such as more advanced syntax (e.g. Aickin 1693Aickin, Joseph 1693 The English grammar. London: John Lawrence.Aickin, Joseph 1693 The English grammar. London: John Lawrence.) and pronunciation (e.g. Miege 1688).
English grammars written in the 17th century were still said to be “heavily influenced by their models — grammars of Latin” (Algeo 1986 1986 “A Grammatical Dialectic”. The English Reference Grammar: Language and linguistics, writers and readers ed. by Gerhard Leitner, 307–333. (= Linguistische Arbeiten, 172). Tübingen: Niemeyer. 1986 “A Grammatical Dialectic”. The English Reference Grammar: Language and linguistics, writers and readers ed. by Gerhard Leitner, 307–333. (=Linguistische Arbeiten, 172). Tübingen: Niemeyer.: 309) and “ancient Greek and Roman literati were still considered authoritative in terms of linguistic understanding” (McCarthy 2020McCarthy, Michael 2020 Innovations and Challenges in Grammar. London: Routledge. McCarthy, Michael 2020 Innovations and Challenges in Grammar. London: Routledge. : 24), although there is more recognition of the autonomy of English and scholars are beginning to have access to more contemporary works.
3.Method
This section outlines the mixed method approach employed in this study. First the corpus data and the selection process of the representative sample collected for the HeidelGram corpus is outlined (3.1). Then the corpus compilation process is explained (3.2). This is supported by a description of the specific challenges historical corpus linguists face when collecting, transcribing, and annotating their texts, in our case a representative corpus of historical English grammars from the 16th and 17th centuries (3.3). Lastly, an in-depth explanation of the extraction and visualization processes as well as the analyses of the onomastic references is given (3.4). The methodology employed in this study makes use of approaches from corpus linguistics and network analysis.
3.1Corpus design
As part of the larger HeidelGram corpus, five grammars from the 16th century and 17 grammars from the 17th century were selected. These numbers represent the few English grammars that were extant in the 16th century as well as the marked increase in publications in the 17th century (see Michael 1970Michael, Ian 1970 English Grammatical Categories and the Tradition to 1800. Cambridge: Cambridge University Press.Michael, Ian 1970 English Grammatical Categories and the Tradition to 1800. Cambridge: Cambridge University Press.: 151; Yáñez-Bouza 2016Yáñez-Bouza, Nuria 2016 “Early and Late Modern English Grammars as Evidence in English Historical Linguistics”. The Cambridge Handbook of English Historical Linguistics ed. by Merja Kytö & Päivi Pahta. Cambridge: Cambridge University Press. Yáñez-Bouza, Nuria 2016 “Early and Late Modern English Grammars as Evidence in English Historical Linguistics”. The Cambridge Handbook of English Historical Linguistics ed. by Merja Kytö & Päivi Pahta. Cambridge: Cambridge University Press. : 166–168). The texts were selected in accordance with the selection criteria for the full corpus using higher- and lower-level strata. More specifically, the higher-level stratum of popularity was considered first. This means that the number of published editions, distribution, and mentions in the primary (i.e. the grammars themselves) and secondary literature (such as Görlach 1998Görlach, Manfred 1998 An Annotated Bibliography of Nineteenth Century Grammars of English. Amsterdam & Philadelphia: John Benjamins. Görlach, Manfred 1998 An Annotated Bibliography of Nineteenth Century Grammars of English. Amsterdam & Philadelphia: John Benjamins. ; Alston 1965Alston, R. C. 1965 A Bibliography of the English Language from the Invention of Printing to the Year 1800. Vol. 1. Leeds: Arnold and Son.Alston, R. C. 1965 A Bibliography of the English Language from the Invention of Printing to the Year 1800. Vol. 11. Leeds: Arnold and Son.) were considered to assess the influence a grammar may have had at the time. Among these popular works, the lower-level strata were considered to provide variety in function, audience, and text type (Busse et al. 2020 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. : 137). Variety in function was achieved by including teaching and scholarly grammars, whereas different text types included the representation of the grammar material in e.g. prose, catechisms, and verse. The audience of the grammars varied as they were intended for novice and advanced learners, youth and adults, natives and foreigners. This stratified sampling approach ensures that the sample includes as much variety within the genre of grammar writing as possible, which leads to a higher degree of representativeness (see Biber 1993Biber, Douglas 1993 “Representativeness in Corpus Design”. Literary and Linguistic Computing 8:4. 243–257. Biber, Douglas 1993 “Representativeness in Corpus Design”. Literary and Linguistic Computing 8:4. 243–257. , Egbert et al. 2022Egbert, Jesse, Douglas Biber & Bethany Gray 2022 Designing and Evaluating Language Corpora: A practical framework for corpus representativeness. Cambridge: Cambridge University Press. Egbert, Jesse, Douglas Biber & Bethany Gray 2022 Designing and Evaluating Language Corpora: A practical framework for corpus representativeness. Cambridge: Cambridge University Press. ). We follow the definition of genres as “categories of texts which are determined by both formal and functional criteria” (Busse 2015Busse, Beatrix 2015 “Genre”. The Cambridge Handbook of Stylistics ed. by Peter Stockwell & Sara Whiteley, 103–116. Cambridge: Cambridge University Press.Busse, Beatrix 2015 “Genre”. The Cambridge Handbook of Stylistics ed. by Peter Stockwell & Sara Whiteley, 103–116. Cambridge: Cambridge University Press.: 103). The inclusion of grammar books that serve different purposes enables us to have a broader overview of who the English grammar scholars considered to be experts on the topic at the time, and how they interacted with one another.
Table 2 lists the grammars that were selected for the 16th- and 17th-century components of the HeidelGram corpus sorted by year of publication.
| Author | Title | Year of Publication | Word count |
|---|---|---|---|
| Richard Sherry | A Treatise of the Figures of Grammer and Rhetorike | 1577 | 27,368 |
| Richard Mulcaster | The First Part of the Elementarie Which Entreateth Chefelie of the Right Writing of our English Tung | 1582 | 10,1047 |
| William Bullokar | Brief Grammar for English | 1586 | 17,606 |
| Gabriel Meurier | The Coniugations in Englishe and Netherdutche | 1586Meurier, Gabriel 1586 The Coniugations in Englishe and Netherdutche, According as Gabriel Mevrier Hath Ordayned the Same, in Netherdutche, and Frenche. Leyden: Thomas Basson.Meurier, Gabriel 1586 The Coniugations in Englishe and Netherdutche, According as Gabriel Mevrier Hath Ordayned the Same, in Netherdutche, and Frenche. Leyden: Thomas Basson. | 7,131 |
| Edmund Coote | The English Schoole-Maister Teaching all his Schollers | 1596 | 29,476 |
| Total word count 16th century | 182,628 | ||
| Alexander Hume | Orthographie and Congruitie of the Britan Tongue | 1617Hume, Alexander 1617 Orthographie and Congruitie of the Britan Tongue. London: Early English Text Society.Hume, Alexander 1617 Orthographie and Congruitie of the Britan Tongue. London: Early English Text Society. | 18,220 |
|
John Hewes
[Huise] |
A Perfect Survey | 1624Hewes, John 1624 A Perfect Survey. London: Edw: All-de.Hewes, John 1624 A Perfect Survey. London: Edw: All-de. | 42,189 |
| John Brinsley | The Posing of the Parts | 1612 | 34,398 |
| Charles Butler | English Grammar | 1633Butler, Charles 1633 English Grammar. Oxford: William Turner.Butler, Charles 1633 English Grammar. Oxford: William Turner. | 34,484 |
| Ben Jonson | The English Grammar | 1640Jonson, Ben 1640 The English Grammar.Jonson, Ben 1640 The English Grammar. | 15,495 |
| Joshua Poole | The English Accidence | 1646Poole, Joshua 1646 The English Accidence. London: E. Cotes.Poole, Joshua 1646 The English Accidence. London: E. Cotes. | 17,189 |
| Francis Lodowyck | A Common Writing | 1647Lodowyck, Francis 1647 A Common Writing.Lodowyck, Francis 1647 A Common Writing. | 4,117 |
| Jeremiah Wharton | The English Grammar | 1654Wharton, Jeremiah 1654 The English Grammar. London: William Du-Gard.Wharton, Jeremiah 1654 The English Grammar. London: William Du-Gard. | 15,113 |
| James Howell | A New English Grammar | 1662 | 46,488 |
| John Wilkins | An Essay towards a Real Character, and a Philosophical Language | 1668Wilkins, John 1668 An Essay towards a Real Character, and a Philosophical Language. London: Royal Society.Wilkins, John 1668 An Essay towards a Real Character, and a Philosophical Language. London: Royal Society. | 216,806 |
| John Newton | School Pastime for Young Children: or the Rudiments of Grammar | 1669Newton, John 1669 School Pastime for Young Children: or the Rudiments of Grammar. London: Robert Walton.Newton, John 1669 School Pastime for Young Children: or the Rudiments of Grammar. London: Robert Walton. | 12,200 |
| Mark Lewis | Plain & Short Rules | 1675Lewis, M. 1675 Plain & Short Rules.Lewis, M. 1675 Plain & Short Rules. | 3,088 |
| John Newton | The English Academy, or, a Brief Introduction to the Seven Liberal Arts | 1677 1677 The English Academy, or, a Brief Introduction to the Seven Liberal Arts. London: W. Godbid. 1677 The English Academy, or, a Brief Introduction to the Seven Liberal Arts. London: W. Godbid. | 23,025 |
|
Christopher
Cooper |
The English Teacher (English translation of Grammatica Linguæ Anglicanæ) | 1687Cooper, Christopher 1687 The English Teacher (English translation of Grammatica Ling. Angl.). London: John Richardson.Cooper, Christopher 1687 The English Teacher (English translation of Grammatica Ling. Angl.). London: John Richardson. | 29,159 |
| Guy Miège | The English Grammar | 1688Miège, Guy 1688 The English Grammar. London: J. Redmaine.Miège, Guy 1688 The English Grammar. London: J. Redmaine. | 15,055 |
| Maurice Wheeler | The Royal Grammar | 1690Wheeler, Maurice 1690 The Royal Grammar. London: J. Heptinstall.Wheeler, Maurice 1690 The Royal Grammar. London: J. Heptinstall. | 51,960 |
| Joseph Aickin | The English Grammar | 1693 | 10,952 |
| Total word count 17th century | 589,938 | ||
Following common practice in earlier works, for these texts, first editions, or the earliest available editions, of the books have been traced and scanned. Further processing of the digitized texts is described in the methodology Section (4.1).
Within the HeidelGram corpus, and therefore also in these sub-corpora, a whole text approach is followed. This means that the grammar texts are included in the corpus in full length. Although the grammars vary significantly in size, it is essential to include the full text in order to attain a holistic understanding of the genre of English grammar writing. Selecting same-size samples for the benefit of balance would lead to the loss of domain characteristics (Hunston 2008Hunston, Susan 2008 “Collection Strategies and Design Decisions”. Corpus Linguistics ed. by Anke Lüdeling & Merja Kytö, 154–167. Berlin: De Gruyter.Hunston, Susan 2008 “Collection Strategies and Design Decisions”. Corpus Linguistics ed. by Anke Lüdeling & Merja Kytö, 154–167. Berlin: De Gruyter.: 165–166). The drawback of the different length of the texts are mitigated through normalized frequencies in the analyses.
3.2Corpus compilation
While it is not the focus of this research paper, we want to briefly address some of the crucial steps in the compilation of the HeidelGram corpus. For the corpus data, PDF scans of the first editions, if available, of the grammar books were collected from online archives and through library scans. Where possible, OCR was used to transcribe the texts into machine-readable format. For each grammar text, two corpus files were created, a clean text file and an XML-annotated text file. However, as mentioned above, the historical nature of these texts complicates the compilation process. In lieu of a standardized printing process or standardized spelling at the time period under investigation, the books exhibit a vast range of irregularities, be it in fonts, page structure, figures, spelling, or even special characters. Thus, manual transcription and correction was necessary to allow for clean transcriptions of the grammars. The raw texts were then manually annotated for visual, structural, and project-specific features, such as typeface, text placement, and references to personal names, respectively. Annotations are added in TEI-adjacent XML and the full guidelines will be published alongside the corpus once fully compiled. For example, in Howell (1662)Howell, James 1662 A New English Grammar. London: T. Williams, H. Brome, and H. Marsh.Howell, James 1662 A New English Grammar. London: T. Williams, H. Brome, and H. Marsh. a reference to Ben Jonson is made, which is annotated as: “[...] it cannot be in the compas of human brain to compile an exact regular Syntaxis thereof, <person>Mr. Ben. Johnson</person> a great Wit, who was as patient as he was elaborat [...].”
These manual processes of transcription and annotation were performed by the project’s team members, that is the present authors and trained student assistants. The project-specific annotation of references to personal names is crucial to the extraction process of onomastic references outlined in the following section.
In order to have a broad overview of all references to persons in the corpus, we decided to tag any mentions of a person when annotating the corpus and this was only filtered out once the concordances were extracted and it was apparent that it was indeed a reference to a specific individual.
Throughout the process of compiling the corpus, occasionally there were instances where the words were illegible. In these cases, later editions of the book and facsimiles were consulted in order to ensure the corpus was transcribed as accurately as possible. However, with historical texts there are times where it is very difficult to decipher whether there was an error made by the printer or the author. Howard-Hill (2006Howard-Hill, Trevor H. 2006 Early Modern Printers and the Standardization of English Spelling. London: Modern Humanities Research Association.Howard-Hill, Trevor H. 2006 Early Modern Printers and the Standardization of English Spelling. London: Modern Humanities Research Association.: 16) outlines the importance of comprehending that “early modern printers–unlike present-day printers–did not follow the spelling of their copies. Consequently, the printed works of spelling reformers and dictionary-makers used spellings that were not sanctioned by their authors”. Such variations in spelling could also occur within onomastic references. Additionally, different realizations of referring to the same person occurred. For these reasons, we decided to collect all instances of an author’s name as it occurred in the text during the extraction stage, and in later stages of the analysis, determine whether they refer to the same person.
3.3Working with historical (corpus) data
Diachronic data, by definition, means there is an element of change or stability occurring over the course of a set time frame. We therefore refer to Bakró-Nagy’s explanation on the specificity of historical linguistic data as:
the sum of statements made about a conceptionally described subset of linguistic phenomena described with specific preconceptions in mind. Data are inconsistent in the sense that their truth value is not constant, that is, they can change as a result of possible change in our knowledge of processes of change or of past states, or even induce processes of chain-like reinterpretation.(Bakró-Nagy 2010Bakró-Nagy, Marianne 2010 “The Data in Historical Linguistics: On utterances, sources, and reliability”. Sprachtheorie und germanistische Linguistik 20:2. 133–195.Bakró-Nagy, Marianne 2010 “The Data in Historical Linguistics: On utterances, sources, and reliability”. Sprachtheorie und germanistische Linguistik 20:2. 133–195.: 136)
Fischer addresses the question of whether data in historical linguistics should be considered in terms of formalist grammar change or language change and suggests that:
[a] historical linguist who bases himself too exclusively on a particular theory may come to suggest an interpretation of the historical data that can only be called an oversimplification of its complex nature or even worse can lead to the neglect of relevant and by no means incidental facts.(Fischer 2004Fischer, Olga 2004 “What Counts as Evidence in Historical Linguistics?” Studies in Language 28:3. 710–740. Fischer, Olga 2004 “What Counts as Evidence in Historical Linguistics?” Studies in Language 28:3. 710–740. : 713)
One of the main challenges of working with historical data is the lack of standardized lexical forms or structural norms. Aside from variations in spelling and diacritical marks i.e., macrons, use of blackletter font can make digitized copies difficult to visually inspect and interpret, particularly if the scans are of a lower quality.
Jenset and McGillivray (2017)Jenset, Gard B. & Barbara McGillivray 2017 Quantitative Historical Linguistics: A corpus framework. Oxford: Oxford University Press. Jenset, Gard B. & Barbara McGillivray 2017 Quantitative Historical Linguistics: A corpus framework. Oxford: Oxford University Press. claim that unlike contemporary texts, historical texts present an additional challenge of potentially missing metadata such as information about the author, title, publisher, publication date, etc. This is particularly important when drawing comparisons between different editions of the grammar books. Moreover, additional data about the authors themselves might show, for example, whether scholars with a similar background might refer more often to other authors, or not.
Focusing on a specific genre, in this case grammar books, meant that in addition to collecting information about the author’s academic endeavours, we also added information in a header above each grammar transcript about the target audience and what topics, i.e. which aspects of grammar, are covered in the book. Unlike literary works, the authors often addressed their intended readers in the preface and specified what the purpose of the book was. For instance, grammars could be addressed to younger audiences or adults, and the purpose of the text could be the use in a classroom or for self-learning at home. Additional information about the authors, such as their education and their date of birth and death, was collected in separate metadata files for each author in the corpus. All additional information on the authors and texts were collected using the relevant secondary literature, bibliographies such as the ones by Görlach (1998)Görlach, Manfred 1998 An Annotated Bibliography of Nineteenth Century Grammars of English. Amsterdam & Philadelphia: John Benjamins. Görlach, Manfred 1998 An Annotated Bibliography of Nineteenth Century Grammars of English. Amsterdam & Philadelphia: John Benjamins. and Alston (1965)Alston, R. C. 1965 A Bibliography of the English Language from the Invention of Printing to the Year 1800. Vol. 1. Leeds: Arnold and Son.Alston, R. C. 1965 A Bibliography of the English Language from the Invention of Printing to the Year 1800. Vol. 11. Leeds: Arnold and Son., and online sources such as the Eighteenth-Century English Grammars database online (ECEG) (Yáñez-Bouza & Rodríguez-Gil 2013Yáñez-Bouza, Nuria & M. E. Rodríguez-Gil 2013 “The ECEG Database”. Transactions of the Philological Society 111:2. 143–164. Yáñez-Bouza, Nuria & M. E. Rodríguez-Gil 2013 “The ECEG Database”. Transactions of the Philological Society 111:2. 143–164. ).
3.4Extracting onomastic references using corpus and network analysis
Once the data were fully transcribed and annotated, the annotation for onomastic references enabled an automated extraction of all references to personal names. This extraction process will be the focus of this methodological section. For this purpose, we have built a custom Python script, which will be made available on the project website once the corpus is completed. The onomastic references were automatically extracted using project-specific annotation together with a concordance window (150 characters to the left and right) and the source text from which it was extracted. An example of such an extracted concordance can be seen in Table 3. For each source text the original year of publication as well as the year of publication of the edition in the corpus precede the name of the grammar author, e.g. 1662–1662-Howell indicates that Howell’s grammar was originally published in 1662 and the edition in the corpus is the first edition. The concordance is extracted as in the original text and can therefore contain spelling errors, such as in the case in Table 3, where Ben Jonson’s last name is misspelled.
| Source text | Left context | Reference | Right context |
|---|---|---|---|
| 1662–1662-Howell | and having such varieties of incertitudes, changes and Idioms, it cannot be in the compas of human brain to compile an exact regular Syntaxis therof, | Mr. Ben. Johnson | a great Wit, who was as patient as he was elaborat in his re- sherches and compositions, as he was framing an English Syntaxis, confess’d the further |
The resulting concordances were then manually coded by two independent raters, that is two members of the project team. Two rounds of inter-rater reliability coding were performed, first for the reference types and second for the person types. A normalized version of the person’s name was added in this stage as well. This workflow is visualized in Figure 1.
Many references had to be excluded from further analysis based on the following criteria:
-
mentions of names within an example,
-
mentions of the author’s, printer’s, bookseller’s name on the book cover,
-
mentions of names recurring in a boilerplate or other structural element of the book,
-
mentions of names in advertisements or other parts of the book that are not part of the original grammar text,
-
ambiguous references that could not be clearly identified,
-
or mentions of names in non-English language sections of the book.
These onomastic mentions are not considered as references. References to titles that were held by persons e.g. the King of England were excluded because they are not name-based and have been held by different people throughout the years. Thus, only those name-based references were considered which were presumably made intentionally by the authors themselves and which we were able to identify and analyse.
The grounded theory approach (Bryant & Charmaz 2007Bryant, Anthony & Kathy Charmaz 2007 “Introduction: Grounded Theory Research: Methods and Practices”. In The SAGE Handbook of Grounded Theory ed. by Anthony Bryant & Kathy Charmaz, 1–28. London: SAGE. Bryant, Anthony & Kathy Charmaz 2007 “Introduction: Grounded Theory Research: Methods and Practices”. In The SAGE Handbook of Grounded Theory ed. by Anthony Bryant & Kathy Charmaz, 1–28. London: SAGE. ) was used both when analysing the type of person referenced and establishing what type of reference was used, building on our previous work for the 19th-century grammar books (Busse et al. 2020 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. 2020 “A Corpus-Based Analysis of Grammarians’ References in 19th-Century British Grammars”. Variation in Time and Space: Observing the world through corpora ed. by Anna Čermáková & Markéta Malá, 133–172. Berlin: De Gruyter. : 143–144). The extraction of concordances enabled us to analyse not only who the persons being referenced were, such as Benjamin Johnson in Table 3 or other figures such as Cicero, God, and John Milton. It also provided us with the context in which they were being referred to, which would not have been possible had we only worked with search terms. The independent raters conducted two different assessments — for both the persons and how they were being written about. These assessments were then quantified using inter-rater reliability (IRR) scores. Both Cohen’s Kappa and Krippendorff’s Alpha were used for the evaluation. These metrics allow for a quantification of agreement, i.e. whether the raters applied the same categories consistently. A high score in these metrics therefore ensures applicability and robustness of the categories. After evaluating the IRR scores, those categories where the raters’ allocations did not align were re-evaluated within the team according to the category descriptions, resulting in the final categorization of authors and references. ’
The obtained frequency and category information is represented in network graphs (see Figure 2 and 3 below), which then allow for further analysis. This representation stores the number of references made from the sources to the targets, as well as the attributes of the targets and the connections. In this case, the graphs represent the number of references made from the grammar authors in the corpus to the persons being referenced as well as the information about the person and references types. The representations were achieved by processing and visualizing the data using a custom R script, which takes a so-called intercitation matrix as input. The scripts for reference extraction and network visualization can be found in the appendices. The citation information is processed using the igraph package (Csárdi et al. 2025Csárdi, Gábor, Tamás Nepusz, Vincent Traag, Szabolcs Horvát, Fabio Zanini, Daniel Noom & Karsten Müller 2025 igraph: Network Analysis and Visualization in R. R package version 2.1.4. Csárdi, Gábor, Tamás Nepusz, Vincent Traag, Szabolcs Horvát, Fabio Zanini, Daniel Noom & Karsten Müller 2025 igraph: Network Analysis and Visualization in R. R package version 2.1.4. ). Both the width of the edges as well as the size of the nodes have been linearly scaled for legibility. The positionality of the nodes is determined by a physics force calculation. This is employed to ensure legibility only, meaning that there is no further purpose to the nodes’ positions.
4.Results
A total of 1394 onomastic references were extracted from the corpus consisting of five English grammar books from the 16th century. For the 17 texts from the 17th century a total of 3636 onomastic references were extracted. After exclusions (see Section 3.4), 233 and 1393 name-based references remained for the 16th- and 17th-century data, respectively. These were further categorized and analysed.
| Source Text | 1668-1668-Wilkins |
| Left Context | Beside several of our own Country-men, Sir Thomas Smith, Bullokar, Alexander Gill, and |
| Reference | Doctor Wallis |
| Right Context | ; the last of whom, amongst all that I have seen published, seems to me, with greatest Accuratene ss and subtlety to have considered the Philosophy of |
| Normalized Name | John Wallis |
| Author Type | Grammar Author |
| Primary Reference Type | Opinion |
| Secondary Reference Type | Comparison/ Contrast |
In the remainder of this chapter the following results are presented. First a categorization of person types has been achieved by means of a grounded theory approach and is shown to be well suited for the data at hand. Existing categories of reference types, established for 19th-century grammar references, were reevaluated for the 16th- and 17th-century grammars. The construction of scholarly networks of grammarians’ references allows for qualitative observations, while further network and corpus analyses provide quantitative results.
4.1Categories of persons
A first result of this study is the categorization system of the different kinds of persons or figures referred to by the grammarians in the data set. The seven person categories that emerged from the data are listed and described in Table 5.
| Person type | Description |
|---|---|
| Grammar Author | Authors of grammar books of any language. |
| Ancient Scholar | Includes philosophers, orators, artists, etc. |
| Poet / Literary Author | Authors of fictional work. |
| Political Figure | Includes kings, statesmen, warriors, knights, etc. |
| Religious Figure | Biblical figures, including God, and figures from Roman or Greek mythology. |
| Contemporary Scholar | Contemporary scholars for the time period of the grammar Authors, who do not fit any other category. |
| None | Generic names and fictional characters. |
It is interesting to note here which types of figures were considered authoritarian on the use of the English language. Not only are grammar authors, poets, and literary authors referred to, but also famous scholars from other disciplines as well as famous orators, politicians, and religious figures. We will see in the following sections that these other scholars and famous figures are even more prominently referenced than linguistic or literary scholars.
4.2Network visualization
The network in Figure 2 displays the references made by the five grammarians in the 16th- century component of the corpus, while the network in Figure 3 displays the references made by the grammarians in the 17th-century component of the corpus.
The rectangular nodes represent the grammars in the corpus, while the size of their node is normalized to provide a representation of the size (i.e., word count) of the grammar text. The circular nodes represent the authors and figures referred to in the grammars. The colours of the nodes refer to the types of persons while the colours of the edges refer to the reference category applied. Thicker edges display larger amounts of references of this type made to this figure. Please note, that this network visualization introduces some room for misinterpretation. For instance, there is no meaning to the positionality of the nodes. Also, larger nodes do represent longer grammars or more references to certain authors, but these sizes are normalized, i.e., a node of double size does not represent a grammar twice as long or twice as many references made. Nevertheless, the network is a meaningful representation of the data, as it allows the observation of patterns such as interconnectedness, centrality, and isolation, as will be elaborated in the following section.
4.3Network observations
Figure 2 clearly shows that two of the 16th-century grammar authors, Richard Sherry and Richard Mulcaster, are the main contributors to the reference network. The other grammars contain very few, in the case of William Bullokar and Edmund Coote, or even no onomastic references at all, in the case of Gabriel Meurier. In that sense, Bullokar and Coote are isolated, Meurier does not occur, and Sherry and Mulcaster are central in the network.
It is also interesting to note that, just as was expected in advance, none of the grammar authors refer to each other. Generally, little references are made to other grammarians of any language. The authorities referred to most in these texts go back to ancient scholars, politicians, and other orators, who are particularly famous for their eloquence in the Latin tongue. Similarly, this can also be seen by means of indegree centrality, which counts the number of incoming edges to a particular author node (Prell 2012Prell, Christina 2012 Social Network Analysis: History, Theory, and Methodology. London: SAGE.Prell, Christina 2012 Social Network Analysis: History, Theory, and Methodology. London: SAGE.: 99). Scaled from zero to one, the ten highest indegree centrality scores throughout the 16th-century network are presented in Table 6.
| Person | Normalized indegree centrality score | Person Type |
|---|---|---|
| Marcus Tullius Cicero | 1 | Political figure |
| Gaius Iulius Caesar | 0.61 | Political figure |
| Marcus Fabius Quintilianus | 0.56 | Ancient Scholar |
| Plato | 0.53 | Ancient Scholar |
| John | 0.31 | Religious figure |
| Aristotle | 0.28 | Ancient Scholar |
| Terence | 0.28 | Poet/Literary Author |
| Virgil | 0.28 | Poet/Literary Author |
| Livie | 0.17 | Ancient Scholar |
| Cato Uticensis | 0.14 | Political figure |
The scores again confirm our observations. Across all five grammar books considered here, the most central figures of reference were not grammar authors, but rather famous orators such as Cicero, political figures such as Caesar, or famous scholars from ancient Greece or Rome, such as Quintilian or Plato.
Closer inspection of the 16th-century network further indicates that Sherry and Mulcaster are not only the main contributors to the network, but they also display the most prominent interconnectedness, i.e. they are connected through persons they both refer to. Both authors frequently reference Caesar, Quintilian, and Cicero, whom they consider great orators and rhetoricians. However, the network also implies that both grammarians considered different figures as authoritative with regard to language, since both exhibit clusters of persons that none of the other grammarians refer to.
The indegree centrality scores for the 17th-century component of the corpus (Table 7) shows a similar image. Rather than including grammar authors or contemporary scholars, most persons referenced are political figures, poets and literary authors from ancient Rome or Greece, such as Cicero, Seneca, Ovid, and Virgil. This finding confirms that 17th-century grammars of English were still largely impacted and influenced by the Latinate tradition.
| Person | Normalized Indegree centrality score | Person Type |
|---|---|---|
| Marcus Tullius Cicero | 0.660 | Political Figure |
| Hieronymus Megiser | 0.261 | Contemporary Scholar |
| Plautus | 0.176 | Poet/Literary Author |
| Terentianus Maurus | 0.176 | Grammar Author |
| God | 0.151 | Religious Figure |
| Ovid | 0.135 | Poet/Literary Author |
| Seneca | 0.123 | Ancient Scholar |
| Horace | 0.110 | Poet/Literary Author |
| Virgil | 0.107 | Poet/Literary Author |
| John Gower | 0.101 | Poet/Literary Author |
The network of the 17th-century references in Figure 3 further displays similar patterns as the one for the 16th century in Figure 2. Most grammarians in the 17th-century data refer to only few other persons, while two grammarians stand out by referring to many different figures, namely Hewes and Wilkins. We can also observe that, as before in the 16th century, there are only few persons that are referred to by multiple grammar authors. Instead, for most grammarians, there is a cluster of persons only they refer to while none of the other authors do.
Quantifiable scores, such as the centrality scores above, enable us to quickly recognize the most central persons referred to within the network. Especially once the networks grow even larger and more complex, quantitative measures will be crucial in order to track influence and referencing activities.
4.4Frequency observations
4.4.1Distrubution of person types
For the assignment of the person types two independent raters sorted the mentioned figures into the corresponding groups. The following inter-rater reliability scores were achieved in doing so.
| Metric | Score 16th Century | Score 17th Century |
|---|---|---|
| Cohen’s Kappa | 0.82 | 0.94 |
| Krippendorff’s Alpha | 0.82 | 0.94 |
The seven person categories established within this research (see Table 5) were applied to the concordance lines with high agreement scores, which indicates an exhaustive and suitable categorization. In the analyses of the 17th-century data an adjustment was made to these person types. In the initial analyses of the 16th-century data, the seventh category was called Interdisciplinary Scholar. For the analyses of the 17th-century data, the category of Interdisciplinary Scholar was replaced by the category None. The None category was intended to capture all generic names and fictional names, which are typically found in examples of grammar texts. As the IRR scores in Table 8 confirm, this adjustment of the person categories led to an improved agreement among the independent raters and a more robust categorization. Future studies concerning later grammarians’ references will again put these categories to the test of time, validating whether or not these are equally applicable for 18th- and 19th-century references, which will be revisited in more depth.
For the references made within the 16th- and 17th-century grammar corpus, these person categories were assigned to all references made in the texts. The normalized frequencies of the person categories referred to in each of the grammars in the corpus are depicted in Figure 4 and 5 respectively. Note that these frequencies refer to the total number of references made, not the number of unique authors of this category. This is especially remarkable in the case of Sherry, whose many references to political figures are in fact mostly references to Caesar and Cicero. This implies how large the impact of just these two figures is on Sherry’s grammar.
Moreover, the chart shows that Bullokar and Coote, who are more isolated in terms of the references they make, also display a tendency to refer to religious figures, rather than any of the other categories. Especially in Coote, most of the references made are acknowledgements of specific biblical texts, which are referenced in the margin of his text. In these, we find references such as “Matth. 28.19.” or “Psal. 19.7.” (Coote 1596Coote, Edmund 1596 The English Schoole-Maister Teaching all his Schollers, of What Age Soever, the Most Easie, Short, and Perfect Order of Distinct Reading, and True Writing our English-Tongue, that Hath Euer Yet Beene Knowne or Published by any. London: Printed by the Widow Orwin, for Ralph Jackson & Robert Dextar.Coote, Edmund 1596 The English Schoole-Maister Teaching all his Schollers, of What Age Soever, the Most Easie, Short, and Perfect Order of Distinct Reading, and True Writing our English-Tongue, that Hath Euer Yet Beene Knowne or Published by any. London: Printed by the Widow Orwin, for Ralph Jackson & Robert Dextar.: 37–38) to give specific indications of the religious texts referred to.
Sherry and Mulcaster, who are interconnected in the citation network by persons they refer to, both tend to refer to scholars or political figures rather than religious ones. Another interesting observation is that most references are made to figures from ancient Rome or ancient Greece, while fairly little references are made to the grammarians’ contemporaries. This heavy influence of the ancient scholars underlines the Latinate tradition that these early grammars of English were still following. For instance, the references Mulcaster makes to figures such as Cicero and Caesar acknowledge the enduring dominance of Latin: “I will vse no mo examples, where there is no more nede, neither prouf of other tungs, where the Latin is enough (Mulcaster 1586Meurier, Gabriel 1586 The Coniugations in Englishe and Netherdutche, According as Gabriel Mevrier Hath Ordayned the Same, in Netherdutche, and Frenche. Leyden: Thomas Basson.Meurier, Gabriel 1586 The Coniugations in Englishe and Netherdutche, According as Gabriel Mevrier Hath Ordayned the Same, in Netherdutche, and Frenche. Leyden: Thomas Basson., Preface).”
But also references to more recent scholars were still heavily based on their proficiency in Latin. For instance, Sherry refers to Lorenzo Valla, an Italian scholar from the 15th century, praising his service to Latin as a language as follows: “So Laurence Ualla broughte agayne into the olde purenesse, the Latine toungue, whiche thorowe ignoraunce of the Barbarians, was almoste quite loste (Sherry 1577Sherry, Richard 1577 A Treatise of the Figures of Grammer and Rhetorike. London: Ricardi Totteli.Sherry, Richard 1577 A Treatise of the Figures of Grammer and Rhetorike. London: Ricardi Totteli.: f. lii., emphasis added).”
Also remarkable to observe is the relatively low number of references to poets, literary authors, or grammarians. In fact, within the entire dataset, only one reference was made to another grammar author, Priscian, who was a Latin grammarian in around 500 AD. Mulcaster refers to him as an authority in orthography as follows: “[…] I vtter a truth, tho I bring not in a Priscian, or anie Priscianlike ortografer or anie of the twelue old grammarians likned to the nine muses and the thre graces in the Latin tung (Mulcaster 1582Mulcaster, Richard 1582 The First Part of the Elementarie Which Entreateth Chefelie of the Right Writing of our English Tung. London: Thomas Vautroullier.Mulcaster, Richard 1582 The First Part of the Elementarie Which Entreateth Chefelie of the Right Writing of our English Tung. London: Thomas Vautroullier.: 160).”
In the 17th century, most references were made to poets/literary authors and political figures. Still only a few references were made to other grammar authors. This may be observed from the normalized frequencies displayed in Figure 5.
As with the 16th-century data it can be observed that individual grammarians seem to exhibit different preferences with regard to whom they refer to in their writing. It is particularly noticeable that Hewes and Jonson have made significantly more onomastic references than the other grammarians in the 17th-century component of the HeidelGram corpus. Both grammar authors seem to be particularly fond of referring to poets and literary authors and in the case of Hewes also political figures. For instance, Hewes praises Caesar’s oratory skills in the following:
To few men it is giuen to excell in any one kinde, but this to endeauour, or to seeke further to aspire (as was noted in Caius Caesar, a sweetnesse in his Oratory, and in his Warlike stratagems a sharpe and ready dexterity) this euermore standeth with the duty of a good man, when Wisedome and knowledge they seldome burthen vs, they neuer shame vs.(Hewes 1624Hewes, John 1624 A Perfect Survey. London: Edw: All-de.Hewes, John 1624 A Perfect Survey. London: Edw: All-de.: Section Y3)
Similarly, Jonson refers to both Homer and Virgil as the archetypes of poets in the Greek and Latin traditions, respectively, in saying “By the Poet, among the Grecians, Homer: with the Latines, Virgill, is understood” (Jonson 1640Jonson, Ben 1640 The English Grammar.Jonson, Ben 1640 The English Grammar.: 75).
Other grammarians in the 17th-century data, such as Lodowyck, Newton, Lewis, and Aickin still represent the strong influence religion had in the context of grammar writing. For instance, Newton provides the following suggestion to the instructor using his grammar text:
And after they have well acquainted themselves with the usual Prayers of the Church, the Church Catechism, David’s Psalms, and Youths Behaviour, let them proceed to the Grammatical Catechism, by which the former Notions received by the Ear, will be more and more confirmed unto them, and a good foundation laid not only for Reading and Writing the English Tongue, but also for the Learning of any other Language.(Newton 1669Newton, John 1669 School Pastime for Young Children: or the Rudiments of Grammar. London: Robert Walton.Newton, John 1669 School Pastime for Young Children: or the Rudiments of Grammar. London: Robert Walton.: Preface)
The onomastic reference in this paragraph is made to David’s psalms, but the entire excerpt refers to diverse religious texts, which the learner of English ought to read. Newton claims that these texts are beneficial to the learner, not only of English, but any language.
Although references to other grammar writers, and especially authors of English grammars, are still sparse in the 17th-century grammar texts, they seem to become aware of their contemporaries’ works. For instance, Hume in his section on syllables acknowledges that “Ben Jonson spells this word syllabe in his English Grammar” (Hume 1865 [1612]: 40, original emphasis).
4.4.2Distribution of reference categories
The six reference categories, which were originally established by means of 19th-century grammar data, were assigned to the identified references of the 16th- and 17th-century grammars by two independent raters. The following inter-rater reliability scores were achieved.
| Metric | Score 16th Century | Score 17th Century |
|---|---|---|
| Cohen’s Kappa (MASI) | 0.73 | 0.83 |
| Krippendorff’s Alpha (MASI) | 0.72 | 0.81 |
These high agreement scores show that the categories originally defined for the 19th-century corpus of grammars can also be applied to the 16th- and 17th-century grammar data robustly. This implies that within the genre of British grammar writing the same strategies to refer to other persons by name were employed throughout the considered time period. However, it remains to be shown which of these strategies were most salient in which time period. This will be explored in the following. The normalized frequencies of the final, agreed upon reference categories for each book from the 16th century are displayed in Figure 6.
Considering the reference categories applied by the authors, it is interesting to note that across the board fairly few references are accompanied by the author’s opinion. And even when an opinion is uttered, it is typically in favour of the mentioned person. For instance, Sherry in his grammar writes: “In expressyng of these among the Latines, Liuius is very cunning (Sherry 1577Sherry, Richard 1577 A Treatise of the Figures of Grammer and Rhetorike. London: Ricardi Totteli.Sherry, Richard 1577 A Treatise of the Figures of Grammer and Rhetorike. London: Ricardi Totteli.: f. xlvi, emphasis added).”
The most prominent categories, especially in the case of Sherry, are quotations and mentions. However, direct quotations are uncommon in these texts. Instead, indirect elaborations of someone else’s ideas are preferred, as in this section:
This is well handled of Cicero in the preface of the third boke of his Offices: that Scipio was wont to saye, he was neuer lesse ydle then whē he was voyde of the common wealthe matters, and neuer lesse alone, thē whē he was alone.(Sherry 1577Sherry, Richard 1577 A Treatise of the Figures of Grammer and Rhetorike. London: Ricardi Totteli.Sherry, Richard 1577 A Treatise of the Figures of Grammer and Rhetorike. London: Ricardi Totteli.: f. lv., emphasis added)
Rather than quoting individual lines or thoughts, the authors have a tendency to print whole sections of text or full orations, as they consider them highly eloquent but also morally educational for the young scholar, whom their grammar is intended for. Coote in his grammar recites full prayers and psalms for the linguistic but also religious advancement of his students. Sherry, on the other hand, includes a full explanation of Cicero’s oration for Marcus Marcellus, as well as what seems to be a translation of another speech to thank Caesar for the restitution of Marcus Marcellus. This second speech is annotated for rhetorical figures and notes in the margins.
In the 17th-century data, the normalized frequencies illustrated in Figure 7 indicate that most references are quotations, followed by mentions and acknowledgements. The other categories are employed rather sparsely.
Especially Hewes (1624)Hewes, John 1624 A Perfect Survey. London: Edw: All-de.Hewes, John 1624 A Perfect Survey. London: Edw: All-de. and Jonson (1640)Jonson, Ben 1640 The English Grammar.Jonson, Ben 1640 The English Grammar. seem to make use of quotations significantly. Considering these two books in more detail has shown that they frequently quote persons using the quoted sentence as an example or for the purpose of an exercise. For instance, Hewes in his grammar poses the following exercise:
But to whether yee may best incline, the Eupheny, or sweetnesse of the speech it selfe will best direct you, as in these Examples following: And where I wish you onely to obserue the place of the Verbe, and that is:
As it accordeth to the Nominatiue 1. former, 2. Later.
(Hewes 1624Hewes, John 1624 A Perfect Survey. London: Edw: All-de.Hewes, John 1624 A Perfect Survey. London: Edw: All-de.: D2, original emphasis)
Tulliola deliciæ nostra tuum munusculum flagitat. Cic.
Mangnæ diuitiæ sunt lege naturæ composita paupertas. Sen.
Captiui præda milituna fuerunt. Liu.
In this excerpt, Hewes makes use of quotations from Cicero, Seneca, and Livy in order to instruct the learner with regard to the placement of the verb in relation to the nominative case. This paragraph thus not only shows how quotations are used for exemplification and exercises but also emphasizes the remaining strong influence of Latinate authors and their eloquence.
Acknowledgements frequently occur in the margins on the texts, where the grammar authors may indicate where the information from a particular passage is taken from. This is particularly salient in Wilkins’ grammar. However, the grammarians also acknowledge someone else’s work directly in their writing. For instance, Hewes mentions that his grammar “serueth for the more plaine exposition of the Grammaticall Rules and Precepts, collected by LILLIE, and for the more certaine Translation of the English tongue into Latine” (Hewes 1624Hewes, John 1624 A Perfect Survey. London: Edw: All-de.Hewes, John 1624 A Perfect Survey. London: Edw: All-de.: Title page).
The reference category ‘mention’ includes onomastic references in which there is no significant context or evaluation. Frequently, these are intended to provide a time reference, such as the following: “And the Grecians themselves before Homer, as the Romans likewise before Livius Andronicus, had no other Meters (Jonson 1640Jonson, Ben 1640 The English Grammar.Jonson, Ben 1640 The English Grammar.: 54).”
The references in this excerpt to Homer and Livius Andronicus refer to the time in which these persons lived, rather than to the persons themselves.
5.Conclusions and outlook
The reference strategies employed by grammarians to reference other persons show us how they position themselves with regards to certain beliefs and paradigms. The study elucidates the sociolinguistic dynamics of the 16th and 17th centuries by revealing patterns in the selection and representation of names within the grammatical discourse of British sources. The categorization of onomastic references allows for an exploration of the social, cultural, and historical dimensions embedded in the linguistic fabric of the 16th and 17th century.
As the 16th century provides us with fairly few English grammars of the English language, it is difficult to draw specific conclusions about the main influences on the grammar authors. It seems, however, that the approach to this new emerging field of English grammar writing was split in two. While Mulcaster and Sherry accentuate Latinate authors and authorities, in the case of Sherry even writing the book bilingually in Latin and English, the other three authors seem to embrace the novelty of the genre by freeing themselves from any references to such authorities. This is aligned with Dons’s view of the beginnings of grammatication of English being a negotiation with the Latin model: “the beginning of the history of grammar-writing is at the same time the history of the authors’ attempts to mould the vernacular after the Latin model — or to free themselves from its yoke.” (Dons 2004Dons, Ute 2004 Descriptive Adequacy of Early Modern English Grammars. Berlin & New York: De Gruyter. Dons, Ute 2004 Descriptive Adequacy of Early Modern English Grammars. Berlin & New York: De Gruyter. : 242).
Raby and Andrieu (2018)Raby, Valérie & Wilfrid Andrieu 2018 “Norms and Rules in the History of Grammar: French and English handbooks in the seventeenth century”. Standardising English: Norms and margins in the history of the English language ed. by Linda Pillière, Wilfrid Andrieu, Valérie Kerfelec & Diana Lewis, 65–88. Cambridge: Cambridge University Press. Raby, Valérie & Wilfrid Andrieu 2018 “Norms and Rules in the History of Grammar: French and English handbooks in the seventeenth century”. Standardising English: Norms and margins in the history of the English language ed. by Linda Pillière, Wilfrid Andrieu, Valérie Kerfelec & Diana Lewis, 65–88. Cambridge: Cambridge University Press. instead argue that this view portrays the Latin system as oppressive and undermines the fact that the Graeco-Latin tradition created the conditions for other grammarians to base their knowledge on and build other language grammar models.
There were not yet enough grammars published for the 16th-century authors to have access to, and their contemporaries had not yet gained authority as experts on language, thus they are seldom referenced in the books. We interpret the practice of referencing ancient scholars and political figures such as Roman statesman Cicero as evidence that these figures carried significant intellectual and moral authority, therefore portraying a connection between eloquence and civic virtue.
The reliance on authorities from ancient Rome and Greece, as well as famous orators and scholars indicates the wish for a standardized language, based on those who are most eloquent in using it. Additionally, a strong tie to religious texts implies that the authors wished to not only teach their students grammar and language, but also moral standards and good manners. Future studies, tracing the grammatical terminology used will shed more light on the influence and longevity of the Latinate approach.
Few opinions occurred in the 16th-century references and those were more often positive than negative. Instead, quotations and mentions were more frequent. This would imply a tendency towards descriptive approaches rather than prescriptive ones. However, the strong ties to Latinate grammar as well as religion might have a tendency towards prescriptivism after all. In order to place the 16th-century authors more reliably on the prescriptivism spectrum, future studies will trace prescriptive and descriptive practices within the grammars, for instance comparatively to other centuries, but also by investigating their use of verbal hygiene and shifts in grammatical terminology.
There is an increase in referencing persons in the 17th century, particularly in Hewes and Wilkins, indicating the grammar authors have access to more sources due to the spread of the printing press. The expanding print culture also made it easier for the intellectual community to be aware of other experts of language. Moreover, keeping in line with the growing emphasis on academic rigor and credibility, referring to other authoritative figures lent legitimacy to the grammarians’ work and demonstrated their knowledge of the scholarly tradition. The interdisciplinary nature of the persons being referenced is observed in both centuries, and the influence of the Latinate tradition persists in the 17th century.
In comparison with the 16th century, there is an increase in the number of referencing across all grammar books to other grammar authors and contemporary scholars, which is a reflection of the growing scholarly discourse on language. Furthermore, the impact of literature giants such as William Shakespeare and Christopher Marlowe during the Elizabethan era led to poets/literary scholars being influential in the 17th century. Other factors included the rise in vernacular literature and political upheaval in the Restoration period in the latter part of the century. While the 17th century grammar authors still cited ancient scholars and political figures for intellectual legitimacy, they drew from a wider range of influences.
There is a salience of quotations in the 17th century data, particularly in Hewes, Jonson, Newton, and Wheeler. The reason for this might be due to citation practices becoming more formalized (see Blair 2010Blair, Ann M. 2010 Too Much to Know: Managing scholarly information before the Modern Age. New Haven, CT: Yale University Press.Blair, Ann M. 2010 Too Much to Know: Managing scholarly information before the Modern Age. New Haven, CT: Yale University Press.) and authors wanting to provide evidence for their arguments and distinguish their work from that of others. Although quotations are common in some of the 16th-century books such as Sherry and Mulcaster, there is a shift from the exploratory nature to more established scholarly conventions in the 17th century. While the 16th century laid the groundwork for grammatization, the 17th century expanded the discussions on language standardization, usage, and grammar rules.
Methodologically, the use of networks for visualization purposes allows researchers to have a broader overview of all the references thus enabling easier ways of determining links between authors that are mentioned within the network. We are also able to see whether some grammarians perhaps share no common features nor connections to their contemporaries. Our IRR scores for the 17th century were high for both the person types and reference categories, implying that the categories are applicable for both datasets.
Further reference analyses on 18th- and 19th-century grammars, which will be revisited in more depth, will provide insights into whether or not these pioneers of English grammar writing have exerted any influence on their successors. Centrality scores, such as those in Section 5.3 above, will enable a diachronic consideration of influence, showing which authors fade in or out of the spotlight. Additionally, a comparative analysis across all centuries will provide insight into whether the overall number of references increases over time as printing production increases and information becomes more widely accessible. This approach will enable us to examine how perspectives on grammar have shaped contemporary linguistic theories and which figures played a pivotal role in the development of English grammar during the Early Modern English period.
Funding
Open Access publication of this article was funded through a Transformative Agreement with University of Cologne.
Acknowledgements
We would also like to thank Ingo Kleiber for his contributions to the HeidelGram project in its earlier phases as well as our team of research assistants who worked on transcribing the grammar books: Alessia Carrone, Savvas Katsidonis, Charlotte Lüders, Julia Marcus, Elisa Pizzo, Matteo Schmelzer, and Marie Steinbrügge.
References
16th-Century Corpus Data
17th-Century Corpus Data
Appendix 1.Python Scripts for Reference Extraction
”””
This is a script to read annotated grammar text files, clean out annotations within the person / work tags. Within this process, full forms of abbreviations are kept,
line breaks are removed, in order to have as uniform results as possible
The output of this preprocessing step is a csv table of the following shape:
source text | left context | reference | right context
”””
import glob
from logging import getLogger
import os
import re
import numpy as np
from pathlib import Path
base_path = Path(__file__).resolve().parent
wdir = base_path / “Input”
file_list = glob.glob(os.path.join(os.getcwd(), wdir, “*.txt”))
# define results array, which contains: [source text; left context; reference; right context]
references_person = [[‘source text’, ‘left context’, ‘reference’, ‘right context’]]
person_string = ‘<person>([A-Za-z.,\n\’\-ӕœ= ]*)<\/person>’
references_work = [[‘source text’, ‘left context’, ‘reference’, ‘right context’]]
work_string = ‘<work>([A-Za-z.,\n\’\-ӕœ= ]*)<\/work>’
# load in the texts
for file_path in file_list:
with open(file_path, encoding=’UTF-8’) as f_input:
text_anno = f_input.read()
filename_pre = f_input.name
filename = filename_pre.replace(base_path / “Input\\”, “”).replace(”ID-[0-9]*-[0-9]*-”, “”).replace(”-[a-zA-Z0-9\_]*-anno.txt”, “”)
print(filename)
#clean the text
#remove all comments of this format <!--Bodleian Libraries Scan Note-->
text_clean = re.sub(”\<\!\-\-.*?\-\-\>”, “”, text_anno)
#remove the header
text_clean = re.sub(”[^\n]*LEVEL\n\n[^\n]*\n\n\n\n\n”, “ “*300, text_clean)
#remove linebreaks
text_clean = text_clean.replace(”-\n”, “”).replace(”- \n”, “”).replace(”=\n”, “”).replace(”= \n”, “”).replace(” \n”, “ “).replace(”\n”, “ “)
#replace | with; in the texts, as | is our table delimiter in the end
text_clean = text_clean.replace(”|”, “;”)
#remove simple tags (with no additional information) as well as zero tags
text_clean = text_clean.replace(”<lig>”, “”).replace(”</lig>”, “”).replace(”<i>”, “”).replace(”</i>”, “”).replace(” “, “ “)
text_clean = text_clean.replace(”<bl>”, “”).replace(”</bl>”, “”).replace(”<b>”, “”).replace(”</b>”, “”).replace(”<BlankPage/>”, “”)
text_clean = text_clean.replace(”<u>”, “”).replace(”</u>”, “”).replace(”<DoublePagePDF/>”, “”)
text_clean = text_clean.replace(”<?>”, “”).replace(”</?>”, “”).replace(”<MissContent/>”, “”)
text_clean = text_clean.replace(”<fig>”, “”).replace(”</fig>”, “”)
# remove complex tags (including added information)
# remove abbreviation tag, but keep the full form, rather than the original, except if fullform is Unknown or UnknownName
# this is done by removing Unknown tags first (leaving the abbreviated form). Then the remaining <a> tags will be with different content
text_clean = re.sub(”<a FullForm=\”UnknownName\” *>”, “”, text_clean)
text_clean = re.sub(”<a FullForm=\”Unknown\” *>”, “”, text_clean)
text_clean = re.sub(”<a FullForm=\”\” *>”, “”, text_clean)
# now remove all remaining <a> tags (leaving the fullform)
text_clean = re.sub(”<a FullForm=\””, “”, text_clean)
text_clean = re.sub(”\” *>[\.A-Za-zæᶜʳ \:]*<\/a>”, “”, text_clean)
text_clean = re.sub(”</a>”, “”, text_clean)
# macrons opening
text_clean = re.sub(”<m FullForm=\”[A-Za-z ]*\” *>”, “”, text_clean)
# language tags opening
text_clean = re.sub(”<l Language=\”[A-Za-z ]*\” *>”, “”, text_clean)
# image tags opening
text_clean = re.sub(”<im [A-Za-z ]*=\”[-\_A-Za-z0-9\.]*\” *>[A-Za-z ]*<\/im>”, “”, text_clean)
# hard to place tags opening
text_clean = re.sub(”<h LocationOnPage=\”[A-Za-z ]*_*[A-Za-z ]*\” *>”, “”, text_clean)
# comment tags opening
text_clean = re.sub(”<c Description=\”[A-Za-z ]*_*[A-Za-z ]*\” *>”, “”, text_clean)
# recurring parts tag opening
text_clean = re.sub(”<r Form=\”[A-Za-z ]*_*[A-Za-z ]*\” *>”, “”, text_clean)
# closing tags
text_clean = re.sub(”</[m,l,c,h,r]>”, “”, text_clean)
#remove tables (partly, leaving only the table tags for easier manual access and clean-up)
text_clean = text_clean.replace(”<cell>”, “”).replace(”</cell>”, “”).replace(”<row>”, “”).replace(”</row>”, “”).replace(”<head>”, “”).replace(”</head>”, “”)
refs = re.findall(’(?=(.{150})<person>(.*?)<\/person>(.{0,150}))’, text_clean)
#<person>([A-Za-z.,\n\’\-ӕœæ=\: ]*)<\/person>
ref = re.findall(person_string, text_clean)
list_refs = []
list_refs = np.asarray(refs)
list_refs = np.insert(list_refs, 0, filename, axis=1)
print(list_refs.shape)
references_person = np.append(references_person, list_refs, axis=0)
print(references_person.shape)
#print the resulting references table into a csv file
np.savetxt(base_path / “person_references.csv”, references_person, delimiter=’|’, fmt=’%s’, encoding=”utf-8”)
”””
This is a script to read the manually revised output of the preprocessing step. The table (csv) output from the preprocessing step is manually enhanced.
The input data has the following shape:
source text | left context | reference | right context | normalized reference | author type | author country of publication | reference type 1 | reference type 2
The data is visualized in different networks representing:
-
author references to all persons (not examples), displaying author and reference types
-
author references to all persons, displaying person’s country affiliation
”””
from logging import getLogger
from turtle import shape
import numpy as np
from numpy import loadtxt
from pathlib import Path
base_path = Path(__file__).resolve().parent
# load in the data
ext_person_refs = open(base_path / “person-references-16th-extend-agreed-no-examples.csv”, ‘rb’)
data = loadtxt(ext_person_refs, delimiter = “|”, dtype=’str’, encoding=’UTF-8’)
# create new arrays with only the needed information
data_format_extended = [[‘source’, ‘target’, ‘count’, ‘author type’, ‘author country’, ‘reference type 1’, ‘reference type 2’]]
data_format_simple = [[‘source’, ‘target’, ‘count’]]
network_data_extended = np.asarray(data_format_extended)
network_data_simple = np.asarray(data_format_simple)
# fill up array with extended info:
j=1
for i in range(1,data.shape[0]):
if i==1:
network_data_extended = np.append(network_data_extended, [[data[i][0],data[i][4],1,data[i][5],data[i][6],data[i][7],data[i][8]]], axis=0)
elif i>1:
for k in range(1,j+1):
if data[i][0] == network_data_extended[k][0] and data[i][4] == network_data_extended[k][1] and data[i][7] == network_data_extended[k][5]: #if the same source made the same type of reference to the same target before
network_data_extended[k][2] = int(network_data_extended[k][2]) + 1 #increase count by one, rest stays the same
if all([[data[i][0],data[i][4],data[i][7]]] != [[network_data_extended[k][0],network_data_extended[k][1],network_data_extended[k][5]]] for k in range(1,j+1)):
j = j+1
network_data_extended = np.append(network_data_extended, [[data[i][0],data[i][4],1,data[i][5],data[i][6],data[i][7],data[i][8]]], axis=0)
# count references from source to target authors, shape = (rows, columns)
j=1
for i in range(1,data.shape[0]):
if i==1:
network_data_simple = np.append(network_data_simple, [[data[i][0],data[i][4],1]], axis=0)
elif i>1:
for k in range(1,j+1):
if data[i][0] == network_data_simple[k][0] and data[i][4] == network_data_simple[k][1]: #if the same source made reference to the same target before
network_data_simple[k][2] = int(network_data_simple[k][2]) + 1 #increase count by one, rest stays the same
if all([[data[i][0],data[i][4]]] != [[network_data_simple[k][0],network_data_simple[k][1]]] for k in range(1,j+1)):
j = j+1
network_data_simple = np.append(network_data_simple, [[data[i][0],data[i][4],1]], axis=0)
#print the resulting references table into a csv file
np.savetxt(base_path / “person_references_17th_network_counts.csv”, network_data_simple, delimiter=’|’, fmt=’%s’, encoding=”utf-8”)
np.savetxt(base_path / “person_references_17th_network_extended.csv”, network_data_extended, delimiter=’|’, fmt=’%s’, encoding=”utf-8”)
# generate node information on source nodes
source_nodes = [[‘name’, ‘author type’, ‘author country’]]
source_nodes = np.asarray(source_nodes)
source_nodes = np.append(source_nodes, network_data_extended[1:,[0,3,4]], axis=0)
#print(source_nodes)
for i in range(1,source_nodes.shape[0]):
source_nodes[i][1] = ‘Source’
source_nodes[i][2] = ‘America’
# store node information from extended network data
node_data = [[‘name’, ‘author type’, ‘author country’]]
node_data = np.asarray(node_data)
node_data = np.append(node_data, network_data_extended[1:,[1,3,4]], axis=0)
node_data = np.append(node_data, source_nodes[1:][:], axis=0)
# make node information unique
unique_node_data = np.unique(node_data[1:][:], axis=0)
#print the resulting nodes table into a csv file
np.savetxt(base_path / “person_references_17th_network_nodes.csv”, unique_node_data, delimiter=’|’, fmt=’%s’, encoding=”utf-8”)
Appendix 2.Python Script for IRR Evaluation
”””
This is a script to calculate IRR scores for agreement regarding reference and person categories between two independent raters.
”””
from pathlib import Path
import random
import sys
import nltk
from nltk.metrics import masi_distance
from nltk.metrics import jaccard_distance
from nltk.metrics.agreement import AnnotationTask
import pandas as pd
from pathlib import Path
base_path = Path(__file__).resolve().parent
def generate_test_data(n):
authors = [‘AuthorA’, ‘AuthorB’]
categories = [‘C1’, ‘C2’, ‘C3’, ‘C4’]
for i in range(n):
cats = ‘;’.join([random.choice(categories) for _ in range(random.randint(3,4))])
print(f’X;Y;;{random.choice(authors)};{random.choice(authors)};;{cats}’)
def kappa_int(k):
if k <= 0:
return ‘This is really bad! 😖’
if 0.01 <= k <= 0.20:
return ‘This is bad! 😪’
if 0.21 <= k <= 0.40:
return ‘This is still pretty bad! 🤔’
if 0.41 <= k <= 0.60:
return ‘Not great, not terrible! 😐’
if 0.61 <= k <= 0.80:
return ‘This is really good! 😀’
if 0.81 <= k <= 1.0:
return ‘This is absolutely fantastic! 🤩’
def alpha_int(a):
if a <= 0.01:
return ‘This is really bad! 😖’
if 0.01 <= a <= 0.666:
return ‘This is bad! 😪’
if 0.667 <= a <= 0.79:
return ‘This is acceptable! 😀’
if a >= 0.8:
return ‘We\’re good! 🤩’
def irr_author(data):
rater_a = list(data[‘Author’])
rater_b = list(data[‘Author’])
categories = set(rater_a + rater_b)
table = [[‘rater_a’, i, category] for i, category in enumerate(rater_a)] + [[‘rater_b’, i, category] for i, category in enumerate(rater_b)]
rating = AnnotationTask(data=table)
print(’# Author Category\n’)
print(f’Sanity Check: There are {len(categories)} categories ({”, “.join(categories)})\n’)
print(f’AoA: {rating.avg_Ao()}’)
print(f’Kappa: {rating.kappa()}, {kappa_int(rating.kappa())}’)
print(f’Kr. Alpha: {rating.alpha()}, {alpha_int(rating.alpha())}’)
def irr_references(data):
table = []
categories = []
for id, r in data.iterrows():
rater_a_cats = []
if type(r[‘RC1Charlotte’]) == str:
rater_a_cats.append(r[‘RC1’])
categories.append(r[‘RC1’])
if type(r[‘RC2Charlotte’]) == str:
rater_a_cats.append(r[‘RC2’])
categories.append(r[‘RC2’])
rater_a = [‘rater_a’, id, frozenset(rater_a_cats)]
rater_b_cats = []
if type(r[‘RC1Savvas’]) == str:
rater_b_cats.append(r[‘RC1’])
categories.append(r[‘RC1’])
if type(r[‘RC2Savvas’]) == str:
rater_b_cats.append(r[‘RC2’])
categories.append(r[‘RC2’])
rater_b = [‘rater_b’, id, frozenset(rater_b_cats)]
# Check for bad conditions
if len(rater_b_cats) == 0 or len(rater_b_cats) == 0:
print(f’One of the raters did not provide labels for a row!’)
print(f’Questionable Row:\n\n{r}’)
sys.exit(’\nExiting’)
table.append(rater_a)
table.append(rater_b)
categories = set(categories)
rating_jaccard = AnnotationTask(data=table, distance=jaccard_distance)
rating_masi = AnnotationTask(data=table, distance=masi_distance)
rating_bd = AnnotationTask(data=table)
print(’# Reference Categories\n’)
print(f’Sanity Check: There are {len(categories)} categories ({”, “.join(categories)})\n’)
print(f’AoA (Jaccard): {rating_jaccard.avg_Ao()}’)
print(f’AoA (Masi): {rating_masi.avg_Ao()}’)
print(f’AoA (BD): {rating_bd.avg_Ao()}’)
print(f’Kappa (Jaccard): {rating_jaccard.kappa()}, {kappa_int(rating_jaccard.kappa())}’)
print(f’Kappa (Masi): {rating_masi.kappa()}, {kappa_int(rating_masi.kappa())}’)
print(f’Kappa (BD): {rating_bd.kappa()}, {kappa_int(rating_bd.kappa())}’)
print(f’Kr. Alpha (Jaccard): {rating_jaccard.alpha()}, {alpha_int(rating_jaccard.alpha())}’)
print(f’Kr. Alpha (MASI): {rating_masi.alpha()}, {alpha_int(rating_masi.alpha())}’)
print(f’Kr. Alpha (BD): {rating_bd.alpha()}, {alpha_int(rating_bd.alpha())}’)
def irr(data_path, lowercasing=True):
data = pd.read_csv(Path(data_path), delimiter=’|’, skip_blank_lines=True)
# Lowercasing
if lowercasing:
columns_to_lower = list(data.columns)
for c in columns_to_lower:
try:
data[c] = data[c].str.lower()
except AttributeError:
pass
irr_author(data)
print(’-’ * 50)
irr_references(data)
if __name__ == “__main__”:
irr(base_path / “17th-references-Cat-IRR.csv”)
Appendix 3.R Script for Network Visualization
#libraries
install.packages(”igraph”, repos=’http://cran.us.r-project.org’)
install.packages(”tidyverse”)
library(”igraph”)
library(”tidyverse”)
##############################################################################
# load data: full reference information
reference_data_full = read.csv(”/%PATH/person_references_network_extended.csv”,header=TRUE, sep = “|”)
# load data: node only information
node_data_full = read.csv(”/%PATH/person_references_network_nodes_length.csv”,header=FALSE, sep = “|”)
##############################################################################
# CREATE BASE NETWORK DATA
reference_network <- graph_from_data_frame(reference_data_full[, c(1,2,3,6)], directed=TRUE, vertices=node_data_full)
##############################################################################
# SET EDGE COLOR ACCORDING TO REFERENCE TYPE
E(reference_network)$color <- NA
E(reference_network)$color <- ifelse(E(reference_network)$reference.type.1 == “Quotation”, “#C17E9B”,”black”) #sets black for else, i.e. here = Example
E(reference_network)$color <- ifelse(E(reference_network)$reference.type.1 == “Opinion”, “#457A93”, E(reference_network)$color)
E(reference_network)$color <- ifelse(E(reference_network)$reference.type.1 == “Comparison/Contrast”, “#DD7331”, E(reference_network)$color)
E(reference_network)$color <- ifelse(E(reference_network)$reference.type.1 == “Acknowledgement”, “#743881”, E(reference_network)$color)
E(reference_network)$color <- ifelse(E(reference_network)$reference.type.1 == “Mention”, “#7A8A65”, E(reference_network)$color)
E(reference_network)$color <- ifelse(E(reference_network)$reference.type.1 == “Application”, “#454545”, E(reference_network)$color)
# SET NODE(VERTEX) COLOR AND SHAPE ACCORDING TO PERSON TYPE
V(reference_network)$color <- NA
V(reference_network)$color <- ifelse(V(reference_network)$V2 == “Grammar Author”, “#743881”,”black”) #sets black for else, i.e. here = Source
V(reference_network)$color <- ifelse(V(reference_network)$V2 == “Ancient Scholar”, “#454545”, V(reference_network)$color)
V(reference_network)$color <- ifelse(V(reference_network)$V2 == “Poet/Literary Author”, “#C17E9B”, V(reference_network)$color)
V(reference_network)$color <- ifelse(V(reference_network)$V2 == “Political Figure”, “#7A8A65”, V(reference_network)$color)
V(reference_network)$color <- ifelse(V(reference_network)$V2 == “Religious Figure”, “#457A93”, V(reference_network)$color)
V(reference_network)$color <- ifelse(V(reference_network)$V2 == “Contemporary Scholar”, “#DD7331”, V(reference_network)$color)
# SET NODE(VERTEX) COLOR AND SHAPE ACCORDING TO PERSON TYPE
V(reference_network)$shape <- NA
V(reference_network)$shape <- ifelse(V(reference_network)$V2 == “Source”, “square”,”circle”) #sets square for source nodes, circle for all target nodes
##############################################################################
# add scaling for number of references. Scaled into [1,5], i.e. 1 reference = width 1, max references = width 5
max_count <- max(reference_data_full$count)
# add scaling for length of source books. Scaled to [5,8], i.e. targets are 5, smallest book is 5, longest book is 8
max_length <- max(node_data_full$V4)
# SET NODE LABELS TO SHOW ONLY THE SOURCE LABELS (for 17th-century data)
V(reference_network)$label <- V(reference_network)$name
V(reference_network)$label <- ifelse(V(reference_network)$V2 == “Source”, V(reference_network)$label,NA)
V(reference_network)$label
# Layout
nicelayout <- layout_nicely(reference_network)
jpeg(”R_Network.jpeg”, width = 22, height = 22, units = ‘in’, res = 300)
plot(reference_network, layout = nicelayout,
vertex.frame.color = “black”, vertex.label.color = “black”,
vertex.size=(((V(reference_network)$V4)-1)/(max_length-1)*(6-1)+4), #For no scaling use: vertex.size=5,
vertex.label.dist=0.8, vertex.label.degree=pi/2,
vertex.label.family = “sans”, vertex.label.cex=1.5,#1
edge.width=(((E(reference_network)$count)-1)/(max_count-1)*(5-1)+1), #For no scaling use: edge.width=E(reference_network)$count,
edge.arrow.size=0.7, edge.arrow.width=1.2)
# ADD A LEGEND FOR AUTHOR TYPES
legA=c(”Ancient Scholar”, “Contemporary Scholar”, “Grammar Author”, “Poet/Literary Author”, “Political Figure”, “Religious Figure”, “Source Node”)
colA=c(”#454545”, “#DD7331”, “#743881”, “#C17E9B”,”#7A8A65”, “#457A93”, “black”)
legend(x=−1.1,y=−0.925, legend=legA, col=colA, pch=16, title=”Person Types (Nodes)”, cex = 1.7)
#legend=levels(as.factor(V(reference_network)$V2)),
#”bottomleft”,
# ADD A LEGEND FOR REFERENCE TYPES
legB=c(”Acknowledgement”, “Application”, “Comparison/Contrast”, “Mention”, “Opinion”, “Quotation”)
colB=c(”#743881”, “#454545”, “#DD7331”, “#7A8A65”, “#457A93”,”#C17E9B”)
legend(x=−0.72, y=−0.96, legend=legB, col=colB, pch=16, title=”Reference Types (Edges)”, cex = 1.7)
# Close device
dev.off()