To save content items to your account,
please confirm that you agree to abide by our usage policies.
If this is the first time you use this feature, you will be asked to authorise Cambridge Core to connect with your account.
Find out more about saving content to .
To save content items to your Kindle, first ensure no-reply@cambridge.org
is added to your Approved Personal Document E-mail List under your Personal Document Settings
on the Manage Your Content and Devices page of your Amazon account. Then enter the ‘name’ part
of your Kindle email address below.
Find out more about saving to your Kindle.
Note you can select to save to either the @free.kindle.com or @kindle.com variations.
‘@free.kindle.com’ emails are free but can only be saved to your device when it is connected to wi-fi.
‘@kindle.com’ emails can be delivered even when you are not connected to wi-fi, but note that service fees apply.
This chapter sets out the book’s aim: an attempt to map the changing lexicons of Christian expression in the English language from the end of the Middle Ages to the early Victorians. It is thus a contribution to historical theolinguistics, drawing on the one hand on corpus linguistics and, on the other, the history of the many varieties of English Christianity that existed during this period. The methodology to be adopted in the body of the book is then illustrated with reference to two contrasting groups of female writers active in the middle of the seventeenth century, a period of exceptional religious turmoil: Fifth Monarchists and Quakers. The special analytic tools and resources harnessed include the Oxford English Dictionary, Semantic EEBO, Glasgow’s Historical Thesaurus of English, UCREL’s Log-Likelihood Calculator, and Sheffield’s Linguistic DNA. The remainder of the chapter defines notions and addresses issues that will recur later in the book: notions of register and genre, and questions to do with the handling of corpora, and with the structuring of religious ‘communities of practice’.
This chapter discusses how corpus stylistics relates to other fields in terms of methodologies and explanatory purposes. It proposes a specific take on corpus stylistics, which makes it possible to both focus on meanings in individual texts and identify features that are shared across texts. This approach brings considerations of register and style together, and it has clear relevance to literary scholarship. To illustrate this kind of corpus stylistics, the chapter presents Mahlberg (2013) as a case study on patterns of characterization in works by Charles Dickens. Mahlberg (2013) addresses the following question: In a corpus of texts by Dickens, can textual patterns be identified that have discernible functions in the creation of fictional worlds? The study uses the concept of local textual functions to combine corpus linguistic and literary perspectives in the analysis of n-grams, and it proposes a novel approach to characterization in Dickens. What this case study illustrates for corpus stylistics more widely is how the application of relative basic corpus linguistic methods can support the creation of a coherent theoretical approach to explain meanings in texts.
Any research question or application relating to language variation and/or use can be approached from a corpus linguistic perspective. The goals of The Cambridge Handbook of English Corpus Linguistics (CHECL) are to survey the breadth of these research questions and applications in relation to the linguistic study of English. The handbook addresses a range of topics, including chapters on lexical variation, grammatical variation, historical change, online language and social media, multimodal corpora, the linguistic description of dialects and registers, and applications to language teaching and translation. In each chapter, authors assess what we have learned from corpus-based investigations to date and provide detailed case studies that illustrate how corpus analyses can be employed for empirical descriptions, documenting surprising patterns of language use that are often unanticipated previously. By bringing together diverse perspectives and cutting-edge research, this volume serves as a comprehensive resource for researchers seeking to understand and apply a corpus-based approach to the study of English.
This chapter provides a critical overview of how L2 researchers have used measures of lexical and phraseological complexity, with a special focus on their definition and operationalization. Lexical and phraseological complexity are often measured in terms of diversity (the number of different words or multi-word units respectively) and sophistication (the number of ‘sophisticated’ words or multi-word units, with ‘sophisticated’ variously understood as less frequent, more specialized, or in the case of phraseological sophistication, as more strongly associated word combinations). With the development of natural language processing (NLP) tools such as the Lexical Complexity Analyzer (Lu, 2012) or the Tool for the Automatic Analysis of Lexical Sophistication (Kyle et al., 2018), L2 researchers can now analyze lexical complexity in L2 English texts using dozens or even hundreds of measures with a click of a button. However, unresolved questions remain: How accurate, reliable, and valid are these measures across learner samples and research contexts? Can they assess L2 proficiency across registers? How should measures be selected from the many available? By contrast, do the current, limited set of measures of phraseological complexity represent the full construct, or is there a need to develop for further development?
This chapter starts out from the serendipitous observation of patterns where look forward to is complemented not by a verbal -ing form (as would be expected by standard grammars) but by a plain infinitive, as in I’m looking forward to meet you. It investigates the spread and frequency of this construction in a wide range of corpora, some of which also allow a diachronic comparison, and considers several potentially conditioning factors. Grammatically, this pattern can be explained as a case of reanalysis in the process of structural transmission, showcasing structural simplification that may characterise second-language acquisition. The form to, the final element in the phrasal-prepositional verb under discussion, is reinterpreted as a homonymic but functionally different form, the to infinitive marker. The results clearly support the hypothesis that the ’look forward to + plain infinitive’ pattern constitutes more than mere performance errors and appears to represent an ’embryonic’ manifestation of an incipient change, with a strong focus in Asian Englishes, and possibly with Indian English as an epicentre, although the trend may have been reversed recently in formal registers.
This article explores how people imagine a 21st-century persona-indexing register as emanating from a 19th-century person-indexing register. The persona in question is the Philippine conyo, regarded as spoiled, empty-headed, rich kids who speak a distinct style of “Taglish” (Tagalog-English). The person in question is José Rizal, one of the most celebrated Filipino historical figures. Drawing on ethnographic and media data, I trace how Rizal is regarded as “the original conyo,” as its first author or animator. Examining how this type-token interdiscursive link between persona and person plays on the inversion of a chronotopic frame, I consider what conceptualizing elite historical continuity accomplishes socially and economically. I argue that citing Rizal creates a channel to move value.
This chapter examines the nature of tone sandhi and various other tonal mutations in Taiwanese Southern Min (TSM). Each base tone in this language corresponds to a specific sandhi tone, with sandhi resulting from two sets of tonal shifts: smooth tone chain shifts and checked tone chain shifts. Each shift modifies either register or pitch, but not both simultaneously.
Some experimental studies have reported low rates of sandhi application, suggesting limited productivity. However, evidence from both experiments using real words and corpus analysis reveals high rates of appropriate tonal alternations, indicating that productivity is the primary mechanism. Theoretical works have further elaborated the tonal alternations as systematic chain shifts, lending support to this productivity-based view. The evidence suggests that future models of TSM tone sandhi should primarily incorporate productive phonological processes, supplemented by selective lexical storage mechanisms for certain exceptional or high-frequency cases.
In diminutive suffixation, the tone of the pre-á syllable undergoes modification through dextrosinistral spreading of register and/or pitch from the -á suffix, whereby the derived [35] ([Lr, h][Hr, h]) tonal output emerges as a distinctive tone cluster. Conversely, in neutral tone operations, sinistrodextral tone spreading applies to a subsequent function word, which may alternatively acquire a low tone by default in the absence of such spreading.
Different texts have different characteristics. In this chapter, we first explore the concepts of register, genre and style, which are, in the tradition of Biber, linked to communicative functions and situational characteristics. The co-occurrence of register features and dimensions are introduced as the linguistic indicators of communicative functions. A particularly useful approach to register centres around keyness, which we demonstrate with historical Portuguese data. We then introduce discourse traditions as a historical-linguistic concept closely related to genre and register. We use French literary examples to explain stylistic differences and the link with the Labovian distinction between indicators, markers and stereotypes. This leads to a discussion of indexicality and indexical fields more generally, for which we draw on ancient Greek plays. The chapter continues the discussion of the literary representation of language variation on the basis of English texts comprising dialect, and explains the important concept of enregisterment.
This Element presents a computational theory of syntactic variation that brings together (i) models of individual differences across distinct speakers, (ii) models of dialectal differences across distinct populations, and (iii) models of register differences across distinct contexts. This computational theory is based in Construction Grammar (CxG) because its usage-based representations can capture differences in productivity across multiple levels of abstraction. Drawing on corpora representing over 300 local dialects across fourteen countries, this Element undertakes three data-driven case-studies to show how variation unfolds across the entire grammar. These case-studies are reproducible given supplementary material that accompanies the Element. Rather than focus on discrete variables in isolation, we view the grammar as a complex system. The essential advantage of this computational approach is scale: we can observe an entire grammar across many thousands of speakers representing dozens of local populations.
This chapter investigates the diction of the fragments attributed to Ennius’ Saturae by ancient sources and conjecturally by modern editors. While thirty or so transmitted lines naturally do not permit one to paint a conclusive picture of Ennius’ experiment, a little more can be said about the relationship between his Saturae and those of Lucilius, and ultimately about Ennius’ role in the introduction of personal poetry at Rome. Monologic and dialogic utterances and the mixture of metres (iambo-trochaic, hexameter, Sotadean) and registers (comic, informal, mock-epic) will be discussed, using Lucilius as a comparandum. Attention is paid to “early” features of language and style, with reference to Ennius’ diction in his epic and dramatic works.
Chapter 6 aims to help readers understand how variation and change affect language, so that translation practices and decisions are not based on personal biases and lay views about language but, rather, on a principled understanding of how language interacts with society. Another goal is to create awareness of the impact of social and use-related (contextual) factors on language so that translated texts respond to the requirements of the translation instructions. Other sociolinguistic notions reviewed in this chapter, along with their implications for translation are register, dialectal variation, socioeconomic variation, the nature of language change and variation, prestigious varieties vs. stigmatized varieties, and translating in multilingual societies. The discussion of register includes field of activity, medium and level of formality, as well as the implications for translation of not considering these within the context of the translation brief and translation norms. The connection between register selection and linguistic and translation competence is explained. Illustrative examples are used throughout the chapter.
In this article, I analyse the word-prosodic system of Drubea and Numèè, two of the rare tonal Oceanic languages. Building on Rivierre’s (1973) seminal work, I show that the word-prosodic system of these two languages can be analysed as involving only register features: an underlying downstep and a postlexical epenthetic upstep. Drubea and Numèè are thus tonal languages without tones stricto sensu. This new type of word-prosodic system has both theoretical and typological implications: (i) register features, defined as in Snider’s (1999) Register Tier Theory, need not be subordinate to or associated with tones, and may exist in the absence of tone, including in underlying representation; (ii) tonal systems come in two types: tone-based systems in which the tonal contrasts are defined paradigmatically, as in most tone languages, and register-based systems where tonal contrasts are defined syntagmatically, as in Drubea and Numèè.
Recent studies in Construction Grammar have suggested that contracted modals constitute different constructions from their full forms. In this article, we present a corpus-based analysis of the relationship between the modal forms going to and gonna in British English used on the blogging platform LiveJournal. We report a Collostructional Analysis and a Behavioural Profile Analysis based on a logistic regression model of blind annotations, assessing factors of semantic, pragmatic and social meaning on the choice of the variant, in addition to processing factors. The results show that register formality is the only significant meaning predictor for the alternation between going to or gonna in the corpus. We discuss these results in light of recent theoretical debates on isomorphism and synonymy avoidance in Construction Grammar: specifically, our study provides evidence that social meaning drives the distinction between going to and gonna, validating the recently formulated Principle of No Equivalence, and providing further evidence for the constructionhood of contracted modals.
Registers have proved to be powerful proxies for language variation and stylistic change in historical research. This chapter investigates five sub-registers within the domain of scientific discourse: philosophy (humanities), history (social sciences), life sciences and astronomy (natural sciences) and medical texts. With data from the Coruña Corpus of Scientific Writing and the corpus of Late Modern English Medical Texts, we carry out a Multi-dimensional analysis of one million words of eighteenth-century scientific English, this leading to the scaling of the five sub-registers along two main dimensions of variation: ‘Involved/Interpersonal versus Narrative/Abstract’ and ‘Complex/Elaborate versus Non-elaborate’ discourse. The analysis confirms, first, that there are substantial differences among sub-registers in terms of the distribution and pervasiveness of distinctive linguistic features, and, second, that fluctuation in prose discourse is a general characteristic of Late Modern English scientific writing.
Chaucer’s works were written during the late fourteenth century, a period which saw considerable changes in the functions of the English language as it came to replace French and Latin as the languages of written record. As well as being an important source for the scholarly understanding of late Middle English, Chaucer’s works shed light on the status of English and its variety of registers and dialects, enabling scholars to gain a deeper awareness of the sociolinguistic connotations of its different forms and usages. The Canterbury Tales, with its array of pilgrims drawn from a variety of professions, social classes and geographical regions narrating a series of tales reflecting a wide range of genres, is a valuable source of evidence for historical pragmatics. This chapter shows the way in which Chaucer’s text offers insights into the conventions of social interaction, including forms of address, politeness and verbal aggression, and the use of discourse markers.
What counts as scientific writing has undergone massive changes over the centuries. Medical writing is a good representative of the register of scientific English, as it combines both theoretical concerns and practical applications. Ideas of health and sickness have been communicated in English written texts for over a thousand years from the Middle Ages to the present, with different traditions and layers of writing reflecting literacy developments and changing thought-styles. This chapter approaches the topic from the perspective of registers and genres, considering how texts are shaped by their functions and communicative purposes and various audiences. Some genres run throughout the history of English: remedy books were already extant in the Old English period. Another core genre, the case study, mirrors wider scientific developments in response to changes in styles of thinking: medieval scholasticism is gradually replaced by a growing interest in increasingly systematic empirical observation. The establishment of learned societies from the seventeenth century onwards gives rise to new genres like the experimental report, and concomitant disciplinary advances and technological developments in the following centuries gradually pave the way for modern evidence-based medicine. Today medical advances are communicated in digital publications to a worldwide readership.
This chapter provides an overview of the language of religious texts in Old, Middle and Early Modern English. We divide religious language into three spheres: Bible language, the language of prayers and the language of texts of religious instruction and discussion. We then discuss the language of religious texts against the background of the impact of the language of the vernacular Bible, particularly before 1500. We argue that, prior to the publication of the King James Bible, there was no specific ‘religious register’ in Old and Middle English, and even in Early Modern English a typically ‘religious style’ is found only as an additional layer in religious texts, which, by and large, follow the general standardising tendencies of the language at the time.
The vernacular historiographical tradition has evolved since the ninth century through merging core genres like annals, chronicles and historical narrative with empirical antiquarian treatises. It became clearly distinguished from religious and fictional writing only in post-medieval times, thus also adapting its concept of truth and its methods. While its earlier history is best described by way of a discourse tradition (Koch), a Wengerian community of practice emerges in the late modern period. Starting off as a purely narrative text-type, historiographical writing developed into a typical narrative–expository–argumentative conglomerate over the early and late modern periods. The heteroglossia so typical of historiography becomes less literary or dramatic and more evidential in nature, also evolving citation styles and footnotes. The evaluative and ideological potential of historiography is present from the start and realised by such means as group/person labels, evaluative lexis and superlatives.
This chapter provides an overview of ways to study, categorise, and analyse non-canonical syntactic patterns in registers of English. It introduces two distinct approaches to studying the role of discourse and register in determining syntactic variation. The first (‘variationist’) approach looks at non-canonical syntax as a case of grammatical variation with register as the predictor. The second (‘text-linguistic’) approach takes register as its proper object of investigation and looks at non-canonical constructions as frequent and pervasive features of a register. We classify non-canonical syntactic constructions according to their form as either reduced, expanded, or re-ordered versions of canonical clauses. Each of these patterns is exemplified in one of the studies that constitute the section of the volume introduced by this chapter (ellipsis as reduced constructions, clefts as expanded constructions, and particle placement as reordering). Comparing these studies, this chapter also elaborates on the role of corpus methods as well as experimental data in shaping research questions regarding the motivation for non-canonical patterns. A final part discusses trends and open questions, such as problems of register classification for text from media and new challenges presented to the field by generative AI tools.
This article presents an exploratory study of an innovative future adverb construction, going forward, typically meaning ‘in the future, from now on’ (e.g. What does this mean going forward?). Going forward probably originated in the domain of business in or around the 1970s. In this study, the spread of going forward is examined on the basis of over 1,500 examples from six genres of the Corpus of Contemporary American English (COCA), covering the years 1990–2019. The data is analysed in terms of four morphosyntactic variables, and the developments in the frequency of going forward are analysed using variability-based neighbour clustering. The results show that, in the 1990s, going forward had a modest rate of occurrence mainly in texts having to do with business and finance, but its frequency rose sharply in the 2000s and the 2010s. At the same time, the discourse contexts in which it appeared broadened from business and finance to other domains. The syntactic contexts of going forward show that it has become an adverb. The results highlight the need to incorporate social meanings such as domain preferences in the description of grammatical constructions. They also illustrate the need to consider constructional innovations at the lexical end of the grammar–lexicon continuum, in addition to highly grammaticalised constructions.