- Home
- Book Series
- Studies in Corpus Linguistics
Studies in Corpus Linguistics
<p>SCL focuses on the use of corpora throughout language study, the development of a quantitative approach to linguistics, the design and use of new tools for processing language texts, and the theoretical implications of a data-rich discipline. </p>
51 - 100 of 128 results
-
-
Diachronic Corpora, Genre, and Language Change
Editor(s): Richard J. WhittPublication Date November 2018show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This volume provides a state-of-the-art overview of the intersecting fields of corpus linguistics, historical linguistics, and genre-based studies of language usage. Papers in this collection are devoted to presenting relevant methods pertinent to corpus-based studies of the connection between genre and language change, linguistic changes that occur in particular genres, and specific diachronic phenomena that are influenced by genre factors to greater and lesser degrees. Data are drawn from a number of languages, and the scope of the studies presented here is both short- and long-term, covering cases of recent change as well as more long-term alterations.
-
-
-
Discourse in the Professions
Editor(s): Ulla Connor and Thomas A. UptonPublication Date November 2004show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book explores the structure and use of academic and professional discourse through the lens of corpus linguistics. The goal of this book is to show how insights from corpus linguistic analyses can help us better understand how we use academic and professional language and help us find ways to better train newcomers to the genres used in various professional contexts. The contributions to this book show that specialized corpora of specific genres from a variety of fields allow us to make more relevant observations about the function and use of language for particular purposes. The specialized corpora examined include written and spoken academic genres, written and spoken business and legal genres, and written philanthropic genres. The book showcases a variety of approaches to analyzing the discourse of specialized corpora, and each chapter concludes with a reflection on the practical and pedagogical implications of the analysis.
-
-
-
The Discourse of Nurse-Patient Interactions
Author(s): Shelley StaplesPublication Date December 2015show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:The Discourse of Nurse-Patient Interactions: Contrasting the communicative styles of U.S. and international nurses is the first book to quantitatively examine a wide range of linguistic features in a corpus of interactions between nurses and standardized patients. The main goal of this book is to compare the discourse of U.S. (L1 English speaking) and international (L2 English speaking) nurses. The research design relies on a mixed method approach, including both quantitative and qualitative discourse analysis of lexico-grammatical, interactional, prosodic, fluency, and non-verbal features; assessments of interactional effectiveness; and qualitative interviews with nurses. The book offers a detailed description of the situational characteristics of the interactions and compares the discourse of nurses and patients in order to contextualize differences in the communicative styles of the two nurse groups. The results provide new insight into the way that sociocultural and linguistic aspects of nurse discourse contribute to the delivery of patient-centered care.
-
-
-
Discourse on the Move
Author(s): Douglas Biber, Ulla Connor and Thomas A. UptonPublication Date September 2007show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Discourse on the Move is the first book-length exploration of how corpus-based methods can be used for discourse analysis, applied to the description of discourse organization. The primary goal is to bring these two analytical perspectives together: undertaking a detailed discourse analysis of each individual text, but doing so in terms that can be generalized across all texts of a corpus. The book explores two major approaches to this task: ‘top-down’ and ‘bottom-up’. In the ‘top-down’ approach, the functional components of a genre are determined first, and then all texts in a corpus are analyzed in terms of those components. In contrast, textual components emerge from the corpus analysis in the bottom-up approach, and the discourse organization of individual texts is then analyzed in terms of linguistically-defined textual categories. Both approaches are illustrated through case studies of discourse structure in particular genres: fund-raising letters, biology/biochemistry research articles, and university classroom teaching.
-
-
-
Discourse Reflexivity in Linear Unit Grammar
Author(s): Cameron SmartPublication Date July 2016show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Discourse Reflexivity in Linear Unit Grammar: The case of IMDb message boards represents a significant landmark. Not only is it the first in-depth corpus-based study to be based on Linear Unit Grammar, it is also the first study to present a unified model of both Linear Unit Grammar and Linear Unit Discourse Analysis. To illustrate this model, the book focuses on the role of discourse reflexivity in the linear structure of online message board discourse from the Internet Movie Database (IMDb) webpage. It is shown that discourse reflexivity plays a central role in the linear structure and antagonism characteristic of this type of discourse. This book will particularly appeal to those who have an interest in carrying forward the innovations in the description of grammar, lexis and discourse proposed by John Sinclair in his lifetime as well as to those with a specific interest in discourse reflexivity and computer-mediated communication.
-
-
-
Emotive Interjections in British English
Author(s): Ulrike StangePublication Date January 2016show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Emotive Interjections in British English: A corpus-based study on variation in acquisition, function and usage constitutes the first in-depth corpus-based study on the use of emotive interjections in Present Day British English. In a novel approach, it systematically distinguishes between child and adult speakers, providing new insights into how they use Ow!, Ouch!, Ugh!, Yuck!, Whoops!, Whoopsadaisy! and Wow! in everyday spoken language. It studies in detail their acquisition by children and pinpoints changes and developments in their use throughout early childhood. The study highlights particularities displayed by child and adult speakers in general and identifies crucial differences regarding how adults use emotive interjections depending on whether they are interacting with children or other adults. This book thus offers an exhaustive overview on the functions of emotive interjections based on thorough empirical research and will appeal to linguists concerned with pragmatics, child language acquisition, the expression of emotion and interjections.
-
-
-
English Discourse Particles
Author(s): Karin AijmerPublication Date September 2002show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:There are few aspects of language which are more problematic than its discourse particles. The present study of discourse particles draws upon data from the London-Lund Corpus to show how the methods and tools of corpora can sharpen their description. The first part of the book provides a picture of the state of the art in discourse particle studies and introduces the theory and methodology for the analysis in the second part of the book. Discourse particles are analysed as elements which have been grammaticalised and as a result have certain properties and uses. The importance of linguistic and contextual cues such as text type, position in the discourse, prosody and collocation for analysing discourse particles is illustrated.
The following chapters deal with specific discourse particles (now, oh, just, sort of, and that sort of thing, actually) on the basis of their empirical analysis in the London-Lund Corpus. Examples and extended extracts from many different text types are provided to illustrate what discourse particles are doing in discourse.
-
-
-
English General Nouns
Author(s): Michaela MahlbergPublication Date December 2005show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book proposes an innovative approach to general nouns. General nouns are defined as high-frequency nouns that are characterised by their textual functions. Although the concept is motivated by Halliday & Hasan (1976), the corpus theoretical approach adopted in the present study is fundamentally different and set in a linguistic framework that prioritises lexis. The study investigates 20 nouns that are very frequent in mainstream English, as represented by the Bank of English Corpus. The corpus-driven approach to the data involves a critical discussion of descriptive tools, such as patterns, semantic prosodies, and primings of lexical items, and the concept of 'local textual functions' is put forward to characterise the functions of the nouns in texts. The study not only suggests a characterisation of general nouns, but also stresses that functions of lexical items and properties of texts are closely linked. This link requires new ways of describing language.
-
-
-
English Prosody in First and Second Language Speakers
Author(s): Karin McClellanPublication Date November 2024show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Discover the intricate dynamics of L2 prosody with this pioneering study, which examines how advanced learners from Czech, German, and Spanish backgrounds engage with British and American English intonation. By employing a multidimensional approach - spanning phonetic, phonological, discourse-pragmatic, and sociolinguistic perspectives - this book provides a comprehensive overview of L2 prosodic features, highlighting patterns of intonational phrasing, f0 range, and the use of tones and uptalk. Building on foundational works by Pierrehumbert, Mennen, and Gut, this work bridges significant gaps in the field by comparing different L1 and L2 varieties, integrating diverse linguistic variables, and proposing a multifactorial model of L2 prosody. Relevant for linguists, language educators, and researchers in SLA, the findings offer valuable insights for reducing foreign accents and enhancing intelligibility, making it an essential resource for improving language teaching methodologies and learner outcomes. Dive into this essential guide and elevate your understanding of L2 prosody and its impact on effective communication.
-
-
-
Exploring Corpora for ESP Learning
Author(s): Laura GavioliPublication Date December 2005show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book investigates the effects of corpus work on the process of foreign language learning in ESP settings. It suggests that observing learners at work with corpus data can stimulate discussion and re-thinking of the pedagogical implications of both the theoretical and empirical aspects of corpus linguistics. The ideas presented here are developed from the Data-Driven Learning approach introduced by Tim Johns in the early nineties. The experience of watching students perform corpus analysis provides the basis for the two main observations in the book: a) corpus work provides students with a useful source of information about ESP language features, b) the process of "search-and-discovery" implied in the method of corpus analysis may facilitate language learning and promote autonomy in learning language use. The discussion is carried out on the basis of a series of corpus-based "explorations" by students and provides suggestions for developing new tasks and tools for language learners.
-
-
-
Exploring Language and Society with Big Data
Editor(s): Minna Korhonen, Haidee Kotze and Jukka TyrkköPublication Date November 2023show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:As the legislative bodies of democratic nations, parliaments play a fundamental role in society. Consequently the linguistic practices observed in parliamentary discourse are of importance to everyone. This volume brings together leading researchers in areas of corpus linguistics, big data, parliamentary discourse, and historical linguistics in a truly interdisciplinary exploration at the vanguard of big data and corpus methods with the aim to investigate the intersection between linguistic and social change. Making use of both quantitative and qualitative methods, the studies included in this volume range from a focus on explicitly linguistic phenomena to topics that contribute to our understanding of language and society more generally. It breaks new ground in its critical reflection on the conceptual and methodological challenges of using large corpora of parliamentary discourse to study both the specialised language of parliamentary speech and the societies that the parliaments in question represent and govern.
-
-
-
Exploring Newspaper Language
Editor(s): Gisle AndersenPublication Date March 2012show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book describes new methodological and technological approaches to corpus building and presents recent research based on the Norwegian Newspaper Corpus. This is a large monitor corpus of contemporary Norwegian language, compiled through daily harvesting of web newspapers. The book gives an overview of the corpus and its system architecture, and presents tools used for tasks such as text harvesting, annotation, topic classification and extraction and frequency profiling of new words and phrases. Among the innovative technologies is Corpuscle, a corpus query engine and management system which is flexible enough to handle very large corpora in an efficient way. The individual research contributions based on the corpus explore different aspects of Norwegian, including the occurrence of anglicisms, neologisms and terminology, and the use of metonymy and metaphor in newspaper language. The book also describes an innovative method of applying correspondence analysis and implicational analysis to investigate interdependencies between morphosyntactic variants.
-
-
-
Exploring Second-Language Varieties of English and Learner Englishes
Editor(s): Joybrato Mukherjee and Marianne HundtPublication Date May 2011show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:The articles in this volume are intended to bridge what Sridhar and Sridhar (1986) have called the 'paradigm gap' between traditional SLA research on the one hand and research into institutionalised second-language varieties in former colonial territories on the other. Since both learner Englishes and second-language varieties are typically non-native forms of English that emerge in language contact situations, it is high time that they are described and compared on an empirical basis in order to draw conceptual and theoretical conclusions with regard to their form, function and acquisition. The present collection of articles places special emphasis on empirical evidence obtained from large-scale analyses of computerised corpora of learner Englishes (such as the International Corpus of Learner English) and of second-language varieties of English (such as the International Corpus of English). It addresses questions such as ‘Are the phenomena we find in ESL and EFL varieties features or errors?’ or ‘How common and wide-spread are features across contact varieties of English?’
-
-
-
Exploring the Lexis–Grammar Interface
Editor(s): Ute Römer and Rainer SchulzePublication Date March 2009show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This volume showcases studies that recognize and provide evidence for the inseparability of lexis and grammar. The contributors explore in what ways these two areas, often treated separately in linguistic theory and description, form an organic whole. The papers in Section I (Setting the Scene) introduce some of the key methodological approaches and theoretical positions at the lexis-grammar interface, while Section II (Considering the Particulars) contains papers that report on case studies and show concrete applications of the central methods and theories. Exploring the Lexis-Grammar Interface is a stimulating collection of papers for anyone who wishes to learn more about and get fresh state-of-the-art perspectives on language patterning.
-
-
-
Fluency in Native and Nonnative English Speech
Author(s): Sandra GötzPublication Date March 2013show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book takes a new and holistic approach to fluency in English speech and differentiates between productive, perceptive, and nonverbal fluency. The in-depth corpus-based description of productive fluency points out major differences of how fluency is established in native and nonnative speech. It also reveals areas in which even highly advanced learners of English still deviate strongly from the native target norm and in which they have already approximated to it. Based on these findings, selected learners are subjected to native speakers' ratings of seven perceptive fluency variables in order to test which variables are most responsible for a perception of oral proficiency on the sides of the listeners. Finally, language-pedagogical implications derived from these findings for the improvement of fluency in learner language are presented. This book is conceptually and methodologically relevant for corpus-linguistics, learner corpus research and foreign language teaching and learning.
-
-
-
Frequency, Dispersion, Association, and Keyness
Author(s): Stefan Th. GriesPublication Date July 2024show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book is an attempt to revisit the main specifically corpus-linguistic statistics/measures the field has been relying on for decades: frequency, dispersion, association, and keyness. The book first discusses the purpose of these measures and how they have been measured. Then, the book makes three main proposals: First, that many measures of dispersion, association, and keyness are too confounded with frequency and how to 'take frequency out of them' to obtain conceptually cleaner and more interpretable measures. Second, that many existing measures can be replaced by the simple information-theoretic measure of the Kullback-Leibler divergence and that it, too, can have frequency 'removed' from it. Third, that corpus linguistics should abandon the tradition of trying to describe its findings with a single number and adopt a tupleization approach instead, where we use several separate dimensions of information for description and interpretation. The book is written in an informal, hands-on style and comes with its own R package featuring functions, example data, and several thousand lines of code exemplifying all applications.
-
-
-
Grammar through the Lens of Corpus Linguistics
Editor(s): Javier Pérez-Guerra, Yolanda Fernández-Pena and Ana Elina Martínez-InsuaPublication Date June 2026show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book brings together cutting-edge research on grammatical variation and change in English, showcasing the state of the art in contemporary corpus linguistics. The studies apply corpus-based and variationist methods to a wide range of grammatical categories (nominal, adjectival, verbal, adverbial, prepositional) across synchronic and diachronic perspectives. Its central aim is to demonstrate how empirical corpus evidence can enhance grammatical theory, especially within usage-based, cognitive, constructionalist frameworks. Of special interest is the volume’s diversity of data sources, innovative methodologies and, in general, contemporary treatments of ongoing variation and change, with an emphasis on World Englishes. Continuing the tradition of the Studies in Corpus Linguistics series, this book advances beyond existing research by providing new case studies and insights addressed to researchers, higher-education instructors and advanced students of English linguistics with interests in corpus linguistics, grammar and variation.
-
-
-
Grammatical Change in English World-Wide
Editor(s): Peter CollinsPublication Date February 2015show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:The contributions to this volume apply and extend the techniques of corpus linguistics and diachronic linguistics to the challenge of describing and explaining grammatical change in varieties of English world-wide. The book is divided into two parts, with ten chapters on ‘Inner Circle’ varieties such as Australian, Canadian, and Irish English, and eight on ‘Outer Circle’ varieties such as Philippine, Indian, and Nigerian English. Contributors examine a range of topics including the progressive aspect, modal auxiliaries, do-support, verb morphology, and quotatives, using a wide variety of corpus resources. Overarching research questions addressed include the following: Do diachronic tendencies observed in a particular variety converge with, diverge from, or run in parallel with, those in the parent variety? What are the possible causes of changes observed (e.g. English teaching traditions, Americanisation, internal changes in registers)? This book will appeal to linguists, particularly those interested in grammatical description, corpus linguistics and World Englishes.
-
-
-
How to Do Corpus Pragmatics on Pragmatically Annotated Data
Author(s): Martin WeisserPublication Date April 2018show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book introduces a methodology and research tool (DART) that make it possible to carry out advanced corpus pragmatics research using dialogue corpora enriched with pragmatics-relevant annotations. It first explores the general use of spoken corpora for pragmatics research, as well as issues revolving around their representation and annotation, and then goes on to describe the resources required for such an annotation process. Based on data from three different corpora, ranging from highly constrained, task-oriented, ones (SPAADIA Trainline & Trains 93) to unconstrained dialogues (Switchboard), it next presents an in-depth discussion and illustration of the potential contributions of syntax, semantics, and semantico-pragmatics towards pragmatic force. This is followed by a description of the largely automatic annotation process itself, and finally an analysis of how a set of more than 110 potential speech acts defined in DART contributes towards establishing the specific communicative characteristics of the three corpora.
-
-
-
How to Use Corpora in Language Teaching
Editor(s): John McH. SinclairPublication Date April 2004show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:After decades of being overlooked, corpus evidence is becoming an important component of the teaching and learning of languages. Above all, the profession needs guidance in the practicalities of using corpora, interpreting the results and applying them to the problems and opportunities of the classroom. This book is intensely practical, written mainly by a new generation of language teachers who are acknowledged experts in central aspects of the discipline. It offers advice on what to do in the classroom, how to cope with teachers' queries about language, what corpora to use including learner corpora and spoken corpora and how to handle the variability of language; it reports on some current research and explains how the access software is constructed, including an opportunity for the practitioner to write small but useful programs; and it takes a look into the future of corpora in language teaching.
-
-
-
The Idiom Principle and L1 Influence
Author(s): Ying WangPublication Date October 2016show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book examines delexical verb + noun collocations such as make a decision, give rise to and take care of in Swedish and Chinese learner English. Using a methodological framework that combines learner corpus research with a contrastive perspective, the study is one of the very few in the field to incorporate corpora of the learner’s L1 to investigate the effects of L1 influence. The book provides a highly detailed and multi-faceted analysis of delexical verb + noun collocations in terms of frequency of occurrence, lexical preferences and morphosyntactic patterns. Quantitative and qualitative results on overuse, underuse and errors are presented with linguistically and pedagogically relevant interpretations that include cultural and discourse aspects. More importantly, the book throws light on how L2 learners may alternate between the open-choice principle and the idiom principle as well as the extent and nature of L1 influence on their collocational use.
-
-
-
In Search of Basic Units of Spoken Language
Editor(s): Shlomo Izre'el, Heliana Mello, Alessandro Panunzi and Tommaso RasoPublication Date June 2020show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:What is the best way to analyze spontaneous spoken language? In their search for the basic units of spoken language the authors of this volume opt for a corpus-driven approach. They share a strong conviction that prosodic structure is essential for the study of spoken discourse and each bring their own theoretical and practical experience to the table. In the first part of the book they segment spoken material from a range of different languages (Russian, Hebrew, Central Pomo (an indigenous language from California), French, Japanese, Italian, and Brazilian Portuguese). In the second part of the book each author analyzes the same two spoken English samples, but looking at them from different perspectives, using different methods of analysis as reflected in their respective analyses in Part I. This approach allows for common tendencies of segmentation to emerge, both prosodic and segmental.
-
-
-
Investigating Wikipedia
Editor(s): Céline Poudat, Harald Lüngen and Laura HerzbergPublication Date November 2024show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:The present volume is intended as a reference book on Wikipedia corpus studies, from corpus construction to exploration and analysis. Wikipedia is a complex object, difficult to manipulate for linguists and corpus researchers. In addition to the encyclopedic articles consulted by millions of users, it contains vast spaces of written discussions, aka talk pages, where Wikipedia authors negotiate the collaborative editing of articles, make evaluations, or discuss related topics. The proposed volume covers Wikipedia articles, their revision histories, and discussions, with a focus on discussions, which have not been studied extensively so far and have also been neglected in previous corpus building efforts. Wikipedia discussions are instances of computer-mediated communication (CMC), thus constituting a completely different, interaction-oriented linguistic genre. Sophisticated tools and methods of linguistic annotation and corpus exploration are needed to exploit the huge and valuable corpus resources that can be constructed from the Wikipedia discussions. The present volume aims at encouraging and facilitating Wikipedia corpus studies, providing standards, recommendations, and innovative methods to build and explore Wikipedia corpora, and presenting corpus studies that make the most of the peculiarities of Wikipedia.
-
-
-
The (Ir)reversibility of English Binomials
Author(s): Sandra MollinPublication Date September 2014show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book focuses on binomials (word pairs such as heart and soul, rich and poor, or if and when), and in particular on the degree of reversibility that English binomials demonstrate. Detailed and innovative corpus linguistic analyses investigate the correlates of the degree of reversibility, linguistic constraints that influence the ordering and reversibility of binomials and the diachronic development of reversibility. In addition, judgment data are analyzed for their convergence and divergence with corpus data regarding degrees of reversibility. The book thus establishes reversibility as a complex characteristic of the binomial construction, at the same time throwing light on general questions in phraseology, lexicalization, language structure and language processing.
-
-
-
Keyness in Texts
Editor(s): Marina Bondi and Mike ScottPublication Date November 2010show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This is corpus linguistics with a text linguistic focus. The volume concerns lexical inequality, the fact that some words and phrases share the quality of being key – and thereby reflect or promote important themes – in some textual contexts, while others do not. The patterning of words which differ in their centrality to text meaning is of increasing interest to corpus linguistics. At the same time software resources are yielding increasingly more detailed ways of identifying and studying the linkages between key words and phrases in text databases. This volume brings together work from some of the leading researchers in this field. It presents thirteen studies organized in three sections, the first containing a series of studies exploring the nature of keyness itself, then a set of five studies looking at keyness in specific discourse contexts, and then three studies with an educational focus.
-
-
-
Korean English
Author(s): Glenn HadikinPublication Date July 2014show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:The English language is changing every day and it is us – the individual speakers and writers – that drive those changes in small ways by choosing to use certain strings of words over others. This book discusses and describes some of the choices made by speakers from South Korea by examining the similarities and differences between two Korean communities: one in England and one in South Korea. The book has two overall aims. Firstly, it is intended to begin a discussion about phraseology and Lexical Priming and how these theoretical concepts relate and play out in the context of a New English. Secondly, it provides a model of how a language variety can be explored by detailed analysis of short strings. It delves into a range of areas from World Englishes to phraseology and formulaic language and would be suitable for students, teachers and researchers in all these areas.
-
-
-
Language Acquisition in CLIL and Non-CLIL Settings
Author(s): Verena MöllerPublication Date December 2017show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Language Acquisition in CLIL and Non-CLIL Settings builds a bridge between Second Language Acquisition and Learner Corpus Research (LCR) methodologies to take the evaluation of Content and Language Integrated Learning (CLIL) to a new level. The study innovates in two main ways. First, it is based on a highly diversified L2 database which includes learner corpus data as well as experimental data from the same learners. These linguistic components of the database are complemented with extensive information on learner variables, including cognitive and affective factors, which are rarely studied in LCR. Second, the study relies on multifactorial statistical analyses to assess the effectiveness of CLIL itself as well as the impact of the selectivity inherent in the CLIL system, which has frequently been ignored. The linguistic focus of the study is the English passive, which is investigated in CLIL and non-CLIL teaching materials, and subsequently related to learner output.
-
-
-
Language and Characterisation in Television Series
Author(s): Monika BednarekPublication Date March 2023show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book explores how language is used to create characters in fictional television series. To do so, it draws on multiple case studies from the United States and Australia. Brought together in this book for the first time, these case studies constitute more than the sum of their parts. They highlight different aspects of televisual characterisation and showcase the use of different data, methods, and approaches in its analysis. Uniquely, the book takes a mixed-method approach and will thus not only appeal to corpus linguists but also researchers in sociolinguistics, stylistics, and pragmatics. All corpus linguistic techniques are clearly introduced and explained, and the book is thus accessible to both experienced researchers as well as novice researchers and students. It will be essential reading in linguistics, literature, stylistics, and media/television studies.
-
-
-
The Language of Outsourced Call Centers
Author(s): Eric FriginalPublication Date February 2009show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:The Language of Outsourced Call Centers is the first book to explore a large-scale corpus representing the typical kinds of interactions and communicative tasks in outsourced call centers located in the Philippines and serving American customers. The specific goals of this book are to conduct a corpus-based register comparison between outsourced call center interactions, face-to-face American conversations, and spontaneous telephone exchanges; and to study the dynamics of cross-cultural communication between Filipino call center agents and American callers, as well as other demographic groups of participants in outsourced call center transactions, e.g., gender of speakers, agents’ experience and performance, and types of transactional tasks. The research design relies on a number of analytical approaches, including corpus linguistics and discourse analysis, and combines quantitative and qualitative examination of linguistic data in the investigation of the frequency distribution and functional characteristics of a range of lexico/syntactic features of outsourced call center discourse.
-
-
-
Language Periphery
Publication Date March 2016show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:A full-length study of monocollocable words, i.e. words whose usage is severely restricted to one or a few combinations only (such as English ado in without much/further ado), that brings together corpus-based data from the four languages along with studies analysing, along both general and language-specific lines, monocollocable words in terms of their frequency, lexical as well as morphosyntactic behaviour, and various facets of their peripheral status. Each of the four langauges covered, namely, English, Italian, German and Czech also offers a short introduction of the respective languages written in English, Italian, German and Czech. A rare contribution to our knowledge of an as yet little studied field, the book will attract the attention of, and stimulate a new interest in, all who are ready to acknowledge that collocation is a core phenomenon of language – lexicologists, lexicographers with a focus on phraseology, language typologists, linguists with a contrastive and historical agenda, and language teachers alike.
-
-
-
Learner Corpora and Language Teaching
Editor(s): Sandra Götz and Joybrato MukherjeePublication Date May 2019show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:While native corpora and corpus linguistic tools and methods have been used and applied for quite some time in the development of learning and teaching materials, learner corpora are only just beginning to impact the field of language teaching, testing and assessment. This volume helps to close this still existing gap and highlights the great potential of learner corpus research for language pedagogy by presenting a selection of 11 original studies on learner corpora, conducted by established experts as well as by excellent young researchers. The papers included in the volume present new corpora and methods; studies on written as well as spoken learner corpora and on using data-driven learning scenarios in the classroom.
All papers include sections on practical and concrete language-pedagogical applications. This volume will be of significant interest to researchers working in corpus linguistics, learner corpus research, second language acquisition and English for Academic and Specific Purposes, as well to language teachers and materials developers.
-
-
-
Learner Corpora in Language Testing and Assessment
Editor(s): Marcus Callies and Sandra GötzPublication Date April 2015show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:The aim of this volume is to highlight the benefits and potential of using learner corpora for the testing and assessment of L2 proficiency in both speaking and writing, reflecting the growing importance of learner corpora in applied linguistics and second language acquisition research. Identifying several desiderata for future research and practice, the volume presents a selection of original studies, covering a variety of different languages. It features studies that present very thoroughly compiled new corpus resources which are tailor-made and ready for analysis in LTA, new tools for the automatic assessment of proficiency levels, and new methods of (self-)assessment with the help of learner corpora. Other studies suggest innovative research methodologies of how proficiency can be operationalized through learner corpus data. The volume is of particular interest to researchers in (applied) corpus linguistics, learner corpus research, language testing and assessment, as well as for materials developers and language teachers.
-
-
-
Learning the Language of Dentistry
Author(s): Peter Crosthwaite and Lisa CheungPublication Date August 2019show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book explores the affordances of disciplinary corpora for the teaching and learning of the language of dentistry, within the field of English for Specific Academic Purposes (ESAP). We extract disciplinary register features and vocabulary from three key genres of the dentistry discipline (published experimental research articles, case reports, and novice/professional research reports within the Dental Public Health domain), before integrating these features into ESAP pedagogy in the form of corpus-based ESAP materials that promote student-led direct engagement with disciplinary corpora – an approach known as 'data-driven learning'. This book is a timely and relevant addition to the field of corpus linguistics and ESAP, and is especially targeted at ESAP professionals who are required to teach disciplinary discourses but who may struggle to know what to teach as non-experts of the target discipline.
-
-
-
Lexical Bundles in Native and Non-native Scientific Writing
Author(s): Danica SalazarPublication Date November 2014show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book presents an investigation of lexical bundles in native and non-native scientific writing in English, whose aim is to produce a frequency-derived, statistically- and qualitatively-refined list of the most pedagogically useful lexical bundles in scientific prose: one that can be sorted and filtered by frequency, key word, structure and function, and includes contextual information such as variations, authentic examples and usage notes. The first part of the volumediscusses the creation of this list based on a multimillion-word corpus of biomedical research writing and reveals the structure and functions of lexical bundles and their role in effective scientific communication. A comparative analysis of a non-native corpus highlights non-native scientists’ difficulties in employing lexical bundles. The second part of the volume explores pedagogical applications and provides a series of teaching activities that illustrate how EAP teachers or materials designers can use the list of lexical bundles in their practice.
-
-
-
Lexical Priming
Editor(s): Michael Pace-Sigge and Katie J. PattersonPublication Date August 2017show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Published in 2005, Michael Hoey’s Lexical Priming – A new theory of words and language introduced a completely new theory of language based on how words are used in the real world. In the ten years that have passed, the theory has since gained traction in the field of corpus-linguistics. This volume brings together some of the most important contributions to the theory, in areas such as language teaching and learning, discourse analysis, stylistics as well as the design of language learning software. Crucially, this book introduces aspects of the language that have so far been given less focus in lexical priming, such as spoken language, figurative language, forced primings, priming as predictor of genre, and historical primings. The volume also focuses on applying the lexical priming theory to languages other than English including Mandarin Chinese and Finnish.
-
-
-
Lexis in Contrast
Editor(s): Bengt Altenberg and Sylviane GrangerPublication Date May 2002show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This volume takes stock of current research in contrastive lexical studies. It reflects the growing interest in corpus-based approaches to the study of lexis, in particular the use of multilingual corpora, shared by researchers working in widely differing fields — contrastive linguistics, lexicology, lexicography, terminology, computational linguistics and machine translation. The articles in the volume, which cover a wide diversity of languages, are divided into four main sections: the exploration of cross-linguistic equivalence, contrastive lexical semantics, corpus-based multilingual lexicography, and translation and parallel concordancing. The volume also contains a lengthy introduction to recent trends in contrastive lexical studies written by the editors of the volume, Bengt Altenberg and Sylviane Granger.
-
-
-
Linear Unit Grammar
Author(s): John McH. Sinclair and Anna MauranenPublication Date November 2006show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:People have a natural propensity to understand language text as a succession of smallish chunks, whether they are reading, writing, speaking or listening. Linguists have found that this propensity can shed light on the nature and structure of language, and there are many studies which attempt to harness the potential of natural chunking.This book explores the role of chunking in the description of discourse, especially spoken discourse. It appears that chunking offers a sound but flexible platform on which can be built a descriptive model which is more open and comprehensive than more familiar approaches to structural description. The model remains linear, in that it avoids hierarchies, and it concentrates on the combinatorial patterns of text.
The linear approach turns out to have many advantages, bringing together under one descriptive method a wide variety of different styles of speech and writing. It is complementary to established grammars, but it raises pertinent questions about many of their assumptions.
-
-
-
Linguistic Variation in Research Articles
Author(s): Bethany GrayPublication Date December 2015show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Linguistic Variation in Research Articles investigates the linguistic characteristics of academic research articles, going beyond a traditional analysis of the generically-defined research article to take into account varied realizations of research articles within and across disciplines. It combines corpus-based analyses of 70+ linguistic features with analyses of the situational, or non-linguistic, characteristics of the Academic Journal Registers Corpus: 270 research articles from 6 diverse disciplines (philosophy, history, political science, applied linguistics, biology, physics) and representing three sub-registers (theoretical, quantitative, and qualitative research). Comprehensive analyses include a lexical/grammatical survey, an exploration of structural complexity, and a Multi-Dimensional analysis, all interpreted relative to the situational analysis of the corpus. The finding that linguistic variation in research articles does not occur along a single parameter like discipline is discussed relative to our understanding of disciplinary practices, the multidimensional nature of variation in research articles, and resulting methodological considerations for corpus studies of disciplinary writing.
-
-
-
Local Grammar Approaches to Speech Act Studies
Author(s): Hang SuPublication Date August 2025show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book brings together corpus linguistics and pragmatics by extending the emerging corpus analytic framework of local grammar to speech act research, aiming to enrich the toolkit of corpus-based speech act studies. It outlines four directions in which local grammar can be useful for investigating speech acts, namely, a local grammar approach to annotating speech acts, developing local grammars of speech acts, identifying speech act constructions via the lens of local grammars, and applying local grammars into contrastive speech act studies. These directions are illustrated with studies on apology in contemporary spoken British English, which shows that local grammar can be an innovative approach to advance speech act studies and that such research has significant implications and applications. The book should be of interest to researchers and students in corpus linguistics, pragmatics, construction grammar, and L2 speech act research and teaching.
-
-
-
Metadiscourse in L1 and L2 English
Author(s): Annelie ÄdelPublication Date September 2006show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:The pervasive phenomenon of metadiscourse – commentary on the ongoing discourse – is beginning to take its rightful place among the major topics of discourse studies. This book makes simultaneous contributions to the theory of metadiscourse, corpus-based methods of studying such phenomena, and our knowledge of metadiscourse use in written English. After comprehensively reviewing previous research, it introduces a more rigorous and empirical approach to metadiscourse studies. Ädel presents a new model of metadiscourse based on Jakobson’s functions of language, and other conceptual tools, including explicit features for defining metadiscourse, a taxonomy of the functions it serves, and maps of the boundaries between it and related phenomena. A large-scale study of writing by L1 and L2 university students is presented, in which the L2 speakers’ overuse of metadiscourse strongly marks them as lacking in communicative competence. This work is of interest both to linguists and to educators concerned with writing in English.
-
-
-
Middle and Modern English Corpus Linguistics
Editor(s): Manfred Markus, Yoko Iyeiri, Reinhard Heuberger and Emil ChamsonPublication Date April 2012show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book brings together a variety of approaches to English corpus linguistics and shows how corpus methodologies can contribute to the linking of diachronic and synchronic studies. The articles in this volume investigate historical changes in the English language as well as specific aspects of Middle and Modern English and, moreover, of English dialects. The contributions also discuss the development of English corpus linguistics generally and its potential in the future. Special focus is given to the continuity between Middle and Modern English – much in line with the linking in previous studies of Middle English and Old English under the generic term “medievalism”. This volume highlights the continual development of English from the medieval to modern period.
-
-
-
Multi-Dimensional Analysis, 25 years on
Editor(s): Tony Berber Sardinha and Marcia Veirano PintoPublication Date July 2014show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Approximately a quarter of a century ago, the Multi-Dimensional (MD) approach—one of the most powerful (and controversial) methods in Corpus Linguistics—saw its first book-length treatment. In its eleven chapters, this volume presents all new contributions covering a wide range of written and spoken registers, such as movies, music, magazine texts, student writing, social media, letters to the editor, and reports, in different languages (English, Spanish, Portuguese) and contexts (engineering, journalism, the classroom, the entertainment industry, the Internet, etc.). The book also includes a personal account of the development of the method by its creator, Doug Biber, an introduction to MD statistics, as well as an application of MD analysis to corpus design. The book should be essential reading to anyone with an interest in how texts, genres, and registers are used in society, what their lexis and grammar look like, and how they are interrelated.
-
-
-
Multilingual Corpus Research
Editor(s): Noelia Ramón and María Pérez BlancoPublication Date February 2026show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Multilingual corpora have been used in cross-linguistic research for 30 years. New technologies have dramatically changed the processes of compilation and exploitation of tailor-made corpora for linguistic research. The studies included in this volume showcase current cross-linguistic research utilising parallel, comparable, and novel types of corpora beyond this traditional two-fold distinction. The first part of the volume draws on specialised comparable corpora of newspaper opinion articles, social media texts, and economic discourse. Parallel corpora are the focus of the second part, and are used to shed light on diverse areas such as translation history, bilingual phraseology extraction, and lexico-grammatical contrastive analysis. Recently, the emergence of Artificial Intelligence (AI) has implied a dramatic shift in corpus-based cross-linguistic research. This book offers valuable insights for scholars in contrastive linguistics and translation studies, delineating potential uses of parallel and comparable corpora in Machine Translation, automated translation quality assessment, post-editing, and other AI-enhanced applications.
-
-
-
Multiple Affordances of Language Corpora for Data-driven Learning
Editor(s): Agnieszka Leńko-Szymańska and Alex BoultonPublication Date May 2015show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:In recent years, corpora have found their way into language instruction, albeit often indirectly, through their role in syllabus and course design and in the production of teaching materials and other resources. An alternative and more innovative use is for teachers and students alike to explore corpus data directly as part of the learning process. This volume addresses this latter application of corpora by providing research insights firmly based in the classroom context and reporting on several state-of-the-art projects around the world where learners have direct access to corpus resources and tools and utilize them to improve their control of the language systems and skills or their professional expertise as translators. Its aim is to present recent advances in data-driven learning, addressing issues involving different types of corpora, for different learner profiles, in different ways for different purposes, and using a variety of different research methodologies and perspectives.
-
-
-
Observing Eurolects
Editor(s): Laura MoriPublication Date December 2018show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Focusing on the multi-faceted topic of Eurolects, this volume brings together knowledge and methodologies from various disciplines, including sociolinguistics, legal linguistics, corpus linguistics, and translation studies. The legislative varieties of eleven EU official and working languages (Dutch, English, Finnish, French, German, Greek, Italian, Latvian, Maltese, Polish, Spanish) are analyzed using corpus methodologies in order to investigate the variational dynamics and translation-induced patterns of the different languages. The underlying assumption is that, within the sociolinguistic continua of the EU languages, it is possible to single out specific legislative varieties (Eurolects) that originate at a supra-national level. This research hypothesis is strongly supported by the empirical findings derived from detailed corpus analyses of each language. This work represents the first systematic and comprehensive linguistic research conducted on a wide range of EU languages using the same protocol and applying corpus methodologies to the extensive Eurolect Observatory Multilingual Corpus.
-
-
-
Parallel Corpora for Contrastive and Translation Studies
Editor(s): Irene Doval and M. Teresa Sánchez NietoPublication Date March 2019show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This volume assesses the state of the art of parallel corpus research as a whole, reporting on advances in both recent developments of parallel corpora – with some particular references to comparable corpora as well– and in ways of exploiting them for a variety of purposes. The first part of the book is devoted to new roles that parallel corpora can and should assume in translation studies and in contrastive linguistics, to the usefulness and usability of parallel corpora, and to advances in parallel corpus alignment, annotation and retrieval. There follows an up-to-date presentation of a number of parallel corpus projects currently being carried out in Europe, some of them multimodal, with certain chapters illustrating case studies developed on the basis of the corpora at hand. In most of these chapters, attention is paid to specific technical issues of corpus building. The third part of the book reflects on specific applications and on the creation of bilingual resources from parallel corpora. This volume will be welcomed by scholars, postgraduate and PhD students in the fields of contrastive linguistics, translation studies, lexicography, language teaching and learning, machine translation, and natural language processing.
-
-
-
Pattern Grammar
Author(s): Susan Hunston and Gill FrancisPublication Date February 2000show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This book describes an approach to lexis and grammar based on the concept of phraseology and of language patterning arising from work on large corpora. The notion of 'pattern' as a systematic way of dealing with the interface between lexis and grammar was used in Collins Cobuild English Dictionary (1995) and in the two books in the Collins Cobuild Grammar Patterns series (1996; 1998). This volume describes the research that led to these publications, and explores the theoretical and practical implications of the research. The first chapter sets the work in the context of work on phraseology. The next two chapters give several examples of patterns and how they are identified. Chapters 4 and 5 discuss and exemplify the association of pattern and meaning. Chapters 6, 7 and 8 relate the concept of pattern to traditional approaches to grammar and to discourse. Chapter 9 summarizes the book and adds to the theoretical discussion, as well as indicating the applications of this approach to language teaching. The volume is intended to contribute to the current debate concerning how corpora challenge existing linguistic theories, and as such will be of interest to researchers in the fields of grammar, lexis, discourse and corpus linguistics. It is written in an accessible style, however, and will be equally suitable for students taking courses in those areas.
-
-
-
Patterns and Meanings
Author(s): Alan PartingtonPublication Date November 1998show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Patterns and Meanings consists of case studies which make use of corpora and concordance technology. Each case study elaborates a problem area, makes reference to both the descriptive and applied literature thus far, and then suggests ways of exploiting corpus data to shed light on the problem. Language phenomena investigated include word sense, phraseology and syntax, metaphor and creative use, text reference, idiom, and translation. Emphasis is given to information that usually cannot be found in dictionaries, grammars, language textbooks or other resources, but which the study of corpus data makes available. This work is particularly important not only for its language description insights, but also for pedagogical application. Further useful suggestions are included on setting up a medium-sized corpus on a personal computer.
-
-
-
Patterns and Meanings in Discourse
Author(s): Alan Partington, Alison Duguid and Charlotte TaylorPublication Date April 2013show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:This work is designed, firstly, to both provoke theoretical discussion and serve as a practical guide for researchers and students in the field of corpus linguistics and, secondly, to offer a wide-ranging introduction to corpus techniques for practitioners of discourse studies. It delves into a wide variety of language topics and areas including metaphor, irony, evaluation, (im)politeness, stylistics, language change and sociopolitical issues. Each chapter begins with an outline of an area, followed by case studies which attempt both to shed light on particular themes in this area and to demonstrate the methodologies which might be fruitfully employed to investigate them. The chapters conclude with suggestions on activities which the readers may wish to undertake themselves. An Appendix contains a list of currently available resources for corpus research which were used or mentioned in the book.
-
-
-
Patterns in Contrast
Author(s): Jarle Ebeling and Signe Oksefjell EbelingPublication Date September 2013show More to view fulltext, buy and share links for: show Less to hide fulltext, buy and share links for:Combining the fields of phraseology and contrastive analysis, this book describes how patterns, defined as recurrent word-combinations with semantic unity, behave cross-linguistically. As the contrastive approach adopted in the book relies on translations and a bidirectional corpus model, the first part offers an in-depth discussion of contrastive linguistics, with special emphasis on using translations as tertium comparationis and a parallel corpus as the main source of material. Central to the contrastive analysis is the use of corpus-linguistic methods in the identification of patterns, while a deeper understanding of the phraseological nature of the patterns is closely related to the concept of extended units of meaning. The second part of the book presents five case studies, using an easy-to-follow step-by-step method to illustrate the phraseological-contrastive approach at work. The studies show that patterns weave an intricate web of meanings across languages and demonstrate the potential of exploring patterns in contrast.
-

















































