Challenges in corpus linguistics rethinking corpus compilation and analysis / edited by Mark Kaunisto, Tampere University, Marco Schilk, University of Hildesheim.
| Other author | Kaunisto, Mark. |
| Other author | Schilk, Marco. |
| Other author | International ICAME Conference 2021 : Universität Dortmund) |
| Format | Electronic |
| Publication Info | Amsterdam ; Philadelphia : John Benjamins Publishing Company, 2024. |
| Description | pages cm. |
| Supplemental Content | Full text available from eBooks on EBSCOhost |
| Subjects |
| Series | Studies in corpus linguistics, 1388-0373 ; volume 118 |
| Contents | From fallacies and pitfalls to solutions and future directions: navigating the evolving terrain of corpus linguistics / Mark Kaunisto -- Engaging with bad (meta)data in historical corpus linguistics / Turo Vartiainen & Tanja Säily -- Named entities as potentially problematic items in corpora / Mark Kaunisto -- Challenges in the compilation, annotation and analysis of learner corpus data / Marcus Callies -- Early newspapers as data for corpus linguistics (and Digital Humanities): issues in using the British Library Newspapers database as a corpus / Turo Hiltunen -- Open Corpus Linguistics - Or how to overcome common problems in dealing with corpus data by adopting open research practices / Stefan Hartmann -- Text length and short texts: An overview of the problem / Aatu Liimatta -- Corpus genre categories: issues at the intersection of linguistics and literature / Daniel Ocic Ihrmark -- Modeling fine-grained sociolinguistic variation: the promises and pitfalls of Twitter corpora and neural word embeddings / Filip Miletic, Anne Przewozny-Desriaux & Ludovic Tanguy. |
| Abstract | "This book contributes to the work on discussing the challenges faced in different areas of corpus linguistics, namely the compilation, annotation, and analysis of linguistic corpora. In a field of growing corpus sizes and expanding possibilities of gathering data, some old issues persist, while at the same time new problems have emerged. As the compilation and study of language corpora gets increasingly sophisticated and complex, continuous attention on the ways of dealing with the data in question and the challenges in text selection and interpretation is needed. The contributions to this volume address problems relating to a variety of areas in corpus linguistic study, including corpus annotation, data variability, learner language, social media texts, and database utilization. The authors provide critical overviews and research-based analyses, discuss the nature of some of the common pitfalls, and offer solutions to existing problems"-- Provided by publisher. |
| General note | Most of the chapters of this volume have their origins in the pre-conference workshop of the 42nd ICAME conference held at TU Dortmund University in August 2021. |
| Bibliography note | Includes bibliographical references and index. |
| Access restriction | Available only to authorized users. |
| Technical details | Mode of access: World Wide Web |
| Genre/form | Electronic books. |
| LCCN | 2024029608 |
| ISBN | 9789027215888 (hardcover) |
| ISBN | (pdf) |
Availability
| Library | Location | Call Number | Status | Item Actions |
|---|---|---|---|---|
| Electronic Resources | Access Content Online | ✔ Available |