Designing and Evaluating Language Corpora

Designing and Evaluating Language Corpora

Author: Jesse Egbert

Publisher: Cambridge University Press

Published: 2022-04-14

Total Pages: 299

ISBN-13: 1107151384

DOWNLOAD EBOOK

This volume introduces a new framework for conceptualizing and achieving corpus representativeness in a rigorous, yet practical way.


Book Synopsis Designing and Evaluating Language Corpora by : Jesse Egbert

Download or read book Designing and Evaluating Language Corpora written by Jesse Egbert and published by Cambridge University Press. This book was released on 2022-04-14 with total page 299 pages. Available in PDF, EPUB and Kindle. Book excerpt: This volume introduces a new framework for conceptualizing and achieving corpus representativeness in a rigorous, yet practical way.


Developing Linguistic Corpora

Developing Linguistic Corpora

Author: Martin Wynne

Publisher: Oxbow Books Limited

Published: 2005

Total Pages: 100

ISBN-13:

DOWNLOAD EBOOK

A linguistic corpus is a collection of texts which have been selected and brought together so that language can be studied on the computer. Today, corpus linguistics offers some of the most powerful new procedures for the analysis of language, and the impact of this dynamic and expanding sub-discipline is making itself felt in many areas of language study. In this volume, a selection of leading experts in various key areas of corpus construction offer advice in a readable and largely non-technical style to help the reader to ensure that their corpus is well designed and fit for the intended purpose. This guide is aimed at those who are at some stage of building a linguistic corpus. Little or no knowledge of corpus linguistics or computational procedures is assumed, although it is hoped that more advanced users will find the guidelines here useful. It is also aimed at those who are not building a corpus, but who need to know something about the issues involved in the design of corpora in order to choose between available resources and to help draw conclusions from their studies.


Book Synopsis Developing Linguistic Corpora by : Martin Wynne

Download or read book Developing Linguistic Corpora written by Martin Wynne and published by Oxbow Books Limited. This book was released on 2005 with total page 100 pages. Available in PDF, EPUB and Kindle. Book excerpt: A linguistic corpus is a collection of texts which have been selected and brought together so that language can be studied on the computer. Today, corpus linguistics offers some of the most powerful new procedures for the analysis of language, and the impact of this dynamic and expanding sub-discipline is making itself felt in many areas of language study. In this volume, a selection of leading experts in various key areas of corpus construction offer advice in a readable and largely non-technical style to help the reader to ensure that their corpus is well designed and fit for the intended purpose. This guide is aimed at those who are at some stage of building a linguistic corpus. Little or no knowledge of corpus linguistics or computational procedures is assumed, although it is hoped that more advanced users will find the guidelines here useful. It is also aimed at those who are not building a corpus, but who need to know something about the issues involved in the design of corpora in order to choose between available resources and to help draw conclusions from their studies.


Analysing Representation

Analysing Representation

Author: Frazer Heritage

Publisher: Taylor & Francis

Published: 2024-05-31

Total Pages: 316

ISBN-13: 104001898X

DOWNLOAD EBOOK

Analysing Representation: A Corpus and Discourse Textbook guides readers through the process of researching how people and phenomena are represented in discourse and introduces them to key tools they can use from corpus linguistics and (critical) discourse analysis. This book takes a step-by-step approach to introducing each concept and includes exercises and further reading to help readers check their progress and prepare for independent research. It is unique in introducing readers to a range of experts representing the full range of work in this area. This book is aimed at final-year undergraduate, taught postgraduate and doctoral level students. It wil also be useful to scholars who are new to combining corpus and discourse methods in investigations of representation.


Book Synopsis Analysing Representation by : Frazer Heritage

Download or read book Analysing Representation written by Frazer Heritage and published by Taylor & Francis. This book was released on 2024-05-31 with total page 316 pages. Available in PDF, EPUB and Kindle. Book excerpt: Analysing Representation: A Corpus and Discourse Textbook guides readers through the process of researching how people and phenomena are represented in discourse and introduces them to key tools they can use from corpus linguistics and (critical) discourse analysis. This book takes a step-by-step approach to introducing each concept and includes exercises and further reading to help readers check their progress and prepare for independent research. It is unique in introducing readers to a range of experts representing the full range of work in this area. This book is aimed at final-year undergraduate, taught postgraduate and doctoral level students. It wil also be useful to scholars who are new to combining corpus and discourse methods in investigations of representation.


Multi-Dimensional Analysis

Multi-Dimensional Analysis

Author: Tony Berber Sardinha

Publisher: Bloomsbury Publishing

Published: 2019-03-21

Total Pages: 304

ISBN-13: 1350023841

DOWNLOAD EBOOK

Multi-Dimensional Analysis: Research Methods and Current Issues provides a comprehensive guide both to the statistical methods in Multi-Dimensional Analysis (MDA) and its key elements, such as corpus building, tagging, and tools. The major goal is to explain the steps involved in the method so that readers may better understand this complex research framework and conduct MD research on their own. Multi-Dimensional Analysis is a method that allows the researcher to describe different registers (textual varieties defined by their social use) such as academic settings, regional discourse, social media, movies, and pop songs. Through multivariate statistical techniques, MDA identifies complementary correlation groupings of dozens of variables, including variables which belong both to the grammatical and semantic domains. Such groupings are then associated with situational variables of texts like information density, orality, and narrativity to determine linguistic constructs known as dimensions of variation, which provide a scale for the comparison of a large number of texts and registers. This book is a comprehensive research guide to MDA.


Book Synopsis Multi-Dimensional Analysis by : Tony Berber Sardinha

Download or read book Multi-Dimensional Analysis written by Tony Berber Sardinha and published by Bloomsbury Publishing. This book was released on 2019-03-21 with total page 304 pages. Available in PDF, EPUB and Kindle. Book excerpt: Multi-Dimensional Analysis: Research Methods and Current Issues provides a comprehensive guide both to the statistical methods in Multi-Dimensional Analysis (MDA) and its key elements, such as corpus building, tagging, and tools. The major goal is to explain the steps involved in the method so that readers may better understand this complex research framework and conduct MD research on their own. Multi-Dimensional Analysis is a method that allows the researcher to describe different registers (textual varieties defined by their social use) such as academic settings, regional discourse, social media, movies, and pop songs. Through multivariate statistical techniques, MDA identifies complementary correlation groupings of dozens of variables, including variables which belong both to the grammatical and semantic domains. Such groupings are then associated with situational variables of texts like information density, orality, and narrativity to determine linguistic constructs known as dimensions of variation, which provide a scale for the comparison of a large number of texts and registers. This book is a comprehensive research guide to MDA.


Corpus Linguistics for Health Communication

Corpus Linguistics for Health Communication

Author: Gavin Brookes

Publisher: Taylor & Francis

Published: 2023-12-22

Total Pages: 261

ISBN-13: 1003819796

DOWNLOAD EBOOK

Corpus Linguistics for Health Communication provides an accessible and practical introduction to the use of corpus linguistics methods to analyse health-related language use across various contexts and genres. Offering a critical review of the field, discussion of extended case studies, and practical exercises based on spoken, written, and digital language data, this book: introduces the fields of health communication and corpus linguistics and critically reviews cutting-edge studies in the burgeoning area of corpus-based health communication; describes the processes involved in planning a corpus linguistics study of health communication, including designing and building a corpus, selecting tools, and implementing techniques of analysis; demonstrates how corpus linguistics methods can – and have – been applied to the study of spoken, written, and digital health communication, offering critical reflections and suggesting areas for future development. Corpus Linguistics for Health Communication is essential reading for those working at the interface of corpus linguistics and health communication. Both those with a little or a lot of experience in either field will find value in its pages.


Book Synopsis Corpus Linguistics for Health Communication by : Gavin Brookes

Download or read book Corpus Linguistics for Health Communication written by Gavin Brookes and published by Taylor & Francis. This book was released on 2023-12-22 with total page 261 pages. Available in PDF, EPUB and Kindle. Book excerpt: Corpus Linguistics for Health Communication provides an accessible and practical introduction to the use of corpus linguistics methods to analyse health-related language use across various contexts and genres. Offering a critical review of the field, discussion of extended case studies, and practical exercises based on spoken, written, and digital language data, this book: introduces the fields of health communication and corpus linguistics and critically reviews cutting-edge studies in the burgeoning area of corpus-based health communication; describes the processes involved in planning a corpus linguistics study of health communication, including designing and building a corpus, selecting tools, and implementing techniques of analysis; demonstrates how corpus linguistics methods can – and have – been applied to the study of spoken, written, and digital health communication, offering critical reflections and suggesting areas for future development. Corpus Linguistics for Health Communication is essential reading for those working at the interface of corpus linguistics and health communication. Both those with a little or a lot of experience in either field will find value in its pages.


Corpus Linguistics

Corpus Linguistics

Author: Douglas Biber

Publisher: Cambridge University Press

Published: 1998-04-23

Total Pages:

ISBN-13: 1316582566

DOWNLOAD EBOOK

This book is about investigating the way people use language in speech and writing. It introduces the corpus-based approach to linguistics, based on analysis of large databases of real language examples stored on computer. Each chapter focuses on a different area of linguistics, including lexicography, grammar, discourse, register variation, language acquisition, and historical linguistics. Example analyses are presented in each chapter to provide concrete descriptions of the research methods and advantages of corpus-based techniques. Ten methodology boxes provide clear and concise explanations of the issues in doing corpus-based research and reading corpus-based studies and there is a useful appendix of resources for corpus-based investigation. This lucid and comprehensive introduction to the subject will be welcomed by a broad range of readers, from undergraduate students to professional researchers.


Book Synopsis Corpus Linguistics by : Douglas Biber

Download or read book Corpus Linguistics written by Douglas Biber and published by Cambridge University Press. This book was released on 1998-04-23 with total page pages. Available in PDF, EPUB and Kindle. Book excerpt: This book is about investigating the way people use language in speech and writing. It introduces the corpus-based approach to linguistics, based on analysis of large databases of real language examples stored on computer. Each chapter focuses on a different area of linguistics, including lexicography, grammar, discourse, register variation, language acquisition, and historical linguistics. Example analyses are presented in each chapter to provide concrete descriptions of the research methods and advantages of corpus-based techniques. Ten methodology boxes provide clear and concise explanations of the issues in doing corpus-based research and reading corpus-based studies and there is a useful appendix of resources for corpus-based investigation. This lucid and comprehensive introduction to the subject will be welcomed by a broad range of readers, from undergraduate students to professional researchers.


Learner Corpora in Language Testing and Assessment

Learner Corpora in Language Testing and Assessment

Author: Marcus Callies

Publisher: John Benjamins Publishing Company

Published: 2015-04-15

Total Pages: 228

ISBN-13: 9027268703

DOWNLOAD EBOOK

The aim of this volume is to highlight the benefits and potential of using learner corpora for the testing and assessment of L2 proficiency in both speaking and writing, reflecting the growing importance of learner corpora in applied linguistics and second language acquisition research. Identifying several desiderata for future research and practice, the volume presents a selection of original studies, covering a variety of different languages. It features studies that present very thoroughly compiled new corpus resources which are tailor-made and ready for analysis in LTA, new tools for the automatic assessment of proficiency levels, and new methods of (self-)assessment with the help of learner corpora. Other studies suggest innovative research methodologies of how proficiency can be operationalized through learner corpus data. The volume is of particular interest to researchers in (applied) corpus linguistics, learner corpus research, language testing and assessment, as well as for materials developers and language teachers.


Book Synopsis Learner Corpora in Language Testing and Assessment by : Marcus Callies

Download or read book Learner Corpora in Language Testing and Assessment written by Marcus Callies and published by John Benjamins Publishing Company. This book was released on 2015-04-15 with total page 228 pages. Available in PDF, EPUB and Kindle. Book excerpt: The aim of this volume is to highlight the benefits and potential of using learner corpora for the testing and assessment of L2 proficiency in both speaking and writing, reflecting the growing importance of learner corpora in applied linguistics and second language acquisition research. Identifying several desiderata for future research and practice, the volume presents a selection of original studies, covering a variety of different languages. It features studies that present very thoroughly compiled new corpus resources which are tailor-made and ready for analysis in LTA, new tools for the automatic assessment of proficiency levels, and new methods of (self-)assessment with the help of learner corpora. Other studies suggest innovative research methodologies of how proficiency can be operationalized through learner corpus data. The volume is of particular interest to researchers in (applied) corpus linguistics, learner corpus research, language testing and assessment, as well as for materials developers and language teachers.


Teaching and Language Corpora

Teaching and Language Corpora

Author: Anne Wichmann

Publisher: Routledge

Published: 2014-06-11

Total Pages: 362

ISBN-13: 1317889584

DOWNLOAD EBOOK

Corpora are well-established as a resource for language research; they are now also increasingly being used for teaching purposes. This book is the first of its kind to deal explicitly and in a wide-ranging way with the use of corpora in teaching. It contains an extensive collection of articles by corpus linguists and practising teachers, covering not only the use of data to inform and create teaching materials but also the direct exploitation of corpora by students, both in the study of linguistics in general and in the acquisition of proficiency in individual languages, including English, Welsh, German, French and Italian. In addition, the book offers practical information on the sources of corpora and concordances, including those suitable for work on non-roman scripts such as Greek and Cyrillic.


Book Synopsis Teaching and Language Corpora by : Anne Wichmann

Download or read book Teaching and Language Corpora written by Anne Wichmann and published by Routledge. This book was released on 2014-06-11 with total page 362 pages. Available in PDF, EPUB and Kindle. Book excerpt: Corpora are well-established as a resource for language research; they are now also increasingly being used for teaching purposes. This book is the first of its kind to deal explicitly and in a wide-ranging way with the use of corpora in teaching. It contains an extensive collection of articles by corpus linguists and practising teachers, covering not only the use of data to inform and create teaching materials but also the direct exploitation of corpora by students, both in the study of linguistics in general and in the acquisition of proficiency in individual languages, including English, Welsh, German, French and Italian. In addition, the book offers practical information on the sources of corpora and concordances, including those suitable for work on non-roman scripts such as Greek and Cyrillic.


Multiple Affordances of Language Corpora for Data-driven Learning

Multiple Affordances of Language Corpora for Data-driven Learning

Author: Agnieszka Leńko-Szymańska

Publisher: John Benjamins Publishing Company

Published: 2015-05-15

Total Pages: 321

ISBN-13: 9027268711

DOWNLOAD EBOOK

In recent years, corpora have found their way into language instruction, albeit often indirectly, through their role in syllabus and course design and in the production of teaching materials and other resources. An alternative and more innovative use is for teachers and students alike to explore corpus data directly as part of the learning process. This volume addresses this latter application of corpora by providing research insights firmly based in the classroom context and reporting on several state-of-the-art projects around the world where learners have direct access to corpus resources and tools and utilize them to improve their control of the language systems and skills or their professional expertise as translators. Its aim is to present recent advances in data-driven learning, addressing issues involving different types of corpora, for different learner profiles, in different ways for different purposes, and using a variety of different research methodologies and perspectives.


Book Synopsis Multiple Affordances of Language Corpora for Data-driven Learning by : Agnieszka Leńko-Szymańska

Download or read book Multiple Affordances of Language Corpora for Data-driven Learning written by Agnieszka Leńko-Szymańska and published by John Benjamins Publishing Company. This book was released on 2015-05-15 with total page 321 pages. Available in PDF, EPUB and Kindle. Book excerpt: In recent years, corpora have found their way into language instruction, albeit often indirectly, through their role in syllabus and course design and in the production of teaching materials and other resources. An alternative and more innovative use is for teachers and students alike to explore corpus data directly as part of the learning process. This volume addresses this latter application of corpora by providing research insights firmly based in the classroom context and reporting on several state-of-the-art projects around the world where learners have direct access to corpus resources and tools and utilize them to improve their control of the language systems and skills or their professional expertise as translators. Its aim is to present recent advances in data-driven learning, addressing issues involving different types of corpora, for different learner profiles, in different ways for different purposes, and using a variety of different research methodologies and perspectives.


History, Features, and Typology of Language Corpora

History, Features, and Typology of Language Corpora

Author: Niladri Sekhar Dash

Publisher: Springer

Published: 2018-02-01

Total Pages: 293

ISBN-13: 9811074585

DOWNLOAD EBOOK

This book discusses key issues of corpus linguistics like the definition of the corpus, primary features of a corpus, and utilization and limitations of corpora. It presents a unique classification scheme of language corpora to show how they can be studied from the perspective of genre, nature, text type, purpose, and application. A reference to parallel translation corpus is mandatory in the discussion of corpus generation, which the authors thoroughly address here, with a focus on Indian language corpora and English. Web-text corpus, a new development in corpus linguistics, is also discussed with elaborate reference to Indian web text corpora. The book also presents a short history of corpus generation and provides scenarios before and after the advent of computer-generated digital corpora. This book has several important features: it discusses many technical issues of the field in a lucid manner; contains extensive new diagrams and charts for easy comprehension; and presents discussions in simplified English to cater to the needs of non-native English readers. This is an important resource authored by academics who have many years of experience teaching and researching corpus linguistics. Its focus on Indian languages and on English corpora makes it applicable to students of graduate and postgraduate courses in applied linguistics, computational linguistics and language processing in South Asia and across countries where English is spoken as a first or second language.


Book Synopsis History, Features, and Typology of Language Corpora by : Niladri Sekhar Dash

Download or read book History, Features, and Typology of Language Corpora written by Niladri Sekhar Dash and published by Springer. This book was released on 2018-02-01 with total page 293 pages. Available in PDF, EPUB and Kindle. Book excerpt: This book discusses key issues of corpus linguistics like the definition of the corpus, primary features of a corpus, and utilization and limitations of corpora. It presents a unique classification scheme of language corpora to show how they can be studied from the perspective of genre, nature, text type, purpose, and application. A reference to parallel translation corpus is mandatory in the discussion of corpus generation, which the authors thoroughly address here, with a focus on Indian language corpora and English. Web-text corpus, a new development in corpus linguistics, is also discussed with elaborate reference to Indian web text corpora. The book also presents a short history of corpus generation and provides scenarios before and after the advent of computer-generated digital corpora. This book has several important features: it discusses many technical issues of the field in a lucid manner; contains extensive new diagrams and charts for easy comprehension; and presents discussions in simplified English to cater to the needs of non-native English readers. This is an important resource authored by academics who have many years of experience teaching and researching corpus linguistics. Its focus on Indian languages and on English corpora makes it applicable to students of graduate and postgraduate courses in applied linguistics, computational linguistics and language processing in South Asia and across countries where English is spoken as a first or second language.