<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD with OASIS Tables with MathML3 v1.4 20241031//EN" "https://jats.nlm.nih.gov/archiving/1.4/JATS-archive-oasis-article1-4-mathml3.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:ali="http://www.niso.org/schemas/ali/1.0/" dtd-version="1.4" article-type="review-article" xml:lang="en"><front><journal-meta><journal-title-group><journal-title xml:lang="ru">Вестник Волгоградского государственного университета. Серия 2. Языкознание</journal-title></journal-title-group><issn publication-format="print">1998-9911</issn><issn publication-format="electronic">2409-1979</issn></journal-meta><article-meta><article-id pub-id-type="doi">10.15688/jvolsu2.2025.1.8</article-id><article-categories><subj-group><subject>Other</subject></subj-group></article-categories><title-group><article-title xml:lang="ru">Роль корпусного языкознания в современных лингвистических исследованиях и переводоведении</article-title><trans-title-group xml:lang="en"><trans-title>The Role of Corpus Linguistics in Contemporary Linguistics Research and Translation Studies</trans-title></trans-title-group></title-group><contrib-group><contrib contrib-type="author"><name-alternatives><name xml:lang="ru"><surname>Пэй</surname><given-names>Хайтун</given-names></name><name xml:lang="en"><surname>Pei</surname><given-names>Haitong</given-names></name></name-alternatives><xref ref-type="aff" rid="aff1"/><email>peihaitong@mail.ru</email><contrib-id contrib-id-type="orcid">0000-0002-5031-6586</contrib-id></contrib><aff-alternatives id="aff1"><aff xml:lang="en"><institution>Heilongjiang University (Harbin, China)</institution></aff><aff xml:lang="ru"><institution>Хэйлунцзянский университет (Харбин, Китай)</institution></aff></aff-alternatives></contrib-group><pub-date pub-type="epub" iso-8601-date="2025-04-24"><day>24</day><month>04</month><year>2025</year></pub-date><volume>24</volume><issue>1</issue><fpage>95</fpage><lpage>106</lpage><history><date date-type="received" iso-8601-date="2024-04-28"><day>28</day><month>04</month><year>2024</year></date><date date-type="accepted" iso-8601-date="2024-10-21"><day>21</day><month>10</month><year>2024</year></date></history><permissions><license xlink:href="https://creativecommons.org/licenses/by/4.0/" xlink:title="CC BY 4.0"><ali:license_ref>https://creativecommons.org/licenses/by/4.0/</ali:license_ref><license-p xml:lang="ru">CC BY 4.0</license-p></license></permissions><self-uri xlink:href="https://l.jvolsu.com/index.php/ru/archive-ru/948-science-journal-of-volsu-linguistics-2025-vol-24-no-1/materialy-i-soobshcheniya/2895-pej-khajtun-rol-korpusnogo-yazykoznaniya-v-sovremennykh-lingvisticheskikh-issledovaniyakh-i-perevodovedenii-na-angl-yaz" xlink:title="https://l.jvolsu.com/index.php/ru/archive-ru/948-science-journal-of-volsu-linguistics-2025-vol-24-no-1/materialy-i-soobshcheniya/2895-pej-khajtun-rol-korpusnogo-yazykoznaniya-v-sovremennykh-lingvisticheskikh-issledovaniyakh-i-perevodovedenii-na-angl-yaz">https://l.jvolsu.com/index.php/ru/archive-ru/948-science-journal-of-volsu-linguistics-2025-vol-24-no-1/materialy-i-soobshcheniya/2895-pej-khajtun-rol-korpusnogo-yazykoznaniya-v-sovremennykh-lingvisticheskikh-issledovaniyakh-i-perevodovedenii-na-angl-yaz</self-uri><abstract xml:lang="ru"><p>Статья представляет собой систематизированный обзор работ по корпусной лингвистике как инновационному направлению эмпирического языкознания. Раскрыты теоретико-методологические основы направления, определена специфика корпусной лингвистики по сравнению с компьютерной, обозначены современные направления корпусных исследований и преимущества привлечения данных текстового корпуса в языковедческих работах. Отмечено, что корпусный метод и его ресурсы сегодня активно используют в различных лингвистических исследованиях многих мировых языков; что языковые корпусы различают по объему, типу, структуре, наполнению, назначению и др. В статье обзорно рассматриваются такие корпусы, как Британский национальный корпус (The British National Corpus), Американский национальный корпус (The American National Corpus), Корпус современного американского английского языка (COCA), Национальный корпус русского языка, Корпус современного китайского языка (The Modern Chinese Language Corpus), созданный в Центре китайской лингвистики при Пекинском университете, Сбалансированный корпус китайского языка. Выявлена специфика параллельных и сравнительных корпусов, среди которых CHILDES – корпус детской речи, Международный сравнительный корпус (The International Comparable Corpus), Корпус англоязычной Википедии (Corpus of English Wikipedia), приводятся различия между параллельными и сравнительными корпусами. Установлено, что перспективы развития национальных корпусов связаны с новыми исследованиями практически в каждой области прикладной и теоретической лингвистики, с дальнейшей разработкой и углублением теории и практики перевода. Выделены особенности одноязычных, сравнительных и параллельных корпусов в контексте их роли в лингвистических исследованиях. Показано, что в инструментарий переводчика, помимо параллельных корпусов, входят и одноязычные, дающие дополнительный материал о предмете перевода и актуализирующие фоновые знания переводчика.</p></abstract><abstract xml:lang="en" abstract-type="summary"><p>The article presents a systematic review of research papers on corpus linguistics as an innovative direction in empirical linguistics. It reveals the theoretical and methodological foundations of the field, defines the specifics of corpus linguistics in comparison with computational linguistics, and highlights modern trends in corpus research as well as the advantages of employing textual corpus data in linguistic studies. It is noted that the corpus method and its resources are actively used in various linguistic researches into many world languages today; that language corpora are differentiated by volume, type, structure, content, purpose, etc. The article provides an overview of such corpora as the British National Corpus (BNC), the American National Corpus (ANC), the Corpus of Contemporary American English (COCA), the Russian National Corpus (RNC), and the Modern Chinese Language Corpus created in the Center for Chinese Linguistics at Beijing University, the Balanced Corpus of the Chinese Language. The specifics of parallel and comparative corpora are noted, including the Child Language Data Exchange System (CHILDES), the International Comparable Corpus (ICC), and the Corpus of English Wikipedia (CEW). The differences between parallel and comparative corpora are also outlined. Prospects for national corpora development lie in new research into almost every area of applied and theoretical linguistics, as well as in scrutiny and further development of translation theory and practice. The characteristics of monolingual, comparative, and parallel corpora are highlighted in the context of their role in linguistic research. It is mentioned that, in addition to parallel corpora, translators' tools also include monolingual corpora, which provide additional material on the subject of translation and enhance the translator's background knowledge.</p></abstract><kwd-group xml:lang="ru"><kwd>сравнительный корпус</kwd><kwd>корпусный метод</kwd><kwd>прикладная лингвистика</kwd><kwd>корпусная лингвистика</kwd><kwd>корпус</kwd><kwd>национальный корпус</kwd><kwd>параллельный корпус</kwd></kwd-group><kwd-group xml:lang="en"><kwd>applied linguistics</kwd><kwd>corpus linguistics</kwd><kwd>corpus</kwd><kwd>national corpus</kwd><kwd>parallel corpus</kwd><kwd>comparative corpus</kwd><kwd>corpus method</kwd></kwd-group><funding-group><funding-statement xml:lang="en">The article is written as part of a grant in the field of philosophy and social sciences from Heilongjiang Province for the year 2024, project number 24YYC006, titled “Comparative Study of the Worldviews of the Chinese and Russian Peoples within the Framework of Cognitive Linguistics.”</funding-statement></funding-group></article-meta></front><back><ref-list><ref id="ref1"><mixed-citation publication-type="other" xml:lang="ru">Baker P., 2006. Using Corpora in Discourse Analysis. London, Continuum, 2006. 208 p.</mixed-citation></ref><ref id="ref2"><mixed-citation publication-type="other" xml:lang="ru">Baranov A.N., 2001. Vvedenie v prikladnuyu lingvistiku [Introduction to Applied Linguistics]. Moscow, Editorial URSS Publ., 2001. 360 p.</mixed-citation></ref><ref id="ref3"><mixed-citation publication-type="other" xml:lang="ru">Biber D., 2001. Using Corpus-Based Methods to Investigate Grammar and Use: Some Case Studies on the Use of Verbs in English. Corpus Linguistics in North America, 2001, pp. 101-115.</mixed-citation></ref><ref id="ref4"><mixed-citation publication-type="other" xml:lang="ru">Chesnokova I.D., Manshin M.E., 2018. Natsionalnyy korpus russkogo yazyka kak osnovnoy instrument poiska pri lingvisticheskikh issledovaniyakh (na primere poiska antonimov v publicisticheskikh tekstakh) [National Corpus of the Russian Language as the Main Search Tool in Linguistic Research (Based on the Search for Antonyms in Journalistic Texts)]. Izvestiya VGPU [Ivzestia of the Volgograd State Pedagogical University], no. 5 (128). URL: https://cyberleninka.ru/article/n/natsionalnyy-korpus-russkogo-yazyka-kak-osnovnoy-instrument-poiska-pri-lingvisticheskih-issledovaniyah-na-primere-poiska-antonimov-v</mixed-citation></ref><ref id="ref5"><mixed-citation publication-type="other" xml:lang="ru">Finegan E., 2015. Language: Its Structure and Use. Stamford, CT, Cengage Learning. 575 p.</mixed-citation></ref><ref id="ref6"><mixed-citation publication-type="other" xml:lang="ru">Kolpachkova E.N., 2019. Korpusy kitayskogo yazyka: sovremennoe sostoyanie i osnovnye problemy [Chinese Language Corpora: An Overview and Major Problems]. URL: https://orient.spbu.ru/images/document/2019/Kolpachkova_ Chinese_Corpus_overview.pdf</mixed-citation></ref><ref id="ref7"><mixed-citation publication-type="other" xml:lang="ru">Kornienko A.V., 2023. Natsionalnyy korpus russkogo yazyka kak istochnikovaya baza sotsio-gumanitarnykh issledovaniy [National Corpus of the Russian Language as a Source Base for Social and Humanitarian Research]. Peterburgskaya sotsiologiya segodnya [Petersburg Sociology Today], no. 21, pp. 58-71. DOI: 10.25990/socinstras.pss-21.8sqp-bb68</mixed-citation></ref><ref id="ref8"><mixed-citation publication-type="other" xml:lang="ru">Kotov A.A., Mineeva Z.I., 2013. Ispolzovanie natsionalnogo korpusa russkogo yazyka pri provedenii lingvisticheskoy ekspertizy [Use of Russian National Corpus in Linguistic Examination]. Filologicheskie nauki. Voprosy teorii i praktiki [Philology. Theory &amp; Practice], no. 8 (26): in 2 parts. Pt. 2, pp. 99-103.</mixed-citation></ref><ref id="ref9"><mixed-citation publication-type="other" xml:lang="ru">McEnery T., Hardie A., 2011. Corpus Linguistics: Method, Theory and Practice. Cambridge, Cambridge University Press. 312 р.</mixed-citation></ref><ref id="ref10"><mixed-citation publication-type="other" xml:lang="ru">McEnery T., Hardie A., 2012. Corpus Linguistics: Method, Theory and Practice. Cambridge, Cambridge University Press. 294 р.</mixed-citation></ref><ref id="ref11"><mixed-citation publication-type="other" xml:lang="ru">Mustajoki A., 2007. Rol korpusov v lingvisticheskikh issledovaniyakh yazykov [Role of Corpora in Linguistic Research and Language Teaching]. Natsionalnyy korpus russkogo yazyka i problemy gumanitarnogo obrazovaniya: materialy Mezhdunar. nauch. konf. [National Corpus of the Russian Language and the Problems of Humanitarian Education. Proceedings of the International Scientific Conference]. Moscow, Nats. issled. un-t «Vyssh. shk. ekonomiki», pp. 152-166.</mixed-citation></ref><ref id="ref12"><mixed-citation publication-type="other" xml:lang="ru">Piotrowski T., 2008. The Translator and Polish-English Corpora. Incorporating Corpora. The Linguist and the Translator. Clevedon, Multilingual Matters Ltd., pp. 117-132.</mixed-citation></ref><ref id="ref13"><mixed-citation publication-type="other" xml:lang="ru">Semino E., Short M., 2004. Corpus Stylistics: Speech, Writing and Thought Presentation in a Corpus of English Writing. London, New York, Routledge. 272 p.</mixed-citation></ref><ref id="ref14"><mixed-citation publication-type="other" xml:lang="ru">Tagabileva M.G., 2010. Word Formation Markup of the National Corpus of the Russian Language: Tasks and Methods. Conference Dialogue 2010”. URL: http://www.dialog-21.ru/digests/dialog2010/materials/pdf/73.pdf</mixed-citation></ref><ref id="ref15"><mixed-citation publication-type="other" xml:lang="ru">Teodorescu M.H., 2017. Machine Learning Methods for Strategy Research. HBS Working Paper, no. 18-011. August 2017. (Revised October 2017). 61 p.</mixed-citation></ref><ref id="ref16"><mixed-citation publication-type="other" xml:lang="ru">Yang Erhong National Language Commission Modern Chinese Language Corpus. Encyclopedia of China (Third Edition). Beijing. URL: https://www.zgbk.com/ecph/words?ID=393640&amp; SiteID=1&amp;Type=bkzyb&amp;webview_progress_ bar=1&amp;show_loading=</mixed-citation></ref><ref id="ref17"><mixed-citation publication-type="other" xml:lang="ru">Yu Shiwen et al., 2002. Basic Processing Norms of the Beijing University Modern Chinese Language Corpus. Journal of Chinese Information Science, vol. 16, iss. 5, pp. 51-56.</mixed-citation></ref><ref id="ref18"><mixed-citation publication-type="other" xml:lang="ru">Zakharov V.P., 2011. Korpusnaya lingvistika: ucheb.-metod. posobie [Corpus Linguistics. Tutorial]. Irkutstk. 161 p.</mixed-citation></ref><ref id="ref19"><mixed-citation publication-type="other" xml:lang="ru">CHILDES. URL: https://www.sketchengine.eu/childes-corpora/</mixed-citation></ref><ref id="ref20"><mixed-citation publication-type="other" xml:lang="ru">Corpus of English Wikipedia. URL: https://www.sketchengine.eu/english-wikipedia-corpus/</mixed-citation></ref><ref id="ref21"><mixed-citation publication-type="other" xml:lang="ru">National Corpus of the Russian Language. URL: https://ruscorpora.ru/?ysclid= lxhsi6qgi7535286247</mixed-citation></ref><ref id="ref22"><mixed-citation publication-type="other" xml:lang="ru">The American National Corpus. URL: https://anc.org/</mixed-citation></ref><ref id="ref23"><mixed-citation publication-type="other" xml:lang="ru">The British National Corpus. URL: https://www.english-corpora.org/bnc/</mixed-citation></ref><ref id="ref24"><mixed-citation publication-type="other" xml:lang="ru">The International Comparable Corpus. URL: https://www.researchgate.net/project/International-Comparable-Corpus-ICC</mixed-citation></ref><ref id="ref25"><mixed-citation publication-type="other" xml:lang="en">Baker P., 2006. Using Corpora in Discourse Analysis. London, Continuum, 2006. 208 p.</mixed-citation></ref><ref id="ref26"><mixed-citation publication-type="other" xml:lang="en">Baranov A.N., 2001. Vvedenie v prikladnuyu lingvistiku [Introduction to Applied Linguistics]. Moscow, Editorial URSS Publ., 2001. 360 p.</mixed-citation></ref><ref id="ref27"><mixed-citation publication-type="other" xml:lang="en">Biber D., 2001. Using Corpus-Based Methods to Investigate Grammar and Use: Some Case Studies on the Use of Verbs in English. Corpus Linguistics in North America, 2001, pp. 101-115.</mixed-citation></ref><ref id="ref28"><mixed-citation publication-type="other" xml:lang="en">Chesnokova I.D., Manshin M.E., 2018. Natsionalnyy korpus russkogo yazyka kak osnovnoy instrument poiska pri lingvisticheskikh issledovaniyakh (na primere poiska antonimov v publicisticheskikh tekstakh) [National Corpus of the Russian Language as the Main Search Tool in Linguistic Research (Based on the Search for Antonyms in Journalistic Texts)]. Izvestiya VGPU [Ivzestia of the Volgograd State Pedagogical University], no. 5 (128). URL: https://cyberleninka.ru/article/n/natsionalnyy-korpus-russkogo-yazyka-kak-osnovnoy-instrument-poiska-pri-lingvisticheskih-issledovaniyah-na-primere-poiska-antonimov-v</mixed-citation></ref><ref id="ref29"><mixed-citation publication-type="other" xml:lang="en">Finegan E., 2015. Language: Its Structure and Use. Stamford, CT, Cengage Learning. 575 p.</mixed-citation></ref><ref id="ref30"><mixed-citation publication-type="other" xml:lang="en">Kolpachkova E.N., 2019. Korpusy kitayskogo yazyka: sovremennoe sostoyanie i osnovnye problemy [Chinese Language Corpora: An Overview and Major Problems]. URL: https://orient.spbu.ru/images/document/2019/Kolpachkova_ Chinese_Corpus_overview.pdf</mixed-citation></ref><ref id="ref31"><mixed-citation publication-type="other" xml:lang="en">Kornienko A.V., 2023. Natsionalnyy korpus russkogo yazyka kak istochnikovaya baza sotsio-gumanitarnykh issledovaniy [National Corpus of the Russian Language as a Source Base for Social and Humanitarian Research]. Peterburgskaya sotsiologiya segodnya [Petersburg Sociology Today], no. 21, pp. 58-71. DOI: 10.25990/socinstras.pss-21.8sqp-bb68</mixed-citation></ref><ref id="ref32"><mixed-citation publication-type="other" xml:lang="en">Kotov A.A., Mineeva Z.I., 2013. Ispolzovanie natsionalnogo korpusa russkogo yazyka pri provedenii lingvisticheskoy ekspertizy [Use of Russian National Corpus in Linguistic Examination]. Filologicheskie nauki. Voprosy teorii i praktiki [Philology. Theory &amp; Practice], no. 8 (26): in 2 parts. Pt. 2, pp. 99-103.</mixed-citation></ref><ref id="ref33"><mixed-citation publication-type="other" xml:lang="en">McEnery T., Hardie A., 2011. Corpus Linguistics: Method, Theory and Practice. Cambridge, Cambridge University Press. 312 р.</mixed-citation></ref><ref id="ref34"><mixed-citation publication-type="other" xml:lang="en">McEnery T., Hardie A., 2012. Corpus Linguistics: Method, Theory and Practice. Cambridge, Cambridge University Press. 294 р.</mixed-citation></ref><ref id="ref35"><mixed-citation publication-type="other" xml:lang="en">Mustajoki A., 2007. Rol korpusov v lingvisticheskikh issledovaniyakh yazykov [Role of Corpora in Linguistic Research and Language Teaching]. Natsionalnyy korpus russkogo yazyka i problemy gumanitarnogo obrazovaniya: materialy Mezhdunar. nauch. konf. [National Corpus of the Russian Language and the Problems of Humanitarian Education. Proceedings of the International Scientific Conference]. Moscow, Nats. issled. un-t «Vyssh. shk. ekonomiki», pp. 152-166.</mixed-citation></ref><ref id="ref36"><mixed-citation publication-type="other" xml:lang="en">Piotrowski T., 2008. The Translator and Polish-English Corpora. Incorporating Corpora. The Linguist and the Translator. Clevedon, Multilingual Matters Ltd., pp. 117-132.</mixed-citation></ref><ref id="ref37"><mixed-citation publication-type="other" xml:lang="en">Semino E., Short M., 2004. Corpus Stylistics: Speech, Writing and Thought Presentation in a Corpus of English Writing. London, New York, Routledge. 272 p.</mixed-citation></ref><ref id="ref38"><mixed-citation publication-type="other" xml:lang="en">Tagabileva M.G., 2010. Word Formation Markup of the National Corpus of the Russian Language: Tasks and Methods. Conference Dialogue 2010”. URL: http://www.dialog-21.ru/digests/dialog2010/materials/pdf/73.pdf</mixed-citation></ref><ref id="ref39"><mixed-citation publication-type="other" xml:lang="en">Teodorescu M.H., 2017. Machine Learning Methods for Strategy Research. HBS Working Paper, no. 18-011. August 2017. (Revised October 2017). 61 p.</mixed-citation></ref><ref id="ref40"><mixed-citation publication-type="other" xml:lang="en">Yang Erhong National Language Commission Modern Chinese Language Corpus. Encyclopedia of China (Third Edition). Beijing. URL: https://www.zgbk.com/ecph/words?ID=393640&amp; SiteID=1&amp;Type=bkzyb&amp;webview_progress_ bar=1&amp;show_loading=</mixed-citation></ref><ref id="ref41"><mixed-citation publication-type="other" xml:lang="en">Yu Shiwen et al., 2002. Basic Processing Norms of the Beijing University Modern Chinese Language Corpus. Journal of Chinese Information Science, vol. 16, iss. 5, pp. 51-56.</mixed-citation></ref><ref id="ref42"><mixed-citation publication-type="other" xml:lang="en">Zakharov V.P., 2011. Korpusnaya lingvistika: ucheb.-metod. posobie [Corpus Linguistics. Tutorial]. Irkutstk. 161 p.</mixed-citation></ref><ref id="ref43"><mixed-citation publication-type="other" xml:lang="en">CHILDES. URL: https://www.sketchengine.eu/childes-corpora/</mixed-citation></ref><ref id="ref44"><mixed-citation publication-type="other" xml:lang="en">Corpus of English Wikipedia. URL: https://www.sketchengine.eu/english-wikipedia-corpus/</mixed-citation></ref><ref id="ref45"><mixed-citation publication-type="other" xml:lang="en">National Corpus of the Russian Language. URL: https://ruscorpora.ru/?ysclid= lxhsi6qgi7535286247</mixed-citation></ref><ref id="ref46"><mixed-citation publication-type="other" xml:lang="en">The American National Corpus. URL: https://anc.org/</mixed-citation></ref><ref id="ref47"><mixed-citation publication-type="other" xml:lang="en">The British National Corpus. URL: https://www.english-corpora.org/bnc/</mixed-citation></ref><ref id="ref48"><mixed-citation publication-type="other" xml:lang="en">The International Comparable Corpus. URL: https://www.researchgate.net/project/International-Comparable-Corpus-ICC</mixed-citation></ref></ref-list></back></article>
