|
Lecturer(s)
|
-
Změlík Richard, doc. Mgr. Ph.D.
|
|
Course content
|
Introduction to the subject; overview of the OCR task History of lexicographic work in the Czech lands Czech lexicography in the 20th century Typology of author dictionaries (international author dictionaries) Czech author dictionaries Quantitative linguistics in the Czech context: basic methods, concepts, and laws; applications of statistical measurement in linguistics and literary studies The Czech National Corpus: definition, issues of representativeness and standard language, and the concept of so-called minimal intervention The Corpus of Czech Verse; stylometry (Petr Plecháč) Building an author corpus: lemmatization, tokenization, disambiguation, annotation, and thematic concentration Statistical models and their interpretation in literary studies: examples of corpus analysis using the Jan Čep corpus (thematic concentration measurement) and other author corpora Statistical models and their interpretation in literary studies: examples of corpus analysis using the Jan Čep corpus (statistical analysis of colours) and other author corpora Proofreading and correction of OCR texts of Jakub Arbes's romanettos Proofreading and correction of OCR texts of Jakub Arbes's romanettos
|
|
Learning activities and teaching methods
|
|
Monologic Lecture(Interpretation, Training), Dialogic Lecture (Discussion, Dialog, Brainstorming), Work with Text (with Book, Textbook)
|
|
Learning outcomes
|
The seminar will focus on author corpora and their use in literary studies. As early as the 1980s, Pavel Vašák pointed out the usefulness of author dictionaries for literary scholarship, as well as of more extensive lexicons of period poetics (Romanticism, Realism, the poetics of the Ruch and Lumír literary groups, etc.). The primary aim of such projects is to provide literary history, and potentially literary theory, with a substantial body of textual material that can be subjected to quantitative analysis and used to construct statistical models interpretable within a literary-studies framework. In the seminar, students will learn not only about the history of author lexicography but, above all, will actively participate in the creation of an author corpus of Jakub Arbes's romanettos. The completed corpus will be published within the Czech National Corpus and made available for quantitative analysis of the genre. The course will also involve working with existing author corpora. Particular emphasis will be placed on the interpretation of statistical models from the perspective of literary studies.
The student will gain a basic overview of lexicographic methods, with particular emphasis on the typology of author dictionaries (primarily those produced abroad). They will also become familiar with the main principles of compiling a digital author dictionary.
|
|
Prerequisites
|
unspecified
|
|
Assessment methods and criteria
|
Student performance, Seminar Work
Active participation in seminars and submission of the assigned seminar paper.
|
|
Recommended literature
|
-
ČERMÁK, František - BARTOŇ, Tomáš. (2007). Slovník Karla Čapka. Praha.
-
ČERMÁK, František - BLATNÁ, Renata. (2005). Jak využívat Český národní korpus. Praha.
-
ČERMÁK, František - BLATNÁ, Renata. (2006). Korpusová lingvistika: stav a modelové přístupy. Praha.
-
ČERMÁK, František - BLATNÁ, Renata. Korpusová lingvistika: stav a modelové přístupy. Praha. 2006.
-
ČERMÁK, František - BLATNÁ, Renata. (1995). Manuál lexikografie. Jinočany.
-
ČERMÁK, František - CVRČEK, Václav - SCHMIEDTOVÁ, Věra. (2010). Slovník komunistické totality. Praha.
-
ČERMÁK, František - CVRČEK, Václav. (2009). Slovník Bohumila Hrabala. Praha.
-
ČERMÁK, František - KŘEN, Michal. (2010). Frekvenční slovník češtiny. Praha.
-
ČERMÁK, František - ŠULC, Michal. (2006). Kolokace. Praha.
-
ČERMÁK, František. Frazeologie a idiomatika česká a obecná. Praha. 2007.
-
ČERMÁK, František. (2007). Frazeologie a idiomatika česká a obecná.
-
ČERMÁK, František. "Jazykový korpus: Prostředek a zdroj poznání." In Slovo a slovesnost, roč. 56, č. 2, s. 119-140..
-
ČERMÁK, František. (2005). Korpus, informace a lingvistika." In Přednášky z XLVIII. běhu Letní školy slovanských studií. Praha: Univerzita Karlova, Filozofická fakulta, s. 15-24. Dostupné z WWW: http://korpus.cz/doc/korp-info-lingv.rtf.
-
ČERMÁK, František. (2004). "Korpusová lingvistika: stručný historický přehled." In Český národní korpus. Dostupné z WWW: <http://korpus.cz/doc/2002_cnk.rtf>..
-
ČERMÁK, František. (2010). Lexikon a sémantika. Praha.
|