Závěrečná práce: Petr Zelina, učo 469366: Pretraining and Evaluation of Czech ALBERT Language Model
Bakalářská práce
Pretraining and Evaluation of Czech ALBERT Language Model
Anotace
Tato práce se zabývá novým jazykovým modelem nazvaným ALBERT, který byl vydaný společností Google Research v roce 2019. Architektura ALBERT je velmi úspěšná ve zpracování anglických jazykových problémů a tato práce testuje možnosti jejího využití pro český jazyk. Práce nejdříve rozebírá vývoj a strukturu modelu ALBERT. Dále se zabývá návrhem a trénováním několika českých modelů. Nakonec porovnává jejich …více
Abstract
This thesis explores a new language model called ALBERT, released by Google Research in 2019. The ALBERT architecture has been very successful in English NLP tasks, and this thesis tests its applicability for the Czech language. The work gives an overview of the models leading to ALBERT and their design. Further, the creation process of several Czech ALBERT models is described. Finally, their performance …více
Zadání práce
A new powerful language model called ALBERT was presented by Google in 2019 when it reached state-of-the-art results in various NLP tasks for English (question answering, language understanding). ALBERT is based on the idea of transfer learning, where it requires a huge text dataset for pretraining general language representations and in successive separate training it can be fine tuned for a specific task.
The goal of this thesis is to pretrain a ALBERT model(s) for the Czech language, evaluate the model efficiency in several extrinsic tasks and compare the results to other current state-of-the-art Czech NLP language models.
The submitted thesis will consist of a thesis text and a practical part. The textual part will include a detailed description of ALBERT and other related language modelling approaches and information about available pretrained models suitable for comparison, a description of the pretraining process, and the details of the evaluation and comparison. The practical part will contain all code and data for replicating the results with the respective licenses for publication where possible.
27. 5. 2020 09:31, doc. RNDr. Aleš Horák, Ph.D., učo 1648
- Zadáno/změněno 24. 6. 2020 16:18, Helena Kryštofová
- Záznam založen 30. 4. 2020 12:19, Jana Zemanová, učo 9619
- Zveřejnit od 26. 5. 2020 07:23, Eva Drštková
- Práce převzata 26. 5. 2020 07:23, Eva Drštková
Práce na příbuzné téma
Seznam prací, které mají shodná klíčová slova.
-
Automatic text summarization
Mgr. Adam Hájek -
Classification and Named Entity Recognition in Airline Emails
Mgr. Bruno Húževka -
Domain-specific English-Czech Neural Machine Translation
Mgr. Martin Wörgötter -
Mining Czech Clinical Notes Using the Language Modelling Technology
Mgr. Tomáš Houfek -
Utilisation of language representations for Information Retrieval
Ing. Petr Mička -
Prediction of missing peaks in mass spectra
Bc. Michal Starý -
Data extraction from medical records
Mgr. Tomáš Houfek -
Komunikace s umělou inteligencí: hledání hranic porozumění
Bc. Veronika Urbášková




