KGR10 FastText Polish word embeddings
Użyj poniższego opisu do cytowania zasobu albo wyeksportuj go w wybranym formacie:
Kocoń, Jan, 2018,
KGR10 FastText Polish word embeddings, CLARIN-PL Repository,
http://hdl.handle.net/11321/606.
Autorzy
Data
2018-09-28
Języki
Opis
Distributional language model (both textual and binary) for Polish (word embeddings) trained on KGR10 corpus (over 4 billion of words) using Fasttext with the following variants (all possible combinations):
- dimension: 100, 300
- method: skipgram, cbow
- tool: FastText, Magnitude
- source text: plain, plain.lower, plain.lemma, plain.lemma.lower
The link below leads to the NextCloud directory with all variants of embeddings. If you use it, please cite the following article:
@article{kocon2018embeddings,
author = {Koco\'{n}, Jan and Gawor, Micha{\l}},
title = {Evaluating {KGR10} {P}olish word embeddings in the recognition of temporal
expressions using {BiLSTM-CRF}},
journal = {Schedae Informaticae},
volume = {27},
year = {2018},
url = {http://www.ejournals.eu/Schedae-Informaticae/2018/Volume-27/art/13931/},
doi = {10.4467/20838476SI.18.008.10413}
}
Finansowanie
Ministry of Science and Higher Education (Poland)
Nazwa projektu:CLARIN-PL
Słowa kluczowe
Kolekcje
Dostęp do plików
Pliki w tym zasobie
Pobierz kompletny pakiet albo skopiuj gotowe polecenie do automatycznego pobierania.
- Nazwa
- fast_text_kgr_em.zip
- Rozmiar
- 196 B
- Format
- application/zip
- Opis
- Suma kontrolna MD5
- 0e4ba54d75f6a13c17bfac18b03d113d

Podgląd pliku
- Podgląd niedostępny
Podgląd tego pliku nie został jeszcze wygenerowany. Spróbuj ponownie później albo skontaktuj się z administratorem systemu.
dspace@clarin-pl.eu