Vector representations of polish words (Word2Vec method)
Użyj poniższego opisu do cytowania zasobu albo wyeksportuj go w wybranym formacie:
Kędzia, Paweł; Czachor, Gabriela; Piasecki, Maciej and Kocoń, Jan, 2016,
Vector representations of polish words (Word2Vec method), CLARIN-PL Repository,
http://hdl.handle.net/11321/327.
Autorzy
Data
2016-11-07
Rozmiar
4000000000 tokens
Języki
Opis
Model skip gram with vectors of length 100. Trained on kgr 10, a corpora with over 4 billion tokens. Data preprocessing involved segmentation, lemmatization and mophosyntactic disambiguation with MWE annotation.
Słowa kluczowe
Kolekcje
Pliki w tym zasobie
- Nazwa
- skipgram_v100.zip
- Rozmiar
- 877.62 MB
- Format
- application/zip
- Opis
- Suma kontrolna MD5
- 0497754c5a6cb276dc5b7dc3807ad1da

The file preview has not been generated yet. Please try again later or contact the system administrator dspace@clarin-pl.eu