Vector representations of polish words (Word2Vec method)

Użyj poniższego opisu do cytowania zasobu albo wyeksportuj go w wybranym formacie:
Kędzia, Paweł; Czachor, Gabriela; Piasecki, Maciej and Kocoń, Jan, 2016, Vector representations of polish words (Word2Vec method), CLARIN-PL Repository, http://hdl.handle.net/11321/327.
Data
2016-11-07
Rozmiar
4000000000 tokens
Języki
Opis
Model skip gram with vectors of length 100. Trained on kgr 10, a corpora with over 4 billion tokens. Data preprocessing involved segmentation, lemmatization and mophosyntactic disambiguation with MWE annotation.
Ten zasób jestPublicznie dostępnyi został udostępniony na licencji:GNU LGPL 3.0

Pliki w tym zasobie

Nazwa
skipgram_v100.zip
Rozmiar
877.62 MB
Format
application/zip
Opis
Suma kontrolna MD5
0497754c5a6cb276dc5b7dc3807ad1da
Preview
  Podgląd pliku
    The file preview has not been generated yet. Please try again later or contact the system administrator