Shared Task on Hierarchical Classification of Blurbs - GermEval 2019

This dataset can be used as a benchmark for clustering word embeddings for German. It contains 18'084 unique samples, 28 splits with 177 to 16'425 samples, and 4 to 93 unique classes.