With the CMU Twitter dataset (Tweebank), some kinds of MWEs are already annotated. It would be nice to either (i) choose supersenses subject to the annotated MWE segmentation, or (ii) choose the segmentation and supersenses such that the annotated MWEs, at minimum, are present (additional MWEs may be predicted).
With the CMU Twitter dataset (Tweebank), some kinds of MWEs are already annotated. It would be nice to either (i) choose supersenses subject to the annotated MWE segmentation, or (ii) choose the segmentation and supersenses such that the annotated MWEs, at minimum, are present (additional MWEs may be predicted).