Skip to content

Normalize Unicode when registering auxiliary lemmas #159

Description

@dan-zeman

As normalized Unicode is required in treebank data, we should make sure that people do not use unnormalized Unicode when registering their auxiliaries. The scripts processing data from https://quest.ms.mff.cuni.cz/udvalidator/cgi-bin/unidep/langspec/specify_auxiliary.pl should normalize the lemma as well as any examples.

Metadata

Metadata

Assignees

Projects

No projects

Milestone

No milestone

Relationships

None yet

Development

No branches or pull requests

Issue actions