Final data cleanup needed: a look at the Czech corpus

Hello @SantosCardonaPR and thanks for this.

Indeed, the Czech corpus is clean. The Polish one still has one duplicate: there are two codes both called “Europe”:

The former one is embedded in the Z hierarchy of POPREBEL, but then only used (within POPREBEL) as a parent code (of “Germany”, “Poland” etc.). The latter one was used in actual annotations by @Jirka_Kocian, @Richard and @Wojt . Is that the way you want it?