NameTag 3 Multilingual Model 250203

PID

This is a trained model for the supervised machine learning tool NameTag 3 (https://ufal.mff.cuni.cz/nametag/3/). NameTag 3 is an open-source tool for both flat and nested named entity recognition (NER). NameTag 3 identifies proper names in text and classifies them into a set of predefined categories, such as names of persons, locations, organizations, etc. The model was trained jointly on 21 flat NE corpora of 17 languages: Arabic, Chinese, Croatian, Czech, Danish, Dutch, English, German, Maghrebi Arabic French, Norwegian Bokmaal, Norwegian Nynorsk, Portuguese, Serbian, Slovak, Spanish, Swedish, and Ukrainian. The model documentation can be found at https://ufal.mff.cuni.cz/nametag/3/models#multilingual.

Identifier
PID http://hdl.handle.net/11234/1-5859
Related Identifier https://ufal.mff.cuni.cz/nametag/3/
Metadata Access http://lindat.mff.cuni.cz/repository/oai/request?verb=GetRecord&metadataPrefix=oai_dc&identifier=oai:lindat.mff.cuni.cz:11234/1-5859
Provenance
Creator Straková, Jana; Straka, Milan
Publisher Charles University, Faculty of Mathematics and Physics, Institute of Formal and Applied Linguistics (UFAL)
Publication Year 2025
Rights Creative Commons - Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0); http://creativecommons.org/licenses/by-nc-sa/4.0/; PUB
OpenAccess true
Contact lindat-help(at)ufal.mff.cuni.cz
Representation
Language Arabic; Chinese; Croatian; Czech; Danish; Dutch; Flemish; English; German; Bokmål, Norwegian; Norwegian Bokmål; Norwegian Nynorsk; Nynorsk, Norwegian; Portuguese; Serbian; Slovak; Spanish; Castilian; Swedish; Ukrainian; Undetermined
Resource Type languageDescription
Format application/zip; application/octet-stream; downloadable_files_count: 1
Discipline Linguistics