Multilingual NER Toolkit: Difference between revisions
Appearance
Content deleted Content added
Importing NeoWiki demo data |
Importing NeoWiki demo data |
||
| neo | |||
|---|---|---|---|
| Line 7: | Line 7: | ||
"statements": { |
"statements": { |
||
"Description": { |
"Description": { |
||
" |
"propertyType": "text", |
||
"value": [ |
"value": [ |
||
"Open-source named-entity recognition tooling tuned for historical European languages." |
"Open-source named-entity recognition tooling tuned for historical European languages." |
||
| Line 13: | Line 13: | ||
}, |
}, |
||
"Started": { |
"Started": { |
||
" |
"propertyType": "number", |
||
"value": 2022 |
"value": 2022 |
||
}, |
}, |
||
"Lead institution": { |
"Lead institution": { |
||
" |
"propertyType": "relation", |
||
"value": [ |
"value": [ |
||
{ |
{ |
||
Latest revision as of 22:06, 27 July 2026
The Multilingual NER Toolkit is an open-source named-entity recognition project tuned for historical European languages. Started in 2022, it ships models and tooling that handle the spelling variation, archaic vocabulary, and code-switching common in pre-modern texts.
Led by Sorbonne University, it serves as a foundation for downstream work on entity disambiguation and linking across cultural heritage corpora.