Dr. Livingstone, I presume? Polishing of foreign character identification in literary texts
| dc.contributor.author | Konovalova Aleksandra | |
| dc.contributor.author | Toral Antonio | |
| dc.contributor.author | Taivalkoski-Shilov Kristiina | |
| dc.contributor.organization | fi=kieli- ja käännöstieteiden laitos|en=School of Languages and Translation Studies| | |
| dc.contributor.organization-code | 2602100 | |
| dc.converis.publication-id | 176107951 | |
| dc.converis.url | https://research.utu.fi/converis/portal/Publication/176107951 | |
| dc.date.accessioned | 2022-10-28T12:42:33Z | |
| dc.date.available | 2022-10-28T12:42:33Z | |
| dc.description.abstract | <p>Character identification is a key element for many narrative-related tasks. To implement it, the baseform of the name of the character (or lemma) needs to be identified, so different appearances of the same character in the narrative could be aligned. In this paper we tackle this problem in translated texts (English–Finnish translation direction), where the challenge regarding lemmatizing foreign names in an agglutinative language appears. To solve this problem, we present and compare several methods. The results show that the method based on a search for the shortest version of the name proves to be the easiest, best performing (83.4% F1), and most resource-independent.</p> | |
| dc.format.pagerange | 123 | |
| dc.format.pagerange | 128 | |
| dc.identifier.isbn | 978-1-955917-73-5 | |
| dc.identifier.olddbid | 178391 | |
| dc.identifier.oldhandle | 10024/161485 | |
| dc.identifier.uri | https://www.utupub.fi/handle/11111/43311 | |
| dc.identifier.url | https://aclanthology.org/2022.naacl-srw.16/ | |
| dc.identifier.urn | URN:NBN:fi-fe2022091258602 | |
| dc.language.iso | en | |
| dc.okm.affiliatedauthor | Konovalova, Aleksandra | |
| dc.okm.discipline | 113 Computer and information sciences | en_GB |
| dc.okm.discipline | 616 Other humanities | en_GB |
| dc.okm.discipline | 113 Tietojenkäsittely ja informaatiotieteet | fi_FI |
| dc.okm.discipline | 616 Muut humanistiset tieteet | fi_FI |
| dc.okm.internationalcopublication | international co-publication | |
| dc.okm.internationality | International publication | |
| dc.okm.type | A4 Conference Article | |
| dc.publisher.country | United States | en_GB |
| dc.publisher.country | Yhdysvallat (USA) | fi_FI |
| dc.publisher.country-code | US | |
| dc.relation.conference | Conference of the North American Chapter of the Association for Computational Linguistics | |
| dc.relation.doi | 10.18653/v1/2022.naacl-srw.16 | |
| dc.source.identifier | https://www.utupub.fi/handle/10024/161485 | |
| dc.title | Dr. Livingstone, I presume? Polishing of foreign character identification in literary texts | |
| dc.title.book | Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies: Student Research Workshop | |
| dc.year.issued | 2022 |
Tiedostot
1 - 1 / 1