- FIN-CLARIAH Research Infrastructure
A new national research infrastructure initiative FIN-CLARIAH for...
8.12.2021 8:12 by eahyvone - WarMemoirSampo published on December 3, 2021
A new “Sampo” application, “WarMemoirSampo”...
8.12.2021 8:04 by eahyvone - Five new SeCo papers accepted for the ISWC 2021
The 20th International Semantic Web Conference (ISWC 2021), the...
2.8.2021 6:53 by eahyvone
- Annastiina Ahola, Lilli Peura, Rafael Leal, Heikki Rantala and Eero Hyvönen: Using generative AI and LLMs to enrich art collection metadata for searching, browsing, and studying art history in Digital Humanities
- Eero Hyvönen, Petri Leskinen, Henna Poikkimäki, Heikki Rantala, Jouni Tuominen, Senka Drobac, Ossi Koho, Ilona Pikkanen and Hanna-Leena Paloposki: LetterSampo Finland (1809–1917) Data Service and Portal: Searching, Exploring, and Analyzing Historical Letters and Their Underlying Networks
- Michael Lewis, Eljas Oksanen, Frida Ehrnsten, Heikki Rantala, Jouni Tuominen and Eero Hyvönen: The Impact of Human Decision-making on the Research Value of Archaeological Data
- Tomaž Erjavec, Matyáš Kopp, Nikola Ljubešić, Taja Kuzman, Paul Rayson, Petya Osenova, Maciej Ogrodniczuk, Çağrı Çöltekin, Danijel Koržinek, Katja Meden, Jure Skubic, Peter Rupnik, Tommaso Agnoloni, José Aires, Starkaður Barkarson, Roberto Bartolini, Núria Bel, María Calzada Pérez, Roberts Darģis, Sascha Diwersy, Maria Gavriilidou, Ruben van Heusden, Mikel Iruskieta, Neeme Kahusk, Anna Kryvenko, Noémi Ligeti-Nagy, Carmen Magariños, Martin Mölder, Costanza Navarretta, Kiril Simov, Lars Magne Tungland, Jouni Tuominen, John Vidler, Adina Ioana Vladu, Tanja Wissik, Väinö Yrjänäinen and and Darja Fišer: ParlaMint II: Advancing Comparable Parliamentary Corpora Across Europe
|
Knowledge Extraction from Natural Language
Case Finnish (and Swedish) Texts |
Knowledge Extraction from Finnish Texts
Goals of Research
In Digital Humanities (DH) the data comes in often in textual form (e.g., news, biographies, articles, novels). When publishing such content as Linked Data for DH analysis, the meaning of literal mentions of entities (such as names of persons and places), concepts, relations, events, topics, etc. have to be extracted from unstructured texts and represented as structured semantic data for the computer. In our own work, for example, such knowledge extraction has been needed when developing the Sampo series of semantic portals.
During this research, various natural language processing (NLP) tools and linguistic datasets have been developed. In our view, such tools and resources should be made openly available on the Web as web services. NLP tools and resources would be an important part of the Linked Open Data Infrastructure for Digital Humanities in Finland.
Services for NLP
More information will appear here later.
Contact Persons
Dr. Cand. Rafael Leal, Aalto University
Prof. Eero Hyvönen, University of Helsinki (HELDIG) and Aalto


