Filip Ginter profile picture
Filip
Ginter
Professor, Data analytics
human language technology, natural language processing, machine learning applied to human language, both methodological and resource creation research

Contact

Areas of expertise

natural language processing
human language technology
machine learning
deep learning
resource development

Biography

I am a researcher at the Department of Computing, University of Turku. My research is in the area of natural language processing. I belong to the TurkuNLP (turkunlp.org) research group.

I was born in 1978 in Ostrava, Czech Republic (Czechoslovakia back then). In 2001, I got a M.Sc. (tech) in computer science at the computer science department of VSB - Technical University Ostrava. My major subject was artificial intelligence. I gained a PhD in computer science in 2007. The title of my thesis is Towards Information Extraction in the Biomedical Domain: Methods and Resources.

As of 2022, I am a professor of language technology and as of 2021 the deputy director of the Department of Computing.

Teaching

I have been actively teaching since early on during my PhD studies. I independently prepared my first advanced level NLP course in 2004, and since ca. 2008 I have been teaching at least one course every year, substantially more during my bioinformatics lecturer appointment. While a lecturer in the bioinformatics MSc degree programme, I was lecturing international students in two cities. In 2016, I was tasked with developing and coordinating the introduction of a new 20 ECTS study module on natural language processing. This module is, with modifications, still in use and shared between the departments of Languages and Computing, both in terms of teaching and in terms of students. In 2019-2020 and 2020-2021 I was also co-lecturing, upon invitation, two courses in natural language processing in the Arcada University of Applied Sciences in Helsinki.

Research

My primary field of research is language technology / natural language processing. In my post-PhD career, I have focused on the development of NLP tools and resources primarily for Finnish, but later also numerous other languages via the Universal Dependencies project. My work is heavy on resource development, both in terms of data and machine learning pipelines. Open science and resources play an important role in my research, much of which is carried out in the open on GitHub and as a rule, all resources are openly available for unrestricted use. I work collaboratively, especially with my younger colleagues, rather than striving for deeper, primary author inquiries.

Publications

Sort by:

Explaining Classes through Stable Word Attributions (2022)

Annual Meeting of the Association for Computational Linguistics, Annual Meeting of the Association for Computational Linguistics
Rönnqvist Samuel, Myntti Amanda, Kyröläinen Aki-Juhani, Ginter Filip, Laippala Veronika
(Vertaisarvioitu artikkeli konferenssijulkaisussa (A4))

GEMv2: Multilingual NLG Benchmarking in a Single Line of Code (2022)

Empirical Methods in Natural Language Processing
Gehrmann S., Bhattacharjee A., Mahendiran A., Wang A., Papangelis A., Madaan A., McMillan-Major A., Shvets A., Upadhyay A., Bohnet B., Yao B., Wilie B., Bhagavatula C., You C., Thomson C., Garbacea C., Wang D., Deutsch D., Xiong D., Jin D., Gkatzia D., Radev D., Clark E., Durmus E., Ladhak F., Ginter F., Winata G.I., Strobelt H., Hayashi H., Novikova J., Kanerva J., Chim J., Zhou J., Clive J., Maynez J., Sedoc J., Juraska J., Dhole K., Chandu K.R., Perez-Beltrachini L., Ribeiro L.F.R., Tunstall L., Zhang L., Pushkarna M., Creutz M., White M., Kale M.S., Eddine M.K., Daheim N., Subramani N., Dusek O., Liang P.P., Ammanamanchi P.S., Zhu Q., Puduppully R., Kriz R., Shahriyar R., Cardenas R., Mahamood S., Osei S., Cahyawijaya S., Štajner S., Montella S., Jolly S., Mille S., Hasan T., Shen T., Adewumi T., Raunak V., Raheja V., Nikolaev V., Tsai V., Jernite Y., Xu Y., Sang Y., Liu Y., Hou Y.
(Vertaisarvioitu artikkeli konferenssijulkaisussa (A4))

Textual Paraphrase Dataset for Deep Language Modelling (2022)

Kanerva Jenna, Ginter Filip, Chang Li-Hsin, Skantsi Valtteri, Kilpeläinen Jemina, Kupari Hanna-Mari, Piirto Aurora, Saarni Jenna, Sevón Maija, Tarkka Otto
(Vertaisarvioitu artikkeli kokoomateoksessa (A3))

Fine-grained Named Entity Annotation for Finnish (2021)

Nordic Conference on Computational Linguistics, Linköping Electronic Conference Proceedings
Luoma Jouni, Chang Li-Hsin, Ginter Filip, Pyysalo Sampo
(Vertaisarvioitu artikkeli konferenssijulkaisussa (A4))