Proceed to contents
Natural Language Understanding with Multilingual Data

Lost in Meaning - Found in Translation

16 Oct 2023 16:00 - 17:00

The task of translation involves language understanding and generation and, in this way, naturally combines the two essential challenges in computational linguistics and language technology.

Practical information:

16 Oct 2023 16:00 - 17:00
1h hours
Online
English
Target audience: Researchers

Want to register?

  • Register until: 13 Oct 2023
  • Price: free
More info & registration ⇗

Georganiseerd door:

The FoTran project is interested in the ability of neural translation models to pick up linguistic properties and to generalise to meaningful representations when trained on large amounts of multilingual data. Their focus is on the effect of linguistic diversity on abstraction and generalisation. In order to study this, they need to create the necessary resources and infrastructure.

In this talk, Jörg Tiedemann will first introduce the OPUS ecosystem that fuels his research. In the second part, he will concentrate on the experiments, studies and developments that this ecosystem enables within and outside of FoTran. He welcomes discussions on further directions that can be taken with the multilingual infrastructure the FoTran project builds.

Teacher / speaker

Jörg Tiedemann is professor of language technology at the Department of Digital Humanities at the University of Helsinki. He received his PhD in computational linguistics for work on bitext alignment and machine translation from Uppsala University before moving to the University of Groningen for 5 years of post-doctoral research on question answering and information extraction. His main research interests are connected with massively multilingual data sets and data-driven natural language processing and he currently runs an ERC-funded project on representation learning and natural language understanding.