Publications

Selected research.

A curated selection across my research, with one factual contribution note per paper. The profiles below contain the complete record.

2026
Multilingual NLPFactuality & EvaluationLinguistics & Semantics Computational Linguistics

A Principled Framework for Evaluating on Typologically Diverse Languages

Esther Ploeger, Wessel Poelman, Andreas Holck Høeg-Petersen, Anders Schlichtkrull, Miryam de Lhoneux, and Johannes Bjerva

Introduces an explicit language-sampling framework and metrics, demonstrating that language selection materially affects the generalisability of multilingual evaluation claims.

2026
Factuality & EvaluationMultilingual NLP ACL 2026

How Good is Your Wikipedia? Auditing Data Quality for Low-resource and Multilingual NLP

Kushal Tatariya, Artur Kulmizev, Wessel Poelman, Esther Ploeger, Marcel Bollmann, Johannes Bjerva, Jiaming Luo, Heather Lent, and Miryam de Lhoneux

Audits non-English Wikipedia with quality filters usually applied to noisy web data, showing that multilingual NLP must treat data quality—not only coverage—as a first-class concern.

2025
Security & Privacy TACL

NLP Security and Ethics, in the Wild

Heather Lent, Erick Galinkin, Yiyi Chen, Jens Myrup Pedersen, Leon Derczynski, and Johannes Bjerva

Connects NLP-security research with established cybersecurity ethics, develops the idea of “white-hat NLP,” and offers guidance for harm minimisation and responsible disclosure.

2024
Multilingual NLPFactuality & EvaluationLinguistics & Semantics EMNLP 2024

What is “Typological Diversity” in NLP?

Esther Ploeger, Wessel Poelman, Miryam de Lhoneux, and Johannes Bjerva

Shows that NLP papers use “typological diversity” inconsistently, and provides measures for more defensible language sampling and multilingual evaluation.

2024
Multilingual NLPFactuality & Evaluation TACL

CreoleVal: Multilingual Multitask Benchmarks for Creoles

Heather Lent, Kushal Tatariya, Raj Dabre, Yiyi Chen, Marcell Fekete, Esther Ploeger, Li Zhou, Ruth-Ann Armstrong, Abee Eijansantos, Catriona Malau, Hans Erik Heje, Ernests Lavrinovics, Diptesh Kanojia, Paul Belony, Marcel Bollmann, Loïc Grobol, Miryam de Lhoneux, Daniel Hershcovich, Michel DeGraff, Anders Søgaard, and Johannes Bjerva

Provides benchmarks for eight NLP tasks across as many as 28 Creole languages, including newly developed datasets.

2024
Linguistics & SemanticsMultilingual NLP Computational Linguistics

The Role of Typological Feature Prediction in NLP and Linguistics

Johannes Bjerva

Critically examines whether typological feature prediction serves its claimed purposes in NLP and linguistic documentation, using a survey of typologists to offer recommendations for better interdisciplinary alignment.

2023
Linguistics & SemanticsMultilingual NLP NoDaLiDa 2023

Colex2Lang: Language Embeddings from Semantic Typology

Yiyi Chen, Russa Biswas, and Johannes Bjerva

Learns language representations from cross-linguistic colexification graphs, providing a semantic-typology route to modelling language similarity.

2021
Multilingual NLP *SEM 2021

Inducing Language-Agnostic Multilingual Representations

Wei Zhao, Steffen Eger, Johannes Bjerva, and Isabelle Augenstein

Tests methods for removing language-identity signals from multilingual representations across 19 languages, clarifying when representation alignment improves cross-lingual transfer and when intuitive normalisation methods fail.

2020
Multilingual NLP EMNLP 2020

Zero-Shot Cross-Lingual Transfer with Meta Learning

Farhad Nooralahzadeh, Giannis Bekoulis, Johannes Bjerva, and Isabelle Augenstein

Uses meta-learning to select source-language training instances for zero-shot transfer across 15 languages, with a typological analysis of when learned cross-lingual sharing helps.

2020
Factuality & Evaluation EMNLP 2020

SubjQA: A Dataset for Subjectivity and Review Comprehension

Johannes Bjerva, Nikita Bhutani, Behzad Golshan, Wang-Chiew Tan, and Isabelle Augenstein

Introduces a question-answering dataset with subjectivity annotations for questions and answer spans across six review domains, making subjectivity effects measurable in extractive QA.

2019
Linguistics & SemanticsMultilingual NLP Computational Linguistics

What Do Language Representations Really Represent?

Johannes Bjerva, Robert Östling, Maria Han Veiga, Jörg Tiedemann, and Isabelle Augenstein

Disentangles structural, genetic, and geographical signals in language representations learned from translated text, finding structural similarity strongest while genetic relatedness can act as a confound.

2019
Linguistics & SemanticsMultilingual NLP NAACL 2019

A Probabilistic Generative Model of Linguistic Typology

Johannes Bjerva, Yova Kementchedjhieva, Ryan Cotterell, and Isabelle Augenstein

Models languages and typological features jointly through exponential-family matrix factorisation, connecting feature covariance with generalisation to unseen languages.

2016
Linguistics & SemanticsMultilingual NLP COLING 2016

Semantic Tagging with Deep Residual Networks

Johannes Bjerva, Barbara Plank, and Johan Bos

Introduces semantic tagging as a fine-grained task for multilingual semantic parsing and presents a deep residual architecture for the task.