PaperScope
LIVE · 2026-09-04 05:40 UTC

Typological Feature Prediction with Large Language Models: An In-Context Learning Approach

Qianwen Wang, York Hay Ng, Aditya Khan, En-Shiun Annie Lee

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.03775 v1
Category
Submitted
2026-09-03

Abstract

Typological features are widely used in multilingual NLP, and the prediction of such features holds downstream utility. However, existing methods to predict missing values lack interpretable justifications for predictions, while their performance across resource levels and feature types remains underexplored. Given LLMs' abilities in meta-linguistic reasoning and in providing rationales, we investigate LLMs' performance in typological feature prediction via an in-context learning approach with linguistic data from URIEL+ and Glottolog. We find that zero-shot prompting is insufficient, but when given phylogenetic and geographic neighbour evidence, LLMs substantially outperform all baselines without disadvantaging low-resource languages. We further find that most LLM rationales are consistent with the provided evidence, offering a step toward explainable typological feature prediction.

Comment: Accepted to EMNLP 2026

arXiv abs page · PDF · same-day batch