From Sparse to Sense-Grounded: Wikipedia Training for Ukrainian Visual-WSD
30th Conference on Computational Natural Language Learning · San Diego, California
Extended the benchmark to 381 instances covering 172 unique lemmas and introduced two scalable approaches for constructing sense-grounded Visual-WSD tuning data with limited manual supervision. SenseWiki-UA links dictionary senses to sense-specific Ukrainian Wikipedia pages and harvests aligned image–text pairs; RA-Wiki-UA retrieves Wikipedia images near the benchmark’s visual distribution and pairs them with generated Ukrainian captions.
ACL Anthology paper ↗