Creating and validating multilingual semantic representations for six languages - Research Portal

Associated organisational units

Electronic data

W17-1908
Final published version, 1.01 MB, PDF document
Available under license: CC BY: Creative Commons Attribution 4.0 International License

Creating and validating multilingual semantic representations for six languages: expert versus non-expert crowds

Research output: Contribution in Book/Report/Proceedings - With ISBN/ISSN › Conference contribution/Paper › peer-review

Published

Publication date	3/04/2017
Host publication	Proceedings of the 1st Workshop on Sense, Concept and Entity Representations and their Applications: Proceedings of the 15th Conference of the European Chapter of the Association for Computational Linguistics
Publisher	Association for Computational Linguistics
Pages	61-71
Number of pages	11
ISBN (print)	9781945626500
<mark>Original language</mark>	English

Abstract

Creating high-quality wide-coverage multilingual semantic lexicons to support knowledge-based approaches is a challenging time-consuming manual task. This has traditionally been performed by linguistic experts: a slow and expensive process. We present an experiment in which we adapt and evaluate crowdsourcing methods employing native speakers to generate a list of coarse-grained senses under a common multilingual semantic taxonomy for sets of words in six languages.
451 non-experts (including 427 Mechanical Turk workers) and 15 expert participants semantically annotated 250 words manually for Arabic, Chinese, English, Italian, Portuguese and Urdu lexicons. In order to avoid erroneous (spam) crowdsourced results, we used a novel taskspecific two-phase filtering process where users were asked to identify synonyms in the target language, and remove erroneous senses.

Research

Associated organisational units

Electronic data

Links

Creating and validating multilingual semantic representations for six languages: expert versus non-expert crowds

Abstract

Quick Links

Connect With Us

Faculties & Depts

Contact Us

Research

Associated organisational units

Electronic data

Links

Creating and validating multilingual semantic representations for six languages: expert versus non-expert crowds

Abstract

Related research outputs

AraSAS: The Open Source Arabic Semantic Tagger

Quick Links

Connect With Us

Faculties & Depts

Contact Us