Yayınlanmış 1 Ocak 2025 | Sürüm v1
Konferans bildirisi Açık

A Semi-Automated Approach to the Annotation of Argument Structures in Turkish Datasets

  • 1. Ozyegin Univ, Istanbul, Turkiye
  • 2. Bogazici Univ, Bebek, Turkiye
  • 3. Galatasaray Univ, Istanbul, Turkiye
  • 4. Mudanya Univ, Mudanya, Turkiye

Açıklama

This paper presents a PropBank annotation project for Turkish, focusing on core arguments in matrix clauses. Our dataset comprises of five different corpora with 25,580 sentences labeled for verbal predicates and core arguments. Using a semi-automatic approach, we leveraged a dependency layer to pre-assign some ARG0s and ARG1s, followed by manual corrections. Our work will be used in the development of Abstract Meaning Representations (AMRs), enhancing Turkish NLP resources for semantic parsing and higher-level language tasks.(1)

Dosyalar

bib-c3a3ea91-346a-4ff3-9d7c-08e90bf45d6b.txt

Dosyalar (261 Bytes)

Ad Boyut Hepisini indir
md5:0479c25153eddce8c30c0ba7bbd18739
261 Bytes Ön İzleme İndir