تخطي إلى التنقل الرئيسي تخطي إلى البحث تخطي إلى المحتوى الرئيسي

UniMorph 4.0: Universal Morphology

  • Khuyagbaatar Batsuren
  • , Omer Goldman
  • , Salam Khalifa
  • , Nizar Habash
  • , Witold Kieraś
  • , Gábor Bella
  • , Brian Leonard
  • , Garrett Nicolai
  • , Kyle Gorman
  • , Yustinus Ghanggo Ate
  • , Maria Ryskina
  • , Sabrina Mielke
  • , Elena Budianskaya
  • , Charbel El-Khaissi
  • , Tiago Pimentel
  • , Michael Gasser
  • , William Lane
  • , Mohit Raj
  • , Matt Coler
  • , Jaime Rafael Montoya Samame
  • Delio Siticonatzi Camaiteri, Esaú Zumaeta Rojas, Didier L. Francis, Arturo Oncevay, Juan L. Bautista, Gema Celeste Silva Villegas, Lucas Torroba Hennigen, Adam Ek, David Guriel, Peter Dirix, Jean Philippe Bernardy, Andrey Scherbakov, Aziyana Bayyr-Ool, Antonios Anastasopoulos, Roberto Zariquiey, Karina Sheifer, Sofya Ganieva, Hilaria Cruz, Ritván Karahóǧa, Stella Markantonatou, George Pavlidis, Matvey Plugaryov, Elena Klyachko, Ali Salehi, Candy Angulo, Jatayu Baxi, Andrew Krizhanovsky, Natalia Krizhanovsky, Elizabeth Salesky, Clara Vania, Sardana Ivanova, Jennifer White, Rowan Hall Maudslay, Josef Valvoda, Ran Zmigrod, Paula Czarnowska, Irene Nikkarinen, Aelita Salchak, Brijesh Bhatt, Christopher Straughn, Zoey Liu, Jonathan North Washington, Yuval Pinter, Duygu Ataman, Marcin Woliński, Totok Suhardijanto, Anna Yablonskaya, Niklas Stoehr, Hossep Dolatian, Zahroh Nuriah, Shyam Ratan, Francis M. Tyers, Edoardo M. Ponti, Grant Aiton, Aryaman Arora, Richard J. Hatcher, Ritesh Kumar, Jeremiah Young, Daria Rodionova, Anastasia Yemelina, Taras Andrushko, Igor Marchenko, Polina Mashkovtseva, Alexandra Serova, Emily Prud'Hommeaux, Maria Nepomniashchaya, Fausto Giunchiglia, Eleanor Chodroff, Mans Hulden, Miikka Silfverberg, Arya D. McCarthy, David Yarowsky, Ryan Cotterell, Reut Tsarfaty, Ekaterina Vylomova

نتاج البحث: فصل من :كتاب / تقرير / مؤتمرمنشور من مؤتمرمراجعة النظراء

ملخص

The Universal Morphology (UniMorph) project is a collaborative effort providing broad-coverage instantiated normalized morphological inflection tables for hundreds of diverse world languages. The project comprises two major thrusts: a language-independent feature schema for rich morphological annotation and a type-level resource of annotated data in diverse languages realizing that schema. This paper presents the expansions and improvements made on several fronts over the last couple of years (since McCarthy et al. (2020)). Collaborative efforts by numerous linguists have added 67 new languages, including 30 endangered languages. We have implemented several improvements to the extraction pipeline to tackle some issues, e.g. missing gender and macron information. We have also amended the schema to use a hierarchical structure that is needed for morphological phenomena like multiple-argument agreement and case stacking, while adding some missing morphological features to make the schema more inclusive. In light of the last UniMorph release, we also augmented the database with morpheme segmentation for 16 languages. Lastly, this new release makes a push towards inclusion of derivational morphology in UniMorph by enriching the data and annotation schema with instances representing derivational processes from MorphyNet.

اللغة الأصليةالإنجليزيّة
عنوان منشور المضيف2022 Language Resources and Evaluation Conference, LREC 2022
المحررونNicoletta Calzolari, Frederic Bechet, Philippe Blache, Khalid Choukri, Christopher Cieri, Thierry Declerck, Sara Goggi, Hitoshi Isahara, Bente Maegaard, Joseph Mariani, Helene Mazo, Jan Odijk, Stelios Piperidis
ناشرEuropean Language Resources Association (ELRA)
الصفحات840-855
عدد الصفحات16
رقم المعيار الدولي للكتب (الإلكتروني)9791095546726
حالة النشرنُشِر - 2022
منشور خارجيًانعم
الحدث13th International Conference on Language Resources and Evaluation Conference, LREC 2022 - Marseille, فرنسا
المدة: ٢٠ يونيو ٢٠٢٢٢٥ يونيو ٢٠٢٢

سلسلة المنشورات

الاسم2022 Language Resources and Evaluation Conference, LREC 2022

!!Conference

!!Conference13th International Conference on Language Resources and Evaluation Conference, LREC 2022
الدولة/الإقليمفرنسا
المدينةMarseille
المدة٢٠/٠٦/٢٢٢٥/٠٦/٢٢

ملاحظة ببليوغرافية

Publisher Copyright:
© European Language Resources Association (ELRA), licensed under CC-BY-NC-4.0.

بصمة

أدرس بدقة موضوعات البحث “UniMorph 4.0: Universal Morphology'. فهما يشكلان معًا بصمة فريدة.

قم بذكر هذا