Language: Arabic - LINDAT/CLARIAH-CZ Catalog Search Results

51. IWPT 2021 Shared Task Data and System Outputs

Creator:: Zeman, Daniel, Bouma, Gosse, and Seddah, Djamé
Publisher:: Universal Dependencies Consortium
Type:: text and corpus
Subject:: treebank, dependency, syntax, enhanced universal dependencies, shared task, and parsing
Language:: Arabic, Bulgarian, Czech, Dutch, English, Estonian, Finnish, French, Italian, Latvian, Lithuanian, Polish, Russian, Slovak, Swedish, Tamil, and Ukrainian
Description:: This package contains data used in the IWPT 2021 shared task. It contains training, development and test (evaluation) datasets. The data is based on a subset of Universal Dependencies release 2.7 (http://hdl.handle.net/11234/1-3424) but some treebanks contain additional enhanced annotations. Moreover, not all of these additions became part of Universal Dependencies release 2.8 (http://hdl.handle.net/11234/1-3687), which makes the shared task data unique and worth a separate release to enable later comparison with new parsing algorithms. The package also contains a number of Perl and Python scripts that have been used to process the data during preparation and during the shared task. Finally, the package includes the official primary submission of each team participating in the shared task.
Rights:: Licence Universal Dependencies v2.7, https://lindat.mff.cuni.cz/repository/xmlui/page/license-ud-2.7, and PUB

52. JIRS

Publisher:: Grid and High Performance Computing Group, ITACA, Universidad Politécnica de Valencia and Universidad de Alicante
Type:: toolService
Language:: Arabic, English, French, Italian, Oromo, and Urdu
Description:: JIRS is a Passage Retrieval system specially suited for Question Answering. It could be adapted to others languages very easily. ask (Written Language): Information Retrieval Applications Question/Answering Environment: OS-independent Access: GPLv3
Rights:: Not specified

53. Křesťané v Arábii před islámem :

Creator:: Mikulicová, Mlada,
Type:: text and monografie
Subject:: Afroasijské (hamitosemitské) literatury (o nich), křesťanství, křesťané, poezie arabská, básníci, Arabové, světové dějiny středověku (do r. 1492), literatura, spisovatelé, and církve, sekty
Language:: Czech and Arabic
Rights:: unknown

54. Le compagnon de tous ou dictionnaire polyglotte pour les écoles, et pour ceux qui s'occupent de lnfues étrangères et aux Arabes qui étudient les Langues Occidentales: enrichi des termes nouveaux de sciences et arts , choisis ou approuvés dans une réunion de sceïkhs. par Louis Calligaris

Creator:: Calligaris, Luigi
Type:: model:monograph and TEXT
Language:: French, Arabic, Latin, Italian, Spanish, Portuguese, German, English, and Modern Greek (1453-)
Rights:: http://creativecommons.org/publicdomain/mark/1.0/ and policy:public

55. Lingua::Interset 2.026

Creator:: Zeman, Daniel
Publisher:: Charles University, Faculty of Mathematics and Physics
Type:: tool and toolService
Subject:: morphology, part of speech, conversion, and tagset
Language:: Arabic, Bulgarian, Bengali, Catalan, Czech, Danish, German, Modern Greek (1453-), English, Spanish, Estonian, Basque, Persian, Finnish, Ancient Greek (to 1453), Hebrew, Hindi, Croatian, Japanese, Multiple languages, and Portuguese
Description:: Lingua::Interset is a universal morphosyntactic feature set to which all tagsets of all corpora/languages can be mapped. Version 2.026 covers 37 different tagsets of 21 languages. Limited support of the older drivers for other languages (which are not included in this package but are available for download elsewhere) is also available; these will be fully ported to Interset 2 in future. Interset is implemented as Perl libraries. It is also available via CPAN.
Rights:: Artistic License (Perl) 1.0, http://opensource.org/licenses/Artistic-Perl-1.0, and PUB

56. LMF Arabic characters lexicon

Creator:: Namly, Driss
Publisher:: Ibtikarat team
Type:: text, lexicon, and lexicalConceptualResource
Subject:: alphabet
Language:: Arabic
Description:: An LMF conformant XML-based file containing all Arabic characters (letters, vowels and punctuations). Each character described with a description, different displays (isolated, at the beginning, middle and the end of a word), a codification (Unicode, others could be added later), and two transliterations (Buckwalter and wiki).
Rights:: Creative Commons - Attribution-NonCommercial 4.0 International (CC BY-NC 4.0), http://creativecommons.org/licenses/by-nc/4.0/, and PUB

57. LMF Contemporary Arabic dictionary

Creator:: Namly, Driss
Publisher:: Ibtikarat team
Type:: text, lexicon, and lexicalConceptualResource
Subject:: lexical semantics
Language:: Arabic
Description:: An LMF conformant XML-based file containing the electronic version of al logha al arabia al moassira (Contemporary Arabic) dictionary. An Arabic monolingual dictionary accomplished by Ahmed Mukhtar Abdul Hamid Omar (deceased: 1424) with the help of a working group
Rights:: Creative Commons - Attribution-NonCommercial 4.0 International (CC BY-NC 4.0), http://creativecommons.org/licenses/by-nc/4.0/, and PUB

58. MADED

Creator:: ridouane, tachicart and karim, bouzoubaa
Publisher:: ALELM
Type:: text, computationalLexicon, and lexicalConceptualResource
Subject:: lexicon and moroccan arabic
Language:: Moroccan Arabic and Arabic
Description:: Moroccan Dialect Electronic Dictionary (MDED) is an electronic lexicon containing almost 15000 MSA entries. They are written in Arabic letters and translated to Moroccan Arabic dialect. In addition, MDED entries are annotated useful metadata such as POS, Origin and root. MDED can be useful in some advanced NLP applications such as Machine translation and morphological analyzer.
Rights:: Creative Commons - Attribution-NonCommercial 4.0 International (CC BY-NC 4.0), http://creativecommons.org/licenses/by-nc/4.0/, and PUB

59. Manual Arabic spelling-errors correction for collected documents

Creator:: Saty, Ahmed, Aouragh, Si Lhoussain, and Bouzoubaa, Karim
Publisher:: Sudan University of Science and Technology
Type:: text and corpus
Subject:: Manual Arabic spelling-errors correction for collected documents
Language:: English and Arabic
Description:: The file represents a text corpus in the context of Arabic spell checking, where a group of persons edited different files, and all of the committed spelling errors by these persons have been recorded. A comprehensive representation these persons’ profile has been considered: male, female, old-aged, middle-aged, young-aged, high and low computer usage users, etc. Through this work, we aim to help researchers and those interested in Arabic NLP by providing them with an Arabic spell check corpus ready and open to exploitation and interpretation. This study also enabled the inventory of most spelling mistakes made by editors of Arabic texts. This file contains the following sections (tags): people – documents they printed – types of possible errors – errors they made. Each section (tag) contains some data that explains its details and its content, which helps researchers extracting research-oriented results. The people section contains basic information about each person and its relationship of using the computer, while the documents section clarifies all sentences in each document with the numbering of each sentence to be used in the errors section that was committed. We are also adding the “type of errors” section in which we list all the possible errors with their description in the Arabic language and give an illustrative example.
Rights:: Creative Commons - Attribution-ShareAlike 4.0 International (CC BY-SA 4.0), http://creativecommons.org/licenses/by-sa/4.0/, and PUB

60. NAFIS Arabic Stemming Gold Standard Corpus

Creator:: Namly, Driss
Publisher:: Ibtikarat team
Type:: text, wordList, and lexicalConceptualResource
Subject:: corpus, stemming;, and Gold Standard Corpus
Language:: Arabic
Description:: Normalized Arabic Fragments for Inestimable Stemming (NAFIS) is an Arabic stemming gold standard corpus composed by a collection of texts, selected to be representative of Arabic stemming tasks and manually annotated.
Rights:: Creative Commons - Attribution-NonCommercial 4.0 International (CC BY-NC 4.0), http://creativecommons.org/licenses/by-nc/4.0/, and PUB

51. IWPT 2021 Shared Task Data and System Outputs

52. JIRS

53. Křesťané v Arábii před islámem :

54. Le compagnon de tous ou dictionnaire polyglotte pour les écoles, et pour ceux qui s'occupent de lnfues étrangères et aux Arabes qui étudient les Langues Occidentales: enrichi des termes nouveaux de sciences et arts , choisis ou approuvés dans une réunion de sceïkhs. par Louis Calligaris

55. Lingua::Interset 2.026

56. LMF Arabic characters lexicon

57. LMF Contemporary Arabic dictionary

58. MADED

59. Manual Arabic spelling-errors correction for collected documents

60. NAFIS Arabic Stemming Gold Standard Corpus

Limit your search

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Search

Search Constraints

Search Results

Limit your search

Contributor

Show values starting with

Coverage

Creator

Show values starting with

Format

Language

Show values starting with

Publisher

Show values starting with

Rights

Show values starting with

Subject

Show values starting with

Type

Show values starting with

Date

Original context has metadata only

Harvested from