EnglishFrenchSpanish

Ad


OnWorks favicon

Arabic Corpus download for Linux

Free download Arabic Corpus Linux app to run online in Ubuntu online, Fedora online or Debian online

This is the Linux app named Arabic Corpus whose latest release can be downloaded as Khaleej-2004-utf8.zip. It can be run online in the free hosting provider OnWorks for workstations.

Download and run online this app named Arabic Corpus with OnWorks for free.

Follow these instructions in order to run this app:

- 1. Downloaded this application in your PC.

- 2. Enter in our file manager https://www.onworks.net/myfiles.php?username=XXXXX with the username that you want.

- 3. Upload this application in such filemanager.

- 4. Start the OnWorks Linux online or Windows online emulator or MACOS online emulator from this website.

- 5. From the OnWorks Linux OS you have just started, goto our file manager https://www.onworks.net/myfiles.php?username=XXXXX with the username that you want.

- 6. Download the application, install it and run it.

Arabic Corpus


Ad


DESCRIPTION

The Arabic Corpus {compiled by Dr. Mourad Abbas ( http://sites.google.com/site/mouradabbas9/corpora ) The corpus Khaleej-2004 contains 5690 documents. It is divided to 4 topics (categories). The corpus Watan-2004 contains 20291 documents organized in 6 topics (categories). Researchers who use these two corpora would mention the two main references:
(1) For Watan-2004 corpus
----------------------
M. Abbas, K. Smaili, D. Berkani, (2011) Evaluation of Topic Identification Methods on Arabic Corpora,JOURNAL OF DIGITAL INFORMATION MANAGEMENT,vol. 9, N. 5, pp.185-192.

2) For Khaleej-2004 corpus
---------------------------------
M. Abbas, K. Smaili (2005) Comparison of Topic Identification Methods for Arabic Language, RANLP05 : Recent Advances in Natural Language Processing ,pp. 14-17, 21-23 september 2005, Borovets, Bulgary.

More useful references to check:
-------------------------------------------
https://sites.google.com/site/mouradabbas9/corpora



Audience

Information Technology, Science/Research, Advanced End Users, Developers, Quality Engineers, Engineering


User interface

Win32 (MS Windows), KDE


Programming Language

Python, C++, JavaScript


Database Environment

MySQL



Categories

Machine Translation, Machine Learning

This is an application that can also be fetched from https://sourceforge.net/projects/arabiccorpus/. It has been hosted in OnWorks in order to be run online in an easiest way from one of our free Operative Systems.


Free Servers & Workstations

Download Windows & Linux apps

Linux commands

Ad