PADIC (Parallel Arabic DIalectal Corpus) is a multi-dialectal corpus built in the framework of the National Research Project "TORJMAN", led by Scientific and Technical Research Center for the Development of Arabic Language and funded by the Algerian Ministry of Higher Education and Scientific Research.
PADIC is composed of 6 dialects: two Algerian dialects (Algiers and Annaba cities), Palestinian, Syrian, Tunisian, Moroccan) and MSA.

Mourad Abbas
Computational Linguistics Department, crstdla
https://sites.google.com/site/mouradabbas9

Publications
-----------------
K. Meftouh, S. Harrat, S. Jamoussi, M. Abbas, K. Smaïli, Machine Translation Experiments on PADIC: A Parallel Arabic DIalect Corpus, The 29th Pacific Asia Conference on Language, Information and Computation, PACLIC 2015, Shanghai, 2015.

TORJMAN website:
-------------------------
https://sites.google.com/site/torjmanepnr/6-corpus

Features

  • XML
  • Buckwalter
  • 5 Arabic dialects + Modern Standard Arabic
  • More than 6000 sentences

Project Activity

See All Activity >

License

GNU General Public License version 3.0 (GPLv3)

Follow PADIC

PADIC Web Site

Other Useful Business Software
Auth0 B2B Essentials: SSO, MFA, and RBAC Built In Icon
Auth0 B2B Essentials: SSO, MFA, and RBAC Built In

Unlimited organizations, 3 enterprise SSO connections, role-based access control, and pro MFA included. Dev and prod tenants out of the box.

Auth0's B2B Essentials plan gives you everything you need to ship secure multi-tenant apps. Unlimited orgs, enterprise SSO, RBAC, audit log streaming, and higher auth and API limits included. Add on M2M tokens, enterprise MFA, or additional SSO connections as you scale.
Sign Up Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of PADIC!

Additional Project Details

Operating Systems

Android, Linux, Windows

Languages

Arabic

Intended Audience

Science/Research

Registered

2017-03-24