Dawarich is a command-line tool (likely Ruby-based) for transforming and analyzing Arabic text data with normalization, diacritic handling, segmentation, and morphological tokenization. Designed for text mining and NLP workflows in Arabic-language contexts.

Features

  • Normalizes Arabic script variants and punctuation
  • Removes or processes diacritics for text standardization
  • Tokenization and segmentation suited to Arabic morphology
  • Supports stop word removal and light stemming
  • Command‑line interface for batch NLP preprocessing
  • Output formats compatibility: plain text, CSV/JSON

Project Samples

Project Activity

See All Activity >

Categories

Mapping

License

Affero GNU Public License

Follow Dawarich

Dawarich Web Site

Other Useful Business Software
AestheticsPro Medical Spa Software Icon
AestheticsPro Medical Spa Software

Our new software release will dramatically improve your medspa business performance while enhancing the customer experience

AestheticsPro is the most complete Aesthetics Software on the market today. HIPAA Cloud Compliant with electronic charting, integrated POS, targeted marketing and results driven reporting; AestheticsPro delivers the tools you need to manage your medical spa business. It is our mission To Provide an All-in-One Cutting Edge Software to the Aesthetics Industry.
Learn More
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of Dawarich!

Additional Project Details

Operating Systems

Linux, Mac, Windows

Programming Language

Ruby

Related Categories

Ruby Mapping Software

Registered

2025-07-31