Dolgoročni cilj projekta Obeliks je izdelava in nadgrajevanje najbolj natančnega statističnega označevalnika za slovenski jezik. Oblikoskladenjsko označevanje je proces pripisovanja oblikoslovnih (in deloma skladenjskih) lastnosti besedam v poljubnem besedilu. Tako označeno besedilo je predpogoj za delovanje večine aplikacij, ki temeljijo na analizi naravnega jezika. Označevanje slovenskih besedil je zelo težak problem, saj mora algoritem za označevanje pravilno izbirati med skoraj dva tisoč oznakami (število različnih oznak za označevanje angleškega besedila je zgolj okoli šestdeset). Izvorna koda je na GitHub-u (glej Wiki). // The aim of the Obeliks project is to develop the most accurate statistical tagger for the Slovene language. Morphosyntactic tagging is the process of categorizing a word in a text into a particular part of speech category and describing it with various morphological features related to that category. The source code is on GitHub (see Wiki).

Project Activity

See All Activity >

Categories

Linguistics

License

MIT License

Follow Obeliks

Obeliks Web Site

Other Useful Business Software
Total Network Visibility for Network Engineers and IT Managers Icon
Total Network Visibility for Network Engineers and IT Managers

Network monitoring and troubleshooting is hard. TotalView makes it easy.

This means every device on your network, and every interface on every device is automatically analyzed for performance, errors, QoS, and configuration.
Learn More
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of Obeliks!

Additional Project Details

Operating Systems

Windows

Intended Audience

Advanced End Users, Developers, End Users/Desktop, Information Technology, Science/Research

User Interface

.NET/Mono, Command-line, Web-based

Programming Language

C#

Related Categories

C# Linguistics Software

Registered

2012-05-05