Standalone. Small. Language-neutral. BudouX is the successor to Budou, the machine learning-powered line break organizer tool. It is standalone. It works with no dependency on third-party word segmenters such as Google cloud natural language API. It is small. It takes only around 15 KB including its machine learning model. It's reasonable to use it even on the client-side. It is language-neutral. You can train a model for any language by feeding a dataset to BudouX’s training script.
Features
- BudouX supports HTML inputs
- Documentation available
- Examples available
- You can get a list of phrases by feeding a sentence to the parser
- BudouX supports HTML inputs and outputs HTML strings
- BudouX uses the AdaBoost algorithm to segment a sentence into phrases
Categories
Machine LearningLicense
Apache License V2.0Follow BudouX
Other Useful Business Software
Build Securely on Azure with Proven Frameworks
Moving to the cloud brings new challenges. How can you manage a larger attack surface while ensuring great network performance? Turn to Fortinet’s Tested Reference Architectures, blueprints for designing and securing cloud environments built by cybersecurity experts. Learn more and explore use cases in this white paper.
Rate This Project
Login To Rate This Project
User Reviews
Be the first to post a review of BudouX!