Join/Login
Business Software
Open Source Software
For Vendors
Blog
About
More

For Vendors Help Create Join Login

Business Software

Open Source Software

SourceForge Podcast

Resources

Articles
Case Studies
Blog

Menu

Help
Create
Join
Login

Home
Open Source Software
Artificial Intelligence Software
Search Results

Search Results for "text classification" - Page 4

x

Sort By:

Relevance

Clear All Filters

OS

Linux 108
Mac 107
Windows 106
More...
BSD 46
ChromeOS 45
Desktop Operating Systems 2

Category

Artificial Intelligence 108
Scientific/Engineering 8
Software Development 6
Business 5
Text Editors 5
System 2
Education 1
Internet 1

License

OSI-Approved Open Source 84
Creative Commons Attribution License 3
GNU Free Documentation License 1

Translations

English 7
French 1
Russian 1
Tamil 1

Programming Language

Python 58
Java 13
C++ 7
C 4
More...
JavaScript 4
Go 3
Rust 2
Groovy 1
Perl 1
PHP 1
S/R 1
Scala 1

Status

Production/Stable 9
Beta 4
Alpha 3
Mature 1

Showing 108 open source projects for "text classification"

View related business solutions

Artificial Intelligence Linux Clear Filters & Widen Search

Our Free Plans just got better! | Auth0
With up to 25k MAUs and unlimited Okta connections, our Free Plan lets you focus on what you do best—building great apps.

You asked, we delivered! Auth0 is excited to expand our Free and Paid plans to include more options so you can focus on building, deploying, and scaling applications without having to worry about your security. Auth0 now, thank yourself later.

Try free now
MongoDB Atlas runs apps anywhere
Deploy in 115+ regions with the modern database for every enterprise.

MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.

Start Free
1

fastText

Library for fast text classification and representation

...The goal of text classification is to assign documents (such as emails, posts, text messages, product reviews, etc...) to one or multiple categories. Such categories can be review scores, spam v.s. non-spam, or the language in which the document was typed. Nowadays, the dominant approach to build such classifiers is machine learning, that is learning classification rules from examples.

Downloads: 0 This Week

Last Update: 2023-10-10
See Project
2

Computer Vision Pretrained Models

A collection of computer vision pre-trained models

A pre-trained model is a model created by someone else to solve a similar problem. Instead of building a model from scratch to solve a similar problem, we can use the model trained on other problem as a starting point. A pre-trained model may not be 100% accurate in your application. For example, if you want to build a self-learning car. You can spend years building a decent image recognition algorithm from scratch or you can take the inception model (a pre-trained model) from Google which...

Downloads: 0 This Week

Last Update: 2022-08-18
See Project
3

NLP Best Practices

Natural Language Processing Best Practices & Examples

In recent years, natural language processing (NLP) has seen quick growth in quality and usability, and this has helped to drive business adoption of artificial intelligence (AI) solutions. In the last few years, researchers have been applying newer deep learning methods to NLP. Data scientists started moving from traditional methods to state-of-the-art (SOTA) deep neural network (DNN) algorithms which use language models pretrained on large text corpora. This repository contains examples and...

Downloads: 0 This Week

Last Update: 2022-08-01
See Project
4

StarSpace

Learning embeddings for classification, retrieval and ranking

...The library supports a variety of tasks (text classification, nearest-neighbor search, recommendation, entity linking) with simple configuration. It includes efficient batching, negative sampling strategies, and on-the-fly embedding updates.

Downloads: 0 This Week

Last Update: 2025-10-07
See Project
Secure File Transfer for Windows with Cerberus by Redwood
Protect and share files over FTP/S, SFTP, HTTPS and SCP with the #1 rated Windows file transfer server.

Cerberus supports unlimited users and connections on a single IP, with built-in encryption, 2FA, and a browser-based web client — all deployable in under 15 minutes with a 25-day free trial.

Try for Free
5

Texar

Toolkit for Machine Learning, Natural Language Processing

...Texar-TensorFlow (this repo) and Texar-PyTorch have mostly the same interfaces. Both further combine the best design of TF and PyTorch. Rich Pre-trained Models, Rich Usage with Uniform Interfaces. BERT, GPT2, XLNet, etc, for encoding, classification, generation, and composing complex models with other Texar components!

Downloads: 0 This Week

Last Update: 2022-08-08
See Project
6

ELI5

A library for debugging/inspecting machine learning classifiers

...The library allows users to inspect model weights, analyze decision trees, and compute permutation feature importance for black-box models. It also provides specialized tools such as TextExplainer, which can highlight important words in text classification tasks to explain why a model produced a particular prediction. Additionally, the library integrates explanation algorithms such as LIME to interpret predictions from arbitrary machine learning models.

Downloads: 0 This Week

Last Update: 2026-03-15
See Project
7

Convolutional Recurrent Neural Network

Convolutional Recurrent Neural Network (CRNN) for image-based sequence

...The implementation also integrates the Connectionist Temporal Classification (CTC) loss function, enabling end-to-end training of the model using labeled sequence data. CRNN has been widely used in computer vision tasks that require interpreting text embedded in images, such as reading street signs, documents, or natural scene text.

Downloads: 0 This Week

Last Update: 2026-03-12
See Project
8

JCLTP

A Java Class Library for Text Processing

JCLTP is a class library designed for processing text. JCLTP is free, open source and developed with the Java programming language. JCLTP is distributed under the GNU license. It incorporates several technologies that enable process information while applying AI techniques, in order to build predictive models for text classification. Through a flexible structure of interfaces and classes, the opportunity to extend, adapt and add functionality JCLTP is provided. ...

Downloads: 0 This Week

Last Update: 2017-01-06
See Project
9

sgmweka

Weka wrapper for the SGM toolkit for text classification and modeling.

Weka wrapper for the SGM toolkit for text classification and modeling. Provides Sparse Generative Models for scalable and accurate text classification and modeling for use in high-speed and large-scale text mining. Has lower time complexity of classification than comparable software due to inference based on sparse model representation and use of an inverted index. The provided .zip file is in the Weka package format, giving access to text classification. ...

Downloads: 11 This Week

Last Update: 2016-06-23
See Project
Stop Cyber Threats with VM-Series Next-Gen Firewall on Azure
Native application identity and user-based security for your Azure cloud

Gain integrated visibility across all traffic in a single pass. Deploy Palo Alto Networks VM-Series to determine application identity and content while automating security policy updates via rich APIs.

Get a free trial
10

Speechalyzer

Process large speech data wrt transcription, labeling and annotation

Speechalyzer: a tool for the daily work of a 'speech worker' It is optimized to process large speech data sets with respect to transcription, labeling and annotation. It is implemented as a client server based framework in Java and interfaces software for speech recognition, synthesis, speech classification and quality evaluation. The application is mainly the processing of training data for speech recognition and classification models and performing benchmarking tests on speech-to-text, text-to-speech and speech classification software systems.

Downloads: 2 This Week

Last Update: 2016-04-27
See Project
11

Java Data Mining Package

The Java Data Mining Package (JDMP) is a library that provides methods for analyzing data with the help of machine learning algorithms (e.g. clustering, classification, graphical models, neural networks, Bayesian networks, text processing, optimization).

Downloads: 0 This Week

Last Update: 2015-08-19
See Project
12

JInsect

The JINSECT toolkit is a Java-based toolkit and library that supports and demonstrates the use of n-gram graphs within Natural Language Processing applications, ranging from summarization and summary evaluation to text classiﬁcation and indexing.

3 Reviews

Downloads: 0 This Week

Last Update: 2015-08-25
See Project
13

VADER

Lexicon and rule-based sentiment analysis tool

VADER (Valence Aware Dictionary and sEntiment Reasoner) is a lexicon and rule-based sentiment analysis tool designed for analyzing the sentiment of text, particularly in social media and short text formats. It is optimized for quick and accurate analysis of positive, negative, and neutral sentiments.

Downloads: 18 This Week

Last Update: 2025-02-20
See Project
14

Persica-A new Persian corpus for NLP

This project presents a new corpus for NEWS text analysis in Persian

Lack of multi-application text corpus despite of the surging text data is a serious bottleneck in the text mining and natural language processing especially in Persian language. This project presents a new corpus for NEWS articles analysis in Persian called Persica. NEWS analysis includes NEWS classification, topic discovery and classification, category classification and many more procedures.

Downloads: 6 This Week

Last Update: 2014-08-31
See Project
15

TextBlob

TextBlob is a Python library for processing textual data

Simple, Pythonic, text processing, Sentiment analysis, part-of-speech tagging, noun phrase extraction, translation, and more. It provides a simple API for diving into common natural language processing (NLP) tasks such as part-of-speech tagging, noun phrase extraction, sentiment analysis, classification, translation, and more. TextBlob stands on the giant shoulders of NLTK and pattern, and plays nicely with both.

Downloads: 0 This Week

Last Update: 2021-07-23
See Project
16

TextProcessor

A Java package to preprocess text datasets for posterior text analysis

The TextProcessor Java package is a text processing toolkit, which provides some frequently used text processing functions such as stemming, removing stop-words, generating a term vocabulary, and calculating the term-doc frequency matrix. Basic topic mining models such as LDA and sparse NMF are also supported. The package can also generate feature files from a given text dataset with LDA and LIBSVM format for posterior procedures such as classification or clustering. ...

Downloads: 0 This Week

Last Update: 2015-11-23
See Project
17

Modular Suite of NLP Tools

This project aims to build a suite of Natural Language Processing tools. Modules will include corpus indexing and access tools, a part-of-speech tagger, tokenisers, text classification software, etc.

Downloads: 0 This Week

Last Update: 2014-06-09
See Project
18

text-analysis

This project aims to implement in java the following text mining techniques: Text Language Detection, Keywords and keyphrases extraction, Text Classification, Text Clustering, Single or multiple documents Summarization, Plagiarism Detection.

Downloads: 0 This Week

Last Update: 2014-05-20
See Project
19

iracema

An information extraction library implementing modern algorithms for the extraction of named entities from text.

Downloads: 0 This Week

Last Update: 2013-04-19
See Project
20

Trainable Relation Extraction framework

T-Rex (Trainable Relation Extraction) is a highly configurable machine learning-based Information Extraction from Text framework, which includes tools for document classification, entity extraction and relation extraction.

Downloads: 0 This Week

Last Update: 2013-05-02
See Project
21

Word Vector Tool

The Word Vector Tool is a simple but flexible Java library to create word vector representations of text documents. Word vectors can be used for various text processing tasks, as text classification, text clustering or information retrieval.

Downloads: 0 This Week

Last Update: 2013-04-08
See Project
22

Young Researchers' Induction Foundation

Collection of Statistical Language Processing Tools and Modules for Information Retrieval, Document Classification, Vectorization, Pattern Matching, Knowledge/Text Mining related problems.

Downloads: 0 This Week

Last Update: 2013-04-17
See Project
23

pySPACE

Signal Processing and Classification Environment in Python using YAML

pySPACE is a modular software for processing of large data streams that has been specifically designed to enable distributed execution and empirical evaluation of signal processing chains. Various signal processing algorithms (so called nodes) are available within the software, from finite impulse response filters over data-dependent spatial filters (e.g. CSP, xDAWN) to established classifiers (e.g. SVM, LDA). pySPACE incorporates the concept of node and node chains of the MDP framework. Due...

Downloads: 0 This Week

Last Update: 2014-10-29
See Project
24

t5-base

Flexible text-to-text transformer model for multilingual NLP tasks

t5-base is a pre-trained transformer model from Google’s T5 (Text-To-Text Transfer Transformer) family that reframes all NLP tasks into a unified text-to-text format. With 220 million parameters, it can handle a wide range of tasks, including translation, summarization, question answering, and classification. Unlike traditional models like BERT, which output class labels or spans, T5 always generates text outputs.

Downloads: 0 This Week

Last Update: 2025-07-02
See Project
25

t5-small

T5-Small: Lightweight text-to-text transformer for NLP tasks

T5-Small is a lightweight variant of the Text-To-Text Transfer Transformer (T5), designed to handle a wide range of NLP tasks using a unified text-to-text approach. Developed by researchers at Google, this model reframes all tasks—such as translation, summarization, classification, and question answering—into the format of input and output as plain text strings. With only 60 million parameters, T5-Small is compact and suitable for fast inference or deployment in constrained environments. ...

Downloads: 0 This Week

Last Update: 2025-07-02
See Project

Previous
1
2
3
You're on page 4
5
Next

Related Searches

sentiment analysis

computer vision

predictive text

sgm

transcription

algorithms

text summarization

rapidminer persian text mining

document term matrix in java

nlp

Related Categories

Artificial Intelligence

Scientific/Engineering

Software Development

Business

Text Editors

SourceForge

Create a Project
Open Source Software
Business Software
Top Downloaded Projects

Company

About
Team
SourceForge Headquarters
1320 Columbia Street Suite 310
San Diego, CA 92101
+1 (858) 422-6466

Resources

Support
Site Documentation
Site Status
SourceForge Reviews

© 2026 Slashdot Media. All Rights Reserved.

Terms Privacy Opt Out Advertise