A curated list of data mining papers about fraud detection
ExtractThinker is a Document Intelligence library for LLMs
Data and tools for generating and inspecting OLMo pre-training data
Industrial-strength Natural Language Processing (NLP)
The Classical Language Toolkit
Uma Ferramenta Computacional para Análise e Recuperação de Patentes
Resources, corpora, and tools for Chinese natural language processing
State-of-the-art Multilingual Question Answering research
An open-source NLP research library, built on PyTorch
Explain, analyze, and visualize NLP language models
A toolkit for managing and manipulating text annotations
Chinese synonyms, chat robot, intelligent question and answer toolkit
A model library for exploring state-of-the-art deep learning
NLP made easy
Data repository for pretrained NLP models and NLP corpora
AiLearning, data analysis plus machine learning practice
We describe a simple XML format to share text documents and annotation