ExtractThinker is a Document Intelligence library for LLMs
DataDreamer: Prompt. Generate Synthetic Data. Train & Align Models
Fast and customizable framework for automatic ML model creation
A Heterogeneous Benchmark for Information Retrieval
The library to build & auto-optimize LLM applications
Data and tools for generating and inspecting OLMo pre-training data
Efficient Retrieval Augmentation and Generation Framework
A full spaCy pipeline and models for scientific/biomedical documents
Dealing with all unstructured data, such as reverse image search
Libraries for applying sparsification recipes to neural networks
Data processing for and with foundation models
A curated list of data mining papers about fraud detection
Neural Network Compression Framework for enhanced OpenVINO
Haystack is an open source NLP framework to interact with your data
Pretrained model hub for Keras 3
Efficient few-shot learning with Sentence Transformers
Toolkit for conversational AI
Module for automatic summarization of text documents and HTML pages
A Unified Library for Parameter-Efficient Learning
Easy-to-use and powerful NLP library with Awesome model zoo
Hub of ready-to-use datasets for ML models
Super Comprehensive Deep Learning Notes
Chinese XLNet pre-trained model
Training data (data labeling, annotation, workflow) for all data types
Build AI-powered semantic search applications