Open source libraries and APIs to build custom preprocessing pipelines
WikiChat is an improved RAG
A high-quality PDF to Markdown tool based on large language model
A Heterogeneous Benchmark for Information Retrieval
A curated list of data mining papers about fraud detection
Chinese XLNet pre-trained model
AI-powered tool for generating, optimizing, and translating subtitles
Translate the video from one language to another and embed dubbing
Efficient Retrieval Augmentation and Generation Framework
LLM based data scientist, AI native data application
Neural Network Compression Framework for enhanced OpenVINO
Semantic search and workflows for medical/scientific papers
A system for agentic LLM-powered data processing and ETL
Toolkit for conversational AI
A natural language interface for computers
Efficient few-shot learning with Sentence Transformers
File Parser optimised for LLM Ingestion with no loss
Public opinion analysis system
TextWorld is a sandbox learning environment for the training
Advanced NLP with spaCy: A free online course
A Repo For Document AI
AI tool for automating desktop tasks via natural language input
Zero-code platform for building AI agents from natural language input
Composable building blocks to build Llama Apps
Document (PDF, Word, PPTX ...) extraction and parse API