Efficient Retrieval Augmentation and Generation Framework
Powerful AI language model (MoE) optimized for efficiency/performance
Browse the web, directly from Cursor etc.
The easiest, and fastest way to run AI-generated Python code safely
Large Language Model Text Generation Inference
A guidance language for controlling large language models
Framework for validating and controlling LLM outputs in AI apps
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
Context database designed specifically for AI Agents
ReFT: Representation Finetuning for Language Models
Hackable and optimized Transformers building blocks
tiktoken is a fast BPE tokeniser for use with OpenAI's models
Python & JS/TS SDK for running AI-generated code/code
Feature engineering package with sklearn like functionality
Centralized agent control plane for governing runtime agent behavior
Parse files for optimal RAG
A Web UI for easy subtitle using whisper model
Gemma open-weight LLM library, from Google DeepMind
Make websites accessible for AI agents
Implements the Zettelkasten knowledge management methodology
Lemonade helps users run local LLMs with the highest performance
Evaluation and Tracking for LLM Experiments
A fast and lightweight framework for creating decentralized agents
Adding guardrails to large language models
Python hands on tutorial with 50+ Python Application