SGLang is a fast serving framework for large language models
Scalable and user friendly neural forecasting algorithms.
Trainable, memory-efficient, and GPU-friendly PyTorch reproduction
End-to-End Library for Continual Learning based on PyTorch
A smarter, self-hosted AI assistant — multi-user, multi-agent
A unified library of SOTA model optimization techniques
A Multi-Modal World Model for Reconstructing, Generating, Simulation
Gracefully face hCaptcha challenge with multimodal llms
Reproduction of Poetiq's record-breaking submission to the ARC-AGI-1
Provides convenient access to the Anthropic REST API from any Python 3
Detecting silent model failure. NannyML estimates performance
A Web UI for easy subtitle using whisper model
Open-source framework for intelligent speech interaction
Large Multimodal Models for Video Understanding and Editing
An Open Source text-to-speech system built by inverting Whisper
Speech-AI-Forge is a project developed around TTS generation model
OCR expert VLM powered by Hunyuan's native multimodal architecture
RGBD video generation model conditioned on camera input
SOTA on-device LLMs, small yet powerful
Requirement-driven evaluation harness for AI agents and LLM
TokenSpeed is a speed-of-light LLM inference engine
Kaggle Python docker image
A 0.1B Omni model trained from scratch
Build a modern LLM from scratch. Every line commented
The AI Assistant that actually does things for the trades