Standardized Serverless ML Inference Platform on Kubernetes
A course of learning LLM inference serving on Apple Silicon
Neural Network architecture based on ideas of the original LSTM
Open speech-to-speech models and pipelines by Hugging Face toolkit AI
AI Powered Knowledge Graph Generator
An MCP server for interacting with Google Colab
Chat with it via text and voice
Deploy reasoning AI agents powered by agentic graph RAG in minutes
A simple, performant and scalable Jax LLM
Low-latency AI inference engine optimized for mobile devices
Ultralytics YOLO
Superduper: Integrate AI models and machine learning workflows
A high-performance ML model serving framework, offers dynamic batching
The Modular Platform (includes MAX & Mojo)
Open source AI Agents hosted on the oTTomator Live Agent Studio
A library for deep learning end-to-end dialog systems and chatbots
Gracefully face hCaptcha challenge with multimodal llms
AI agents running research on single-GPU nanochat training
AI agents running research on single-GPU nanochat training
Z80-μLM is a 2-bit quantized language model
Tensor search for humans
Examples of using E2B
Workshop-Level Automated Scientific Discovery via Agentic Tree Search
Definitions for AI/ML tasks like dataset creation
Data Infrastructure providing an approach to multimodal AI workloads