Run frontier MoE models on hardware you already own
Why use many token when few token do trick
C#/.NET binding of llama.cpp, including LLaMa/GPT model inference
Core building blocks for AI apps
Repository for skills to assist AI coding agents with .NET and C#
Project-scoped Lean workflow orchestrator from Math, Inc.
Agent Skills for optimizing web quality based on Lighthouse
Drag & drop UI to build your customized LLM flow
An elegent pytorch implement of transformers
Building a Secure and Interoperable Future for AI-Driven Payments
Flux 2 image generation model pure C inference
Kubernetes native framework for building AI agents
AI-Powered Personalized Learning Assistant
Tiny vision language model
Containerized automation engine for programmable CI/CD workflows
Agentic IM Chatbot infrastructure
Python SDK for Claude Agent
AI-Powered Data Processing: Use LOTUS to process all of your datasets
Deep universal probabilistic programming with Python and PyTorch
The AI toolkit for the AI developer
The agent IDE that builds itself
Cloud-native runtime for agentic AI
Build resilient language agents as graphs
Hundreds of fully solved job interview questions
Z80-μLM is a 2-bit quantized language model