How to optimize some algorithm in cuda
Open-source evaluation toolkit of large multi-modality models (LMMs)
General technology for enabling AI capabilities w/ LLMs and MLLMs
The first AI agent that builds permissionless integrations
A python module to repair invalid JSON from LLMs
Open-source model for program synthesis
Llama Chinese community, real-time aggregation
Large Language Model Principles and Practice Tutorial from Scratch
Qwen3-ASR is an open-source series of ASR models
Run LLM prompts from your shell
Spark-TTS Inference Code
A frontier, first-principles handbook
Fast-stable-diffusion + DreamBooth
A Pragmatic VLA Foundation Model
End-to-end pipeline converting generative videos
OpenTinker is an RL-as-a-Service infrastructure for foundation models
Motion-controllable Video Generation via Latent Trajectory Guidance
A tool to use the Ai2 Open Coding Agents Soft-Verified Agents
Habit Tracker for the AI Coding Workshop
Multimodal embedding and reranking models built on Qwen3-VL
Z80-μLM is a 2-bit quantized language model
The knowledge and task management backbone for AI coding assistants
Build a machine learning model from a prompt
LLM training in simple, raw C/CUDA
Implementation of "MobileCLIP" CVPR 2024