TokenSpeed is a speed-of-light LLM inference engine
Browser userscript that enhances ChatGPT reliability and usability
NeurIPS2025 Spotlight] Quantized Attention
Llama Chinese community, real-time aggregation
Lightning-Fast RL for LLM Reasoning and Agents. Made Simple & Flexible
NVIDIA Isaac Sim is an open-source application on NVIDIA Omniverse
Collaborative & Open-Source Quality Assurance for all AI models
AI discovers 520000 stable inorganic crystal structures for research
A vector index built on TurboQuant, written in Rust with Python
On-device TTS model by Neuphonic
20+ high-performance LLMs with recipes to pretrain, finetune at scale
Extract schema, statistics and entities from datasets
Kubernetes observability and automation
Create beautiful slides on the web using Claude's frontend skills
Framework for validating and controlling LLM outputs in AI apps
Running large language models on a single GPU
Code and models for ICML 2024 paper, NExT-GPT
A Next-Generation Training Engine Built for Ultra-Large MoE Models
Self-supervised visual learning using momentum contrast in PyTorch
Open-source Agent OS for hardware intelligence
Skills for threat modeling, scanning, triage, patching, etc.
A Python package for extending the official PyTorch
Operating LLMs in production
Let Claude (or any LLM) actually watch a video
Find the local LLM that actually runs and performs best