Efficient Triton Kernels for LLM Training
Supercharge Your LLM with the Fastest KV Cache Layer
A python tool that uses GPT-4, FFmpeg, and OpenCV
Style-Bert-VITS2: Bert-VITS2 with more controllable voice styles
Implementation of TurboQuant (ICLR 2026)
Python tool for browser-based interactive data apps in one file
A Multi-Modal World Model for Reconstructing, Generating, Simulation
A security scanner for custom LLM applications
Reproduction of Poetiq's record-breaking submission to the ARC-AGI-1
Management of Yandex Station and other smart home devices
Controllable and fast Text-to-Speech for over 7000 languages
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
Generating Immersive, Explorable, and Interactive 3D Worlds
Proofs, cases, concept supplements, and reference explanations
Framework for building realtime multimodal voice AI agents apps
AI-Powered Personalized Learning Assistant
Open-weight, large-scale hybrid-attention reasoning model
Large-language-model & vision-language-model based on Linear Attention
Qwen3-omni is a natively end-to-end, omni-modal LLM
A python library for easy manipulation and forecasting of time series
Request recommended movies, TV shows and anime to Jellyseer/Overseer
Offline inference engine for art, real-time voice conversations
AI agents running research on single-GPU nanochat training
Go ahead and axolotl questions
Evaluation suite designed to assess the performance of LLMs