Implementation of AudioLM audio generation model in Pytorch
The data structure for multimodal data
Offical Implementation for "Recursive Multi-Agent Systems"
Search all of YouTube from the command line
Public opinion analysis system
Qwen3-ASR is an open-source series of ASR models
Making RAG Simpler with Small and Open-Sourced Language Models
A New Axis of Sparsity for Large Language Models
"Big Model" trains a visual multimodal VLM with 26M parameters
Implementation of "MobileCLIP" CVPR 2024
Scalable data pre processing and curation toolkit for LLMs
Model Context Protocol Server for Apache OpenDAL™
Evaluate and monitor ML models from validation to production
Deep Understanding AI Agents
Foundational model for human-like, expressive TTS
Pretrained model hub for Keras 3
Conversational voice AI agents
Pushing the Frontier of Long Audio-Visual Generation
An open phone agent model & framework
Concatenate a directory full of files into a single prompt
Models for the spaCy Natural Language Processing (NLP) library
Bidirectional token-classification model for identifiable info
Ultra-Efficient LLMs on End Device
Open speech-to-speech models and pipelines by Hugging Face toolkit AI
Advanced NLP with spaCy: A free online course