Why use many token when few token do trick
AI-data warehouse to enrich, transform and analyze unstructured data
An Open Source text-to-speech system built by inverting Whisper
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
An open-source toolkit for BigMac-style pipeline-parallel training
An Open Real-time Video-Language Interaction System
Open Vision Agents by Stream. Build voice and vision agents quickly
Open-source platform for evaluating, observing, and improving LLM
Paste Markdown and AI responses into Word Excel instantly fast
Multilingual Document Layout Parsing in a Single Vision-Language Model
Reinforcement Learning for Humanoid Robot with Zero-Shot Sim2Real
Master the fundamentals of machine learning, deep learning
Python library for portfolio optimization built on top of scikit-learn
Numerical differential equation solvers in JAX
A general fine-tuning kit geared toward image/video/audio diffusion
Python Audio Analysis Library: Feature Extraction, Classification
Minimal reproduction of OneRec
NBA sports betting using machine learning
From nobody to big model (LLM) hero
How to optimize some algorithm in cuda
Maimaibot, a (more focused) multi-platform intelligent agent
Cybersecurity AI (CAI), the framework for AI Security
Large Language Model Principles and Practice Tutorial from Scratch
Memory Management Kit for Agents
This repository contains code released by Google Research