Implementation of Imagen, Google's Text-to-Image Neural Network
Hub of ready-to-use datasets for ML models
Build cross-modal and multimodal applications on the cloud
A library for deep learning end-to-end dialog systems and chatbots
The go-to web for your AI coding agent
A general, three-party dependency-free, cross-platform
C++ DataFrame for statistical, Financial, and ML analysis
PyTorch/TorchScript/FX compiler for NVIDIA GPUs using TensorRT
Free, ultrafast Copilot alternative for Vim and Neovim
A fast image processing library with low memory needs
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
A trainable PyTorch reproduction of AlphaFold 3
Official Repo For "Sa2VA: Marrying SAM2 with LLaVA
High-Fidelity and Controllable Generation of Textured 3D Assets
Multi-modal large language model designed for audio understanding
Open-source framework for intelligent speech interaction
Large Multimodal Models for Video Understanding and Editing
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
On-device Speech-to-Intent engine powered by deep learning
Benchmarking synthetic data generation methods
AIMET is a library that provides advanced quantization and compression
Powering Amazon custom machine learning chips
Model Context Protocol server that integrates AgentQL's data
Recognition and resolution of numbers, units, date/time, etc.