Code for running inference with the SAM 3D Body Model 3DB
A Powerful Native Multimodal Model for Image Generation
Open Vision Agents by Stream. Build voice and vision agents quickly
Synthesizing and manipulating 2048x1024 images with conditional GANs
A Unified Library for Parameter-Efficient Learning
Integrating LLMs into structured NLP pipelines
Framework and no-code GUI for fine-tuning LLMs
A coding-free framework built on PyTorch
A series of math-specific large language models of our Qwen2 series
Qwen-Image is a powerful image generation foundation model
Training Large Language Model to Reason in a Continuous Latent Space
Gracefully face hCaptcha challenge with multimodal llms
State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX
MII makes low-latency and high-throughput inference possible
Ultra-Efficient LLMs on End Device
Library for building type-safe natural language interfaces with LLMs
machine learning tutorials (mainly in Python3)
Implementation for MatMul-free LM
Power CLI and Workflow manager for LLMs (core package)
Performance-optimized AI inference on your GPUs
Retrieval and Retrieval-augmented LLMs
Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass
The open-source data curation platform for LLMs
Agent S: an open agentic framework that uses computers like a human
An Open Source text-to-speech system built by inverting Whisper