HunyuanVideo: A Systematic Framework For Large Video Generation Model
Turn words into colors
Turn words into chords
Automatically translates the text of a video based on a subtitle file
Visual Causal Flow
CLI tool to extract (meta)data from PDF and manipulate PDF files
Stable Diffusion WebUI optimized for AMD GPUs with editing tools
The official repo of Qwen chat & pretrained large language model
Open source libraries and APIs to build custom preprocessing pipelines
Removes 20+ patterns of AI slop from any piece of writing
[CVPR 2026 Oral] VGGT Omega
The data structure for multimodal data
Offical Implementation for "Recursive Multi-Agent Systems"
Public opinion analysis system
Making RAG Simpler with Small and Open-Sourced Language Models
A New Axis of Sparsity for Large Language Models
"Big Model" trains a visual multimodal VLM with 26M parameters
Curl cryptocurrencies exchange rates
Implementation of "MobileCLIP" CVPR 2024
Scalable data pre processing and curation toolkit for LLMs
Collect your thoughts and notes without leaving the command line
Model Context Protocol Server for Apache OpenDAL™
Qwen-Image-Layered: Layered Decomposition for Inherent Editablity
A library for converting HTML into PDFs using ReportLab
borb is a library for reading, creating and manipulating PDF files