Open source AI model for generating full songs from lyrics prompts
A lightweight alternative to Clawdbot / OpenClaw
Extensible, parallel implementations of t-SNE
Semi-Structured Agentic Framework. Workflows build themselves
CLIP, Predict the most relevant text snippet given an image
Foundational video generation model with 13.6B parameters
AI bridge enabling Cursor agents to read and modify Figma designs
GLM-Image: Auto-regressive for Dense-knowledge and High-fidelity Image
A Survey of Large Language Models
Your Fully-Automated Personal AI Assistant
Java interface to OpenCV, FFmpeg, and more
Give Claude the ability to watch and understand videos
A series of math-specific large language models of our Qwen2 series
Handwritten Text Recognition (HTR) system implemented with TensorFlow
Unsupervised Learning for Image Registration
RGBD video generation model conditioned on camera input
CUDA Templates for Linear Algebra Subroutines
One-stop AI digital human system with video voice synthesis tools
Unleashing 10,000+ Word Generation from Long Context LLMs
The best ChatGPT that $100 can buy
Analyze computation-communication overlap in V3/R1
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Trainable, memory-efficient, and GPU-friendly PyTorch reproduction
Chinese and English multimodal conversational language model
A Unified Framework for Image Customization