Contexts Optical Compression
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Open-source multi-speaker long-form text-to-speech model
Open Source Speech Language Model
Bidirectional token-classification model for identifiable info
Diffusion Transformer with Fine-Grained Chinese Understanding
Large-language-model & vision-language-model based on Linear Attention
The official repo of Qwen chat & pretrained large language model
Ultra-Efficient LLMs on End Device
Audio foundation model excelling in audio understanding
OCR expert VLM powered by Hunyuan's native multimodal architecture
Visual Causal Flow
Multi-modal large language model designed for audio understanding
Large Multimodal Models for Video Understanding and Editing
Official implementation of DreamCraft3D
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
AI-powered tool to quickly remove watermarks from images flawlessly
A Conversational Speech Generation Model
Dataset of GPT-2 outputs for research in detection, biases, and more