Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
A Unified Framework for Image Customization
RGBD video generation model conditioned on camera input
CodeGeeX4-ALL-9B, a versatile model for all AI software development
Extensible, parallel implementations of t-SNE
Chinese and English multimodal conversational language model
A fast TTS architecture with conditional flow matching
ChatGLM2-6B: An Open Bilingual Chat LLM
Constrained Value Alignment via Safe Reinforcement Learning
Open Multilingual Multimodal Chat LMs
Repo-local continuity runtime for AI coding agents. Helps them continu
Synchronized Translation for Videos
Chinese safety prompts for evaluating and improving the safety of LLMs
VITS2 backbone with multilingual-bert
Quick guide (especially) for trending instruction finetuning dataset
Official code for Style Aligned Image Generation via Shared Attention
Official release of InternLM series
CSAw is an NLP framework for low-resource languages
Clarity in the current fast-paced mess of Open Source innovation
ChatGPT DAN, Jailbreaks prompt
Neural machine translation and sequence learning using TensorFlow
Libraries for optimizing AI models, inference speed, and GPU usage
Joint Face Detection and Alignment
CPT: A Pre-Trained Unbalanced Transformer
GLIDE: a diffusion-based text-conditional image synthesis model