2D and 3D Face alignment library build using pytorch
Robust recipes to align language models with human and AI preferences
Industry leading face manipulation platform
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
State-of-the-art 2D and 3D Face Analysis Project
Automatic Speech Recognition with Word-level Timestamps
CogView4, CogView3-Plus and CogView3(ECCV 2024)
An alignment auditing agent capable of exploring alignment hypothesis
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
Multimodal Diffusion with Representation Alignment
Technical principles related to large models
Qwen3 is the large language model series developed by Qwen team
A python library that makes AMR parsing, generation and visualization
Multimodal-Driven Architecture for Customized Video Generation
A Lightweight Face Recognition and Facial Attribute Analysis
Data manipulation and transformation for audio signal processing
Multilingual Automatic Speech Recognition with word-level timestamps
A dataset consists of 15,140 ChatGPT prompts from Reddit
GLM-4 series: Open Multilingual Multimodal Chat LMs
3D reconstruction software
Volcano Engine Reinforcement Learning for LLMs
Uncommon Objects in 3D dataset
Open Source Speech Language Model
CodeGeeX4-ALL-9B, a versatile model for all AI software development
High-Performance Face Recognition Library on PaddlePaddle & PyTorch