face recognition using python free download

Hunyuan3D 2.0

High-Resolution 3D Assets Generation with Large Scale Diffusion Models

The Hunyuan3D-2 model, developed by Tencent, is designed for generating high-resolution 3D assets using large-scale diffusion models. This model offers advanced capabilities for creating detailed 3D models, including texture enhancements, multi-view shape generation, and rapid inference for real-time applications. It is particularly useful for industries requiring high-quality 3D content, such as gaming, film, and virtual reality. Hunyuan3D-2 supports various enhancements and is available...

Downloads: 29 This Week

Last Update: 2025-09-23

See Project

Qwen3-Coder

Qwen3-Coder is the code version of Qwen3

... Claude Sonnet. Qwen3-Coder supports an exceptionally long context window of 256,000 tokens, extendable to 1 million tokens using Yarn, enabling repository-scale code understanding and generation. It is capable of handling 358 programming languages, from common to niche, making it versatile for a wide range of development environments. The model integrates a specially designed function call format and supports popular platforms such as Qwen Code and CLINE for agentic coding workflows.

1 Review

Downloads: 27 This Week

Last Update: 2025-09-23

See Project

Qwen2-Audio

Repo of Qwen2-Audio chat & pretrained large audio language model

... classification, emotion, etc.), and offers pretrained models (e.g. 7B) released via ModelScope and Hugging Face. Code & examples provided with Hugging Face transformers, and usage via AutoProcessor, model classes etc. High performance on many standard benchmarks: ASR, speech-emotion recognition, vocal sound classification, speech translation etc.

Downloads: 4 This Week

Last Update: 2025-09-23

See Project

Qwen2.5-Omni

Capable of understanding text, audio, vision, video

...-of-the-art performance in many multimodal benchmarks, particularly spoken language understanding, audio reasoning, image/video understanding, etc. Very strong benchmark performance across modalities (audio understanding, speech recognition, image/video reasoning) and often outperforming or matching single-modality models at a similar scale. Real-time streaming responses, including natural speech synthesis (text-to-speech) and chunked inputs for low latency interaction.

Downloads: 6 This Week

Last Update: 2025-09-23

See Project

Perception Models

State-of-the-art Image & Video CLIP, Multimodal Large Language Models

... integrates with PE to power vision-language modeling, achieving results competitive with leading multimodal systems such as QwenVL2.5 and InternVL3, all while being fully reproducible with open data. The project supports a wide range of research applications, from visual recognition and dense prediction to fine-grained multimodal understanding. Additionally, it includes several large-scale open datasets for both image and video perception.

Downloads: 0 This Week

Last Update: 2025-10-08

See Project

DeepSeek MoE

Towards Ultimate Expert Specialization in Mixture-of-Experts Language

... or LLaMA2 7B using about 40% of the total compute. The repo publishes both Base and Chat variants of the 16B MoE model (deepseek-moe-16b) and provides evaluation results across benchmarks. It also includes a quick start with inference instructions (using Hugging Face Transformers) and guidance on fine-tuning (DeepSpeed, hyperparameters, quantization). The licensing is MIT for code, with a “Model License” applied to the models.

Downloads: 0 This Week

Last Update: 2025-10-03

See Project

Tencent-Hunyuan-Large

Open-source large language model family from Tencent Hunyuan

Tencent-Hunyuan-Large is the flagship open-source large language model family from Tencent Hunyuan, offering both pre-trained and instruct (fine-tuned) variants. It is designed with long-context capabilities, quantization support, and high performance on benchmarks across general reasoning, mathematics, language understanding, and Chinese / multilingual tasks. It aims to provide competitive capability with efficient deployment and inference. FP8 quantization support to reduce memory usage...

Downloads: 0 This Week

Last Update: 2025-09-24

See Project

CSM (Conversational Speech Model)

A Conversational Speech Generation Model

The CSM (Conversational Speech Model) is a speech generation model developed by Sesame AI that creates RVQ audio codes from text and audio inputs. It uses a Llama backbone and a smaller audio decoder to produce audio codes for realistic speech synthesis. The model has been fine-tuned for interactive voice demos and is hosted on platforms like Hugging Face for testing. CSM offers a flexible setup and is compatible with CUDA-enabled GPUs for efficient execution.

Downloads: 0 This Week

Last Update: 2025-03-19

See Project

Search Results for "face recognition using python"

Showing 8 open source projects for "face recognition using python"

Hunyuan3D 2.0

Qwen3-Coder

Qwen2-Audio

Qwen2.5-Omni

Perception Models

DeepSeek MoE

Tencent-Hunyuan-Large

CSM (Conversational Speech Model)

Search Results for "face recognition using python"

Showing 8 open source projects for "face recognition using python"

Hunyuan3D 2.0

Qwen3-Coder

Qwen2-Audio

Qwen2.5-Omni

Perception Models

DeepSeek MoE

Tencent-Hunyuan-Large

CSM (Conversational Speech Model)

Related Searches

Related Categories