Open-source large language model family from Tencent Hunyuan
The open-source data curation platform for LLMs
Synthesizing and manipulating 2048x1024 images with conditional GANs
Gracefully face hCaptcha challenge with multimodal llms
Tools for merging pretrained large language models
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
A python module for scientific analysis of 3D data
Miso TTS is an 8 billion, highly emotive text-to-speech model
Framework and no-code GUI for fine-tuning LLMs
MII makes low-latency and high-throughput inference possible
Capable of understanding text, audio, vision, video
An Open Source text-to-speech system built by inverting Whisper
Graphical User Interface Face Anonymization Tool
Automated Face Blurring, Kinematics Extraction and Leg dystonia Dx
Di♪♪Rhythm: Blazingly Fast & Simple End-to-End Song Generation
3x3x3 Rubik's Cube solver
Powerful open source image generation model
Local Face Tagging Photos in ALL Formats
Image processing App for Windows Desktop
A Conversational Speech Generation Model
AI Suite for upscaling, interpolating & restoring images/videos
MARS5 speech model (TTS) from CAMB.AI
ChatGLM-6B: An Open Bilingual Dialogue Language Model
Unofficial implementation of InstantID for ComfyUI
Synchronized Translation for Videos