Learn to build your Second Brain AI assistant with LLMs
Tools for merging pretrained large language models
Controllable and fast Text-to-Speech for over 7000 languages
Unified Multimodal Understanding and Generation Models
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
Implementation of the Surya Foundation Model for Heliophysics
Interface for OuteTTS models
FAIR Chemistry's library of machine learning methods for chemistry
Miso TTS is an 8 billion, highly emotive text-to-speech model
Capable of understanding text, audio, vision, video
Automated Face Blurring, Kinematics Extraction and Leg dystonia Dx
Graphical User Interface Face Anonymization Tool
ChatGLM-6B: An Open Bilingual Dialogue Language Model
Di♪♪Rhythm: Blazingly Fast & Simple End-to-End Song Generation
3x3x3 Rubik's Cube solver
MARS5 speech model (TTS) from CAMB.AI
Powerful open source image generation model
Unofficial implementation of InstantID for ComfyUI
Local Face Tagging Photos in ALL Formats
Stable Diffusion with Core ML on Apple Silicon
Synchronized Translation for Videos
Openai style api for open large language models
Image processing App for Windows Desktop
Open-source choice to scale, assess and maintain natural language data
A Conversational Speech Generation Model