VITS2 backbone with multilingual-bert
Toolkit for audio, music, and speech generation
Guiding Instruction-based Image Editing via Multimodal Large Language
Tool-integrated Reasoning LLM Agents
Code for the paper Language Models are Unsupervised Multitask Learners
A library for transfer learning by reusing parts of TensorFlow models
PDF Combiner is a user-friendly, GUI-based tool built in
AIlice is a fully autonomous, general-purpose AI agent
Repo for YaYi Chinese LLMs based on LlaMA2 & BLOOM
Multi-Voice and Prompt-Controlled TTS Engine
Obsei is a low code AI powered automation tool
A tool for learning vector representations of words and entities
Converts text input or URL into knowledge graph and displays
Dataset of GPT-2 outputs for research in detection, biases, and more
Transformers4Rec is a flexible and efficient library
AI powered speech denoising and enhancement
A deep learning toolkit for Text-to-Speech, battle-tested in research
DuckDuckGo from the terminal
All-in-one text de-duplication
A repository that contains models, datasets, and fine-tuning
Embed images and sentences into fixed-length vectors
Official code for Style Aligned Image Generation via Shared Attention
textgen, Text Generation models
Framework that is dedicated to making neural data processing