[NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences
Source code for the X Recommendation Algorithm
21 Lessons, Get Started Building with Generative AI
A Customizable Image-to-Video Model based on HunyuanVideo
Build multi-modal Agents with memory, knowledge, tools and reasoning
Python client for the Telegram's tdlib
Enables Jupyter Notebooks to share resources across clusters
MTEB: Massive Text Embedding Benchmark
Streamline your ML workflow
Framework for building AI-powered interactive digital humans and agent
AI-ready web crawler that extracts and structures website content
Python tool for crawling and extracting structured data from news site
Check breached emails and find exposed passwords from public dumps
Open source framework for large scale network reconnaissance and analy
Tools for merging pretrained large language models
Build your own Cowork, AI Scientist and other SoTA Agents
Qwen-Image-Layered: Layered Decomposition for Inherent Editablity
Context data platform for building observable, self-learning AI agents
End-to-end speech processing toolkit
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
Provides convenient access to the Anthropic REST API from any Python 3
GLM-4-Voice | End-to-End Chinese-English Conversational Model
GPT4V-level open-source multi-modal model based on Llama3-8B
GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning
A port of liquid template engine for python