Performance Co-Pilot
Build cross-modal and multimodal applications on the cloud
Full-stack Open-source Self-Evolving General AI Agent
Actionhero is a realtime multi-transport nodejs API server
Text and image to video generation: CogVideoX and CogVideo
A coding-free framework built on PyTorch
Implementing large models into scenario-based applications
Gracefully face hCaptcha challenge with multimodal llms
Diffusion Transformer with Fine-Grained Chinese Understanding
Fancy stream processing made operationally mundane
Deep Research framework, combining language models with tools
Multilingual sentence & image embeddings with BERT
This repo contains the code for 1D tokenizer and generator
A general-purpose AI image generation framework that supports HF
A framework for out-of-core and parallel execution
Simple, powerful and flexible site generation framework
A Storybook Addon, Save the screenshot image of your stories
Composable building blocks to build Llama Apps
Ongoing research training transformer models at scale
Lightning fast C++/CUDA neural network framework
Cluster computing framework for processing large-scale geospatial data
A framework to enable multimodal models to operate a computer
A system for agentic LLM-powered data processing and ETL
Open source headless commerce framework built with TypeScript
A Customizable Image-to-Video Model based on HunyuanVideo