Diffusion Transformer with Fine-Grained Chinese Understanding
Lightning fast C++/CUDA neural network framework
Generate high-definition story short videos with one click using AI
Gracefully face hCaptcha challenge with multimodal llms
Enlightened library to convert HTML and CSS to SVG
Ollama client that simplifies experimenting with LLMs
Plugin and skin collection for DeepSeek Harness (DSH) Web UI
AI-powered Telegram chat backup and semantic search tool system
AI Toolkit for Healthcare Imaging
GPT4V-level open-source multi-modal model based on Llama3-8B
Codex plugin that turns attached object images into code-only
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
2^x Image Super-Resolution
Provides code for running inference with the SegmentAnything Model
Offline inference engine for art, real-time voice conversations
Nexa SDK is a comprehensive toolkit for supporting ONNX and GGML
Use Claude Code as the foundation for coding infrastructure
Focus on prompting and generating
Swift community driven package for OpenAI public API
Sharp Monocular Metric Depth in Less Than a Second
Model Context Procotol(MCP) server for using Amazon Bedrock
Dealing with all unstructured data, such as reverse image search
Capable of understanding text, audio, vision, video
A Unified Framework for Image Customization
High-Resolution Image Synthesis with Latent Diffusion Models