Your CrewAI Powered Video Editing Assistant
ComfyUI wrapper nodes for HunyuanVideo
Expressive Portrait Image Animation for Live Streaming
About 24 Lessons, 12 Weeks, Get Started as a Web Developer
Fast, powerful, git-native ticket tracking in a single bash script
Qwen3-omni is a natively end-to-end, omni-modal LLM
An AI-agent skill that turns Markdown into paste-ready WeChat article
Neural style in TensorFlow
AI framework to autonomously improve the performance of any AI system
GPT Image 2 prompt gallery, image prompt library, agentic skill
An on-premises, OCR-free unstructured data extraction
Handwritten Text Recognition (HTR) system implemented with TensorFlow
Harmonized and Coherent Human Image Animation
Foundation model for image generation
Marrying Grounding DINO with Segment Anything & Stable Diffusion
Motion-controllable Video Generation via Latent Trajectory Guidance
Multimodal embedding and reranking models built on Qwen3-VL
"Big Model" trains a visual multimodal VLM with 26M parameters
Stable Virtual Camera: Generative View Synthesis with Diffusion Models
Modular quant framework
A theme for Sublime Text 3 by Mattia Astorino
Cross-platform API testing client for humans
Gracefully face hCaptcha challenge with multimodal llms
Open multimodal web agent built by Ai2
Zero-code platform for building AI agents from natural language input