GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
GPT Image 2 prompt gallery, image prompt library, agentic skill
Open source web scraping system for automated data collection tasks
Multilingual Document Layout Parsing in a Single Vision-Language Model
"Big Model" trains a visual multimodal VLM with 26M parameters
A Claude Code Skill that turns prompts into magazine-style HTML decks
Zero dependency docking layout manager supporting tabs
A skill that turn any brand into a scrollable 3D world
ExDARK dataset is the largest collection of low-light images
Build your timeless portfolio with Once UI's Magic Portfolio
UI component library for React, Next.js, and other JSX frameworks
Guiding Instruction-based Image Editing via Multimodal Large Language
Reference project for creating clear, standardized Git commit messages
Temporal-Consistent Diffusion Model for Real-World Video
An AI-agent skill that turns Markdown into paste-ready WeChat article
A cinematic Git commit replay tool for the terminal
macOS app to manage your Codex skills
Open-source AI agent command center for Claude Code agent teams
Free NextJS Landing Page Template written in Tailwind CSS 3
Handwritten Text Recognition (HTR) system implemented with TensorFlow
Curated directory of thousands of generative AI tools by category
High performance and delightful way to play with APNG format in iOS
A command-line utility for taking automated screenshots of websites
Inactive project
Open-source ThreeUI Community catalog with live interactive components