Open-Source Financial Large Language Models
PyTorch code and models for the DINOv2 self-supervised learning
One-click local MCP server installation in desktop apps
A Customizable Image-to-Video Model based on HunyuanVideo
A 0.1B Omni model trained from scratch
26m function call model that runs on incredibly small devices
Video Object and Interaction Deletion
Phi-3.5 for Mac: Locally-run Vision and Language Models
Qwen-Image-Layered: Layered Decomposition for Inherent Editablity
Qwen2.5-VL is the multimodal large language model series
1B text generation model based on the HRM architecture
HY-Motion model for 3D character animation generation
Generate Any 3D Scene in Seconds
Repo of Qwen2-Audio chat & pretrained large audio language model
Tongyi Deep Research, the Leading Open-source Deep Research Agent
Diffusion Bee is the easiest way to run Stable Diffusion locally
Advancing Open-source World Models
A Systematic Framework for Interactive World Modeling
code for Mesh R-CNN, ICCV 2019
GLM-4-Voice | End-to-End Chinese-English Conversational Model
Generating Immersive, Explorable, and Interactive 3D Worlds
Open image model at the forefront of design
Hunyuan Translation Model Version 1.5
Tool for exploring and debugging transformer model behaviors
CLIP, Predict the most relevant text snippet given an image