A Unified Framework for Text-to-3D and Image-to-3D Generation
Multimodal Diffusion with Representation Alignment
Bidirectional token-classification model for identifiable info
Genome modeling and design across all domains of life
Project Lyra: Open Generative 3D World Models
Achieving 3+ generation speedup on reasoning tasks
Ultra-Efficient LLMs on End Device
Generate Any 3D Scene in Seconds
Fast and Universal 3D reconstruction model for versatile tasks
A collection of high-quality models for the MuJoCo physics engine
A Production-ready Reinforcement Learning AI Agent Library
Official implementation of DreamCraft3D
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Global weather forecasting model using graph neural networks and JAX
FAIR Sequence Modeling Toolkit 2
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
Provides convenient access to the Anthropic REST API from any Python 3
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
High-Fidelity and Controllable Generation of Textured 3D Assets
RGBD video generation model conditioned on camera input
The Clay Foundation Model - An open source AI model and interface
Use ChatGPT to summarize the arXiv papers
Community plugin marketplace for Claude Cowork and Claude Code
AI PPT Track Terminator, the strongest PPT Skill ever
An Efficient Agentic Model for Computer Use