MiniMax H3 is a general-purpose, omni-modal generative system
Industrial-level controllable zero-shot text-to-speech system
Awesome multilingual OCR toolkits based on PaddlePaddle
AlphaFold 3 inference pipeline
From Images to High-Fidelity 3D Assets
Native and Compact Structured Latents for 3D Generation
An Open Real-time Video-Language Interaction System
Visual Causal Flow
Video Object and Interaction Deletion
A theoretical reconstruction of the Claude Mythos architecture
AI cognitive-enhancement Skills based on Anthropic's J-space
Genome modeling and design across all domains of life
Reproduction of Poetiq's record-breaking submission to the ARC-AGI-1
A collection of high-quality models for the MuJoCo physics engine
Contexts Optical Compression
Controllable & emotion-expressive zero-shot TTS
Audio foundation model excelling in audio understanding
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Jev-like family of decision models built on top of Qwen3.5/3.8
Robust Speech Recognition Across Languages, Dialects
FAIR Sequence Modeling Toolkit 2
Python SDK for Claude Agent
Open Source Speech Language Model
A Multi-Modal World Model for Reconstructing, Generating, Simulation
Bidirectional token-classification model for identifiable info