AI cognitive-enhancement Skills based on Anthropic's J-space
first person shooter, space invader
A browser agent with a dynamic, indexed action space
CLIP, Predict the most relevant text snippet given an image
PyTorch code and models for VJEPA2 self-supervised learning from video
AI-Driven Exploration in the Space of Code
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
Text-space optimizer that trains reusable natural-language skills
Foundational video generation model with 13.6B parameters
PyTorch code and models for V-JEPA self-supervised learning from video
A Family of Open Sourced Music Foundation Models
Recovering the Visual Space from Any Views
Implementation of the Surya Foundation Model for Heliophysics
One-ink editorial print image skill
Multimodal Agents as Smartphone Users, an LLM-based multimodal agent
A collective list of free APIs
1B text generation model based on the HRM architecture
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
Generate Any 3D Scene in Seconds
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
Open source file indexing & storage analytics powered by Elasticsearch
Generate high-definition story short videos with one click using AI
Semi-Structured Agentic Framework. Workflows build themselves
Open-Source Dual-Arm Mobile Robot with Motorized Lift
Latent Collaboration in Multi-Agent Systems