"VideoRAG: Chat with Your Videos
Capable of understanding text, audio, vision, video
Framework for building, orchestrating, and deploying AI agents
Hackable and optimized Transformers building blocks
Synthetic data generators for tabular and time-series data
Anthropic's educational courses
An on-premises, OCR-free unstructured data extraction
Production-grade platform for building agentic IM bots
Open-source industrial-grade ASR models
A frontier, first-principles handbook
End-to-end pipeline converting generative videos
OpenTinker is an RL-as-a-Service infrastructure for foundation models
Motion-controllable Video Generation via Latent Trajectory Guidance
Official implementation of Watermark Anything with Localized Messages
Video understanding codebase from FAIR for reproducing video models
A general fine-tuning kit geared toward image/video/audio diffusion
Llama Chinese community, real-time aggregation
SimpleMem: Efficient Lifelong Memory for LLM Agents
Qwen3-omni is a natively end-to-end, omni-modal LLM
Collaborative & Open-Source Quality Assurance for all AI models
The official Python SDK for the xAI API
Sparsity-aware deep learning inference runtime for CPUs
DeepMind model for tracking arbitrary points across videos & robotics
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
Generating Immersive, Explorable, and Interactive 3D Worlds