Elyra extends JupyterLab with an AI centric approach
Full-stack AI Red Teaming platform
No-code LLM Platform to launch APIs and ETL Pipelines
Self-supervised visual learning using momentum contrast in PyTorch
Wan2.1: Open and Advanced Large-Scale Video Generative Model
GitLab automatic code review tool based on large models
Lets make video diffusion practical
AI tool that converts GitHub repositories into interactive diagrams
Suite of reference architectures for building GPU-accelerated vision
Claude code for everything except coding
Welcome the Era of One-shot Long-horizon Parsing
Python inference and LoRA trainer package for the LTX-2 audio–video
Official Repo For "Sa2VA: Marrying SAM2 with LLaVA
Fast, powerful, git-native ticket tracking in a single bash script
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Multilingual Document Layout Parsing in a Single Vision-Language Model
"Big Model" trains a visual multimodal VLM with 26M parameters
Python package for AutoML on Tabular Data with Feature Engineering
Guiding Instruction-based Image Editing via Multimodal Large Language
Temporal-Consistent Diffusion Model for Real-World Video
Taming Stable Diffusion for Lip Sync
OCR expert VLM powered by Hunyuan's native multimodal architecture
Handwritten Text Recognition (HTR) system implemented with TensorFlow
Inference script for Oasis 500M
ICLR2024 Spotlight: curation/training code, metadata, distribution