Skywork-R1V is an advanced multimodal AI model series
Code and models for ICML 2024 paper, NExT-GPT
CV, NLP, LLM project applications, and advanced engineering deployment
Multimodal Agents as Smartphone Users, an LLM-based multimodal agent
A simple yet powerful agent framework for personal assistants
Natural language workflows for AI agents
Bringing the Unsloth experience to Mac users via Apple's MLX framework
Document Image Parsing via Heterogeneous Anchor Prompting”
MobileLLM Optimizing Sub-billion Parameter Language Models
ImageBind One Embedding Space to Bind Them All
A Powerful Native Multimodal Model for Image Generation
Training framework for Stable Baselines3 reinforcement learning agents
Monte Carlo tree search in JAX
The Library for LLM-based multi-agent applications
Efficient few-shot learning with Sentence Transformers
Framework for rapid prototyping and development of real-time rendering
A fast and lightweight framework for creating decentralized agents
Enables Jupyter Notebooks to share resources across clusters
Code to accompany "A Method for Animating Children's Drawings"
Autonomous research from idea to paper. Chat an Idea. Get a Paper 🦞
Official implementation of Watermark Anything with Localized Messages
Video understanding codebase from FAIR for reproducing video models
Agent framework and applications built upon Qwen>=3.0
Framework and no-code GUI for fine-tuning LLMs
A text-to-speech, speech-to-text and speech-to-speech library