Learning agent trained in a diffusion world model
The goal of CLAIMED is to enable low-code/no-code rapid prototyping
Framework for building realtime multimodal voice AI agents apps
Hypernetworks that adapt LLMs for specific benchmark tasks
Driving with Graph Visual Question Answering
Chat with any codebase in under two minutes | Fully local
Unified KV Cache Compression Methods for Auto-Regressive Models
Scalable RL solution for advanced reasoning of language models
TigerBot: A multi-language multi-task LLM
Overcoming Group Chat Scenarios with LLM-based Technical Assistance
Implementation for MatMul-free LM
DepGraph: Towards Any Structural Pruning
StarVector is a foundation model for SVG generation
Implement a concise and clear Deep Search Agent from 0
One API call, pull Claude agent, completely sandboxed
A curated collection of skills for AI coding agents
Large Audio Language Model built for natural interactions
Specification and documentation for Agent Skills
Document Image Parsing via Heterogeneous Anchor Prompting”
Framework for building neural networks
ICLR2024 Spotlight: curation/training code, metadata, distribution
The open-source data curation platform for LLMs
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
A personal context-agent that learns how you work
Learn to build your Second Brain AI assistant with LLMs