Qwen3-omni is a natively end-to-end, omni-modal LLM
LLM inference server with continuous batching & SSD caching
Pre & Post-training & Dataset & Evaluation & Depoly & RAG
The first AI agent that builds permissionless integrations
A high-performance ML model serving framework, offers dynamic batching
A lightweight framework for building LLM-based agents
Overcoming Group Chat Scenarios with LLM-based Technical Assistance
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
Structured data extraction and instruction calling with ML, LLM
MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training
Capable of understanding text, audio, vision, video
Build AI WhatsApp Bots with Pure Python
ChatGLM-6B: An Open Bilingual Dialogue Language Model
ChatGLM2-6B: An Open Bilingual Chat LLM
A state-of-the-art open visual language model
Automatic question answering for local knowledge bases based on LLM
Did you say you like data?
Code for Language models can explain neurons in language models paper
Visual Instruction Tuning: Large Language-and-Vision Assistant
Ship RAG based LLM web apps in seconds
Run any Llama 2 locally with gradio UI on GPU or CPU from anywhere
Codes for "Chameleon: Plug-and-Play Compositional Reasoning