Towards Efficient Self-Evolving Agent System
Chat with any codebase in under two minutes | Fully local
Unified KV Cache Compression Methods for Auto-Regressive Models
Learning to Reason with Search for LLMs via Reinforcement Learning
A tension reasoning engine over 131 S-class problems
Bringing BERT into modernity via both architecture changes and scaling
Scalable RL solution for advanced reasoning of language models
Unleashing 10,000+ Word Generation from Long Context LLMs
An agentless approach to automatically solve software development
A simple, performant and scalable Jax LLM
The Cradle framework is a first attempt at General Computer Control
LISA: Reasoning Segmentation via Large Language Model
The Security Toolkit for LLM Interactions
Enhances Tesseract OCR output using LLMs (local or API)
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System
Leaderboard Comparing LLM Performance at Producing Hallucinations
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
Code and models for ICML 2024 paper, NExT-GPT
Run PyTorch LLMs locally on servers, desktop and mobile
Examples and tutorials to help developers build AI systems
Build a large language model from 0 only with Python foundation
Instruction-tuning LLM with Chinese Medical Knowledge
Accelerate local LLM inference and finetuning
Chat with it via text and voice
Long-form streaming TTS system for multi-speaker dialogue generation