Learning to Reason with Search for LLMs via Reinforcement Learning
Benchmark LLMs by fighting in Street Fighter 3
Cache-Augmented Generation: A Simple, Efficient Alternative to RAG
Recipes to train reward model for RLHF
An Efficient Web-enhanced Question Answering System
Unleashing 10,000+ Word Generation from Long Context LLMs
Make your agents learn from experience
An agentless approach to automatically solve software development
Empowering Code Generation with OSS-Instruct
Neural Network architecture based on ideas of the original LSTM
A simple, performant and scalable Jax LLM
A lightweight framework for building LLM-based agents
TigerBot: A multi-language multi-task LLM
The SOTA Open-Source Browser Agent
Overcoming Group Chat Scenarios with LLM-based Technical Assistance
LISA: Reasoning Segmentation via Large Language Model
The Security Toolkit for LLM Interactions
Enhances Tesseract OCR output using LLMs (local or API)
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System
Implementation for MatMul-free LM
Leaderboard Comparing LLM Performance at Producing Hallucinations
Skywork-R1V is an advanced multimodal AI model series
DepGraph: Towards Any Structural Pruning
Code and models for ICML 2024 paper, NExT-GPT
High-performance Inference and Deployment Toolkit for LLMs and VLMs