The RF and reverse engineering framework for everyone
AI Agent Evaluator & Red Team Platform
Designed for training LLM/VLM agents via RL
Driving with Graph Visual Question Answering
Benchmark LLMs by fighting in Street Fighter 3
Recipes to train reward model for RLHF
An agentless approach to automatically solve software development
A simple, performant and scalable Jax LLM
LISA: Reasoning Segmentation via Large Language Model
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System
Skywork-R1V is an advanced multimodal AI model series
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
CV, NLP, LLM project applications, and advanced engineering deployment
Multimodal Agents as Smartphone Users, an LLM-based multimodal agent
Natural language workflows for AI agents
Document Image Parsing via Heterogeneous Anchor Prompting”
Training framework for Stable Baselines3 reinforcement learning agents
Monte Carlo tree search in JAX
The Library for LLM-based multi-agent applications
Helps developers deploy LangChain runnables and chains as a REST API
Evaluate your LLM's response with Prometheus and GPT4
Run PyTorch LLMs locally on servers, desktop and mobile
Enterprise AI agent platform for workflows, models, and RAG apps
A dataset consists of 15,140 ChatGPT prompts from Reddit
Open-Source Financial Large Language Models