Reinforcement Learning Algorithms

View 30 business solutions

Browse free open source Reinforcement Learning Algorithms and projects below. Use the toggles on the left to filter open source Reinforcement Learning Algorithms by OS, license, language, programming language, and project status.

  • Build Agents and Models on One Platform Icon
    Build Agents and Models on One Platform

    Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

    Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
    Try It Free
  • $300 Free Credits to Build on Google Cloud Icon
    $300 Free Credits to Build on Google Cloud

    New to Google Cloud? Get $300 in credits to explore Compute Engine, BigQuery, Cloud Run, Gemini Enterprise Agent Platform, and more.

    Start your next project with $300 in free Google Cloud credit. Spin up VMs, run containers, query petabytes in BigQuery, or build agents with Gemini Enterprise Agent Platform. Once your credits are used, keep building with 20+ always-free tier products including Compute Engine, Cloud Storage, GKE, and Cloud Run functions. No commitment required—just sign up and start building.
    Claim $300 Free
  • 1
    AirSim

    AirSim

    A simulator for drones, cars and more, built on Unreal Engine

    AirSim is an open-source, cross platform simulator for drones, cars and more vehicles, built on Unreal Engine with an experimental Unity release in the works. It supports software-in-the-loop simulation with popular flight controllers such as PX4 & ArduPilot and hardware-in-loop with PX4 for physically and visually realistic simulations. It is developed as an Unreal plugin that can simply be dropped into any Unreal environment. AirSim's development is oriented towards the goal of creating a platform for AI research to experiment with deep learning, computer vision and reinforcement learning algorithms for autonomous vehicles. For this purpose, AirSim also exposes APIs to retrieve data and control vehicles in a platform independent way. AirSim is fully enabled for multiple vehicles. This capability allows you to create multiple vehicles easily and use APIs to control them.
    Downloads: 40 This Week
    Last Update:
    See Project
  • 2
    Bullet Physics SDK

    Bullet Physics SDK

    Real-time collision detection and multi-physics simulation for VR

    This is the official C++ source code repository of the Bullet Physics SDK: real-time collision detection and multi-physics simulation for VR, games, visual effects, robotics, machine learning etc. We are developing a new differentiable simulator for robotics learning, called Tiny Differentiable Simulator, or TDS. The simulator allows for hybrid simulation with neural networks. It allows different automatic differentiation backends, for forward and reverse mode gradients. TDS can be trained using Deep Reinforcement Learning, or using Gradient based optimization (for example LFBGS). In addition, the simulator can be entirely run on CUDA for fast rollouts, in combination with Augmented Random Search. This allows for 1 million simulation steps per second. It is highly recommended to use PyBullet Python bindings for improved support for robotics, reinforcement learning and VR. Use pip install pybullet and checkout the PyBullet Quickstart Guide.
    Downloads: 11 This Week
    Last Update:
    See Project
  • 3
    Brax

    Brax

    Massively parallel rigidbody physics simulation

    Brax is a fast and fully differentiable physics engine for large-scale rigid body simulations, built on JAX. It is designed for research in reinforcement learning and robotics, enabling efficient simulations and gradient-based optimization.
    Downloads: 4 This Week
    Last Update:
    See Project
  • 4
    Godot RL Agents

    Godot RL Agents

    An Open Source package that allows video game creators

    godot_rl_agents is a reinforcement learning integration for the Godot game engine. It allows AI agents to learn how to interact with and play Godot-based games using RL algorithms. The toolkit bridges Godot with Python-based RL libraries like Stable-Baselines3, making it possible to create complex and visually rich RL environments natively in Godot.
    Downloads: 3 This Week
    Last Update:
    See Project
  • Fully Managed MySQL, PostgreSQL, and SQL Server Icon
    Fully Managed MySQL, PostgreSQL, and SQL Server

    Automatic backups, patching, replication, and failover. Focus on your app, not your database.

    Cloud SQL handles your database ops end to end, so you can focus on your app.
    Try Free
  • 5
    Physical Symbolic Optimization (Φ-SO)

    Physical Symbolic Optimization (Φ-SO)

    Physical Symbolic Optimization

    Physical Symbolic Optimization (Φ-SO) - A symbolic optimization package built for physics. Symbolic regression module uses deep reinforcement learning to infer analytical physical laws that fit data points, searching in the space of functional forms.
    Downloads: 3 This Week
    Last Update:
    See Project
  • 6
    ViZDoom

    ViZDoom

    Doom-based AI research platform for reinforcement learning

    ViZDoom allows developing AI bots that play Doom using only the visual information (the screen buffer). It is primarily intended for research in machine visual learning, and deep reinforcement learning, in particular. ViZDoom is based on ZDOOM, the most popular modern source-port of DOOM. This means compatibility with a huge range of tools and resources that can be used to create custom scenarios, availability of detailed documentation of the engine and tools and support of Doom community. Async and sync single-player and multi-player modes. Fast (up to 7000 fps in sync mode, single-threaded). Lightweight (few MBs). Customizable resolution and rendering parameters. Access to the depth buffer (3D vision). Automatic labeling of game objects visible in the frame. Access to the list of actors/objects and map geometry.ViZDoom API is reinforcement learning friendly (suitable also for learning from demonstration, apprenticeship learning or apprenticeship via inverse reinforcement learning.
    Downloads: 3 This Week
    Last Update:
    See Project
  • 7
    Tensorforce

    Tensorforce

    A TensorFlow library for applied reinforcement learning

    Tensorforce is an open-source deep reinforcement learning framework built on TensorFlow, emphasizing modularized design and straightforward usability for applied research and practice.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 8
    AI4U

    AI4U

    Multi-engine plugin to specify agents with reinforcement learning

    AI4U is a multi-engine plugin (Godot and Unity) that allows you to design Non-Player Characters (NPCs) of games using an agent abstraction. In addition, AI4U has a low-level API that allows you to connect the agent to any algorithm made available in Python by the reinforcement learning community specifically and by the Artificial Intelligence community in general. Reinforcement learning promises to overcome traditional navigation mesh mechanisms in games and to provide more autonomous characters. AI4U can be integrated into Imitation Learning through Behavioral Cloning or Generative Adversarial Imitation Learning present on stable-baslines. Train using multiple concurrent Unity/Godot environment instances. Unity/Godot environment partial control from Python. Wrap Unity/Godot learning environments as a gym.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 9
    Jittor

    Jittor

    Jittor is a high-performance deep learning framework

    Jittor is a high-performance deep learning framework based on JIT compiling and meta-operators. The whole framework and meta-operators are compiled just in time. A powerful op compiler and tuner are integrated into Jittor. It allowed us to generate high-performance code specialized for your model. Jittor also contains a wealth of high-performance model libraries, including image recognition, detection, segmentation, generation, differentiable rendering, geometric learning, reinforcement learning, etc. The front-end language is Python. Module Design and Dynamic Graph Execution is used in the front-end, which is the most popular design for deep learning framework interface. The back-end is implemented by high-performance languages, such as CUDA, C++. Jittor'op is similar to NumPy. Let's try some operations. We create Var a and b via operation jt.float32, and add them. Printing those variables shows they have the same shape and dtype.
    Downloads: 1 This Week
    Last Update:
    See Project
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • 10
    ML for Trading

    ML for Trading

    Code for machine learning for algorithmic trading, 2nd edition

    On over 800 pages, this revised and expanded 2nd edition demonstrates how ML can add value to algorithmic trading through a broad range of applications. Organized in four parts and 24 chapters, it covers the end-to-end workflow from data sourcing and model development to strategy backtesting and evaluation. Covers key aspects of data sourcing, financial feature engineering, and portfolio management. The design and evaluation of long-short strategies based on a broad range of ML algorithms, how to extract tradeable signals from financial text data like SEC filings, earnings call transcripts or financial news. Using deep learning models like CNN and RNN with financial and alternative data, and how to generate synthetic data with Generative Adversarial Networks, as well as training a trading agent using deep reinforcement learning.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 11
    Machine Learning PyTorch Scikit-Learn

    Machine Learning PyTorch Scikit-Learn

    Code Repository for Machine Learning with PyTorch and Scikit-Learn

    Initially, this project started as the 4th edition of Python Machine Learning. However, after putting so much passion and hard work into the changes and new topics, we thought it deserved a new title. So, what’s new? There are many contents and additions, including the switch from TensorFlow to PyTorch, new chapters on graph neural networks and transformers, a new section on gradient boosting, and many more that I will detail in a separate blog post. For those who are interested in knowing what this book covers in general, I’d describe it as a comprehensive resource on the fundamental concepts of machine learning and deep learning. The first half of the book introduces readers to machine learning using scikit-learn, the defacto approach for working with tabular datasets. Then, the second half of this book focuses on deep learning, including applications to natural language processing and computer vision.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 12
    Project Malmo

    Project Malmo

    A platform for Artificial Intelligence experimentation on Minecraft

    How can we develop artificial intelligence that learns to make sense of complex environments? That learns from others, including humans, how to interact with the world? That learns transferable skills throughout its existence, and applies them to solve new, challenging problems? Project Malmo sets out to address these core research challenges, addressing them by integrating (deep) reinforcement learning, cognitive science, and many ideas from artificial intelligence. The Malmo platform is a sophisticated AI experimentation platform built on top of Minecraft, and designed to support fundamental research in artificial intelligence. The Project Malmo platform consists of a mod for the Java version, and code that helps artificial intelligence agents sense and act within the Minecraft environment. The two components can run on Windows, Linux, or Mac OS, and researchers can program their agents in any programming language they’re comfortable with.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 13
    Pwnagotchi

    Pwnagotchi

    Deep Reinforcement learning instrumenting bettercap for WiFi pwning

    Pwnagotchi is an A2C-based “AI” powered by bettercap and running on a Raspberry Pi Zero W that learns from its surrounding WiFi environment in order to maximize the crackable WPA key material it captures (either through passive sniffing or by performing deauthentication and association attacks). This material is collected on disk as PCAP files containing any form of handshake supported by hashcat, including full and half WPA handshakes as well as PMKIDs. Instead of merely playing Super Mario or Atari games like most reinforcement learning based “AI” (yawn), Pwnagotchi tunes its own parameters over time to get better at pwning WiFi things in the real world environments you expose it to. To give hackers an excuse to learn about reinforcement learning and WiFi networking, and have a reason to get out for more walks.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 14
    Stable Baselines3

    Stable Baselines3

    PyTorch version of Stable Baselines

    Stable Baselines3 (SB3) is a set of reliable implementations of reinforcement learning algorithms in PyTorch. It is the next major version of Stable Baselines. You can read a detailed presentation of Stable Baselines3 in the v1.0 blog post or our JMLR paper. These algorithms will make it easier for the research community and industry to replicate, refine, and identify new ideas, and will create good baselines to build projects on top of. We expect these tools will be used as a base around which new ideas can be added, and as a tool for comparing a new approach against existing ones. We also hope that the simplicity of these tools will allow beginners to experiment with a more advanced toolset, without being buried in implementation details.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 15
    TorchRL

    TorchRL

    A modular, primitive-first, python-first PyTorch library

    TorchRL is an open-source Reinforcement Learning (RL) library for PyTorch. TorchRL provides PyTorch and python-first, low and high-level abstractions for RL that are intended to be efficient, modular, documented, and properly tested. The code is aimed at supporting research in RL. Most of it is written in Python in a highly modular way, such that researchers can easily swap components, transform them, or write new ones with little effort.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 16
    TradeMaster

    TradeMaster

    TradeMaster is an open-source platform for quantitative trading

    TradeMaster is a first-of-its-kind, best-in-class open-source platform for quantitative trading (QT) empowered by reinforcement learning (RL), which covers the full pipeline for the design, implementation, evaluation and deployment of RL-based algorithms. TradeMaster is composed of 6 key modules: 1) multi-modality market data of different financial assets at multiple granularities; 2) whole data preprocessing pipeline; 3) a series of high-fidelity data-driven market simulators for mainstream QT tasks; 4) efficient implementations of over 13 novel RL-based trading algorithms; 5) systematic evaluation toolkits with 6 axes and 17 measures; 6) different interfaces for interdisciplinary users.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 17
    This project provides a framework for testing and comparing different machine learning algorithms (particularly reinforcement learning methods) in different scenarios. Its intended area of application is in research and education.
    Downloads: 3 This Week
    Last Update:
    See Project
  • 18
    PIQLE is a Platform Implementing Q-LEarning (and other Reinforcement Learning) algorithms in JAVA. Version 2 is a major refactoring. The core data structures and algorithms are in piqle-coreVersion2. Examples are in piqle-examplesVersion2. A complete doc
    Downloads: 1 This Week
    Last Update:
    See Project
  • 19
    festival3os

    festival3os

    mods to the Festival sokoban solver to run on OSX + Win + linux

    Mods to the Festival sokoban solver that allow building on OSX, Linux, & Windows
    Downloads: 1 This Week
    Last Update:
    See Project
  • 20
    In this Project, We solved 8-puzzle problem, very famous problem in AI, by using reinformcemnt learning concepts.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21
    AgentUniverse

    AgentUniverse

    agentUniverse is a LLM multi-agent framework

    AgentUniverse is a multi-agent AI framework that enables coordination between multiple intelligent agents for complex task execution and automation.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22
    Alibi Explain

    Alibi Explain

    Algorithms for explaining machine learning models

    Alibi is a Python library aimed at machine learning model inspection and interpretation. The focus of the library is to provide high-quality implementations of black-box, white-box, local and global explanation methods for classification and regression models.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23
    Best-of Machine Learning with Python

    Best-of Machine Learning with Python

    A ranked list of awesome machine learning Python libraries

    This curated list contains 900 awesome open-source projects with a total of 3.3M stars grouped into 34 categories. All projects are ranked by a project-quality score, which is calculated based on various metrics automatically collected from GitHub and different package managers. If you like to add or update projects, feel free to open an issue, submit a pull request, or directly edit the projects.yaml. Contributions are very welcome! General-purpose machine learning and deep learning frameworks.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 24
    BindsNET

    BindsNET

    Simulation of spiking neural networks (SNNs) using PyTorch

    A Python package used for simulating spiking neural networks (SNNs) on CPUs or GPUs using PyTorch Tensor functionality. BindsNET is a spiking neural network simulation library geared towards the development of biologically inspired algorithms for machine learning. This package is used as part of ongoing research on applying SNNs to machine learning (ML) and reinforcement learning (RL) problems in the Biologically Inspired Neural & Dynamical Systems (BINDS) lab.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 25

    CLSquare

    Closed Loop Simulation System

    Closed Loop Simulation System (CLSquare) is an integrated architecture to train, test and compare reinforcement learning controllers on different plants. CLSquare provides simulated plants as well as interfaces to real plants.
    Downloads: 0 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • 2
  • 3
  • 4
  • Next

Open Source Reinforcement Learning Algorithms Guide

Open source reinforcement learning algorithms are machine learning methods that enable artificial intelligence systems to improve decision-making through repeated interaction with an environment. Instead of relying only on predefined rules or labeled datasets, these algorithms learn by receiving feedback based on the outcomes of their actions. Their open source nature allows organizations, researchers, and developers to inspect the underlying methods, adapt them for specialized use cases, and contribute improvements through collaborative development. As a result, they have become an important foundation for experimentation and innovation across a wide range of industries.

These algorithms are commonly used to solve sequential decision-making problems where an agent must determine the best action to maximize long-term rewards. They support applications involving robotics, autonomous systems, industrial automation, finance, gaming, logistics, healthcare, and scientific research. Many frameworks provide implementations of popular reinforcement learning approaches, making it easier to build, train, evaluate, and refine intelligent agents. Flexible deployment options also allow organizations to integrate reinforcement learning into research environments, cloud infrastructure, or on-premises environments.

As adoption continues to grow, open source reinforcement learning algorithms are benefiting from advances in computational performance, simulation environments, and scalable training methods. Businesses can experiment with different learning strategies while maintaining greater visibility into how models are developed and optimized. Access to community-driven improvements also helps accelerate innovation without being limited to proprietary approaches. For organizations exploring advanced artificial intelligence capabilities, these algorithms provide a flexible foundation for creating adaptive systems that continuously improve through experience.

Features of Open Source Reinforcement Learning Algorithms

  • Flexible training workflows: Supports custom environments, reward functions, and learning objectives for varied reinforcement learning tasks.
  • Multiple algorithm options: Includes value-based, policy-based, and actor-critic methods for different problem requirements.
  • Environment compatibility: Connects with simulation environments through standardized interfaces for consistent training and evaluation.
  • Hyperparameter configuration: Allows adjustment of learning rates, exploration settings, and optimization values to improve performance.
  • Model checkpointing: Saves training progress for recovery, comparison, and continued experimentation without restarting.
  • Parallel training support: Uses multiple environments simultaneously to accelerate data collection and improve learning efficiency.
  • Performance evaluation: Measures rewards, episode lengths, and other metrics to monitor training effectiveness over time.
  • Hardware acceleration: Takes advantage of modern processors and graphics hardware to reduce training duration.
  • Experiment tracking: Records configurations, outcomes, and performance metrics to simplify reproducibility and result comparison.

Types of Open Source Reinforcement Learning Algorithms

  • Value-based algorithms: Estimate action values to identify decisions that maximize long-term rewards in environments with discrete action spaces.
  • Policy-based algorithms: Learn decision-making policies directly, making them suitable for continuous or complex action environments.
  • Actor-critic algorithms: Combine value estimation and policy learning to improve training stability and learning efficiency.
  • Model-based algorithms: Build predictive environment models that support planning before selecting actions.
  • Model-free algorithms: Learn through repeated interactions without constructing an internal representation of the environment.
  • Offline reinforcement learning algorithms: Train using previously collected datasets instead of requiring continuous interaction with live environments.
  • Multi-agent reinforcement learning algorithms: Enable multiple intelligent agents to cooperate or compete while learning within shared environments.

Open Source Reinforcement Learning Algorithms Advantages

  • Encourages customization: Teams can adapt learning methods for specialized objectives without depending on closed development models.
  • Promotes transparency: Accessible source code helps users inspect decision logic, implementation details, and training workflows.
  • Supports innovation: Developers can extend existing frameworks and introduce new reinforcement learning techniques more efficiently.
  • Reduces licensing expenses: Organizations avoid recurring licensing fees while expanding research or production environments.
  • Improves flexibility: Solutions can operate across different infrastructures, deployment strategies, and hardware configurations.
  • Strengthens collaboration: Communities contribute improvements, documentation, and testing that enhance overall reliability.
  • Enables educational value: Students and researchers gain practical experience by examining real implementations and modifying algorithms.
  • Simplifies experimentation: Teams can compare approaches, adjust parameters, and validate performance using their own datasets.

Types of Users That Use Open Source Reinforcement Learning Algorithms

  • AI researchers: Evaluate learning methods, compare training approaches, and explore new reinforcement learning techniques for academic and commercial research.
  • Machine learning engineers: Build, test, and refine intelligent decision-making models for production environments and experimental projects.
  • Robotics developers: Train autonomous machines to improve navigation, movement, and task completion through repeated interactions.
  • Autonomous vehicle teams: Develop decision-making systems that adapt to changing road conditions and operational scenarios.
  • Game developers: Create adaptive characters, optimize gameplay balance, and improve non-player behaviors using reinforcement learning techniques.
  • Industrial automation teams: Enhance operational efficiency by training systems to improve manufacturing workflows and resource allocation.
  • Financial analysts: Develop decision-making models for portfolio optimization, trading strategies, and risk evaluation using historical and simulated data.
  • Healthcare researchers: Investigate treatment optimization, scheduling improvements, and medical decision support through reinforcement learning methods.

How Much Do Open Source Reinforcement Learning Algorithms Cost?

Open source reinforcement learning algorithms are generally available without licensing fees, making them an attractive option for researchers, developers, and organizations looking to reduce upfront costs. While the algorithms themselves can be downloaded and used at no cost, the overall expense depends on the computing resources required for training and deployment. Simple projects may run on standard hardware, but more advanced models often require powerful GPUs, cloud infrastructure, or distributed computing environments that can significantly increase operational costs.

Organizations should also account for expenses beyond infrastructure. Implementation, customization, integration with existing tools, ongoing maintenance, and technical expertise all contribute to the total cost of ownership. Teams without in-house machine learning experience may need to invest in training or consulting services to successfully deploy and optimize reinforcement learning solutions. Evaluating both infrastructure and labor costs provides a more accurate understanding of the long-term investment.

What Software Do Open Source Reinforcement Learning Algorithms Integrate With?

Open source reinforcement learning algorithms can integrate with machine learning platforms that manage model training, experimentation, and deployment. They also connect with data processing tools that prepare datasets, transform inputs, and organize training pipelines. Integration with simulation environments allows models to learn through repeated interactions before being used in real-world scenarios. Many organizations also combine these algorithms with analytics platforms to monitor performance, evaluate outcomes, and identify opportunities for improvement. Cloud infrastructure, container orchestration platforms, and workflow automation tools help streamline training, scaling, and deployment across different environments. In addition, reinforcement learning algorithms can work with robotics platforms, Internet of Things systems, gaming engines, and business applications that provide continuous feedback for decision-making tasks.

Trends Related to Open Source Reinforcement Learning Algorithms

  • More teams adopt reinforcement learning for robotics, simulation, and autonomous decision-making across diverse industries.
  • Improved scalability supports larger environments, faster training cycles, and increasingly complex learning objectives.
  • Better compatibility with machine learning frameworks simplifies deployment, testing, and ongoing model refinement.
  • Community collaboration accelerates feature development, documentation improvements, and broader algorithm validation.
  • Growing interest in multi-agent learning expands research into coordinated decision-making across dynamic environments.
  • Greater emphasis on efficiency reduces training costs while improving resource utilization and practical adoption.
  • Enhanced benchmarking encourages consistent evaluation methods, making performance comparisons more meaningful across different approaches.
  • Increasing focus on safety promotes responsible training techniques, reliable behavior, and stronger evaluation standards.

How Users Can Get Started With Open Source Reinforcement Learning Algorithms

Selecting the right open source reinforcement learning algorithms starts with identifying the problem you want to solve. Different algorithms perform better depending on whether the environment is discrete, continuous, deterministic, or highly unpredictable. Matching the algorithm to the task improves learning efficiency and overall performance.

Next, evaluate training requirements, scalability, and hardware compatibility. Some algorithms demand significant computing resources and long training times, while others are better suited for smaller datasets or limited infrastructure. Consider whether the algorithm supports distributed training, parallel processing, or acceleration through modern hardware.

Review documentation quality, community activity, and update frequency to ensure long-term usability. Strong documentation and active development can simplify implementation and troubleshooting. Also examine customization options, evaluation methods, integration capabilities, and licensing terms. Testing several algorithms with representative data and comparing accuracy, stability, convergence speed, and resource consumption will help identify the most suitable option for your objectives.