{"repo":"opendilab/DI-engine","free":true,"listed":false,"github":"https://github.com/opendilab/DI-engine","clone":"git clone https://github.com/opendilab/DI-engine.git","description":"OpenDILab Decision AI Engine. The Most Comprehensive Reinforcement Learning Framework B.P.","language":"Python","stars":3639,"topics":["reinforcement-learning","multiagent-reinforcement-learning","self-play","imitation-learning","inverse-reinforcement-learning","exploration-exploitation","distributed-system","python","impala","smac"],"license":"Apache-2.0","category":"machine-learning","readme_excerpt":"--- Updated on 2024.12.23 DI-engine-v0.5.3 Introduction to DI-engine Documentation 中文文档 Tutorials Feature Task & Middleware TreeTensor Roadmap DI-engine is a generalized decision intelligence engine for PyTorch and JAX. It provides python-first and asynchronous-native task and middleware abstractions, and modularly integrates several of the most important decision-making concepts: Env, Policy and Model. Based on the above mechanisms, DI-engine supports various deep reinforcement learning algorithms with superior performance, high efficiency, well-organized documentation and unittest: - Most basic DRL algorithms: such as DQN, Rainbow, PPO, TD3, SAC, R2D2, IMPALA - Multi-agent RL algorithms: such as QMIX, WQMIX, MAPPO, HAPPO, ACE - Imitation learning algorithms (BC/IRL/GAIL): such as GAIL, SQIL, Guided Cost Learning, Implicit BC - Offline RL algorithms: BCQ, CQL, TD3BC, Decision Transformer, EDAC, Diffuser, Decision Diffuser, SO2 - Model-based RL algorithms: SVG, STEVE, MBPO, DDPPO, DreamerV3 - Exploration algorithms: HER, RND, ICM, NGU - LLM + RL Algorithms: PPO-max, DPO, PromptPG, PromptAWR - Other algorithms: such as PER, PLR, PCGrad - MCTS + RL algorithms: AlphaZero, MuZero, please refer to LightZero - Generative Model + RL algorithms: Diffusion-QL, QGPO, SRPO, please refer to GenerativeRL DI-engine aims to standardize different Decision Intelligence environments and applications , supporting both academic research and prototype applications. Various training pipelines and ","default_branch":null,"files":null,"tree":[],"storefront":"/r/opendilab","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/opendilab/DI-engine/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}