2-day longest streak
-
Popular-RL-Algorithms ★ PINNED
PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT-Opt, PointNet..
Jupyter Notebook ★ 1.4k 1y agoExplain → -
tensorlayer ★ PINNED ⑂
Deep Learning and Reinforcement Learning Library for Scientists
Python ★ 0 7y agoExplain → -
MARS ★ PINNED
MARS is shortened for Multi-Agent Research Studio, a library for mulit-agent reinforcement learning research.
Jupyter Notebook ★ 52 2y agoExplain → -
Consistency_Model_For_Reinforcement_Learning ★ PINNED
Official implementation for: Consistency Models as a Rich and Efficient Policy Class for Reinforcement Learning ICLR'24
Python ★ 27 1y agoExplain → -
Reinforcement_Learning_for_Traffic_Light_Control
Apply deep reinforcement learning methods including DQN, DDPG for traffic light control in simulation (discrete environment), to prove the 'Green Wave' phenomenon in intelligent traffic system.
Python ★ 89 7y agoExplain → -
QT_Opt
Q-network with cross-entropy (CE) method for reinforcement learning.
Jupyter Notebook ★ 58 7y agoExplain → -
Cascading-Decision-Tree
Open-source code for paper CDT: Cascading Decision Trees for Explainable Reinforcement Learning
Jupyter Notebook ★ 40 9mo agoExplain → -
Benchmark-Efficient-Reinforcement-Learning-with-Demonstrations
Benchmark present methods for efficient reinforcement learning. Methods include Reptile, MAML, Residual Policy, etc. RL algorithms include DDPG, PPO.
Python ★ 32 3y agoExplain → -
Robotic_Door_Opening_with_Tactile_Simulation
Official code (simulation part) for paper Sim-to-Real Transfer for Robotic Manipulation with Tactile Sensory Zihan Ding, Ya-Yen Tsai, Wang Wei Lee, Bidan Huang International Conference on Intelligent Robots and Systems (IROS) 2021
Python ★ 26 4y agoExplain → -
nash-dqn
Official code of Nash-DQN for paper: Nash-DQN algorithm for two-player zero-sum Markov games, details see our paper: A Deep Reinforcement Learning Approach for Finding Non-Exploitable Strategies in Two-Player Atari Games. Zihan Ding, Dijia Su, Qinghua Liu, Chi Jin
Python ★ 22 4y agoExplain → -
RL_RLBench
Reinforcement Learning for RLBench
Python ★ 6 6y agoExplain → -
shortcut-torch
No description.
Python ★ 3 1y agoExplain → -
Consistency-Trajectory-Model
No description.
Python ★ 3 2y agoExplain → -
On_board_FNN_qubit_discrimination
Sigle qubit state discrimination with machine learning method (neural networks); An on board implementation with vivado FPGA + ARM, for fast qubit discrimination and feedback control in real physics system.
VHDL ★ 3 7y agoExplain → -
marl_tf ⑂
A simple OpenAI Gym environment for single and multi-agent reinforcement learning
★ 3 5y agoExplain → -
RL-with-AutoEncoder-for-Learning-from-Image-Pixels
No description.
Python ★ 3 7y agoExplain → -
DiffusionWorldModel
Official implementation of "Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning" (arXiv:2402.03570)
Python ★ 2 2mo agoExplain → -
FinRL ⑂
FinRL: The first open-source project for financial reinforcement learning. Please star. 🔥
Jupyter Notebook ★ 2 2y agoExplain → -
franka_panda_robot_control
No description.
C ★ 2 5y agoExplain → -
robolite ⑂
This is a fork of Surreal Robotics Suite that enables support to domain randomization and inverse kinematics (IK).
Python ★ 2 4y agoExplain → -
Robosuite-Panda-IK
No description.
Python ★ 2 6y agoExplain → -
Meta-Learning-for-Reinforcement-Learning
No description.
Python ★ 2 7y agoExplain → -
UPESI
Code for paper Not Only Domain Randomization: Universal Policy with Embedding System Identification.
Jupyter Notebook ★ 2 4y agoExplain → -
PointNet_Landmarks_from_Image
No description.
Python ★ 2 7y agoExplain → -
webpage ⑂
No description.
HTML ★ 1 9d agoExplain → -
Open-Sora-Plan ⑂
This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.
★ 1 2y agoExplain → -
transformers-pytorch
No description.
Python ★ 1 11mo agoExplain → -
adversarial-robustness-toolbox
No description.
Python ★ 1 2y agoExplain → -
human_prior_games
No description.
Jupyter Notebook ★ 1 2y agoExplain → -
Machine_Learning_Basic_Algorithms-Pytorch
Pytorch based codes. Implementations of most popular algorithms at present.
Jupyter Notebook ★ 1 7y agoExplain → -
Glints-detection
No description.
C++ ★ 1 8y agoExplain → -
marl_torch
No description.
Jupyter Notebook ★ 1 5y agoExplain → -
ion-trap-tomography-experiment
No description.
Python ★ 1 8y agoExplain → -
Robot_Learning2
No description.
Python ★ 1 3y agoExplain → -
PyRep ⑂
A toolkit for robot learning research.
Python ★ 1 6y agoExplain → -
Robot_Learning
No description.
ASP ★ 1 7y agoExplain → -
Store
No description.
ASP ★ 1 3y agoExplain → -
multigame-dt ⑂
Implementation of Multi-Game Decision Transformers in PyTorch
★ 0 3y agoExplain → -
Flash-Loan-Arbitrage ⑂
No description.
★ 0 3y agoExplain → -
scholar_ai
No description.
Python ★ 0 4mo agoExplain → -
dollar
No description.
HTML ★ 0 1y agoExplain → -
Video-Eval-Metrics
No description.
Python ★ 0 1y agoExplain → -
AI-Agent-Comparison
No description.
★ 0 1y agoExplain → -
DexterousHands ⑂
This is a library that provides dual dexterous hand manipulation tasks through Isaac Gym
★ 0 3y agoExplain → -
VBench ⑂
[CVPR2024 Highlight] VBench - We Evaluate Video Generation
Python ★ 0 1y agoExplain → -
Open-Sora-Plan-v1.2
No description.
Python ★ 0 2y agoExplain → -
diffusion_imitation
No description.
Python ★ 0 2y agoExplain → -
gensim ⑂
Topic Modelling for Humans
Python ★ 0 7y agoExplain → -
street-fighter-ai ⑂
This is an AI agent for Street Fighter II Champion Edition.
Python ★ 0 3y agoExplain → -
Flash_Loans_V3 ⑂
Code to borrow as much { WETH, USDC, DAI, USDT } as you want from Aave and make an arbitrage transaction with Uniswap up to V3
★ 0 3y agoExplain → -
consistency_models ⑂
Unofficial Implementation of Consistency Models in pytorch
Python ★ 0 3y agoExplain → -
Diffusion-Policies-for-Offline-RL ⑂
No description.
Python ★ 0 3y agoExplain → -
Muesli-lunarlander ⑂
Muesli RL algorithm implementation (PyTorch) (LunarLander-v2)
★ 0 3y agoExplain → -
ElegantRL ⑂
Cloud-native Deep Reinforcement Learning. Please star. 🔥
Python ★ 0 3y agoExplain → -
ensemble-dqn
No description.
Python ★ 0 4y agoExplain → -
Robotics_Modules
A collection of basic modules used in robotics.
Jupyter Notebook ★ 0 4y agoExplain → -
cleanrl ⑂
High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)
★ 0 4y agoExplain → -
taichi ⑂
Productive & portable high-performance programming in Python.
★ 0 4y agoExplain → -
nn_pde
No description.
Jupyter Notebook ★ 0 4y agoExplain → -
openmlsys-zh ⑂
《机器学习系统:设计与实现》
TeX ★ 0 3y agoExplain → -
COS513_project
No description.
Jupyter Notebook ★ 0 4y agoExplain → -
python_autocomplete ⑂
A simple neural network for python autocompletion
★ 0 6y agoExplain → -
poker ⑂
🃏♠️♥️♦️♣️
★ 0 6y agoExplain → -
safety_rl
No description.
Jupyter Notebook ★ 0 4y agoExplain → -
StepNeverStop ⑂
No description.
★ 0 4y agoExplain → -
quantumiracle
No description.
★ 0 4y agoExplain → -
ai-deadlines ⑂
:alarm_clock: AI conference deadline countdowns
HTML ★ 0 7y agoExplain → -
cb-trading ⑂
Code to trade the financial markets using Contextual Bandits
★ 0 6y agoExplain → -
rocket-recycling ⑂
Rocket-recycling with Reinforcement Learning
★ 0 4y agoExplain → -
mdp ⑂
Make it easy to specify simple MDPs that are compatible with the OpenAI Gym.
★ 0 7y agoExplain → -
hpc_beginning_workshop ⑂
No description.
Shell ★ 0 4y agoExplain → -
introcmdline ⑂
No description.
★ 0 5y agoExplain → -
NJ_mvc_appointment
No description.
Python ★ 0 5y agoExplain → -
sphinx_rtd_theme ⑂
Sphinx theme for readthedocs.org
★ 0 5y agoExplain → -
pytorch-nfsp ⑂
Implementation of Deep Reinforcement Learning from Self-Play in Imperfect-Information Games (Heinrich and Silver, 2016)
Jupyter Notebook ★ 0 5y agoExplain → -
distribuuuu ⑂
The pure and clear PyTorch Distributed Training Framework.
★ 0 5y agoExplain → -
minimalRL ⑂
Implementations of basic RL algorithms with minimal lines of codes! (pytorch based)
★ 0 6y agoExplain → -
awesome-deep-rl ⑂
This project is for learning and researching on Deep RL. Maintained by University AI researchers.
★ 0 7y agoExplain → -
stas00 ⑂
No description.
★ 0 5y agoExplain → -
PRML ⑂
PRML algorithms implemented in Python
★ 0 6y agoExplain → -
Kalman-and-Bayesian-Filters-in-Python ⑂
Kalman Filter book using Jupyter Notebook. Focuses on building intuition and experience, not formal proofs. Includes Kalman filters,extended Kalman filters, unscented Kalman filters, particle filters, and more. All exercises include solutions.
★ 0 6y agoExplain → -
crypto-rl ⑂
Deep Reinforcement Learning toolkit: record and replay cryptocurrency limit order book data & train a DDQN agent
★ 0 6y agoExplain → -
Soft-Decision-Tree ⑂
Distilling a Neural Network Into a Soft Decision Tree., Nicholas Frosst, Geoffrey Hinton., 2017.
Jupyter Notebook ★ 0 6y agoExplain → -
neural-backed-decision-trees ⑂
Making decision trees competitive with neural networks on CIFAR10, CIFAR100, TinyImagenet200, Imagenet
★ 0 6y agoExplain → -
shap ⑂
A game theoretic approach to explain the output of any machine learning model.
★ 0 6y agoExplain → -
interpretable-ml-book ⑂
Book about interpretable machine learning
★ 0 6y agoExplain → -
d2l-zh ⑂
《动手学深度学习》:面向中文读者、能运行、可讨论。英文版即伯克利“深度学习导论”教材。
★ 0 6y agoExplain → -
TCN ⑂
Sequence modeling benchmarks and temporal convolutional networks
★ 0 6y agoExplain → -
MinAtar ⑂
No description.
★ 0 6y agoExplain → -
research-and-coding ⑂
研究资源列表 A curated list of research resources
★ 0 6y agoExplain → -
website ⑂
No description.
★ 0 6y agoExplain → -
Differentiable_Activation_Function
No description.
Jupyter Notebook ★ 0 6y agoExplain → -
awr ⑂
Implementation of advantage-weighted regression.
★ 0 6y agoExplain → -
bindsnet ⑂
Simulation of spiking neural networks (SNNs) using PyTorch.
Python ★ 0 7y agoExplain → -
rl_a3c_pytorch ⑂
A3C LSTM Atari with Pytorch plus A3G design
★ 0 7y agoExplain → -
Nips2017Learning2Run
No description.
Python ★ 0 6y agoExplain → -
arp ⑂
Autoregressive policies for continuous control reinforcement learning
★ 0 7y agoExplain → -
store3
No description.
Jupyter Notebook ★ 0 7y agoExplain → -
store2
No description.
Jupyter Notebook ★ 0 3y agoExplain → -
CirclePacking
No description.
Jupyter Notebook ★ 0 7y agoExplain → -
real_robots ⑂
Gym environments for Robots that learn to interact with the environment autonomously
Python ★ 0 7y agoExplain → -
huskarl ⑂
Deep Reinforcement Learning Framework + Algorithms
Python ★ 0 7y agoExplain → -
Vrep
No description.
Python ★ 0 7y agoExplain → -
D4PG ⑂
Tensorflow implementation of a Deep Distributed Distributional Deterministic Policy Gradients (D4PG) network, trained on OpenAI Gym environments.
Python ★ 0 7y agoExplain → -
neurips2019_disentanglement_challenge_starter_kit ⑂
Starter Kit for the NeurIPS 2019 Disentanglement Challenge
Python ★ 0 7y agoExplain → -
Evolutionary-Algorithm ⑂
Evolutionary Algorithm using Python
Python ★ 0 7y agoExplain → -
PyTorch-Tutorial ⑂
Build your neural network easy and fast
Jupyter Notebook ★ 0 7y agoExplain → -
tutorials ⑂
机器学习相关教程
Python ★ 0 7y agoExplain → -
PILCO ⑂
Bayesian Reinforcement Learning in Tensorflow
Python ★ 0 7y agoExplain → -
osim-rl ⑂
Reinforcement learning environments with musculoskeletal models
Python ★ 0 7y agoExplain → -
2017-learning-to-run ⑂
The Winning Solution for the Learning To Run Challenge 2017
Python ★ 0 8y agoExplain → -
bayesian_benchmarks ⑂
A community repository for benchmarking Bayesian methods
Jupyter Notebook ★ 0 7y agoExplain → -
marathon-envs ⑂
A set of high-dimensional continuous control environments for use with Unity ML-Agents Toolkit.
C# ★ 0 7y agoExplain → -
ptan ⑂
PyTorch Agent Net: reinforcement learning toolkit for pytorch
Python ★ 0 7y agoExplain → -
Deep-Reinforcement-Learning-Hands-On ⑂
Hands-on Deep Reinforcement Learning, published by Packt
Python ★ 0 7y agoExplain → -
sawyer_control ⑂
Python3 ROS Interface to Rethink Sawyer Robots with OpenAI Gym Compatibility
Python ★ 0 7y agoExplain → -
Photorealistic-Style-Transfer ⑂
High-Resolution Network for Photorealistic Style Transfer
Jupyter Notebook ★ 0 7y agoExplain → -
oyster ⑂
Implementation of Efficient Off-policy Meta-learning via Probabilistic Context Variables (PEARL)
Python ★ 0 7y agoExplain → -
sensenet ⑂
No description.
Python ★ 0 8y agoExplain → -
single-parameter-fit ⑂
Real numbers, data science and chaos: How to fit any dataset with a single parameter
Jupyter Notebook ★ 0 7y agoExplain → -
neural-mmo ⑂
Neural MMO - A Massively Multiagent Game Environment
Python ★ 0 7y agoExplain → -
RL-Adventure-2 ⑂
PyTorch0.4 implementation of: actor critic / proximal policy optimization / acer / ddpg / twin dueling ddpg / soft actor critic / generative adversarial imitation learning / hindsight experience replay
Jupyter Notebook ★ 0 8y agoExplain → -
spinningup ⑂
An educational resource to help anyone learn deep reinforcement learning.
Python ★ 0 7y agoExplain → -
softlearning ⑂
Softlearning is a reinforcement learning framework for training maximum entropy policies in continuous domains.
Python ★ 0 7y agoExplain → -
learning_note ⑂
A collection of my learning notes
C++ ★ 0 7y agoExplain → -
awesome-meta-learning ⑂
A curated list of Meta-Learning resources/papers.
★ 0 7y agoExplain → -
TD3 ⑂
PyTorch implementation of TD3 and DDPG for OpenAI gym tasks
Python ★ 0 7y agoExplain → -
BayesianOptimization ⑂
A Python implementation of global optimization with gaussian processes.
Python ★ 0 7y agoExplain → -
progressive_growing_of_gans ⑂
Progressive Growing of GANs for Improved Quality, Stability, and Variation
Python ★ 0 7y agoExplain → -
pysc2 ⑂
StarCraft II Learning Environment
Python ★ 0 7y agoExplain → -
batch-ppo ⑂
Efficient Batched Reinforcement Learning in TensorFlow
Python ★ 0 7y agoExplain → -
pytorch-a2c-ppo-acktr ⑂
PyTorch implementation of Advantage Actor Critic (A2C), Proximal Policy Optimization (PPO) and Scalable trust-region method for deep reinforcement learning using Kronecker-factored approximation (ACKTR).
Python ★ 0 7y agoExplain → -
deepcolor ⑂
Automatic coloring and shading of manga-style lineart, using Tensorflow + cGANs
Python ★ 0 8y agoExplain → -
openai-cartpole ⑂
random search, hill climbing, policy gradient
Python ★ 0 8y agoExplain → -
pytorch-maml ⑂
PyTorch implementation of MAML: https://arxiv.org/abs/1703.03400
Jupyter Notebook ★ 0 7y agoExplain → -
pytorch-maml-rl ⑂
Reinforcement Learning with Model-Agnostic Meta-Learning in Pytorch
Python ★ 0 7y agoExplain → -
maml_rl ⑂
Code for RL experiments in "Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks"
Python ★ 0 9y agoExplain → -
AI_hack_IC
No description.
Python ★ 0 7y agoExplain → -
RL-Adventure ⑂
Pytorch Implementation of DQN / DDQN / Prioritized replay/ noisy networks/ distributional values/ Rainbow/ hierarchical RL
Jupyter Notebook ★ 0 8y agoExplain → -
adversarial-autoencoder ⑂
Chainer implementation of adversarial autoencoder (AAE)
Python ★ 0 8y agoExplain → -
Tensor-GAN ⑂
Tensor GAN
Python ★ 0 7y agoExplain → -
IMCL-nips
No description.
Python ★ 0 7y agoExplain → -
dopamine ⑂
Dopamine is a research framework for fast prototyping of reinforcement learning algorithms.
Jupyter Notebook ★ 0 8y agoExplain → -
TensorFlow-Tutorials ⑂
TensorFlow Tutorials with YouTube Videos
Jupyter Notebook ★ 0 8y agoExplain → -
FPGA_feedforward-neural-network_for_qubit_discrimination
No description.
C++ ★ 0 8y agoExplain → -
convolutional_variational_autoencoder
No description.
Python ★ 0 8y agoExplain → -
VAE-CNN ⑂
Variational Auto-Encoder with Convolutional Neural Network
Python ★ 0 8y agoExplain → -
verilog_modules ⑂
verilog modules
Verilog ★ 0 8y agoExplain → -
ion-trap-tomography-simulation
No description.
Python ★ 0 8y agoExplain → -
gym ⑂
A toolkit for developing and comparing reinforcement learning algorithms.
Python ★ 0 8y agoExplain → -
generative-models ⑂
Collection of generative models, e.g. GAN, VAE in Pytorch and Tensorflow.
Python ★ 0 8y agoExplain → -
qubit-discrimination-paper
No description.
Python ★ 0 8y agoExplain → -
arxiv-sanity-preserver ⑂
Web interface for browsing, search and filtering recent arxiv submissions
Python ★ 0 8y agoExplain → -
tensorly ⑂
TensorLy: Tensor Learning in Python.
Python ★ 0 8y agoExplain → -
rnnlib ⑂
RNNLIB is a recurrent neural network library for sequence learning problems. Forked from Alex Graves work http://sourceforge.net/projects/rnnl/
C ★ 0 8y agoExplain → -
WorldModels ⑂
An implementation of the ideas from this paper https://arxiv.org/pdf/1803.10122.pdf
Jupyter Notebook ★ 0 8y agoExplain → -
qubit-discrimination
No description.
Python ★ 0 8y agoExplain → -
netket ⑂
Machine learning algorithms for many-body quantum systems
C++ ★ 0 8y agoExplain → -
Tensorflow-Tutorial ⑂
Tensorflow tutorial from basic to hard
Python ★ 0 8y agoExplain → -
TensorFlow-Examples ⑂
TensorFlow Tutorial and Examples for Beginners with Latest APIs
Jupyter Notebook ★ 0 8y agoExplain → -
DeepLearningTutorials ⑂
Deep Learning Tutorial notes and code. See the wiki for more info.
Python ★ 0 8y agoExplain → -
DeepLearning ⑂
Deep Learning (Python, C, C++, Java, Scala, Go)
Java ★ 0 8y agoExplain → -
website-autovisitor
No description.
Python ★ 0 8y agoExplain → -
reinforcement-learning ⑂
Implementation of Reinforcement Learning Algorithms. Python, OpenAI Gym, Tensorflow. Exercises and Solutions to accompany Sutton's Book and David Silver's course.
Jupyter Notebook ★ 0 8y agoExplain → -
models ⑂
Models and examples built with TensorFlow
Python ★ 0 8y agoExplain → -
Pupil-detection
No description.
C++ ★ 0 8y agoExplain → -
Neural-Network-in-Ion-Trap
No description.
Python ★ 0 8y agoExplain → -
ustcthesis ⑂
LaTeX template for USTC thesis v3.0
TeX ★ 0 8y agoExplain → -
git
No description.
Python ★ 0 8y agoExplain → -
TF-Tutorials ⑂
A collection of deep learning tutorials using Tensorflow and Python
Jupyter Notebook ★ 0 9y agoExplain → -
reinforcement-learning-an-introduction ⑂
Python implementation of Reinforcement Learning: An Introduction
Python ★ 0 9y agoExplain → -
thoughts
No description.
★ 0 8y agoExplain → -
rllab ⑂
rllab is a framework for developing and evaluating reinforcement learning algorithms, fully compatible with OpenAI Gym.
Python ★ 0 9y agoExplain → -
DeepRLHacks ⑂
Hacks for training RL systems from John Schulman's lecture at Deep RL Bootcamp (Aug 2017)
★ 0 9y agoExplain → -
Reinforcement-learning-with-tensorflow ⑂
Reinforcement learning tutorials
Python ★ 0 9y agoExplain →
No repos match these filters.