Agent

This repository will host all initial machine learning efforts applying the Neodroid platform.

Neodroid is developed with support from Research Council of Norway Grant #262900. (https://www.forskningsradet.no/prosjektbanken/#/project/NFR/262900)

Contents Of This Readme

Algorithms
Requirements
Usage
Results
- Target Point Estimator
- Perfect Information Navigator
Contributing
Other Components

Algorithms

REINFORCE (PG)
DQN
DDPG
PPO
TRPO, GA, EVO, IMITATION...

Algorithms Implemented

Deep Q Learning (DQN) _{^{(Mnih et al. 2013)}}
DQN with Fixed Q Targets _{^{(Mnih et al. 2013)}}
Double DQN (DDQN) _{^{(Hado van Hasselt et al. 2015)}}
DDQN with Prioritised Experience Replay _{^{(Schaul et al. 2016)}}
Dueling DDQN _{^{(Wang et al. 2016)}}
REINFORCE _{^{(Williams et al. 1992)}}
Deep Deterministic Policy Gradients (DDPG) _{^{(Lillicrap et al. 2016 )}}
Twin Delayed Deep Deterministic Policy Gradients (TD3) _{^{(Fujimoto et al. 2018)}}
Soft Actor-Critic (SAC & SAC-Discrete) _{^{(Haarnoja et al. 2018)}}
Asynchronous Advantage Actor Critic (A3C) _{^{(Mnih et al. 2016)}}
Syncrhonous Advantage Actor Critic (A2C)
Proximal Policy Optimisation (PPO) _{^{(Schulman et al. 2017)}}
DQN with Hindsight Experience Replay (DQN-HER) _{^{(Andrychowicz et al. 2018)}}
DDPG with Hindsight Experience Replay (DDPG-HER) _{^{(Andrychowicz et al. 2018 )}}
Hierarchical-DQN (h-DQN) _{^{(Kulkarni et al. 2016)}}
Stochastic NNs for Hierarchical Reinforcement Learning (SNN-HRL) _{^{(Florensa et al. 2017)}}
Diversity Is All You Need (DIAYN) _{^{(Eyensbach et al. 2018)}}

Environments Implemented

Bit Flipping Game _{^{(as described in Andrychowicz et al. 2018)}}
Four Rooms Game _{^{(as described in Sutton et al. 1998)}}
Long Corridor Game _{^{(as described in Kulkarni et al. 2016)}}
Ant-{Maze, Push, Fall} _{^{(as desribed in Nachum et al. 2018 and their accompanying code)}}

Requirements

pytorch
tqdm
Pillow
numpy
matplotlib
torchvision
torch
Neodroid
pynput

(Optional)

visdom
gym

To install these use the command:

pip3 install -r requirements.txt

Usage

Export python path to the repo root so we can use the utilities module

export PYTHONPATH=/path-to-repo/

For training a agent use:

python3 procedures/train_agent.py

For testing a trained agent use:

python3 procedures/test_agent.py

Results

Target Point Estimator

Using Depth, Segmentation And RGB images to estimate the location of target point in an environment.

REINFORCE (PG)

DQN

DDPG

PPO

GA, EVO, IMITATION...

Perfect Information Navigator

Has access to perfect location information about the obstructions and target in the environment, the objective is to navigate to the target with colliding with the obstructions.

REINFORCE (PG)

DQN

DDPG

PPO

GA, EVO, IMITATION...

Contributing

See guidelines for contributing here.

Licensing

This project is licensed under the Apache V2 License. See LICENSE for more information.

Citation

For citation you may use the following bibtex entry:

@misc{neodroid-agent,
  author = {Heider, Christian},
  title = {Neodroid Platform Agents},
  year = {2018},
  publisher = {GitHub},
  journal = {GitHub repository},
  howpublished = {\url{https://github.com/sintefneodroid/agent}},
}

Other Components Of the Neodroid Platform

neo
droid

Authors

Christian Heider Nielsen - cnheider

Here other contributors to this project are listed.

Name		Name	Last commit message	Last commit date
Latest commit History 232 Commits
.github		.github
benchmark/memory		benchmark/memory
docs		docs
neodroidagent		neodroidagent
requirements		requirements
samples		samples
scripts		scripts
tests		tests
.ascii		.ascii
.coveragerc		.coveragerc
.dccache		.dccache
.flake8		.flake8
.gitattributes		.gitattributes
.gitignore		.gitignore
.nojekyll		.nojekyll
.pre-commit-config.yaml		.pre-commit-config.yaml
.pyup.yml		.pyup.yml
.readthedocs.yaml		.readthedocs.yaml
.travis.yml		.travis.yml
KEYWORDS.md		KEYWORDS.md
LICENSE.md		LICENSE.md
MANIFEST.in		MANIFEST.in
README.md		README.md
_config.yml		_config.yml
environment.yaml		environment.yaml
pyproject.toml		pyproject.toml
pytest.ini		pytest.ini
requirements.txt		requirements.txt
setup.cfg		setup.cfg
setup.py		setup.py
tox.ini		tox.ini

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Agent

Contents Of This Readme

Algorithms

Algorithms Implemented

Environments Implemented

Requirements

Usage

Results

Target Point Estimator

REINFORCE (PG)

DQN

DDPG

PPO

GA, EVO, IMITATION...

Perfect Information Navigator

REINFORCE (PG)

DQN

DDPG

PPO

GA, EVO, IMITATION...

Contributing

Licensing

Citation

Other Components Of the Neodroid Platform

Authors

About

Releases

Sponsor this project

Packages

Contributors 3

Languages

License

sintefneodroid/agent

Folders and files

Latest commit

History

Repository files navigation

Agent

Contents Of This Readme

Algorithms

Algorithms Implemented

Environments Implemented

Requirements

Usage

Results

Target Point Estimator

GA, EVO, IMITATION...

Perfect Information Navigator

GA, EVO, IMITATION...

Contributing

Licensing

Citation

Other Components Of the Neodroid Platform

Authors

About

Topics

Resources

License

Code of conduct

Stars

Watchers

Forks

Sponsor this project

Languages