Andrew White 🐦‍⬛ @andrew.diffuse.one · Dec 31

A lot of effort in this work was framing the learning problem of agents. We settled on defining agents using stochastic compute graphs and splitting the environment and agent according to what we want to train. Here are some components of well-known agents as compute graphs:

1 likes 1 replies

?

Replies

Andrew White 🐦‍⬛ · Dec 31

The environments in Aviary truly require multiple steps of cycles of observation and action. Here you can see multiple trajectories of how the agent solves problems differently than its demonstrations