engineer

My Git Set Up

The following is setting up for my os.

My os Set Up

The following is setting up for my os.

Vim Set Up

The following is setting up for my vim.

Back to Top ↑

papers

Vanilla policy gradient

The value-function approach on reinforcement learning has worked well in many applications. It obtains the policy by selecting the action in each state with highest estimated value iteratively.

Back to Top ↑

awesome site

Awesome site

The fllowing is the papers I am reading and have read.

Back to Top ↑