바닥부터 배우는 강화학습

1.[바닥부터 배우는 강화학습] - Chapter 1. 강화 학습이란

post-thumbnail

2.[바닥부터 배우는 강화학습] - Chapter 2. 마르코프 결정 프로세스

post-thumbnail

3.[바닥부터 배우는 강화학습] - Chapter 3. 벨만 방정식

post-thumbnail

4.[바닥부터 배우는 강화학습] - Chapter 4. MDP를 알 때의 플래닝

post-thumbnail

5.[바닥부터 배우는 강화학습] - Chapter 5. MDP를 모를 때 밸류 평가하기

post-thumbnail

6.[바닥부터 배우는 강화학습] - Chapter 6. MDP를 모를 때 최고의 정책 찾기

post-thumbnail

7.[바닥부터 배우는 강화학습] - Chapter 7. Deep RL 첫걸음

post-thumbnail

8.[바닥부터 배우는 강화학습] - Chapter 8. 가치 기반 에이전트

post-thumbnail

9.[바닥부터 배우는 강화학습] - Chapter 9. 정책 기반 에이전트 1 (Policy Gradient)

post-thumbnail

10.[바닥부터 배우는 강화학습] - Chapter 9. 정책 기반 에이전트 2 (Actor-Critic)

post-thumbnail