블로그
- 강화학습 ) 마르코프 결정 프로세스(MDP) Markov Decision Process 정책(policy) 가치(value) 상태 가치 함수(State value function) 벨만 수식 (Bellman equation) 상태-행동 가치함수(state-action value function) 할인 누적 보상액(discounted accumulating reward)
- 청소년에 대한 처벌을 강화 시킬까 - 반대 (Should the punishment for youth be more powerful and stronger - Disagree) no point of strengthening punishment because the youth detention center has a loss of preventive function due to their mistakes and strengthening punishment, it is the act of stigmatizing the youth by the state
- [Visual Studio] .pdb .map msdn.microsoft.com/en-us/library/yd4f8bd1(v=vs.71).aspx A program database (PDB) file holds debugging and project state 파일을 만들때마다 컴파일러는 VC70.PDB(VCx0.PDB) 파일에 디버그 정보를 덮어 넣는다. type inforamtion without symbol information(function
← 이전
2 페이지