Index of /internal/projects/Giancarlo/Reinforce Learning

[ICO]NameLast modifiedSizeDescription

[PARENTDIR]Parent Directory  -  
[TXT]Applications - David Silver.html2022-06-22 13:34 147K 
[DIR]Applications - David Silver_files/2022-10-13 10:16 -  
[   ]Chapter 2c in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:25 128K 
[   ]Chapter 2d in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:26 95K 
[   ]Chapter 2e in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:28 145K 
[   ]Chapter 2f in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:29 46K 
[   ]Chapter 2g in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:30 49K 
[   ]Chapter 2h in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:31 71K 
[   ]Chapter 2i in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:32 48K 
[   ]Chapter 2 in Reinforcement Learning2 An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:24 108K 
[   ]Chapter 2 in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:25 128K 
[   ]Chapter 4b in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:37 216K 
[   ]Chapter 4c in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:37 143K 
[   ]Chapter 4 in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:36 190K 
[   ]Chapter 5b in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:39 150K 
[   ]Chapter 5c in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:40 115K 
[   ]Chapter 5d in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:41 94K 
[   ]Chapter 5 in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:39 246K 
[   ]Chapter 6b in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:42 49K 
[   ]Chapter 6b in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto. copy.pdf2022-06-22 12:46 107K 
[   ]Chapter 6c in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:47 103K 
[   ]Chapter 6 in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:42 76K 
[   ]Chapter 6 in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto. copy.pdf2022-06-22 12:45 419K 
[   ]Chapter 7b in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:50 164K 
[   ]Chapter 7c in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:50 75K 
[   ]Chapter 7d in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:51 62K 
[   ]Chapter 7e in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:52 73K 
[   ]Chapter 7 in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:48 133K 
[   ]Chapter 8b in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:54 1.1M 
[   ]Chapter 8c in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:55 47K 
[   ]Chapter 8 in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:53 123K 
[   ]Chapter 9b in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 13:21 66K 
[   ]Chapter 9c in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 13:22 70K 
[   ]Chapter 9d in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 13:23 31K 
[   ]Chapter 9 in Reinforcement Learning An Introduction by Richard S. Sutton and Andrew G. Barto..pdf2022-06-22 12:56 69K 
[   ]DP.pdf2022-06-22 13:30 805K 
[TXT]Deep reinforcement learning - Wikipedia.html2022-06-22 13:28 206K 
[DIR]Deep reinforcement learning - Wikipedia_files/2022-10-13 10:16 -  
[   ]Easy21-Johannes.pdf2022-06-22 13:32 226K 
[   ]FA.pdf2022-06-22 13:31 1.9M 
[   ]GetTiles_Mex.C.pdf2022-06-22 12:54 32K 
[   ]GetTiles_Mex_Script.m.pdf2022-06-22 12:53 29K 
[TXT]Gym Documentation.html2022-06-22 13:36 28K 
[DIR]Gym Documentation_files/2022-10-13 10:16 -  
[   ]MC-TD.pdf2022-06-22 13:31 1.4M 
[   ]MDP.pdf2022-06-22 13:30 816K 
[   ]RLbook2020.pdf2022-06-22 12:20 70M 
[   ]R_learn_acq.m.pdf2022-06-22 12:47 34K 
[   ]R_learn_acq_Script.m.pdf2022-06-22 12:46 30K 
[VID]Reinforcement Learning in 3 Hours Full Course using Python.mp42022-06-22 13:44 312M 
[   ]XX.pdf2022-06-22 13:32 1.3M 
[   ]binary_bandit_exps.m.pdf2022-06-22 12:25 37K 
[   ]binary_bandit_exps_Script.m.pdf2022-06-22 12:25 27K 
[   ]blocking_mz_Script.m.pdf2022-06-22 13:21 32K 
[   ]cmpt_P_and_R.m.pdf2022-06-22 12:35 31K 
[   ]cmpt_arms_err.m.pdf2022-06-22 12:43 28K 
[   ]cmpt_bj_value_fn.m.pdf2022-06-22 12:38 33K 
[   ]control.pdf2022-06-22 13:31 1.4M 
[   ]determineReward.m.pdf2022-06-22 12:38 27K 
[   ]do_ex_9_1_exps.m.pdf2022-06-22 12:56 32K 
[   ]do_mnt_car_Exps.m.pdf2022-06-22 12:55 32K 
[   ]dyna.pdf2022-06-22 13:31 2.1M 
[   ]dynaQ_maze.m.pdf2022-06-22 12:55 39K 
[   ]dynaQ_maze_Script.m.pdf2022-06-22 12:56 32K 
[   ]dynaQplus_maze.m.pdf2022-06-22 13:20 40K 
[   ]dynaQplus_maze_Script.m.pdf2022-06-22 13:21 32K 
[   ]eg_6_2_learn.m.pdf2022-06-22 12:43 31K 
[   ]eg_7_5_Script.m.pdf2022-06-22 12:52 30K 
[   ]eg_7_5_episode.m.pdf2022-06-22 12:51 28K 
[   ]eg_7_5_learn_at.m.pdf2022-06-22 12:51 31K 
[   ]eg_7_5_learn_rt.m.pdf2022-06-22 12:52 30K 
[   ]eg_rw_batch_learn.m.pdf2022-06-22 12:43 35K 
[   ]ex_4_2_sys_solv.m.pdf2022-06-22 12:34 27K 
[   ]ex_4_5_Script.m.pdf2022-06-22 12:36 31K 
[   ]ex_4_5_policy_evaluation.m.pdf2022-06-22 12:36 33K 
[   ]ex_4_5_policy_improvement.m.pdf2022-06-22 12:36 34K 
[   ]ex_4_5_rhs_state_value_bellman.m.pdf2022-06-22 12:36 29K 
[   ]ex_5_4_Script.m.pdf2022-06-22 12:40 35K 
[   ]ex_9_4_dynaQplus.m.pdf2022-06-22 13:22 41K 
[   ]ex_9_4_dynaQplus_Script.m.pdf2022-06-22 13:22 34K 
[   ]exam-rl-answers.pdf2022-06-22 13:33 326K 
[   ]exam-rl-questions.pdf2022-06-22 13:33 292K 
[   ]exercise_2_5.m.pdf2022-06-22 12:25 34K 
[   ]exercise_2_7.m.pdf2022-06-22 12:28 34K 
[   ]exercise_2_7_Script.m.pdf2022-06-22 12:28 28K 
[   ]exercise_2_11.m.pdf2022-06-22 12:31 34K 
[   ]exercise_2_11_Script.m.pdf2022-06-22 12:31 27K 
[   ]gam_Script.m.pdf2022-06-22 12:37 32K 
[   ]gam_rhs_state_bellman.m.pdf2022-06-22 12:37 28K 
[   ]games.pdf2022-06-22 13:32 3.0M 
[   ]gen_rt_episode.m.pdf2022-06-22 12:40 35K 
[   ]get_ctg.m.pdf2022-06-22 12:54 28K 
[   ]gw_w_et.m.pdf2022-06-22 12:50 36K 
[   ]gw_w_et_Script.m.pdf2022-06-22 12:51 31K 
[   ]handValue.m.pdf2022-06-22 12:38 27K 
[   ]init_unif_policy.m.pdf2022-06-22 12:40 29K 
[   ]intro_RL.pdf2022-06-22 13:30 2.9M 
[   ]iter_poly_gw_inplace.m.pdf2022-06-22 12:34 35K 
[   ]iter_poly_gw_not_inplace.m.pdf2022-06-22 12:34 35K 
[   ]jcr_example.m.pdf2022-06-22 12:35 31K 
[   ]jcr_policy_evaluation.m.pdf2022-06-22 12:35 33K 
[   ]jcr_policy_improvement.m.pdf2022-06-22 12:35 33K 
[   ]jcr_rhs_state_value_bellman.m.pdf2022-06-22 12:35 29K 
[   ]learn_cw.m.pdf2022-06-22 12:46 40K 
[   ]learn_cw_Script.m.pdf2022-06-22 12:45 32K 
[   ]linAppFn.m.pdf2022-06-22 12:52 28K 
[   ]mcEstQ.m.pdf2022-06-22 12:41 29K 
[   ]mc_es_bj_Script.m.pdf2022-06-22 12:39 36K 
[   ]mk_arms_error_plt.m.pdf2022-06-22 12:42 28K 
[   ]mk_batch_arms_error_plt.m.pdf2022-06-22 12:42 28K 
[   ]mk_ex_9_1_mz.m.pdf2022-06-22 12:55 28K 
[   ]mk_ex_9_2_mz.m.pdf2022-06-22 12:56 28K 
[   ]mk_ex_9_3_mz.m.pdf2022-06-22 13:21 28K 
[   ]mk_fig_6_6.m.pdf2022-06-22 12:42 28K 
[   ]mk_rt.m.pdf2022-06-22 12:40 28K 
[   ]mnt_car_learn.m.pdf2022-06-22 12:54 36K 
[   ]mnt_car_learn_Script.m.pdf2022-06-22 12:53 29K 
[   ]n_armed_testbed.m.pdf2022-06-22 12:22 34K 
[   ]n_armed_testbed_softmax.m.pdf2022-06-22 12:23 34K 
[   ]next_state.m.pdf2022-06-22 12:55 29K 
[   ]nips-tutorial-policy-optimization-Schulman-Abbeel.pdf2022-06-22 13:34 44M 
[TXT]notation.tex2022-06-22 13:27 9.7K 
[   ]opt_initial_values.m.pdf2022-06-22 12:29 34K 
[   ]opt_initial_values_Script.m.pdf2022-06-22 12:29 27K 
[   ]persuit_method.m.pdf2022-06-22 12:32 35K 
[   ]persuit_method_Script.m.pdf2022-06-22 12:32 27K 
[   ]pg.pdf2022-06-22 13:31 1.8M 
[   ]plot_cw_policy.m.pdf2022-06-22 12:46 31K 
[   ]plot_gw_policy.m.pdf2022-06-22 12:45 32K 
[   ]plot_mz_policy.m.pdf2022-06-22 12:56 31K 
[   ]reinforcement_comparison_methods.m.pdf2022-06-22 12:30 34K 
[   ]reinforcement_comparison_methods_Script.m.pdf2022-06-22 12:30 28K 
[   ]ret_q_in_st.m.pdf2022-06-22 12:54 29K 
[   ]rr_action_bellman.m.pdf2022-06-22 12:33 32K 
[   ]rr_state_bellman.m.pdf2022-06-22 12:32 31K 
[   ]rt_pol_mod.m.pdf2022-06-22 12:41 29K 
[   ]run_all_gw_Script.m.pdf2022-06-22 12:43 27K 
[   ]rw_accumulating_vs_replacing_Script.m.pdf2022-06-22 12:51 31K 
[   ]rw_episode.m.pdf2022-06-22 12:48 29K 
[   ]rw_offline_ntd_learn.m.pdf2022-06-22 12:48 31K 
[   ]rw_offline_ntd_learn_Script.m.pdf2022-06-22 12:48 31K 
[   ]rw_offline_tdl_learn.m.pdf2022-06-22 12:49 32K 
[   ]rw_offline_tdl_learn_Script.m.pdf2022-06-22 12:49 31K 
[   ]rw_online_ntd_learn.m.pdf2022-06-22 12:48 32K 
[   ]rw_online_ntd_learn_Script.m.pdf2022-06-22 12:48 32K 
[   ]rw_online_tdl_learn.m.pdf2022-06-22 12:49 31K 
[   ]rw_online_tdl_learn_Script.m.pdf2022-06-22 12:49 31K 
[   ]rw_online_w_et.m.pdf2022-06-22 12:50 30K 
[   ]rw_online_w_et_Script.m.pdf2022-06-22 12:50 31K 
[   ]rw_online_w_replacing_traces.m.pdf2022-06-22 12:51 30K 
[   ]sample_discrete.m2022-06-22 12:21 962  
[   ]shufflecards.m.pdf2022-06-22 12:38 27K 
[TXT]slides - Google Drive.html2022-06-22 13:25 493K 
[DIR]slides - Google Drive_files/2024-01-30 12:28 -  
[   ]soft_policy_bj_Script.m.pdf2022-06-22 12:39 39K 
[   ]stateFromHand.m.pdf2022-06-22 12:39 28K 
[   ]stp_fn_approx_Script.m.pdf2022-06-22 12:53 32K 
[   ]targetF.m.pdf2022-06-22 12:53 27K 
[   ]tiles.C.pdf2022-06-22 12:54 35K 
[   ]tiles.h.pdf2022-06-22 12:54 26K 
[   ]velState2PosActions.m.pdf2022-06-22 12:41 30K 
[   ]wgw_w_kings.m.pdf2022-06-22 12:44 37K 
[   ]wgw_w_kings_Script.m.pdf2022-06-22 12:44 29K 
[   ]wgw_w_kings_n_wind.m.pdf2022-06-22 12:44 36K 
[   ]wgw_w_kings_n_wind_Script.m.pdf2022-06-22 12:44 30K 
[   ]wgw_w_stoch_wind.m.pdf2022-06-22 12:45 37K 
[   ]wgw_w_stoch_wind_Script.m.pdf2022-06-22 12:44 30K 
[   ]windy_gw.m.pdf2022-06-22 12:44 35K 
[   ]windy_gw_Script.m.pdf2022-06-22 12:43 30K 

Apache/2.4.52 (Ubuntu) Server at docs.mbari.org Port 80