Commit Graph

  • 149b8f4bd8 Merge branch 'master' of https://github.com/kono94/refo master kono 2020-04-19 20:56:21 +02:00
  • 19d5a87ce0 add multiple food scenario kono 2020-04-19 20:55:42 +02:00
  • 4590562a4c add new every visit results, fix rngEnv NullPointer kono 2020-04-19 19:13:38 +02:00
  • 7d3d097599 add opening dialog to select all learning settings kono 2020-04-07 11:03:17 +02:00
  • 9d1f8dfd46 apply code improvements suggested by intelliJ kono 2020-04-05 14:44:48 +02:00
  • 7de2a5d1af spawn start field of antGame in the same spot everytime DinoSampling kono 2020-04-05 14:10:47 +02:00
  • 94ad976a1f spawn start of antgame constant kono 2020-04-05 14:07:24 +02:00
  • ff6807dabd spawn start of antgame constant antWorldRewardAnalysis kono 2020-04-05 14:07:24 +02:00
  • bbccef1e71 removed unnecessary stuff from sampling branches kono 2020-04-05 13:37:38 +02:00
  • 0300f3b1fd Merge branch 'antWorldRewardAnalysis' kono 2020-04-05 13:21:20 +02:00
  • ad07c1da8f remove DinoSampling stuff kono 2020-04-05 13:10:13 +02:00
  • 737d78c6da mend kono 2020-04-05 12:54:14 +02:00
  • e8f4fa06b6 add specific environment RNG kono 2020-04-05 12:52:49 +02:00
  • 5b82e7965d rename MC class and improve specific analysis of antGame examples kono 2020-04-05 12:29:44 +02:00
  • 42dfebb048 Merge branch 'epsilonBehavior' into DinoSampling kono 2020-04-05 12:07:39 +02:00
  • 4402d70467 Merge remote-tracking branch 'origin/antWorldRewardAnalysis' into antWorldRewardAnalysis kono 2020-04-05 12:03:23 +02:00
  • b9be640284 add multiple folders to organize results kono 2020-04-05 12:00:16 +02:00
  • a08b8160a3 add new results of needed timestamps in total kono 2020-04-04 17:06:15 +02:00
  • 595451e88b add new results of needed timestamps in total kono 2020-04-04 17:06:15 +02:00
  • a40e279f48 change reward function for antgame to match BA kono 2020-04-04 14:41:58 +02:00
  • 3bdcbb39bc reset DinoSampling to advanced Every Visit kono 2020-04-02 18:45:36 +02:00
  • 9a3452ff9c add Every-Visit Monte-Carlo kono 2020-04-02 15:56:11 +02:00
  • b0ca634b64 add every visit no jump results kono 2020-04-02 17:07:15 +02:00
  • f2aa7487af Merge remote-tracking branch 'origin/epsilonBehavior' into epsilonBehavior kono 2020-04-02 15:57:32 +02:00
  • 6477251545 add Every-Visit Monte-Carlo kono 2020-04-02 15:56:11 +02:00
  • 740289ee2b add constant for default reward kono 2020-04-02 14:01:37 +02:00
  • e7404a8d24 add improved result graphs kono 2020-03-31 17:49:15 +02:00
  • 0fde1bd962 Merge remote-tracking branch 'origin/antWorldRewardAnalysis' into antWorldRewardAnalysis kono 2020-03-29 17:22:56 +02:00
  • f4b50627d1 add antGame analysis data and R Scripts and images kono 2020-03-29 17:22:01 +02:00
  • 78955a9521 add antGame analysis data and R Scripts and images kono 2020-03-29 17:22:01 +02:00
  • 328fc85214 modify q Learning to sample results and update R script kono 2020-03-28 12:35:33 +01:00
  • 28c40c58dd add updated R script kono 2020-03-27 17:09:03 +01:00
  • eca0d8db4d create Dino Sampling state kono 2020-03-26 19:22:50 +01:00
  • 58f9900f3c Delete con.txt kono 2020-03-17 18:33:54 +01:00
  • ee1d62842d split Antworld into episodic and continuous task kono 2020-03-15 16:58:53 +01:00
  • 4641f50b79 add results for convergence for advanced dino jumping kono 2020-03-05 13:17:54 +01:00
  • b1d06293fe add shadowJar kono 2020-03-05 12:25:42 +01:00
  • 1f743cf8f2 fix eps/sec stat kono 2020-03-05 12:09:36 +01:00
  • e67f40ad65 split DinoWorld between simple and advanced example kono 2020-03-05 11:58:57 +01:00
  • 18d6e32f64 split DinoWorld between simple and advanced example kono 2020-03-05 11:58:57 +01:00
  • cffec63dc6 apply threading changes to master branch and clean up for tag version kono 2020-03-05 11:49:51 +01:00
  • 9b54b72a25 add epsilon convergence test and will remove unnecessary multithreaded learning kono 2020-03-03 02:52:39 +01:00
  • 6613e23c7c Fixed new method name for MC kono 2020-03-02 23:19:54 +01:00
  • 33f896ff40 Merge remote-tracking branch 'origin/epsilonTest' kono 2020-03-02 23:10:01 +01:00
  • 18a702ba62 add BlackJack environment and fix save and load kono 2020-03-01 13:51:47 +01:00
  • 0e4f52a48e first epsilon decaying method kono 2020-02-27 15:29:15 +01:00
  • cff1a4e531 add isJumping info to dinoState kono 2020-02-26 17:14:28 +01:00
  • 77898f4e5a add TD algorithms and started adopting to continous tasks kono 2020-02-17 13:56:55 +01:00
  • f4f1f7bd37 add QTableFrame and clickable states that display a gui kono 2020-01-01 23:54:18 +01:00
  • a8f8af1102 add gradle wrapper and jar building kono 2020-01-01 18:58:25 +01:00
  • 295a1f8af0 remove javaFX dependency in favour of org.javaTuples kono 2020-01-01 18:25:22 +01:00
  • b7d991cc92 render 5 frames for every RL step kono 2020-01-01 18:05:59 +01:00
  • ec86006a07 enhance hashCode and equals methods kono 2020-01-01 14:57:08 +01:00
  • 518683b676 split GUI parts from controller into sub class kono 2019-12-31 14:43:40 +01:00
  • 195722e98f enhance save/load feature and change thread handling kono 2019-12-29 01:12:11 +01:00
  • 64355e0b93 add javadoc kono 2019-12-27 00:50:59 +01:00
  • b2c3854b3a change RL-Controller initialization process and action space iterable kono 2019-12-24 19:38:35 +01:00
  • 5a4e380faf add dino jumping environment, deterministic/reproducable behaviour and save-and-load feature kono 2019-12-22 23:33:56 +01:00
  • b1246f62cc add features to gui to control learning and moving learning listener interface to controller kono 2019-12-22 17:06:54 +01:00
  • 34e7e3fdd6 distinguish learning and episodic learning, enable fast-learning without drawing every step to reduce lag kono 2019-12-21 00:23:09 +01:00
  • 7db5a2af3b add fix RNG, add extended interface EpsilonPolicy and move rewardHistory to model instead of view kono 2019-12-20 16:51:09 +01:00
  • e0160ca1df adopt MVC pattern and add real time graph interface kono 2019-12-18 16:48:24 +01:00
  • 7f18a66e98 add random policy test kono 2019-12-12 00:20:38 +01:00
  • 584d6a1246 add javaFX gradle plugin and switch to java11 and add system.outs for error detecting kono 2019-12-10 15:37:20 +01:00
  • 55d8bbf5dc add Random-, Greedy and EGreedy-Policy and first implementation of monte carlo method kono 2019-12-09 23:21:48 +01:00
  • 0100f2e82a remove the Action interface in favour of Enums kono 2019-12-09 17:30:14 +01:00
  • 8a533dda94 change ActionSpace interface temporarily to quickly fit antWorld test and improve gui of walking ant kono 2019-12-09 13:41:00 +01:00
  • 2fb218a129 add separate class for intern Ant representation and adopt gui cell size to panel size kono 2019-12-09 12:08:53 +01:00
  • c11cc2c3f2 add two simple scroll panes to represent environment and ant brain kono 2019-12-09 01:09:34 +01:00
  • db9b62236c add logic to handle ant action and compute rewards kono 2019-12-08 16:03:00 +01:00
  • ec67ce60c9 add default structure for AntAgent kono 2019-12-08 13:15:20 +01:00
  • 581cf6b28b Merge remote-tracking branch 'origin/master' kono 2019-12-07 22:08:20 +01:00
  • 87f435c65a add basic core structure and first parts of antGame implementation kono 2019-12-07 17:31:30 +01:00
  • 431ae4d3df add basic core structure and first parts of antGame implementation kono 2019-12-07 17:31:30 +01:00
  • 66ee33b77f init the gradle project kono 2019-12-06 13:11:29 +01:00