← Haadhi Irfan

Deep Q-Learning in Continuous Control

Kyushu University · Advisor: Prof. Danilo Vasconcellos Vargas · February 2023 – March 2023

Prototype

?
?

A simplified 2-D re-creation of the Gymnasium Pusher task: a jointed arm pushes the object (white) onto the goal (red) from randomized spawn cells, trained live in your browser with tabular Q-learning over (fingertip, object) states. Tune the episode budget and reward weights, retrain, and compare runs below — the same reward-convergence analysis documented in the research.

Highlights

Technologies