Skip to content
Gyanateet Dutta
Work·Academic·CV

Work

Vision

Deep RL & Hugging Face

This repository contains PPO, DQN, and related implementations from the Hugging Face Deep RL course, including experiments in Doom and PyBullet.

A PPO agent playing Doom: the view swings across a canyon while the HUD tracks health and ammunition.
Fig. 1The PPO agent in ViZDoom, eight seconds of one rollout. One recorded rollout from a course exercise in ViZDoom.
A PyBullet Ant robot with four coloured legs walking across a blue and white checkered floor.
Fig. 2The PyBullet Ant, five seconds of one rollout. One recorded rollout from a course exercise, cropped and slowed to 8 fps.
Scope

Course exercises using established algorithms.

Stack
Deep RL, PPO, DQN, Hugging Face
Links
Code

Also

  • 2026MVA Rare Disease Hackathon 2026
  • 2026Causal-JEPA reproduction
  • 2026GOT-JEPA surgical tool tracking
  • 2025AIMS: Surgical Phase Detection
  • 2025MSc Thesis: Surgical Video Prediction
GitHub· ORCID· Google Scholar· LinkedIn· Hugging Face· X