All work Reinforcement Learning

RL Playground

DeepMind-style RL from scratch in JAX

Pure JAX/Flax implementations of the RL algorithm ladder — DQN, A2C, and DDPG — with no Stable-Baselines or RLlib dependencies. Every update step is a JIT-compiled pure function. The pipeline includes multi-modal observation engines, Gymnasium and MuJoCo environments, and an ONNX export path targeting NVIDIA Jetson Orin and Thor for real-world robotics edge deployment via TensorRT.

Capabilities applied

Reinforcement Learning Robotics Edge AI ONNX Export

Technology

JAX Flax NNX Optax Gymnasium MuJoCo jax2onnx NVIDIA Jetson

Interested in similar work?

Start a conversation