Inference-time dynamic planning budgets for TD-MPC2 with paired return and latency evaluation.
-
Updated
Aug 9, 2026 - Python
Inference-time dynamic planning budgets for TD-MPC2 with paired return and latency evaluation.
Reinforcement learning (PPO, TD-MPC2) and behavior cloning for closed-loop harvest control of a simulated Spirulina photobioreactor, with a curriculum-gated training pipeline and held-out validation.
To associate your repository with the td-mpc2 topic, visit your repo's landing page and select "manage topics."