TeslaCoding·45 minMembers
Speed-Limit RL Reward from Trajectory Samples
Members only
Implement a reward function that penalizes speed-limit violations from raw trajectory samples shaped `[batch, num_waypoint, 2]` sampled at 10 Hz, then reason about state-dependent speed limits.
MLE
RS
reinforcement-learning
numpy
simulation
medium
Frequency
Single report
Last asked
2026-02-03
Stage
tech-screen
Log in to continue reading the full content
Comments
Sign in to join the discussion
Loading...
