Tesla logoTesla
Coding·45 minMembers

Speed-Limit RL Reward from Trajectory Samples

Members only

Implement a reward function that penalizes speed-limit violations from raw trajectory samples shaped `[batch, num_waypoint, 2]` sampled at 10 Hz, then reason about state-dependent speed limits.

MLE
RS
reinforcement-learning
numpy
simulation
medium
Frequency
Single report
Last asked
2026-02-03
Stage
tech-screen

Log in to continue reading the full content

Comments

Sign in to join the discussion
Loading...