RL Is Bottlenecked by Inference. Scale It Independently (skypilot.ai) 11 points by alex000kim 1mo ago ↗ HN
[–] efiop 1mo ago ↗ what did gpu hours look like here? with 3 replicas for a 1.8x speedup, the cost tradeoff isn’t obvious.
2 comments
[ 4.0 ms ] story [ 15.6 ms ] thread