3 comments

[ 3.9 ms ] story [ 18.3 ms ] thread
(comment deleted)
Looks really cool and useful. Seems like GPT-4o it's a lot better than 4.
BTW, how did you manage all of the throughput to the models and navigate the various throttling strategies for all the models you mentioned?