I just got it a few mins ago
> much of the work is repetitive, but it comes with its edge cases for the repetitive stuff, just use copilot embedded in whatever editor you use. the edge cases are tricky, to actually avoid these the model would need…
can someone explain how his costs went to $1? he essentially just replaced GPT4 with a tuned variant of mixtral 8x7b which requires multiple GPUs to run. even if he quantized the model himself it would still need to pay…
his username checks out
Did you now?? I'll have you know that I wrote the full word2vec paper on a roll of shabby two-ply tissue paper during my time in a Taco Bell stall. Sadly, it was then used to mop up my dietary regrets and was…
This guy got me through my engineering degree
How did you come up with 40b for the memory? specifically, why 0.7 * total params?
lit
Bible study just got more lit
I just got it a few mins ago
> much of the work is repetitive, but it comes with its edge cases for the repetitive stuff, just use copilot embedded in whatever editor you use. the edge cases are tricky, to actually avoid these the model would need…
can someone explain how his costs went to $1? he essentially just replaced GPT4 with a tuned variant of mixtral 8x7b which requires multiple GPUs to run. even if he quantized the model himself it would still need to pay…
his username checks out
Did you now?? I'll have you know that I wrote the full word2vec paper on a roll of shabby two-ply tissue paper during my time in a Taco Bell stall. Sadly, it was then used to mop up my dietary regrets and was…
This guy got me through my engineering degree
How did you come up with 40b for the memory? specifically, why 0.7 * total params?
lit
Bible study just got more lit