Yeah. LLaMA-7B can already run on a phone, and mobile RAM/AI compute is scaling pretty well.
I dunno if it will be "common," as most companies will keep their models and tuning in the cloud. There needs to be a concerted community effort to run stuff locally.
1 comment
[ 2.6 ms ] story [ 22.0 ms ] threadI dunno if it will be "common," as most companies will keep their models and tuning in the cloud. There needs to be a concerted community effort to run stuff locally.