Ask HN: Will the coming GPT revolution require significantly more energy?
Just something I am thinking about, wondering if we can even know. If we completely replaced and/or augmented, e.g., web search with GPT(-like) models, would there be substantially more energy needed in order to sustain this? (Or less, or about the same?)
15 comments
[ 3.9 ms ] story [ 47.7 ms ] threadhttps://arxiv.org/abs/2303.06219
it will actually save energy.
As an example of the latter, I can answer questions out of my long term memory and GPT-4 does the same. I can answer more questions with more accurate results I’d I search for documents and use the capabilities encoded in my long term memory to evaluate, summarize and otherwise process those documents and no doubt that GPT-4 could be embedded in a system that does this as well and the large attention window would be a big help.
(E.g. a few papers a week are submitted to arXiv about this now)
One of the most fun things I did in grad school was when, for my A exam, I was asked to write a summary in the literature for giant magnetoresistance which involved looking at 50 or so papers. One would imagine it could take 50x the effort (queries, energy, etc.) for GPT-4 to do this than it would take to speak extemporaneously on the subject. (Funny I’d have a hard time talking about the physics w/o looking it up now but I could definitely talk about applications off the cuff.)
https://www.euronews.com/green/2020/02/17/is-playing-video-g...
and
https://gizmodo.com/downloaded-games-have-a-larger-carbon-fo...
You have to consider not just the server and client costs but how to attribute the cost of all the routers and networking parts between them that are invisible. Networking hardware can be particularly wasteful when it is sitting there waiting for a packet which has to be attributed somehow.
If someone in Google propose a new search method that cost more money (in average) than they money they get (in average), the proposal will be ignored.
Google can show better ads, but the user will escape and not see the adds in the second page, so Google compensate the money.
Google can show better ads, and the users will migrate from Bing to Google, and the advertisers too.
Think of that training a teacher take more than 10 years at minimum, and still they do not perform well.
In America, 10 years is quite alot of schooling if we are just talking about K-12 education.
And regardless, this is just one of the many ways we will be firing up these GPU arrays..