Ask HN: If ChatGPT surpasses Google search, do data sources become the moat?
If the moat to building chatGPT-like apps is training on up-to-date primary source data from sites like reddit, and the algorithms quickly become copied and commodified (like stable diffusion etc did to dall-e), doesn’t all the value accrue to primary information sources like reddit? They can sell access to their information and ban/sue anyone who tries to scrape their data wholesale, and essentially determine the best AI?
2 comments
[ 5.8 ms ] story [ 23.8 ms ] thread