Skip to content

Ask HN: If ChatGPT surpasses Google search, do data sources become the moat?

2 pointsbased692 comments
On HN

If the moat to building chatGPT-like apps is training on up-to-date primary source data from sites like reddit, and the algorithms quickly become copied and commodified (like stable diffusion etc did to dall-e), doesn’t all the value accrue to primary information sources like reddit? They can sell access to their information and ban/sue anyone who tries to scrape their data wholesale, and essentially determine the best AI?

Comments

Personally I like the idea of something like chatgpt replacing google search. But, and a huge but, I miss the sources its trained on. Are these copyrighted sources, who owns the rights to the content it spews out? I like to know whos content I am reading and or using (when opensource) to reference them when needed.

Surpass in what? The main point of search is still link retrieval which ChatGPT doesn't do.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.