"because GPT is good enough to convince users of the site that the answer holds merit, [the] signals the community typically use to determine the legitimacy of their peers’ contributions frequently fail to detect severe issues with GPT-generated answers. As a result, information that is objectively wrong makes its way onto the site."
Basically it is much faster to generate incorrect answers algorithmically than it is for humans to evaluate their accuracy and flag them.
Given that people already know that GPT derivatives are inaccurate and given that it is trivial to detect it (see https://huggingface.co/openai-detector/), I don't think it is true.
This is what happens next. The web becomes full of AI-generated stuff, be it art or text or whatever, and people scrape that and train models and it starts consuming its own shit and regurgitating the same kind of stuff. And then human-created content becomes valuable again, and people find new sources of inputs, etc and we go on..
And obviously platforms will be built for creators/artists/writers/likenesses/whatever to be used by AI (and they get paid for training data) if they’re not being built already.
They won't. It is easy to detect AI generated answers (link in the parent thread).
Also the algorithms for detecting fake facts (e. g. "politician_name was born at <random date>") also exist; for knowledge graphs like Google Knowledge Graph it is fully automated; if a website consistently outputs answers that contradict other sources, GKG increases penalty coefficients for all facts from this source.
Comments
"because GPT is good enough to convince users of the site that the answer holds merit, [the] signals the community typically use to determine the legitimacy of their peers’ contributions frequently fail to detect severe issues with GPT-generated answers. As a result, information that is objectively wrong makes its way onto the site."
Basically it is much faster to generate incorrect answers algorithmically than it is for humans to evaluate their accuracy and flag them.
Given that people already know that GPT derivatives are inaccurate and given that it is trivial to detect it (see https://huggingface.co/openai-detector/), I don't think it is true.
That detector doesn't work very well, it says text from ChatGPT is 99.98% Real (Prediction based on 148 tokens).
and what happens if incorrect AI generated answers are then fed back into the system?
This is what happens next. The web becomes full of AI-generated stuff, be it art or text or whatever, and people scrape that and train models and it starts consuming its own shit and regurgitating the same kind of stuff. And then human-created content becomes valuable again, and people find new sources of inputs, etc and we go on..
And obviously platforms will be built for creators/artists/writers/likenesses/whatever to be used by AI (and they get paid for training data) if they’re not being built already.
They won't. It is easy to detect AI generated answers (link in the parent thread).
Also the algorithms for detecting fake facts (e. g. "politician_name was born at <random date>") also exist; for knowledge graphs like Google Knowledge Graph it is fully automated; if a website consistently outputs answers that contradict other sources, GKG increases penalty coefficients for all facts from this source.