← Back to context

Comment by NateEag

5 years ago

Well... you can read it.

I might be wrong, but so far I think it's been pretty easy to tell if text came from a human or a deep-learning system.

Granted, that probably doesn't scale well.

That has been true up until very recently, but lately there has emerged an uncanny valley that has confused the distinction between "underpaid freelance (ESL) writer farm" and "shoddy but roughly convincing neural language model" so that you may be convinced the blogspam you may happen upon across the internet may just as easily be computer-generated as anything else.

  • If GPT-3 et al. are producing results on par with "underpaid freelance (ESL) writer farm" then wouldn’t the next step be to focus only on better writing? i.e. books, newspapers, magazines, etc.

    Obviously the corpus (perhaps 1 trillion to to 10 trillion words?) will be exhausted by a sufficiently large model so there’s an upper bound.