I may be misreading, but I think the earlier post was about the LLM being used to edit (clean up) their writing.
It's quite different than just reading a lot of text from other authors and passively absorbing diction and style. Instead, it means getting detailed reinforcement signals in the context of your own active writing effort.
This can start an ironic reinforcement learning pattern where the writer gets tuned to write in the LLM's "corrected" form. Imagine a traditional professional writer, in a long-term partnership with a particular copy editor who gates their output.
Comments
I may be misreading, but I think the earlier post was about the LLM being used to edit (clean up) their writing.
It's quite different than just reading a lot of text from other authors and passively absorbing diction and style. Instead, it means getting detailed reinforcement signals in the context of your own active writing effort.
This can start an ironic reinforcement learning pattern where the writer gets tuned to write in the LLM's "corrected" form. Imagine a traditional professional writer, in a long-term partnership with a particular copy editor who gates their output.