Comment on Self-Adapting Language ModelsparentComments−kovek1yWhat if you can check if the user responds positively/negatively to the output, and then you train the LLM on the input it got and the output it produced?
Comments
What if you can check if the user responds positively/negatively to the output, and then you train the LLM on the input it got and the output it produced?