ml research, both brainstorming and running experiments
eg I'd throw a hypothesis at it in the evening, and overnight it would write the code, do a sanity check, start a run, monitor the metrics, identify and fix a bug, propose a new hypothesis based on the results, write the code and start the second experiment, etc
still not that good at generating ideas or drawing conclusions, but much better at criticism than opus
Comments
ml research, both brainstorming and running experiments
eg I'd throw a hypothesis at it in the evening, and overnight it would write the code, do a sanity check, start a run, monitor the metrics, identify and fix a bug, propose a new hypothesis based on the results, write the code and start the second experiment, etc
still not that good at generating ideas or drawing conclusions, but much better at criticism than opus
even the api price is not that expensive for this