Comment on Browser Agent Benchmark: Comparing LLM models for web automationparentComments−djohnston7moYeah but then their own product might not score the highest.−pixel_popping7moExactly why I'm pointing it out, which feels a bit corrupt, but understandable.−djohnston7motbh i was a bit cranky yesterday - even if they are #2 on a legit benchmark that would be impressive
Comments
Yeah but then their own product might not score the highest.
Exactly why I'm pointing it out, which feels a bit corrupt, but understandable.
tbh i was a bit cranky yesterday - even if they are #2 on a legit benchmark that would be impressive