Comment on DeepSeek-V3Comments−deyiao1yThe benchmark results seem unrealistically good, but I'm not sure from which angles I should challenge them.−ai-christianson1yI think they're real. The model is performing better than claude-3-5-sonnet-20241022 on the claude leaderboard:https://aider.chat/docs/leaderboards/
Comments
The benchmark results seem unrealistically good, but I'm not sure from which angles I should challenge them.
I think they're real. The model is performing better than claude-3-5-sonnet-20241022 on the claude leaderboard:
https://aider.chat/docs/leaderboards/