Skip to content

Comment on A human metaphor for evaluating AI capabilityparent

Comments

BTW, have you tried out your challenges with a LLM?

I have they whiff pretty hard.

getting quite good at code synthesis

There was a post yesterday about vibe coding basic. It pretty much reflects my experience with code for SBC's.

I run home assistant, it's got tons of sample code out there so lots for the LLM to have in its data set. Here they thrive.

It's a great tool with some known hard limits. It's great at spitting back language but it clearly lacks knowledge. It clearly lacks the reason to understand the transitive properties of things, leaving it lacking at the edges.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.