Skip to content

GPT-2 implementation in Modular MAX

github.com/tommyip
2 pointsred2awn1 comment
On HN

Comments

I am learning to write LLM pipelines using the Modular MAX inference framework. As a starting point I got GPT-2 working after reading through "The Illustrated GPT-2", Karpathy's nanoGPT codebase and existing models in the Modular repo. The MAX framework does require a lot of boilerplates and not designed to be very flexible, but you do gain awesome performance out of the box.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.