Skip to content

Comment on Show HN: GPU-Accelerated Inference Hostingparent

Comments

Well, I guess I know where I am going to host GPT-J-6B then. I don't think it is sustainable.

How are you planning to put a gpt whatever when the service clearly have a model size limit?!

The size limit is very close to allowing it (12GB vs 10GB). I imagine you can reduce it somewhat further and get it to fit.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.