Skip to content

Comment on Intel Gaudi2 chips outperform Nvidia H100 on diffusion transformers

Comments

Has anyone been running LLMs on TPUs in prod? Curious to hear experiences.

Yeah they train well and very stably even int8, maxtext now has LLaMA and mistral support too, pytorch xla gets 50% MFU with spmd and you have some nice stacks like levanter

Haven't been too impressed with inference versus tensor rt llm for example though

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.