Skip to content

Comment on Nemotron-4-340B

Comments

Why does nvidia release models that compete with its customers businesses but don’t make any money for nvidia?

Are they commodotising their complements?

[commoditizing] their complements

That's exactly what this would be.

compete with its customers businesses

I suspect most of their business comes from a few massive corporate spenders, not a "long tail" of smaller businesses, so it seems like a questionable goal to disrupt those customers without a clear path to new customers. Then again, few have the resources to run this model, so I guess this just ensures that their big customers are all working with some floor in model size? Probably won't impact anything realistically.

Nvidia offers AI Enterprise suite with NeMo, NIMS and many other services and consultancy to enterprise customers. These customers than can either use any AI models or Nvidia models.

Nvidia has no intention to earn money on models but to offer foundation models and extending their SW products which require their HW platform.

Basically, just like CUDA costs you nothing, it costs you nothing to use Nvidia models. And since you're on it you might want to use Nvidia HW for better performance and then you might want security and get interested in Nvidia SW enterprise.

They target this model at generating synthetic data. Data is the lifeblood of LLM training; quality synthetic data means more training can occur which means more demand for GPUs.

The model is big enough that you need expensive Nvidia GPUs to run it effectively

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.