Skip to content

What 1,200 Production Deployments Reveal About LLMOps in 2025

zenml.io
1 pointstrickvl1 comment
On HN

Comments

Author here. I work at ZenML, where we maintain the LLMOps Database (https://www.zenml.io/llmops-database) — a collection of production LLM case studies we've been cataloguing for a while now. This post summarises patterns from the latest 400+ entries (database just crossed 1,200 total).

The findings that surprised me most:

- Context engineering has become a distinct discipline — teams are treating the million-token window as a ceiling to stay under, not a feature to exploit

- Software engineering skills matter more than AI expertise for production success (durable execution, distributed systems, infrastructure work)

- The "wait for the next model" strategy keeps not working — teams shipping today are constraining models, not unleashing them

There's also a shorter TL;DR version if you prefer: https://www.zenml.io/blog/the-experimentation-phase-is-over-...

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.