Fair context. Still, those prices are at least somewhat subsidized by you selling your plaintext session data which is a non-starter for me. Especially when dealing with code or data that is under NDA and not mine to sell in exchange for cheaper tokens.
Also, when I want 1M context I have 118GB of usable vram on my Strix Halo, or I can combine 4 r9700s and have 128gb and can run 1M context models like laguna or deepseek, but in practice lots of smaller sessions is better for my workflow in most cases.
Qwen 3.8 27b is smart enough that all I want is to speed-max and paralell-max on that.
Sure privacy (or legality) considerations are valid, and depending on the subscription or the API you use, you may or may not get that.
But we are not talking about the same product anymore. Your $12k homelab does not provide the same product as a $20/month subscription (to say nothing of a $200/month one). And the trade-offs of the homely may be worth it (or necessary) for you, but it may not be worth it (or even be feasible) for others.
Well at the very least lets agree that if you are going to use cloud models, there are much better options than OpenAI and Anthropic prices for most people. One does not have to give Sam or Dario money.
Personally though, I would sooner trade my car for GPUs than let a third party be in control of the tools I use to do my job.
Comments
Fair context. Still, those prices are at least somewhat subsidized by you selling your plaintext session data which is a non-starter for me. Especially when dealing with code or data that is under NDA and not mine to sell in exchange for cheaper tokens.
Also, when I want 1M context I have 118GB of usable vram on my Strix Halo, or I can combine 4 r9700s and have 128gb and can run 1M context models like laguna or deepseek, but in practice lots of smaller sessions is better for my workflow in most cases.
Qwen 3.8 27b is smart enough that all I want is to speed-max and paralell-max on that.
Sure privacy (or legality) considerations are valid, and depending on the subscription or the API you use, you may or may not get that.
But we are not talking about the same product anymore. Your $12k homelab does not provide the same product as a $20/month subscription (to say nothing of a $200/month one). And the trade-offs of the homely may be worth it (or necessary) for you, but it may not be worth it (or even be feasible) for others.
Well at the very least lets agree that if you are going to use cloud models, there are much better options than OpenAI and Anthropic prices for most people. One does not have to give Sam or Dario money.
Personally though, I would sooner trade my car for GPUs than let a third party be in control of the tools I use to do my job.