Posted by sbochins 7 hours ago
If you hunt in the settings you can restrict your account to only use EU servers for inference... Which means you can't use a lot of the US frontier models, but you can use all the Chinese ones, albeit within EU GDPR, etc.
This to me is a good compromise between privacy and cost.
I realize this text is just slop but it never stops being a "real bargain" at any point.
And it's more like $200/mo for $4000+/mo in tokens. You can also buy additional subscriptions.
There's no sense in running local models or doing anything else as long as VCs (and soon the public markets) are willing to pay your bill.
At the end of the day, AI models are relatively small files that we run little CUDA programs on.
If you still need more tokens, odds that you're vibecoding unmaintainable throwaway trash.
No clue what y'all are doing, perhaps because I'm hobbying, and also I'm old and can perhaps do more of this by hand.
But I'm basically just doing what I did before, plus ollama self hosted and sometimes gemini and I feel like I'm going lightspeed beyond what I've ever done.
And I suppose this is still very fine-grained. I have it make a draft, then just have them fix/change it step by step?
I tried one of the bigger boys that can one-shot apps, which I guess is cool, but I'm finding it's just as hard to modify as if I just grabbed someone elses repo on github.
As usual, an extraordinary claim without an extraordinary evidence: https://stephen.bochinski.dev/apps/