upvote
Those tokens aren’t guaranteed (esp. with RE and security tasks - rooted my own TV last week, Claude crapped out on “cyber safety” grounds; but also no guarantees about the model served - providers can pull a switcheroo on weights or quantization at any moment, and new options may not work for you), and you’re throwing money at entities that aren’t aligned with your interests instead of entities who are interested in actually empowering you.
reply
> Claude crapped out on “cyber safety” grounds.

A $10/mo subscription to OpenCode Go would have done the job for you.

They have models like Kimi K3, Grok 4.6 , GLM-5.3, Mimo 2.6 Pro (launched today, already available) which are happy to follow your orders without accusing you of being a terrorist.

https://models.dev/providers/opencode/

reply
The token price isn't the only reason to run a model locally though. You can do additional training to specialize or remove censorship that may be a no-no per TOS with cloud GPUs.
reply
Cloud GPUs have ToS about purposes you’re allowed to crank numbers for?

I thought this only applies to LLM inference providers, but not raw GPU rentals.

reply
some GPU-rental providers have restrictions that aren't really enforcable, like bans on crypto mining
reply
$2200 for a 64GB VRAM machine if you are willing to do a bit of work.

* https://imgur.com/mpdorVJ

reply