upvote
For readers wondering, OpenRouter isn’t capable of caching as effectively as DeepSeek is because they will, for instance, switch inference providers in the middle of a session.
reply
The way I use openrouter is I find a model/provider combination I like then pin all requests for that model to that single provider.
reply
If you disable all other providers but DeepSeek in your OpenRouter guardrails, is that effectively the same thing?
reply
How do you do this?
reply
provider: { order: ['deepinfra/turbo'], allowFallbacks: false, },

https://openrouter.ai/docs/guides/routing/provider-selection

reply
What pests said. And you can make a preset and pass "model": "@preset/deep-seek"
reply
Fireworks directly. At the time they were the best value of cost, speed, ZDR. They got slower on me though, but I think they are retooling, so maybe things have or will get better again. I think fireworks is primarily for when you want to do your own training on top, which I wasn't doing.
reply