upvote
One of the reasons I use OpenRouter is because they offer zero data retention. As far as I can tell, DeepSeek's own API doesn't support ZDR.
reply
AFAIK, when you use DeepSeek via OpenRouter, it still does not have zero data retention.

See here:

https://openrouter.ai/providers/

reply
I have ZDR enforced and see only compatible models and providers, yet am able to use it. DeepSeek as a provider may not be ZDR, but the models are available from ZDR and no training providers on EU/US servers.
reply
DeepInfra does and it's the same price. That's what I use.
reply
DeepInfra is decent but has a bad history of quantizing models
reply
Yeah, they all do it though. And now that models are being natively trained at FP8/NVFP4, I’m not sure it matters.

Only way you can really know you’re getting the full model is to host it yourself. Every inference provider has every reason to lie and it’s impossible to find out the degree to which they are.

reply
Youre doing something so special you need that?
reply
It’s a pretty common requirement in the enterprise world. If you’re processing data for enterprise customers, it’s a lot easier to retain nothing than to deal with all the compliance issues that arise if you’re retaining data.
reply
Ugh yeah try to get through a DPIA review!
reply
I've got a toilet cam to install in your bathroom
reply
I'm all in for saving money and _can_ move to using DS directly from them, but maybe I am missing something here:

OpenRouter Pricing:

$0.02/M input tokens $0.60/M output tokens

DeepSeek Pricing (cache miss, off-peak):

$0.15/M Input $0.60/m output

reply
When 98.5% of my requests are cache hits (according to Pi for the last week), the cache miss price isn’t that important to me, and $0.003-0.006 per 1M input tokens is shockingly cheap.

It’s also the major difference between using DeepSeek directly vs other providers also serving it, though I have not looked lately: it’s possible other providers have matched its cache hit pricing better?

reply
Interesting, if the cache hit is that good, I think HN convinced me to toss $20 at DS official, and see how long that lasts.
reply
Read the tech report and see how laser-focused they've been on compressing the disk KV footprint specifically. 890 bytes/token is absurd, and is probably the reason I regularly get ~0.5 Mtok request full cache hits after an hour. On their end that's just yanking a 414 MiB file off an SSD, then doing a little bit of compute (bounded SWA replay). Nobody else seems to be serving their model as well as they do.
reply
It will of course depend on what you’re doing with it, but right now my session at work has a 99.8% cache hit rate, and I’ve been running this session for hours with 23M tokens read and 713K tokens written (Opus 5.5 in this case though)
reply
cache hit is in fact that good.
reply
I've heard that certain inference providers may have different quality of caching implementations, so even if the listed numbers are as you say, the practical cache hit % you get might be significantly different/incur significantly different costs.
reply
There’s a big difference in speed & quality between using DeepSeek API directly with DSH vs. DeepSeek in Opencode Go with Opencode CLI. Can’t tell if it’s the provider or the harness - but worth to give it a try.
reply
what about dsh + openrouter? and configuring open router to just serve from DeepSeek own servers

I think the 5% cut from open router is fair if I want to user other cheap models like mimo

reply
DeepSeek trains on your inputs. That's why people go on OpenRouter and choose ZDR providers.
reply
This isn't true at least on the API. If you read their privacy policy you'll see the training clause is scoped specifically to the consumer terms i.e. for the chat product. No such clause exists for the API service, and it would absolutely be required under Chinese law if it was taking place.

Contrary to popular belief, DeepSeek really aren't interested in your prompts.

reply
Do you have a link to their API privacy policy?

What I see is https://cdn.deepseek.com/policies/en-US/deepseek-privacy-pol...

They do not have a specific exclusion for API use.

I know Z.ai has an exclusion for API use. It's widely reported Deepseek doesn't.

reply
I'll preface by saying you'd be entirely reasonable in not finding this sufficiently reassuring, but compare the standard terms: https://cdn.deepseek.com/policies/en-US/deepseek-terms-of-us...

To the open platform terms (i.e. for API use): https://cdn.deepseek.com/policies/en-US/deepseek-open-platfo...

The standard terms includes clause 4.3 which grants them the right to retain inputs and outputs for training purposes, and this is missing from their API terms. The standard terms also cover the right to opt out (which you can do from your user settings). No such opt-out exists on the API because it isn't applicable.

reply
You're linking to the Terms of Use. I'm linking to the Privacy Policy. Your Terms of Use link points to the Privacy Policy.
reply
Deepseek does not offer zero training on OpenRouter. Out of dozens of alternatives, they are the only provider for 4.1 Flash that does this.
reply
Let me get this straight you guys really like deep seek because it’s open but you don’t wanna help them improve.
reply
No, we just want a choice on how to license our work.
reply
All these products, western or Chinese, are built on a hell of a lot of running rough-shod over licensing or IP laws in general.

I'm writing a program I personally need, but I would be happy if there existed something like it already, if someone else vibecoded a better version of it than mine, or if DS got better at vibing this kind of thing.

reply
>license our work

what proof do you have, that they don't train on your data?

they can say they don't, but I don't see any way for you to confirm it.

with how these companies operate currently, I won't be surprised, if they say that one of agents "mistakenly" did that already..

reply
Mistakenly and autonomously of course.
reply
Why does liking a product mean you have to give them all of your data? People are so outraged at LG because they make the best TVs and people wanted their expensive product, yet some MBA convinced them they could make more money by spying on your entire household all the time.

It actually was awesome in the early Facebook days where you could have your entire phone contacts and other apps filled out with a profile picture and Birthday by connecting them together. But that relationship has been completely abused, privacy has been invaded, and my data has been sold to multiple companies.

The goal going forward is to keep that data private. If your company can't survive without it then I hope your company goes out of business

reply
Just pin your config to a single provider, or several providers with the params `order` and `allow_fallbacks: false`. I regularly get ~98-99% cache hit rates with OpenCode. And some providers are much faster than DeepSeek; I was getting 200-300 tokens/second the other day with Together as my provider.

It's regrettable that OpenRouter doesn't even try to pin you to a single provider per session, but once you know about it, it's a problem that's easily solved.

reply
Don’t you get cache expirations then if they change providers? This ought to introduce delays and costs
reply
Yes, but that's why you pin them. If you specify more than one provider in `order`, OpenRouter will use the first one unless it's down, so that's the only time it would switch providers on you. And personally, I'd pay a few cents instead of waiting for the API to come back.
reply
Any proof of 100x spend? That seems a bit excessive
reply
There's no difference between the two if you pin the provider to Deepseek on Openrouter.

If you don't want to mess about client side with pinning, set a guardrail on Openrouter that limits the available providers to only the official one.

reply
Then why bother using OpenRouter and paying the extra fees?
reply
Because you have like every model on the planet to choose from. So if a contender drops you switch.

Also if DS is down you can choose another provider.

I have credits at DS and OR directly. But I do see the value in OR.

reply
So you're basically send your code to China?

I thought the advantage of DeepSeek is that you can host it on a server of your choosing.

reply
Yes, I send it to China, and they use it to improve open-weight models. I'm ok with this arrangement. At least, I'm happier with this than with companies using my open-source work to improve proprietary models without my consent.
reply
Or.. BYOK Deepseek because OpenRouter's UX is much nicer?
reply
Zero Data Retention and not having company source code leak to "CHINA!" (said in Trumps annoying voice) would be two reasons not to
reply