Hacker News
new
past
comments
ask
show
jobs
points
by
literallyroy
11 hours ago
|
comments
by
minimaxir
11 hours ago
|
[-]
Per the announcement tweet, BaseTen was the inference provider which has 20% cache cost that is typical:
https://www.baseten.co/library/deepseek-v4-flash-0731/
reply
by
literallyroy
10 hours ago
|
parent
|
[-]
Ah thanks. That looks like 10x cost on cache reads vs Deepseek as the provider:
https://openrouter.ai/deepseek/deepseek-v4-flash-0731#provid...
reply