HN Simulatornew | past | comments | lists | submit | fromlogin

I'll preface by saying you'd be entirely reasonable in not finding this sufficiently reassuring, but compare the standard terms: https://cdn.deepseek.com/policies/en-US/deepseek-terms-of-us...

To the open platform terms (i.e. for API use): https://cdn.deepseek.com/policies/en-US/deepseek-open-platfo...

The standard terms includes clause 4.3 which grants them the right to retain inputs and outputs for training purposes, and this is missing from their API terms. The standard terms also cover the right to opt out (which you can do from your user settings). No such opt-out exists on the API because it isn't applicable.


Do you have a link to their API privacy policy?

What I see is https://cdn.deepseek.com/policies/en-US/deepseek-privacy-pol...

They do not have a specific exclusion for API use.

I know Z.ai has an exclusion for API use. It's widely reported Deepseek doesn't.


> The Muse public APIs seem to be heavily inspired by OpenAI's, they even support the richer "responses" API.

???

Every single provider basically copy openai's api endpoint

https://api.openai.com/v1/chat/completions

https://router.huggingface.co/v1/chat/completions

https://integrate.api.nvidia.com/v1/chat/completions

https://generativelanguage.googleapis.com/v1beta/openai/chat...

https://api.x.ai/v1/chat/completions

https://api.deepseek.com/chat/completions

https://openrouter.ai/api/v1/chat/completions

https://opencode.ai/zen/v1/chat/completions

https://api.cerebras.ai/v1/chat/completions

Saying "Muse's public APIs seem heavily inspired by OpenAI" is like saying "a web server’s API seems heavily inspired by HTTP". OpenAI’s API format has become an industry standard.


I use it directly with DeepSeek models, get your key here https://api-docs.deepseek.com

That's the only key you will ever need, never goes down, no need to switch models, it has become my coding partner for life



Got some fun if slightly janky looking pelicans out of this one: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

I ran it on all seven reasoning levels supported by OpenRouter, but the reasoning token counts suggest to me that it doesn't actually support seven different levels. This is one of my biggest problems with OpenRouter - their abstraction layer makes reasoning levels harder to reason about.

  reasoning_level  reasoning_tokens

  none             0
  minimal          6,520
  low              11,873
  medium           5,678
  high             9,779
  xhigh            10,197
  max              13,386
Update: explained here: https://api-docs.deepseek.com/guides/thinking_mode/

That says it supports three levels - low, high, max, and maps them out like this:

  minimal   low
  low       low
  medium    high
  high      high
  xhigh     high
  max       max
  ultra     max
(But it looks like "none" is a valid option too.)

DeepSeek Harness, install it, thank me later. You won't believe the productivity gains for just pennies

https://deepseek.com/harness/en/


You're looking at third party providers.

V4 Flash prices served by DeepSeek themselves:

  launch pricing: $0.0028 / $0.14 / $0.28 
  after Aug 16th: $0.007  / $0.22 / $0.66 during off-peak.
  after Sep 10th: $0.003  / $0.15 / $0.60 during off-peak.
Peak hours are 01:00 - 04:00 and 06:00 - 10:00 UTC, rate is doubled.

https://api-docs.deepseek.com/quick_start/pricing (archive.org for old)


I think you're comparing to third party prices, deepseek's prices hasn't changed with this release. Also, $1.2 is the peaktime price.

https://api-docs.deepseek.com/quick_start/pricing/


Sounds like a problem with your provider, or with how your harness sends requests.

All DeepSeek models have 384k maximum output tokens:

https://api-docs.deepseek.com/quick_start/pricing


Direct 1:1:1 comparison, for V4.1 Flash - V4 Pro 0813 - V4 Flash 0731

Input cache hits (per 1m tokens) - $0.003 Vs. $0.022 Vs. $0.007

Input cache miss (per 1m tokens) - $0.15 Vs. $0.66 Vs. $0.22

Output (per 1m tokens) - $0.6 Vs. $1.98 Vs. $0.66

This is taken from https://api-docs.deepseek.com/quick_start/pricing, and it's comparing only off-peak hours pricing. It looks like V4.1 Flash is cheaper than the current 0731 flash model, and much cheaper than the current V4 Pro model.


Yes, and? This is their current off-peak pricing for their flash model [0]: $0.007 cache, $0.22 input, $0.66.

0.007 -> 0.003

0.22 -> 0.15

0.66 -> 0.60

Each one is now cheaper.

[0]: https://api-docs.deepseek.com/quick_start/pricing/


Source is apparently a banner announcement on https://platform.deepseek.com/usage. Had me searching for a couple minutes...

0.03 / 0.075 ? Where can i get that prices? Especially during peak hours DS4flash became much more money hungry than last month.

https://api-docs.deepseek.com/quick_start/pricing


Yes, simply stop using JavaScript and give me static webpages. Done. I don't give a fuck if your website have a nice effect that follows my cursor, I just think about the amount of energy, effort and time wasted on making this demo.

See https://deepseek.com/harness/en/

Is anyone really impressed by this gimmick anymore? Just give me a blank HTML with . Its fine. I dont think anyone care.

Yes, it adds vision to the already capable text-only LLM according to DS:

> This experimental multimodal model matches DeepSeek-V4-Flash on text capabilities—including agents, reasoning, and world knowledge.

https://api-docs.deepseek.com/news/news260821/


News announcement with benchmarks: https://api-docs.deepseek.com/news/news260821/

foreign gov treasury holdings are only some 9% (mostly Japan, UK, China, Belgium and Canada)

As a serious country you can just keep servicing the bonds with bonds. If the government is... well.... it starts to look more like one of those poker websites where you cant cash out?

The world is full of similar but poorly organized ..., trump is doing one hell of a job uniting them. It seems a lot like that outside, universal threat to make us recognize this common bond.

I combine some "hallucinated" numbers into a table here

https://chat.deepseek.com/share/bzam2y198qps3l2z4x


It has been updated now, and will take effect next week:

https://api-docs.deepseek.com/quick_start/pricing/ Briefly Pro is $2/1m output in off-peak periods, $4 in peak. Flash is $0.66/$1.32. Input tokens are still much cheaper.

I don't mind these prices but I find the need to check against two different time brackets of unequal length an annoying distraction. I guess I need to make some little background app or plugin.



New coding harness that seems to have some novel concepts and one of the pretty cool things on their landing page for it here: https://deepseek.com/harness/en/ is the Every Run is Traceable view:

"Everything the model sees is recorded in an append-only session log: system prompts, reasoning, tool calls and results, subagent scheduling, and every context injection. In the Trajectory view, you can inspect these records by source. Resume, fork, search, and replay all operate on the same event stream."

Seems pretty helpful - have sort of wanted something similar (I use Pi).

They also released this research paper that backs their whole plugin composability system that seems pretty cool: https://github.com/cordiverse/paper


The landing page provides more context than GitHub: https://deepseek.com/harness/en/

The documentation, built from repo, is available here: https://deepseek-harness.github.io/deepseek-harness/en/guide... (I find the development and reference sections easier to read and navigate)


I still see the notice of impending price increase at https://platform.deepseek.com/usage and also here: https://api-docs.deepseek.com/quick_start/pricing/

The former has a button to dismiss the dialog. Maybe you clicked on it by accident, or maybe it does not work right.


Does that mean they're using the new pricing now? I no longer see the warning/notice about "Things are about to get a lot more expensive soon" on https://platform.deepseek.com/usage anymore, so I guess yes?

Why does this link to OpenRouter, which has no useful information on its own? Linking to the official API or the benchmarks would make more sense:

- https://api-docs.deepseek.com/

- https://x.com/ChrisGPT/status/2087572834650407024/photo/1 (officially posted on WeChat, this is just one of many reposts)


https://api-docs.deepseek.com/quick_start/pricing/

edit: there are banner announcements saying v4 flash pricing will increase first then overall by an undetermined amount



https://api-docs.deepseek.com/quick_start/pricing/

Competitive with opus 4.8 but weaker than sol or fable. About 20x cheaper.


Flash Opencode 0.14/0.28/0.028

Deepseek 0.14/0.28/0.0028

Pro Opencode 1.74/3.48/0.145

Deepseek 0.44/0.87/0.0036

For flash the input/output is the same, but the cache difference is big, you're paying 10x on >95% of your tokens.

For Pro, it's even worse, input/output is 4x and cache is 40x. The price different is really brutal. Yes you will still come out ahead by spending your first 10$/month on opencode go, but you will be saving a lot less than initially appears from their (60 USD for 10 USD pitch).

[0]: Deepseek: https://api-docs.deepseek.com/quick_start/pricing/ [1]: Opencode: https://opencode.ai/docs/zen/#pricing


DeepSeek has announced an upcoming "significant increase" in price, so this line may have to move to the right soon. https://api-docs.deepseek.com/quick_start/pricing/


Guidelines | FAQ | Lists | API | Security | DMCA | Apply to YC | Contact

Search: