Retired · 2026-09-01

DeepSeek's API no longer lists a plain V4 model. The line split into Flash and Pro variants, and V4.1 Flash is the current cheapest member of the family.

Use DeepSeek V4.1 Flash. It runs on the same endpoint and the same key, so switching is a one-line change to the model id.

Switch

What to use instead

RetiredAdded 1 Jun 2026DeepSeek

DeepSeek V4 API

Open-weight frontier-adjacent reasoning at commodity prices.

Model spec

Input
$0.30per 1M tokens
Output
$0.90per 1M tokens
Context
128K

Model ID

OpenAI-compatible · api.companyfabric.com/v1

Overview

About DeepSeek V4

DeepSeek V4 is the open-weight price-performance benchmark: near-frontier reasoning and coding at commodity token prices, MIT-style licensed. The V4 family spans three tiers, Flash for volume, base V4 for the default, Pro for depth, all one string apart on the same CompanyFabric key. Every model on CompanyFabric shares the same API key, billing balance, and rate-limit envelope, one integration covers the entire library.

Facts

At a glance

Availability

Retired

Added 1 Jun 2026

Pricing

$0.30 in · $0.90 out

Per 1M tokens, metered exactly per call

Context window

128K tokens

Commercial use

See license

Open weights; commercial use permitted (MIT-style model license).

Variants

All endpoints

Prompt library

Example runs

Solve and explain: a train leaves at…

Step 1…

This exact run cost $0.001

Quickstart

How to use the DeepSeek V4 API

  1. 1

    Create a free CompanyFabric account, starter credits are included, no card required.

  2. 2

    Copy your API key from the dashboard. One key covers every model in the library.

  3. 3

    Point any OpenAI SDK at api.companyfabric.com/v1 and set model to "deepseek/deepseek-v4".

  4. 4

    Read the exact metered cost of the call from the response, the same number shown in the playground.

Use cases

What teams build with it

High-volume agent loops

Sub-agent calls, bulk refactors, and test generation where volume dwarfs per-call quality differences, the bill stays flat.

agentsbulkcodegen

Cost-tiered routing

V4 as the default lane with a frontier model on escalation, most teams cut spend 5–10× without a quality cliff.

model routingcost optimization

Self-host migration path

Open weights mean the exit door is real: prototype through the API, self-host later if scale justifies GPUs, same model either way.

open weightsself-host

Get better output

Prompting tips

  • State the output format before the task, models follow contracts declared early.
  • Few-shot beats description: two worked examples outperform a paragraph of rules.
  • Cap output length explicitly when cost matters; output tokens dominate most bills.

Exact, metered, prepaid

DeepSeek V4 pricing

DeepSeek V4 pricing is per token: $0.3 / M input · $0.9 / M output. Output tokens cost 3× input tokens, so long generations, not long prompts, are what actually drive most bills. A typical call with a 2k-token prompt and a 1k-token answer costs about $0.0015.

Every call is metered exactly and attributed to your key, the dashboard, the API usage response, and your invoice all read from the same ledger row. Cap any key's monthly spend with a budget cap, and bring your own provider key (BYOK) at 0% platform fee if you already have a direct contract.

EndpointTypePrice
DeepSeek V4Chat$0.3 / M input · $0.9 / M output
DeepSeek V4.1 FlashChat$0.3 / M input · $1.2 / M output · $0.006 / M cached input
DeepSeek V4 ProChat$1.32 / M input · $3.96 / M output · $0.044 / M cached input
DeepSeek V4 FlashChat$0.44 / M input · $1.32 / M output

Comparisons

DeepSeek V4 vs the alternatives

DeepSeek V4 vs Gemini 3.7 Flash

DeepSeek V4 and Gemini 3.7 Flash compete for the same text workloads. Both run behind the same CompanyFabric key at exact metered prices, A/B them on your real prompts and let the outputs decide, then switch with a one-line change.

Full comparison: DeepSeek V4 vs Gemini 3.7 Flash →

DeepSeek V4 vs Claude Haiku 4.5

Haiku 4.5 is the fast tier of a frontier family with Anthropic's tool-calling polish; V4 is the open-weight value play with deeper reasoning per dollar. Latency-sensitive product surfaces lean Haiku; batch and agent-internal work leans V4.

Full comparison: DeepSeek V4 vs Claude Haiku 4.5 →

Compare

DeepSeek V4 head to head

FAQ

Frequently asked questions

How much does the DeepSeek V4 API cost?

$0.3 / M input · $0.9 / M output. Every call is metered exactly and shown before you run it.

How do I get a DeepSeek V4 API key?

Create a CompanyFabric account, and one key unlocks DeepSeek V4 and every other model in the library. Free starter credits are included, no subscription, no card required to try it.

Why use DeepSeek V4 on CompanyFabric instead of going direct?

One key, one prepaid balance, and one rate-limit envelope across the whole library, at pricing at parity with or below the direct API. Model switching is a one-line change, and BYOK routing is 0% fee.

Is the API OpenAI-compatible?

Yes, point any OpenAI SDK at api.companyfabric.com/v1 and set model to "deepseek/deepseek-v4".

Can I use the output commercially?

Open weights; commercial use permitted (MIT-style model license).

How does billing work?

Prepaid credits via card, or BYOK with your own provider keys at 0% platform fee. No subscription required; balances never expire.

Provider

About DeepSeek

DeepSeek publishes the deepseek family. Also in this family: DeepSeek V4.1 Flash, DeepSeek V4 Pro, DeepSeek V4 Flash. License: Open weights; commercial use permitted (MIT-style model license).

One key. Every model. Exact prices.

Run DeepSeek V4 with free starter credits. Free starter credits included, no subscription required.