Malaysian Tech Wiki
Malaysian Tech Wiki_
← Back to Artificial Intelligence
DeepSeek official whale mark

Artificial Intelligence / Official Model Sites

DeepSeek

DeepSeek offers a chat assistant and developer APIs for its reasoning models. API capabilities and prices below were checked on 4 October 2026; the web assistant can have different access limits.

Logo: DeepSeek API documentation

tool · Sources verified · Checked 2026-10-04

Overview

DeepSeek offers a chat assistant and developer APIs for its reasoning models. API capabilities and prices below were checked on 4 October 2026; the web assistant can have different access limits.[1][3]

Use cases

  • Build conversational applications and coding assistants using an OpenAI-compatible or Anthropic-compatible API.[1][2]
  • Use visual inputs with V4.1 Flash, or connect model responses to application tools.[1]

Features

  • Thinking and non-thinking modes, JSON output, tool calling and Responses API support are documented for the current API models.[1]

Limitations

  • V4 Pro does not support vision. FIM completion is restricted to non-thinking mode.[1]
  • The Responses endpoint is stateless: applications must send conversation history. File inputs are not supported by that endpoint.[2]

Access and availability

  • The official changelog retains V4 Pro API access after 14 September 2026, despite earlier phase-out announcements. Legacy Flash names route to V4.1 Flash.[3]

Pricing

  • USD per million tokens at peak rates: Flash input $0.30 uncached or $0.006 cached, output $1.20; Pro input $1.32 uncached or $0.044 cached, output $3.96. Off-peak rates are half. These are API rates, not a chat subscription price.[1]

Specifications

Verified 2026-10-04. Versions and access terms can change.

API model identifiersdeepseek-flash (V4.1-Flash); deepseek-v4-pro (V4-Pro-0813)[1]
Context / maximum output1 million / up to 384,000 tokens; endpoint defaults can be lower[1]
API compatibilityOpenAI and Anthropic formats; Responses API[1]

Public user feedback

Attributed external experiences. These are separate from verified specifications and vendor claims.

V4 Flash text-generation experience

ConsequenceInside167 reported a positive day-long text-generation test in Cherry Studio and praised value for money. This April 2026 report concerns the earlier V4 Flash, not the current V4.1 release.

u/ConsequenceInside167 · Reddit · r/DeepSeek · 2026-04-28 · Checked 2026-10-04

V4.1 Flash first impressions

ClearRabbit605 praised response speed and some UI results, but also reported unwanted actions and the need to keep coding work within scope. This is one person's early experience, not a benchmark.

u/ClearRabbit605 · Reddit · r/DeepSeek · Publication date unavailable · Checked 2026-10-04

A contrasting V4.1 coding experience

sirjoee liked the speed and price but reported forgotten instructions while building a game engine with DSH. The author explicitly questioned whether the coding harness contributed to the problems.

u/sirjoee · Reddit · r/DeepSeek · Publication date unavailable · Checked 2026-10-04

Links

Sources and evidence gaps

  1. DeepSeek models and API pricingdocumentation · checked 2026-10-04
  2. DeepSeek Responses API referencedocumentation · checked 2026-10-04
  3. DeepSeek API changelogdocumentation · checked 2026-10-04
  • Current web-assistant quotas were not established; the old catalog's unlimited claim is retained only as historical catalog information.
  • The two V4.1 discussions expose relative timestamps in the fetched pages, so exact publication dates are omitted.