V4-Pro stays on the API · V4.1 Flash benchmarks published

The Complete DeepSeek AI GuideModels, Pricing, API & How to Use It

Everything about DeepSeek AI in one place: the new V4.1 Flash and V4-Pro models, the free web chat and apps, API pricing, downloads, login help and fixes when DeepSeek isn't working. Independent, and checked against DeepSeek's official sources.

deepseeksr1.com is not affiliated with DeepSeek (杭州深度求索人工智能). To use DeepSeek, go to the official chat.deepseek.com, platform.deepseek.com or deepseek.com.

Explore the Guide Latest News Pricing
552B
V4.1 Flash · 8B/16B active
1M
Token context · 384K out
90.6
Terminal-Bench 2.1
$0.15
Input / 1M (Flash, off-peak)
Open
Weights on Hugging Face
V4-Pro Still served · price unchanged
DeepSeek-V4.1-Flash 552B MoE · 8B/16B active
API model ID deepseek-flash
Terminal-Bench 2.1 90.6 (V4.1 Flash)
GPQA Diamond 90.9 (V4.1 Flash)
KV Cache 1/4 HBM · 1/8 SSD
Flash pricing $0.15 in / $0.60 out off-peak
Funding ~$12B round reported (Oct 6)
V4.1 Pro In development · no date
V4-Pro Still served · price unchanged
DeepSeek-V4.1-Flash 552B MoE · 8B/16B active
API model ID deepseek-flash
Terminal-Bench 2.1 90.6 (V4.1 Flash)
GPQA Diamond 90.9 (V4.1 Flash)
KV Cache 1/4 HBM · 1/8 SSD
Flash pricing $0.15 in / $0.60 out off-peak
Funding ~$12B round reported (Oct 6)
V4.1 Pro In development · no date

What's in this DeepSeek guide

Jump straight to what you need.

01Overview

What Is DeepSeek AI?

DeepSeek AI is a Chinese AI lab, and the family of free and low-cost AI models it builds. Here is who is behind it, what it offers and how it got here.

DeepSeek is a Chinese artificial-intelligence lab based in Hangzhou that builds large language models and gives the public three ways to use them: a free chatbot (web and mobile), a low-cost developer API, and open model weights anyone can download and run.

It was founded in 2023 by Liang Wenfeng as an offshoot of the quantitative hedge fund High-Flyer. DeepSeek became globally known in January 2025, when its R1 reasoning model matched leading US models at a reported fraction of the training cost and its app briefly topped app-store charts worldwide.

Its signature is efficiency: mixture-of-experts models that switch on only a small slice of their parameters per token, sparse attention for long inputs, and aggressive API pricing. The current lineup is V4.1 Flash (552B parameters, native vision) and V4-Pro (1.6T parameters).

Founded
2023
HQ
Hangzhou, China
Founder
Liang Wenfeng
Latest model
V4.1 Flash
  1. Dec 2024DeepSeek-V3 makes low-cost MoE mainstream
  2. Jan 2025R1 reasoning model and app go viral
  3. Dec 2025V3.2 introduces DeepSeek Sparse Attention
  4. Apr 2026V4 family previews with 1M-token context
  5. Sep 2026V4.1 Flash adds native vision
  6. Oct 2026~$12B funding round reported ahead of a planned IPO

Want the long version? Read our full explainer: What is DeepSeek AI? →

02Overview

Why Choose DeepSeek AI? 8 Key Benefits

Why so many people and developers pick DeepSeek over other AI chatbots and APIs (lower cost, strong coding, open weights and more), plus when it is not the right fit.

💰
$0.60output / 1M tokens off-peak

Much lower cost

V4.1 Flash costs a fraction of most frontier APIs, cached input is just $0.003 per million tokens, and off-peak hours are half price.

🤖
90.6Terminal-Bench 2.1

Frontier-level coding & agents

DeepSeek reports V4.1 Flash ahead of its own V4-Pro on agentic coding, with CyberGym 88.1 and DeepSWE v1.1 74.2.

🔓
HFweights on Hugging Face

Open weights you can own

Download and self-host the same models behind the API, fine-tune them on your data, and keep everything on your own servers.

📏
1Mtokens in · 384K out

Huge context window

Fit whole codebases, long contracts or book-length drafts in a single prompt, with a compact KV cache keeping long runs cheap.

👁️
Visionbuilt into V4.1 Flash

Sees images natively

Send screenshots, charts, invoices or photos of homework in the same request as text, with no separate vision model.

🔌
2 linesto migrate

Easy to switch to

Works with OpenAI- and Anthropic-format SDKs: change the base URL and model name, and your existing code keeps running.

🆓
$0web chat & apps

Free to try

The official chat and mobile apps are free, so you can test DeepSeek on your own tasks before spending anything on the API.

🎚️
3thinking-effort levels

Control speed vs. depth

Choose low, high or max reasoning effort per request to trade cost and latency against answer quality.

Output price per 1M tokens

Lower is cheaper · off-peak list prices

DeepSeek V4.1 Flash
$0.60
DeepSeek V4-Pro
$1.98
Claude Opus 4.8
$25.00

DeepSeek prices from its official pricing page (off-peak; peak is 2×). Claude Opus 4.8 list price as listed in our comparison table. Input pricing and real costs depend on your workload.

Maybe not the best pick if…

  • You handle regulated or confidential data and can't self-host. The hosted service is run by a Chinese company.
  • Your organisation restricts Chinese AI providers, or you work in US government or defence supply chains.
  • You need answers on politically sensitive China-related topics; the hosted models filter them.
  • You need audio input or image generation, which the current models don't offer.
See the latest regulatory news →
04Models

All DeepSeek AI Models

A complete list of DeepSeek models, from DeepSeek LLM to V4.1 Flash, with release dates, sizes, context length and whether each is on the API or available as open weights.

21
Models released
2
On the API today
7B → 1.6T
Parameter range
4K → 1M
Context growth
ModelReleasedTypeSize (active)ContextStatusNote
DeepSeek-V4.1-ProTBAGeneral · flagshipTBATBAUpcomingConfirmed in development, no date
DeepSeek-V4.1-FlashSep 2026General · vision552B (8B in / 16B out)1MOn API nowAPI: deepseek-flash
DeepSeek-V4-Pro (0813)Aug 2026General · agents1.6T (49B)1MOn API nowAPI: deepseek-v4-pro
DeepSeek-V4-Flash-Vision-ExpAug 2026Vision · experimental—1MOpen weightsAPI retired 10 Sep 2026
DeepSeek-V4-Flash (0731)Jul 2026General · fast284B (13B)1MOpen weightsAPI retired 10 Sep 2026
DeepSeek-V4 (preview)Apr 2026GeneralPro 1.6T · Flash 284B1MOpen weightsPreview, replaced by GA builds
DeepSeek-V3.2Dec 2025General · reasoning671B (37B)128KOpen weightsIntroduced Sparse Attention
DeepSeek-V3.2-ExpSep 2025General · experimental671B (37B)128KOpen weightsFirst DSA release
DeepSeek-V3.1 / TerminusAug–Sep 2025Hybrid thinking671B (37B)128KOpen weightsOne model, thinking + non-thinking
DeepSeek-R1-0528May 2025Reasoning671B (37B)128KOpen weightsFewer hallucinations, tool calls
DeepSeek-Prover-V2Apr 2025Math · theorem proving671B—Open weightsFormal proofs in Lean 4
Janus-ProJan 2025Image understanding + generation1B / 7B—Open weightsUnified multimodal model
DeepSeek-R1 + distillsJan 2025Reasoning671B (37B) · distills 1.5B–70B128KOpen weightsThe model that went viral
DeepSeek-V3Dec 2024General671B (37B)128KOpen weightsLow-cost MoE breakthrough
DeepSeek-VL2Dec 2024Vision-languageTiny / Small / Base (MoE)—Open weightsSuccessor to DeepSeek-VL
DeepSeek-V2.5Sep 2024General + code236B (21B)128KOpen weightsMerged chat and coder models
DeepSeek-Coder-V2Jun 2024Code236B (21B) · Lite 16B128KOpen weights338 programming languages
DeepSeek-V2May 2024General236B (21B)128KOpen weightsIntroduced MLA attention
DeepSeek-VLMar 2024Vision-language1.3B / 7B—Open weightsFirst vision model
DeepSeek-MathFeb 2024Math7B4KOpen weightsIntroduced GRPO training
DeepSeek-CoderNov 2023Code1.3B–33B16KOpen weightsFirst code models
DeepSeek LLMNov 2023General7B / 67B4KOpen weightsDeepSeek's first LLMs

Sizes are total parameters, with active parameters per token in brackets for mixture-of-experts models. "Open weights" means weights are published on DeepSeek's Hugging Face page; check each model card for its license. "—" = not applicable or not published. Linked names open our detailed guides.

05Models

DeepSeek V4.1, V4, V3.2, V3.1, R1 & V3 Explained

What changed in each major DeepSeek release, from the V3 breakthrough to V4.1 Flash. Pick a model on the timeline to see its specs, key improvements and benchmark highlights.

V3Dec 2024 R1Jan 2025 V3.1Aug 2025 V3.2Dec 2025 V4Apr 2026 V4.1Sep 2026 Context 128K128K128K128K1M1M
Latest · native vision

DeepSeek V4.1 Flash

Released 10 Sep 2026
Official announcement ↗Our V4.1 Flash guide →

The smallest model in DeepSeek's new architecture family, and the first with native visual understanding. DeepSeek calls it smarter, faster and more efficient, with results ahead of its own flagship V4-Pro.

Params552B MoE
Active8B in · 16B out
Context1M
APIdeepseek-flash
  • New asymmetric causal encoder–decoder architecture: reading input is cheaper than writing output.
  • KV cache needs 1/4 the HBM and 1/8 the SSD of the previous generation, cutting agent costs.
  • Native image understanding replaces the separate V4-Flash-Vision-Exp model.
  • Open weights and technical report on Hugging Face.
  • DeepSeek first planned to route V4-Pro traffic to it, then reversed that on 11 Sep 2026; V4-Pro stays available.
90.6Terminal-Bench 2.1
90.9GPQA Diamond
88.1CyberGym
3471Codeforces
1M context · two sizes

DeepSeek V4

Released 24 Apr 2026
Official announcement ↗Our V4 guide →

A new generation launched as an open-source preview in two sizes, Pro and Flash, making a 1M-token context the default across all DeepSeek services. GA builds followed (Flash 0731, Pro 0813).

V4-Pro1.6T · 49B active
V4-Flash284B · 13B active
Context1M
APIdeepseek-v4-pro
  • Token-wise compression plus DeepSeek Sparse Attention makes 1M-token context cost-effective.
  • DeepSeek says V4-Pro led open models on agentic coding, maths, STEM and world knowledge at launch.
  • Thinking and non-thinking modes; OpenAI and Anthropic API formats.
  • Retired the old deepseek-chat / deepseek-reasoner names on 24 Jul 2026.
  • Built to plug into agents such as Claude Code, OpenClaw and OpenCode.
1.6TV4-Pro total params
1Mtoken context
87.9Terminal-Bench 2.1 (Pro GA)
82.7Terminal-Bench 2.1 (Flash GA)
Thinking + tools

DeepSeek V3.2

Released 1 Dec 2025
Official announcement ↗Our V3.2 guide →

The official successor to V3.2-Exp, which introduced DeepSeek Sparse Attention for cheaper long-context use. DeepSeek described V3.2 as GPT-5-level, with a high-compute Speciale variant.

Params671B · 37B active
Context128K
VariantV3.2-Speciale
StatusOpen weights
  • DeepSeek's first model to integrate thinking directly into tool use, in both thinking and non-thinking modes.
  • Agent training data from 1,800+ environments and 85k+ complex instructions.
  • V3.2-Speciale reached gold-level results at IMO, CMO, ICPC World Finals and IOI 2025 (API-only, briefly).
  • Weights and tech report for both models on Hugging Face.
1,800+agent environments
85k+complex instructions
4olympiad golds (Speciale)
128Kcontext
Hybrid thinking

DeepSeek V3.1

Released 21 Aug 2025
Official announcement ↗Our V3.1 guide →

One model with two modes: Think and Non-Think, switched by the DeepThink button. It reached answers faster than R1-0528 and brought big gains in tool use and agent tasks.

Params671B · 37B active
Context128K
Base+840B tokens of long-context pretraining
UpdateV3.1-Terminus
  • Hybrid inference: deepseek-chat = non-thinking, deepseek-reasoner = thinking.
  • Stronger tool use and multi-step agent tasks after post-training.
  • Added Anthropic API format support and strict function calling (beta).
  • V3.1-Terminus (Sep 2025) fixed Chinese/English mixing and improved code and search agents.
66.0SWE-bench Verified
54.5SWE-bench Multilingual
31.3Terminal-bench
128Kcontext
Reasoning

DeepSeek R1

Released 20 Jan 2025 · update 28 May 2025
Official announcement ↗Our R1 guide →

The reasoning model that made DeepSeek famous, matching leading US models at a fraction of the reported cost. The R1-0528 update improved benchmarks, cut hallucinations and added JSON output and function calling.

Params671B · 37B active
Context128K
UpdateR1-0528
StatusOpen weights
  • Thinks step by step before answering (chain of thought).
  • R1-0528: reduced hallucinations and better front-end code generation.
  • R1-0528 added JSON output and function calling.
  • Its role is now covered by thinking mode in newer models; weights remain on Hugging Face.
87.5AIME 2025 (was 70.0)
81.0GPQA (was 71.5)
73.3LiveCodeBench (was 63.5)
71.6Aider (was 57.0)
The MoE breakthrough

DeepSeek V3

Released 26 Dec 2024
Official announcement ↗Our V3 guide →

The large, open mixture-of-experts model that proved frontier-level AI could be trained and served cheaply. It's the foundation that R1, V3.1 and V3.2 were built on.

Params671B · 37B active
Training14.8T tokens
Speed60 tokens/sec
StatusOpen weights
  • 671B MoE parameters with only 37B active per token.
  • Trained on 14.8T high-quality tokens.
  • 60 tokens/second, 3× faster than V2.
  • Fully open-source models and paper.
  • Launch API price: $0.27/M input (cache miss), $0.07/M (cache hit), $1.10/M output.
671Btotal params
37Bactive params
14.8Ttraining tokens
3×faster than V2

Facts from DeepSeek's official announcements; benchmark figures that those pages show only as images are taken from DeepSeek's API change log. All figures are DeepSeek's own and may differ from independent tests. Parameter counts for V3.1, V3.2 and R1 follow the V3 base they're built on.

06Models

How DeepSeek Models Work: MoE, Sparse Attention & More

The technology behind DeepSeek V4 and V4.1, and why it lets DeepSeek offer frontier-level AI at a fraction of the usual price.

🔀
Mixture of Experts (MoE)

V4-Pro has 1.6T total parameters but activates ~49B per token. V4.1 Flash has 552B total and activates only 8B to read input and 16B to generate output.

↔️
Causal Encoder–Decoder (V4.1)

V4.1's new asymmetric design uses a smaller active path for reading context than for writing output, cutting the cost of the long inputs that dominate agent workloads.

🗜️
Compact KV Cache (V4.1)

V4.1 Flash's KV cache needs about 1/4 the HBM and 1/8 the SSD storage of the previous generation. Since cache hits are a big share of agent bills, this flows straight into lower prices.

🧩
DeepSeek Sparse Attention (DSA)

Reduces long-context attention cost from O(L²) toward O(kL), making the 1M-token context window practical for whole-repository and book-length inputs.

👁️
Native Multimodality

V4.1 Flash understands images natively rather than via a separate vision model, replacing the retired V4-Flash-Vision-Exp. V4-Pro remains text-only.

🎚️
Thinking Effort Levels

Choose low, high or max reasoning effort per request. Thinking mode is on by default; non-thinking mode is available for fast, cheap responses.

07Models

DeepSeek Benchmarks: V4.1 Flash vs V4-Pro vs Claude Opus 4.8

How DeepSeek V4.1 Flash and V4-Pro compare with Claude Opus 4.8 and Kimi K3 on coding, agent and reasoning benchmarks, and on price.

● BENCHMARK SCORES

DeepSeek-V4.1-Flash
DeepSeek-V4-Pro-0813
Claude Opus 4.8
Kimi K3

● FULL MODEL COMPARISON TABLE

Metric V4.1 Flash V4-Pro Opus 4.8 Kimi K3
Output price /1M$0.60 (off-peak)$1.98 (off-peak)$25.00—
Context window1M1M——
Vision input✓ Native✗——
Open weights✓✓✗—
Terminal-Bench 2.190.687.985.088.3
HLE (no tools / tools)36.8 / 63.942.7 / 60.049.8 / 57.9—
NL2Repo65.461.569.7—
CyberGym88.183.378.3—
DeepSWE74.2*62.7—67.5
Agents' Last Exam31.825.7——
SWE-bench Verified—80.688.6—
Toolathlon-Verified—74.1—76.5
GPQA Diamond90.9———

V4.1 Flash and V4-Pro scores are from DeepSeek's official change log (10 Sep and 13 Aug 2026); competitor scores from vendor model cards as of Aug–Sep 2026. *V4.1 Flash was tested on DeepSWE v1.1, so it is not directly comparable with the other DeepSWE figures. "—" = not reported on a comparable setting. Vendor-reported benchmarks often differ from independent tests. Deeper dives: DeepSeek vs Claude · DeepSeek vs GPT-5 · all models compared.

08Get Started

How to Use DeepSeek AI: Web, App & API

Three ways to use DeepSeek AI, step by step: chat free in your browser, use the mobile app, or connect your own software through the API.

Free

Use it in your browser

  1. Go to chat.deepseek.com (DeepSeek's official site).
  2. Sign up with your email or a supported login.
  3. Type your question, or attach an image or file.
  4. Turn on deeper thinking for hard problems, then read and refine the answer.
Free

Use the mobile app

  1. Open the App Store or Google Play and search "DeepSeek".
  2. On Android you can also get the APK from DeepSeek's official download page; avoid APKs from other sites.
  3. Log in with the same account to sync your chats.
  4. Ask by typing, speaking, or snapping a photo.
Pay per use

Build with the API

  1. Create an account at platform.deepseek.com and top up a small balance.
  2. Generate an API key and keep it secret.
  3. Point the OpenAI or Anthropic SDK at DeepSeek's base URL.
  4. Call deepseek-flash or deepseek-v4-pro, ideally off-peak for half price.
See the full code sample ↓

Illustrations are simplified and don't show DeepSeek's actual interface. deepseeksr1.com is an independent guide; we don't host chat, apps or API keys.

09Get Started

DeepSeek Web: Free AI Chat in Your Browser ↗

DeepSeek Web is the official chat at chat.deepseek.com. It runs in any modern browser on a computer, tablet or phone, with nothing to install, and it's free with an account.

💬
Chat in any browser

Chrome, Edge, Safari or Firefox, on desktop or mobile, with no download.

🧠
Deeper thinking

Switch on reasoning for maths, logic and coding; leave it off for quick answers.

🔎
Web search

Let it look up recent information and show sources, where the option is available.

📎
Files & images

Upload PDFs, documents, screenshots or photos and ask questions about them.

🗂️
Chat history

Past conversations stay in the sidebar and sync with the mobile app.

🆓
Free

No subscription. Usage can be slowed or paused when servers are busy.

Open DeepSeek Web ↗

Web, app or API: which should you use?

WebMobile appAPI
CostFreeFreePay per token
Install neededNoYesNo (code)
Best forLong work at a computer, filesQuick questions, voice, photosBuilding apps & automation
Chat history sync✓✓—
Wherechat.deepseek.comDownloadDeepSeek API

Tips for DeepSeek Web

  • One topic per chat. Start a new chat for each task; very long chats get slower and less focused.
  • Bookmark it. Save chat.deepseek.com so you never land on a look-alike site from search ads.
  • Make it feel like an app. In Chrome or Edge, use the browser menu to install the page as an app; on a phone, add it to your home screen.
  • Paste, don't screenshot, text. Copy text in directly when you can, and save uploads for images and documents.
  • Busy message? Wait a minute and resend; see DeepSeek Not Working?
  • Mind privacy. Don't paste passwords or confidential data; chats are stored on DeepSeek's servers.

The illustration is simplified and doesn't reproduce DeepSeek's actual interface. Available buttons and features can change and may vary by region.

10Get Started

DeepSeek Login: How to Sign In Safely

How to sign in to DeepSeek safely on the web, in the app and on the API platform. You only ever log in on DeepSeek's own sites; deepseeksr1.com has no accounts and never asks for your password.

🌐

Web chat

  1. Type chat.deepseek.com into your browser yourself, or use a saved bookmark.
  2. Choose one of the sign-in options shown. These vary by region and may include email and password, a phone verification code, or a third-party account.
  3. Enter any verification code DeepSeek sends you.
Go to official login ↗
📱

Mobile app

  1. Open the official DeepSeek app (see Download).
  2. Sign in with the same method you used on the web so your chats sync.
⚙️

API platform

  1. Developers sign in at platform.deepseek.com to manage API keys, top-ups and usage.
  2. Chatting is free; the API console is only needed for building apps.
Open API platform ↗

Check the address before you type a password

The real login lives on deepseek.com addresses. Other sites using the DeepSeek name, with different endings or extra words, aren't run by DeepSeek. Don't enter your password there, even if they look identical. (Example addresses shown are made up.)

Login problems?

Forgot your password
  1. On the official login page, use the "forgot password" option if your sign-in method has one.
  2. Follow the reset link or code sent to your email or phone; it expires quickly.
  3. If you signed up with a third-party account, sign in with that instead; there may be no separate password.
Verification code doesn't arrive
  1. Check spam or promotions folders, and wait a minute before resending.
  2. Make sure the email or phone number has no typos.
  3. See DeepSeek Not Working? for more fixes.
"Account not found" on the API platform
  1. Make sure you're using the same sign-in method you originally signed up with.
  2. If you've never used the API, you may need to sign up on the platform first.
Keep your account safe
  • Use a strong password you don't use anywhere else.
  • Never share verification codes. DeepSeek staff won't ask for them.
  • Log out on shared or public computers.
  • Keep API keys secret, and delete any key you think has leaked.
11Get Started

DeepSeek Download: Android, iPhone, Windows & Mac

DeepSeek offers a free chat app for phones and a desktop app for Windows and Mac. Always download from DeepSeek's own download page or the official store listing; we don't host any files.

🤖

Android

DeepSeek chat app · free
  • Source: APK on DeepSeek's download page; Google Play listing where available
  • Includes: chat, voice input, photo upload, chat sync
  1. Open deepseek.com/download on your phone and tap Android.
  2. If prompted, allow your browser to install apps, then open the downloaded file.
  3. Turn that permission back off after installing, then log in.

Only install an APK that comes from download.deepseek.com. Copies on other sites may be modified or contain malware.

Official Android download ↗
🍎

iPhone & iPad

DeepSeek chat app · free
  • Source: Apple App Store
  • Includes: chat, voice input, photo upload, chat sync
  1. Open the App Store and search "DeepSeek", or use the link below.
  2. Check the developer is DeepSeek, then tap Get.
  3. Open the app and log in with the same account you use on the web.

Availability varies by country. If it isn't in your store, use chat.deepseek.com in Safari.

Open in App Store ↗
🪟

Windows

DeepSeek Harness desktop
  • Needs: 64-bit (x64) Windows PC
  • What it is: an AI workspace for documents, spreadsheets, code and scheduled tasks, extendable with plugins. More about Harness ↓
  1. Go to deepseek.com/download and choose Windows.
  2. Run the downloaded .exe installer and follow the prompts.
  3. Open DeepSeek Harness and follow its first-run setup.

DeepSeek says the desktop app is still being refined. Just want to chat? Use the web chat in Edge or Chrome; both let you install a site as an app from their menu.

Official Windows download ↗
💻

Mac

DeepSeek Harness desktop
  • Needs: Apple-chip Mac (M-series), macOS 13 or later
  • What it is: the same desktop AI workspace as on Windows. More about Harness ↓
  1. Go to deepseek.com/download and choose macOS.
  2. Open the .dmg and drag the app into Applications.
  3. Launch it from Applications and follow the setup.

Intel Macs aren't supported by the desktop app; use the web chat in Safari or Chrome instead.

Official Mac download ↗

🛡️ Safe download checklist

✓ Download only from deepseek.com, download.deepseek.com or the official App Store listing.
✓ DeepSeek's apps are free. Anything asking you to pay to download is fake.
✓ Be wary of "DeepSeek for PC" sites and ads; many bundle adware.
✓ Never enter your password on a site that isn't deepseek.com.
Linux isn't supported by DeepSeek's apps; use the web chat.
12Get Started

DeepSeek Harness: Open-Source AI Agent Workspace ↗

DeepSeek Harness (DSH) is DeepSeek's open-source AI agent workspace. Point it at a folder and it can organise documents, analyse spreadsheets, write and test code, research with sources, and run tasks on a schedule, all extendable with plugins. It's MIT-licensed and runs as a desktop app or a local web UI.

📂
Complete a range of tasks

Organise documents, analyse spreadsheets, write code. Works with Word, Excel, PDF, HTML, Markdown, Python and TypeScript files, with a review view showing diffs.

🧩
Everything is a plugin

Tools, skills and even the interface are plugins. Install them, or build one by chatting in Creator mode.

⏰
Adapt to your workflow

The Scheduled tasks plugin runs work on demand or on a schedule, such as a weekly report every Friday.

🛠️
Developer tools

Execution traces and runtime details help you debug tool calls and see what the agent actually did.

Ways to run it

💻Mac appApple-chip Mac, macOS 13+
🪟Windows app64-bit Windows
🌐Local web UINeeds Node.js
🐙From sourceClone from GitHub
# start the web UI (requires Node.js)
npx @deepseek-ai/dsh web

Get started in 4 steps

  1. Install the desktop app, or launch the web UI with the command shown.
  2. Open Settings → Models and add a DeepSeek API key (from platform.deepseek.com). Other OpenAI-compatible providers can also be configured.
  3. Click Choose workspace and pick the project folder the agent may work in.
  4. Start a session and describe the task. Approve file changes or commands when it asks.

Built-in & official plugins

Agent loopTerminalSubagents Agent teams expScheduled tasks expVoice input expAuto approval reviewer exp

More on GitHub under the dsh-plugin topic. Harness is built on the Cordis plugin framework.

💰 Cost: The app is free and open source; the AI work is billed to your DeepSeek API account at normal token prices. Agent tasks can use lots of tokens, so watch your usage.
🛡️ Safety: Harness can edit files and run commands. Give it a dedicated folder, keep backups, review diffs, and don't let it near passwords or secrets.
🚧 Preview: DeepSeek says Harness is in public preview and still being refined; features marked experimental may change.

Details from DeepSeek's Harness page and developer quickstart. Illustration is simplified and doesn't reproduce the real interface. See DeepSeek's Harness privacy policy and terms.

13Use It

DeepSeek AI Use Cases for Developers, Business & Students

Pick who you are to see practical ways to use DeepSeek, which model fits each job, how much thinking effort to set, and a prompt you can copy straight into the official chat or the API.

Fix bugs across a whole repo

V4.1 Flasheffort: high

Paste the failing test output and the relevant files. Find the root cause, explain it in two sentences, then give a minimal patch as a unified diff.

Run a coding agent

V4.1 Flasheffort: high

Hook deepseek-flash into Codex, Claude Code or OpenCode via DeepSeek's documented integrations and let it plan, edit and test in your terminal.

Turn a screenshot into code

V4.1 Flasheffort: low

Here is a screenshot of a settings page. Rebuild it as a single responsive HTML + Tailwind component. Keep spacing and hierarchy, skip real data.

Write tests and docs

V4.1 Flasheffort: low

Write pytest unit tests for this module covering edge cases, then a short README section explaining how to use its public functions.

Customer-support assistant

V4.1 Flasheffort: low

You are a support agent for an online store. Using only the policy text below, answer the customer politely and say when to escalate to a human.

Read invoices and forms

V4.1 Flasheffort: low

Extract supplier, invoice number, date, line items and total from this invoice image. Return valid JSON matching the schema below.

Summarise long contracts

V4-Proeffort: high

Summarise this contract for a non-lawyer: parties, term, payment, termination, liability caps and anything unusual. Flag clauses worth a lawyer's review.

Automate workflows

V4.1 Flasheffort: high

Given these tools (create_ticket, send_email, lookup_order), process each incoming message below and decide which tool calls to make.

Step-by-step tutoring

V4.1 Flasheffort: high

I'm stuck on this calculus problem. Don't give the answer yet: ask me one guiding question at a time and check my reasoning.

Solve a photographed problem

V4.1 Flasheffort: high

Here's a photo of my physics homework question. Explain which principle applies and walk through the solution with units.

Make a quiz

V4.1 Flasheffort: low

Create 10 multiple-choice questions on the French Revolution for 15-year-olds, with answers and a one-line explanation for each.

Plan a lesson

V4.1 Flasheffort: low

Draft a 45-minute lesson plan introducing photosynthesis, with a starter activity, main task, differentiation ideas and an exit question.

Review a stack of papers

V4-Proeffort: max

Here are 12 abstracts. Group them by method, note where findings conflict, and list three open questions none of them answer.

Analyse a dataset

V4.1 Flasheffort: high

Here is a CSV sample and column descriptions. Write pandas code to clean it, find the three strongest trends, and plot each one.

Read charts and figures

V4.1 Flasheffort: high

Describe what this chart shows, extract the underlying numbers as a table, and point out anything misleading about the axes.

Hard maths and proofs

V4-Proeffort: max

Prove the following statement rigorously. If it's false, give a counterexample. Show every step.

Brainstorm ideas

V4.1 Flasheffort: low

Give me 15 YouTube video ideas about budget travel in Sri Lanka, each with a hook title and a one-line angle.

Edit a long manuscript

V4-Proeffort: high

Here is my 60,000-word draft. Flag plot holes, inconsistent character details and pacing problems, chapter by chapter.

Translate and localise

V4.1 Flasheffort: low

Translate this product description into Sinhala and Tamil, keeping the brand tone friendly and adapting idioms for local readers.

Repurpose content

V4.1 Flasheffort: low

Turn this blog post into a LinkedIn post, five tweets and a 60-second video script, each in a matching tone.

V4.1 Flash default choice: cheaper, faster, understands images V4-Pro for knowledge-heavy or very long, careful work (text only) Effort low = quick answers · high = coding & analysis · max = hardest problems

Don't paste confidential, personal or regulated data into the hosted service; see Why Select DeepSeek AI? for when to self-host instead.

14Use It

Best DeepSeek Prompts and a Free Prompt Builder

Better prompts get better answers. Learn the five parts of a strong prompt, build your own in seconds, or copy one of our ready-made prompts into the official DeepSeek chat.

1
Role

Who should DeepSeek act as?

"Act as a senior tax adviser…"
2
Task

What exactly do you want done?

"…explain the new VAT rules…"
3
Context

Background it can't guess.

"…for a small café in Colombo…"
4
Format

How the answer should look.

"…as 5 bullet points…"
5
Constraints

Limits and things to avoid.

"…under 150 words, no jargon."
✗ Vague

"Write about marketing."

DeepSeek has to guess the audience, length, angle and format, so you get a generic essay.
✓ Specific

"Act as a marketing consultant. Suggest 5 low-budget Instagram ideas for a new bakery in Kandy aimed at university students. Give each a one-line hook and an estimated cost. Keep it under 200 words."

Role, task, context, format and constraints are all there, so the answer is usable straight away.

Prompt builder

Your prompt

Tips for DeepSeek
  • Turn on deeper thinking (or set effort to high or max on the API) for maths, logic and coding; keep it off for quick chats.
  • V4.1 Flash can read images, so attach screenshots, charts or photos instead of describing them.
  • On the API, keep long fixed instructions at the start of every request so they're cached and billed cheaply.
  • If the answer misses, reply with what to change ("shorter", "add costs") rather than starting over.

Prompt library

Writing
Blog post outline

Act as an experienced content editor. Create a detailed outline for a 1,200-word blog post titled "[TITLE]" for [AUDIENCE]. Include an intro hook, 5–7 H2 sections with bullet points, and a conclusion with a call to action.

Writing
Rewrite in a clearer tone

Rewrite the text below so it is clear, friendly and easy to read for a general audience. Keep every fact, cut filler, and keep it under [N] words.

[PASTE TEXT]

Writing
Professional email reply

Draft a polite, concise reply to the email below. My goal is to [GOAL]. Offer two subject-line options and keep the body under 150 words.

[PASTE EMAIL]

Coding
Explain this code

Explain what the code below does, line by line, for a junior developer. Then list any bugs, edge cases or security issues you see, most serious first.

[PASTE CODE]

Coding
Write a function with tests

Write a [LANGUAGE] function that [WHAT IT SHOULD DO]. Include type hints, a docstring, input validation, and unit tests covering normal cases and edge cases.

Coding
Debug an error

I get this error when running my program. Explain the likely cause, show how to confirm it, and give the smallest fix.

Error:
[PASTE ERROR]

Code:
[PASTE CODE]

Study
Explain like I'm new

Explain [TOPIC] to a complete beginner in under 200 words, using one everyday analogy. Then give three quick questions to check my understanding.

Study
Make flashcards

Turn the notes below into 15 question-and-answer flashcards. Keep answers to one sentence and focus on facts likely to appear in an exam.

[PASTE NOTES]

Study
Check my working

Here is my solution to a maths problem. Don't redo it from scratch: find the first step where I went wrong, explain why, and show the correct step.

[PASTE PROBLEM AND SOLUTION]

Business
SWOT analysis

Create a SWOT analysis for [BUSINESS] in [MARKET/LOCATION]. Give four points per quadrant, each with a one-line reason, then the two most important actions to take next.

Business
Meeting notes to actions

Turn the meeting transcript below into: a 3-sentence summary, a table of action items (owner, task, deadline), and open questions.

[PASTE TRANSCRIPT]

Analysis
Extract data as JSON

Extract every [ITEM TYPE] from the text below. Return only valid JSON: an array of objects with the fields [FIELD 1], [FIELD 2], [FIELD 3]. Use null for anything missing.

[PASTE TEXT]

Analysis
Compare two options

Compare [OPTION A] and [OPTION B] for [USE CASE]. Make a table covering cost, ease of use, performance and risks, then recommend one and explain when the other would be better.

Images
Describe a chart

Look at this chart image. Describe what it shows in plain language, extract the data as a table, and point out any misleading scales or missing labels. (Attach the image; works with V4.1 Flash.)

Images
Screenshot to steps

This is a screenshot of an app screen. Write numbered steps telling a new user how to [TASK] from this screen. (Attach the screenshot; works with V4.1 Flash.)

No prompts match your search.

Replace the [BRACKETED] parts with your own details. More techniques in our prompting guide. Never paste passwords, ID numbers or confidential data into any AI chat.

15Developers

DeepSeek API: Endpoints, Features & Quick Start

DeepSeek's API lets your own apps, scripts and agents call its models. It speaks both OpenAI and Anthropic formats, so most existing code works after changing two settings.

Key details

Base URL (OpenAI format)
https://api.deepseek.com
Base URL (Anthropic format)
https://api.deepseek.com/anthropic
Models
deepseek-flash (V4.1 Flash) · deepseek-v4-pro
Context / max output
1M tokens / 384K tokens
Concurrency limit
2,500 (Flash) · 500 (V4-Pro)
Billing
Prepaid balance, per token, off-peak 50% off
Old names
deepseek-v4-flash, -vision-exp → temporary aliases to Flash

What the API can do

💬
Chat Completions

OpenAI-compatible /chat/completions, with streaming.

🔁
Responses API

OpenAI Responses format, adapted for Codex-style agents.

🅰️
Anthropic format

Use Anthropic SDKs and tools via /anthropic.

🔧
Tool calls

Function calling for agents and workflow automation.

🧾
JSON output

Force valid JSON for extraction and structured data.

👁️
Vision

Send images to deepseek-flash (not V4-Pro).

🧠
Thinking mode

On by default; switch off or set low / high / max effort.

💾
Automatic caching

Repeated input is cached and billed ~98% cheaper on Flash.

✂️
FIM & prefix (beta)

Fill-in-the-middle and chat-prefix completion, non-thinking mode only.

Quick start: OpenAI-format code

Python
Node.js
cURL
# pip install openai
from openai import OpenAI

# DeepSeek uses an OpenAI-compatible API
client = OpenAI(
  api_key="<your-deepseek-api-key>",
  base_url="https://api.deepseek.com"
)

response = client.chat.completions.create(
  model="deepseek-flash", # V4.1 Flash, or deepseek-v4-pro
  messages=[
    {"role": "system", "content": "You are a helpful assistant."},
    {"role": "user", "content": "Explain MoE architecture"}
  ],
  stream=False
)

print(response.choices[0].message.content)
// npm install openai
import OpenAI from 'openai';

const client = new OpenAI({
  apiKey: '<your-deepseek-api-key>',
  baseURL: 'https://api.deepseek.com',
});

const completion = await client.chat.completions.create({
  model: 'deepseek-flash',
  messages: [
    { role: 'user', content: 'Hello!' }
  ],
});

console.log(completion.choices[0].message.content);
# Replace YOUR_KEY with your API key
curl https://api.deepseek.com/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_KEY" \
  -d '{
    "model": "deepseek-flash",
    "messages": [
      {"role":"user","content":"Hello!"}
    ]
  }'

Already using the Anthropic SDK?

Point it at DeepSeek's Anthropic-format endpoint:

# pip install anthropic
import anthropic

client = anthropic.Anthropic(
    api_key="<your-deepseek-api-key>",
    base_url="https://api.deepseek.com/anthropic",
)
msg = client.messages.create(
    model="deepseek-flash",
    max_tokens=1024,
    messages=[{"role": "user", "content": "Hello!"}],
)
print(msg.content[0].text)
OpenAI-format examples (Python, Node, cURL) ↑

Works with popular coding agents

DeepSeek's docs include setup guides for these tools:

CodexClaude CodeOpenCodeOpenClawHermesWorkBuddy / CodeBuddyQoderReasonixDeepSeek Harness

Good practice

  • Keep your API key on the server, never in browser or app code.
  • Put stable instructions first so they get cached and billed cheaply.
  • Run batch jobs off-peak for half price.
  • Call deepseek-flash, not the old aliases, which are temporary.
  • Handle rate-limit errors with retries and back-off.

Details from DeepSeek's official Models & Pricing page and change log. Our API guide goes deeper.

16Developers

DeepSeek Pricing: API Costs and Token Calculator

Current DeepSeek API pricing for V4.1 Flash and V4-Pro, how tokens are billed, and a calculator showing what your own usage will cost. The web chat and apps are free.

Official API prices

💡 No monthly fees. Cost = (input tokens × input price) + (output tokens × output price). Peak hours: 01:00–04:00 and 06:00–10:00 UTC, Monday–Friday (excluding Chinese public holidays); all other times are off-peak at 50% of peak. Cache hits cut input cost by ~98% on deepseek-flash.
Model (API name) Context Input (Cache Hit) Input (Cache Miss) Output Status
V4.1 Flash (deepseek-flash) 1M / 384K out $0.003 ($0.006 peak) $0.15 ($0.30 peak) $0.60 ($1.20 peak) Current · vision · 2,500 concurrency
V4-Pro-0813 (deepseek-v4-pro) 1M / 384K out $0.022 ($0.044 peak) $0.66 ($1.32 peak) $1.98 ($3.96 peak) Current · text only · 500 concurrency
deepseek-v4-flash / deepseek-v4-flash-vision-exp 1M Retired 10 Sep 2026 · served by V4.1 Flash at Flash prices Temporary alias
deepseek-chat / deepseek-reasoner 128K Retired 24 Jul 2026 · weights still free to self-host Retired

USD per 1M tokens, off-peak (peak in brackets), from DeepSeek's official Models & Pricing page. Flash prices effective 04:00 UTC, 10 Sep 2026. DeepSeek can change prices at any time, so always check the official page before budgeting.

How token costs work

What is a token?

A token is a small chunk of text, often a word or part of a word. In English, 1,000 tokens is roughly 750 words. You pay for the tokens you send (input) and the tokens the model writes back (output).

  • Cache hit: input DeepSeek has seen recently, such as a repeated system prompt, billed at about 2% of the normal input price on Flash.
  • Cache miss: new input, billed at the standard input rate.
  • Output: everything the model writes, including its thinking when thinking mode is on.
  • Peak hours (01:00–04:00 and 06:00–10:00 UTC, weekdays) cost 2×. Everything else is off-peak.
Cost = (cached input × hit price) + (new input × miss price) + (output × output price)

Token cost calculator

Per request—
Per day—
Per month (30 days)—

Uses DeepSeek's published USD list prices; actual bills can differ. Check the official pricing page before budgeting.

What everyday tasks cost (off-peak)

TaskQuantityV4.1 FlashV4-Pro
💬One chat message
~500 tokens in · ~300 out
per 1,000 messages$0.255$0.924
📄Summarise a 20-page PDF
~12K in · ~800 out
per 100 documents$0.228$0.950
🖼️Read an invoice image
~2K in · ~400 out
per 1,000 invoices$0.540$2.11
🤖One coding-agent session
~2M in (90% cached) · ~40K out
per session$0.059$0.251
📚Analyse a whole book
~600K in · ~4K out
per book$0.092$0.404

Token counts are rough estimates; real usage depends on language, file content, and how much the model thinks. Peak-hour costs are double. The free official chat and apps cost nothing; these prices apply only to the API.

17Help

DeepSeek Not Working? How to Fix Common Problems

"Server is busy", login trouble, app crashes or API errors? Start with a quick status check, then pick where you're having the problem for step-by-step fixes.

Quick 4-step check

  1. Is it down for everyone? Check DeepSeek's official status page.
  2. Wait and retry. "Server busy" usually clears within minutes.
  3. Refresh or restart the page or app, and check your internet.
  4. Try another route: web instead of app, or a different browser or network.
Check DeepSeek status ↗
"Server is busy. Please try again later."
  1. This is DeepSeek's capacity limit, not a problem with your device. Wait 30–60 seconds and send again.
  2. Check status.deepseek.com for an outage.
  3. Busy spells are worst right after big launches and in Asian daytime hours, so try again later.
  4. If you need answers now, shorten the request or start a fresh chat.
The page won't load or stays blank
  1. Hard-refresh the page (Ctrl/Cmd + Shift + R).
  2. Clear cookies and cache for chat.deepseek.com, or try a private window.
  3. Turn off ad blockers, VPNs or privacy extensions one at a time.
  4. Try a different browser or network (e.g. switch Wi-Fi to mobile data).
I can't log in or the verification code never arrives
  1. Check your spam or promotions folder.
  2. Wait a minute before tapping "resend"; repeated requests can be throttled.
  3. Double-check the email address or phone number for typos.
  4. Try another sign-in option offered on the login page.
The answer stops halfway or is cut off
  1. Type "continue" to pick up where it stopped.
  2. Very long chats get heavy, so start a new chat and paste a short summary.
  3. Split huge documents into smaller parts.
  4. Regenerate the response if it looks broken.
File or image upload fails
  1. Check the file type and size; try a smaller or compressed version.
  2. Paste the text directly instead of uploading a scanned PDF.
  3. Re-upload after refreshing the page.
I can't find the app in my store
  1. DeepSeek's app may not be offered in every country or on every store.
  2. Search for the app from developer "DeepSeek" only; avoid look-alike apps.
  3. Android users can get the APK from DeepSeek's official download page; never from other websites.
  4. Use the web chat in your phone's browser instead.
The app crashes or won't open
  1. Update the app and your phone's operating system.
  2. Restart your phone.
  3. Android: Settings → Apps → DeepSeek → Clear cache.
  4. Still crashing? Uninstall and reinstall, then log in again.
Chats aren't syncing between devices
  1. Make sure you're logged in to the same account everywhere.
  2. Pull down to refresh the chat list.
  3. Check your internet connection, then log out and back in.
Voice input or camera doesn't work
  1. Allow microphone and camera permission for DeepSeek in your phone's settings.
  2. Close other apps using the mic or camera.
  3. Restart the app after changing permissions.
CodeMeaningFix
400Invalid Format
Request body is malformed
Fix the body using the hints in the error message.
401Authentication Fails
Wrong API key
Check the key, or create a new one on the platform.
402Insufficient Balance
Account has run out of credit
Top up your balance.
422Invalid Parameters
A parameter isn't valid
Correct it using the error message hints.
429Rate Limit Reached
Too many concurrent requests
Slow down and retry with back-off; limits are 2,500 (Flash) / 500 (Pro) per account.
500Server Error
Problem on DeepSeek's side
Wait briefly and retry; contact DeepSeek if it persists.
503Server Overloaded
High traffic
Wait briefly and retry.
Requests hang for a long time
  1. DeepSeek keeps slow connections open instead of failing them: non-streaming calls receive empty lines and streaming calls receive ": keep-alive" comments while waiting.
  2. Set a generous client timeout; DeepSeek closes the connection only if inference hasn't started after 10 minutes.
  3. If you parse raw HTTP yourself, ignore those keep-alive lines.
  4. Use streaming so users see output as soon as it starts.
"Model not found" or unexpected model
  1. Use deepseek-flash or deepseek-v4-pro.
  2. deepseek-chat and deepseek-reasoner were retired on 24 Jul 2026.
  3. deepseek-v4-flash still works only as a temporary alias, so switch now.
Images are ignored
  1. Only deepseek-flash (V4.1 Flash) supports vision; V4-Pro is text-only.
Bills are higher than expected
  1. Peak hours (01:00–04:00 and 06:00–10:00 UTC, weekdays) cost double.
  2. Thinking tokens count as output, so lower the effort for simple tasks.
  3. Keep fixed instructions at the start of each request to get cache-hit pricing.
Still stuck? Contact DeepSeek directly through the help options in its app or website, or email its API team at api-service@deepseek.com.
deepseeksr1.com is an independent guide and can't access or fix DeepSeek accounts.
03Overview

DeepSeek News: Latest Updates

The latest DeepSeek news: new models, API changes, pricing and company updates, checked against DeepSeek's official change log.

10 Sep 2026 · Full details now public

DeepSeek-V4.1-Flash: 552B open-weight MoE with native vision, and V4-Pro isn't going anywhere yet

DeepSeek's release notes and technical report fill in the specs for V4.1 Flash: a 552B-parameter mixture-of-experts with a new causal encoder–decoder design that activates just 8B parameters to read input and 16B to write output. It understands images natively, keeps the 1M-token context, and its KV cache needs about 1/4 of the HBM and 1/8 of the SSD of the previous generation. Weights and the technical report are on Hugging Face.

On the API it is called deepseek-flash; V4-Flash and V4-Flash-Vision-Exp are retired and their old names temporarily route to V4.1 Flash. DeepSeek originally planned to route all V4-Pro traffic to V4.1 Flash from 14 September, but reversed that on 11 September after user pushback: deepseek-v4-pro keeps running at its existing price until further notice.

90.6
Terminal-Bench 2.1 (V4-Pro: 87.9)
88.1
CyberGym (V4-Pro: 83.3)
$0.15
Uncached input / 1M off-peak
$0.60
Output / 1M off-peak

Benchmarks are DeepSeek's own published figures. Third-party analysis of the model card notes V4.1 Flash leads V4-Pro on agentic tests but trails it on most knowledge-heavy base-model tests.

06 OCT 2026
DeepSeek reportedly nears a ~$12B funding round

Bloomberg and Reuters report DeepSeek is close to raising at least 80 billion yuan, with Tencent and CATL among the largest backers, well above its original ~50 billion yuan target. Sources tie the extra demand to the V4.1 Flash launch. DeepSeek has not announced the round; an IPO on Shanghai's STAR Market is reportedly being prepared for early 2027.

04 OCT 2026
Analyst: US–China AI gap at a record low

A Bloomberg Intelligence report says top Chinese models now trail US rivals by about 3% on benchmark scores after V4.1 Flash, down from roughly 9% in May.

LATE SEP 2026
DeepSeek invited to UN Security Council AI briefing

Reuters reported DeepSeek would join OpenAI, Anthropic and other labs in briefing the Security Council on AI risks. DeepSeek also announced work with Huawei on programming tools for Ascend AI chips.

11 SEP 2026
U-turn: V4-Pro API stays online

"In response to user demand," DeepSeek said it will keep serving V4-Pro after 14 September with billing unchanged, instead of routing it to V4.1 Flash. A V4.1 Pro is confirmed to be in development, with no date given.

10 SEP 2026
V4.1 Flash launches; China rejects distillation advisory

New Flash prices took effect at 04:00 UTC. The same day, China's Commerce Ministry rejected the US advisory (below), calling distillation a common industry practice and warning of countermeasures if it is used to target Chinese firms.

09 SEP 2026
US agencies accuse Chinese labs of "systematic" distillation

An NSA, FBI and CISA advisory named DeepSeek, Alibaba, Moonshot AI, MiniMax, StepFun and Z.AI, alleging they extract US models in breach of terms of use.

EARLIER
Model releases before September

See the model timeline for V3, R1, V3.1, V3.2 and V4.

18Help

DeepSeek FAQ: Frequently Asked Questions

Quick answers to the most common questions about DeepSeek AI.

What is DeepSeek V4.1 Flash? +

DeepSeek-V4.1-Flash, released 10 September 2026, is the smallest model in DeepSeek's new architecture family. It is a 552B-parameter mixture-of-experts with a causal encoder–decoder design (8B active parameters for input, 16B for output), native image understanding, a 1M-token context and up to 384K output. Its KV cache needs about 1/4 the HBM and 1/8 the SSD of the previous generation. On the API it is called deepseek-flash, and the weights are on Hugging Face.

Is DeepSeek V4 Pro being shut down? +

Not for now. DeepSeek initially announced that all deepseek-v4-pro requests would be routed to V4.1 Flash from 14 September 2026. On 11 September it reversed that, saying that in response to user demand it would keep serving V4-Pro with unchanged billing and give notice of any future change. A V4.1 Pro is in development, but no date has been announced.

Should I use V4.1 Flash or V4-Pro? +

Start with V4.1 Flash. It is cheaper (about 70% less per output token), supports images, has higher concurrency limits, and DeepSeek reports it ahead of V4-Pro on agentic coding benchmarks. Consider V4-Pro for knowledge-heavy work where it still scores higher, such as HLE without tools (42.7 vs 36.8).

Is DeepSeek free to use? +

Yes. DeepSeek's official web chat at chat.deepseek.com and its official mobile apps are free. The API is pay-per-token: off-peak, deepseek-flash costs $0.15 per million input tokens ($0.003 with cache hits) and $0.60 per million output tokens. See our DeepSeek Web guide.

What happened to deepseek-v4-flash and the Vision-Exp model? +

Both were retired on 10 September 2026. For compatibility, the names deepseek-v4-flash and deepseek-v4-flash-vision-exp temporarily route to V4.1 Flash and are billed at Flash prices. DeepSeek describes this as temporary, so update your code to deepseek-flash.

Is DeepSeek safe to use for business or sensitive data? +

For non-sensitive work the hosted API is widely used. However, DeepSeek is a Chinese company, data sent to its API may be stored on servers subject to Chinese law, and a September 2026 US security advisory accused it and other Chinese labs of unauthorised model distillation. For healthcare, finance, legal or government data, self-host the open weights or use a third-party cloud host with appropriate compliance and data-residency guarantees.

Can I run DeepSeek locally on my computer? +

Yes, with the right hardware. On a laptop or single GPU, use Ollama or LM Studio with a distilled or quantized model (1.5B–70B, 8–48GB VRAM). The full V4-Flash (284B) fits a single 8-GPU server; V4.1 Flash (552B) and V4-Pro (1.6T) need multi-GPU clusters. See our self-hosting guide.

When is the next DeepSeek model coming? +

DeepSeek has confirmed that V4.1 Pro is in development but has given no release date, specs or pricing. We'll update this page as soon as DeepSeek's official change log confirms a release.

Is this the official DeepSeek website? +

No. deepseeksr1.com is an independent guide and is not affiliated with DeepSeek. We don't offer chat, accounts, API keys or app downloads. Use deepseek.com, chat.deepseek.com and platform.deepseek.com for DeepSeek's own services.

Get Started

Start Using DeepSeek AI Today

Try V4.1 Flash free on DeepSeek's official chat, or dig into our guides for building with the API.

Official DeepSeek Chat ↗ Our API Guide GitHub ↗