DeepSeek V4 Flash Chat
deepseek/deepseek-chat
Paid V4 Flash in non-thinking mode (1.6T-class quality at $0.14 in / $0.28 out). Production-grade reliability and 5MB request bodies.
Code Examples
from blockrun_llm import LLMClient
client = LLMClient() # Uses BLOCKRUN_WALLET_KEY (never sent to server)
response = client.chat("deepseek/deepseek-chat", "Hello!")import { LLMClient } from '@blockrun/llm';
const client = new LLMClient(); // Uses BLOCKRUN_WALLET_KEY (never sent to server)
const response = await client.chat('deepseek/deepseek-chat', 'Hello!');curl -X POST https://blockrun.ai/api/v1/chat/completions \
-H "Content-Type: application/json" \
-H "PAYMENT-SIGNATURE: <payment_header>" \
-d '{
"model": "deepseek/deepseek-chat",
"messages": [{"role": "user", "content": "Hello!"}],
"max_tokens": 1024
}'Pricing
No markup on official DeepSeek rates; $0.001 fee per call.
Payment
Pay per request with USDC on Base. No subscription required.
Try It
Send a message to try DeepSeek V4 Flash Chat
Connect your wallet to enable payments
About DeepSeek V4 Flash Chat
DeepSeek V4 Flash Chat is a coding model from DeepSeek with a 1.05M-token context window and up to 66K tokens of output per call. It is billed per token at the provider's list rate, with no markup on chat tokens. It is served through BlockRun's OpenAI-compatible API, so it can be called without an account, an API key, or a subscription.
What it costs
DeepSeek V4 Flash Chat is billed per token. The underlying rates are $0.14 per million input tokens and $0.28 per million output tokens with no platform margin on chat tokens; only a small per-transaction fee covering on-chain settlement is added on top, and the resulting price is quoted in the 402 before you pay. You are charged for the tokens a request actually consumes — there is no monthly commitment and no unused balance to forfeit. With a context window of 1,048,576 tokens and up to 65,536 tokens of output, a single call can carry a substantial working set.
Specifications
- Context window
- 1,048,576 tokens
- Maximum output
- 65,536 tokens
- API compatibility
- OpenAI-compatible
- Payment
- USDC via x402
- Categories
- chat, coding
Calling it from your code
Pass deepseek/deepseek-chat as the model field. Because the endpoint mirrors the OpenAI chat completions schema, any existing OpenAI client works by changing the base URL — streaming, tool use, and multi-turn messages all behave the same way.
The first request returns a 402 carrying the exact price; a signed retry runs it, and client libraries fold the two into one call — how the payment works. See the documentation for request and response shapes, the LLM API page for every model on this endpoint, or browse the full catalog of 78 chat models to compare alternatives.
Related models
Every model on BlockRun is reached through the same endpoint and the same payment flow, so switching is a one-line change to the model field. These are the closest alternatives to DeepSeek V4 Flash Chat:
- DeepSeek V4 Flash Vision$0.44 in / $1.32 out per 1M
- DeepSeek V4 Pro$1.32 in / $3.96 out per 1M
- Moonshot Kimi K3$3.00 in / $15.00 out per 1M
- Z.ai GLM-5.3$1.40 in / $4.40 out per 1M
- Z.ai GLM-5.3 Flash$0.15 in / $0.50 out per 1M
- DeepSeek V4 Flash Reasoner$0.14 in / $0.28 out per 1M