Deepseek V4 Flash

deepseek/deepseek-v4-flash

DeepSeek-V4-Flash is a lightweight model meticulously designed by DeepSeek to deliver the ultimate combination of lightning-fast response times and unmatched cost-effectiveness. Engineered with fewer parameters and significantly lower activation overhead, V4-Flash provides an exceptionally fast and economical API service. At its core, V4-Flash demonstrates outstanding reasoning capabilities that closely rival the V4-Pro model. While featuring a slightly streamlined repository of world knowledge, it remains highly capable of satisfying the demands of most application scenarios. In Agentic applications, V4-Flash performs on par with the Pro version when handling standard and fundamental tasks. As the premier choice for developers prioritizing high concurrency, low latency, and cost efficiency, DeepSeek-V4-Flash serves as the optimal solution for deploying large-scale, high-frequency, and lightweight AI workloads.

機能

サーバーレス API

ドキュメント

deepseek/deepseek-v4-flash is available via Novita's serverless API, where you pay per token. There are several ways to call the API, including OpenAI-compatible endpoints with exceptional reasoning performance.

利用可能なサーバーレス

クエリをすぐに実行し、使用した分だけお支払い

入力$0.14 / M Tokens

キャッシュ読み取り$0.028 / M Tokens

出力$0.28 / M Tokens

以下のコード例を使用して、当社の API と統合してください:

1from openai import OpenAI
2
3client = OpenAI(
4    api_key="<Your API Key>",
5    base_url="https://api.novita.ai/openai"
6)
7
8response = client.chat.completions.create(
9    model="deepseek/deepseek-v4-flash",
10    messages=[
11        {"role": "system", "content": "You are a helpful assistant."},
12        {"role": "user", "content": "Hello, how are you?"}
13    ],
14    max_tokens=393216,
15    temperature=0.7
16)
17
18print(response.choices[0].message.content)

情報

プロバイダー

DeepSeek

量子化

fp8

サポートされている機能

コンテキスト長

1048576

最大出力

393216

Serverless

サポートされています

Function Calling

サポートされています

Structured Output

サポートされています

Reasoning

サポートされています

Anthropic API

サポートされています

入力機能

text

出力機能

text

本番環境向けAIを構築するために必要なすべて。

200以上のモデル、オンデマンド GPUs、安全なエージェントランタイムを、1つの API に統合。無料で始められ、成長に合わせてスケールできます。