Models

DeepSeek V4 Flash (Reasoning, Max Effort)

DeepSeek V4 Flash is a fast-response model optimized for reasoning tasks. It is designed to deliver high-quality reasoning outputs with maximum computational effort while maintaining low latency.

FastReasoning
Input / 1M tokens
$0.14
Output / 1M tokens
$0.28
Output tokens/s
102.4
First-token seconds
0.87s
Supported plans
40

Benchmark history

Evaluations

11

Tau Banking

Measured Jul 10, 2026Source

Score

0.23

TAU2

Measured Jul 10, 2026Source

Score

0.95

Terminalbench V2 1

Measured Jul 10, 2026Source

Score

0.62

Terminalbench Hard

Measured Jul 10, 2026Source

Score

0.36

Lcr

Measured Jul 10, 2026Source

Score

0.63

Ifbench

Measured Jul 10, 2026Source

Score

0.79

Scicode

Measured Jul 10, 2026Source

Score

0.45

Hle

Measured Jul 10, 2026Source

Score

0.32

Gpqa

Measured Jul 10, 2026Source

Score

0.89

Artificial Analysis Coding Index

Measured Jul 10, 2026Source

Score

56.2

Artificial Analysis Intelligence Index

Measured Jul 10, 2026Source

Score

40.3

Plan availability

Products and plans that support this model

12
Apertis Coding Plan

Apertis Coding Plan

Apertis Coding Plan is a subscription-based AI coding service providing unified access to 30+ AI models (GPT-5.4, Claude Opus 4.6, Gemini 3.1 Pro, and more) through a single API key. Designed for developers using coding agents like Claude Code, Cursor, Cline, and OpenCode, it offers predictable monthly pricing, free prompt caching, auto-failover, and quota-based billing across OpenAI, Anthropic, Google, and other providers.

OpenCode Go

OpenCode Go

OpenCode Go is a low-cost monthly subscription that provides reliable access to a curated set of powerful open-source coding models, such as GLM-5.1, Kimi K2.6, and DeepSeek V4 Pro, for use with AI coding agents like OpenCode. Priced at $10/month with a $5 first month, it offers generous request limits to support developers.

SenseNova Token Plan

SenseNova Token Plan

SenseNova Token Plan is a developer subscription offering curated models and generous usage quotas for office productivity and agentic workflows. It provides access to SenseNova's own multimodal models (SenseNova 6.7 Flash-Lite, SenseNova U1 Fast) plus DeepSeek V4 Flash via an OpenAI-compatible API endpoint at https://token.sensenova.cn/v1.

User ratings

Loading ratings...

Discussion

Thinking... Make sure you are connected to GitHub server