whichllm — Browse and compare AI model specs and pricing

NanoGPT

DeepSeek V4.1 Flash Thinking model ID, context window & pricing

deepseek

Quick facts

Model ID deepseek/deepseek-v4.1-flash:thinking
Source NanoGPT
Context Window 1000000
Pricing $0.16 input / $0.31 output per 1M tokens
Capabilities tool calling, reasoning, structured output

Model overview

DeepSeek V4.1 Flash Thinking is an AI model from NanoGPT with 1000000 token context window and text, image input support.

Published pricing is $0.16 input and $0.31 output per 1M tokens.

  • Workloads that use text, image inputs with text outputs.
  • Agent and tool workflows that need function calling.
  • Reasoning-heavy prompts where stepwise problem solving matters.
Model ID deepseek/deepseek-v4.1-flash:thinking
Provider NanoGPT
Family deepseek
Status -
Knowledge Cutoff -
Release Date 2026-09-08
Input Modalities text, image
Output Modalities text
Context Window 1000000
Input Limit 1000000
Output Limit 384000
Tool Calling Yes
Reasoning Yes
Structured Output Yes
Temperature Control -
Open Weights No
Input Cost / 1M tokens $0.16
Output Cost / 1M tokens $0.31
Reasoning Cost / 1M tokens -
Cache Read Cost / 1M tokens $0.03
Cache Write Cost / 1M tokens -