whichllm — Browse and compare AI model specs and pricing

InferX

Qwen3-Coder-Next-FP8-no-thinking model ID, context window & pricing

qwen

Quick facts

Model ID Qwen3-Coder-Next-FP8-no-thinking
Source InferX
Context Window 260000
Pricing -
Capabilities tool calling, structured output, temperature control, open weights

Model overview

Qwen3-Coder-Next-FP8-no-thinking is an AI model from InferX with 260000 token context window and text input support.

Public token pricing is not listed for this model in the current catalog source.

  • Workloads that use text inputs with text outputs.
  • Agent and tool workflows that need function calling.
Model ID Qwen3-Coder-Next-FP8-no-thinking
Provider InferX
Family qwen
Status -
Knowledge Cutoff 2025-04
Release Date 2026-02-03
Input Modalities text
Output Modalities text
Context Window 260000
Input Limit -
Output Limit 65536
Tool Calling Yes
Reasoning No
Structured Output Yes
Temperature Control Yes
Open Weights Yes
Input Cost / 1M tokens -
Output Cost / 1M tokens -
Reasoning Cost / 1M tokens -
Cache Read Cost / 1M tokens -
Cache Write Cost / 1M tokens -