whichllm — Browse and compare AI model specs and pricing

InferX

Gemma 4 31B IT FP8 model ID, context window & pricing

gemma

Quick facts

Model ID gemma-4-31B-it-fp8
Source InferX
Context Window 262144
Pricing -
Capabilities tool calling, reasoning, structured output, temperature control, open weights

Model overview

Gemma 4 31B IT FP8 is an AI model from InferX with 262144 token context window and text, image input support.

Public token pricing is not listed for this model in the current catalog source.

  • Workloads that use text, image inputs with text outputs.
  • Agent and tool workflows that need function calling.
  • Reasoning-heavy prompts where stepwise problem solving matters.
Model ID gemma-4-31B-it-fp8
Provider InferX
Family gemma
Status -
Knowledge Cutoff -
Release Date 2026-04-02
Input Modalities text, image
Output Modalities text
Context Window 262144
Input Limit -
Output Limit 32768
Tool Calling Yes
Reasoning Yes
Structured Output Yes
Temperature Control Yes
Open Weights Yes
Input Cost / 1M tokens -
Output Cost / 1M tokens -
Reasoning Cost / 1M tokens -
Cache Read Cost / 1M tokens -
Cache Write Cost / 1M tokens -