DeepSeek V4 Flash is free to use until Aug 12. Get an Endpoint →
    All models
    DeepSeekDeploy

    DeepSeek-V4-Flash-0731

    Native Responses API support and Codex compatibility for upgraded agent and coding workflows.

    Context length
    GPU requirementNot published
    Concurrency
    API model IDDeepSeek-V4-Flash-0731
    Input pricingNot published
    Output pricing$0.14 / 1M
    Cached input pricing$0.0014 / 1M

    Recommended use cases

    Where this model fits.

    Supported features

    Ready for production.

    OpenAI compatible
    Scale to zero

    API example

    Use the interface you already know.

    Every published text endpoint uses an OpenAI-compatible request surface.

    curl
    curl https://api.inferx.com/v1/chat/completions \
      -H "Authorization: Bearer $INFERX_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{ "model": "DeepSeek-V4-Flash-0731", "messages": [...] }'