All models
    DeepSeekReady now

    DeepSeek V4 Flash Preview

    Context length1,000,000
    GPU requirementNot published
    Concurrency
    API model IDdeepseek-v4-flash
    Input pricing$0.14$0 / 1MFree
    Output pricing$0.28$0 / 1MFree
    Cached input pricing$0.028$0 / 1MFree

    Recommended use cases

    Where this model fits.

    Supported features

    Ready for production.

    OpenAI compatible
    Scale to zero

    API example

    Use the interface you already know.

    Every published text endpoint uses an OpenAI-compatible request surface.

    curl
    curl https://api.inferx.com/v1/chat/completions \
      -H "Authorization: Bearer $INFERX_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{ "model": "deepseek-v4-flash", "messages": [...] }'