Deployments

    Model Guides

    Select, configure, and access models through the InferX Console.

    Published endpoints

    Published endpoint pages expose model metadata and identifiers. Sign in when you are ready to copy the tenant-scoped URL, API key, and runnable integration flow.

    Dedicated model deployment

    The existing Console supports Deploy From Catalog and Build From Scratch. The configuration flow includes the model identifier, vLLM image, runtime parameters, GPU count, and vRAM.

    Gated model access

    For gated Hugging Face models, accept the model license and provide HF_TOKEN through the Console’s advanced environment configuration.