---
title: How to scale dedicated Generative APIs deployments
description: This page explains how to scale dedicated Generative APIs deployments in size
tags: generative-apis-dedicated-deployment ai-data ip-address
dates:
  validation: 2026-04-16
  posted: 2025-06-03
---
import Requirements from '@macros/iam/requirements.mdx'


You can scale your dedicated Generative APIs deployment up or down to match it to the incoming load of your deployment.

<Message type="important">
This feature is currently in [Public Beta](https://www.scaleway.com/en/betas/).
</Message>

<Requirements />

  - A Scaleway account logged into the [console](https://console.scaleway.com)
  - A [dedicated Generative APIs deployment](/generative-apis/how-to/create-deployment/)
  - [Owner](/iam/concepts/#owner) status or [IAM permissions](/iam/concepts/#permission) allowing you to perform actions in the intended Organization

1. Click **Generative APIs** in the **AI** section of the side menu in the [Scaleway console](https://console.scaleway.com/) to access the dashboard. The list of models displays.
2. Select the **Deployments** tab.
3. From the drop-down menu, select the geographical region you want to manage.
4. Click a deployment name to access the deployment's dashboard.
5. Click the **Settings** tab and navigate to the **Scaling** section.
6. Click **Update node count** and adjust the number of nodes in your deployment.
    <Message type="note">
      High availability is only guaranteed with two or more nodes.
    </Message>
7. Click **Update node count** to update the number of nodes in your deployment.
    <Message type="note">
      Changes may take 15-30 minutes to apply depending on the model size. Your deployment remains available during the process. 
    </Message>
